In the end an "idea" is by definition contrarian which is the opposite of the training objective of LLMs. The question is how far fine-tuning and tree-search can go to extrapolate from the data-manifold. And the answer is probably not that far, currently.