Self-play GPT (by bots in a rich simulation) similar to Alpha Go Zero?
We might end up with more regularised language, and a more consistent model of the world, but that would come at the expense of accuracy and faithfulness (two things which are already lacking).