The search space for the game of Go was also thought to be too large for computers to manage.
[1] https://www.vice.com/en/article/a-human-amateur-beat-a-top-g...
This is far from unsolvable. It just means that the "apply RL like AlphaGo" attitude is laughably naive. We need at least one more trick.
As you said brute forcing the search space as the starting procedure would take way too long for the AI to build intuition.
But if we could give it a million or so lemmas of human math, that would be a great starting point.