The body of literature on learning theory, and beyond that on specific types of learning and specific mediums such as learning from text is so rich there are way more useful models to draw from. Believe it or not, prellm, researchers in the textual learning field had already demonstrated you can achieve performance equal or better than novice tutors using pretty basic computer aids that follow specific hint/pump interaction structures. Guiding an LLM to use these findings has evidence backing it and is way better than telling it "i guess be like socrates". The problem is, to realize there might be richer more effective and highly researched ways of tackling the problem beyond the first fart of a thought you had one afternoon requires the deep respect for expertise and specialization that precisely basically everyone in the AI space right now fundamentally lacks.
You need much more time and guidance.
TLDR: Actual human tutoring sessions were recorded and analyzed and Socratic questioning was barely used at all. Instead the following pattern was observed:
Pump — "Uh huh?" "What else?" Costs nothing, so try it first.
Hint — points at the region of the answer. "What about the pumpkin's motion sideways?"
Prompt — fishes for one specific word, with the sentence frame supplied. "The pumpkin keeps moving forward at the same ___?"
Assertion — just says it. "It keeps the runner's horizontal velocity."
This statement would out you as someone who didn't attend an elite school.
Trying to diagrams/animations didn't yield good results even with frontier models. But pure text, any model does a decent job.
So it may be very slow or become unavailable, back end can't handle that, no caching whatsoever.