It’s a task much harder to RL and much more subjective. I don’t want to say we won’t get there, but let’s just say that LLMs could “write” well enough since gpt3.5 era and I don’t think the pleasantness of the prose improved dramatically since then.
And subjectively the explanation LLMs currently provide are usually horrible, horrible enough that I usually just instruct them to provide me human written literature I can read.