It did a really good -- surprisingly good -- job. That incident has been a reference point for me. Even if it is anecdotal.
It feels like they've mastered language, but it's looking more and more like they've actually mastered canon. Which is still impressive, but very different.
We are warned in statistics to be careful when extrapolating from a regression analysis.
I think LLMs do great summaries. I am not able to come up with anything where I could criticize it and say "any human would come up with a better summary". Are my tasks not "truly novel"? Well, then I am not able, as a human, to come up with anything novel either.
Depending how unique the text is determines how accurate the summarization is likely to be.