Ultimately its pretty difficult to distinguish between the two, so these sort of reports do need to be taken seriously. We have no idea what sort of caching mechanisms they have with ChatGPT, nor do we know how they handle memory.
I tried it via api two ways, both via system message and no user message and one with the word rewrite at the beginning and one without. Both gave similar long mishmashed outputs that seemed like a hallucination.
This has happened since the beginning and sometimes too does happen if you ask it to repeat the same thing 1000 times
It’s similar to how sometimes people just spew bullshit and then once they hear it they’re like no, wait, it’s like….
There isn’t a proponer closed loop on LLMs yet so we sometimes end up with terrible answers.