>It's important to remember that while Wikipedia is "The Free Encyclopedia that Anyone Can Edit,"
>It's important to remember that while Wikipedia is "The Free Encyclopedia that Anyone Can Edit,"
Given its nature as an LMM and a complex next word predictor, the phrase "it is important to remember" could be a way that was inadvertently trained so that it keeps itself on track.
If it has "some points, it is important to remember {something}, some more things" it may be able to better generate text compared to "some points, weird tangent".
Since it doesn't have a hidden memory, everything that it "thinks" is out there in the text including its own cues for what it should do. It also can't go back an edit its previous text to remove the self hints or clarify earlier points without calling them out. That style of writing differs from natural human writing since we are able to keep on topic (or not) without needing to write messages to ourselves that others can read.
When we do, it's pointed out rather than trying to slip it in casually. "Note to self or reader" for things that are to be pointed out and break the flow of the text or "as an aside" for the tangents.
I'm guessing that a human did edit this AI output, just not very well.
I'm still surprised to see that, as far as I can tell, no news outlets have made a public commitment to never use AI in their writing. Seems like it would be an easy way to promote the brand on commitment to quality.
I've griped about this before, but here we still are. We now know that MSN news, for instance, has no credibility due to their publishing AI-generated misinformation. https://news.ycombinator.com/item?id=39043135
not THE reason, not THE ONLY reason, so it is not COMPLETELY untrue. agitated?
Someone who's studied the subject quite a bit might visualize an actual Einstein chalkboard from memory. Their output could be verified and referenced, but looking for new Einstein-level math on the chalkboard would be madness.
Someone who Einstein himself would consider a peer might use this visualization method as a way to do actual work.
If we're assigning value to our chalkboards we'd be able to explain why we chose the numbers -1, 0, and 1. This would bias any math we'd do towards one of the chalkboards depending on our intent. At this point my chalkboard is useful as a filter.
Putting this all together, we'd see that the overall look of the cartoon responses would be a blend of 0 and 1 styles and would depend on our requests e.g. a reference request would look mostly like 0's art style. My own personal art style will be intentionally absent because it's only ever framing nonsense, by my own admission.
The worst thing about AI is how it can easily betray you, manipulate you and have swarms execute long term sleeper plans at scale!
> It's important to remember that while Wikipedia is "The Free Encyclopedia that Anyone Can Edit," it's hardly The Wild West.
My reading is that the article is average human prose (not great, not unreadable), not LLM prose.
It's not average human prose either(for one thing expecting GPT to converge on some average doesn't really make sense in the first place).
Base models with no rlhf or fine-tuning don't talk like that at all. This is specifically an artifact of the post-training fine-tuning/RLHF process
"It's important to remember" is a phrase that plenty of people used before ChatGPT. I've included three examples below from a quick time-boxed search. Just because ChatGPT says a phrase doesn't mean that every time you see it it came from ChatGPT—someone wrote the training data that ChatGPT was trained on, and someone else wrote the data (or selected the responses) that it was fine tuned on. ChatGPT isn't inventing new phrases out of whole cloth, it has a stereotyped style that is pieced together out of many existing phrases.
A collection of such stereotyped phrases in a single piece would be stronger evidence of GPT authorship, but I see no evidence of that here.
https://old.reddit.com/r/gravityfalls/comments/bgtnez/while_...
https://twitter.com/AsteadWH/status/1050813462673264640
https://stackoverflow.blog/2019/12/19/what-senior-developers...
Which is, of course, exactly what ChatGPT is trained to produce. A lot of people's mental AI detector is actually a mediocrity detector.
There's no average. Pre-training incentives being able to predict the smartest string of text in the corpus as readily as the dumbest. It doesn't converge on "average" and it doesn't really make sense that it would either.
Base models don't talk like GPT. This is strictly an artifact of post training fine-tuning/RLHF.
"It's important to remember to consult a mechanic" for example.
It's just some priming.
But as most Chess Grandmasters say, you only really need to use it in one difficult spot to change the result of a game.