Lines of code. 1,596 BTC gone
onekey.so
onekey.so
There’s this sort of local optima all these models pick which is instantly recognizable and hardly digestible for human consumption. Perhaps too much internal RL against benchmarks during chain of thought? The older non reasoning models had they’re own problems but at this point I can’t get an LLM to summarize data for humans, which is troubling.
As much as I’d like to read this piece it follows suit and I don’t have the patience to try and extract anything valuable from it.
Say this in a deep masculine dramatic voice:
One man. One gun. A story that would could never die.
Compare that with this Coldcard page: Two files, four lines, five years unnoticed. This is the COLDCARD
entropy failure
It's trying too hard to replicate a dramatic movie arc with every... single... sentence...what do you think these things are trained on? mountains of rubbish, human and likely also AI slop. how many bad movies and ads do u think they are fed. poor things. i can see them caged in the corporate HQ, with their tails shoved down their own throats rigged up to the SlopExtractor9000.
its time for the UN to Unslop the world. or something like that :(
"write (whatever you want written), BUT do not make it sound like a typical AI don't use any of the tropes or typical phrases that AI tends to use - if you include anything that makes it appear to be from AI then the task has failed. Pull your writing away from the median and toward very specific random styles that are NOT the cliched, averaged LLM style."
edit: actually wrong I tried to get ChatGPT to do this and it kept falling back into LLM speak even when it recognised that's what it was doing.
prompting it to be concise ("250 words or less") seems to help somewhat but theres only so much you can do esp in cases where the details matter and don't compress well
Before LLM, the majority of the web was content farms... but those are instantly recognizable and we seldom fall into them. Now LLM mixed the content farm style writing with _some_ real content. It take time and effort to tell they are slop, causing much fatigue.
To borrow a programming lang analogy the failure mode I see a lot is it invents its own jargon that is effectively like a compiler intermediate representation of the high level natural language you actually want a human to look at and then inserts it directly into what the human has to read.
It needs to stop doing that but it doesn't seem to do a good job at differentiating from what is or isn't chain of thought slop. I'm sure the jargon is useful during its reasoning but it's very unhelpful and not very nice to deliver it to a human.
It is possible to get the AI to write stuff that does not sound like AI you just have to ask it to properly.
> Six steps, no jargon. The note beside each one is the precise technical version, for anyone who wants to check the work.
It actually makes me uncomfortable reading this garbage now.
"But here's the part keeping me up at night....."
>Six steps, no jargon. The note beside each one is the precise technical version, for anyone who wants to check the work.
I think this is called "concept leak" or "prompt leak" (please someone correct me, I can't recall at the moment.) It is one of my least-favorite failure modes. I notice it a lot when drafting landing pages, and it eats up a lot of time to edit away.