But I suppose it doesn’t harm its reasoning!
But I suppose it doesn’t harm its reasoning!
Checking code is (relatively) easy, you can use static type checks, linters, and execute it to see if it's correct.
Fact checking is harder. A RAG can only check what's in the database, so you have to know what to know beforehand.
A global database of facts would make easier for AI to stay factual
Also ironically it would also make it easier to align AI to do things like consistently censor or distort some political facts
In contrast to human intelligence, there is an underlying mechanism that propels intelligent behaviour. A person is no less intelligent just because they lose sight, sound or inner voice.
I see this a lot in Claude Code. I assume it has to do with the training structure.
Example is “fallbacks”. Claude constantly sprinkles “fallbacks” in the code, even when I ask it not to. That is, write multiple candidate implementations into the same code with some kind of switch.
This is a problem because you only need one, and it would seem to have inflated the code for no reason. (The madness accelerates with code volume, so you must push back.) anyway, few people would do it this way.
But I thought, what could be the benefit?
If you’re being conditioned to pass evals one-shot with code that will be discarded and never read, it’s a great strategy. If you have more than one way to solve it, you can just put both. The behavior would easily be reinforced, if trained that way.
But in any case, again, a certain nature and certain conditioning.
I think we’ll learn to accept it as AGI but also that no intelligence is fully divorced from context, limits and conditioning.
I mean, if our position is that a hyper advanced statistical model is going to ultimately struggle with the concept of something being unlikely to be true then the statisticians may as well give up in despair. There is no theoretical obstacle here.
If your last experience with frontier LLMs (read: not models that Google and ChatGPT are giving away for free, but models you have to pay for) was over a year ago, you may not realize that.
Though I haven't read GEB so I'm not sure how the strange loop thing ties in with either of those.
I don't remember a single thing! (Which is unusual for me, I usually remember much of what I read.)
I shall have to read it again :)