And no I came up with the metaphor all on my own, send me the chat of you getting the LLM to come up with it. Why not argue based on merit instead of strawman and ad hominem attacks?
And no I came up with the metaphor all on my own, send me the chat of you getting the LLM to come up with it. Why not argue based on merit instead of strawman and ad hominem attacks?
Harnesses (and the concept of agents before them) presuppose competence in LLMs which simply doesn’t exist.
0. https://www.businessinsider.com/sam-altman-ai-utility-electr...
His idea of metering is predicated on the thing he’s selling being AGI, it is not, and all his predictions have turned to dust.
Also that isn’t how metaphors work - they illuminate by comparison, if the comparison is not close they are not useful.
I don’t believe in AGI, but that doesn’t mean I don’t find AI useful. I just understand that the correct harness can take them to the next level.
Then how do you explain the wild success at using them for development?
That doesn’t make them intelligent agents which think independently.
I have a system that entirely reverse engineers old arcade games. Creates semantic symbol mappings that were considered impossible just a couple years ago.
Granted, it took me a couple weeks to build the system.
From impossible to a couple weeks in just a couple years.
Would you like to see it or continue to pretend these things don't exist? Your call.
(It's finding the coolest stuff - the anti-tampering hacks they put into the old machines is fascinating.)
Mostly I felt they reproduced games from the training data, some with more examples available worked better than others.
I use them most days for work, and for that reason don’t trust them that much and certainly don’t worry or fantasise about AGI.
Another of my projects is to incorporate Pixar's ideas from RenderMan into a 3d printer slicer. Displacement shaders, in a 3d printer, have never been done before. Would you like to see that? I had the idea 10 years ago but it was too tedious to implement. I have a working system now in just a couple weeks AGAIN.
Anyone claiming agents aren't profoundly useful is WRONG. If you disagree, please let's discuss it.
In my experience the use of agents and harnesses and other scaffolding around them doesn’t improve the performance of LLMs much, which is adequate for some tasks under supervision but nothing like general intelligence (or electricity for that matter). But good that it works for you.
Which I've always found so strange. Language was always considered the pinnacle of the human mind - right up until we created LLMs. Now it's the world model - the things that animals always had and the thing we used to look down on.
From my perspective, language is still at the top. I'm not a fickle lover of the gaps.
I have to say - that seems straight-up crazy to me. Would you mind digging into that discussion?
As just one example - in a harness I can ask an agent to confirm everything it says via a second sub agent - which dramatically improves the output. It almost entirely solves the problem of hallucinations. You don't see that as an improvement?