> when I can have it do something like write entire working kernel module fixes for old MacBooks on a whim
They’re not saying it can’t do that, and that’s not proof it doesn’t hallucinate. In fact, having used 6-8 agents at a time for a year plus while writing AI tooling for an AI startup, I can definitely surely tell you that they’re almost inversely correlated as in models that hallucinate a lot sometimes also put out the best most impressive solutions.
I’m definitely not anti AI and I definitely have found a way to make it work very well and I’m content with the work I get out of it (again maxing out several max 20x subs), but I have had sol definitely hallucinate this week and I’m a bit shocked you’re trying to say otherwise.
Listen I know it’s going to be I’m holding it wrong too, but I’ve been reading white papers and research on LLMs for a long time and was definitely at the cutting edge of context engineering, implementing features in our tooling harness a year before they were in codex or Claude.
maybe I am holding it wrong still but but like at some point if I’m holding it wrong who else will be holding it right? Dozens of people? At some point, the technology has to be approachable enough for everyone to have your point of view automatically.
I'm no expert on the inner workings/harnesses/etc beyond a basic understanding of the architecture. Maybe I've just developed a good sense for effective prompts? I could share some recent sessions.
Fact is, those with the foresight to see how this can go badly are also seemingly the types of people who don't end up in a position to prevent harms by it. Furthermore, it seems inevitable in a sense, because greed for power seems to necessitate development of automated weapons in ASAP in spite of the risks, and the fact it basically renders traditional warfare pointless.
It appears we'll have to learn the hard lessons, same as our forebearers with. I just hope we can avoid having to regress back to sticks and stones on account of fucking ourselves by overdoing our capability to destroy on account of not being willing0able to peacefully coexist.
In other news, New Orleans 911 is using AI to triage localized incident duplication calls from unique emergencies. I'm not saying that it's good or their only practical choice, but it's happening.
The biggest dangers I see are the outsourcing of supervisory control, appeal to authority (when used to summarize content or answer a question), and hallucinated mistakes.
0. (PDF) https://docs-library.unoda.org/General_Assembly_First_Commit...
1. PDRMUAIA https://www.state.gov/bureau-of-arms-control-deterrence-and-...
2. REAIM 2023 Call to Action https://www.government.nl/documents/2023/02/16/reaim-2023-ca...
3. REAIM 2023 Endorsing Countries and Territories https://www.government.nl/documents/2023/02/16/reaim-2023-en...