"Claude, prepare me a presentation on XYZ."
I get to work, go straight to the meeting room, and pull up what it made to present.
It's barely coherent nonsense. Lots of irrelevant details, buzz words, wrong charts or confusing phrasing. Obviously LLM output.
I read it out.
When I'm done, I get a question about one of Claude's incorrectly inferred details.
The shame instantly kills me.
This is a scenario I've seen play out with coworkers. Except that last part, instead of dying or owning up to the mistake of trusting LLM output they waffle. Their shame circuit is broken.
"What does this bit mean?" "I dunno."
It was all awful.
I don't really know what the right response is, though. Walking out would be seen as way too hostile
Plenty of that even before LLMs too though.
This year, a good 40% of workshop slides are fully LLM generated, incomprehensible, and almost not matching what the speaker is talking about at all. Last year, while there were some poor presentations, standards were much higher.
Now that people are generating decks with AI, they’re basically worthless. I don’t need a bunch of bullet points you didn’t take anytime to think about on the screen while you talk about something you didn’t prepare for.
All the Apple Intelligence commercials were shamelessly this too. Felt like such poor branding for apple
If you are just presenting it for the first time after seeing it you are going to just be reading off the slides, which is a waste of everyones time.
I suspect it’s because writers don’t normally have jobs and the only time they are ever in offices it’s to deliver a pitch for a project.
I see a version of that all the time and would never, ever do it.
one of the best parts of vacations is actually planning it! trying to find restaurants, organizing your days, trying to fit activities -- that makes you look forward to your time off.
really, i want to automate the boring parts of my life (did i really pay rent this month?) not the things that make me happy.
For some. Others hate or fear it. For others it just feels like work.
We all know the feeling of wanting to be successful, or to go on a great vacation. Showing these scenarios creates strong emotional response in the audience - we're now the people winning at work, going on a great vacation. This is especially important when you're selling something dull. See: insurance.
Claude: Would you like to go to Bermuda instead?
I make restaurant reservations 7-8 times a year? My wife and I go on dates and you simply aren't getting a table where I live if you don't. Fyi we typically spend $120 or so on those dinners.
Most restaurants have a web UI which is very simple and easy to use.
At least in this city, reserving a table using the apps you already have on your phone is faster than using an LLM. Maybe these products are for people who know about the existence of Claude and ChatGPT but not Google Maps?
(They have iOS and Android apps too, but I've never bothered to install them.)
My github repos all have CONSTITUTION.md files that keep multi-agent dev grounded. Production runs trigger github actions that automatically check logs daily, file bug reports, etc. I launch Orca and literally just type "checking in" and the CONSTITUTION.md file governs the scanning of github issues and prioritization of which issues need to be addressed. Another process prioritizes and bundles them. When I type "checking in" the Orca worktree handles the launching of sub-agents that fix things. Occasionally I weigh in with an opinion or pick a recommendation. If I'm in the mood to pay attention I'll ask it to find another round of bugs and let's keep going. Often I just let it close after the first round. When I tell it "done?" it does a full regression and a production box review. Code auto-deploys to production twice/day. I don't regularly push to production manually.
On Fridays the CONSTITUTION.md kicks off a strategy and roadmap review when I type "checking in". As I have ideas I add them to the roadmap in one of the worktrees. Generally nothing happens until they're reviewed on Fridays together, unless I push one manually.. which happens.
Hermes runs on the production box. I have a few skills and quick commands that lets me check in on production runs and the status of things. It's read-only by design, basically my version of a dashboard. If I have a thought, idea, research link, or question it'll write to a github issue and we'll deal with it in the next check-in.
Orca lets me do all this remotely from an iPhone, and I talk to Hermes via Telegram. I generally don't INITIATE new work using either of these. I report issues/ideas and let the process do the things. But I can, and have.. I just tend not to. The whole thing churns, and so my inclination to "jump in and do a thing" is less than before. My inclination now is to toss things into the machine and let the machine work the schedule. I'm at a point where I could easily automate 80% of this and do my manual things 1x/2x week and I think progress would stay steady. I'll get there at some point, but I like the level of engagement I'm participating in now.
Occasionally I use Claude Code as a watchtower review of things, or do a wholesale code review, or do a review of logs. But deepseek is much better at building this machine -- I built something manually over time and through a ship of theseus process it got ugly. Deepseek reviewed the situation and wrote the CONSTITUTION.md and designed the processes independently. It favors deterministic scripts for process but launches them via LLM and monitors for exceptions; also, no memory system which also disqualifies Claude from being at the center of this. Memory plugins didn't work well because instructions start accumulating all through the chain of instruction files (AGENTS.md, memory, etc) and that gets very undeterministic very quickly. Instructions are written once, reviewed/audited intentionally.. it's best not to let LLMs learn and write their own dynamically (yet).
I've since used it as a template for other projects I've started -- they work the same. Currently building the GTM agent that will handle all SEO, marketing, keywording, etc for my projects -- it'll be a significantly autonomous Hermes agent. Don't really need Claude for any of this and my Deepseek bill is $50-100/mo.
Two years ago, I had no commute, and presentations were tedious ( i was NOT a good google slides user ), I kinda prefer this world for now