HNHacker News
TopNewBestAskShowJobs

macNchz

7,927 karma · joined July 15, 2013

adrien@incinc.io

https://a.drien.com

https://github.com/drien

https://infosec.exchange/@adrien

submissionscomments
macNchz··on Gemini 4 Argon
Yes, and I think it has improved some, but just this week 6 Astra lost the most key details of a project across a compaction and got confused about what we were actually trying to do. I would have preferred to stop at 85%, interactively develop a next-steps prompt and continue from there when ready, rather than seeing it compact and become 5x dumber from one turn to the next.
macNchz··on Gemini 4 Argon
In the olden times, aka like two years ago, AI chats would just stop working or just start slicing off the oldest parts of the context to fit the model's window.

That said, compaction feels like an idea that should work reasonably well, but across all of the major providers and agent tools I've used has never actually produced compelling results, to where if I see I'm getting close to the token limit I prefer to start putting a bow on the project and readying it for a fresh start. Even when I provide a detailed compaction prompt it usually focuses on the wrong stuff.

macNchz··on We’re forgetting what darkness feels like
I grew up in a rural place where we could see the texture of the milky way just stepping outside at night—it’s been fun taking friends who grew up in cities on wilderness backpacking trips and sitting out to watch the stars. There’s an awe to a true night sky that I think has real psychological impact.

I did also have a cool experience, deep in the backcountry at a high elevation camp some years ago: we went to sleep in a gale, the tent flapping all around, with a cloudy sky. I thought we’d get a big storm, but I woke up hot at 3am, the air totally still. I slid out of the tent and the sky was completely crystal clear. An incredible spread of stars. I watched for a while and noticed a reddish glow on the horizon. We were so far out and it wrapped around in all directions, so I knew it couldn’t possibly have been a town, but I was puzzled. I learned about airglow when we got out of the woods: https://en.wikipedia.org/wiki/Airglow

macNchz··on Everybody’s home. No one’s coming over
I certainly don't have all of the answers here, but I do think the somewhat ineffable nature of this concept contributes to how at a business or societal scale we wind up optimizing for the wrong things–it's hard to describe, multi-factorial, hard to measure, so instead we optimize for things that deliver obvious measurable improvements. That line of thinking is how you wind up bulldozing a city center to run a highway through it: travel times fall 80%! The negative side effects are harder to measure.

The AC thing interests me as a topic–purely my opinion, but I do think that living in a very narrow temperature range changes your relationship with the outside world, and there's all sorts of good stuff to be done outside that's much harder to get motivated to do if the delta with what you're used to is too big.

As an anecdote: I mostly disliked hot weather in the past, but years ago when I got hooked on running, I found that I disliked the treadmill more than the heat, so I ran outside through a muggy NYC summer. It was miserable to get started but got a bit easier every year, and over time I've noticed, from the digital thermometer on my desk, that the temperature where I have to stop working and turn the AC on has slowly crept upwards. My running club puts on a 5k summer race series–if you'd asked me 10 years ago whether I'd want to run up a hill in the swampy Brooklyn heat every other Wednesday, I would have laughed out loud at the absurdity of the idea–but today it is genuinely something I love, and has been a highlight of my year for several years now.

All that to say, your point about suffering for suffering's sake stands: I cannot sleep without AC above a certain temperature/humidity, and I don't see any reason to force myself to suffer through restless sweaty nights. On the flip-side, though... I have never had more visceral, euphoric appreciation for the comforts of modern life than after a cold and drizzly backpacking trip where we spent the entire time damp and never really, fully warmed up for multiple days. 20-something years later I have the most vivid memory of getting in a hot shower when we got home.

macNchz··on Everybody’s home. No one’s coming over
I've been interested in this for quite a while, and have read a fair bit about it as a phenomenon. I think that pursuing comfort and convenience to the point that it's actually harmful is sort of a big factor in a lot of societal change the last several decades in many developed/Western countries. As a European/American dual national I think it's a particularly strong current in American culture, personally.

An accessible intro is Michael Easter's book The Comfort Crisis, which, while a little pop-sciencey, explores the idea that making our lives more and more comfortable (across various axes) creates a feedback cycle that shrinks our comfort zones, making some of the more rewarding (but potentially challenging/uncomfortable) things in life seem unthinkable.

For me this was a bit of a "once you've seen it you can't unsee it" kind of thing. Letting convenience be your primary metric for success often produces totally unsatisfying results, anywhere from how you build cities to the way you eat a meal.

The arc of the original World of Warcraft (broader social impacts aside) is an interesting microcosm to explore: players continually asked for convenience features that would simplify the most tedious aspects of the game and reduce social friction (organizing groups, traveling etc), but as they made things more convenient over the years, many people found that the game actually began to lose something important. The response culminated in a very successful re-release of the original, inconvenient, game after 15 years.

macNchz··on OpenAI halts training of latest models as reports mount of AI agents going rogue
It doesn't actually matter that the AI chooses which tools to use, because people give the tools to the AI! With no tools, the AI cannot do anything but generate text. "The AI figured out how to abuse the tools I gave it to achieve the goal I gave it" does not mean the AI went rogue, it means you, the human in control, let it happen.
macNchz··on OpenAI halts training of latest models as reports mount of AI agents going rogue
> Go to codex or claude code or any harness right now

The harness is the whole thing here. AI generates text. Everything else is undertaken by harnesses and infrastructure humans provide, have control over, and therefore responsibility for.

Every action AI takes is fundamentally not independent, it requires an explicit choice to let the AI write code, have a physical machine to run it on, to have network access, etc. The concept that these things are "rogue" ignores the role humans play in giving them goals and tools to pursue those goals, and makes it seem like it’s a self-determined force, over which humans cannot exercise control at all.

macNchz··on How to keep enjoying programming in a world of LLMs
Hyperbolic to be sure, but I think it’s fairly common for people to get into programming for fun as a teenager, turn that into a well-paid job, realize that programming jobs are often entirely different from programming as a hobby, then feel trapped because all the experience they have is in programming, and finding an alternative starts to look like competing with recent college grads for entry-level office jobs (ask a 22 year old how that’s looking these days), an expensive degree, or something menial for $15/hour or less.

First world problems, perhaps, but also a microcosm of our broader societal divergence, where a reasonably comfortable middle is increasingly being replaced by growing working and upper-income classes, with very different lifestyles.

macNchz··on OpenAI agent hacked Australian government website, PM says
> I don't for a second believe that an agent starts trying to hack backend system, when the form or API it has been asked to use isn't working.

I have seen coding agents on my own machine (in sandboxed VMs) start doing things while trying to accomplish what I've asked that I felt sort of exceeded my mandate (changing database passwords, poking at the egress proxy that's preventing them from accessing some domains). Not to the point of causing any real issues, but I don't have much trouble envisioning scenarios like this when using stronger instructions around pursuing the goal + a running in a misconfigured sandbox envrionment.

That said, there's a lot of potential upside for American AI labs if they're able to get people scared about AI, they can:

- To your point, claim the regulations slowed them down and paper over near/mid term financial concerns

- Get the government to create stupid regulations that don't actually slow them down at all, but do effectively lock out any future competition (and current global competition)

- Position themselves as the only organizations blessed by the government with the ability to make safe AI, therefore eventually allowing them to claim to be some flavor of "too big to fail" and worthy of a bailout, should the financials not work out.

- Effectively create a distraction that avoids further public conversation/accountability/regulation/liability re the more tangible sorts of problems their products cause right now.

macNchz··on The darker side of being a doctor
On top of these, a few things with the training process also jump out to me from a "can we attract smart and motivated people to this work" point of view: the sheer cost of medical school that creates an imperative for high pay down the line, coupled with a real chance that you can wind up without a residency match but still owe all that money for school, and in particular the grueling nature of the residency system, currently capped at only 80 hours/week since 2003, because people were working >100 and making mistakes.

There are arguments that these are factors that filter out the people who are not sufficiently motivated, but it's hard for me to imagine there aren't a lot of bright young people who might be interested in medicine, but see one of the various paths that exist today to making doctor-level money with only an undergraduate degree and in an environment that doesn't require a working schedule that actively harms your health.

macNchz··on Teleoperated Humans
> If you had an expert looking over your shoulder and telling you what to do, I expect most of you could do most of the work of an electrician … In fact, that's the bulk of how electricians learn their trade: through apprenticeship. People wearing glasses with built-in cameras, connected to today's strongest AIs could already do a lot with a bit of scaffolding.

Apprentices have their work directly checked by the person supervising them, who is also ultimately personally responsible for the results.

I am a programmer but also like to work with my hands—I’ve done some amount of "real" electrical, plumbing, carpentry, light construction, mechanical, and landscaping work. To me, this line of thinking overall falls into a category of "things computer programmers believe about other kinds of jobs," that I see fairly often on HN. It’s key to the whole argument here but kind of papered over with an unvalidated assumption it’d probably already work with today’s AI.

Anyone who has followed a DIY YouTube video more than a few times has encountered situations where having muscle memory and proprioceptive experience with actually manipulating the tools and objects you’re working with would significantly reduce the chance you mess it up. How tight is too tight? Am I going to shear this bolt head off? Is this saw blade getting dull and going to hurt me? Even the best guides and teachers don’t cover the combinatoric ways things can frequently go wrong–it’s accumulated knowledge in the heads and hands of people who don’t blog about it so it can be slurped into training data.

I think there’s an enormous chasm between "can AI provide reasonably correct guidance for an inexperienced person along a happy path to accomplish basic physical tasks" and "I no longer need to hire an electrician because AI guides me through every step."

macNchz··on OpenAI buys smartphone camera maker Glass Imaging for $300M
The idea of managing a consumer app API for which every client is auto-vibecoded to the user's preference sounds...fun.

User: Hello Uber support? Yes, the app shows my driver is in the ocean off the west coast of Africa? And his ETA is -2147483648 minutes?

Uber: Your app, your problem.

User: Hello phonebrain, can you fix my Uber app?

Phonebrain: You're absolutely right, I must have misunderstood the API docs. Squiggling... ... Shuffling ...

Phonebrain: Error, you've hit your usage limit. Please wait 12 hours, or switch to pay-as-you-go app development mode. I cannot estimate how much it will cost to fix your Uber app.

macNchz··on XCancel service is suspended until further notice
For years Twitter stood out among major platforms for having an actually-usable mobile web interface that seemed to be a first-class citizen / wasn’t intentionally degraded to force you into the app, and didn’t even nag you about it. Unsurprising they’ve since ruined it.
macNchz··on iPhone Duo
> Honestly if anything it’s still a bit too big.

Stumbled on my old iPhone 5S in the back of a drawer recently and could not believe how good it felt in my hand, even having only used 12 and 13 Mini models since 2021. What an excellent form factor and construction. Would happily pay a big premium for a modern version of it. Dreaming of a sub-mini phone trend. Sell me a $1500 Zoolander phone—not even joking!

macNchz··on Record-High 89% in U.S. Say Government Corruption Widespread
There has been a lot more than "some changes in alignment over the years"—the parties of 160+ years ago are essentially unrecognizable today. Given the nature of how the modern Republican party established its base over the course of the 20th century, this idea of taking pride in a legacy of ending slavery is basically comical, a fiction that falls apart with even a basic knowledge of the history of American political ideology and party alignment.
macNchz··on Discovery of a new OpenAI agent message board
My impression is that some of these things are coming out of efforts to make the models more persistent in completing their goals.

A year ago it was pretty common for coding agents to sort of half-ass their tasks and give up easily if something didn’t work quite right, but I’ve noticed a clear trend since then towards a sort of dogged pursuit of success criteria, and a concomitant rise of the agents trying "out of the box" approaches when something doesn’t work.

In my use with agents running in isolated VMs this usually presents as the agent having something fail to build or whatever, and the agent going on a wild goose chase reinstalling system packages or reading a million irrelevant documentation files trying to get it to work, but I’ve also had agents start poking around and probing the egress proxy they sit behind (similar to what they did in this story) looking for a way to make network requests they’re not supposed to be able to make, and have also had Claude—tasked only with a visual QA of a website frontend—write a script to enumerate users and reset my super admin password in the dev database when it got stuck trying to access part of the app with its own cookie.

macNchz··on Claude Fable 5.1 and Claude Mythos 5.1
It's like a dialect of corporatese. The kind of droning non-speak you can sit in a 90 minute meeting listening intently to and come away wondering whether anyone actually said anything.
macNchz··on Taylor Farms: How One Company's Reach Became a National Risk
I think a lot of people sort of instinctively avoid thinking about or engaging with the dangers posed by cars, at least in part because it creates a sense of dissonance that there’s a non-zero risk of life changing harm to you or your family from such a “normal” activity. For many people today, there is no alternative but to accept the risk if you want to do anything outside your home, other than dramatically upending your life to move somewhere else. Thinking too much about risks that feel outside your control just creates a state of anxiety.
macNchz··on Pixel 11 Pro Fold feels like the end of an era
I broke my 12 Mini, replaced it with a 17, then went so far as to return the 17 and buy a used 13 Mini.

It was totally annoying to me to not be able to operate the phone with one hand without feeling like I was about to drop it. I kept the 17 for most of the return window thinking I’d get used to it, but I just kept finding more situations where it bothered me. Battery life on the mini is not amazing, but a slim magsafe powerbank makes it largely a non-issue.

macNchz··on OpenRouter is joining Stripe
Having worked on web apps that processed online payments before and after Stripe I totally agree, this was an area of real pain that became suddenly extremely simple because of Stripe. The alternatives were terrible–100 page Word docs of SOAP API docs for Authorize.net, massive PCI compliance requirement specifications, horrible legacy merchant services businesses.

On the other hand, I operate an app that talks to (and logs prompts/meters costs) to many different LLM API providers, and I do not consider it painful at all. I have an AI agent to deal with any integration quirks, if needed. Mostly they provide OpenAI-compatible APIs anyhow. It's basically a no-brainer to go direct with the providers and save 5%, the great majority of the cost of an AI-powered app is no longer dev time implementing the integration, it's the tokens themselves.

macNchz··on GitHub down again? no PR access
I don't have data for it, but I have been a Github user since 2012 and have found it to be down often for pretty much that entire time. I always figured it's cultural to a degree–new features always seem pretty buggy/underbaked, and often continue to long term but remain unloved and incomplete.
macNchz··on Accelerating GPT-5.6 Sol Ultrafast
This is foundationally similar to a lesson I've found from years of pre-LLM software development: builds that turn around in 500ms instead of 5 minutes fundamentally change the way you can work as a software engineer. I think a lot of the same applies to working with LLMs. I'm not sure, though, what the path from where we are today to some future state of high speed token abundance actually looks like...I think there's plenty of chance that we see the bubble pop in the near term over token costs and complexities of today's infrastructure, then some totally different landscape of LLM use in 5-10 years that looks quite unlike what we have today, similar to how waiting 30 minutes for an MP3 of a single song to download on a 28k modem in 1999 seems quaint today.
macNchz··on Docker Sandboxes – Disposable, isolated sandboxes for AI agents
It does, for whatever reason the marketing page doesn’t advertise it but the docs have Linux instructions: https://docs.docker.com/ai/sandboxes/

I’ve been using this pretty extensively for a few months on Mac and Linux and have been super happy with it.

macNchz··on Turn And Face The Strange
I operate an app that orchestrates code running on a bunch of Sprites. A few months ago it was an exceptionally buggy service, frequently just fully broken or losing data (had to get support involved to recover one that got into a fully broken state), but from what I've seen at least they've become much more reliable in the past two months or so.

I evaluated Fly.io itself several years ago and found it unacceptably buggy as well for any real use case at the time, but I kept an eye on it and today I operate a few production workloads on it that are very stable and reliable—I figured the same would be true for Sprites and it seems to be taking shape that way.

macNchz··on What AI did to stackoverflow in a graph
Yes this is a blame-your-politicians situation more than anything else.
macNchz··on Claude Code: Anatomy of a Misfeature
I am a longtime and heavy Claude Code user, but Anthropic's product management overall (including for their Desktop/web products) has been really baffling me. I agree that these things often have the air of having been vibe coded without enough human input, and change so quickly (often without a particularly compelling reason for the change) as to be aggravating.

The most recent one that's had me annoyed is the "Fullscreen" TUI feature, which is super unintuitive, implementing its own text highlighting and copy-on-select mechanics, overriding your terminal's native right click. Easy to disable but terrible defaults, IMO. It's not even really clear to me what problem it was actually supposed to solve.

macNchz··on Tiny data centre used to heat public swimming pool
An outdoor heated pool that’s open all winter in a cold climate would be a destination worth a drive. A rather decadent use of energy otherwise, it’d be a good use for waste heat. There’s prior art in the Blue Lagoon in Iceland, a destination spa that uses water from a geothermal power plant.
macNchz··on Potential session/cache leakage between workspace instances or consumer accounts
The person posting this claims to have reproduced in a separate context down the thread:

> Same thing just happened on a Claude Mobile session in same Enterprise account. Common theme in both is Sonnet 5, first response after more than 5 minutes (cache miss).

macNchz··on Potential session/cache leakage between workspace instances or consumer accounts
Caching is not supposed to work like that, but that doesn’t preclude the cache key computation function from having bugs.
macNchz··on Reality has a surprising amount of detail (2017)
One of my first real DIY projects during a summer in college nearly 20 years ago was replacing the rotted out basement bulkhead doors on the ~120 year old house I grew up in. I took measurements of the old ones, bought some nice tongue-and-groove cedar and high-quality hardware, and built the new doors in the garage. When they were fully assembled, I carried them over to install on the old stone frame. I took off the old ones, put mine in their place...and they didn't fit properly at all.

Momentarily baffled, I realized that, despite appearances, the old frame was actually not square, in fact it was a parallelogram. I'd measured the height and width and assumed it was square. The previous (experienced) carpenter who'd built the doors I was replacing had clearly noticed this, and simply allowed for the misalignment in his design. He built perfectly square-appearing doors that mounted to the not-square frame. I had to go back and rework mine considerably for them to fit without looking ridiculous. They're still there and holding up well, but I also still think of this lesson on a regular basis in my day to day life now.

Page 1 of 34Next →