HNHacker News
TopNewBestAskShowJobs

loufe

2,378 karma · joined March 24, 2017

submissionscomments
loufe··on Gemini 4 Argon
What? You mean the technique they had turned OFF during all training run where the agents they are responsible for hacked huggingface?
loufe··on Gemini 4 Argon
And a stake in SpaceX
loufe··on It's Time to Investigate the AI Labs
Listening to Dwark's interview of Noam last week was bewildering. The gall of the figureheads for these companies to say things like "chain of thought monitoring cannot be relied on without potentially poisoning the well", while also saying "we need to pace the frontier but please still trust us [and so called" independent " 3rd party firms which at least in one case is actually a significant stakeholder in Anthropic] to self police" and THEN "chain of thought monitoring was turned off but we have 'more, robust monitoring tools in place now'" is absolutely insane as it is coming out live that your model in training is still committing felonies.

Corporate influence in government in the USA is wild.

loufe··on HomelabFest will be in St. Louis in September 2027
I wasn't aware he's publicly said anything like this, but I've always assumed he was quite religious so I'm not particularly surprised.

To be fair, this is a pretty tame and measured expression of disapproval of the modern societal acceptance of homosexuality. I honestly don't know what more I could ask for, in modern societal discourse, for someone I disagree with in terms of how they go about expressing it: the links I read seem to imply he was honest, polite, and clear on his feelings, not inciting hate, violence, or spewing vitriol. Jeff's views are his own, and maybe the crux here is that I don't see politely disagreeing about the righteousness of living homosexually as remotely severe enough to warrant this much attention considering the quality of character of a lot of public figures visible to those of us outside the USA.

I admit seeing a response like this irks me as I feel it needlessly detracts from the mission. That said, it's both A) fair and a good thing to judge and critique any public figure's character and I do think public figures and their views deserve scrutiny, and B) a means to continue to allow powerful people and groups to assassinate the character of and dilute the message of people who run contrary to their goals. I'm genuinely torn on how to feel about how to feel here. I worry we allow ideas to be discarded at the first sign of something we disapprove of in people who share them -which is wrong-, yet several of my best friends are queer and I also could not stand to see them trod upon. Those same queer friends all have views and personality flaws which would be judged just as harshly by others.

Regardless, I don't like you using my excerpt that way. Not everyone's views of "community" mean the classic cancel culture era idiosyncratic definition whereby anyone in any role remotely resembling leadership must share the same specific set of homogeneous views on political and cultural issues to be given any member's "consent" to lead a group. I stand by what I said and still do believe Jeff is an excellent advocate for these ideas. There are obvious limits, to be clear, I would never go to Toastmasters if the local club head spoke about killing jews, but this isn't that and IMO there are so, so many more important battles to go fight out there than this.

loufe··on HomelabFest will be in St. Louis in September 2027
For those who don't watch his YouTube channel or read his blog, Jeff is a fantastic spokesperson for for everything self hosting (enjoying the NTP binge of late!). I really hope this is a success as he's an excellent advocate for a healthy and positive community and framework of understanding this stuff, not to mention his quality as an orator. It would be great to see a cultural push back against surveillance capitalism and centralization of compute and we need good voices and responsible advocates sharing their thoughts to those learning the ropes.
loufe··on Using LLMs to trace alchemical knowledge and decode 17th century letters
I've been using AI to help with my genealogical research, and it's been fantastic. I am loving pushing the family history back and catching mistakes that someone who many others share as a common ancestor have made.
loufe··on One year of sponsored Servo development
The more I actually look at projects I care about the more and more I'm noticing just how extensive their support is. I wish more governments and philanthropists would do like them.
loufe··on Australia says it could follow Canada in forging deeper ties with EU
Identify in what way? What do you mean by "identity with"? The other poster is speaking of their own experience
loufe··on GrapheneOS' rewritten Messages app is released
Don't give up the good fight! Took some time but I got the whole family to adopt it as we had no common shared platform before, and it's been great.
loufe··on GrapheneOS' rewritten Messages app is released
This is completely false. Tap the call, tap "call details" and its there. Took be 2 second to disprove. This is lazy to the point of being seemingly malicious commentary.
loufe··on Ask HN: Resources to get good at soldering?
Exactly, I solder outside or under the range hood which in confident is enough for a rare job I do
loufe··on GrapheneOS says Pixel 11 has MTE support after all
I am almost certainly going to live with whatever drawbacks in terms of camera quality, battery life, etc. Come with their Motorola phone when it's time to upgrade. MTE is such a non-negotiable for modern digital security on phones it's crazy Google would be so okay with this regression.

What's especially stuck in my mind lately is how insecure basically all desktop OS' feel. In at the point of buying a second and third GPU for my desktop to run my email and browser in dedicated VMs because everything feels as watertight as a sieve. Qubes seems more and more appealing in a world where every open source software supply chain is under seige, corporate software underprioritizes security, and most sites will stop at almost nothing to surveil you.

I truly lament this new reality where MY computers I PURCHASED feel to use like I'm reaching blind into a paper bag filled with razor blades.

loufe··on Nancy Grace Roman Space Telescope
I always thought this was a ridiculous part of the movie but it truly stuck with me more anything.
loufe··on AWS Acquires DuckLabs
Not to mention Angular
loufe··on Nitter and XCancel receive cease and desist notices
I'm not sure why Threads makes your cut here, trading one walking negative externality of person in control for another is not really a sell to most folks.
loufe··on Harvest hikes bills by 1500% after purchased by Bending Spoons
"What's Changing"

That warning infobox triggers an AI flag in my brain instantly.

loufe··on The AI Credit Resale Economy
The comments here are baffling. Is nobody seeing the easy opportunity for gathering amazing high-quality training data by inserting yourself as a MITM? If I were a competing lab, criminal, or opportunist I'd lie/cheat/steal/simply pay the difference to get the chance to listen into real life scenarios of usage of modern models in a high-impact business.

Seems a huge part of the story completely absent to me.

loufe··on Accelerating GPT-5.6 Sol Ultrafast
Going back between two different company's AI tools when facing a tricky architecture question often surfaces holes in an approach I'd been building.

Similarly, if I ever get a bit too vibey and don't carefully review code changes myself, the blast radius is generally significantly resolved by a carefully tuned "did you consider x, y, and z" skill after a first draft partnered with a "deploy an adversarial review agent for the worktree".

loufe··on Using an open model feels surprisingly good
This is only a thinly veiled ad. It's fine, I was curious about this exact setup, anyways.

What would be useful is a cost metric. I'm curious how much I'd be willing to spend as a premium to not have those companies piping my conversations directly to the NSA. Maybe only some conversations? Claude and OpenAI are heavily subsidized, by all accounts, so Kimi K3 on a private endpoint might end up costing more or less - that's what I want to know.

loufe··on IRGC claims it destroyed Amazon's Bahrain data center
As an aside: I lack the vocabulary, but man those interactive "stats" cards scream LLM-written to me. No problem with it, necessarily, just something about that style seems to raise a flag in my mind.
loufe··on Codex Resets
Edit:

The $200 plan is explicitly 4x the $100 plan[1] only for "per session". That's so vague. I initially pushed back against your claim, but reading now Anthropic is not at all clear, in fact.

[1] https://support.claude.com/en/articles/11049741-what-is-the-...

loufe··on Transcribe.cpp
author of the blogpost is the maintainer of Handy, so almost guaranteed!
loufe··on Grok Build is open source
I wonder if releasing this may have been on the roadmap, but been prioritized as a bit of whiplash following the "you forfeit the entirety of your working directory as a condition of working with this tool" upset from a few days ago.
loufe··on GPT-5.6
I agree to an extent but it needs to be balanced. Receiving a half-baked, extremely verbose recap of thinking on benign details with Opus 4.8 or GPT 5.5 feels like an extraordinary loss of quality of experience compared with fable 5.

Yes it shares less, but I think the trade-off is you pay less in tokens and hopefully it's truly just not needing to say things because it truly does just better get what you're saying, think to read X markdown file or GH issue which contains the info, etc.

As long as I can still push back and get it to share its thinking on demand and I'm confident the model isn't actually basing things on poor premises, this is okay for me. I am more productive when not inundated with time-wasting check-ins.

That said, I absolutely lament the loss of the ability to access the thinking - I would happily read the "DANGER DANGER DANGER" internal gremlin thoughts fable 5 makes to verify something if they were accessed, and prefer that to a recap presented only for my benefit.

loufe··on Claude Science
First party support would be nice since this is not a high-trust in the AUR period, but fair point, I'll probably use it, thank you!
loufe··on Claude Science
unfortunately no arch based distro support. I'm curious why it's not packaged as a flatpak.
loufe··on Claude Sonnet 5
Not to single you out, parent commenter, but I really hope the quality of discourse on HN will move past these basic comparisons eventually. It seems like every thread on every model release has the exact same comments.

"Wow, X models is Y% better or worse than Claude Z model on T benchmark"

"That's irrelevant, they're just benchmaxing."

"Not useable for daily coding or agentic workloads, the vibes are totally wrong."

"It's almost as good, and costs a lot less, so I will absolutely use it."

"I cannot imagine justifying using these, as the step change means open models lower costs do not make up for the productivity loss"

I'm an unhappy Anthropic customer and really rooting for open models and non-gatekept intelligence, but how do we move on from this now meme-like model release discourse rigamarole. I do not know what that would be. I don't design LLMs nor benchmarks, and I genuinely appreciate that people do their best to provide information, even if non-perfect here. I'm sure most of you who actively read these comment pages on announcements must feel similarly, though, right?

loufe··on Claude Code is steganographically marking requests
Genuine question though, why would I care about this if I'm paying for a subscription and adhering to TOS. I'm very skeptical about their privacy policy, business practices, and so on, but am curious what the negative about this is. Seems like it would work to my favour as a customer pushing back any date of the cutting of subsidies.

That said, these fraudulent proxies are helping Chinese labs keep up, which might be to my advantage long term in eventually having a high quality private AI I fully control on my own hardware. That's not support, but I do recognize the incentive, for whatever that's worth.

loufe··on Previewing GPT‑5.6 Sol: a next-generation model
"Next generation model"

If it was the next generation, why isn't it a major version change..?

loufe··on SpaceX to buy Cursor for $60B
For what it's worth, flickering in CC has been fixed since around the beginning of the year.
Page 1 of 20Next →