HNHacker News
TopNewBestAskShowJobs

troupo

5,745 karma · joined May 24, 2023

Opinions on things I know nothing about

https://dmitriid.com

You may know me from these:

- Everything around LLMs is still magical and wishful thinking(2025) https://dmitriid.com/everything-around-llms-is-still-magical-and-wishful-thinking

- Prompting LLMs is not engineering (2025) https://dmitriid.com/prompting-llms-is-not-engineering

submissionscomments
troupo··on Most data centers refusing to say how much water, electricity they use
Sadly, the EU is also lax when it comes to enforcing regulations. See GDPR.
troupo··on Livenerf: Has Opus 5.5 been nerfed yet?
> And even though it's super straight forward to collect the evidence, there seems to be zero substantiating the vibe bro conspiracy theories.

Until shit like this: https://www.anthropic.com/engineering/april-23-postmortem

Where people pointed out issues early and en masse, and Anthropic denied it was happening, gaslighted anyone claiming this was an issue, then begrudgingly admitted it was an issue, and then spent another two weeks "fixing it".

Or shit like this: https://www.anthropic.com/engineering/a-postmortem-of-three-...

Anthropic is in a perpetual state of "oops, these 'bugs' degraded our model quality" and only admit the issues when it's immediately obvious and visibly affects a large number of customers.

Otherwise all open benchmarks can be (and are) gamed. And it's quite hard to judge the output of a non-determenistic black box that Anthropic (or OpenAI) constantly tweak.

troupo··on Coding is not solved
> I've written 100s of thousands of lines of difficult code.

How do you know it's difficult if you say you don't understand it?

> Anyway all of this reads like someone who is not actually using LLMs to build software or hasn't tried them in a while.

I run all the "latest and greatest" models the moment they become available to me. The amount of insanely bad code they produce remains largely the same, and largely in the same areas. And it cannot be caught by tests unless you know that bad code is there and end up with extremely bad tests anyway. I wrote about it here: https://dmitriid.com/adding-to-i-dont-read-ai-code-discourse

Main one is, of course, "to get a record from a database read all records from it, and filter in memory".

troupo··on Improving site performance by shipping more CSS
Read Performance Inequality Gap https://infrequently.org/2025/11/performance-inequality-gap-...
troupo··on Improving site performance by shipping more CSS
Read Performance Inequality Gap https://infrequently.org/2025/11/performance-inequality-gap-...

Download speed isn't the only thing that matters

troupo··on SNL Weekend Update: Anthropic CEO Dario Amodei on A.I.'S Threat to Humanity [video]
To be honest I struggle to understand why SNL endures. It has never been funny with very very few exceptions.
troupo··on What About Rails?
> People should stop building UIs, nobody wants to interact with a UI. Everyone's app should just be an API that you can use with a chatbot.

The last thing people want is interacting with a chatbot for anything that doesn't require a chatbot.

> in terms of functionality and value to users, software ought to be malleable and composable.

No, it shouldn't. The last thing people want is for software to change from under them, or be replaced with the bullshit that is a chatbot.

> plaster a ‘programmable’ interface on top of unstructured data/interface meant for humans

Ah yes. I can imagine everyone "plastering programmable interfaces" on top of daily stuff they want to do. Like buying a ticket to a museum. Ordering Doorsash. Sending a meme to a friend. Renting a car...

troupo··on Zelensky says Russia has widened attacks to hit Ukraine's data centres
Remember when Internet was supposed to withstand nuclear war?
troupo··on Zelensky says Russia has widened attacks to hit Ukraine's data centres
> And destroying civilian infrastructure does not play well with what Russia is trying to sell for the audience.

They sell "we're only targeting military infra" and "we only escalate after unwarranted escalation by Ukraine".

They hit Ukraine's energy grid during literally the coldest days in winter. The audience's reaction? "YES LET THEM FREEZE".

Russia couldn't care less about what they need to sell. They've been hitting civilian infra right and left.

troupo··on Anthropic resumes charging for requests blocked by safeguards
The "<0.1%" they flagged for me:

- as biology: "Use Unicode graphemes" (already used by the project)

- as generally unsafe: English text I typed without switching from Russian layout (so, gibberish, which it previously just converted to English and executed)

troupo··on Early rogue AI agent activity and attempts to hack found on urlquery.net
You are not a multibillion-dollar company with friends in high places.

You are going to jail.

troupo··on Early rogue AI agent activity and attempts to hack found on urlquery.net
> have heard from several lawyers that at least in US, CFAA[1] in unlikely to be sufficient because it requires intent.

1. What about negligence?

2. Every follow up to every story after the news cycle moved on shows both intent and negligence. To the point of "we opened internet access and told it to hack"

troupo··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
> Cuz OpenAI has been secretly downgrading models on many accounts, including mine lately.

Same with Anthropic. On top of that Anthropic rarely or ever admits any issues, and even if they do, you get like 6 hours of reset. Rmemeber March?

troupo··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
Strange then that Anthropic answers to all user issues with complete derision
troupo··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
> Anthropic has the most information and made the choice they made.

Anthropic is the last company I would trust to make any decisions. Look at any discussions surrounding their "Claude is a tiny game engine" idiocy, numerous bugs that a junior can discover, a full "plugin system" in which they neeeed a dozen files in the worst Clean Code manner to read one of two files etc.

troupo··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
> The AGENTS.md support was implemented via our new extensibility system for CC, called Mods,

AKA "we need 100~ish files wrtitten in the most horrible Clean Code style replete with no two files agreeing on the same naming of the same feature... to read one of two files, one of which has been a de-facto industry standard for over two years"

troupo··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
> Sorry folks, this is a rollout artifact

Aka: "an issue even a junior would've spotted if we didn't rely on Claude of 100% of our tasks"

troupo··on Apple has added persistent 'ads' to iOS, and it's driving users crazy
Read the original comment: "I doubt John Ternus will change direction anytime soon, since it will look like he’s reversing the plethora of eyesore ads that Tim Cook added over his tenure.".

I mean. John Ternus and other execs and senior management had nothing to do anything with it. They were all new, picked up randomly out of the box and only obligingly did the bidding of Tim Cook (or Alan Dye, or both).

troupo··on Apple has added persistent 'ads' to iOS, and it's driving users crazy
Sorry. For some reason got him mixed up with Steve Lemay lol.

But still.

Everyone keeps blaming just one guy for enshittifying Apple, be it Alan Dye or Tim Cook. As if literally no other executive or senior manager has any say in anything that happens.

Even here.

Joh Termus was just picked up randomly out of a box. He never knew what was happening at the company, is entirely new to Apple, and never weighed in on any decision besides hardware. That is why he was picked to checks notes lead all of Apple.

troupo··on Apple has added persistent 'ads' to iOS, and it's driving users crazy
> I doubt John Ternus will change direction anytime soon, since it will look like he’s reversing the plethora of eyesore ads that Tim Cook added over his tenure.

Ternus was one of the main people driving these eyesores. He wasn't a nobody.

troupo··on Robin Williams' Daughter to Fans Creating AI Videos: 'Have Some Shame'
Yes. He did it. In person. Putting in, you know, actual work and character studies.
troupo··on English: A vs. An
Most people can ignore APA rules and just wing it. E.g. English comma is more of a "there's a pause" in most communication.

Don't know about Bulgarian, but Russian punctuation is notoriously horrendously complicated. I still remember a dictation [1] which had a sentence that had a comma after nearly every word. (And this last sentence would have three)

> English spelling is better than, let's say, Greek or French.

Can't speak for Greek, but IIRC in French a collection of letters usually conforms to a rule. "eu" or "ou" will rarely represent different sounds.

Whereas one of the famous spellings of "fish" in English is "ghoti": https://en.wikipedia.org/wiki/Ghoti?wprov=sfti1

[1] Common in Soviet and post-Soviet schools: https://en.wikipedia.org/wiki/Dictation_(exercise)?wprov=sft...

troupo··on Meta bans ads for Virginia Woolf play in Spain
Yes, exactly: META blames the EU for its failures. (Or useful gullible fools blame the EU because "poor supranational corporations can never do anything wrong")
troupo··on Meta bans ads for Virginia Woolf play in Spain
No. Supranational corporations will find any and all ways to be maliciously compliant and blame the EU.
troupo··on I am often wrong
I feel like we should call the behaviour of people out.

That is the modus operandi of Anthropic and all of its employees (at least those active on social media).

More here: https://news.ycombinator.com/item?id=49785996

This behaviour is regardless of whether the issue is an edge case or it affects most of their users who are shouting from the rooftops

troupo··on I am often wrong
> To be fair, isn't this the standard response to issues everywhere?

No, not really. Anthropic is notorius for ignoring issues, pretending they don't exist, gaslighting and blaming users, then spending weeks fixing things that end up being obvious even to juniors.

At the same time they proudly tell everyone how they don't look at code anymore and only run a gazillion Claude sessions.

troupo··on I am often wrong
> The goal of this post was to communicate to the team the way I think

For that you use internal communication channels.

Though we've known for a while that no one at Anthropic is capable of doing any proper communication.

troupo··on I am often wrong
The typical Anthropic/bcherny reaponse. "I've never noticed this, can you tell me more".

See literally every isssue, even those widely reported.

troupo··on Measure internet censorship
I don't know if it's possible to improve every kind of voting in the same way.

But we could improve some parts. For example, voting in elections could be improved by introducing Single Transferrable Vote https://en.wikipedia.org/wiki/Single_transferable_vote (CGP Grey had a bunch of videos on this: https://youtube.com/playlist?list=PL3897F608FAD61E88&si=TzEP...)

troupo··on ChatGPT now knows what you do on other websites via ad collector
Ads don't require pervasive and invasive tracking, and surveillance that Stasi would have wet dreams about.
Page 1 of 34Next →