HNHacker News
TopNewBestAskShowJobs

einrealist

1,402 karma · joined December 16, 2013

Former CTO at a startup. Now a senior consultant at Evil Corp, again.
submissionscomments
einrealist··on AI Robots – When will they be in our homes
Did I miss ‘they are loud and noisy’ somewhere?
einrealist··on P(doom)
Can't imagine that RSI will result in something useful in practice. There are serious obstacles like alignment drift and model collapse. All while we still cannot solve the accuracy problem with current frontier models.
einrealist··on Measuring the sloppiness of code
And there is another problem: LLMs generating too much code, code that is doing more than was asked. And that cannot be fixed by tests. Usually, we create tests for wanted behavior and expected exceptions. But we don't create tests for undesired behavior.
einrealist··on Ubisoft's FOR HONOR will block SteamOS / Linux players starting September 10
Client-side anti-cheat is a band-aid that will always punish legit players and merely hinder cheaters.

'The client cannot be trusted' applies to simple HTTP exchanges as well as to game state.

Of course, if the server has to track and validate everything, it would cost more to operate game servers....

einrealist··on Malicious Rust crate Arrayref runs a build-time payload
A dedicated status code + Location header allows to decouple the way an advisory is presented completely. At that location, there can be any representation, HTML or structured data like JSON-LD. The RFC extension would be small and clean.
einrealist··on My Business Is Dying
There is something I'd call Subscription Fatigue. People and businesses are asked to pay subscription fees left and right.

If I convert my bank statements only each quarter in a spike or once or twice a year, why should I pay for each month or the entire year? Of course I don't know the real usage patterns of this application, but that's one scenario I'd think of.

einrealist··on Malicious Rust crate Arrayref runs a build-time payload
But I want to specifically signal a security concern to the requesting party. For example, a artifact proxy (like Artifactory or CodeArtifact) can use this response to warn developers.
einrealist··on Malicious Rust crate Arrayref runs a build-time payload
If anyone wants to create a revision to RFC 9110 :)

  HTTP/1.1 309 Security Advisory
  Location: https://acme/aaargh-another-advisory
einrealist··on Norway should buy OpenAI
Or 2x less. 800B valued is not 800B invested. And it wouldn't be the first time investors sell at a loss if sentiment declines.
einrealist··on The Amazon tax
This is so true!

I rarely order anything from Amazon these days, maybe once or twice a year. Everything feels like a rip-off or fake. I don't trust the search function at all. Nor do I trust the customer ratings. The only exception is if I buy from a brand I already know that has its own brand-shop, owned directly by the manufacturer.

einrealist··on UEFA and its national associations will not participate in FIFA competitions
Yes, but isn't football already overly commercialised? Why is advertising for sports betting or alcohol any better than the FIFA issue? UEFA should take a look at itself, too.
einrealist··on DSpark: Speculative decoding accelerates LLM inference [pdf]
Yet another band aid.
einrealist··on There is a shadow hanging over this Fable thing
Nice summary. Reading this reminds me about the strong encryption discussion.

> We optimize what we can measure, not what we actually want to achieve. We hope and pray that these are the same thing, but they often aren’t.

He points out the core problem with LLMs. I believe it is impossible (or extremely expensive) to ensure that the models are aligned safely for everyone and any intention. And 'safe' can mean different things for a different audience.

einrealist··on Why is Vivado 2026.1 dropping Linux support for free tier?
Maintaining support for Windows is free?
einrealist··on SpaceX S-1
"We do not anticipate declaring or paying any cash dividends to holders of our common stock in the foreseeable future."

Sounds like 'never' to me.

einrealist··on OpenAI Is Preparing to File for an IPO Soon
Must be retail investors believing: big number == good.
einrealist··on Google changes its search box
So good SEO will require prompt injection now?
einrealist··on Who will buy your services if you fire us all?
I strongly doubt it will even provide us with a roof over our heads. In an unconstrained market, the pressure to extract as much as possible from the UBI will be enormous. The amount of UBI will probably always lag years behind the actual amount required to create a liveable situation, and increasing its amount will be a constant political struggle.

UBI in an unconstrained market is nothing else than enslavement.

Fair and progressive taxation and proper social systems are far more efficient. UBI is just an excuse to get rid of social systems and leave everyone individually stranded with problems no one can solve alone.

einrealist··on Every AI Subscription Is a Ticking Time Bomb for Enterprise
Those price increases will increase the pressure to use cheaper / free models (commoditization), thus cutting into the revenue projections of the frontier model vendors. Its going to be exciting to see what happens to these huge investments and valuations.
einrealist··on Teaching Claude Why
Isn't alignment a dilemma?

Because what is aligned, how and for whom? And who decides how that alignment should look like? There are probably many domains in which required alignment is in conflict with each other (e.g. using LLMs for warfare vs. ethically based domains). I can't imagine how this can be viable on the required scale (like one model per domain) for the already huge investments.

einrealist··on A recent experience with ChatGPT 5.5 Pro
Compute in science was already subsidized by public funding or by donations. Most supercomputers are financed this way. And that's a good thing. If you have a good science problem that can be computed, apply for compute time. There is nothing wrong to apply that to LLMs as well, like I wrote in my initial post. The human is still required to identity problems that are worth to be computed, to create prompts that the LLM can act on, and to verify results. But, OpenAI providing compute for basically free is still tied to a different incentive: to fuel the hype and to capture the market, while distorting/obfuscating the real costs. That's also the reason for why we cannot claim that 'economics on LLMs is just unbeatable'. It depends on the problem, the reason for a prompt.
einrealist··on A recent experience with ChatGPT 5.5 Pro
Did I praise our animal agriculture anywhere?
einrealist··on A recent experience with ChatGPT 5.5 Pro
"After 16 minutes and 41 seconds, it came back" ... "further 47 minutes and 39 seconds" ... "After 13 minutes and 33 seconds" ... "After 9 minutes and 12 seconds" ... "After 31 minutes and 40 seconds" ... plus other computations

Anyone spotting the issue here? What did that really cost?

I am not against compute being used for scientific or other important problems. We did that before LLMs. However, the major LLM gatekeepers want to make all industries and companies dependent on their models. And, at some point, they need to charge them the actual, unsubsidized costs for the compute. In the meantime, companies restructure in the hopes that the compute costs remain cheap.

einrealist··on Microsoft and OpenAI end their exclusive and revenue-sharing deal
I wonder how this figure was settled. Is it based on consumer pricing? Can't Microsoft and OpenAI just make a number up, aside from a minimum to cover operating costs? When is the number just a marketing ploy to make it seem huge, important and inevitable (and too big to fail)?
einrealist··on An AI agent deleted our production database. The agent's confession is below
Also funny how people (including LLM vendors, like Cursor) think that rules in a system prompt (or custom rules) are real safety measures.
einrealist··on An update on recent Claude Code quality reports
Is 'refactoring Markdown files' already a thing?
einrealist··on Claude Design
True. I didn't expect it to provide novel designs. Maybe Anthropic should find a better replacement for 'Design'.

In my example, I expected it to create UI elements for a business application / expert system. And it did fine. In fact, I believe its perfect for creating average and functional designs. Its a better way to test variations of UIs for expert systems. But I want to know what the actual costs are.

einrealist··on Claude Design
Good for crunching out some prototypes, ideas and getting inspirations I guess. Two prompts - the initial one and one refinement - took about ten minutes and used up 90% of the token budget. I wonder what the real costs are. After the IPO, they will no longer be able to subsidize token costs. The question will then be whether it's still cheap enough just for prototypes, ideas and inspiration.
einrealist··on Claude Opus 4.7
They are trying to optimize the circus trick that 'reasoning' is. The economics still do not favor a viable business at these valuations or levels of cost subsidization. The amount of compute required to make 'reasoning' work or to have these incremental improvements is increasingly obfuscated in light of the IPO.
einrealist··on Sam Altman may control our future – can he be trusted?
I can slow down the compute by a factor of a thousand. It would not change the result. But it changes the economics. We only call it intelligent, because we can do the backpropagation, the inference (and training) fast enough and with enough memory for it to appear this way.
Page 1 of 19Next →