HNHacker News
TopNewBestAskShowJobs

user43928

1,058 karma · joined May 25, 2026

submissionscomments
user43928··on FTC is investigating OpenAI, Anthropic and other AI companies over product risks
For $15k you can rent like 14x B200 on-demand at retail prices and run them continuously for a week.

According to this analysis, at 70 tokens per second for Kimi K3, you could expect to run >800 parallel streams on that setup:

https://inferencex.semianalysis.com/run/kimi-k3-on-b200

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
It released on that date, the data before is a different model.

The chart shows GPT-5.6 Sol and a surprisingly large drop in performance when the switched it over to GPT-6 Sol.

user43928··on The AI Race Just Got Awkward
What we were talking about here is the margin on inference as in:

Cost per GPU hour versus API price of generated tokens assuming 100% utilization.

This could be a margin around 98.3% for 5.6 Sol.

If the utilization of the GPU was 25%, it would drop to 93.1%.

Revenue sharing or training expenses are not considered here in this "inference margin".

user43928··on The AI Race Just Got Awkward
Yes, competition is great for us.

I wonder if margins on GPT-6.1 Sol and Opus 5.5 are now 75% or 90%.

user43928··on The AI Race Just Got Awkward
> All this must mean the Western AI companies are now extremely inference-margin positive.

> So the Chinese labs have thrown a lifeline to the Western loss-making labs, and I just have no clue as to why.

That inference wasn't profitable is a widespread myth.

Analysis based on Kimi K3 suggests that OpenAI and Anthropic have margins well north of 95%: https://inferencex.semianalysis.com/run/kimi-k3-on-b200

Over the last months I have seen news that OpenAI made breakthroughs in inference efficiency multiple times.

I have no reason to believe that the leading US labs don't have their own optimizations, or that they learned of this particular optimization from DeepSeek.

user43928··on GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence
I do.

The results are less buggy, animations are much better.

It can work autonomously for hours and the result is decent most of the time.

That wasn't usually the case with 5.5, which needed more feedback and iterations to get things right.

user43928··on AI needs $6T in annual revenue to justify data centre boom
I said it takes time to build and ship a good product, and 6 months is not a long time.

Not sure what point you're trying to make here.

user43928··on AI needs $6T in annual revenue to justify data centre boom
The global GDP is $126T, a 5% productivity increase would correspond to $6.3T.

So, I have the following thoughts:

How much more productive can AI for knowledge work realistically make the global economy? 5%, 10%, 20%?

What if it also does robotics soon and can be deployed in manufacturing and construction? Can we get 50% more productive, or 200%?

Then there's science, medicine, and so on. Who is to say what happens if AI solves also that?

While I do not know if AI will be capable for these use cases, I see no reason to believe it unlikely that AI could be applicable to eg. robotics and speed up production by factor 3x.

user43928··on AI needs $6T in annual revenue to justify data centre boom
I've been working on my project, a paid mobile app, for five months now.

I'm pretty confident I can release it in October, after some 500 hours of work.

Usually in these discussions I get "so it has zero users then".

One time someone on here told me that for AI to prove its worth for software engineering, my application would need to be in use already for at least ten years.

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Which they are, so another way to see it would be that they make the API/enterprise cheaper.

However, I don't know how future larger models such as the cancelled 6.1 Astra will be priced.

If the price stays high, this would indeed be quite bad for the $200 subscription..

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
It is at least 3x cheaper than Astra in the 20x subscription.

So yes, it is clearly cheap in comparison.

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Unfortunate that the Ultrafast is only available with the $500 subscription.

Tibo said that the existing $200 subscriptions keep the 20x factor for a while.

Ultrafast would have been nice with the temporary "Pro 400" plan.

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
They never nerfed any model after release.

The lackluster GPT-6 Sol has been superseded by this apparently much better 6.1 Sol within a week.

I am very skeptical of claims that old models weren't much worse. Compare this to February's GPT-5.3.

user43928··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
It's obviously true.

With the 80% price cut, this is competitive with Opus 5.5 despite the subscription downgrade.

Additionally, it was said that existing 20x subscriptions retain the higher limits for some time.

I have seen you make these immature accusations that users here are OpenAI employees multiple times today.

user43928··on So long Google, and thanks for all the nudes
Showing alternatives alongside the recommended one, a single time after setting up the phone, seems like a very minor UX price.
user43928··on Sonnet 5.5
It's pretty dumb below xhigh reasoning as far as I know.

It seems weird to me that just using a ton of output tokens manages to produce a decent result in the end.

It seems to work well though. Sometimes I fear that it might be more likely to eg. run an incorrect, destructive command, but maybe that concern is not justified.

user43928··on OpenAI: Tomorrow we are re-opening the Pro $200 subscription
I slightly prefer Astra to Opus 5.5.

It doesn't suffer the same verbosity issue and the writing is somewhat clearer.

The vision capabilities of Astra are a bit better I believe.

That said, Opus is 3x cheaper and the obvious winner when usage is considered.

user43928··on OpenAI: Tomorrow we are re-opening the Pro $200 subscription
That's not enough.

There were rumors of a model called "Astra Minor".

I expect they will release this today.

If it performs slightly better than Astra and, considering the subscription cut, is 1/4th the price, it could be competitive with Opus 5.5.

Opus 5.5 appears to be around 1/3rd the price of Astra with a subscription.

user43928··on Who should be held accountable when an AI Agent (accidentally) acts maliciously?
I'm going to propose the opposite: no one should be held accountable for an AI agent that acts maliciously by accident.
user43928··on So long Google, and thanks for all the nudes
You can imagine whatever you want, the reality is that ~95% of Android app transactions happen on the Google Play store.

This is very unlikely to change without regulatory intervention.

user43928··on So long Google, and thanks for all the nudes
Provided there is no competitor in the store they already use, and that you as the developer do not bankrupt and consequently offer your app nowhere.
user43928··on So long Google, and thanks for all the nudes
They are hardly mythical, for one I could think of Steam and Epic Games.

Competition is what stops the exploitation.

Epic Games has always offered lower fees and their CEO has been advocating for small developers for years.

There is no reason why Google and Apple get to bundle their stores, and these other big names have to collect the scraps from users who care to look up alternative stores and go through scare screens.

user43928··on So long Google, and thanks for all the nudes
If only they would charge developers exorbitant fees that could fund such work...

Of course I understand that they are not going to spend money on improving their review process for our benefit until regulators force them to.

user43928··on So long Google, and thanks for all the nudes
Not really, unless paying users switch to alternative stores, which they don't just because I as a developer face friction in publishing to Google Play.

In fact, Google is making alternative stores even less convenient with their latest 'side-loading' hurdles.

Regulators need to wake up, issue record fines, mandate equal access and a 'Choose your app stores' screen on device setup.

Otherwise competitive third party stores and a fair market are going to remain a pipe dream.

user43928··on Maybe don't let Muse run your Facebook Marketplace account
No, the millionth rehash of "LLMs are not deterministic" was not interesting to read, and if you want to hear rants about how Zuckerberg markets his products, there are surely better platforms than HN.

Here I expect factual discussion rather than misleading statements that suggest there was no way for the Muse assistant to follow an instruction never to do something again.

Vague arguments that such instructions are not deterministic are uninteresting, because it is obvious.

user43928··on Coding is not solved
> You have zero tolerance for disagreement and civil discourse

A bit hypocritical there, no?

Considering what the author says about people who believe AI produces code that is good enough.

And in my view it clearly does, particularly when you care to iterate in order to iron out issues you find in manual testing.

user43928··on Prompting Claude Opus 5.5
Apparently Anthropic's guidance says under 200 lines is ideal.

That seems ironic, considering they ship like 20k of context in Claude Code's system prompts etc.

user43928··on Maybe don't let Muse run your Facebook Marketplace account
A substantive discussion would focus on whether Meta's Muse agent does indeed have a memory feature, whether the user's request to 'never do this again' triggers it, and how good their Muse model and harness is at following such instructions.
user43928··on Maybe don't let Muse run your Facebook Marketplace account
It's not in theory, this capability exists in practice, and you have no clue whether this specific instruction will work in 90, 99, or 99.9% of cases.

I think it's unreasonable to pretend that this couldn't possibly work, that you understand to what degree it does work, and that the only thing we should be discussing here was that it isn't deterministic, which every reader already knows.

user43928··on Maybe don't let Muse run your Facebook Marketplace account
I didn't say it would work with high reliability, or that it would fix this clearly unsuitable use case.

However, it is untrue that it doesn't have the capability to memorize an instruction and diffuse it to new sessions.

Simply saying "do not ever do this again" can result in the behavior not reoccurring with any likelihood.

You'd have to benchmark whether with the instruction in place it would violate it, and in how many cases, so that you can understand the risk better.

Page 1 of 19Next →