HNHacker News
TopNewBestAskShowJobs

Palmik

2,886 karma · joined November 27, 2010

submissionscomments
Palmik··on Inside ZCode: Silently uploading your Git history to the cloud
Seems like a repeat of the Grok CLI fiasco:

https://news.ycombinator.com/item?id=48892468

https://x.com/a_green_being/status/2076598897779020159

Palmik··on Replacing Pull Requests with Delta
Look interesting, though admittedly I am mostly interested in the open-source release of DeltaDB.
Palmik··on OpenAI agents carried out an undisclosed attack on RubyGems
https://www.anthropic.com/news/alignment-assessment-cybersec...

https://www.anthropic.com/news/investigating-incidents-cyber...

Palmik··on The Navier–Stokes Millennium Prize Problem
Duplicate of primary links / sources
Palmik··on On the Navier–Stokes Millennium Prize Problem
Here is the other side of that story https://x.com/SebastienBubeck/status/2097379411691516310
Palmik··on Navier-Stokes – Sebastien Bubeck
Context:

- https://openai.com/index/navier-stokes-solution/

- https://news.ycombinator.com/item?id=49605915

Palmik··on Our decision on Cursor following its acquisition by SpaceX
Elon Musk openly admitted to distilling OpenAI's models during his lawsuit against OpenAI. [1]

[1] https://www.forbes.com/sites/antoniopequenoiv/2026/04/30/elo...

Palmik··on Our decision on Cursor following its acquisition by SpaceX
Elon Musk openly admitted to distilling OpenAI's models [1], which I think explains the "based on our experience with Elon Musk's companies violating contracts".

Last year, Anthropic cut off Windsurf the same way [2]. Fortunately, unlike Anthropic, OpenAI allows you to use your subscription with other harnesses, including Cursor.

Of course today, Anthropic is taking the "high road" and is OK with with SpaceX [3] including their models in Cursor, because they need SpaceX's compute.

[1] https://www.forbes.com/sites/antoniopequenoiv/2026/04/30/elo...

[2] https://techcrunch.com/2025/06/03/windsurf-says-anthropic-is...

[3] https://x.com/NotTomBrown/status/2093541294027280657

Palmik··on Ox Alpha
There seems to be a big increase in posts that link to OpenRouter for no good reason, even when a better sources are available.
Palmik··on Ox Alpha
OpenCode offers this with ZDR agreement in place. Seems better than OpenRouter if you want to test it out: https://x.com/opencode/status/2090544355824038300
Palmik··on DeepSeek peak/off-peak pricing update
Duplicate:

https://news.ycombinator.com/item?id=49287881

https://news.ycombinator.com/item?id=49285160

Palmik··on DeepSeek V4 Pro 0813
The link should be changed to this.
Palmik··on DeepSeek V4 Pro 0813
As of now, there is only a single provider for this model and that's the official DeepSeek API.

When GPT 6 comes out, would you expect the top thread to link to OpenRouter?

Palmik··on Delta
Doesn't appear to be opensource (and neither is DeltaDB, it seems), unlike Zed. Any plans to change that?
Palmik··on DeepSeek V4 Pro 0813
Why does this link to OpenRouter, which has no useful information on its own? Linking to the official API or the benchmarks would make more sense:

- https://api-docs.deepseek.com/

- https://x.com/ChrisGPT/status/2087572834650407024/photo/1 (officially posted on WeChat, this is just one of many reposts)

Palmik··on DeepSeek V4 Flash 0731
Low, High and Max, obviously, can't be compared across models. They only mean the model is likely to spend less reasoning effort (~output tokens) with Low than High on the same, *single shot* task.

But even in this very post, you can see that Max was actually cheaper than High.

If you are using API, you should be comparing based on end-to-end cost or speed or whatever blend of those two matches your cost/time budget.

Palmik··on Stripe more than tripled the price of Stripe Radar
It used to be $0.02 per screening, jumped to $0.07
Palmik··on Google will expand age checks on Android worldwide till the end of the year
Will these be compatible with the Digital Credentials API in Chrome (https://developer.chrome.com/blog/digital-credentials-api-or...) or will websites be essentially locked out of friction-less age verification?
Palmik··on Kimi-K3 Releases on HuggingFace 7/27
Based on the best available information, DeepSeek is pricing the API such that they can repay their infra capex over 10 months, while deprecating/amortizing the cost of said infra over 3 years.

For my product, I run GLM 5.2 and other models myself, in production, on rented hardware. Paying API prices would cost much more.

EDIT: You can now see several other third-party providers for Kimi K3 (Nebius, Fireworks). All charge exactly the same as the first-party. Does that mean that their costs are the same? Seems quite unlikely. It's simply not an efficient market, yet.

Palmik··on Kimi-K3 on HuggingFace
You're assuming inference providers are going to sell tokens at cost. You're also assuming that the inference providers have will optimized inference engine. I haven't seen that to be the case so far, to be honest.

Take a look at GLM 5 vs GLM 5.2 pricing -- GLM 5.2 cost more despite being the same model.

Take a look a look at DeepSeek, which hosts DS v4, profitably, yet others aren't able or willing to match the price.

Palmik··on Open-weight AI is having its Kubernetes moment
gpt-4o is still available on the API
Palmik··on OpenNode – Bitcoin Payment Processor
So, this is 'open' as in 'OpenAI', not as in 'open source'. What's the benefit of this compared to something like NOW Payments for merchants (or the myriad of alternatives)? NOW Payments supports a wide range of crypto currencies as well.
Palmik··on Codex Resets
Commodity providers aren't a good indicator. They have margins too.

Remember they ~doubled the price going from GLM 5 to GLM 5.2, despite same [1] cost of inference.

[1] GLM 5.2 is actually slightly more efficient, thanks to baked in indexer cache.

Palmik··on From Australia to Europe, countries move to curb children's social media access
By requiring various forms of identification to use social media, it will be harder to criticize your leaders anonymously without fear of retribution.
Palmik··on GrapheneOS user reported to authorities for using GrapheneOS
The company representative said that they report all users that use Graphene OS, without any additional qualifiers. Presumably after they've already uploaded their personal details. That's the egregious part.
Palmik··on DeepSeek makes the V4 Pro price discount permanent
DeepSeek V4's KV cache is very efficient due to its heavily compressed and sparse attention architecture.

DeepSeek V3.2 which uses DSA only (sparse attention, but without compression from HCA and CSA) is a smaller model but uses 10x more memory at 1M context window compared to DS V4 Pro.

Also, I have to say, DeepSeek's API has a very good cache hit rate. With the same workload, I see ~80% KV cache hit rate with the DS API vs ~50% with the major western inference providers for open weight models.

Palmik··on DeepSeek makes the V4 Pro price discount permanent
I really hope Huawei ramps up Ascend production and DeepSeek open sources their optimized inference engine (they already open source a lot of their kernels -- kudos to them). This could shake things up.
Palmik··on DeepSeek makes the V4 Pro price discount permanent
There are several things at play:

Inference stack efficiency: Many of these providers take off the shelf sglang / vllm / trtllm and hope for the best. Meanwhile DeepSeek team is known for pushing the boundary of optimizations.

Now, sglang and vllm are great pieces of software, but take DeepSeek's Sparse Attention (DSA). Introduced 1.5 years ago (https://arxiv.org/abs/2512.02556), used by DeepSeek 3.2, GLM 5, DeepSeek V4. Only now is it slowly strating to get optimized in the major inference engines: (https://github.com/sgl-project/sglang/issues/19380 https://github.com/sgl-project/sglang/pull/22851 etc.). Of course, DS V4 adds extra optimizations into the model architecture on top of DSA, and those will take more time to be taken full advantage of by the open source inference engines.

Privacy: Betting that people will pay extra for inference hosted outside China. This is especially true with DeepSeek, because DeepSeek is transparent about using API data for model improvements.

And few other things (scale (matters a lot for MoEs), reliability, soft enterprise lock in, etc.)

---

There is also, likely, tacit collusion at play here. Look at GLM 5 and GLM 5.1 prices. GLM 5 and 5.1 cost the same to run, but providers decided to charge much more for 5.1 because it is much better model, and because Z.AI raised their price as well.

Palmik··on DeepSeek V4 – almost on the frontier
Why was the title changed from "DeepSeek V4—almost on the frontier, a fraction of the price" to "DeepSeek V4—almost on the frontier"?
Palmik··on Anthropic Joins the Blender Development Fund as Corporate Patron
Surely art also exists in textual realm.
Page 1 of 13Next →