HNHacker News
TopNewBestAskShowJobs

wg0

4,455 karma · joined May 21, 2020

submissionscomments
wg0··on Introducing System One Models and Jev
Can I put it as Air Traffic Controller? With similar error rates as humans?

That would be the litmus test.

"Does not hallucinate" is not the same as "is never wrong".

So the ATC test could be the benchmark.

wg0··on CSS-Tricks in Limbo
I am concerned that what will happen to detailed technical write ups that defined much of the "technical internet space"? There used to be "CodeProject" (still is around probably) that used to have great articles on C, Windows Programming, Windows System programming and what not. Similar other sources on Web as well such as some TutPlus network.

Now with LLMs, they won't gain attention and all knowledge would be stuck in LLMs but there would be no knowledge base to train the LLMs on so... would this all collapse in a decade or two?

PS: Not anti AI rant, just a genuine open question.

wg0··on Fable 5.1 Solves the Cyphral Distich, a 370-year-old cipher
Has anyone tried it with other models.?
wg0··on Anthropic CEO Says It's Time to Slow AI Model Advances
Then slow down your bots too please. I deployed a website last night. 10 out of 11 GB of bandwidth went to Claude bots alone with more than 90k requests.

Hypocrites hyping for the IPO.

PS: True story.

wg0··on DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
What I mean is that price has been effectively halved.

During off-peak hours, the unit price is reduced from $0.007 for input cache hits to $0.003, $0.22 for input cache misses to $0.15, and $0.12 for output to $0.6

wg0··on DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
DeepSeek v4 Flash with high is already a really great work horse. Reliable. But this time, not only that it is better but they are reducing the price by 50% so that's great.

I also find the DeepSeek models to be more precise than Claude models (last I used 4.7) in that I yet had not the occasion where model did something unintentional that I did not direct it to.

EDIT: Updated percentage reduction.

wg0··on Quantum battery upends the rules of charging
Not happening.
wg0··on LinkedIn CringeBot 3000
Would love to see what was it.
wg0··on US strikes $1.2B deal to pay German firm to halt offshore wind projects
Man - What exactly are these people?
wg0··on DeepSeek-V4-Flash Update
Don't know about the bench marks but I am getting Opus 4.7 level performance at fraction of cost with DeepSeek V4 Flash set to high. It is a reliable workhorse.
wg0··on SQLite in Production: Optimizing WAL Mode, Concurrency, and VFS Layers
I am obsessed with the idea of per tenant databases. But I am afraid of migrations. Has anyone tried that?
wg0··on Our position on open-weights models
Basically:

"We are in favour of 3D printers but there should be a body that tests and certifies that a 3D printer cannot print anything that can be used as weapon. Anything pointy or with a spring and recoil or... or..."

wg0··on The Kimi K3 Moment
> tied to the corrupt Trump administration that are neither the highest quality nor the cheapest

Someone is calling corrupt as corrupt. Surprising.

wg0··on SpaceX bond worth 10% less than issue price – heading for junk bond status
Is it the bond or is it the share price that also tanked below IPO level?
wg0··on Amber the programming language compiled to Bash/Ksh/Zsh
Much needed. I have hard time understanding bash.
wg0··on GLM 5.2 and the coming AI margin collapse
Everyone is declaring GLM 5.2 as something that's really a big deal.

I don't know about that but based on my own experience with Deepseek v4 Lite alone (with high effort) I have no doubt in my mind that anyone claiming such great things about GLM 5.2 must be true because Deepseek v4 already is really awesome.

wg0··on Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5
This administration has damaged the US soft power. For decades, US was a reliable trade partner. Not any more.

EU is looking and charting its course already. Yeah, we can joke about it, we can mock it but it is in momentum already, one step at a time.

wg0··on DSpark: Speculative decoding accelerates LLM inference [pdf]
That's why I pay them. Regularly. Without fail. Despite my token usage isn't that much.

But I vote for these heroes with my wallet. Just yesterday did again.

wg0··on Anthropic says Alibaba illicitly extracted Claude AI model capabilities
It's just like web scraping is impossible to guard against.

Change my mind.

wg0··on Founding a company in Germany: €9600, 152 days and I still can't send an invoice
Try firma.de
wg0··on Deno Desktop
I hope bun desktop is coming soon?
wg0··on Apertus – Open Foundation Model for Sovereign AI
You might dismiss it as nothing but the Linux analogy does not work here either. It is more than that and direct threat to commercial AI labs and their business model. These labs are milking bunch of foundational papers for years now and the end is near.

Going forward would be such open source, open data and open recipe models possibly someday even with the training being crowd sourced if not inference like the BitTorrent model.

Lastly, even Chinese models (GLM, Deepseek, MiMax) work really really good and any user would testify that they do not miss OpenAI/Anthropic/Gemini at all if they're using those Chinese models which is argument enough that with such models, no one is going to miss Chinese models as well.

wg0··on Fable Converted Pylint to Rust
I think open source is dead. Basic issue is - if your product is open source or even open core, building a business around it would be impossible because someone else would point an AI agent at it and would have similar thing to offer.

Hence, closed source is what's next probably. Unfortunately.

wg0··on Leaked financial docs show OpenAI is losing billions of dollars a year
Yeah insane that people think it'll be okay in the long run but wondering how much different the financial status of other such company would be? Not much I guess.
wg0··on DeepSeek v4 Pro 1.6T model post-trained by Huawei on 1000 Ascend 910C chips
>As for the claim coming out of Shenzen, it carries no benchmarks, gives no figure for how long the run took, how it compared to the same job on Nvidia hardware, or how efficiently the 1,000-chip cluster was used. It’s ultimately just another addition to a series of dubious claims that have come from the Chinese state without anything to back them up; DeepSeek itself hasn’t commented.

As if OpenAI and Anthropic are giving us ball to ball commentary on how their training runs go. Deepseek did train it on domestic hardware, model might be out in public soon (open weights or not) and then anyone can see what is it about.

wg0··on Applying Brevity and Language Efficiency in Prompt Engineering
Snake oil. Tell me something that has hard irrefutable and reproducible evidence.
wg0··on EU Commission looking at practical consequences of Anthropic decision
Flip the rules put a blanket ban on US digital services and see what comes out of Europe within couple of years.

The only problem is - when US services are available, there's no incentive to bring anything to the market.

wg0··on How to Earn a Billion Dollars
I would not compare to lottery because a lottery ticket is bought at pure chance whereas these startup are taken onboard after very thorough audit.
wg0··on How to Earn a Billion Dollars
Does that work always without fail with anyone anywhere anytime?
wg0··on Don't trust large context windows
Not specific to Opus but yes it would make mistakes. I usually try to keep context window under 10%
← PreviousPage 2 of 34Next →