HNHacker News
TopNewBestAskShowJobs

ggcr

388 karma · joined January 8, 2023

submissionscomments
ggcr··on Best LLM for every budget, updated daily
Oh yeah totally, good point. I remember there was controversy around Claude tokenizer generating more tokens lol
ggcr··on Best LLM for every budget, updated daily
What's impeding a lab from releasing its own new model, pricing it really low for the beautiful Pareto plot, accompanied by phrases VC love like "establishing a new frontier in cost", to just then raise prices back up?
ggcr··on GPT-6 Sol and Luna
Live notification in Codex:

> GPT-5.6-Sol is retiring. This conversation will automatically switch to GPT-6-Sol

I don't recall OAI retiring a model so early lol. Similar arch?

ggcr··on Introducing System One Models and Jev
Interesting. Perhaps I can see this being quickly adopted in LLMs-as-a-judge, where you normally need (a) a structured answer, say, with lots of different fields (metrics) and (b) you want the judge to be fast, not being a bottleneck.
ggcr··on I've operated petabyte-scale ClickHouse clusters for 5 years
The CTO himself reported 600 commits and 300 PRs each day. And they can do so because they have a massive CI [1]:

> Every day CI runs about 20..80 million tests in 600 commits and 300 pull requests

> Last year, ClickHouse spent 360 years of machine time for CI

I am no user of CH so I can't talk about their product. But we are talking about a company with 686 employees as per their LinkedIn, where ClickHouse is clearly the core of their business. Considering all of this, is 50+ commits a day that much?

[1] https://presentations.clickhouse.com/2026-openhouse-sf/great...

ggcr··on Apple Unveils iPhone Duo
> I am so used to using one hand to operate phones.

I even forgot about this. God damn it, I enjoyed the SE so much because of this.

ggcr··on iPhone Duo
Seems like a big engineering feat by Apple, but for real, who asked for this?

Edit: Is there that many demand for a foldable iPhone that justifies Apple making such engineering investment?

ggcr··on Navier-Stokes – Tristan Buckmaster [pdf]
Reminds me, kinda, to when Astra was launched and OpenAI announced an improvement to the bounded prime gap. Which BTW, Prof. Julia Stadlmann had published an independent result only a few days earlier

Stadlmann improved it from 246 to 240, OpenAI later claimed 186 I think?

Maybe someone can help clarify? I am no expert at all, but I can't help but see similarities.

[0] https://arxiv.org/abs/2608.31126

ggcr··on Navier-Stokes – Tristan Buckmaster [pdf]
> Let me make plain what I have said to colleagues in private: in view of this body of work, I believe Luis Martínez-Zoroa deserves a Fields Medal.
ggcr··on Show HN: GET Together – A social network where you don't need POST to Post
This would go hard if I was an OpenAI agent
ggcr··on The efficient frontier of LLM inference
2026 has been the year where spec-dec has matured, it has been adopted by all big OSS engines and i'm sure it's present in quite a lot of inference providers as the default

I feel like P/D dissaggregation will be the next big one for providers, as prefill tends to be compute bound while decode mem bound which I guess each will have a different type of node

ggcr··on Nvidia agrees to acquire Hugging Face for $13B
Here is an employee of HF at a time answering how do they make money, supposedly [1]

> We make money via compute credits + Enterprise Hub + HF Pro subs [...]

I guess this plus custom inference deployments, external inference providers, partnerships with the big cloud AWS, Azure, etc.

[1] https://x.com/reach_vb/status/1928050126498713706

ggcr··on "Famous Deep Learning Papers", David Bau
Agree. I'm not into interpretability but I always try to follow what Bau and his team are up to.
ggcr··on Ox-Alpha Is GLM?
Z.ai founders were one of the pioneers in MM-LLMs with Cog-VLM years ago, back when LLaVA emerged. I wouldn't be surprised if they added multi-modal capabilities
ggcr··on Ox-Alpha Is GLM?
Ziphu has that many resources to be able to serve capacity for 1 quadrillion tokens per day on Nous portal? My bet is that it's a Composer model from Cursor running on xAI cluster, they already did a Composer based on Kimi-K2.5
ggcr··on Follow live the open training of a 535B (23B activated) LLM
Announcement tweet by Stanford prof Percy Liang and Marin: https://x.com/percyliang/status/2090918065634684997 https://xcancel.com/percyliang/status/2090918065634684997

It will include Pre-training and Mid-training. ETA 3 months.

Live wandb dashboard: https://wandb.ai/marin-community/marin_moe/reports/535B-A23B...

GitHub issue tracking progress: https://github.com/marin-community/marin/issues/8435

23T tokens (not deduped) directly in S3 for anyone to see: s3://marin-us-east-02a/marin/datakit/store_4d2e363d

Intermediate checkpoints also in S3

ggcr··on Ornith-1.5: From Self-Scaffolding to Self-Improvement
They should've included Qwen3.5-397B-A17B in the benches then :/
ggcr··on YC startups are abandoning .com
TIL that alphabet.com belongs to BMW group
ggcr··on Nvidia Nemotron 3.5 Lightning
> nvda will have an incentive to continue this kind of releases, even if other parties slowly abandon the open release of models

This! It's literally in their best interest for open-weights models to succeed

ggcr··on Nvidia Nemotron 3.5 Lightning
Nice cadence of releases by the Nemotron team :)
ggcr··on Discovery Loop
> Oriol Vinyals, Sanjay Ghemawat, Jeff Dean, Quoc Le

as founding members is crazy !

ggcr··on ClickHouse Labs, a new research group led by Andy Pavlo
Blog: https://clickhouse.com/blog/andy-pavlo-joins-clickhouse?utm_...
ggcr··on Qwen3.8-Max: A New Bar for Coding and Cowork
Waiting for Qwen3.8-27B :)

Their base models and architecture has quickly become the go-to for local inference and fine-tuning, even when they introduced some tricky things like GDN, so many people use it, that it was matter of days/weeks until lots of OSS frameworks adopted it.

ggcr··on DeepSeek-V4-Flash Update
Woah, a 200B model competing with GLM-5.2 and getting close to Opus 4.8. Quite impressive.

If those numbers translate well to its general capabilities, with the great caching DeepSeek has, I feel like this model will get tons of usage.

ggcr··on Inkling: Our Open-Weights Model
My personal bet is that this model should really shine in Autoresearch NanoGPT-style speedruns because its first-class integration with Tinker
ggcr··on YouTrackDB is a general-use object-oriented graph database
It's a fork of OrientDB, isn't it?
ggcr··on Ants: Who looks after the injured in a colony?
> I’m surprised they don’t just eject the injured worker from the colony

Wonder if this has something to do due with space constraints. If the study was done in a controlled nest, it must be space bounded one way or another. Dynamics might change when in real-world?

ggcr··on Ants: Who looks after the injured in a colony?
Fascinating. Hidden on the bottom of the article seems to be a video [1] showcasing how they track each ant out of the six colonies of 110 each.

I'd like to read the paper to skim over the methodology but it's not open-access :(

[1] https://www.uni-wuerzburg.de/fileadmin/uniwue/2026/0702Ameis...

ggcr··on OpenAI unveils its first custom chip, built by Broadcom
With Reinforcement Learning, inference is very present in post-training stages now too
ggcr··on Show HN: Lightweight C++23 S3 client with no extra deps (just curl and OpenSSL)
Fair point. Should've titled it "minimal dependencies" perhaps

Since C++ has no HTTP client in its std lib, I really had no other choice but to use curl. Same with OpenSSL. It'd be quite naïve of me to re-implement the whole HTTP stack and SHA256 from scratch =)

Page 1 of 2Next →