But as you pointed out, while that absolves ChatGPT, it makes Opus look worse.
10,135 karma · joined July 16, 2012
If anything I've written on this site seems interesting, or confusing, or you think I'd be interested in something you've written/read, please let me know: hn@justinblank.com.
But as you pointed out, while that absolves ChatGPT, it makes Opus look worse.
Technically true, extremely misleading. The law placed vastly fewer restrictions on who was a legal immigrant for much of that period. Of course, it is true starting in 1921, and 1924, legislation was put in place to limit immigration by "undesirables" such as Asians, Italians and Slavs.
Curious...which one of those regimes do you want to bring back?
Your mileage may vary. I live without air conditioning, but when it’s 30, I’m pretty tempted to defect. But I also tend to run hot.
I realize the idea that rules exist even if you dislike the people they (momentarily) protect, even if those people have violated them, is a tough pill to swallow.
I don't think blurring faces fully addresses that (though I'm not commenting on legality in Amsterdam).
I can see the argument that this video is in the public interest (hey, I'm Americanese), but it does seem plausible some people would find it a violation of a norm.
This is evidence, but it would be a mistake to think that high profile online contests would have the same dynamics as low profile online polling. 4chan is characterized by an insane level of effort on the specific jokes that they care about. But they don't care about a random poll nearly as much as they care about big stunts like this.
I do think it’s plausible this is a smaller release. Not that this was the real point of the discussion, but I think it’s still just a good idea to have more than one release a year. It keeps things moving smoothly, and lowers the cost of missing a release, which has beneficial effects.
Admittedly, the terminology here is almost designed to be maximally confusing, and I’ve never read a good post that laid out how everything relates.
Also, in this case, the game name is not “Game A” but something like “Deep Seek v4 Pro”, which they have previously chosen to use to describe Deep Seek v4 Pro, not Deep Seek v4.1 Flash.
However, we have to distinguish a few hypotheses:
1. No careful readers will notice when a piece is AI written.
2. Careful readers will generally not notice AI writing.
3. Everyone who writes comments on HN will reliably classify writing as AI or not.
Yes, 3 is not true, but Bryan’s point depends on something in the area of 2.
The ability to distinguish AI writing depends on having a good ear. For people who lack it, they either don’t notice and don’t care, or they make paranoid accusations against anything that is remotely non-standard (“you used an em-dash, you must be AI!”).
But I absolutely think it is bad for the reader, and deserves to be called out. Authors should know that it’s not good enough.
The summary I shared is much more straightforward. It’s mediocre, but just barely good enough to extract the message without making me super annoyed.
> So the difference has to be in what the JIT generated, and the profiler gives us exactly that.
> That is the whole vocabulary. Let’s read some code.
> Decoding the x86 version instruction by instruction is out of scope here.
> What is not architecture specific is the logic.
Here’s a segment flagged by Pangram: https://www.pangram.com/history/87e25169-30a4-4030-a65a-dba8...
Slightly less annoying summary from ChatGPT free: https://chatgpt.com/share/6a9ac7a3-15a0-83eb-8c2a-6f72cd9beb....
Caveat emptor: it makes high level sense, but I haven’t thought about it in detail.
I agree that you cannot do it reliably, in principle.
However, Bobby Fischer peaked at 2785, and Gary Kasparov peaked at 2851. These are not far from what informed observers suggest--maybe 50 or 100 points off. They are well into super-grandmaster territory. Kasparov would be 1st today, Fischer would be 4th.
But on goratings.org, the top player of the 90s would be roughly 30th today.
My point is that the go ratings are much more unstable than the chess ratings. With chess ratings, you'll be wrong in the details. With go ratings, you'll be catastrophically wrong.
(This is especially true if they don't get briefed "this is time-traveling Lee Changho, he doesn't know contemporary joseki, play a trap variation").
However, you can't compare goratings over time, the top ranks are not nearly stable enough. https://www.goratings.org/en/history/ (I think it's believable Shin Jinseo is better than Lee Changho, but not that there has been steady progress since the days of Lee Changho, so that there are now 20 players stronger than him).
It's not a comparison of the worth of the games (I play both, though I'm better at Go, and prefer it), but the dynamic range of Go is larger.
That said, any cross-game/sport comparisons of this kind are pretty tough to do properly.
[0] not technically distillation. https://thomasdullien.github.io/posts/2026-06-15-rl-economic...
- Having the feature available seems good.
- Having the feature default on is debatable, but not outside the realm of possibility.
- Having the feature default on, without any kind of release notes that say that it happens and link to the option to disable it is clownshoes "bring back the ghost of Steve Jobs to yell at the product team" type behavior.