HNHacker News
TopNewBestAskShowJobs

cwyers

12,226 karma · joined December 14, 2013

[ my public key: https://keybase.io/colinwyers; my proof: https://keybase.io/colinwyers/sigs/oY3s_sY1T5jtSxebr0LrQSuJ8_6mFLfNR5_RjA_D6yU ]
submissionscomments
cwyers··on Verizon is about to break our Gizmo watches
```The associate agreed with me that they should not deprecate the old app until the new app can handle this configuration. I asked them to raise this up the chain: they need to push back the deprecation date.```

There's no way a CSR has any power over this.

cwyers··on What happens if Japan takes in zero immigrants?
I am really shocked at the tone of so many of the comments here. Did HN become a breeding ground for xenophobia at some point? Has it always been that and is it just way more mask off now?
cwyers··on From Rust to Ruby
> and then when the work was done they were happy with the result

It's worse than that -- they don't even know the result! They never tried to run it!

cwyers··on The Biochemical Beauty of Retatrutide: How GLP-1s Work
I was on a GLP-1 a few years ago and lost 70 lbs. After I got off, I kept a ton of diet changes (no more Pepsi or Gatorade and a lot of water instead, switching to whole grains and fiber/protein variants on pasta, etc.) and gained the weight back in a year and a half. The literature backs this up: keeping up weight loss is hard.
cwyers··on Rewrite Bun in Rust has been merged
Because it passes the existing test suite? And he knows what's in the test suite?
cwyers··on LLMs Are Not a Higher Level of Abstraction
If people can figure out how to write RFCs about IP over carrier pigeons for April Fools, they can figure out how to conceive of LLMs as a layer of abstraction beneath a protocol as well.
cwyers··on John Bradley, author of xv, has died
I was wondering why HN was linking to Vox Day. The answer "because the alternative is worse" is probably the most justifiable one I can think of.
cwyers··on OpenAI agrees with Dept. of War to deploy models in their classified network
There's a lot of people in this thread that assume that Sam Altman is the one who is being dishonest here, and I kind of understand, but the other two parties who could just as easily be lying are Pete Hegseth and Donald Trump, and of the three of them if you think sama is the _most_ likely to lie I feel like you have not been paying attention.
cwyers··on GPT-5.3-Codex
Codex now lets you tell the LLM tgings in the middle of its thinking without interrupting it, so you can read the thinking traces and tell it to change course if it's going off track.
cwyers··on Only 5 Sears stores remain in the U.S.
Ironically they had the foresight, they were just too early/didn't execute. They ran an online service (co-owner with IBM and CBS) called Prodigy that competed with AOL and CompuServ, and they tried to do online shopping there.
cwyers··on Ruby 4.0.0
Just use WSL2 and Docker Desktop. VS Code has DevContainer support so you can standardize on a Docker image for your project.
cwyers··on I have to give Fortnite my passport to use Bluesky
It's as much of a stretch as describing using an Azure service as "I have to use Halo" or AWS as "I have to use Rings of Power."
cwyers··on Hacker News front page now, but the titles are honest
"Slop" is at _least_ as fair a description of "we had an LLM rewrite HN headlines" as "we rewrote it in Rust so you have to upvote it" is of "we removed our biggest source of crashes on Android by getting rid of Go FFI issues."
cwyers··on Jonathan Blow has spent the past decade designing 1,400 puzzles
I mean, he's _allowed_. The Compiler Police aren't going to roll up to his house and take away his Jai compiler if there isn't a quorum of HN users blessing his efforts. But people can point out they don't feel the juice is worth the squeeze. Also, Blow is certainly an advocate for his position, which means this kind of public debate is germane to the question of if _other_ people should adopt Jai.
cwyers··on Jonathan Blow has spent the past decade designing 1,400 puzzles
If you have bills to pay, it really is.
cwyers··on Jonathan Blow has spent the past decade designing 1,400 puzzles
The point is that Blow has two blockbuster hits under his belt and can afford to take a decade to ship a single game. Most people would go broke never having shipped a game if they tried to do things Blow's way.
cwyers··on I fed 24 years of my blog posts to a Markov model
Yeah, there's only two differences between using Markov chains to predict words and LLMs:

* LLMs don't use Markov chains, * LLMs don't predict words.

cwyers··on Perl's decline was cultural
Node.js is the most popular web framework/technology in the StackOverflow developer survey. Express is more popular than FastAPI, Django, Flask and Rails in the same survey. Just... what are you talking about?
cwyers··on Perl's decline was cultural
I was surprised by that, too, and assumed it was a decade-old article until I saw the date at the bottom. Both being mentioned before Python is wilder, as is the total exclusion of JavaScript.
cwyers··on GrapheneOS is the only Android OS providing full security patches
You can actually look at history and see what happens when IBM tries to wrest control of the PC platform back with the PS/2, which was a flop with consumers because it wasn't backwards compatible enough with IBM's own previous PCs or the wider PC market that developed. A bunch of PC clone manufacturers got together and came up with the EISA bus standard so they wouldn't have to pay IBM license fees for MCA, and made it backwards-compatible with ISA cards people already had. It was successful enough that IBM ended up adopting EISA for some of their PCs.

The other notable thing about the situation is that three companies ended up simultaneously responsible for a large part of the PC platform, originally -- IBM, Microsoft and Intel. They all worked in various ways to encourage competition to each other -- the reason we see OS competition on the PC platform is that IBM and Intel both found it in their interests to allow other OSes on the platform to reduce Microsoft's leverage over them. IBM in fact created one of the competing PC OSes out the gate, OS/2, which was originally an IBM/Microsoft joint project until they started feuding. Now, OS/2 is dead, but IBM's interest in being able to support their own OS instead of Microsoft's is a big reason the PC platform was built in an OS agnostic way. People criticize UEFI for locking down the PC platform more than the previous BIOS implementations, but UEFI is still _way_ more open than basically any other platform, most of which don't have a standard for bootloaders at all. It's really the absense of a standard for bootloaders that keeps most Android phones locked down. Two Android phones from the same OEM might have different bootloaders, much less two phones from different manufacturers. We've yet to see an alternate OS with the resources to support implementing their own bootloaders for a majority of Android phones.

cwyers··on GrapheneOS is the only Android OS providing full security patches
Because the original IBM PC was designed to be cheap and built in a hurry. IBM had a mandate for the original PC to use off the shelf components as much as possible. They also neglected to secure an exclusive license from Microsoft for DOS. 95% of building an IBM PC clone was buying the same parts and getting a DOS license from Microsoft (which they were very happy to sell you). Everyone saw what happened to IBM and just didn't do it that way again.
cwyers··on MinIO is now in maintenance-mode
MinIO is absolutely not a passion project, it's a business.
cwyers··on Composer: Building a fast frontier model with RL
I'm not saying SWE-Bench is perfect, and there are reports that suggest there is some contamination of training sets for LLMs with common benchmarks like SWE-Bench. But they publish SWE-bench so anyone can run it and have an open leaderboard where they attribute the results to specific models, not just vague groupings:

https://www.swebench.com/

ARC-AGI-2 keeps a private set of questions to prevent LLM contamination, but they have a public set of training and eval questions so that people can both evaluate their modesl before submitting to ARC-AGI and so that people can evalute what the benchmark is measuring:

https://github.com/arcprize/ARC-AGI-2

Cursor is not alone in the field in having to deal with issues of benchmark contamination. Cursor is an outlier in sharing so little when proposing a new benchmark while also not showing performance in the industry standard benchmarks. Without a bigger effort to show what the benchmark is and how other models perform, I think the utility of this benchmark is limited at best.

cwyers··on Keep Android Open
The short version is: the PC is a historical accident. By "the PC" I mean "the Windows-Intel platform on which most consumer PCs were built." Linux and BSD were both able to exist in the form they did because there was a commodity hardware platform that was standardized (ad-hoc standardization, mind you) and _somewhat_ open. IBM, Microsoft and Intel were all best frenemies, able to exert enough power to standardize the PC platform but also able to exert enough power against each other to prevent them from locking the platform down too much. There is no standard "smartphone" platform like there is with the PC, really the only standard is Android AOSP. Because of this, it's a lot harder to do a third-party phone platform without adopting large parts of Android's code.
cwyers··on Composer: Building a fast frontier model with RL
The lack of transparency here is wild. They aggregate the scores of the models they test against, which obscures the performance. They only release results on their own internal benchmark that they won't release. They talk about RL training but they don't discuss anything else about how the model was trained, including if they did their own pre-training or fine-tuned an existing model. I'm skeptical of basically everything claimed here until either they share more details or someone is able to interpedently benchmark this.
cwyers··on A conspiracy to kill IE6 (2019)
If you read the article, one of the buttons on the bar prompted people to upgrade to the latest version of IE.
cwyers··on Just talk to it – A way of agentic engineering
LLMs are good at pursuing objectives, but they aren't necessarily good at juggling competing objectives at once. So you can picture doing the following, for instance:

- "Here is a spec for an API endpoint. Implement this spec."

- "Using these tools, refactor the codebase. Make sure that you are passing all tests from (dead code checker, cyclomatic complexity checker, etc.)"

The clankers are very good at iteratively moving towards a defined objective (it's how they were post-trained), so you can get them to do basically anything you can define an objective for, as long as you can chunk it up in a way that it fits in their usable context window.

cwyers··on Doctorow: American tech cartels use apps to break the law
The Doctorow school argument, as best I can tell, would go 'the regulations on black car service were meant for things like limo services that don't compete directly with taxis, and once Uber started competing directly with taxis, regulators and authorities should have moved more aggressively to write new regulations/laws that regulated Uber the same way taxis are regulated.' They would not agree with "the reason why taxis are tightly regulated are for reasons that mostly do not apply to Uber."

And this is exactly why I think the question of "what is the correct way to regulate car ride services" shouldn't hinge on incumbency bias towards taxis, but actually ask the question of what is best for participants in the market (which doesn't just include taxis and Ubers but also includes public transportation and its users, for instance). But that doesn't fit neatly into Doctorow's enshitification narrative.

cwyers··on Doctorow: American tech cartels use apps to break the law
Yeah, it's a real thing that happened to me, to. In multiple US cities. And I'm sure we're far from alone.
cwyers··on Doctorow: American tech cartels use apps to break the law
The opening of the article is laying out the case that the laws are good -- they make the market legible to participants. As he says:

``` To navigate all of these technical minefields, you need the help of a third party. In a modern society, that third party is an expert regulator who investigates or anticipates problems in their area of expertise and then makes rules designed to solve these problems.

To make these rules, the regulator convenes a truth-seeking exercise, in which all affected parties submit evidence about what the best rule should be and then get a chance to read what everyone else wrote and rebut their claims. Sometimes, there are in-person hearings, or successive rounds of comment and counter-comment, but that’s the basic shape of things.

Once all the evidence is in, the regulator—who is a neutral expert, required to recuse themselves if they have conflicts—makes a rule, citing the evidence on which the rule is based. This whole system is backstopped by courts, which can order the process to begin anew if the new rule isn’t supported by the evidence created while the regulator was developing the record.

This kind of adversarial process—something between a court case and scientific peer review—has a good track record of producing high-quality regulations. You can thank a process like this for the fact that you weren’t killed today by critters in your tap water or a high-voltage shock from one of your home’s electrical outlets. ```

And this is central to Doctorow's point, right? The narrow question of the legality of Uber's current service offerings is actually pretty well litigated, and if Uber was as flagrantly illegal as he claims, "we're an app" wouldn't have kept them in business. Doctorow argues that this is happening through regulatory capture -- the case isn't primarily that Uber is violating the currently existing set of laws, regulations, court precedents, etc. It's that Uber is violating what the regulations _would be_ in a world where they had less market power with which to influence regulations.

And so it's not enough to argue about how the apps get around _current_ laws. By Doctorow's own arguments, we're debating the merits of a counterfactual set of different regulations that we would have if you changed current conditions. And at that point, it is absolutely fair game to ask if this counterfactual set of different regulations is actually better for market participants.

Page 1 of 34Next →