HNHacker News
TopNewBestAskShowJobs

dTal

18,061 karma · joined June 2, 2013

submissionscomments
dTal··on Maryland becomes first state to ban surveillance pricing in grocery stores
Well that's easy enough - don't apply sneaky pricing when there's two people looking.
dTal··on Maryland becomes first state to ban surveillance pricing in grocery stores
E-ink price tags are not uncommon. Technology to track individual customers through the store based on smartphone RF is already deployed in many supermarkets. Some stores even do scan-as-you-shop, where the customer scans the item at the shelf, rather than at the front of the store. There are certainly a lot of i's to dot and t's to cross, but it's hardly a theoretical impossibility - find the right store and you could do it today with no more than a software update.
dTal··on Talkie: a 13B vintage language model from 1930
404
dTal··on Talkie: a 13B vintage language model from 1930
On the contrary, I think a "simulator of common objections" that reflects all the blind spots and biases of wider society is an extraordinarily valuable tool for exploring and evaluating new ideas. You might find that, in neatly summarizing what "everyone" thinks and justifying it as best it can, the LLM inadvertantly shines a spotlight on common misconceptions. Look there for the novelty.

In this case - the concept of using automated computing devices to manipulate numbers that represent ideas at arbitrary levels of abstraction was, by 1930, nearly an entire century old. Talkie's myopic viewpoint does not represent the most farsighted viewpoint, merely the average. So if, in 1930, you had read the writings of Ada Lovelace, gotten very excited, and wanted to figure out how to pitch it to investors - Talkie might have been very useful.

dTal··on How ChatGPT serves ads
The word I have heard is "bullshitting". Lies at least orient themselves with regard to the truth, bullshit floats free
dTal··on Localsend: An open-source cross-platform alternative to AirDrop
I love this software for its reliability (as compared to, say, KDE Connect, which I gave up on after years of frustrated use after it became clear that the developers did not believe there was an issue and it would never improve).

I do not love that it is a heavy electron app that takes many seconds to launch on my mid-spec machine and burns 20% of an entire CPU core the entire time it is running.

Why can't we have a simple command line tool that works?

dTal··on The quiet resurgence of RF engineering
Bit older than 8 years but even cramming a working GPS reciever into a phone was a huge, nontrivial achievement.
dTal··on Humpback whales are forming super-groups
>Every number so far has been wrong

No it hasn't.

dTal··on Humpback whales are forming super-groups
>strongly predicting what will happen in 30 years has always been wrong so far

No it hasn't, this is climate change denialist nonsense. In fact no less a figure than ExxonMobil correctly predicted the trajectory of global CO2 levels and corresponding increase in warming as far back as the 1970s and their predictions remain accurate today.

dTal··on US Department of Justice has officially reclassified cannabis as less dangerous
It's a good way to frame the discussion, with the caveat that for some things that subset is 5% of the population, and for other things that subset is 95%.

Is there a threshold? Can we define a principle that covers the entire range?

It seems clear that in the ideal scenario, people's freedoms should not be curtailed merely because there exist other people who would do unproductive things with that freedom. And on the other hand it seems clear that "freedom" to engage or not engage with deliberately targeted highly addictive things is not meaningful, and "individual responsibility" as an organizing principle of society only takes you so far.

dTal··on Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model
Ooh, car analogy time!

It's kinda like saying a car with a 6L engine will always outperform a car with a 2L engine. There are so many different engineering tradeoffs, so many different things to optimize for, so many different metrics for "performance", that while it's broadly true, it doesn't mean you'll always prefer the 6L car. Maybe you care about running costs! Maybe you'd rather own a smaller car than rent a bigger one. Maybe the 2L car is just better engineered. Maybe you work in food delivery in a dense city and what you actually need is a 50cc moped, because agility and latency are more important than performance at the margins.

And if you're the only game in town, and you only sell 6L behemoths, and some upstart comes along and starts selling nippy little 2L utility vehicles (or worse - giving them away!) you should absolutely be worried about your lunch. Note that this literally happened to the US car industry when Japanese imports started becoming popular in the 80s...

dTal··on Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model
Why? No one else was. The discussion was about OpenAI / Anthropic's lack of moat when there are open weights models that are almost as good. You can host them anywhere you like. Pay a US company to do so if you want.
dTal··on Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model
It's the same user and they already answered you: "If you read I’m talking about their service only models."

But yes this is a non-sequitor. The original question was "What competitive advantage does OpenAI/Anthropic has when companies like Qwen/Minimax/etc are open sourcing models that shows similar (yet below than OpenAI/Anthropic) benchmark results?"

Even if you don't trust Chinese companies, and you want a hosted model, you can always pay a third party to host a Chinese open weight model. And it'll be a lot cheaper than OpenAI.

dTal··on I don't want your PRs anymore
Won't be much "raw material" left before long, if everyone takes that view.
dTal··on Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving
The general case is that our own current relative ignorance on the best way to use and adapt pretrained weights is a short-lived anomaly caused by an abundance of funding to train models from scratch, a rapid evolution of training strategies and architectures, and a mad rush to ship hot new LLMs as fast as possible. But even as it is, the things you mentioned are not impossible, they are easy, and we are only going to get better at them.

>What if you need to reduce number of layers

Delete some.

> and/or width of hidden layers?

Randomly drop x% of parameters. No doubt there are better methods that entail distillation but this works.

> would the process of "layers to add" selection be considered training?

Er, no?

> What if you still have to obtain the best result possible for given coefficient/tokenization budget?

We don't know how to get "the best result possible", or even how to define such a thing. We only know how to throw compute at an existing network to get a "better" network, with diminishing returns. Re-using existing weights lowers the amount of compute you need to get to level X.

dTal··on Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving
None of that is true, at least in theory. You can trivially change layer size simply by adding extra columns initialized as 0, effectively embedding your smaller network in a larger network. You can add layers in a similar way, and in fact LLMs are surprisingly robust to having layers added and removed - you can sometimes actually improve performance simply by duplicating some middle layers[0]. Tokenization is probably the hardest but all the layers between the first and last just encode embeddings; it's probably not impossible to retrain those while preserving the middle parts.

[0] https://news.ycombinator.com/item?id=47431671 https://news.ycombinator.com/item?id=47322887

dTal··on Hyperscalers have already outspent most famous US megaprojects
We are in such different universes that I fear that this will not be a productive discussion; to my eyes LLMs are the most obviously socially transformative technology in my lifetime, up there with "internet" and "smartphones".

You say the largest niche is software production. Okay, let's talk about that. If the jury is still out then the jury is asleep. When ChatGPT first came out - the GPT3 days, years ago, before "vibe code" was even a term - an artist friend of mine who never wrote a line of code in his life straight-up vibe coded 3d visuals to accompany a performance of the band he was in. In Processing, which he'd never heard of until ChatGPT suggested it to him. Do you realize what this means? Normies can use computers now. Actually use, not just consume. You can describe what you want and the computer will do it - will even ask you for clarification if your specification is too ambiguous. Hell, it will even educate you about the subject matter, meeting you at exactly your level, in your favorite writing style.

If you are still thinking in terms of whether vibe coded software is "copyrightable" or whether LLMs are useful for "selling software", you are a blacksmith scoffing that cars are pointless because they don't need horseshoes. Your entire framework is obsolete.

dTal··on Hyperscalers have already outspent most famous US megaprojects
>There is no indication that LLMs are a pathway to meaningful and transformative AI.

Reality check, they are already astoundingly meaningful and transformative AI. They can converse in natural language, recall any common fact off the top of their heads, do research online and synthesize new information, translate between different human languages (and explain the nuances involved), translate a vague hand wavey description into working source code (and explain how it works), find security vulnerabilities, and draw SVGs of pelicans on bicycles. All in one singularly mind-blowing piece of tech.

The age of computers that just do what you tell them to, in plain language, is upon us! My God, just look at the front page! Are we on the same HN?

dTal··on Wacli – WhatsApp CLI
Practically speaking, it isn't secure; no closed app can be. It receives regular compulsory updates (old versions refuse to work) and there's nothing at all stopping Zuck from sneaking in backdoors targeted at you personally.
dTal··on Sam Altman's home targeted in second attack
>Altman's death would change nothing about that fundamental calculus. You'd have to kill probably tens of thousands of people to really put a dent in AI development

Your analysis seems to assume that people will remain more afraid of being "outcompeted" than of being murdered, even after a campaign of terrorism that would make 9/11 look minor.

>it often also creates new [problems], often surprising ones

Let's reframe this to remove the negative bias: murder has the obvious direct first-order effect of removing the target from existence, but also a host of non-obvious higher-order effects resulting from people's response to that violence. These can be counterproductive to the goals of the murderer, but they can also work in favor of it. That is why "terrorism" is a real thing - the higher-order effects are essentially a force multiplier, and if you have nothing to lose then the calculus of causing a major disruption begins to look favorable; any disruption, because regression to the mean is good if you're at the shitty end of the bell curve.

dTal··on Nowhere is safe
*Aryan - Ayran is Turkish buttermilk :)
dTal··on Nowhere is safe
I keep seeing comments that refer to Iranians as "brown people" - usually to emphasize their perceived "otherness" by the ignorant, as in this case. But Iranians aren't brown, or Arab apart from a small minority, and relatively speaking their culture isn't even that "other" - it would probably feel more familiar to the average American than some European countries even.

Do Americans really hear "Iran" and think of durka-durka from Team America?

dTal··on Claude mixes up who said what
>Fixed input-to-output mapping is determinism. Prompt instability is not determinism by any definition of this word

It really depends on your perspective.

In the real world, everything runs on physics, so short of invoking quantum indeterminacy, everything is deterministic - especially software, including things like /dev/random and programs with nasty race conditions. That makes the term useless.

The way we use "determinism" in practice depends contextually on how abstracted our view of the system is, how precise our description of our "inputs" can be, and whether a chunked model can predict the output. Many systems, while technically a fixed input/output mapping, exhibit an extreme and chaotic sensitivity to initial conditions. If the relevant features of those initial conditions are also difficult to measure, or cannot be described at our preferred level of abstraction, then actually predicting ("determining") the output is rendered impractical and we call it "non-deterministic". Coin tosses, race conditions, /dev/random - all fit this description.

And arguably so do LLMs. At the "token" level of abstraction, LLMs are indeed deterministic - given context C, you will always get token T. But at the "semantic" level they are chaotic, unstable - a single token changed in the input, perhaps even as minor as an extra space after a period, can entirely change the course of the output. You understand this, of course. You call it "prompt instability" and compare it to human performance. But no one would call humans deterministic either!

That is what people mean when they say LLMs are not deterministic. They are not misusing the word. It just depends on your perspective.

dTal··on April 2026 TLDR Setup for Ollama and Gemma 4 26B on a Mac mini
Ah, 'twas a mere jest, a sarcastic jab that of all the manifold builds provided, the most useful is missing - doubtless for good and practical reasons.

Nevertheless, worth looking at the Vulkan builds. They work on all GPUs!

dTal··on Show HN: Moon simulator game, ray-casting
I love this gorgeous and evocative little time waster and come back to it every now and then. Notes:

It starts out buttery smooth but over time its performance slows to a crawl. Changing window geometry seems to do some sort of garbage collection and it speeds back up. I just hit F11 twice real quick.

The optimal strategy is to try and make the trip parabolically with a single large burn at liftoff.

Gravity physics is of course symmetrical on ascent and descent, so the optimum time to start your deceleration burn is approximately when your downward velocity is equal to whatever your upward velocity was when you stopped burning.

dTal··on Show HN: Moon simulator game, ray-casting
The "car-like handling" is still physically accurate - thrusters automatically align your velocity vector to match your view direction. You can think of it as simply an interface - view direction is both a command and a display.
dTal··on Claude mixes up who said what
Sort of. They are deterministic in the same way that flipping a coin is deterministic - predictable in principle, in practice too chaotic. Yes, you get the same predicted token every time for a given context. But why that token and not a different one? Too many factors to reliably abstract.
dTal··on ML promises to be profoundly weird
Nobody said anything about Europeans having a "natural right". Bad enough to derail a conversation with irrelevant political nitpicking, unforgiveable to use a strawman to do so. Boo.
dTal··on Are We Idiocracy Yet?
I'm afraid you are misremembering. The movie is explicitly eugenicist. The people of the future are explicitly biologically stupid. The opening transcript is unambiguous:

[Man Narrating] As the 21st century began… human evolution was at a turning point.

Natural selection, the process by which the strongest, the smartest… the fastest reproduced in greater numbers than the rest… a process which had once favored the noblest traits of man… now began to favor different traits.

[Reporter] The Joey Buttafuoco case-

Most science fiction of the day predicted a future that was more civilized… and more intelligent.

But as time went on, things seemed to be heading in the opposite direction.

A dumbing down.

How did this happen?

Evolution does not necessarily reward intelligence.

With no natural predators to thin the herd… it began to simply reward those who reproduced the most… and left the intelligent to become an endangered species.

dTal··on Caveman: Why use many token when few token do trick
Huh okay, there was a major gap in my mental model. Thanks for helping to clear it up.
← PreviousPage 7 of 34Next →