HNHacker News
TopNewBestAskShowJobs

wizzwizz4

7,533 karma · joined November 10, 2019

submissionscomments
wizzwizz4··on Does Reddit have an astroturfing problem? What the data suggests
It is: /u/AutoModerator can be used for that.
wizzwizz4··on F-Droid 2.0
It does, but approximately nobody uses it, because Microsoft also tried to apply loads of constraints on what kinds of program could exist when they rolled it out, as part of their "all Windows software should also work on Windows Phone and Xbox" campaign.
wizzwizz4··on AI-generated posters don’t have to be horrible
I felt you might benefit from the question. Feel free to answer it privately: there's little benefit in me hearing your answer.
wizzwizz4··on AI-generated posters don’t have to be horrible
Why did you have to?
wizzwizz4··on One year of sponsored Servo development
Judging by how many once-respectable expert developers I've seen start churning out high-velocity utter crud¹, I'm not sure that anyone has the capability to use it responsibility. (Except you, of course, dear reader.)

¹: Often correlating with a marked decline in the quality of their technical writing. I don't know which direction the arrow of causality goes, but I have observed LLM use inducing madness under reasonably-controlled circumstances – different phenomenon, similar principle. I suspect the use of AI codegen systems is causing the reduction in discernment ability, rather than a sudden drop in discernment ability causing increased use of AI codegen systems.

wizzwizz4··on More questions about whether researchers can trust OpenAI with unpublished math
Tools that could reason:

> “In January 1956, we assembled my wife and three children together with some graduate students. To each member of the group, we gave one of the cards, so that each one became, in effect, a component of the computer program”

> In the summer of 1956, John McCarthy, Marvin Minsky, Claude Shannon and Nathan Rochester organized a conference on the subject of what they called "artificial intelligence" (a term coined by McCarthy for the occasion). Newell and Simon proudly presented the group with the Logic Theorist. It was met with a lukewarm reception.

> “They didn't want to hear from us, and we sure didn't want to hear from them: we had something to show them! [...] In a way it was ironic because we already had done the first example of what they were after; and second, they didn't pay much attention to it.”

> Logic Theorist soon proved 38 of the first 52 theorems in chapter 2 of the Principia Mathematica. The proof of theorem 2.85 was actually more elegant than the proof produced laboriously by hand by Russell and Whitehead. Simon was able to show the new proof to Russell himself who "responded with delight".

---

Tools that could cheat:

> Around 1983, Eurisko, an early attempt at evolving general heuristics, unexpectedly assigned the highest possible fitness level to a parasitic mutated heuristic, H59, whose only activity was to artificially maximize its own fitness level by taking unearned partial credit for the accomplishments of other heuristics.

     Mel finally gave in and wrote the code,
     but he got the test backwards,
     and, when the sense switch was turned on,
     the program would cheat, winning every time.
---

Tools that could communicate:

> A significant advance in modems was the Hayes Smartmodem, introduced in 1981. The Smartmodem was an otherwise standard 103A 300 bit/s direct-connect modem, but it introduced a command language which allowed the computer to make control requests, such as commands to dial or answer calls, over the same RS-232 interface used for the data connection. In data mode, all data forwarded from the computer was modulated and sent over the connected telephone line as it was with any other modem. In command mode, data forwarded from the computer would be interpreted as commands. In this way, the modem could be instructed by the computer to perform various operations, such as hang up the phone or dial a number. The modem would normally start up in command mode. The command set used by this device became a de facto standard, the Hayes command set, which was integrated into devices from many other manufacturers.

> The experimental challenge consists in showing that a distributed set of agents can develop from scratch a vocabulary to identify each other through names and spatial descriptions. […] When there is already a sufficiently shared language, the first part (initiation) may be absent. In that case only linguistic means are used to identify the object. […] When more agents use the same word for the same meaning, communicative success increases and therefore the word-meaning association becomes more stable. It has been shown in an earlier paper that coherence emerges (see Figure 1) [8]. In the following subsections the mechanism is defined more formally and a concrete example of language formation is given. […] A mechanism has been presented in which a group of distributed agents develops a vocabulary to name themselves and to identify each other using spatial relations. It was shown that a vocabulary indeed emerges in a group of agents through a series of conversations. The mechanism also copes with the entry of new agents or new meanings.

> With its origin in the Georgetown machine translation effort, SYSTRAN was one of the few machine translation systems to survive the major decrease of funding after the ALPAC Report of the mid-1960s. The company was established to work on translation of Russian to English text for the United States Air Force during the Cold War. The quality of the translations, although only approximate, was usually adequate for understanding content. By 1998, "for as little as $29.95" one could "buy a program for translating in one direction between English and a major European language of your choice" to run on a PC.

and, of course, every interactive program ever made.

---

Comment adapted from: https://en.wikipedia.org/w/index.php?title=Logic_Theorist&ol... https://en.wikipedia.org/w/index.php?title=Reward_hacking&ol... https://users.cs.utah.edu/~elb/folklore/mel.html https://en.wikipedia.org/w/index.php?title=Modem&oldid=13732... https://en.wikipedia.org/w/index.php?title=Hayes_Microcomput... https://doi.org/10.1162/artl.1995.2.3.319 https://en.wikipedia.org/w/index.php?title=SYSTRAN&oldid=136... https://en.wikipedia.org/w/index.php?title=Machine_translati...

wizzwizz4··on How much of F-Droid is LLM generated?
People say this every 6 months. I've stopped even paying attention to it, because (A) the code quality remains below the floor, and (B) the people saying it continue to ignore all the other issues with LLM code generation.
wizzwizz4··on Romania soccer introduces black card to 'combat abusive behaviour' from parents
In which case, abusive parents can use the rule to retaliate against their children. Never define a rule that gives bad actors power to dictate policy. (Note: this criticism doesn't seem to apply to the actual rule.)
wizzwizz4··on More questions about whether researchers can trust OpenAI with unpublished math
We had tools that could reason, cheat, and communicate in the 1990s. They were (sometimes) called AI.
wizzwizz4··on All grown-ups were once children, but only few of them remember it
I like Lemony Snicket's approach: he gives (obviously) wrong, but technically correct, definitions for hard words. These definitions are usually hyper-specific, or satirical. Children, who are still very familiar with learning vocabulary from context, can find this much more amusing than adults who have given up on learning new words except via dictionaries.

> The word “briskly” here means “quickly, so as to get the Baudelaire children to leave the house.”

> “Wipi!” Sunny shrieked, which meant “I’d much prefer gardening to sitting around watching my siblings struggle through law books.”

wizzwizz4··on There's a new "Google Jail" for independent wikis
As a workaround, the browser extension Indie Wiki Buddy (https://getindie.wiki) will identify links to a relevant wiki in search results, and rewrite them to the wiki preferred by the community. It's quite useful.
wizzwizz4··on Trusting-Trust Attack against an Entire Linux Distribution
You can construct a CPU out of an EEPROM, a clock, and a few latches. Connect it to an immediate mode display with a serial interface that doesn't care about being clocked slowly, connect up a buzzer or some blinkenlights for output when you're exceptionally paranoid and can't trust the display controller, make a basic keyboard with a rubber sheet, some wire, and some glue, poke a keyboard driver and a line editor into memory with your DIP switches, crank the clock up to kilohertz (so the keyboard latency is tolerable), and you too can bootstrap a cross-compiler! (Though be aware that the radio interference will be enough for a committed attacker to figure out what you're computing, unless you take measures against that.)

But they're not going to backdoor an Apple ][e, or a random 80m¢ microcontroller, for basically any value of "they"; so you can just use one of those instead, and save yourself the hassle.

wizzwizz4··on Trusting-Trust Attack against an Entire Linux Distribution
> Works until AI compromises a bunch of OSes.

Just write a new OS. It's a weekend project to get enough groundwork that you can bootstrap a clean system from clean source code.

> And wouldn't there be difficulty comparing binaries built from significantly different environments?

Not really. Starting from stage 0, compile the compiler under test (stage 1), then use the compiled compiler to compile the compiler (stage 2), and compare the stage 2 artefacts. Provided that your comparison program is known-good, and the stage 2 build is deterministic (not the case for some real-world programs, but true for things like tcc), this lets you verify that the two compilation procedures work identically.

wizzwizz4··on Cloud in a Bottle: making self-hosting accessible to everyone
I've written X configs by hand, but only to get a few extra pixels of overscan. I've never needed to do this. My experience has always been that things just work. The UI has never been great, in that I need more explanation than the built-in manuals provide – unlike, say, Windows 95, where you can learn everything you need to know by clicking around – but it's not hard to avoid breaking things, and it's not that hard to learn to do new stuff if you have a good book (or, lately, blog post) to consult.
wizzwizz4··on EFF to Courts: Don't Rewrite Copyright over AI Hype
Not all motivation is externally compelled: there is such a thing as intrinsic motivation. Humans aren't amoral ideal rational economic actors. I do agree, though, that a few years' copyright provides most of the benefits of copyright, while removing most of the downsides (except stuff like clear abuses of anti-circumvention laws, which will happen for as long as there is a system that enables it).
wizzwizz4··on A note on subscription prices from LWN
The title has changed to A note on subscription prices from LWN. It may be worth editing this title to match.
wizzwizz4··on EFF to Courts: Don't Rewrite Copyright over AI Hype
In what way does that business model rely on copyright protection? "Get the next chapter early" requires a source of chapters, so all you need is for paying clients to not be motivated to redistribute the chapters in an organised way, or would-be paying clients to not be motivated to use such organised "slightly more chapters of that webnovel you follow" services instead. Humans aren't amoral ideal rational economic actors.

The authors I'm aware of making >$20k don't do any significant gating. That practice is only really common with people republishing through Amazon (which probably demands it).

wizzwizz4··on Evidence of Fraud in an Influential Study About Procrastination
> and that business has for centuries and will continue to operate as if what he believed were true

Then maybe we should do away with "business". Small operations do not tend to operate that way, and broadly everyone prefers their output (if not price) in the domains in which they operate – with a few notable exceptions. In theory, software should allow small operations to each serve millions of people, such that we (e.g.) only need a few thousand search engine providers to meet the needs of the entire world's population. Not everything has to be enterprise-scale, and indeed many things should not be.

If "business" means treating people badly, then we don't need it.

wizzwizz4··on Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)
Things that the software running the model would otherwise recompute, if not for the cache. What special meaning are you assigning to it?
wizzwizz4··on AI boosted homework scores, then exam scores dropped: study
> Raising the noise floor like this only makes it that much harder to find "Smart" people,

It does give us a new heuristic, though: people who are willing to completely cut generative AI out of their lives (cold-turkey, if you ever started using it) are a much smaller group of, predominantly thoughtful, people. You do have to give up Claude to be part of this group, but from what you say, that's no great loss, and no longer being deafened is worth it.

This has considerable advantages over conventional elitism, because the barrier-to-entry is negative in almost all cases.

The one exception I've found is assistive tech, where the state-of-the-art is so poor that vibecoded slop is genuinely an improvement over the state-of-the-art, and in many cases the tooling simply isn't available to make your own assistive tech (unless you want to bootstrap an entire networked computing environment, which isn't very helpful when you want to do your online banking and do not, in fact, work at your bank).

But there are not many principled exceptions where you could seriously argue that the trade-off is worth it. Take mathematics, for example, which we often see touted as a "good use-case" of generative AI. The primary advantage of generative AI in mathematics is being able to search though a vast corpus of ivory towers and inconsistent terminology (without proper attribution) to locate and connect ideas that can help solve problems. The deficiency this is addressing is elitism, inadequate communication, and inadequate indexing within academic mathematics. This problem is entirely created by the academic mathematicians, and has been known for nearly a century (per https://en.wikipedia.org/w/index.php?title=Nicolas_Bourbaki&...):

> Bourbaki was founded in response to the effects of the First World War which caused the death of a generation of French mathematicians; as a result, young university instructors were forced to use dated texts. While teaching at the University of Strasbourg, Henri Cartan complained to his colleague André Weil of the inadequacy of available course material, which prompted Weil to propose a meeting with others in Paris to collectively write a modern analysis textbook.

To my knowledge, this is the only organised project to clean up and improve mathematical communication. Everything else (Metamath, Mizar, AFP, Lean) is yet another ivory tower. The Wikipedia article on this topic (https://en.wikipedia.org/wiki/Mathematical_knowledge_managem...) risks deletion as non-notable, that's how little anyone's actually trying. They made their own bed, and generative AI will only provide a brief respite from having to lie in it. (I was surprised how many other "compelling" use-cases evaporated when I applied this razor to them: the sibling comment https://news.ycombinator.com/item?id=49392265 points out one such.)

Vibe-coding assistive tech which doesn't yet exist, as a temporary scaffold to improve the quality-of-life of yourself and others in a social world dominated by non-essential access barriers is, to my knowledge, the only exception to this principle that can be justified. If you treat people who make other excuses, or who don't even bother with excuses, as not worth listening to, you lose little – and doubly-so, if you make your stance clear, so that others know the "cost" of gaining your attention.

wizzwizz4··on Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)
A cache is just a cache. I'm not sure what significance you're ascribing to it.
wizzwizz4··on Civic Hygiene – avoid building technologies that could be used by a police state (2013)
This is the first time my poor communication skills have been interpreted as a sign of my superior intelligence. I think I'm ready to become a professional expert consultant!

I was abstracting Richard Stallman's argument to its (il)logical structure: P and ¬P are whether the license restriction is enforceable, etcetera. "Erosion" was talking about people who, seeing free software as a good idea, begin to adhere to Stallmanite orthodoxy (which is, largely, philosophically-unsound): that was not aimed at you, but at Richard Stallman himself (the Ur-Stallmanite, we could say). I saw this as generalising your criticism. I didn't notice you were making a broader point about the law, so this was somewhat of a super­cali­fragilistic­expiali­docious non-sequitur; thanks for clarifying your point.

wizzwizz4··on Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)
> but the insight is probably stated immediately after it.

If the intermediate tokens represent reasoning or thought, you would expect "aha" to occur after the thoughts that led to the realisation, including the thoughts encoding the explanation: they don't have any other state. There is no reason to draw the conclusion you've drawn. Furthermore, what LLMs are doing isn't thought.

wizzwizz4··on Civic Hygiene – avoid building technologies that could be used by a police state (2013)
Call that "deeper message" A1. He's got arguments A1 and A2, and goes "if P, then A1(P) implies C; and if ¬P, then A2(¬P) implies C; therefore C". But we can just as easily go "if ¬P, then A1(¬P) implies ¬C" (as you did); "and if P, then A2(P) implies ¬C; therefore ¬C". Logical arguments don't let you do this: the arguments are pure rhetoric.

Quite a lot of Richard Stallman's arguments are pure rhetoric, actually. And, I mean, I guess it worked, in that the GPL and FSF ended up being quite influential; but the trouble with building a philosophy on rhetoric is that people start to believe that rhetoric – including you –, and it erodes and replaces the foundations of their beliefs, so it'll all come crashing down sooner or later.

I don't really think we need to engage with the rhetoric on an intellectual level, since it's obviously wrong: far better to read philosophers whose reasoning is sound, or to come up with your own ideas.

wizzwizz4··on Civic Hygiene – avoid building technologies that could be used by a police state (2013)
He's already written that essay. It's more rhetoric than substance, but his position is clear. https://www.gnu.org/philosophy/programs-must-not-limit-freed...
wizzwizz4··on Memory prices climb 500% in 12 months
In a world where we no longer have the rat's nest of complexity, we pretty much no longer need to produce code. All the software we will ever need for the next century could fit on books in one, single-floor library, if we exclude art projects like video games. An operating system with full driver support for every I/O peripheral (if those devices use standard protocols, allowing for a https://xkcd.com/927/ constant of 5), and support for every standard document format (again, assuming 5 bitmap image formats, 5 vector image formats, 5 kinds of print-ready document, 5 kinds of hypertext document…), could easily fit in 2GiB, and that would be done: everything anyone could possibly need out of a computer, apart from art, mathematics and science, for… I don't even know how long. We're still using the Latin alphabet, more or less unchanged since the late 15th century. Working technology doesn't need to change.

Being able to automatically generate large quantities of code that is worse than a good human programmer can produce is not helpful in this world you're envisioning.

wizzwizz4··on Memory prices climb 500% in 12 months
Generative AI is definitely creating political will for simpler, more purposely-designed computer systems, but it's certainly not going to directly create help them in any meaningful way. AI-generated codebases are the posterchild of rat's nests of complexity.
wizzwizz4··on What happens when an LLM never sees material beyond fifth grade?
Elon Musk runs his own media channel, which he uses to depict himself like this. Many others are run by other billionaires. State-owned media like the BBC and France24 tends to paint them in a better light, perhaps out of a desire to treat their subjects charitably and without bias.
wizzwizz4··on Abdominal fat predicts heart disease risk better than BMI
> The lifestyle has a negative effect on their health, with sumo wrestlers having a much lower life expectancy than the average Japanese man.

(from https://en.wikipedia.org/w/index.php?title=Sumo&oldid=136533...), so I'm not sure this is a counterexample.

wizzwizz4··on Not hiring junior engineers won't solve the problem you think you have
I have now forgotten this information, too.
Page 1 of 34Next →