HNHacker News
TopNewBestAskShowJobs

red75prime

1,378 karma · joined September 24, 2015

submissionscomments
red75prime··on “Math 2.0” will need to value mathematical progress more holistically
I think determinism has nothing to do with it. If you mean sensitivity to word ordering and such, it's a generalization failure.
red75prime··on “Math 2.0” will need to value mathematical progress more holistically
> The core technology of an LLM is sampling from a distribution so there is literally no way to make it deterministically robust (only probabilistically).

An LLM mostly deterministically (except parallel processing nondeterminism that can be mitigated) produces a probability distribution that can be sampled deterministically: just take the highest probability token or use beam search.

red75prime··on “Math 2.0” will need to value mathematical progress more holistically
The problem is that people strongly believe that this is an insurmountable problem that will persist indefinitely (or for a long time) and plan accordingly, while this, most likely, will be fixed soon by adding RLCAF (RL on conversational agent feedback) or something like that.
red75prime··on Navier–Stokes Lost in Translation
You can't consider them monkeys if you need 1000 of them instead of 2^1000.
red75prime··on Sharing AI progress in mathematics
"If you have nothing to say, don't post a chatbot's responses, because anyone can ask the chatbot directly if they wanted to"-principle? Well, people can't ask their chatbot directly, because it's not public.
red75prime··on In Ukraine, distributed renewables foil Russia's assaults
I'm absolutely not a fan of Trump and of politically motivated liars in general regardless of their affiliations.
red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
Whatever, then. It's totally in line with what I described as a typical usage of "category error".
red75prime··on Sharing AI progress in mathematics
> we can clearly specify what AGI or ASI is

We'll have plenty of time for this, while living off UBI.

red75prime··on ChatGPT is adding real cartoonists' signatures to fake New Yorker cartoons
"You've made a machine that can be easily tuned to produce almost precisely my exact hammer and people who use the machine get freaked out that I can sue them."
red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
The original claim is "They’re fundamentally not suited to thinking like we do."

> but training does not produce such LLMs

If we are talking about fundamental limitations it should not be an empirical observation: "does not produce" (which is factually wrong, BTW). It should be a fundamental limitation: "can not produce in principle."

I still don't understand what you are talking about when you say "category mistake." I was talking about computational capabilities of LLMs with CoT that their training can exploit, not about Befunge-98.

red75prime··on Anthropic reported diary entry to police, woman faces felony charge
Aren't high trust societies staying that way by kicking out the incorrigible violators? Pondering violence might not be treated lightly even in a high-trust society.
red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
"Category mistake" outside of kindergarten usually means "I don't like your definitions, but I can't bother to explain which definitions I use."

What do you mean exactly? "Thinking is not an algorithm."?

red75prime··on A 40ms Go garbage collector pause caused by swap
You just need to remember that garbage (unreachable reference cycles) isn't collected.
red75prime··on Why the Bronze Age Collapsed
Or a single disease the New World inhabitants have adapted to. Someone could have been bitten by a bat or something like that.
red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
Turing completeness applies to a model of computation, not to a physical instantiation of a machine. The stumbling block of "figure out how the brain works" applies more to the argument like the one I was responding to. How a person can know that a general model of computation can't implement the way people think, if we don't know how people think?

The existing LLM training methods on the other hand give the results that are hard to distinguish from "thinking like people," judging by the end results.

red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
"They are fundamentally not suited to thinking like we do" stays wrong nevertheless. They are fundamentally suited to everything not proven to be outside their modelling ability.
red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
> They’re fundamentally not suited to thinking like we do

LLMs with CoT are Turing-complete. So, theoretically, they can implement any kind of finitely describable algorithm (barring super-Turing computations).

red75prime··on LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents
LeCun calling them "world models" gives a high-level description of the desired functionality. They are Joint Embedding Predictive Architectures (with SIGReg). They might produce more useful world models, but it's yet to be seen.
red75prime··on In Ukraine, distributed renewables foil Russia's assaults
> The last part is why they will remain illegal

Where are intermittent energy sources illegal?

red75prime··on Inside Anthropic's Quest to Instill Morality into Its A.I. Models
We aren't dealing with individuals coming from intelligent species that have managed not to self-destruct (yet). AIs don't have the corresponding evolutionarily selected mechanisms. Those mechanisms need to be created de novo.
red75prime··on Several vulnerabilities have been discovered in the Linux kernel
Provable correctness is much harder than focusing on the happy path. If a pipeline doesn't include a look-for-the-vulnerabilities stage to save costs, a model wouldn't go out of its way to do it. The models are trained to do what they are asked to do.

I forget to add the obvious: some garbage gets through despite the mitigations. In the limit of pure garbage input, you'll get a model that internalized garbage generation. And you'd be better off throwing it away and starting over.

red75prime··on Several vulnerabilities have been discovered in the Linux kernel
This is simplistic to the point of being blatantly wrong. Training data isn't garbage. It's programs that do their job, but that are sprinkled with errors. Uncorrelated errors gets averaged out during autoregressive pretraining. Correlated errors can be somewhat suppressed during post-training. Hallucinations (of the generalization-error kind) can be dealt with using synthetic data that improves the model's generalization.
red75prime··on Why the Bronze Age Collapsed
It seems to be just a quirk of fortune that it wasn't the other way around. That is that Europe wasn't decimated by American deceases.
red75prime··on Vermont replacing power plants with home batteries
Batteries compete with peaker plants for diurnal variance, but they can't replace peaker plants for seasonal variance. It's an interesting dynamic.
red75prime··on Thinking fast and slow in AI: The role of metacognition (2021)
It reminds me of "At the time we drew boxes labeled 'perception', 'cognition' with arrows between them." An imprecise quote that I can't place.

I guess my box labelled 'subconsciousness' is trying to say that low-level mechanisms that give rise to the observed cognitive phenomena might have nothing to do with neat boxes.

red75prime··on When did Google get so weird?
Astronomically (or better to say combinatorically) large Markov chain can be used to describe a foundational model, but it doesn't capture generalization ability of the foundational model, which is demonstrated by post-training.
red75prime··on How to keep enjoying programming in a world of LLMs
Programmers don't work reliably too.
red75prime··on Google's first Suncatcher orbital data center test launches October 1
_Radiative cooling_ is the only option because it's in vacuum. But you can transfer heat to the radiator either passively (using thermal conductivity), or actively (using coolant). The second option allows to effectively use larger radiators.

BTW, we seem to have Noisesphere instead of Noosphere.

red75prime··on Google's first Suncatcher orbital data center test launches October 1
They use passive cooling in their MVP satellite (for the sake of simplicity, probably). This, naturally, severely limits the heat rejection capacity.
red75prime··on Yes, Claude can do nine loops
Yep. There goes "hallucinations are inevitable" crowd.
Page 1 of 34Next →