HNHacker News
TopNewBestAskShowJobs

Zondartul

142 karma · joined May 13, 2022

Freelance C++ dev and EE enthusiast
submissionscomments
Zondartul··on Show HN: Axe – A 12MB binary that replaces your AI framework
It is easier to trust in the correctness and reliability of an LLM when you treat it as a glorified NLP function with a very narrow scope and limited responsibilities. That is to say, LLMs rarely mess up specific low level instructions, compared to open-ended, long-horizon tasks.
Zondartul··on Kilo Code: Speedrunning open source coding AI
In my unqualified opinion, LLMs would do better at niche languages or even specific versions of mainstream languages, as well as niche frameworks, if they were better at consultig the documentation for the language or framework, for example, the user could give the LLM a link to the docs or an offline copy, and the LLM would prioritise the docs over the pretrained code. Currently this is not feasible because 1. limited context is shared with the actual code, 2. RAG is one-way injection i to the LLM, the LLM usually wouldn't "ask for a specific docs page" even if they probably should.
Zondartul··on Why the weak nuclear force is short range
At some point our understanding of fundamental reality will be limited not by how much the physicists have uncovered but by how many years of university it would take to explain it. In the end each of us only has one lifetime.
Zondartul··on Cognitive load is what matters
Comments that explain the intent, rather than implementation, are the more useful kind. And when intent doesn't match the actual code, that's a good hint - it might be why the code doesn't work.
Zondartul··on The unbearable slowness of being: Why do we live at 10 bits/s?
Ask stupid questions, receive stupid answers.
Zondartul··on Making memcpy(NULL, NULL, 0) well-defined
What does "speculative" mean in this case? I understand it as CPU-level speculative execution a.k.a. branch mis-prediction, but that shouldn't have any real-world effects (or else we'd have segfaults all the time due to executing code that didn't really happen)
Zondartul··on NASA Investigates Laser-Beam Welding in a Vacuum for In-Space Manufacturing
Normal welding needs heat to melt the metals. Cold welding happens without heat. Two metal parts will cold-weld on any smooth, touching faces if the air molecules that keep the two separated disappear.
Zondartul··on NASA Investigates Laser-Beam Welding in a Vacuum for In-Space Manufacturing
Cold welding is unintentional, spontaneous joining of two metal parts in vacuum. You don't want that to happen, especially if the parts are meant to move.

Normal welding is intentional application of heat to partially melt two parts at the seam, so that they "mix" in semi-liquid state and become one part when they solidify. Welding may or may not use a third material (solder) to aid the process.

Zondartul··on The surprising effectiveness of test-time training for abstract reasoning [pdf]
ARC problems are too hard for me. I'm no longer sure I'm generally intelligent.
Zondartul··on Show HN: LlamaPReview – AI GitHub PR reviewer that learns your codebase
By "learns" do you mean "just shove the entire codebase into the context window", or does actual training-on-my-data take place?
Zondartul··on The Copenhagen Book: general guideline on implementing auth in web applications
To be fair, once someone has physical access to the machine, them having full access is just a matter of time and effort. So at that point it's security-through-too-much-effort-to-bother.
Zondartul··on It's Time to Stop Taking Sam Altman at His Word
Future NeuroLink collab? Grab the experience of qualia right from the brains of those who do the experiencing.
Zondartul··on It's Time to Stop Taking Sam Altman at His Word
If we figure out AGI, that still doesn't mean a singularity. I'm going to speak as though we're on the brink of AI outthinking every human on earth (we are not) but bear with me, I want to make it clear we're not going jobless any time soon.

For starters, we still need the AI (LLMs for now) to be more efficient, i.e. not require a datacenter to train and deploy. Yes, I know there are tiny models you can run on your home pc, but that's comparing a bycicle to a jet.

Second, for an AGI it meaningfully improve itself, it has to be smarter than not just any one person, but the sum total of all people it took to invent it. Until then no single AI can replace our human tech sphere of activity.

As long as there are limits to how smart an AI can get, there are places where humans can contribute economically. If there is ever to be a singularity, it's going to be a slow one, and large human AI vompanies will be part of the process for many decades still.

Zondartul··on GPTs and Hallucination
You can say "Bullshit". LLMs bullshit all the time. Talk without regard to the truth or falsity of statements. It also doesn't pressupose that the trueness is known, nir deny it, so it should satisfy both camps; unlike hallucination which implies that truth and fiction are separate.

I wonder if there is some sort of transition between recalling declarative facts (some of which have been shown to be decodable from activations) on one hand and completing the sentence with the most fitting word on the other hand. The dream that "hallucination" can be eliminated requires that the two states be separable, yet it is not evident to me that these "facts" are at all accessible without a sentence to complete.

Zondartul··on A Word, Please: Coffee-shop prompt stirs ChatGPT to brew up bland copy
My hunch is that since LLMs are trained on a per word basis (okay, per-token), vacuus verbosity is overrepresented.

If you have one normal sentence and one overly verbose, the latter will have more tokens and therefore more weight.

Zondartul··on How the Higgs field gives mass to elementary particles
Rule #1 of talking about the aether is "don't call it aether". Nowadays it's "spacetime this" and "mass-energy tensor that" and "properties of vacuum something else".... and we still end up with empty space behaving like a funky fluid.
Zondartul··on Maker Skill Trees
I applaud the effort, but, these templated skill trees are just not very good imo. I have only one issue but it's a big one:

It's missing a tree structure, so there is no ordering of skills (learn easy stuff before hard stuff because there is a learning curve to anything).

They might be good as prints to hang on your wall, but in the current state they're more "achievement lists" rather than anything resembling a tech tree.

Zondartul··on AI solves International Math Olympiad problems at silver medal level
As far as I understand, and I may be wrong here, the system is composed of two networks: Gemini and AlphaZero. Gemini, being an ordinary LLM with some fine-tunes, only does translation from natural to formal language. Then, AlphaZero solves the problem. AlphaZero, unburdened with natural language and only dealing with "playing a game in the proof space" (where the "moves" are commands to the Lean theorem prover), does not hallucinate in the same way an LLM does because it is nothing like an LLM.
Zondartul··on NASA's Curiosity rover discovers a surprise in a Martian rock
It's cool how some minerals are just lying out in the open on Mars. On Earth this would have been washed away or buried under the soil.
Zondartul··on Disney's Internal Slack Breached? NullBulge Leaks 1.1 TiB of Data
With how big and aggressive Disney is I'd expect it to be under ongoing litigation 24/7/365.
Zondartul··on Computational Life: How self-replicating programs emerge from simple interaction
It's convienient because any random collection letters is a "valid" brainfuck program.
Zondartul··on Disney's robots use rockets to stick the landing
I'm now convinced Disney has better engineers than Tesla and SpaceX. Maybe even on par with Boston Dynamics.
Zondartul··on Continuous Social Media Scrolling Negatively Impacts Eye Movement
There is no benefit to ultra-processed fast food, but a lot of people eat it anyway. Social media is just fast food for the brain.
Zondartul··on How Chain-of-Thought Reasoning Helps Neural Networks Compute
Sorry, I meant the information that is inferred (from scratch on every token) from the entire context, and is then reduced to that single token. Every time a token is generated, the LLM looks at the entire context, does some processing (and critically, this step generates new data that is inferred from the context) and then the result of all that processing is reduced to a single token.

My conjecture is that the LLM "knows" some things that it does not put into words. I don't know what it is, but it seems wasteful to drop the entire state on every token. I even suspect that there is something like a "single logic step" of some conclusions from the context. Though I may be committing the fallacy of thinking in symbolic terms of something that is ultimately statistical.

Zondartul··on How Chain-of-Thought Reasoning Helps Neural Networks Compute
The tokens are also necessary to store information, or at least off-load it from neuron activations.

E.g. if you asked an LLM "think about X and then do Y", if the "think X" part is silent, the LLM has a high chance of:

a) just not doing that, or

b) thinking about it but then forgetting, because the capacity of 'RAM' or neuron activations is unknown but probably less than a few tokens.

Actually, has anyone tried to measure how much non-context data (i.e. new data generated from context data) a LLM can keep "in memory" without writing it down?

Zondartul··on Why it's so challenging to land upright on the moon
Some Moon ideas: 1) have a small robot bulldozer flatten a landing pad/polygon so there is loads of safe landing area. 2) use same dozer to make a moon highway to other sites of interest around the moon. 3) Moon GPS or laser-based local navigation beacons, so the spacecraft can rely on those if instruments fail. 4) Just stupid-wide landing legs that let the craft ski around the moon even when landing at an angle and with horizontal velocity.
Zondartul··on Wintergatan Marble Machine (2016) [video]
I won't say anything about the viability of the design #2 vs #3, but from purely entertainment point of view, it was fun and relaxing to watch his regular tinkering videos while he was working on the second machine, but once he stopped, his channel became an emotional rollercoaster. It's just too emotionally draining to watch the later videos involving machine #3, so I stopped.
Zondartul··on OTP at a High Level (2019)
Quoting the article: "OTP stands for _Open Telecom Platform_, which is literally a meaningless name that was used to get the stuff open-sourced back in the old days of Erlang at Ericsson."
Zondartul··on Sora: Creating video from text
I got reminded of an even older sci-fi story: https://qntm.org/responsibility
Zondartul··on Inside the proton, the ‘most complicated thing you could possibly imagine’
A silly thought I had while reading that article: it presupposes that "nothing" is a noun. In doing so, it assumes that in the sentence "the <noun> <verbs>" you can substitute "nothing" and it would mean "<nothing> is an entity that does the <verbing>" instead of "<verbing> simply does not happen", and I feel that is a meaningful distinction.
Page 1 of 3Next →