HNHacker News
TopNewBestAskShowJobs

benob

425 karma · joined December 6, 2017

submissionscomments
benob··on Relicensing with AI-Assisted Rewrite
I don't think this would qualify as clean room (the Library was involved in learning to generate programs as a whole). However, it should be possible to remove the library from the OLMO training data and retrain it from scratch.

But what about training without having seen any human written program? Coul a model learn from randomly generated programs?

benob··on Relicensing with AI-Assisted Rewrite
What about doing that with movies and music?
benob··on I built a demo of what AI chat will look like when it's “free” and ad-supported
Also, don't forget to install our local ad blocking LLM. Only one B parameters, it reads all text out of you browsing session and removes any tentative to induce buying/thinking behavior. Best in town...
benob··on Smallest transformer that can add two 10-digit numbers
I had that in mind too. What if you handcraft a subnetwork with (some subset of) Turing machine capability? Do those kinds of circuits emerge naturally during training? Can reasoning use them for complex computation?
benob··on What Claude Code chooses
What is the need for dependencies when you can code them from scratch?
benob··on Mark Zuckerberg to testify in landmark social media trial
Same with asbestos, I mean what could go wrong in 20 years?
benob··on ai;dr
You could totally make a believable timing generation model from a few (hundreds) recordings of human writing. Detecting AI is hard...
benob··on Claude’s C Compiler vs. GCC
Give me self hosting: LLM generates compiler which compiles LLM training and inference suite, which then generates compiler which...
benob··on Claude’s C Compiler vs. GCC
Why don't LLMs directly generate machine code?
benob··on Beyond agentic coding
I really like the "file lens" example:

> “Focus on…” would allow the user to specify what they're interested in changing and present only files and lines of code related to their specified interest.

> “Edit as…” would allow the user to edit the file or selected code as if it were a different programming language or file format.

benob··on Show HN: LocalGPT – A local-first AI assistant in Rust with persistent memory
What local models shine as local assistants? Is there an effort to evaluate the compromise between compute/memory and local models that can support this use case? What kind of hardware do you need to not feel like playing with a useless shiny toy?
benob··on France dumps Zoom and Teams as Europe seeks digital autonomy from the US
Too bad it doesn't have zooms echo cancelation / isolation
benob··on GPTZero finds 100 new hallucinations in NeurIPS 2025 accepted papers
Maybe it will also change the whole publication as evaluation of science.
benob··on Nanolang: A tiny experimental language designed to be targeted by coding LLMs
An LLM targeting language and no token efficiency?
benob··on Our approach to advertising
It was over when they named the company Open AI
benob··on A linear-time alternative for Dimensionality Reduction and fast visualisation
Is there a pip installable version?
benob··on Show HN: Gemini Pro 3 imagines the HN front page 10 years from now
What, no reference to quantum or crypto?
benob··on Show HN: Gemini Pro 3 imagines the HN front page 10 years from now
Will there ever be a llama12? Is it going to go the yolo route?
benob··on Shai-Hulud Returns: Over 300 NPM Packages Infected
How was the attack detected in the first place?
benob··on Ask HN: How are Markov chains so different from tiny LLMs?
Was it this one? https://infini-gram.io/
benob··on Gemini 3 Pro Model Card [pdf]
What does it mean nowadays to start from scratch? At least in the open scene, most of the post-training data is generated by other LLMs.
benob··on Supercookie: Browser Fingerprinting via Favicon (2021)
If caching is bounded in time, can't you use other fingerprinting methods to seal the gaps?
benob··on Supercookie: Browser Fingerprinting via Favicon (2021)
Why doesn't this apply to any kind of cached content?
benob··on Android developer verification: Early access starts
Google's move is very good for the web. By pushing app makers away from walled platforms, you turn them to standardized, open ones such as the web.
benob··on Omnilingual ASR: Advancing automatic speech recognition for 1600 languages
> Bring Your Own Language

Few-shot new languages is going to be a game changer for linguists

benob··on Omnilingual ASR: Advancing automatic speech recognition for 1600 languages
Not tested on that particular model, but the idea has been flying around for some time: https://arxiv.org/abs/2509.04166v1
benob··on Radiant Computer
It comes with a basic interpreter, but the thriving community develops plenty of stuff: python, lua, forth, lisp... have nice ports which you can play with on device. There also is a library of software developed for rp2040 which have been ported, such as a mac classic emulator. If you feel that a micro controller is too low power, you can plug in a luckfox lyra which runs a proper linux with 128M of RAM.
benob··on Radiant Computer
I have been having a lot of fun with PicoCalc. It's not targeted at end users but is fun for developers alike who want a taste of developing things from first principles. More than anything it can live independently from your other devices.
benob··on Rouille – Rust Programming, in French
This is very lenient French: "fetchez le dico"
benob··on Neural audio codecs: how to get audio into LLMs
Audio tokenization consumes at least 4x tokens versus text. So there is an efficiency problem to start with. Then is there enough audio data to train a LLM from scratch?
← PreviousPage 2 of 7Next →