HNHacker News
TopNewBestAskShowJobs

JoelEinbinder

404 karma · joined July 16, 2021

submissionscomments
JoelEinbinder··on GPT-5.5
My understanding is that existing rail lines aren't flat/straight enough for high speed rail. There's no point to a bullet train if it has to constantly slow down for corners/hills.
JoelEinbinder··on Show HN: Lightpanda, an open-source headless browser in Zig
The benchmark shows lower ram usage on a very simple demo website. I expect that if the benchmark ran on a random set of real websites, ram usage would not be meaningfully lower than Chrome. Happy to be impressed and wrong if it remains lower.
JoelEinbinder··on Show HN: Lightpanda, an open-source headless browser in Zig
When I've talked to people running this kind of ai scraping/agent workflow, the costs of the AI parts dwarf that of the web browser parts. This causes computational cost of the browser to become irrelevant. I'm curious what situation you got yourself in where optimizing the browser results in meaningful savings. I'd also like to be in that place!

I think your ram usage benchmark is deceptive. I'd expect a minimal browser to have much lower peak memory usage than chrome on a minimal website. But it should even out or get worse as the websites get richer. The nature of web scraping is that the worst sites take up the vast majority of your cpu cycles. I don't think lowering the ram usage of the browser process will have much real world impact.

JoelEinbinder··on We're forking Flutter
Google's monorepo of closed source code.
JoelEinbinder··on Waymo and Hyundai enter multi-year, strategic partnership
Seems like Hyundai own 33% of Kia, rather than it just being a brand under the same company like Lexus/Toyota. They share some things and compete on others.
JoelEinbinder··on Using the Infinite Bookspace to Reason About Language Models
I think Hacker News might appreciate some of the behind the scenes of this post.

Getting this page to load quickly was not trivial. The initial dataset of books starting sentences was over 20 megabytes. By only sending the unique prefix of each book, I was able to get that to be much smaller. Using a custom format, sorting the prefixes, and gzipping got the size down to 114kb. About 3 bytes per book. The full first sentences are downloaded on demand as the books are filtered down.

Rendering the books requires 5 million triangles. I used WebGL 2's drawArraysInstanced method. This allows me to define the book geometry only once, and each book is just defined by it's rotation/position/color. Then it's just a matter of keeping the fragment shader simple.

Going into this project, I wasn't sure if it was possible. But I have left feeling really impressed with how capable the web is these days if you are willing to push a bit.

JoelEinbinder··on Ask HN: Would competitive chess be better with one allowed take-back per turn?
It would make competitive chess even more draw-ish. It is much easier to see when you accidentally get into a losing position than when you miss a winning idea. So the take back would be used defensively.
JoelEinbinder··on Are you better than a language model at predicting the next word?
On the full set of 1000 questions, the language models are getting 30-35% correct. With patience, humans can do 40-50%.

The language models were prompted with the text + each candidate answer, and the one with the lowest perplexity was picked. I tried to avoid instruction tuned models wherever possible to avoid the "voice" problem.

JoelEinbinder··on Are you better than a language model at predicting the next word?
What scores are you getting using this technique?
JoelEinbinder··on Are you better than a language model at predicting the next word?
After the quiz, the source is linked along with the full comment.
JoelEinbinder··on Are you better than a language model at predicting the next word?
The prompts you see in the quiz are from real hacker news comments. Whatever word the commenter said next is the "correct" word.
JoelEinbinder··on Are you better than a language model at predicting the next word?
Temperature doesn't play a role here, because the LLM is not being sampled (other than to generate the candidate answers). Instead the answer the llm picks is decided by computing the complexity for the full prompt + answer string.
JoelEinbinder··on Are you better than a language model at predicting the next word?
If I used old comments then it's likely that the models will have trained on them. I haven't tested if that makes a difference though.
JoelEinbinder··on Are you better than a language model at predicting the next word?
The language model generating the candidate answers generates tokens until a full word is produced. The language models picking their answer choose the completion that results in the lowest perplexity independent of the tokenization.
JoelEinbinder··on Are you better than a language model at predicting the next word?
The LLM didn’t generate the next word. Hacker News commenters did. You can see the source of the comment on the results screen.
JoelEinbinder··on Are you better than a language model at predicting the next word?
If you want to practice it one question at at time, you set the question count to 1. https://joel.tools/smarter/?questions=1

When I tested it this way it resulted in less of an emotional reaction.

JoelEinbinder··on Are you better than a language model at predicting the next word?
That isn't how it's supposed to work. I mean sometimes you get a supper annoying prompt like ">", but if you guess the right answer it should give you the point. I just checked the two prompts like that, and they seem to work for me.
JoelEinbinder··on Are you better than a language model at predicting the next word?
I made a little game/quiz where you try to guess the next word in a bunch of Hacker News comments and compete against various language models. I used llama2 to generate three alternative completions for each comment creating a multiple choice question. For the local language models that you are competing against, I consider them having picked the answer with the lowest total perplexity of prompt + answer. I am able to replicate this behavior with the OpenAI models by setting a logit_bias that limits the llm to pick only one of the allowed answers. I tried just giving the full multiple choice question as a prompt and having it pick an answer, but that led to really poor results. So I'm not able to compare with Claude or any online LLMs that don't have logit_bias.

I wouldn't call the quiz fun exactly. After playing with it a lot I think I've been able to consistently get above 50% of questions right. I have slowed down a lot answering each question, which I think LLMs have trouble doing.

JoelEinbinder··on Puppeteer Support for Firefox
You can set `pipe` to true in puppeteer (default false) here https://pptr.dev/api/puppeteer.launchoptions

By default, Playwright launches this way and you have to specifically enable the tcp listening.

JoelEinbinder··on Prevention of HIV
I'm not going to invest a drug company with a 90% chance of failure unless I can expect to get a 10x return if it succeeds.
JoelEinbinder··on Implementing Vertical Form Controls
Interesting to me that WebKit gets vertical form controls before MacOS. I don't have any experience with vertically written languages. How important are these controls to computer usage in Japan? Does Windows have them?
JoelEinbinder··on JSR: The JavaScript Registry
While it's not quite C++, here is Chromium's implementation: https://source.chromium.org/chromium/chromium/src/+/main:v8/...
JoelEinbinder··on WebKit switching to Skia for 2d graphics rendering
It's been a few years since I was looking into it, so this might be out of date. But WPE is targeted at kiosks and places you might want to display web content but not have a full web browser. Igalia can sell consulting services to these companies, as opposed to webkitgtk which has a small number of non-paying users. So WPE serves as a place for more active development of webkit-on-linux while not breaking webkitgtk which powers the web browser "Web" on gnome. Things from WPE tend to slowly make it into the webkitgtk build eventually. It's all maintained by the same people.

Looking at https://webkit.org/wpe/, the first design goal is the one that justifies WPE vs other webkit ports: "To provide a no-frills, straight to the point, web runtime for embedded devices."

The other goals, like standards compliant and hardware acceleration, are there to differentiate WPE from non-webkit and ancient-webkit browser engines that people might use on embedded devices.

JoelEinbinder··on Ask HN: What stocks / indices does HackerNews invest in?
Assuming you are working in tech, I'd avoid investing in tech. You don't want a situation where tech does badly and you are both out of a job and out of your savings.
JoelEinbinder··on Queues don't fix overload (2014)
Real life queues can be scary too. I think of how complicated Disney's fast pass system got https://www.youtube.com/watch?v=9yjZpBq1XBE. Luckily with software it is way easier to get more servers than it is to build more theme park rides.
JoelEinbinder··on Linear transformers are faster after all
Developer tools point to MathJax https://www.mathjax.org/. If you disable javascript you can see some LaTex.
JoelEinbinder··on Show HN: Halloween game to show off my new Terminal
I tried my best to. What is your user agent?
JoelEinbinder··on Show HN: Halloween game to show off my new Terminal
I made it so that it short circuits the frame creation of if there aren't going to be any files to show in that directory.

I'm honestly not sure if the web sandbox I made has enough realism to solve that puzzle though programming. There is a feature of snail that I was using to get past it.

JoelEinbinder··on Show HN: Halloween game to show off my new Terminal
Well I fixed it for the next person at least. It is still a bit laggier than it should be but it will no longer crash your browser. Thanks for playing!
JoelEinbinder··on Show HN: Halloween game to show off my new Terminal
:O yeah I should fix that. That is going to create 6000 <iframe> elements. I don't think your web browser would be happy.
Page 1 of 2Next →