HNHacker News
TopNewBestAskShowJobs

yu3zhou4

1,438 karma · joined May 19, 2022

Hello to you, person, AI, bot or perhaps another kind of being I haven’t identify yet. Here are some links to follow if you have some more time to spend

https://jedrzej.maczan.pl

https://github.com/jmaczan

https://x.com/jedmaczan

https://scholar.google.com/citations?user=VC4YmSAAAAAJ

jedrzej@maczan.pl

Wishing you a good day (or whatever timescale you operate on!)

submissionscomments
yu3zhou4··on Kolibri – Tech Report [pdf]
Much thanks for the explanation and for your help Daniel!
yu3zhou4··on Kolibri – Tech Report [pdf]
It’s not a dupe, it’s a full training and infra report
yu3zhou4··on "As a Language Model": Chat Template Switches LLM Self-Referential Voice
As far as I know we don't know much about metacognition in LLMs, though? Not sure
yu3zhou4··on "As a Language Model": Chat Template Switches LLM Self-Referential Voice
Same! I believe that you could actually train a LoRA on top of a model to get results close to that
yu3zhou4··on "As a Language Model": Chat Template Switches LLM Self-Referential Voice
The strange thing is that the base models (before RLHF) use the "experiential" voice, even though they are not incentivized to do that.
yu3zhou4··on "As a Language Model": Chat Template Switches LLM Self-Referential Voice
Thanks for pointing out, maybe I should be more explicit in the wording - I mean we don't fully know what drives the voice in LLMs. Models that are post trained as instruct models are expected to have the disclaimers, but what about base models (those that are trained on just a lot of text)? How do they talk about themselves? What happens when you strip off the chat template from instruct model's prompt? I hope the rest of the paper makes the questions clearer, but I will try to do better in the abstract next time, as you point out this sentence is kind ambiguous. Thank you!
yu3zhou4··on Best LLM for every budget, updated daily
I’d love this but for every hardware
yu3zhou4··on Show HN: Drop – a rootless Linux sandbox with gVisor support
Gratulacje Jan! Looks like something critical to gain adoption these days, security-wise. For others who also wonder how it works, I find this docs page a bit more informative than the landing page https://droprun.sh/docs/sandbox-overview/
yu3zhou4··on The Secret Life of Circuits
Michał Zalewski (the author of this book) is goated. His book Silence on the Wire was very inspiring to me and the techniques he described often come to my mind years after reading the book
yu3zhou4··on Ask HN: What are you working on? (September 2026)
I study stats and probability
yu3zhou4··on LLMs and Self-Referentiality
I found that in LLMs the self-reference defined as "referring in generated text to itself" (so a bit different than in Scott's blog) comes largely from the chat template

https://openreview.net/pdf?id=3O2A23MhNi

yu3zhou4··on PageRank explained
Also PageRank as a Markov chain

[0] https://math.libretexts.org/Bookshelves/Linear_Algebra/Under...

yu3zhou4··on Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025)
I’m happy it helps you!
yu3zhou4··on Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025)
If you prefer C++ and CUDA, then there's also tiny-vllm of mine [0] - recently we broke 1k gh stars

[0] https://github.com/jmaczan/tiny-vllm

yu3zhou4··on Paged Out #9 [pdf]
My fav zine
yu3zhou4··on Learn WebGPU for C++
It's a very useful guide, I used it when working on torch-webgpu some time ago. Elie also published few packages that help with using some WebGPU related things, I don't remember exactly what was that but I recall it saved me lots of time

https://github.com/jmaczan/torch-webgpu

yu3zhou4··on Exapunks (2018)
Print ondemand can be ordered for less than $5 per zine in Lulu, since 2022 or so. I guess Zach doesn’t make a cut from it since this $5 probably barely covers the cost of print and Lulu fee :(
yu3zhou4··on Advanced Compilers: The Self-Guided Online Course
PyTorch?
yu3zhou4··on Show HN: Tiny-vLLM – high performance LLM inference engine in C++ and CUDA
README is in my opinion (author here) the most interesting - I wrote it to help others build useful mental model to be able to recreate the project yourself, without need to even read my code
yu3zhou4··on The vi family
There was onivim that was a bit hyped a few years ago but unfortunately it died
yu3zhou4··on Ask HN: What are you working on? (May 2026)
Just learning math and trying myself in ml research
yu3zhou4··on Poland is now among the 20 largest economies
Laughed hard about Collegium Tumanum
yu3zhou4··on Poland is now among the 20 largest economies
Poland was sort of occupied until 1989
yu3zhou4··on HERMES.md in commit messages causes requests to route to extra usage billing
Good idea, too! Why do you see my explanation as cynical, though?
yu3zhou4··on HERMES.md in commit messages causes requests to route to extra usage billing
Maybe it’s in order to have an external provider to blame for failures and shift the blame/responsibility?
yu3zhou4··on TorchTPU: Running PyTorch Natively on TPUs at Google Scale
Adding a support for new hardware to PyTorch is actually quite convenient. I did that with WebGPU using the same PrivateUse1 mechanism TorchTPU used. Every hardware has its own slot and identifier, and when you want to add a support for a new one without merging it into PyTorch, PrivateUse1 works essentially like plug-in slot

https://github.com/jmaczan/torch-webgpu

yu3zhou4··on TorchTPU: Running PyTorch Natively on TPUs at Google Scale
They write that they use PrivateUse1, so it’s a custom out-of-tree backend
yu3zhou4··on All phones sold in the EU to have replaceable batteries from 2027
The system like iOS is still closed. Once Apple ends support for a device, being able to swap a battery won’t help much
yu3zhou4··on Guide.world: A compendium of travel guides
Thx, great read about Yemen from Maciej Cegłowski https://idlewords.com/2014/07/sana_a.htm
yu3zhou4··on Ask HN: What Are You Working On? (April 2026)
An open course on building high performance LLM inference engine! Hope to finish by the end of April

https://github.com/jmaczan/tiny-vllm

Page 1 of 9Next →