416 karma · joined August 8, 2022
Systems @ modal.com
(disclaimer: I work with the author)
Heard on everything else, thank you!
I'm not worried about this at Modal, but I am worried about this in the greater OSS community. How can I reasonably trust that the tools I'm using are built in a sound manner, when the barrier to producing good-looking bad code is so low
On the other hand, the other responsibilities of being an engineer have become quite a bit less appealing.
If not, we could see that LLMs of tomorrow struggle to keep up with the bloat of today. The "interest on tech debt" is a huge unknown metric w.r.t. agents.
I find that the flaws of agentic workflows tend to be in the vein of "repeating past mistakes", looking at previous debt-riddled files and making an equivalently debt-riddled refactor, despite it looking better on the surface. A tunnel-vision problem of sorts
IMO we shouldn't strive to make an entire codebase pristine, but building anything on shaky foundations is a recipe for disaster.
Perhaps the frontier models of 2026H2 may be good enough to start compacting and cleaning up entire codebases, but with the trajectory of how frontier labs suggest workflows for coding agents, combined with increasing context window capabilities, I don't see this being a priority or a design goal.
The tragedy, for me, is that the bar has been lowered. What I consider to be "good enough" has gone down simply because I'm not the one writing the code itself, and feel less attachment to it, as it were.
Kata runs atop many things, but is a little awkward because it creates a "pod" (VM) inside which it creates 1+ containers (runc/gVisor). Firecracker is also awkward because GPU support is pretty hard / impossible.
I think it's worth including these things in a future update to the post, but I didn't have the time / need to explore it back then.
In the meantime, I'd point to the following post on Unicode that remains very nice to read >20 years later: https://www.joelonsoftware.com/2003/10/08/the-absolute-minim...
I wrote this post mostly out of interest for a personal project and thus it's not actually a very holistic exploration of the topic. May revisit and update it in the future :)
Just sprung for one at a good price.
Related: In notation, one thing that I used to struggle with is how addresses (e.g. 0xAB_CD) actually have the bit representation of [0xCD, 0xAB]. Wonder if there's a common way to address that?
More generally, I'm not surprised at the symtab bloat from statically-linking given the absolute size increase of the binary.
General idea was just to highlight some of the dangers of vector registers. I believe the same is true of ymm (256) to a lesser extent.
https://stackoverflow.com/questions/56852812/simd-instructio...
It might be more concise, e.g., in the Newtype case, but the sacrifice seems to be quite a bit more cognitive complexity, which I would personally value over conciseness
Apple seems to be gearing up for significant advances in on-device inference using this LLMs
> I decided to look for open Developer positions, to work with a team of experienced developers so I can learn even faster.
Seeing as you're in Europe and not in the U.S. (with exorbitant tuition costs), I would actually recommend against skipping out on College. I can confidently say that while experience in the industry is great (having worked part-time at a startup, as well as various internships during College), I think the knowledge gained from University classes is often underrated.
Not only that, but College is definitely a life experience that I would recommend not to skip lightly.
Regardless, this is super impressive work!