HNHacker News
TopNewBestAskShowJobs

juliangoldsmith

1,063 karma · joined December 6, 2013

A 33-year-old developer.
submissionscomments
juliangoldsmith··on People who grew up with high economic connectedness earn more
Impulse control issues, like ADHD, can factor in too.
juliangoldsmith··on People who grew up with high economic connectedness earn more
A possible "third variable" could be intelligence. It's been shown to be strongly correlated with success in general, and it seems likely people of similar intelligence would find it easier to relate to each other.
juliangoldsmith··on Ornith-1.0: Self-scaffolding LLMs for agentic coding
The claim isn't so wild when it's a generalist versus a finetune trained specifically on the tasks being benchmarked.
juliangoldsmith··on Ornith-1.0: self-improving open-source models for agentic coding
That benchmark ranks Kimi K2.6 and K2.7 Code near the bottom. Both are below Ornith 35B. It ranks Gemma 4 26B much higher than GLM-5.2. The results don't make much sense.
juliangoldsmith··on Ornith-1.0: Self-scaffolding LLMs for agentic coding
It looks like they're comparing Orinth 9B to Qwen 3.5 35B, not Qwen 3.6. I guess it kind of makes sense since it's a finetune of 3.5, but I totally missed until I looked closely.

In my brief tests, Ornith 35B performed quite well. It won't replace DeepSeek V4 Flash for me, but if it was fast and cheap enough it might.

I don't remember being super impressed with Ornith 9B, but I could see it being on par with Qwen 3.5 35B.

juliangoldsmith··on Qwen-AgentWorld: Language World Models for General Agents
It should improve agents' action selection by allowing them to evaluate actions' effects before performing them.

An agent using only a regular LLM has no real way to predict the results of its actions. It has to just take an action based on its training data and hope it's the right one. With a world model like this, it could do a second pass before each action to catch mistakes.

I don't know if this actually delivers yet, but if it does it might help make agents more usable.

juliangoldsmith··on What is Z-Angle Memory and why is Intel developing it?
One can only imagine how much money Intel would have made from Optane during the ongoing RAM shortages. It would be absolutely perfect for warm KV cache, and potentially good for MoE expert offloading.
juliangoldsmith··on Your phone is about to stop being yours
No true Scotsman would ever use binary blobs.
juliangoldsmith··on Parental controls aren't for parents
Blocking by age rating takes out the majority of the classic Disney movies and shows. They only consider the newer CGI stuff "child-friendly".
juliangoldsmith··on Rust--: Rust without the borrow checker
The silenced errors aren't guaranteed to be memory leaks or use after frees. There are some situations where memory is being handled properly, but the borrow checker isn't able to prove it.

One example might be a tree-like struct where a parent and child have references to each other. Even if everything is cleaned up properly, the borrow checker has no way to know that when the struct is created. Solving it requires unsafe at some point, usually through something like RefCell.

juliangoldsmith··on Why We Abandoned Matrix (2024)
It sounds like you were stuck between a rock and a hard place there. Hope the Rust integration goes well.
juliangoldsmith··on Why We Abandoned Matrix (2024)
>trading off for speed

If speed is a concern, why did you all stick with Synapse (essentially single-threaded due to the GIL) over moving to Dendrite? As far as I can tell, Dendrite is, for all intents and purposes, abandoned.

juliangoldsmith··on Why We Abandoned Matrix (2024)
It doesn't appear to be open source, so users have no control or lasting guarantees of privacy.
juliangoldsmith··on Windows GUI – Good, Bad and Pretty Ugly (2023)
As someone who as attempted to use React Native for Windows, I can tell you that the "native" XAML doesn't make things any better. If it was using web technologies I wouldn't need to manually modify RNSVG to fix segfaults when an SVG goes offscreen.
juliangoldsmith··on Strudel REPL – a music live coding environment living in the browser
"Hard Refresh" and "Airglow" made it onto my "On Repeat" playlist almost immediately.
juliangoldsmith··on 25L Portable NV-linked Dual 3090 LLM Rig
I'd use caution with the Mi50s. I bought a 16GB one on eBay a while back and it's been completely unusable.

It seems to be a Radeon VII on an Mi50 board, which should technically work. It immediately hangs the first time an OpenCL kernel is run, and doesn't come back up until I reboot. It's possible my issues are due to Mesa or driver config, but I'd strongly recommend buying one to test before going all in.

There are a lot of cheap SXM2 V100s and adapter boards out now, which should perform very well. The adapters unfortunately weren't available when I bought my hardware, or I would have scooped up several.

juliangoldsmith··on The LLM Lobotomy?
I've been using Azure AI Foundry for an ongoing project, and have been extremely dissatisfied.

The first issue I ran into was with them not supporting LLaMA for tool calls. Microsoft stated in February that they were working on it [0], and they were just closing the ticket because they were tracking it internally. I'm not sure why they've been unable to do what took me two hours in over six months, but I am sure they wouldn't be upset by me using the much more expensive OpenAI models.

There are also consistent performance issues, even on small models, as mentioned elsewhere. This is with a rate on the order of one per minute. You can solve that with provisioned throughput units. The cheapest option is one of the GPT models, at a minimum of $10k/month (a bit under half the cost of just renting an A100 server). DeepSeek was a minimum of around $72k/month. I don't remember there being any other non-OpenAI models with a provisioned option.

Given that current usage without provisioning is approximately in the single dollars per month, I have some doubts as to whether we'd be getting our money's worth having to provision capacity.

juliangoldsmith··on An overview of gradient descent optimization algorithms (2016)
What is it that makes higher order derivatives less useful at high dimensionality? Is it related to the Curse of Dimensionality, or maybe something like exploding gradients at higher orders?
juliangoldsmith··on ROCm Device Support Wishlist
It works out of the box without jumping through any hoops, and the fact that it has an OpenCL backend means it can run on a wide variety of hardware.

I don't know of any other autograd libraries with a non-CUDA backend, but I'd be interested to learn about them.

juliangoldsmith··on Build a tiny CA for your homelab with a Raspberry Pi
If we were dealing with pure cosmic background radiation, or inside a Faraday cage, sure.

What I'm referring to are things like radio broadcasts, 60 Hz hum from power lines, noise put out by switching power supplies, and that sort of thing.

Just having a bias, as in your example, would be still truly random. If you knew that every tenth roll you'd get a 3, it would no longer be random. When your random number generator can be influenced by the outside world, it's no longer suitable for cryptographic use.

juliangoldsmith··on ROCm Device Support Wishlist
How does Tinygrad fall short? Performance is fine [0]. It's much smaller than Pytorch and all, but that's kind of in the name.

I've been hearing about MLIR and OpenXLA for years through Tensorflow, but I've never seen an actual application using them. What out there makes use of them? I'd originally hoped it'd allow Tensorflow to support alternate backends, but that doesn't seem to be the case.

0: https://cprimozic.net/notes/posts/machine-learning-benchmark...

juliangoldsmith··on ROCm Device Support Wishlist
AMD's hardware might be compelling if it had good software support, but it doesn't. CUDA regularly breaks when I try to use Tensorflow on NVIDIA hardware already. Running a poorly-implemented clone of CUDA where even getting Pytorch running is a small miracle is going to be a hard sell.

All AMD had to do was support open standards. They could have added OpenCL/SYCL/Vulkan Compute backends to Tensorflow and Pytorch and covered 80% of ML use cases. Instead of differentiating themselves with actual working software, they decided to become an inferior copy of NVIDIA.

I recently switched from Tensorflow to Tinygrad for personal projects and haven't looked back. The performance is similar to Tensorflow with JIT [0]. The difference is that instead of spending 5 hours fixing things when NVIDIA's proprietary kernel modules update or I need a new box, it actually Just Works when I do "pip install tinygrad".

0: https://cprimozic.net/notes/posts/machine-learning-benchmark...

juliangoldsmith··on Build a tiny CA for your homelab with a Raspberry Pi
>How does the disconnected audio input of any random PC or thinclient compare?

That will give you RF noise, which isn't really random.

juliangoldsmith··on Chinese Innovations Spawn Wave of Toll Phishing via SMS
Most debit cards in the US can be run either as debit or credit. Debit transactions require a PIN, but credit transactions don't.
juliangoldsmith··on A new learning experience on MDN
The search engine they linked to happens to provide a significant portion of Mozilla's revenue.
juliangoldsmith··on France's most powerful nuclear reactor connected to grid after 17-year build
That Wikipedia page explicitly states that the public has never subsidized damage from a nuclear accident. Nuclear energy companies are required to have $450M in private insurance for each reactor. For amounts over that, the Price-Anderson Act requires all nuclear energy companies to pay up to $121M per reactor, for a total of $12B in coverage. The public would potentially cover anything after that $12B, but that has never happened.

If every nuclear reactor in the US simultaneously had an accident requiring the $70M paid out for Three-Mile Island, we'd be around 1.2% of the way to needing Treasury funds. Three-Mile Island's operator was responsible for cleaning it up, and they paid the entire $1B required to do so.

juliangoldsmith··on KlongPy: High-Performance Array Programming in Python
Rank in the array language terminology seems to be synonymous with tensor order (sometimes called tensor rank). Ken Iverson (APL's creator) was a mathematician, so I'm not that surprised to learn the term came from a branch of math.
juliangoldsmith··on KlongPy: High-Performance Array Programming in Python
The array languages aren't super popular because of the sharp learning curve. They're a lot different than most other languages, and they have a lot of operators that simply don't exist in something like C++.

A few years ago there was an article about K's use in high-frequency trading. I'm not sure about usage of APL and J, though. BQN is still fairly new, so it will take a while to see much production usage.

If you've ever written code using NumPy, Tensorflow, or PyTorch, you're doing array programming. Those libraries are heavily influenced by the array languages, including taking a lot of terminology (rank, etc.). I've personally found that playing with J and BQN helped me understand Tensorflow, and vice versa.

juliangoldsmith··on Toddler's backyard snake bite bills totaled more than a quarter-million dollars
The title is somewhat misleading, since it implies the family paid that much. From near the bottom of the article:

>Brigland’s family paid $7,200, their plan’s out-of-pocket maximum.

juliangoldsmith··on Cramming Solitaire onto a Nintendo E-Reader card
Dupe: https://news.ycombinator.com/item?id=42010136
Page 1 of 17Next →