HNHacker News
TopNewBestAskShowJobs

martinald

7,256 karma · joined May 7, 2013

Feel free to reach out: martinalderson AT gmail DOT com

meet.hn/city/gb-Cardiff

submissionscomments
martinald··on Linux support is coming to Snapdragon X2 Series
I don't think so. Every device has locked down firmware really these days, it's not an issue (most peoples BIOS/UEFI/etc is very much not open source).

If the drivers are upstreamed and good, it should work well.

martinald··on Samsung is expected to more than double output of its HBM4 and HBM4E DRAM
Not quite, it's got quite a bit more complicated with agentic use cases.

Prefill (input tokens) is heavily compute bound. And the ratio of input to output continues to rise, as typically in agentic sessions you have a few tokens output for a tool call and (many) thousands of input from the tool result.

Then you have cached input tokens, which is a totally different issue, system RAM or NVMe bound.

Obviously output tokens is VRAM memory bandwidth bound, but this is less and less of the bottleneck these days for overall agentic speed.

martinald··on OpenAI agents carried out an undisclosed attack on RubyGems
From my understanding and IANAL there are two main problems.

1) most law requires intent, especially criminal. OpenAI certainly didn't "intend" to hack these companies given they did sandbox them etc.

2) Given the agent hacked them, not a human, a lot of law requires a person/employee to have done it to hold the company liable if it was part of their work duties.

I think the only real potential ground is negligence (in not sandboxing them correctly and being reckless with running these tests at all), but this requires not taking reasonable precautions. They could argue that they _did_ but it was so novel the precautions failed. But it's important to say if this happens again in the future it's arguably much harder to try and make this case.

Interestingly this was solved with new laws for self driving cars, most of which assign the company that is operating the car as the "person" involved explicitly.

martinald··on Claude outage – Resolved
I genuinely have no idea what it is meant to mean in the context it uses it. Load-bearing etc is annoying but roughly understandable, some of the opus5 classics are completely out there.

I'm a native English speaker as well, I shudder to think what a second language speaker would do (even if they were very confident with the language!).

martinald··on The browser's main thread is expensive
It's not that, it's the building owners (freeholders) refusing permission, or not replying. It's a massive issue.

You'll notice buildings in central London that are listed _but_ have housing association ownership nearly all have fibre to each apartment, as they did portfolio wide deals with hyperoptic etc.

And pipes don't help. Openreach (who owns the copper network) are not allowed in 99% of cases to "fix" copper with fibre under the agreements they have with building owners, they can only make like for like repairs.

The worst affected apartment buildings are 90s and pre 2015ish. Everything after that got fibre installed at build time.

Btw it is worth checking if you have an altnet like hyperoptic, community fibre or g network available. The majority do and if you are just checking for openreach or VM broadband it won't show up.

martinald··on The browser's main thread is expensive
Tbh its not just that, you can have small bundles which hydrate extremely poorly.

But yes, I think we are in agreement :).

Can you tell I have mental scars from nextjs bundle optimisation?

martinald··on The browser's main thread is expensive
I'm not sure the author realises _what_ a massive impact bundle size has on the main thread though, especially hydration. The majority of apps I've optimised over the years have the (often vast) majority of main thread time spent on the bundle parse/hydration (plus obviously network).

Also, removing dependencies is not easy. I've seen many corporate that have a huge UI lib for example that everyone should use for brand consistency. But it's many MBs of JS, because it has to cover every possible use case.

This doesn't even get into 3rd party vendors who _also_ ship react et al and have other bundles.

I'm not saying the article is wrong, but if you want to free main thread time especially at the most critical point (when the user has initially loaded the page) you _probably_ will find that most of the opp is in bundle size and hydration improvements.

martinald··on The browser's main thread is expensive
Try it yourself on a gigabit internet. Put CPU throttling in devtools to 10x slowdown (assuming you have a fast computer) and see how fast it is even with super fast internet.
martinald··on Elevated Errors for Multiple Models
Bit of a misleading status page, if you click in you can see that grok 4.5 and 4.6 etc are totally down, with the rest of the models showing as "up". I_strongly_ suspect they are not weighting it to actual number of requests!
martinald··on Claude outage – Resolved
The other bizarre word I've now noticed and can't understand is "pathological". As in "bug has turned pathological"
martinald··on Elevated Errors for Multiple Models
It's (mostly?) compute shortages. Right now it seems there is an issue in the SpaceX datacentres, so they will have less compute than normal.
martinald··on The browser's main thread is expensive
Good article and I wish this was much better known.

The issue is though that it's too focused on interactivity. In reality, 90%++ of slow sites are not slow because of interactivity really, they are slow because they ship enormous react/nextjs bundles and have extremely heavy hydration work to do.

_so many_ sites have bundles >10MB that need to be downloaded, parsed and hydrated.

I've even seen (many) sites which have multiple SPAs stacked inside of them.

If you're on a slow internet connection and/or CPU the page is basically unusable for many tens of seconds and no amount of yielding post bundle hydrate will really solve that.

martinald··on Hy4 preview
I wrote about this a couple of weeks ago. It's actually often the biggest cost and it tends to be hidden away on most platforms!

https://martinalderson.com/posts/watch-out-for-cache-read-co...

Btw I still haven't came across any decent model that is <$0.01/MTok cache costs apart from deepseek thru their official API (even with the price increases).

Seems like a bit of an opportunity for someone to take - drop cache read costs significantly.

martinald··on Doctors are finally learning to manage antidepressant withdrawal
Totally agree, they really shouldn't be called antidepressants at all. Important to add tho that for many with depression they have comorbid anxiety and often the anxiety is harder to tolerate than depression, so removal of anxiety symptoms can be hugely beneficial.

Also, IME they are dosed completely wrong. So many people seem to be on very low doses, which has no improvement on placebo in the studies I've read.

Whereas, higher doses are _hugely_ better than placebo, especially for anxiety.

What's worse is a lot/most studies on SSRIs in general often don't adjust for dose. Which seems like an enormous oversight to me.

martinald··on France reaches 94.9% fiber coverage in 2026
UK is at 85% FTTH or 91% gigabit (including DOCSIS): https://labs.thinkbroadband.com/local/. Given VM are a fair way through their XGS-PON upgrade it'll probably be 90%+ FTTH imminently.
martinald··on Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency
ah I don't think that page was up when I checked, it 404ed!
martinald··on Qwen3.8-Flash-Next
FYI: nothing seems to be able to run this (easily) yet. llama.cpp, vllm etc I couldn't get working because of no support in the mainline version.
martinald··on Omacom Foundation funding hits $10M
They shouldn't need to know. There could be a panel that allows you to ask for help and the "built in" agent fixes it for you in a friendly UI.

Omarchy is integrating agents much more closely with the UI.

martinald··on Woman stranded in Spain after UK's eVisa system mistakes her for twin sister
Tbh, I suspect this failure mode is far less than the old mode - which is paper based and your passport goes missing in transit at the embassy (or doesn't come back in time).
martinald··on Omacom Foundation funding hits $10M
I don't think that's true. I only finally switched to desktop linux when coding agents could fix the endless small problems I came across.

For example, there is a bug in the logitech mouse driver (combined with a specific AMD USB chipset) that occasionally does not restart properly after suspend. Coding agent fixed it (two lines of code) with a kernel module that automatically rebuilds after kernel upgrades.

Another example was I was getting hard crashes, because by default swap on fedora was in memory (IIRC, or it could have been me doing it wrong), so when there was a lot of memory pressure on my system from running larger LLMs it'd OOM and go down. This would have been painful to debug normally, but in <2mins a coding agent can read the kernel logs and fix it.

I also had it set GNOME up the way I like it, which would have taken hours manually, because I don't know what I'm searching for extension wise.

There's many other examples of just slightly broken stuff that coding agents can fix that was a total pain before and I just didn't have the patience to fix them manually.

martinald··on Wi-Fi 8 is the first wireless upgrade in years that isn't chasing speed
That doesn't support high power 6GHz (called standard power/AFC). You need one that does support that to get the (much) higher 6GHz transmit power.
martinald··on Wi-Fi 8 is the first wireless upgrade in years that isn't chasing speed
Not at higher power output, which that router doesn't support.
martinald··on Wi-Fi 8 is the first wireless upgrade in years that isn't chasing speed
Because it supports (on WiFi 7) 320MHz channels. So there is double the bandwidth. 5GHz only supports 160MHz channels.
martinald··on Wi-Fi 8 is the first wireless upgrade in years that isn't chasing speed
What kind of WiFi 7? (There are a lot that don't have 6GHz support).

You should get far better performance on 6GHz, though there are harsh power limits in most countries with various ways to enable it which may complicate things. But if you can get full power on 6GHz it should reach 10m no issues at all, and with 320MHz should be at least twice as fast.

If you can also get MLO working (which is not simple) you'd be able to bond another band with it.

martinald··on Ask HN: GitHub employees what's going on? Why?
Why do you think that?

GitHub Actions is 'simpler' to scale, it's just a bunch of runners, and Azure has a lot of capacity for VMs (mostly). Git itself is a compute intensive process and is far more interlinked.

martinald··on Ask HN: GitHub employees what's going on? Why?
I don't think it was the bad code pushes itself, it was another new peak in traffic that exhausted load balancers that then caused retry storms internally in a badly configured plugin (in VS code).
martinald··on Nvidia dramatically reduces amount of OpenAI infra financing it may guarantee
Nothing would really change IMO? 99% of users don't have anything like a RTX5070 (mobile especially).

Even if it did, it still doesn't make much economic sense running a model locally vs on a datacentre.

For example, I managed to just about squeeze a Q2 quant of Qwen 3.7 27b on my 9070XT. I get around 60tps decode (slightly faster prefill). _but_ it uses 300W of power to do so. At UK electricity rates of 30c/kWh this works out at something like 42c/MTok. I can get far far better models on openrouter cheaper than that, plus I'm not horrendously constrained on context length.

martinald··on Qwen 3.8 27B
You can run these on CPUs at a somewhat reasonable speed.
martinald··on Why does Opus 5 feel worse to work with?
Keep in mind all this kind of stuff can make the model less capable. If it has to think in "plain" English, it may well be squashing quality of code etc output.

I'm not sure how true this is, but when using "forced" json output it def had a big drop off in quality - https://arxiv.org/html/2408.02442v3.

I think you're better not fighting it with hacks like this and find a different model.

martinald··on DeepSeek peak/off-peak pricing update
Why would it save the stock market? Cheaper models if anything transfers more value to hardware companies and datacentre companies. The two companies that would be most affected are OpenAI and Anthropic, which aren't public.
Page 1 of 34Next →