HNHacker News
TopNewBestAskShowJobs

nessex

156 karma · joined January 24, 2018

submissionscomments
nessex··on Necker, 1832: "An optical phænomenon on viewing a figure of a geometrical solid"
(Old) English had æ first!
nessex··on Open-source AI and open models reading list
Many of the Nemotron datasets are gated behind approval, and a license agreement.

The preamble on these datasets is: "This repository is publicly accessible, but you have to accept the conditions to access its files and contents."

I don't know what others have experienced, but I requested access to multiple Nemotron datasets and those requests were ignored for months before all but one request was rejected. There's no explanation for why, nor anything I can see which would lead to a rejection. So it's purely anecdotal and YMMV, but I don't see these as being particularly open.

nessex··on “We have information that Moonshot distilled Fable for the development of K3”
Paying the list price Anthropic decided upon for outputs from an LLM, then using those outputs for your work? Or sourcing those outputs from others that paid for the service, and using those? If that's "distillation", it seems fine to me. And a lot like what most are doing with these same providers.

Anthropic: training AI is "transformative", it's not copyright infringement if we don't re-transmit the copyrighted books we trained on

Anthropic: training AI models on our outputs is stealing our secret sauce, outputs that could only be produced by us

Anthropic: AI model outputs are unreliable and do not reflect Anthropic's views, we are not liable if they harm you

(not exact quotes, they're "distilled")

So Anthropic "distills" knowledge of others with reckless abandon, packages it up, sells it to you, claims it's your fault if anything bad happens but then also lobbies to treat you as a criminal if the outputs you paid for end up being transformed into any sort of competition for them. By you, or others that use the outputs you paid for.

nessex··on There is minimal downside to switching to open models
Forgot to mention, but it was after the 4.7 release when I was still using 4.6 that I saw those loops too... Before that, 4.6 had been a pretty seamless experience.
nessex··on There is minimal downside to switching to open models
They would quantize the model. That'd make it cheaper to run, and have slightly worse output but it would still generate outputs with a similar feel, derived from a compressed version of the same knowledge base etc.

They wouldn't even need to do this uniformly, quantized versions of the model could be routed only a subset of the requests. They could do this to nerf the old model, or more likely just to give themselves more hardware to run the new one on by handling more requests on less hardware. Or to handle increased request volume as traffic ramps up faster than hardware can be provisioned.

Playing with local models at various quants, the degradation can be hard to spot. Sometimes it's only noticeable in aggregate. And even then, you never really know if you just got unlucky with a bad response due to RNG.

I've had Opus 4.6 fall into some weirdly incoherent loops that I rarely see from even Sonnet, that felt like the kind of thing I got frequently with Qwen3.5 9B on local. And the above applies... Was that just bad RNG? Or was my request to Opus routed to some lower quality variant? There's no great way for me to tell for any given request, nor any way to guarantee Anthropic _didn't_ do that.

nessex··on Local Qwen isn't a worse Opus, it's a different tool
This is a great post that covers a lot of the recent ground. I have a very similar setup after a very similar journey, minus the RTX6000. Worth noting though that a lot of the recent changes make a single 3090/4090 much more viable here too. MTP and the recent improvements to kv quantization in particular, as well as model-specific template & quant fixes. I run a 4090 with the 4-bit quantized variant of the same model now and have had a great experience. Qwen3.5 was already a big step up, but with 3.6 and the rest of the improvements it's substantially more reliable as a daily use tool and I find myself reaching for hosted models a lot less. Feels like I could work entirely without them if they were to disappear without going back to typing every line of code myself.

To make 4-bit fit on one card with reasonable (100k+) context needs a bit more care though. And tuning can be highly specific to your machine, gpu and use-case. But I use a headless server, offload multi-modal to CPU, use fit-target to reduce wasted memory and use q8_0 kv since the 4090 performs well with it... In addition to most of the same config as the author elsewhere. I get 50-60tps generation with a power limit of 275W (450W is default), more than enough to offer a roughly an Opus-speed feedback loop.

I haven't seen many of the issues with looping the author mentions. But I did with Qwen3.5 and in particular other 4-bit quants in the past. But the difference is probably a mix of the improvements above, as well as habits changing to avoid cases where models will loop. For what I'm doing, it seems like I loop Qwen3.6 on the same kind of prompts I'll make Haiku or Sonnet loop on (the latter hide some of their existential loops behind "thinking"). Usually it's cause I was too vague about some aspect of what I'm wanting them to do or I forgot to include some context that smaller models just don't have access to in their smaller knowledge base. But at least for what I'm doing (Rust, React, kubernetes) it's not been a notable problem at all with the latest iteration of this whole stack. And knowledge of standard libraries and default k8s resource kinds has been almost flawless.

There's still plenty of more complex stuff where I'll choose to jump straight to Claude or GLM-5.2, but if it's not worth that jump I've stopped paying for the middle ground as it's usually not much better than just one more iteration through qwen.

All this to say, if you have a 3090/4090, feel free to give the same setup a go. It's come a long way in recent weeks.

nessex··on Redis is fast – I'll cache in Postgres
I've got a similar setup with a k3s homelab and a bunch of small projects that need basic data storage and caching. One thing worth considering is that if someone wants to run both redis and postgres, they need to allocate enough memory for both including enough overhead that they don't suddenly OOM.

In that sense, seeing if the latency impact of postgres is tolerable is pretty reasonable. You may be able to get away with postgres putting things on disk (yes, redis can too), and only paying the overhead cost of allocating sufficient excess RAM to one pod rather than two.

But if making tradeoffs like that, for a low-traffic service in a small homelab, I do wonder if you even need a remote cache. It's always worth considering if you can just have the web server keep it's own cache in-memory or even on-disk. If using go like in the article, you'd likely only need a map and a mutex. That'd be an order of magnitude faster, and be even less to manage... Of course it's not persistent, but then neither was Redis (excl. across web server restarts).

nessex··on M8.7 earthquake in Western Pacific, tsunami warning issued
Sure, but if you insist it's like a tide you downplay the risk of the initial hit of the wavefronts and the potential for it to slam up the coast or a seawall becoming a larger local wave. And if you insist it's like a wave, you downplay the persistent risk of both follow-up waves and ongoing flooding that won't subside quickly.

So saying it's not waves is dangerous, and saying it's not a sea level rise is dangerous. It's not useful to try and delineate between a tsunami being one of the two when it's in reality an event that consists of both.

(Ignoring that a sea level rise and a long-wavelength wave are the same thing)

nessex··on M8.7 earthquake in Western Pacific, tsunami warning issued
It's a distinction without value I think. There are waves, and many of them. There is a rise in the sea level. For anywhere affected, both certainly matter. Like you mentioned, tsunami isn't a brief event. And here in Japan, they are talking about tsunami waves, not a singular tsunami. And talking about sea level rise and checking the local power poles for sea level indicators from previous tsunami events and floods.
nessex··on M8.7 earthquake in Western Pacific, tsunami warning issued
There are planes, buoys and other things being mentioned on the news here in Japan as ways things are being tracked. Maybe not what you meant, but tracking the wave isn't necessarily correct. There are many waves, and the initial wave is often (in this case also) not the largest.

The news mentioned a previous similar event where the largest wave was 4 hours later.

nessex··on M8.7 earthquake in Western Pacific, tsunami warning issued
Not true. As the news reporters here in Japan are repeating every few minutes, there will be many waves and they can get bigger over time. They already have, 20-30cm initial waves had 40-60cm later waves.

Waves can get bigger due to earthquakes not being instantaneous or necessarily a single movement, due to amplification by geography, by reflections, by aftershocks, and many other things. The news is suggesting waves lasted about a day for a previous event in a similar area.

nessex··on Apple says it will add 20k jobs, spend $500B, produce AI servers in US
If you haven't tried OpenAI's advanced voice mode, it's a mind blowing version of exactly what things like Siri really ought to become with a little more development. If that's what you mean by LLM Siri, I totally agree.

Being able to chat casually with low latency, correct yourself, switch languages mid-sentence, incorporate context throughout a back-and-forth conversation etc. turns talking to these kinds of systems from a painful chore into something that can actually add value.

nessex··on How Tokyo became an anti-car paradise
> Because there are so few of them. Most of the time you can walk in the middle of the street, so rare is the traffic

This is a bit of a stretch. There are cars everywhere you go in Tokyo, it's pretty well set up for driving given the size and population. That said, I've lived here without a license for years, and rarely had the need to hop into any cars. Speed limits are generally low, and lights are everywhere, which often makes the train or sometimes even a bike a faster option than a car or a bus.

It's only when you have multiple connections in your route or when you get well outside the city that you start to see consistent benefit from a car.

One reason it's so easy to get by without a car in most of Tokyo is that the shops and attractions are distributed well. Zoning means there are tiny shops everywhere, and the bigger shops are present at many of the train stations in the city. You are pretty much always within walking distance of everything essential. One more reason is that the postal system is excellent. If you order something from Amazon, you'll receive it about 24 hours later in most parts of Tokyo. Who needs a car when you pick up everything essential on a 10 minute walk from your house, and everything you want will be at your door tomorrow?

nessex··on How Tokyo became an anti-car paradise
I ride in Tokyo and surrounding areas often. There are bike lanes scattered around, though some places have far more than others. In general though, there is a lot of patience on the part of drivers here in the city (at least relative to Australia) that makes riding a bike on roads without a bike lane generally a fairly safe option.

Also, there are plenty of terraces to have a nice coffee on. The difficult thing in Tokyo is you need to search in advance for what you want. It's so crammed full of options, that you might never realize all the interesting things you're walking past. It's a very three-dimensional city, there's stuff hidden everywhere.

nessex··on We updated our RSA SSH host key
It's not mentioned in the blog post or keys page, but the _old_ value[1] you'll find in known_hosts is:

  github.com ssh-rsa AAAAB3NzaC1yc2EAAAABIwAAAQEAq2A7hRGmdnm9tUDbO9IDSwBK6TbQa+PXYPCPy6rbTrTtw7PHkccKrpp0yVhp5HdEIcKr6pLlVDBfOLX9QUsyCOV0wzfjIJNlGEYsdlLJizHhbn2mUjvSAHQqZETYP81eFzLQNnPHt4EVVUh7VfDESU84KezmD5QlWpXLmvU31/yMf+Se8xhHTvKSCZIFImWwoG6mbUoWf9nzpIoaSjB+weqqUUmpaaasXVal72J+UX2B+2RPW3RcT0eOzQgqlJL3RKrTJvdsjE3JEAvGq3lGHSZXy28G3skua2SmVi/w4yCE6gbODqnTWlg7+wC604ydGXA8VJiS5ap43JXiUFFAaQ==
You can search for this in your codebases, hosts etc. to see if there are any areas that need updating. The new value is linked from the blog post, you can find it here: https://docs.github.com/en/authentication/keeping-your-accou...

[1] https://github.blog/changelog/2022-01-18-githubs-ssh-host-ke...

nessex··on Resource efficient Thread Pools with Zig
This is not the case, it measures the time between two points in the code at runtime and prints that[1].

[1] https://github.com/kprotty/zap/blob/blog/benchmarks/rust/ray...

nessex··on Scaleway announces measures against abusive Chia plotting and farming
Most of the work happens in C++[1]

[1] https://github.com/Chia-Network/chiapos

nessex··on I’ve had the same supper for 10 years
That's pretty much the exact philosophy I live by. I've definitely found no bed frame to be a hard-sell to family and friends, and it's hard to see why once you've tried all the options. A mattress makes a lot of sense, but a bed frame adds little value unless you are short on storage and one has storage built in, or you aren't mobile enough to get to the ground. But maybe I'm missing some utility that others have found in their bedframes!

Living in Japan now, I had a few months with a padded mat + quilt on the floor as is tradition (and a damn cheap one), but upgraded to a mattress on the floor because the floor was too cold in winter as you mentioned.

There's as much to be gained from taking stuff away that isn't useful, as there is from adding useful stuff to your life.

nessex··on I’ve had the same supper for 10 years
Sorry if you saw my original comment, I misread this as a dismissal through exaggeration, but after double checking my comprehension I realise I was both wrong and missing the fact that I can relate to most of this. I've tried many of the things you mention, and while I don't do all of those things still, many of them do make my life easier and more stress free. It's interesting how many of the things I've just stopped thinking about as I tried them and subsequently rid my conscious mind of other more time consuming or stressful options.

There are so many better things to spend time on than the mundane parts of life.

nessex··on I’ve had the same supper for 10 years
I'm in Japan so I use nosh.jp. It's decent and surprisingly cheap, not much more than food from the supermarket here which is expensive regardless.
nessex··on I’ve had the same supper for 10 years
Right, it's about eliminating the mundane parts, not about having nothing in my life. It's a balance that will be different for everyone.
nessex··on I’ve had the same supper for 10 years
Yeah absolutely, that's kind of what I'm talking about. Though even then, managing different contract durations across many different companies for many different bills each month is annoying, and there are companies that can do that part for you as well. Haven't ever tried it, nor checked the cost, but it sounds like something that might be beneficial to not worry about. They can send me a summary each month to make sure I'm not spending too much.

The index fund investing with scheduled transfers is exactly what I meant by automated financial services. I probably micro-manage it a little too much right now for no real benefit.

nessex··on I’ve had the same supper for 10 years
I've found a lot of freedom in similar decisions. Not sure I could take it to the same level, but even just having a small set of meals to eat every week makes shopping, cooking and planning around expiry dates so much easier. Clothes can be similarly hacked such that everything goes together and every combination is something you are comfortable wearing, leaving you never needing to consider what to wear. I've optimised these to the point that they take up nearly zero mental space and generate no stress. In my case, I use pre-prepared frozen meal delivery service, but I know some meal preppers who find similar freedom that way. Don't cook or order anything you won't eat at any arbitrary time, and you'll never be stuck with wasted food or indecision. And for clothes I found a small set that works for me and can be worn in any given situation (except formal, though that doesn't impact me in any way).

I see a lot of comments that seem to see all the things you miss out on in this situation. But in my mind, it frees up a lot of mental effort, time and stress. If I ever get bored I can go to a restaurant and eat something wild and it will be all the more exciting given I don't optimize for excitement or luxury in my everyday steady-state.

When Soylent came out I was super excited about this idea. Don't think about three meals a day that you normally fuss over, and instead have two predictable, quick meals and optimize to make the third one amazing. Soylent was OK, and DIY soylent offered some hope too. The third meal WAS always amazing, in a relative sense, and tasted better somehow than when I had the same thing before this diet. Unfortunately liquid diets are just not satisfying to me and so frozen meals won out.

I'd love to find other areas of my life that can be similarly optimized. I have hope for bill management services to take the annoyance out of juggling payments etc., and roboinvestors or similar automated financial services. Doing these things manually offers no excitement and no added value beyond the transitively provided service so I don't think they should take up my life.

The amount of time wasted across the whole human population on things like preparing meals, choosing outfits and managing everyday responsibilities must be huge and that is all time that could be spent doing other exciting or valuable things.

nessex··on Dave Herman’s contributions to Rust
Possibly, when working with big codebases I'm typically working on Kubernetes controllers in both Golang and Rust. That makes for extra slow golang compiles, and the rust incremental compiles are significantly quicker in comparison. Otherwise the codebases tend to be quite small, and for those golang full compiles (the only option?) and rust incremental compiles are similarly nearly instant.

It is absolutely apples to oranges, but if you just care about the everyday local workflow and ability to iterate and test, it's close enough most of the time to not be much of a problem in either case.

I'm sure this experience isn't guaranteed for all codebases, and it certainly helps that I make heavy use of crates, which would minimize the work required during incremental compilation. Though I'm not actively going out of my way to optimize for incremental compilation really, beyond the config linked above.

nessex··on Dave Herman’s contributions to Rust
When I use rust, I find compile times faster and more manageable than other languages due to the speed of iterative compiles. Compiling from scratch is very slow, but iterative compiles are faster than most of my golang compiles and faster than running a JS builder in most projects. To make it extra fast, I follow the instructions from the bevy game engine[1]. With that setup, the feedback loop is quick.

[1] https://bevyengine.org/learn/book/getting-started/setup/#ena...

nessex··on Fines up to $66k or five years prison for Australians returning home from India
As an Australian citizen, I have plenty of sympathy for those who have dual/multiple citizenship. They've been treated particularly harshly by both the federal government and their fellow Australian citizens. Some notes:

- There is an income test for permanent residency, which is required for most pathways to citizenship. Apart from family-stream visas which require a direct relationship with an Australian citizen, most visas eligible for permanent residency are work-related and you won't get one of those without an income test[1].

- There is an employment test, as above

- There is a language test, see the section on Language Ability[2], this has been a significant bottleneck for personal friends who speak English perfectly well

- Student visas don't make you eligible for permanent residency, let alone citizenship[2]

- Australia's community values explicitly state that dual/multiple citizenship doesn't exempt you from being Australian[3]

- Australia's pathways to citizenship (via PR, work visa etc.) are some of the most expensive and time-consuming in the world

[1] https://immi.homeaffairs.gov.au/visas/permanent-resident/vis...

[2] https://immi.homeaffairs.gov.au/citizenship/become-a-citizen...

[3] https://immi.homeaffairs.gov.au/citizenship-subsite/Pages/Le...

nessex··on Fines up to $66k or five years prison for Australians returning home from India
This sort of thing has been brought up for more than a year. The federal government has shirked all responsibility from the start of this whole affair. States have requested federal support for establishing these sorts of quarantine facilities, and been rejected on multiple occasions. The federal government claims it is a state responsibility, as health is generally a state responsibility. However, constitutionally quarantine is a Commonwealth responsibility and therefore the federal government should be handling it. As a citizen overseas, states do not harbor much responsibility for my situation as I'm not a resident of said state, which is fair enough in my mind. But that means the federal government's dereliction of responsibility leaves me and many others out to dry. I'm just glad the country that invited me in hasn't been so incompetent and cruel.
nessex··on Ryzen 5800X vs. M1: Programming benchmarks
The most impressive part to me is how the m1 compares to the 3900X. I’ve got a mere 3600X and every laptop I owned or worked on over the past year is noticeably and painfully slower than the 3600X. It’s been a relief to get home and turn on my desktop. It doesn’t matter if the laptops I’m using are very recent i7’s and i9’s, the desktop is always very noticeably faster.

I got my m1 MacBook Pro 16GB yesterday and was pretty confused to find that Rust compilation felt faster than the desktop. To go from recent Intel laptops taking two to three times longer to compile Rust than my budget desktop despite the laptops' whiny fans and extreme heat, to having my desktop be sightly outclassed by an ice cold laptop on battery (which got at least 15 hours of use, including said compiling, with no need to charge) is a world changer. I can’t remember the last time a laptop was this close to desktop performance for my everyday workflow.

Now that I think about it, I haven’t even tried optimising compile times on the MacBook like I usually do with Rust projects. My desktop would have been running lld at least to make its compilation significantly faster, and the MacBook more than kept up in spite of the handicap.

nessex··on We can do better than DuckDuckGo
I really wish Google would prioritize English results for English searches consistently. I'm living in Japan as a native English speaker, and have my OS, browser and logged in Google account all configured for English only. Despite that, Google search results always prioritize Japanese language content. Every now and then (though not consistently) it gives me a yellow popup asking if I'd like English results instead, which is a bit disappointing given they already have all the information they should need to make a judgement call about that. Maybe the individual experience here depends on the languages and regions involved.