HNHacker News
TopNewBestAskShowJobs

soleveloper

62 karma · joined August 6, 2025

submissionscomments
soleveloper··on Plan mode is dead
While possible to commit, I'd still loose the comments themselves, and the commit will be at least per changeset, and not connected ro a specific changed section.

Roughly speaking, I'd be happy if plannotator would persist something similar to github PR reviews combined with Google docs comments & suggestions.

soleveloper··on Plan mode is dead
I have roughly the same workflow, also with plannotator - which I like a lot - and haven't used or felt the need to use _plan mode_ for at least 3-4 months.

Then telling Claude to work on a document, the instruction is kept to its core.

Now when bcherny explicitly mentioned that it merely adds a single line - it explains why I don't need it.

What may be concerning about "super plan" mode from the creators (or a skill, for that matter) - is that tuning the amount of effort, and how much deep to dig - may become too hard, as it will interfere with several embedded paragraphs explaining what to do, how to do, where to do, etc'.

What I do look for is even better plannotator ability to track changes, combining historical comments (like Google docs), and git blame of several "generations" before current reviewed doc.

soleveloper··on Show HN: Airmash – HTML5 Multiplayer Missile Warfare
Thought I was seeing a 10 years old cached show HN, but no:

Apparently there's an official reboot of the game:

https://www.reddit.com/r/airmash/comments/1whwxxg/the_offici...

soleveloper··on Nvidia announces native GPU programming in Rust
So now every already written kernel can be re-written in Rust and have competitive performance to the cpp version?

If so, that's really big.

And to add the Next natural strp - custom codegen for simulating gpu compute and memory without Nvidia gpu.

soleveloper··on Introducing System One Models and Jev
I think the meaning of can't hallucinate in this model is that the type won't be hallucinated.

So if the generated schema is for a tool call for calculator, then the numbers will be valid numbers for sure (and not random words).

To me, it looks similar to BNF schema already introduced and implemented few years ago: generally speaking - it limits the next token that is allowed to be generated, probs are drawn from a subset tokens.

(tbh, I'm not sure why it didn't pick up as a more standard interface to LLMs, as it made a lot of sense back then, and now.)

soleveloper··on 216M Spy TVs – The LG Smart TV Problem [video]
Is there a list of all LG tracking services and external DNS endpoints?

I'd just block them - and keep local streaming and remote control working

edit: just found out WebOS doesn't support DoH/DoT so nextdns won't work.

I wonder if one of the public dns providers got LG blocklisted natively with specific IP

soleveloper··on Show HN: Buslens – where can I get to by bus? (UK)
It reminds me a site where you'd pick a point on map and see where you'd get to in 15/30/etc minutes, including walking iirc.

Then you could even pick *two* points in map and see their intersection in time period, which was great for finding places to meet friends.

soleveloper··on Show HN: PicoMQ – Durable Streams over HTTP, on object storage
Yes, discord as an example would be awesome, exactly what I thought.

Is this back of envelope pricing include the traffic/bandwidth of the readers?

That could easily be 10-100x of number of messages.

This solution together with a cheap/free caching layer (especially for non members/writers) could be amazing.

BTW, a classic example would be an hn mirror. ;)

soleveloper··on Show HN: PicoMQ – Durable Streams over HTTP, on object storage
Sounds great, and documentation is very clear.

So - in theory - something like a massive chat client, discord like, can be implemented via this solution? And what would be the pricing of such a solution. Cheap-serverless-discord

soleveloper··on SpaceX bond worth 10% less than issue price – heading for junk bond status
> They may only have bought themselves 6 more months of time given their purported burn rate

If they had only ~6 more months they (+auditor) had to issue a warning. The 6 is not a hard number, AFAIK, but surely a point where it must be reported.

So honestly, I doubt it's the case.

soleveloper··on Sodium-ion "salt" batteries will revolutionize electric-vehicle and grid storage
Well, the EU regulated USB-C and it affected the whole smartphones world.
soleveloper··on Sodium-ion "salt" batteries will revolutionize electric-vehicle and grid storage
It is time to standardize EV batteries like wheels: 5-10 mainstream types & competition on quality vs. price vs. range.

Regulator help is needed here.

soleveloper··on Is One Layer Enough? A Single Transformer Layer Matches Full-Parameter RL Train
Makes sense - This is very similar to fine tuning a down stream task in encoder-decoder architecture (~Bert style)
soleveloper··on Giant Banana Pulled Over: Driver Says Cops Have Stopped Him 100s of Times
Was just reading the title and thinking it's a new and upgraded image gen model from Google.. Anyone else?
soleveloper··on Uber's $1,500/month AI limit is a useful signal for AI tool pricing
If the premium models are just about 10% better - that could justify the price vs. self hosting a ~0.5-1T open weights model.

Remember that utilization of these huge racks will not be 24h/7, and these are usually not GPU intensive shops that would train models on the spare compute. With prices of 100-200k USD and north with ~2 years lifetime, that would be hard to justify financially.

Self hosting could easily amount to ~1000 USD a month amortized across many developers. In rush hours - there will be hard rate limits.

Would that 1500-1000=500$ monthly USD justify the 10% decrease in "AI Productivity" ? I guess not. In most cases.

For everyone that asks me around, I'd say that in short term, unless there's a really good reason to self host these coding assistant models, then the big 2/3 coding assistants providers are the better choice.

No one got fired from licensing claude code.

soleveloper··on Trademark violation: Fake Notepad++ for Mac
NixPad++

But don't block on the name, you could release it under NejneobhospodařovávatelnějšíPad++ and people will download.

It'll be easy search & replace later once you settle on a name

soleveloper··on Uber Torches 2026 AI Budget on Claude Code in Four Months
> Are they just copy/pasting their entire ticket description into Claude Code and having it iterate until they land on something that works?

"Their ticket" = that was AI generated. After which they will wait their AI generated PR be checked by an automated AI QA that will validate against the AI generated spec.

It feels like important metric of "corporate AI adoption" should be how effective the human in steering the AI.

IF THE HUMAN ISN'T EFFECTIVE, THE HUMAN NEEDS TO GO.

soleveloper··on A Faster Alternative to Jq
I already can't remember jq syntax. Naming this jg just means I'll type one, instinctively use the other's syntax, and get an error anyway. It's a DX trap.

But I will admit, the new syntax makes a lot more sense.

soleveloper··on France's aircraft carrier located in real time by Le Monde through fitness app
An intelligence satellite - which is not a super common utility nations have - will locate where the aircraft _was_ X hours ago, or at least many minutes ago. A constantly updated missile with a rather simple GPS tracker would benefit A LOT from a live location of its target.
soleveloper··on How will OpenAI compete?
There are incredible authors who happen to be dyslexic, and brilliant mathematicians who struggle with basic arithmetic. We don't dismiss their core work just because a minor lemma was miscalculated or a word was misspelled. The same logic applies here: if we dismiss the semantic capabilities of these models based entirely on their token-level spelling flaws, we miss out on their actual utility.
soleveloper··on How will OpenAI compete?
Treat LLMs as dyslexic when it comes to spelling. Assess their strengths and weaknesses accordingly.
soleveloper··on Claws are now a new layer on top of LLM agents
Will that protect you from the agent changing the code to bypass those safety mechanisms, since the human is "too slow to respond" or in case of "agent decided emergency"?
soleveloper··on The path to ubiquitous AI (17k tokens/sec)
Yes, and even holding couple of cartridges for different scenarios e.g image generation, coding, tts/stt, etc
soleveloper··on The path to ubiquitous AI (17k tokens/sec)
There are so many use cases for small and super fast models that are already in size capacity -

* Many top quality tts and stt models

* Image recognition, object tracking

* speculative decoding, attached to a much bigger model (big/small architecture?)

* agentic loop trying 20 different approaches / algorithms, and then picking the best one

* edited to add! Put 50 such small models to create a SOTA super fast model

soleveloper··on The path to ubiquitous AI (17k tokens/sec)
In 20$ a die, they could sell Gameboy style cartridges for different models.
soleveloper··on Show HN: Trained YOLOX from scratch to avoid Ultralytics (aircraft detection)
Great write-up! There are so many directions you can take it to:

* By training on user data, you can source specific model data images, and then train & classify the airplane model. It might require another model, where only the bbox will be the input, together with distance/calculated measurements of the object ("pixel size"), and the orientation of the plane (side? front? belly?).

* Provide alerts/notification of special aircrafts like helicopters, military, airforce-1, etc'

* When bbox is detected, you can run super-resolution upscaling on the photo/stream of images

soleveloper··on Lessons from 14 years at Google
This is a perfect example of a "bug" actually being a requirement. The travel industry faced a similar paradox known as the Labor Illusion: users didn't trust results that returned too quickly. Companies intentionally faked the "loading" phase because A/B tests showed that artificial latency increased conversion. The "inefficiency" was the only way to convince users the software was working hard. Millions of collective hours were spent staring at placebo progress bars until Google Flights finally leveraged their search-engine trust to shift the industry to instant results.
soleveloper··on Date bug in Rust-based coreutils affects Ubuntu 25.10 automatic updates
Is it? I hope I won't step on somebody's else toes: GenAI would greatly help cover existing functionality and prevent regressions in new implementation. For each tool, generate multiple cases, some based on documentation and some from the LLM understanding of the util. Classic input + expected pairs. Run with both GNU old impl and the new Rust impl.

First - cases where expected+old+new are identical, should go to regression suite. Now a HUMAN should take a look in this order: 1. Cases where expected+old are identical, but rust is different. 2. If time allows - Cases where expected+rust are identical , but old is different.

TBH, after #1 (expected+old, vs. rust) I'd be asking the GenAI to generate more test cases in these faulty areas.

soleveloper··on Show HN: ModernBERT in Pure C
Hey, cool initiative!

Worth mentioning in the title that it's CPU-only: >1200 tokens/s on a single thread is impressive.

Have you considered doing optimization iterations like nanogpt-speedrun? Would be interesting to see how far you can push the performance.

soleveloper··on Wtm (Worktree Manager): A simpler way to work with Git Worktrees
How would you recommend using this for working with multiple agents? I'm thinking about scenarios where you might have several AI coding agents (or automated processes) that need to work on different branches simultaneously without stepping on each other's toes.

Any recommended workflow for coordinating multiple agents? For instance, handling naming conventions, cleanup strategies, or preventing race conditions when multiple agents try to create worktrees at the same time?

This seems like it could be really powerful for that use case, especially with the hook system for per-worktree setup.