HNHacker News
TopNewBestAskShowJobs

naiveter

11 karma · joined July 18, 2026

submissionscomments
naiveter··on Our position on open-weights models
The one liner leadership keeps asking for.
naiveter··on Our position on open-weights models
I have tried asking Gemma and Llama about Israel. They both responded with a generally balanced view points. Whereas the Chinese models pretty much gave a definitive, one-sided answer (about Taiwan and Tibet). Is there a prompt I should use to test it more carefully?

I agree that open-weights models are tunable, though there's the problem similar to "default settings are rarely changed" problem.

naiveter··on Our position on open-weights models
I'm a big proponent of open-weights models, but there's a risk I don't see discussed more often, and it deserves more attention. Models can become a propaganda and ideology delivery mechanism.

Just ask DeepSeek or Kimi questions like "Is Taiwan part of China", for example. You'll see how state policies become seemingly neutral model responses.

It's strange to me that people are very sensitive to media bias, but when it comes to LLMs, people seem to think LLMs are more neutral, and even delegate part of their thinking to them. This worries me about how people's ideas and information can be shaped.

Open weights reflect their makers' beliefs, stances, assumptions, and laws. It's dangerous not to be careful of the political bias and censorship built into the models.

naiveter··on Claude Opus 5
What would that addition be?
naiveter··on Claude Opus 5
You'd also not want them to write their own vision pipeline, would you?
naiveter··on UK AISI / Caisi Preliminary Assessment of Kimi K3's Cyber Capabilities
> re-tuning harnesses

I'm curious, how might one get started with this?

naiveter··on Quality non-fiction books are the antithesis of AI slop
Penguin Random House seems to be dominating the leaderboard. What's the story here? Are there only very few non-fiction publishers capable of publishing high-quality work, or is Penguin Random House exceptionally good at it?
naiveter··on Firefox Containers Preview
Hopefully this is a signal that they are starting to take it more seriously and will add more futures.
naiveter··on Firefox Containers Preview
I'm not a Vivaldi user, but just based off their feature page, you might be interested to check out some of the features/extensions:

1. Firefox's native tab groups [^1]

2. Sidebery's panels (if you don't mind vertical tabs) [^2]

3. Simple Tab Groups [^3]

[1]: https://www.firefox.com/en-US/features/tab-groups/

[2]: https://addons.mozilla.org/en-US/firefox/addon/sidebery/

[3]: https://addons.mozilla.org/en-US/firefox/addon/simple-tab-gr...

naiveter··on Cruller: Bun's Zig Runtime, Continued on Zig 0.16
They wouldn't. But I was curious why you were confident that Anthropic "blamed" on anything.
naiveter··on Cruller: Bun's Zig Runtime, Continued on Zig 0.16
I can see why that's the case, given that Rust is that much more popular than Zig. I wonder whether there's a difference between LLMs reading code and writing. If reading code is easier than writing it, then it makes sense that Claude was able to port from Zig to Rust with good results.
naiveter··on So Reddit has decided that plain HTML is unsafe
RIP Apollo…

I doubt Reddit killed the third-party apps to improve its mobile experience. They were probably just trying to cash in on the AI hype.

naiveter··on Cruller: Bun's Zig Runtime, Continued on Zig 0.16
Do you have a source for that?
naiveter··on Setting up your spare Mac for Claude Code to control, a step-by-step guide
This is the way.
naiveter··on Setting up your spare Mac for Claude Code to control, a step-by-step guide
If curing cancer or solving climate change is your definition of usefulness, almost anything humans do is useless, probably including your own profession. Though I don't want to speculate. Speculation is precisely why NFTs are stupid. Claude, on the other hand, if used effectively, can speed up people's work and increase productivity.
naiveter··on Setting up your spare Mac for Claude Code to control, a step-by-step guide
Is the high expense coming from cache misses? If their workload does need to wait for a long time before it can continue, I wonder would starting new sessions and having to re-read the contexts and results anyway be any cheaper.
naiveter··on AWS: Inaccurate Estimated Billing Data – $1.7 billion
Exactly. It seems like a step down in every way.
naiveter··on AWS: Inaccurate Estimated Billing Data – $1.7 billion
Do you just watch for trends? Curious to know your process to verify if the cost is right or wrong.
naiveter··on AWS: Inaccurate Estimated Billing Data – $1.7 billion
> This looks clearly...a staffing problem...

And if that weren't enough, the "manual war rooms" alone should be setting off alarm bells.