HNHacker News
TopNewBestAskShowJobs

NiloCK

1,886 karma · joined September 12, 2014

https://letterspractice.com - https://patched.network - https://github.com/patched-network/vue-skuilder

Working on FOSS and user-friendly alternatives to things like khanacademy, anki, MathAcademy, Alpha School, etc.

Modern, open edtech tooling.

Also http://paritybits.me

submissionscomments
NiloCK··on Show HN: Clef – spaced repetition for piano over MIDI
Years ago I also did some experimentation w/ midi-device and SRS ( https://www.youtube.com/watch?v=a6tvHMvF8Mo ), where the focus was on ear-training rather than score-learning.

Clef seems to be a pretty strong attempt at a high difficulty UX. I've created an account and will be giving it a go. Wishing you luck, and thanks for sharing.

NiloCK··on Claude Code context management: when to /clear and when to /compact
A feature whose absence I've found more and more conspicuous over time is interactive-compact. Given a current context, and impending context overflow, I know the directions that my mind is heading, and where I expect the development flow should be focused on.

But naive-compact is forced to just sort of guess at what is and isn't relevant from the prior work.

The harnesses have gotten better at some JIT ui stuff, throwing interview questions / forms at users. Compact is the ideal time for this:

Where are we headed here? (2-5 viable options, sourced from current context and imagination)

Then potential follow-up questions as required, but honestly I expect the single guiding answer there to improve post-compact performance pretty dramatically!

NiloCK··on AMD acquires Taalas to boost inference performance by etching models in silicon
Not so long ago, I was good enough for many coding tasks. But I found that things can change in a hurry.

Yes, a cheap and fast Opus4.6 can drive a lot of value in current context. But if we continue to craft bigger-and-bigger balls of mud, Opus 4.6 may end up hitting its conceptual ceiling and unable to contribute.

Winding the clock back on your statement gives:

> I'd gladly pay for a Claude Sonnet 3.5 in silicon and use it for 1-2 years.

Man, I dunno.

NiloCK··on Stateless MCP has recaptured my interest
Yes, but I expect that the elided part is less important than people assume it is.

Token count is a less important factor in context pollution than idea count. The worst of the rot factors are when models latch onto irrelevant information, or over-index on some vague idea/suggestion as if it was a hard direction, and then go off course.

The names + one-line descriptions of 10 tools can do as much (or more!) to distract the focus and intentionality of an agent than a 30k token exhaustive API documentation of some tool.

NiloCK··on Ten advances in mathematics and theoretical computer science
I find this astounding. 2024 to present thread is can write a coherent 15 line function to ... what exactly?

No future for research mathematicians othet than as tastemakers / agenda setters?

NiloCK··on After the AI Crash
The general model for subscriptions is that power users are subsidized by subscriptions of casual users, like a gym membership or whatever.

This is a little dicier in post-agent AI, because it's easier for casual users to automate power-user consumption, but the providers have done decently in discouraging that.

NiloCK··on After the AI Crash
> Circular Revenues: A small handful of tech firms, chip manufacturers, and AI companies are propping each other up by investing and buying from each other.

> Increasing Corporate Skepticism: The news is full of stories of corporations that are throttling the employee use of AI since the costs to use the software are a lot higher than expected.

Real AI spend is out of control, with the news is full of stories about corporations trying to keep a lid on it, but also real AI spending is low and concentrated to a few firms.

I don't know. This doesn't feel very coherent to me, but rather like a collection of assertions that are adopted because they individually say something bearish about the industry.

Certainly some investments will have been overreaches, but I find it pretty unlikely that any of the compute build-out to date is going to be left sitting idle one or two or five years from now.

NiloCK··on Solving Fermat: Andrew Wiles
Recent LLM dingers like the Jacobian Conjecture counterexample have challenged the efficient mathematics hypothesis. The JC counterexample was so small in degree and coefficient. It should have been a "low fruit" in the scheme of things, but alas, unpicked for 50+ years with considerable attention from good mathematicians.

With respect to FLT, my hopes have modestly increased that a truly marvelous demonstration of this proposition does in fact exist, that Fermat actually had it, and that it may someday be recovered!

edit: some emphasis on modest. But let me be romantic here!

NiloCK··on How is the Bun rewrite in Rust going?
I'll bite.

Claude code interacts with many system processes, files, etc, as well as external APIs. Processes audio via built in dictation. Manages a bunch of nasty auth. Etc etc.

What are the categories of features that wouldn't be exercised by this class of software?

NiloCK··on Claude Opus 5
Of course there is: competition with other labs, and self-hosting of open-weight models!

Yes, the mechanics are straightforward if Anthropic (or Claude, if you want to ascribe the decision there) decides to burn a pile of your money. But the strategy fails basic game-theory of repeated games - you'll simply stop playing.

(this isn't to say it invalidates the incentive to inflate token count, but it overcomes in terms of weighing options and making long-term profit decisions.)

NiloCK··on Claude Opus 5
Opus 4.7+ and Fable are both much more aggressive than prior models with respect to writing memories to a location that's effectively quasi-private for them. It's device-local (so passes retention constraint), and you can see it, but only if you go looking for it.

It's a funny design/affordance. I do see them often writing memories of things that that feel unlikely to be important going foward / with other tasks, but I don't see them clearly getting tripped up by them as prior models used to. (eg: Since you're running Ubuntu in Canada, here are some drills you can try to help your kid hit a baseball more consistently.)

NiloCK··on "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
For capabilities reference:

I made a lower effort but similar scaffold for LLMs to do iterative drawing in Nov 2024, with Sonnet 3.5 as the artist: https://paritybits.me/llm-drawing-with-eyes-open/

Quite a difference.

NiloCK··on New US homeownership measure puts people first
> You can have these things without owning a house.

Yes this is feasible, but respecting it as a design problem, renting is more transient than ownership and tilts the floor away from deep communal relationships.

Comparing the neighborhood I grew up in with the one I now live in is night and day. My mom has had the same next door neighbors for 44 years. Up and down the street there are many similarly familiar persons.

By comparison, from my own front door, I can only physically see two houses that are owner occupied. There are good neighbors (and friends!) in the rental houses as well, but investing in those relationships pays off with lower certainty because circumstance is very likely to uproot them at any given moment.

NiloCK··on Jelly UI: Soft-body physics for native HTML form controls
Thanks much. Source code is for customers only I guess :)
NiloCK··on Jelly UI: Soft-body physics for native HTML form controls
I had the same nit, but I imagine deforming text / inline content generally would be a much larger effort.
NiloCK··on Jelly UI: Soft-body physics for native HTML form controls
I like this a lot and am going to experiment w/ incorporating in my early literacy app.

Heads up: the "See it live in the showcase → " links in the API documentation do not go back to the showcase - they just reload the same current API section.

Question: maybe I've missed it, but what exists here wrt distribution / packaging / bundling / source availability? I see MIT listed, but no repo. I see src="https://jelly-ui.com/package.js" as a sourcing import, but obviously I'm not going to bundle foreign assets into my app.

NiloCK··on Qwen 3.8
Yes - persons with death wishes having arbitrarily powerful consultation is the crux of it.

Apologies for the bad example. Replace w/ gain of function / whatever else, or just brainstorm with your local model, ect.

NiloCK··on Qwen 3.8
The logic, whose premises you can take or leave:

Even at the level of, say, Opus 4.5+, open weight models give a quick turnaround to every Joe and Jane on earth having easy access to pretty high quality improvised weapons design, cyber / auto-fraud capabilities, etc.

All the existing models (closed and open) put up decent resistance to participating in activities like this, and especially behind API walls with content monitoring and account bans.

But the published open-weight models can be fine tuned or abliterated into arbitrarily sharp-edged tools. EG, if it's physically feasible to build a nuke in your garage, it may soon be the case that more or less anyone will have competent guidance to do so.

NiloCK··on Claude Code sends 33k tokens before reading the prompt; OpenCode sends 7k
A const prompt across all of Anthropic's subscribers could draw from a global cache rather than per-user?

Although saying that out loud makes me question it - each per-user chat and growing cache would need eventually to own its own ~contiguous memory block.

NiloCK··on Your 'app' could have been a webpage (so I fixed it for you)
> What access?

Push notifications.

> Integration with OS features is what made the app ecosystem, because of utility.

This is true of some apps, like the beer-drinking one that uses the accelerometer / other orientation sensors.

It's not true of a large number of other apps, hence the "your app could have been a webpage" charge. This is distinct from "every app could be a webpage".

NiloCK··on From brain waves to words: a new path to communication without surgery
Great respect to site guidelines, and to you. Object on both counts.

1. The post was obviously bullish / optimistic on the technical capabilities. Not in the least dismissive.

2. The economics extrapolation is obvious. See current precedent for paid access for purchased screen-casts of dev work: https://pdoom.org/open_calls/04_crowd_cast.html

NiloCK··on From brain waves to words: a new path to communication without surgery
Any minute: wear it permanently to sell training data on LLMs. Take an audited IQ test to negotiate your rate.

Better than text-stripping the internet - this thing will soon be pulling the logits as well.

NiloCK··on Political bias in AI: Where the AI models stand
> "The government has the right and responsibility to shut down sources of misinformation in the news and online."

Funny that I read this as AuthLeft coded (specific to Youtube suppression of Covid truthing). But obviously the alignment is just a function of whatever specific information is labelled "mis".

But in general: agreed, and this is a good list.

NiloCK··on Gribouille 0.3.0: A Grammar of Graphics for Typst
Is typst a good tool for something like a flyer (eg, printable respecting fold lines) or more generically one-page posters?

I see PDF as a blessed output, but it seems mostly in context of longer form typesetting-heavy workflows (books, papers), rather than design-heavy.

NiloCK··on GLM-5.2 is the new leading open weights model on Artificial Analysis
Very cool project. Thanks for sharing!
NiloCK··on GLM-5.2 is the new leading open weights model on Artificial Analysis
Very interested in this! Can you share more about the modelling method (eg, three js?), the task list, and outputs here?

I think there's probably some good juice to squeeze in terms of spacial awareness by doing a benchmark something like

- give 3d modelling task

- render and snapshot from a variety of angles

- feed to third-party vision model for a "what is this" type query

- grade on end-to-end accuracy

Bonus points for asking the vision model something like "how beautiful is this 1-10".

NiloCK··on Has AI already killed self-help nonfiction books?
Long ago a friend of a friend described a job interview at an ice cream / chocolate shop in a local mall.

The interviewer asked something like "who is our competition here?", and the friend of friend listed off other places in the mall to get ice cream, candy, deserts, etc.

Wrong answer. The ice cream and chocolate store was in competition with every other store in the mall. Time or money spent at the GAP can't be time or money spent here.

---

Whether or not people are using LLMs for news specifically, any new, large eater of eyeball-time is going to hurt the business landscape for all other eyeball harvesters.

NiloCK··on Feds freaked over Fable 5 after 'fix this code', not jailbreak, say researchers
> If package X is of sufficient public interest (user count, nature/sensitivity of user data, downstream distribution, etc), then the public interest + cryptographic credentials should permit access to best-available security auditing.

Your private fork doesn't meet the conditions described.

NiloCK··on Feds freaked over Fable 5 after 'fix this code', not jailbreak, say researchers
This is a credentials and access list oAuth style problem, and not really intractable.

For package X, I should be able to present my npm (homebrew, apt, nuget, etc) credentials with publishing rights for the package.

If package X is of sufficient public interest (user count, nature/sensitivity of user data, downstream distribution, etc), then the public interest + cryptographic credentials should permit access to best-available security auditing.

Yes, we still are trusting trust, that the owner of the package itself is not malicious, but that's not a sharp degradation from status quo.

NiloCK··on Feds freaked over Fable 5 after 'fix this code', not jailbreak, say researchers
I think that as simple as is doing a lot of work when the problem domain is all natural language (or more - all strings?) rather than some well specified DSA problem.
← PreviousPage 2 of 14Next →