HNHacker News
TopNewBestAskShowJobs

NiloCK

1,886 karma · joined September 12, 2014

https://letterspractice.com - https://patched.network - https://github.com/patched-network/vue-skuilder

Working on FOSS and user-friendly alternatives to things like khanacademy, anki, MathAcademy, Alpha School, etc.

Modern, open edtech tooling.

Also http://paritybits.me

submissionscomments
NiloCK··on Mechanical Watch (2022)
On-demand ciechanow.ski caliber articles are a pretty good AGI indicator. All the work on that site is wonderful.
NiloCK··on Ask HN: What are you working on? (June 2026)
I'm working on a framework for general purpose interactive tutoring systems. An SRS background process over a pluggable system of pedagogy protocols over a given curriculum. This is at https://github.com/patched-network/vue-skuilder, or https://patched.network/skuilder

With this framework, I'm making (among other things) an early literacy app at https://letterspractice.com. My aim here is to hit >= 75% efficacy of Mentava at <= 1% of the price.

The app is near to production readiness, and I'd be happy to share access now with anyone who has verbal but non-literate kids. Be in touch if interested at colin at letterspractice.com

NiloCK··on There is a shadow hanging over this Fable thing
People may perceive you to be cheaply mischaracterizing the argument.

Nobody believed or suggested that GPT2 could do longform or produce novel text that stood up against careful scrutiny as insightful or well informed. But because the capabilities were novel, people would have no strong alternative than to believe some person wrote it.

You current tripping over LLMisms is irrelevant. You have years of antibodies, both personal and herd-immunity (eg, the many, many articles and comments that describe LLMisms).

NiloCK··on There is a shadow hanging over this Fable thing
At that time, nobody believed a dead internet was technically feasible. Maybe this is hard to remember now.

The "danger" was in terms of spam / misinformation proliferation, not the same category of capabilities adjacent risks current discussed.

You can hold your own opinions on spam/misinformation as a problem, but to say there was no credibly anticipated outsized downside to a sudden jump in human-passing text generation feels pretty off to me.

NiloCK··on Claude Fable is relentlessly proactive
... so the mechanic produced an invoice, itemized.

changing the CSS - $0.05

knowing which CSS to change - $30

NiloCK··on Lines of code got a better publicist
Sure, live in shame, but don't let go of the humor in it all :)
NiloCK··on The better the autopilot the worse the pilot
> Automation doesn't make operators more careful. It makes them forget how to be. The more reliable the system, the less ready the human.

The entire premise of a system is that it removes the need for careful attention.

system: signal lights tell me whether or not I can pass through an intersection, so that I do not have to attend to potentially high speed traffic from a variety of directions.

system: the side my knife blade sits on my arched guide fingers, so that I do not have to attend to the edge of the blade or the location of my fingers.

etc etc.

NiloCK··on They’re made out of weights
replication: https://en.wikipedia.org/wiki/Quine_(computing)

autonomous replication: https://en.wikipedia.org/wiki/Computer_worm

nb that writing your own quine remains in general terms a fun and challenging exercise in many programming languages, but not python.

NiloCK··on The ways we contain Claude across products
I'm no decision theorist but I think they should wait for the rewards outweigh the expected harms in expectation rather than being statistically equal.
NiloCK··on Artificial intelligence is not conscious – Ted Chiang
Once or twice I've experienced extreme pain, and it was downstream of a bright light shining on a wet rock for millions of years.

I try to imagine myself long ago, on the outside looking in, with someone explaining to me that extreme pain, wondrous art, hunger, triumph, and despair would all unfold in due time where the rocks were wet and the lights bright enough.

I can imagine myself calling this clear nonsense.

NiloCK··on Artificial intelligence is not conscious – Ted Chiang
> You're assuming that because Claude produces text that appears to express these qualities, Claude must have them.

Not to be confrontational, but the OP assumed no such thing. OP asserted that it's important for Claude to have the qualities - not that it's important for Claude to present as-if it had them.

NiloCK··on Expanding Project Glasswing
I find this line of reasoning highly dubious.

Yes, Anthropic is compute constrained, even after the SpaceX Colossus deal.

But supply constraints are the normal operating mode of any market. Anthropic could choose to serve whatever models it pleases at whatever price points it chooses and let the market decide where the value is.

If Mythos at $X overwhelms their capacity, they could just charge $X+1. If still overwhelmed, there are larger prices as well.

NiloCK··on Anthropic surpasses OpenAI to become most valuable AI startup
https://github.com/anthropics/claude-code/issues/6235

to be clear, I don't mean that Claude refuses to read AGENTS.md files, but that Anthropic refuses to bake AGENTS.md into the harness as a first-class feature.

Claude the model by now has learned what an AGENTS file is, and itself is not petty enough to ignore it for partisan purposes, but claude-code the harness doesn't natively support it.

NiloCK··on Anthropic surpasses OpenAI to become most valuable AI startup
Notable and persistent and extraordinarily petty is their refusal to read AGENTS.md files, forcing the inclusion of their branding into the source of repos directly.
NiloCK··on Anthropic surpasses OpenAI to become most valuable AI startup
What do you mean with this?

I doubt there is any large demographic of users paying subscription fees for the joy of abusive role play.

NiloCK··on Please Use AI
A great many today find themselves surrounded by people staring at phones, who express irritation any time they are asked to look up.

I saw a video a while back on one social media site or another where someone sitting in a car recorded three young men shotgunning some beers on an apartment balcony. The insinuation being that hanging out was cringe, and that the poster had caught some losers in the act.

It's hard to gauge "real" general sentiment from social media, but if having a beer in a slightly silly way is the level of vulnerability at which you can be recorded for public ridicule, it's not hard to empathize with a generation reluctant to reach out for connection.

NiloCK··on Orchestrating AI code review at scale
Like it or not, the "merge request" (eg, open a PR) is the Schelling point of relevant information. I expect that At scale here refers to size of software projects, and not only code velocity. Software projects of large enough size have CI configuration that don't typically fully-run on each dev machine.
NiloCK··on The $500K AI Film That "Premiered at Cannes" Was Not in the Official Festival
No no no no no. Big misunderstanding here.

We just meant in the city of Cannes.

NiloCK··on Claude Opus 4.8
I appreciate the generosity, but you're gonna want to meet me first.
NiloCK··on Claude Opus 4.8
I think it's telling how split the opinions are around all of this. A lot of people distinctly disliked 4.7.

Are the dividing lines around personality? Working domains? Opinionated software stuff?

Who knows?

NiloCK··on Claude Opus 4.8
A rambling comment:

I think this is the first time we've had a third minor version bump on a frontier Anthropic model. (I count the 0.5s as major here, because they've been issued non-sequentially and also corresponded to massive capability leaps, eg, Sonnet 3.5, Opus 4.5).

So now the Opus 4.5 family has successors 4.6, 4.7, and 4.8, each posting fairly modest claimed gains. My own experience w/ 4.6 and 4.7 are that I don't firmly grasp any capabilities improvements over my memory of 4.5, but it's all so fuzzy that it's truly difficult to tell.

Maybe my own tastes are saturated now (it's smarter than me?) and I'll never again perceive model progress. Maybe the incrementalism is such that I'd notice immediately if my 4.7 workflows were redirected now to 4.5.

Difficult spot for the labs to be in because, if they have a stronger product, I'd prefer they release it and that I can use it.

But as this dynamic continues, the improvements are going to be less and less legible for end-users, who will complain about the churn-without-payoff, even when the payoff may actually be real.

NiloCK··on Last.fm is now independent
I'm mostly unfamiliar with the current offering of last.fm, but the name is familiar from way back. Glad to see something well-liked reclaim some independence.

At a glance, they're providing an interface to YT sourced content with some value adds around tracking or categorizing listening.

A quick question for users: can the site itself be configured as a listener without streaming / displaying the video? In general, YT has a lot of music, but the perf hit of streaming typically high-quality video as well is a blocker when doing dev work on my main machine.

NiloCK··on All of human cooking compressed into 2 megabytes
Ahh - the dependency graph recipe card. These are excellent. I've imagined something like this forever. Always annoyed that recipes put ingredients in a giant undifferentiated list and then give an instruction like "mix the dry ingredients in a deep bowl".

For a while I expected there could be a good return on a good implementation of this, but now as soon as a strong interface itself is created it seems easy to copy.

NiloCK··on Microsoft starts canceling Claude Code licenses
The thing that drove me away from manual edits was that I found myself confusing the LLM all the time. It would read or write, some code, I'd twiddle with things, and then the LLM's future references to the same code would be a mess.

On balance, and via dictation, it feels likely to be faster overall to just enact the changes I want 'inline' of the conversation thread.

Is this stuff any better now? I think current harnesses probably do have things like file change listeners that automatically inform agents before they act on a file they've previously engaged with if it has changed in the meantime.

NiloCK··on Waymo pauses Atlanta service as its robotaxis keep driving into floods
80-90% reduction, over the course of 170 million miles driven on the famously very controlled city streets of LA, SF, Austin and Phoenix.

On average, I wouldn't expect the regulatory agencies to be very friendly toward outright fraudulent reporting from Waymo. On the very outside, maybe these 80-90% reductions are optimistic roundups from 50-65% reductions. Or do you believe that Waymo is secretly running people down and scooping corpses into their trunks?

What is a sedentary pace of driving?

NiloCK··on AI has a multiplying effect on existing technical skills
It's great that you believe this, but are you hiring?

I don't intend this to read as pure snark, but someone's abstract value isn't much good to them if the job market itself can't / won't recognize it.

NiloCK··on Waymo pauses Atlanta service as its robotaxis keep driving into floods
Over a given driving distance, compared to humans, Waymos produce a 90% reduction in serious injury, 90% reduction in pedestrian strikes, 83% reduction in airbag deployments, 85% reduction in cyclist strikes [1].

We currently sit in the ballpark of 300,000 pedestrian deaths per year worldwide [2]. You should be relieved every time they deploy to a new city.

[1] - https://waymo.com/safety/impact/

[2] - https://ourworldindata.org/data-insights/more-than-a-million...

NiloCK··on Waymo pauses Atlanta service as its robotaxis keep driving into floods
It's about many things, including reaction speed, visual awareness, specific expertise and informed decision making wrt braking or acceleration power. All of these are better in a modern self-driving car (I do not know whether Tesla falls into this category) than in a human.

https://waymo.com/safety/impact/

Over a given driving distance, compared to humans, Waymos produce a 90% reduction in serious injury, 90% reduction in pedestrian strikes, 83% reduction in airbag deployments, 85% reduction in cyclist strikes.

NiloCK··on Was my $48K GPU server worth it?
Doesn't it benefit me if the models I use improve?
NiloCK··on Shunning AI is the human choice
Maybe you're sitting pretty right now, but try posting this from your deathbed, or that of your kid.

The lack of compassion that people display here is shocking to me.

"Don't automate science, because there are junior scientists could be denied the thrill of specific discoveries."

Cancer patients are not accessories to anyone's self-actualization.

← PreviousPage 3 of 14Next →