886 karma · joined April 15, 2017
A common form of this failure is the model picking up on random wordings from earlier in the session (e.g. some comment it made to me in the middle of a response, that I never explicitly endorsed) and then treating these as hard commitments. Or over-interpreting a specific word choice or clumsy phrasing as if it were a "load-bearing" constraint on the task.
None of this clumsiness would be so problematic if the model didn't have such a strong drive toward autonomy. It's much like with people: there's no shame in not understanding what you're being asked to do, provided you ask clarifying questions. There's no shame in ignorance if it's wedded to curiosity. Benchmaxing has RLVRed curiosity and clarification straight out of these models. It sucks.
Humans fail at good style more radically than LLMs do, and LLMs are on average better writers than people. But good style requires particularity and history in a way that a model could never achieve with just a filter.
The closest I’ve gotten to getting an LLM to write well was when I gave it an intellectual biography along with its writing assignment. But even that fails over longer context, because attention makes it very hard to build up narrative (even just conceptual narrative) in a compelling way that extends past the length of a short blogpost. Past a certain point you always end up back in the uncanny valley with that weird lack of cohesion and the unjustified contrastives, tricolons, and empty cliches.
I’m slowly rereading Wittgenstein’s Philosophical Investigations.
I am perpetually rereading the works of Richard Rorty.
> These projects were quite useful, in that they sort of cured me of my AI mania... because after I built these things, I got to look at them and ask “does the world need a slow, buggy, half-baked Python JavaScript interpreter?”
> I don’t think the world does.
Very relatable.
For the past year I’ve been yo-yo-ing in and out of existential despair about the future of civilization depending on how I feel the answer to this question looks. It’s emotionally exhausting, on top of everything else, and I wonder how others are coping with it aside from denial and cynicism.
To draw a broad strokes historical parallel you could say this: in pre-Enlightenment Europe, responsibility for reasoning was signed away to various authorities (the church, the prince, etc.), "liberating" common people from the need to reason or ask questions or make self-directed use of their intellectual faculties. The vision of enlightenment Kant puts forward gives us an alternative ideal: maturity and freedom come from the self-directed use of one's intellectual faculties, the choice to take on responsibility for one's beliefs oneself, even if the freedom to think and discuss is constrained within the confines of public law. The ideal is inquiry, questioning, the desire to understand things from first principles and the refusal to accept the dictates of external authorities as given without being able to participate in the reasoning that justifies their claims.
If Huang, Altman, etc., are offering us all a future in which nobody is indepdendent minded enough to be able to perform arithmetic or even know their own address, what they're offering is "liberation" backwards into the world of intellectual immaturity, into a kind of serfdom disguised as leisure or carelessness. Maybe some people want that, and maybe they can even argue that it's good, but it's worth being clear about the fact that "liberation" from the need or desire to know or reason about the basic facts of one's life is not empowering, and does not yield greater personal independence or maturity. It is, in Kant's terms, anti-Enlightenment.
> Enlightenment is man's emergence from his self-imposed immaturity. Immaturity is the inability to use one’s understanding without guidance from another. This immaturity is self-imposed when its cause lies not in lack of understanding, but in lack of resolve and courage to use it without guidance from another. Sapere Aude! “Have courage to use your own understanding!”--that is the motto of enlightenment.
I know there's been discussion about whether pelicanmaxxing is happening, but this is at least evidence that Claude was explicitly exposed to this problem.
The threat to mathematics isn't that suddenly the profitability of their profession (lol) is going to go away, it's that people are thinking of AI as a replacement for the human social and intellectual practices that constitute the discipline.