Guardian Angels: LLM Personalization for Productivity and Security
gwern.net
gwern.net
I’m not sure if I want to be inspired and try the fine-tuning model part too!
May be Gwern is right that this static personalization approach may not be long-lived. Still, I find it fascinating, and I’m seeing a lot of interest in this area from others too.
When I started, my personal interest was on how to multiply/enhance what I already know and touch adjacent topics/experiences and not really become pseudo expert or the know-all-generalists in things I have absolutely no clue.
Now, shamelessly plugging my idea that I started early this year
> Above all, a GA should amplify the principal, and not simply substitute for them for someone else’s purposes or benefit.
[…]
> A GA must be aligned with its principal. It should not be designed to manipulate or control or guide the principal in any way which does not derive from the principal themselves. “Constitutional AI”, “Terms of Service”, “social harmony” etc. may all have their place, particularly for widely deployed superintelligent systems—but inside the privacy of a GA, the principal must have freedom from optimization pressure.
…I read this to suggest that it should amplify a scammer’s scamming, a thinker’s thinking, a tinkerer’s tinkering, a cop’s sleuthing… and I’d imagine it implies amplifying a person’s capability to avoid being scammed, too…
One man’s scam is another man’s “pro-social nudge” and another man’s “attractive opportunity” and another’s “advertisement for a delightful consumer wonder” and another’s “patriotic duty to sustain demand to prop up the too-big-to-fail ideas we bet the whole economy on.”
When you fix and operationalize all values centrally, universally, and externally to the principal… that’s current-gen frontier chatbots, not Gwern’s GA concept.
It sounds like the context is that Gwern is a writer who wants writing assistance, and in 2026, all the AI labs are working on coding. Writing style is much less important when working on code.
Perhaps if at least one AI lab focused on writing style, we would see how much better they can be at writing? I'm not sure it requires a "digital twin."
I've always assumed Gwern was anonymous because the norm on the internet generally was to try to be anonymous, especially those who've done research around crypto, darknet markets and other spicy areas.
Maybe I'm just getting old and losing touch, but is the expectation now that you can only write long articles and ask for donations or contributions if you publish your real name somewhere?
Could I ask what sort of age group you are?