Behaviorally fingerprinting Ox Alpha's provenance
ctgt.ai
ctgt.ai
Matching the tokenizer is interesting tho
Comparatively, you can't be 100% sure that Z.ai isn't able to host some other lab's model (although in this case, the hosting errors still support the GLM theory).
You can finetune an LLM to a new tokenizer by nudging just a few layers (even wildly different kinds of tokens), and there's nothing stopping a lab from using someone else's tokenizer.
Ox-Alpha Is GLM? - https://news.ycombinator.com/item?id=49422226 - Aug 2026 (65 comments)
A mysterious free AI model is impressing developers. Nobody knows who made it - https://news.ycombinator.com/item?id=49406289 - Aug 2026 (4 comments)
Ox Alpha - https://news.ycombinator.com/item?id=49381896 - Aug 2026 (202 comments)
Pretty cool model, especially since it's not a nanny, if you want to unlock your own devices, like rooting an Android, it will happily help instead of flagging you.
Available for free via OpenRouter and Nous free tier. Also via OpenCode Go, but you have to pay a $5–$10 subscription. APIs are hammered now, so service is bumpy.
https://openrouter.ai/stealth/ox-alpha#uptime
https://openrouter.ai/z-ai/glm-5.3#uptime
Screenshot of a recent blip: https://files.catbox.moe/haq90y.png
Both as in "Why is there a stealth launch like that in the first place?" but also "Why does it matter? Is it very good in something?"
Okay, fair enough. Thanks!
I like Ox Alpha. I'm getting some actual work done with it, I'm really enjoying working with it, and I've already cut back my Anthropic budget in anticipation of working with this model instead.
But I'm terrified that when the model is revealed, I'll discover that it was Grok 4.7 all along.