Fable is by Anthropic, and this is too expensive, GLM 5.2 is roughly the same quality at a much cheaper price.
(I mantain a client with llama.cpp and 101 models across 14 companies by http)
(I mantain a client with llama.cpp and 101 models across 14 companies by http)
Having said that, the safety system on Fable makes it an extremely unattractive model. It feels that half of the time you're paying double for Opus level performance.
I finally bumped into a task that Codex would refuse to work on.
Was I attempting to reverse-engineer a GPU driver? Yes. Was I trying to hack into the DoD? No.
I wasn't doing anything wrong, but that's not what OpenAI's safety mechanisms thought.