I mean, it's a structured output model that (apparently) can't hallucinate. I don't mind the name.
I think that this specific part is not super interesting if your harness just recovers from invalid LLM outputs.
The latency and cost - yes, those are super interesting.
Would like to have something like in the original post but open weights.