How does this work since LLM outputs aren't deterministic? Most users don't have full control of their LLMs.
Naturally, I hope that OpenAI and Anthropic will one day offer deterministic inference; see the "Weights" page for what this might look like.