Some things I'd like to see solved in this space:
- Versioning. Don't change the LLM model behind my application's back. Always provide access to older versions.
- Freedom. Allow me to take my business elsewhere, and run the same model at a different cloud provider.
- Determinism. When called with the same random seed, always provide the same output.
- Citation/attribution. Provide a list of sources on which the model was trained. I want to know what to expect, and I don't want to be part of an illegal operation.
- Benchmarking. Show me what the model can and cannot do, and allow me to compare with other services.