Good problem to take on, you're absolutely right that there is a lot of demand for reliability. Curious how effective the learning + tuning really is.
I presume that's the reason for the limited selection of models - i.e. only some are tunable? I think that's my biggest issue with this solution, if I'm going to be dropping this into my LLM pipeline I need to have more control over how the model is used.