> how exactly would you evaluate how well they show empathy?
How would you want yours rated? By someone you have communicated with, or some data centre somewhere?
How would you want yours rated? By someone you have communicated with, or some data centre somewhere?
I suppose you could do that with the survey as well. It'd be an interesting study to see which is more reliable.
By a knowledgeable/skilled person who listens to the call. (Which the AI solution provides).
Can you point me to the information you evidently have about which models Kaiser is using? All I can find is that they're using innovaccer, which can use any of anthropic, openai, and meta models on AWS or azure. Even their published papers don't seem to specify a particular model or capability level, just "AI". For all we know it's a gpt mini or similarly cost-effective model that has the context awareness of a Labrador hearing the word "walk".We don't know, so let's not pre-judge.
Are you saying that the AI is the same as a knowledgable/skilled person?
Sorry, I parsed this as claiming ‘the AI solution provides a quality of results the same as a human.’
Are you actually saying that the AI solution should provide a human with the calls it identifies as needing a human review?