It is stunning that Amazon wasn't the forerunner in using ChatGPT/LLM-based tech for Alexa. A mistake in the tens, if not hundred billion dollars when looking at the price MS paid for OpenAI.
Why aren't they hooking up a LLM right now?
It is stunning that Amazon wasn't the forerunner in using ChatGPT/LLM-based tech for Alexa. A mistake in the tens, if not hundred billion dollars when looking at the price MS paid for OpenAI.
Why aren't they hooking up a LLM right now?
I'm a current scientist at Amazon. Amazon has little ability to invest in longer term research. While you have scope to pursue projects without immediate product application to some extent (though far less than you would at Alphabet/Meta/Microsoft), that is less valued than at the other Big Tech companies. There's also a far more bureaucratic hurdles to jump through to publish at Amazon. The lack of care for non-product research shows at every instance. Amazon also has extremely low -- or I should say non-existent standards for hiring scientists. If you're putting a person with a decent and non-existent pub record on the same level, that is not going to yield great results.
Luckily for Amazon, their bread and butter -- Ads via retail and AWS -- are not immediately threatened by LLMs.
Honestly if they had taken the billions they burnt on Alexa and put it in research they would be far far far better off today. But try explaining that to the bean counter mentality of executives that Bezos cultivated...
And that's ignoring the fact that all the publicly accessible LLMs have been "jailbroken" and you can get them to say all manner of wild things, and Amazon is definitely aware that all it takes is one or two viral clips of Alexa saying something unhinged for their sales to start dropping, and the tech isn't currently sufficiently understood for them to avoid that.
I am also much less concerned about jailbroken LLMs, because Alexa is already so broken it can't get much worse. Normal users don't seem to care much about potential misuses (compare with search autocomplete biases of Google) as long as a service works for them. But unless you widely deploy a tool, you won't uncover how it is abused (see Sidney's launch).
ChatGPT was, while less widely deployed, being abused (very publicly) in very similar manners to the way Sydney was; Microsoft could have learned about those abuses by paying attention to the public information about ChatGPT.
I really struggle to think how it would have worthwhile. My problem with smart speakers isn't that they aren't smart enough - it's that for anything slightly more complex, my phone has 100x better UX and information density. An Alexa LLM sounds cool in a sci-fi sort of way, but still feels like it would be useless.
“Alexa, be my friend and talk with me for a while, ok?”
Of course I only need to do this because I might listen to loud music during the day and if I forget to turn it down it'll blast "OK TURNING OFF THE MASTER BEDROOM LIGHT" at max volume right as the family is going to bed.
It's the input side of these things that needs a ton of improvements and the use of an LLM could be basically invisible to the user, just "summarize what the user said into a command"; my most common issues with voice assistants are if I say things out of the expected order or change my mind. "Alexa turn on the bedroom lights actually turn on the office lights" or somesuch, or something like "Remind me to make a reservation for Friday tomorrow" not distinguishing that "make a reservation for Friday" is the text of the reminder and "tomorrow" is the desired time.