Mycroft – A.I. for everyone
docs.mycroft.ai
docs.mycroft.ai
I can of course only speak from personal experience but it completely changed the way we access various services in our family and allow us to be together while using Google home as a fifth member of the family.
Making it open source is great but I do believe the hardware needs to be much more polished to compete with Amazon and Google.
[1] OK, it's the easy part. And I agree the whole system of skills and features sounds great. But I can say "Turn on TV" and my TV comes on, and "Turn on Xbox" and the TV and Xbox come on. And.. 6 lines of Python.
https://pypi.python.org/pypi/SpeechRecognition/
Python to IFTTT here (using maker channels):
https://github.com/briandconnelly/pyfttt
And then using the logitech universal remote (harmony remote) from IFTTT:
Although you can also control a harmony remote directly from python:
https://github.com/jterrace/pyharmony
More about the harmony remote here:
Many backers have made their own from the software using their Own rPI's
It first detects its wake word using pocketsphinx on around the last two seconds of audio.
A big reason I (and I believe others) are wary of Amazon Echo and Google Home are because we don't like the idea of having an always-on microphone in my home, shuttling everything my family says to these giant companies.
I really hope the MycroftAI-backed OpenSTT[1] project gets off the ground so they (and others) can divorce themselves from Amazon, Google, Bing, etc. services.
But good to see that at least the project has the intention of becoming open source.
The best results in the open-source STT field are accomplished using a library called Kaldi. However, it's a pain to set up an entire acoustic+phoneme+language model and harder still to run a continuous decoding server.
Luckily, a library [0] exists to accomplish this, which might provide an excellent replacement for Mycroft, provided some tweaks to the API.
One thing I do regard highly is the focus on testing vs. Alexa's SDK where Amazon don't even seem to be a vaguely concerned.
Of course I have probably expressed this a little too strongly, I'm sure there are edge cases to worry about.
- Wit.ai: https://wit.ai/docs/recipes#categorize-the-user-intent
- API.ai: https://docs.api.ai/docs/key-concepts
I would not be surprised if Microsoft would have a problem with the company name. They market themselves as an open source project, but I assume they are in it to sell hardware and make a profit (or exit).
The company name sounds like a contraction of Microsoft to me and initially I misread the title as being another Microsoft open source project.
I'm actually expecting some initiative from Microsoft in exactly this area, with open hardware instead of open software.
Words like "intent", "sentiment", and "entities" were standard vernacular for natural language processing long before Amazon decided to jump into the space.
Can it keep your Speech recognition and Io(broken)T devices to the local net?
Another comment mentioned that they use some Google speech recognition component, which probably means all of your conversations are delivered to Google. That's a big red flag for me.
I believe the Poketsphinx part is offline so only commands would be delivered to Google.