Picovoice – Embed private voice AI into any product instantly
picovoice.ai
picovoice.ai
The keyword generating application: https://github.com/Picovoice/Porcupine/tree/master/tools/opt...
The keyword generating application is free for personal use only and requires a license for commercial use, but there's no pricing available and it only provides an email address for contacting about purchasing a license.
If anyone from the team is reading this, the information above should be front and center on the landing page. I would guess that 99% of your site visitors are going to bounce on your landing page because the relevant information is buried so deeply.
Thanks again for sharing
You said it needed to be addressed, and was a priority yet the question remains unanswered. It gives me the cognitive-dissonance.
In reality, it is a lengthy decision tree. I do not want to put that here or on the company's website as it will just maximize the confusion. But it makes sense to put the common easy cases and then ask for contact in the rest. Which is probably the path we take as it will save us some time.
That being said I challenge you to find a company who offers similar tech and have pricing on their website. I suspect what I mentioned is the reason. I could be wrong.
Would that type of stuff be viable with your technology?
Though would be a way of automated nappy changing, but an automated rattle or such toy triggered by the baby may well make some tasks less impacting and equally more engaging for the baby. Though identification of needs and with that recognition via audio would be the start.
Mycroft is all FOSS
Kitt seems all FOSS too
Snips has a big CTA to 'Contact Enterprise Team'
It's clear you're trying to discover the right price and of course it's complex. Please be upfront about it, what you've said here could be copy-pasta to your site already.
I've been watching Pico for a while, it's very interesting but I'm feeling like you keep teasing more FOSS and being unclear on costs.
I want to like your project too but currently the others are doing a bit better on presentation.
Something as simple as a few basic use cases, android app, no engineering support, under 10,000 unite, x price for example maybe.
I'm a small developer and tinkerer, so my choice to explore more on not is really price sensitive. However I do also consult with other groups, and may suggest your product as a fit if it meets other criteria I look for "private voice AI" - you've already checked a few boxes!
However, any time I see a "contact us for price inquiries" - I shut down. I know if they can't tell me the price on the page I can't afford it. At that point I don't bookmark the site, I don't research any further, and it reinforces the awesomeness of other projects I've put into my memory for use.
This is not just you, it's a lot of projects/ site on the web.
I assume these businesses are mainly looking for those high volume / high price clients and or people with that kind of money to buy them out completely.. all the while making a cheap plan for people to tinker with, build an MVP and maybe scale up.. and perhaps get enough mid range sales to prove they are worth something to others..
Of course that's not the game plan for every service out there. Anyhow good luck to ya, glad to see people working on ways to do things more privately one way or another.
That would indeed be great.
Also, if complex decision tree is your real objection, then consider putting a price range. "Depending on X, Y and other factors, the price ranges from A for minimal deployment low on Z, to B for large-scale solutions.". Or something like that.
Exact prices are always the best, but estimates for typical cases and a price range are second-best. It lets a potential user decide whether to even bother checking your service out.
Oh wow, thank you!
I also saw a reference that this is all open-source but https://github.com/Picovoice/ does not have a Picovoice repository. Is it https://github.com/Picovoice/Porcupine?
Not sure what the plans for open sourcing the rest of the components are though.
1- Speech-to-Intent: It allows you to issue complex voice commands in a specific domain and in turn returns the intent. For example, in the case of a coffee maker you can say "Please may I have a single shot espresso with no milk and two sugars". The engine returns a JSON-like object with {"product": "espresso", "milk": "no", "sugar", "two", "# shots": "2"}. It is a tightly coupled domain-specific speech recognition and NLU. It is small (less than 3MB and 8% CPU usage on RPi3) and ideal for home automation, industrial application, service industry, etc.
2- Speech-to-Text: It is large vocabulary speech recognition software that runs locally. It will support all platforms currently being supported. It allows you to do large vocabulary transcription with high accuracy locally on an embedded platform.
It's apparently only talking about their benchmarking tool: https://github.com/Picovoice/stt-benchmark
The engine itself is called Cheetah, and appears to be a closed source, "inquire for pricing/licensing" product.
[1] http://snips.ai/
With Picovoice you can use the voice control engine to accomplish this. Maybe something similar to this demo?
https://picovoice.ai/#voice-control-demo
The cool thing about this engine is that it is tiny. It uses less than 8% CPU on RPi3 and altogether it is less than 2MB (code, model, etc). Technically you can run it on something much smaller and cheaper than RPi.
Alternatively, the speech-to-intent engine could be a good candidate. More information on this along with an interactive demo will be released this weekend.
We do work with a couple of SoC manufacturers and will disclose some of the results when our partners are ready. In general, we can run on any MCU with a C compiler and 200KB of RAM (maybe less if there is fast FLASH available). We already of models working on ARM Cortex-M and Cadence's HiFi4.
> Runs in real-time with only 5.6 MB of memory and 25% CPU usage on a Raspberry Pi 3.
The voice control comes in two variations standard and tiny. The tiny one consumes even fewer resources. I provided metrics for the standard one. You can check the benchmark repo on benchmark it yourself as well :) https://github.com/Picovoice/wakeword-benchmark
Snips looks like a promising prospect!
"You can turn off your internet connection and it will keep working."
That's nice!
We will add another demo for speech-to-intent module fairly soon (over this weekend). Stay tuned!
However i wonder why there is no way (visible?) to generate words for Javascript? Or at least a documentation on how the format for those byte arrays is build.
Assuming this is a licensing thing, i would really suggest to not put limit in that way. On first impression i assumed that this is useable for free for everything except commerical projects.
I went to the website, and had questions, so I tried to email them. I got a reply to the email telling me to join the Discord. I did, and asked in the Discord, and the only person who bothered to reply told me to read the website. ...I wouldn't have asked the questions if they were answered on the website.
E: You should probably try in the late night or early morning time for the States (EST) since they are located in Europe I believe.
Obviously, you can just grab the audio stream after the phrase is detected and route it to whatever you like. Google ASR or even a local one running on the device!
I have literally 3 projects, one a current big that depend or would heavily benefit from this. Can't wait to give it a try!
1- the business model 2- in some cases, it actually needs some engineering. for example a new brand name, etc.
Pretty impressive tech by the way.