Show HN: AI companions stack – create and host your own AI companions
github.com
github.com
That said, I suspect summarization, translation, etc will take time. I’d suspect under 1 year.
Like which one?
And there are some very new 65b finetunes I have not tried.
Huggingface is full of finetunes now, and I believe a 33b model can be finetuned on a single 3090.
Llama.cpp is developing some kind of training, but I have no idea what the requirements will be.
Maybe a year before we are the level of stablediffusion?
If you're looking to self-host chat memory rather than go all in on Supabase, there's Zep: https://github.com/getzep/zep
Full disclosure: I'm a co-author.
We wouldn’t bat an eye at using S3, EC2, and RDS as a host your own setup. The only difference here is that startups are moving faster than incumbents.
FWIW that’s one reason why Steamship (disclaimer: I’m the founder) aggregates all AI services under a single API key and interface. It’s to deal with the insane glue-code hassle of running this stuff on your own.
13b will work on 16GB RAM, and 33b on 32GB RAM, with pretty much any dGPU for a little acceleration and RAM offloading.
Doubly so if you host it as an AI Horde node (so you have priority access to many models through the web browser).
P.s nobody will *sms the companion
But there's still a good argument for a hybrid solution. Buy GPT4 access through the API and get a native UI to query it. Much cheaper to pay as you go, and someone else is still handling the heavy lifting. But if you want an uncensored model, you're out of luck.
Anything going through the API on the other hand has a commitment to not do this and to purge the history after a month.
We’re in the nascent stages but I think there will probably always be a community of folks who want to add more nuance to the communication, whether it’s reveries that enact a mood or goal, tie-ins to other services, etc.
Eg imagine wanting to have your ChatGPT DnD master also keep some kind of score. It may be ultimately easiest to put a wrapper around a themed GPT window that imposes a predictable way to do that rather than require everyone to figure out how to prompt it correctly.
Probably also tiny corp for self hosting: https://geohot.github.io/blog/jekyll/update/2023/05/24/the-t...
I think there's a tremendous concern about letting data exfiltrate to any AI SaaS offering among business execs. If you could offer an experience minus the cloud that's easy to use and has compliance/logging features, I think you'd find success from industries that are reticent sharing customer data(e.g. Banking, government) who benefit greatly from the NLP workflows that LLMs enable
> Shortcomings
> Oh, there are so many.
Smart. A VC firm that heavily invests in software platforms will make much better decisions if they have first hand experience using the products.