Stability AI releases StableVicuna, a RLHF LLM Chatbot
stability.ai
stability.ai
In terms of getting work done 30B and under llamas aren't going to be very helpful. Maybe 65B behaves differently but the 30B and under are just amazing toys; not tools. And, again, if you fine-tune them they get significantly stupider. Think of fine tuning as a lobotomy. The model is more likely to stay on the rails but it loses emergent capabilities.
- RWKV (raven 14B)
- GPT-NeoX-Chat-Base (20B)
- Flan-T5-xxl
- Fairseq Dense (13B)
- Pythia Chat Base (7B)
- Codegen (16B)
- Bloomz (7B)
Some are general, some are more specific. But I could get something out of all of them.
- WizardLM was released a few days ago and showed promising performance for 7b, are you going to fine tune it next?
A startup with $100M in funding shouldn't focus on these optimizations.
The ambiguous licensing poisons all usage up the stack...
We will swap the base model out for StableLM as soon as we can and iterate from there. We just thought the community would enjoy this research artifact :)
Would love to see how it performs against ChatGPT though.