OpenAI Lifeboat – Proxy OpenAI Code to Llama 70B on Replicate
lifeboat.replicate.dev
lifeboat.replicate.dev
If you have paying customers who expect GPT-4-level quality, is using LLaMa 70B to generate responses really better than just telling users that the OpenAI API is temporarily down, from a reputational standpoint?
It's not as smart as GPT-4, but you can fine-tune Llama to do things that aren't possible with larger models (https://replicate.com/blog/fine-tune-llama-2).