I see 6 ways to improve foundation LLMs other than cost. If your product is best at one of the below, and has parity at the other 5 items, then customers will switch. I'm currently using GPT-4-8k. I regularly run into the context limit. If Claude-100K is close enough on "intelligence" then I will switch.
Six Dimensions to Compare Foundation LLMs:
1. Smarter models
2. Larger context windows
3. More input and output modes
4. Lower time to first response token and to full response
5. Easier prompting
6. Integrations