I don't get this. It's not like the models are running on their own GPUs.
So if you're running open models on AWS GPUs, you might as well run Claude (which AWS supports, and doesn't share any data with Anthropic).
Same with Azure/OpenAI.
So if you're running open models on AWS GPUs, you might as well run Claude (which AWS supports, and doesn't share any data with Anthropic).
Same with Azure/OpenAI.
I want to believe (pinky promises from terms of service don't count)
One engineer's salary to accelerate a team of twelve is so cheap you can't afford not to.
Open models on-prem is the future, not a single doubt in my mind.
The biggest problem we had with on-prem was maintenance as it took a lot of staff and time to ensure decent reliability.
Ah... oh.
Well, it's a nice thought.