My email is richard at our website's domain if you'd like to get in touch!
Trivial issue because it saves me A LOT of time, but it could be an issue for new people testing it.
I would love to test this approach. Are you guys fine tuning for each codebase?
Their cost is $0.7 per 1M token.
DeepSeek is $0.14 / 1M tokens ( cache miss)
1. Data is used for training
2. Context window is rather small and doesn't fit as well large codebase
I keep saying this over and over in all the content I create, the valu of coding with AI will come from working on big, complex, legacy codebases. Not from flashy demo where you create a to-do app.
For that you need solid models with big context and private inference.
https://api-docs.deepseek.com/quick_start/pricing
Running it locally is quite a bit beyond the scope of being productive while coding with AI.
Beside that 128k is still significantly less than Claude
[1] https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct
Whenever using a model to be more effective as a developer I don't particularly care if the model is open source or closed source.
I would love to use open source models as well, but the convenience to just plug an API against some endpoints in unbeatable.