how is Jev cheaper if I can run locally. 0.5% prefill, 0.1% decode, 99.4% cached, latency is <20ms
11 karma · joined November 8, 2020
vs
Press "Tab"
NVIDIA itself is also training foundation models (and open-sourcing them). If there is excess compute available, NVIDIA can increase the scale of such models.
- Training SW [x]
- Inference SW [x]
- Evaluation SW [x]
- Data [x]
Output:
- Weights []
DeepSeek is closed-source with *open-weights*