I am getting 98.6% cache hit ratio on deepseek-v4-flash with opencode
On the sheer performance it’s comparable to Opus ?
./cost.py amount-2026-5.csv 0.3 3.75 15
input_cache_hit_tokens: 472,971,520 tokens -> $141.8915
input_cache_miss_tokens: 13,299,013 tokens -> $49.8713
output_tokens: 3,334,962 tokens -> $50.0244
cache hit rate: 97.27% (472,971,520/486,270,533)
cache miss rate: 2.73% (13,299,013/486,270,533)
total: $241.7872
All of this usage was with an OpenCode subagent exclusively.Total input token = input + cache read + cache write Cache hit rate = cache read / total input token.
That is 71% in my very limited use of opencode.