HNHacker News
TopNewBestAskShowJobs

popopanda

11 karma · joined December 27, 2025

submissionscomments
popopanda··on Show HN: LLM Inference Calculator – Estimate VRAM, Latency, and Throughput
Yeah, tokens/s varies a lot based on workload. However, I’ve calibrated the estimator against public benchmarks, so it stays within a 30% error margin!

I'll definitely explore how to estimate vision models next!