HNHacker News
TopNewBestAskShowJobs

mcptokensaver

30 karma · joined July 31, 2026

submissionscomments
mcptokensaver··on Show HN: Mcptoon – Token-efficient MCP CLI client
The format preserves all fields — name, description, and input schema are all there, just encoded with pipes instead of braces and quotes. It's lossless, not a truncation. I should have made that clearer in the post.
mcptokensaver··on Show HN: Mcptoon – Token-efficient MCP CLI client
Fair point. I used tiktoken (cl100k_base) for all measurements. The 2,034 token count is from the actual JSON tool listing returned by 5 MCP servers (filesystem, memory, sequential-thinking, sqlite, time). The benchmark script is in the repo under /benchmarks if anyone wants to verify.
mcptokensaver··on Show HN: Mcptoon – Token-efficient MCP CLI client
Good point - Claude Code does defer tool loading when definitions exceed 10% of context. That helps a lot.

But they are solving different problems. Deferred loading is "don't load tools until you need them." mcptoon is "when you do load them, the listing is 5x smaller." They are complementary - you can defer loading AND compress what gets loaded.

The scenario where mcptoon helps most is when you actually need all your tools loaded (e.g., a coding session where the agent might call any of 96 tools). Claude Code's deferral would not kick in if you are actively using tools from all 5 servers.