Is it considerably more cost effective than cline+sonnet api calls with caching and diff edits?
Same context length and throughput limits?
Anecdotally I find gpt4.1 (and mini) were pretty good at those agentic programming tasks but the lack of token caching made the costs blow up with long context.