To me this is clearly a skill issue. Several millions of tokens per day is peanuts, even if uncached. gpt-5.5 is $5 per million of input tokens.
Anybody doing things seriously understand how to optimize their workflows for smaller models once they start to lock in processes.