> It’s an illusion, folk. You’re being played.
How are they "being played" if Claude 5 isn't even out yet
How are they "being played" if Claude 5 isn't even out yet
There are plenty of ways to reduce inference cost for a high-intelligence model. Making sparser weights, for example, can increase the parameter count while reducing the inference cost and time.
Let’s see what happens :)
Claude 3 Opus: $15.00 (Input) / $75.00 (Output) per 1M tokens
Claude 4 Opus: $15.00 (Input) / $75.00 (Output) per 1M tokens
Claude 4.1 Opus: $15.00 (Input) / $75.00 (Output) per 1M tokens
Claude 4.5 Opus: $5.00 (Input) / $25.00 (Output) per 1M tokens