To satisfy the experiments vs operations criteria, all they have to do is deploy a software update to the inference library for slightly higher tokens per second, and compare, you know, in an experiment
This is something we all do
pitchforks down, its the reporter and the non technical accountants that are wrong and both groups should be replaced