The model we are comparing against makes 10X as many errors.
I hadn't imagined someone would argue that's not a meaningful difference.
Though the difference is statistically significant too.
I hadn't imagined someone would argue that's not a meaningful difference.
Though the difference is statistically significant too.
Predictive accuracy is measured on 1000 samples, not 20.