I had a hunch that opus 4.7 hedged more than other models - and it turns out it's true
model total_claims hedged_count hedged_pct
claude-opus-4-7 1000 451 45.1
sonar-pro 1000 391 39.1
gpt-5.4 1000 277 27.7
gemini-3-retrieval 1000 129 12.9
gemini-3-pro 1000 60 6.0
datasette query herehttps://lite.datasette.io/?csv=https%3A%2F%2Fstatic.simonwil...