Sonnet 5.5 scores just behind Opus 5.5 on Artificial Analysis Intelligence Index
artificialanalysis.ai
artificialanalysis.ai
Human curated benches aren't accurate enough
You go from one prompt to the final solution, whereas for many of us it is about the experience of iterative, multi turn working.