I think the big difference is dimensionality. If the dimensionality is low, then taking account of the 2nd derivatives becomes practical and worthwhile.
In practice, clever optimisation algorithms that use the 2nd derivative won't actually form this matrix.