I assume it's because such large context takes lots of memory, so you might as well have smarter model if you are not gonna fit in small vram anyway
And I think the optimized backends should implement that sliding 16k context soon...
Anyway, point is a huge context really helps certain types of queries, and VRAM usage is reasonable with a 7B model.