Why aren't modern systems just statically divided up, with known throughput-oriented junk homed on "efficiency cores" and known latency-sensitive tasks isolated on "performance cores", where each class of CPUs can use an optimal thread scheduler for its tasks? I know thread scheduling isn't the only thing users of lowlatency want, but I bet it solves the bulk of interactivity issues for desktop users.