Your reply contains some weird non-sequiters. Nvidia is right on target when they say there is no "magic" autoparallelizing compiler. The autoparallelization they talk about, and the scheduling issues you talk about, are completely unrelated except that they both contain the word "compiler" :/
As for double-precision performance, it is also well known that the GTX 680 is only for consumers. The HPC version of Kepler is due later this year and I will not be surprised to see it offer much improve double precision performance compared to current Teslas. Intel MIC is not a consumer product, so it is unfair to compare it to consumer products like GTX 680.