PathScale hackers reverse-engineer Nvidia GPUs to write a better compiler
hpcwire.com
hpcwire.com
- http://blog.langly.org/2009/02/12/cuda-hacking-ptx-code/ -- Breakdown of some CUDA code compiled down to PTX assembly.
- http://wiki.github.com/laanwj/decuda/ -- Assembler and disassembler for CUDA binaries.
I'm really excited to see what people come up with compiler-wise here. GPGPU stuff is still very new; I can't wait to see what people come up with.
That said, they've got a long road ahead of them. GPU development iterates much more quickly that that of CPUs; there has been an entirely new hardware architecture (Fermi), three of four driver updates, and two major CUDA API revisions released by NVIDIA in the 6 months since I did my independent study on the topic.
That should be easier, given the specs for those are freely available.
Also your perception of what AMD has released isn't entirely correct. The low level details for things around launching a compute program on ATI Evergreen hasn't been made public yet. They've only released the graphics details and we're hoping they release the compute stuff soon. This isn't to say we couldn't figure it out, but we're focused on what customers want.