Does this simply run your java code on the GPU, or does it parallelize your code automatically? The latter would be really cool.
So I guess you still end up writing your algorithm in OpenCL / Cuda and maybe use the serialization provided by this lib.
Update: (Just read the hpcc_rootbeer.pdf slides.) You write your _parralellized_ implementation of an Algorithm in Java - and it will be executed on the GPU.