The slide stated that one Epiphany core is 27 times slower than one Xeon core. clearly the goal here is to ramp..the 324 cores all came online today. If we don't get to >10,000 cores, then the effort has failed.
Will you send out any test load before May 30? Would be better to see those cores to work if they are online, otherwise it feels kinda wasted. Wouldn't be surprised if people go offline after a while if there's no work done.
Another question: how are you planning to implement fault tolerance? If you're running across hundreds of nodes via the internet, the probability of one failing while my job is running is high. Are you going to run a fault-tolerant scheduler?
And: how are you going to do file I/O? Does the user have to run the master MPI process on his/her own machine and do I/O there?