The current leader uses neural networks (LSTMs I believe) so that's obviously a good strategy. But since the size of the decoder gets counted, vanilla tensorflow is probably a bad choice.
you can submit your compressor compressed too, if tensorflow networks lend themselves to compression.
For now GPU is prohibited, as high performance GPGPU capable of ML is not that widespread.