Reading the description of the submission process is pretty interesting: http://predictioncenter.org/casp13/doc/CASP13_Abstracts.pdf. There were multiple submissions and manual synthesis of different model results in some cases.
I’m curious how this competition works exactly — it seems like a set of label predictions are submitted and some form of accuracy result feedback is provided (a single accuracy score for the whole prediction set?). And that there are a certain number of allowed submissions ...? How much of the ultimate strategy for playing this game at a high-level ends up being around optimizing for receiving as much leaked information from the test set as possible — is best guess at this point that this result is likely to be a good indicator of a true increase in prediction capability ...?