Edit: from the other HN article on this topic:
> We trained this system on publicly available data consisting of ~170,000 protein structures from the protein data bank together with large databases containing protein sequences of unknown structure. It uses approximately 128 TPUv3 cores (roughly equivalent to ~100-200 GPUs) run over a few weeks
https://deepmind.com/blog/article/alphafold-a-solution-to-a-...