Yeah, they say it will be "chinchilla-optimal", which means that it probably will be < 70B, might be actually much less as I've seen some recent work that 20B models are able to compete with GPT task [0] so I guess it might be using it but probably isn't, so there's further room for improvement.
[0]https://www.reddit.com/r/MachineLearning/comments/y4tp4b/r_u...