> "Google, similar to OpenAI, didn’t provide a lot of the technical details about how it trained this next-gen model, including parameter counts (PaLM 2 is a 540-billion parameter model, for what it’s worth). The only technical details Google provided here are that PaLM 2 was built on top of Google’s latest JAX and TPU v4 infrastructure."
I'm sad but not really surprised that these companies aren't publishing and bragging about all of the technical details of their model architecture, size, and training anymore.