> "We developed this model using a combination of an improved base LLM (PaLM 2 [4]), medical domain-specific finetuning and a novel prompting strategy that enabled improved medical reasoning."
> "The PaLM 2 pre-training corpus is composed of a diverse set of sources: web documents, books, code, mathematics, and conversational data. The pre-training corpus is significantly larger than the corpus used to train PaLM"
It sparks joy in my heart that this AI doctor, who answers medical questions better than actual physicians, is almost certainly trained on 4chan greentexts and shitposts.
> "Interestingly, we see a drop in performance between GPT-4-base and the aligned (production) GPT-4 model on these multiple-choice benchmark"
This is not interesting if you know about GPT-4. When they lobotomized it, it became worse at everything except refusing to answer some questions.