Those samples are from the large model (GPT-2)! Regarding memorization vs. generalization, see our paper for more analysis.
For example, the generated text about the Civil War mentions that Thomas Jefferson Randolph [0] was named after his grandfather, the president. But is the wording mostly influenced by articles talking about that specific fact, or does it draw from more general examples of someone being named after their grandfather?
From a safety perspective, it would be useful to see what prompt a piece of text might have been generated with...