The way it works is you first give it a few sample lines, and those do include both the Q and A parts, and then GPT-3 generates consistent output based on the pattern of your samples. But all the answers/monologs in my comment were 100% GPT-3!
Here are some more fun examples that I posted on my Twitter: https://twitter.com/blixt/status/1285274259170955265
The fact that it can take vague English and turn it into fully functional code or shell commands is quite amazing. It's not just that it finds the correct result from its training data, it can also tweak parts of the output to correctly match your free text query.
Were they the first output, or did you run it multiple times? (And if so, about how many?)
It is indeed pretty interesting to read.