… Isn't it possible that it understands the innuendo and is going along with making the joke?
What a wonderful new world.
I would guess the youtubers in question also know this, because you wouldn't ask a joke like this if you didn't know the punchline.
This is from 2018 so presumably it has made it into some LLM training data set by now.
"A Massive Object Devastated Uranus A Long Time Ago And It Never Fully Recovered"
https://www.bgr.com/science/uranus-collision-early-solar-sys...
How many D's are there in Pluto?
There's one D in "Pluto".
How many Fs are there in Mars?
There is 1 F in "Mars".
https://chatgpt.com/share/6aba6fb8-085c-83e8-9d0e-c0eed1e59c...
I guess it could think this is some kind of "give an f" joke but seems like a stretch.
Some future "AI" could be a billion benchmark-hacks and a way to tell which one is needed.
I've got no problem with an AI doing something similar
Byte Latent Transformer - https://arxiv.org/abs/2412.09871
1.1% vs 99.9% on a vanilla vs byte latent transformer on a CUTE Spelling benchmark. Char and Word manipulation benchmarks also saw huge gains.
If it was top priority, every company that can't find a post training fix would go disable half their tokenizer code and it would be solved in the next model.