But those links are Googled after the model started to answer, they are not the links to the training data
Imagine an artificial “librarian” that read all the books and spits hallucinated quotes for you
But doesn’t let you enter the library, open a single book or even see the sources for those hallucinated quotes
But instead Googles some sources based on hallucinations after generating them ;-)
It’s better than nothing but you can Google them, too, while training data (the library) is completely hidden from you, even the public domain parts of it - zero attribution