It's not good.
For example, Wikipedia licenses all encyclopedic text under CC BY-SA 3.0[1] and GFDL[2], both of which permit free-as-in-libre redistribution.
[1]: https://en.wikipedia.org/wiki/Wikipedia:Text_of_the_Creative...
[2]: https://en.wikipedia.org/wiki/Wikipedia:Text_of_the_GNU_Free...
We really need to stop muddying the water surrounding "AI" and their use of source materials. Those so-called "AIs" are just software like any other program, they are tools like any other computer application, and the copyright and licensing legal precedences that apply to software apply to them.
Obligatory IANAL.
If you use any copyrighted source material without appropriate licensing (note: Fair Use is a form of licensing) in any way, that's copyright infringement. This is very simple and the legal world has demonstrated that fact time and time again long before computers ever came onto the scene.
Some "AI" drawing a picture derived from copyrighted materials is no different from a game using copyrighted assets from another game. If there is no appropriate licensing, it's copyright infringement.
Obviously it won’t have anything post 2021 but for some subjects it’s fine.
"As a language model, I don't have the knowledge to answer your question 'who won 2024 world cup?'. These are the search result from Bing engine: ..."
Then you never need to actually access google.com or bing.com.