How can open source models respectful of robots.txt possibly perform equally if they are missing information that the other models have access to?
Don't focus too much on a single variable, especially when all the variables have diminishing returns.
[0] the ultimate, of course, being profit.
I'm not sure if we're thinking of the same field of AI development. I think I'm talking about the super-autocomplete with integrated copy of all of digitalized human knowledge, while you're talking about trying to do (proto-)AGI. Is that it?
You just listed possible options in the order of their relative probability. Human would attempt to use them in exactly that order