I wouldn’t even be surprised if a law is passed requiring sites to provide equal access to humans whether accessed directly or via these models.
It’s too important an innovation to stall, especially considering the US’s competitors (China) won’t respect robots.txt either.
They do say "Currently, deep research can access the open web...", so maybe "open" there implies something significant. Like, "websites that have agreements with OpenAI and/or do not enforce norobot policies".
Amazon listings are blocked from google shopping and other price comparison sites.
I see Amazon results there all the time. 3 of the visible 8 sponsored results are Amazon, in the non-sponsored results an Amazon listing is either first or second in every category.
How would you know its a crawler?