I think I came up with this idea in the world where you literally browsed the web. Like go to Yahoo, see what new sites there were, and read 'em ;) That is ... not what we do anymore.
Oddly, this idea seems to share some of the same problems of the "Metaverse" idea. Facebook's Metaverse is a videogame that doesn't realize it's a videogame, it's a centralized service that you're trying to layer on top of the internet, and it ends up being totally empty and with no community, because it's for "everyone", which means it's for nobody. Discord and Reddit both understood the problem and partition themselves into user-moderated interest-oriented spaces; Facebook's idea of a metaverse doesn't do this, and a pan-web comment system sort of can't by design.
Mobile users are the ones by far the most in need of agency. Google saying no here is a top existential threat to their dominance. Hopefully now that Firefox is starting to re-allow webextensions on mobile broadly, pressure can mount. Keep trying folks!
Beyond "can mobile run extensions", I think the main thing missing so far in collaborative web extensions is that few are actually interoperable in any significant way. They're all isolated services. That's helpful for content discovery, but I don't think a centralized service is going to cut it for letting a new depth to the web really form.
I also would say that Hypothesis has been a fairly successful fairly great collaborative web extension. I don't often run into other people commenting on sites or articles or comments, but it happens, and the effort in general is well loved. It somewhat uses standard protocols, but is designed for their service. https://web.hypothes.is/
You might be right that they wouldn't take off even if they would work really well.
But I'd like to think they could :)
With a good UI, a frictionless login, clear personal data separation and a sufficiently widespread browser extension platform available on mobile, this would be awesome...
Maybe Safari would be a good first target with their WebExtension support.
There is of course also the technical problem of dynamic rendering and unpredictable DOM structure.
And the whole problem of reaching a social media user base and providing valuable interaction tools to them.
Especially when considering ads, the main question is how to locate these DOM elements — for example, the filter lists shipped with uBlock origin do a great job here.
I'm wondering if maybe having more flexible locators available could help identify DOM subtrees by their content instead of a CSS selector.
Similar to the ones available in e2e testing frameworks, but maybe even more powerful things could be conceived, or matching snapshots etc...
Could certainly imagine this taking off, but imagination is not reality :)
And if people want to talk about some page that doesn't have comments, that's what Twitter and Reddit, historically, are for. The hard parts are not technical, they're social.