There's no reason a running service couldn't also opensource its back end code. It does provide an avenue for people to tinker, potentially improve, and self-host if they have the resources available or the service itself goes under.
There are also a lot of open projects that distribute the work of handling tweaks to parsers and source lists when the upstreams drift.
Even globally I can't imagine the minutely data from reporting stations everywhere being "big data", maybe in the 10s of gigabytes a day (very rough guess, I haven't looked into this specifically)? I'd still be willing to bet API / web requests dwarf the processing and bandwidth requirements of the raw data for a public service like this.
That probably does put it pretty reasonably in the realm of self-hosting if you put a threshold on how much historical data you want to keep and geofence the region you care about.