This is the data gathering phase. When we're able to release a database of these hospital prices with high data quality I think it'll be a pretty big deal, just because it's so much work.
It's hard as hell because of how inconsistently formatted these price sheets are. We'll need to develop a robust ML tool to process all of them in a consistent way, or just put a lot of man hours in. That's something that One Fact is working on. DoltHub's main interest is in producing the source databases, which is just a lot of grind work.
We're crowd sourcing the data collection via a "data bounty." It's like a scavenger hunt where you get paid for the data you input. I designed the table and I'm who reviews the data going in, via pull requests.
Incidentally we do have a hospital price database here (the only open one of its kind) but with mixed data quality. https://www.dolthub.com/repositories/dolthub/hospital-price-...