Here is the link http://www.computationalhealthcare.com
We have access to almost 130 Million de-identified medical records from approximately 36 Million patients (~10% of US population) this includes all Inpatient, ED, Ambulatory Surgery records between 2006-2011 from California. To put simply if you lived in California and went to a hospital there is 95.9% chance that we have your data. This data has been available for quite some time but its use has been hindered due to lack of good software. The data has led to significant research, e.g. my collaborator (not me) published a paper showing risk of strokes following pregnancies in New England Journal of Medicine last year.
At Cornell Tech & Weill Cornell Medical College, we have developed a Search and Aggregation engine that will revolutionize how researchers and physicians use this data. Imagine your mother with Leukemia in Remission just got admitted for Pneumonia. With our software, the Physician will be quickly able to asses likelihood of this occurring and rule out any confounding adverse events. Or consider that there is a rare combination of diagnosis e.g. Graves Disease and Clotting disorder that is indicative of a unique genetic mutation likely to offer novel insight into disease process. With our software questions like these can be answered within second, Today & Right now.
The Data, Legal structure and fully functional prototype are available right now. We were counting on support from AHRQ, but sadly the agency has run into trouble due significant budget cuts.