In that case, shouldn't the training data be open source? Seems like somebody got an unfair advantage here.
They intentionally built an unfair advantage so that they could sell it.
Here's[1] a patent they have filed towards the system. Claims 1-18 and 20 are focused on the training of the neural network. Looks like Claims 1-18 are going to be granted soon largely in that form also from looking at PAIR[2].
[1] https://patents.google.com/patent/US20160292856A1/en?q=AI,ar... [2] https://portal.uspto.gov/pair/PublicPair