What is your startup and what research are you relying on? It sounds interesting.
We believe this is a very promising way to generate large volumes of labelled data, which we know is what ML loves. The trick is to structure both the biochemistry and the ML properly so as to avoid artifacts and generate useful molecules reliably.
Happy to talk about it in more detail, give me a shout at ntilmans at anagenex.com!