How do you deal with different dataset train/validation/test? How do you measure the degradation of the model? Is there any way to select the metric you target (accuracy, f1-score or any other)?
https://demo.postgresml.org/models/1 https://demo.postgresml.org/models/15
The short term goal would be to expose more metrics from the toolkit.