ParentFull threadagnosticmantis·Isn’t the current best practice to train highly over-parametrized models to zero training error? That’d be a global optima, no?Unless we’re talking about the optima of test error.View on HN