Machine Learning Models Are Missing Contracts
gradio.app
gradio.app
This is because the data handling and pre-processing nitty gritty is not actually interesting from an academic perspective.
I cannot count the times I went to look at a paper's implementation source code on a benchmark dataset to find it cut-off annotated sequences to a fixed max length, essentially turning it into a different dataset making comparison to previous work invalid.
Good documentation costs time, time academics don't have.