Sounds like a great library to use for automatically testing all the new models being released everyday and finding out if a new open source model significantly performs better on your custom dataset.
1. What's the largest model (number of parameters) that you've tested the library with?
2. Will MoE models work as well? They're known to have more unstable training and need some custom techniques to stabilize