Basically there, with what kind of VRAM and processing requirements? I doubt anyone running on a CPU can fine tune in a time frame that doesn't give them an obsolete model when they're done.
This would make fine tuning a qantized 13B model achievable in ~0.3 seconds per training example on a CPU.
Which is my personal holy grail towards making myself unnecessary; it'd be amazing to be doing some light gardening while the bot handles my coworkers ;)
Or it handles their bots ;)
"As a limitation, MeZO takes many steps in order to achieve strong performance."
I completely misread that!