The Jev alternative, laya is only 334m parameters and you can post train it as the workflow is open source. The downside is that laya is only 1024 tokens, and Jev is ~32,000 tokens. I think for many tasks 32k tokens might be too small. You could probably design a task complexity router with something like qwen 0.8b with similar speed and a much larger context window (260k token). As always, "it depends". 0.8B is almost as fast as Jev without any of the limitations.