The python project is https://github.com/ml-explore/mlx and the converted project is https://github.com/frost-beta/node-mlx
I wrote a long prompt: https://github.com/frost-beta/node-mlx/blob/main/tests/promp...
The first result was almost always bad, but after manually modifying the assistant's answer, following generation usually went much better.