I got it working on my laptop! M2 64GB and needs 38GB of disk space:
uv run --with 'numpy<2.0' --with mlx-vlm python \
-m mlx_vlm.generate \
--model mlx-community/QVQ-72B-Preview-4bit \
--max-tokens 10000 \
--temp 0.0 \
--prompt "describe this" \
--image pelicans-on-bicycles-veo2.jpg
Output of that command here: https://simonwillison.net/2024/Dec/24/qvq/#with-mlx-vlm