I'm really looking for a multi-modal image capable version of Jev.
If we could get machine learning type results on images without training, that would be fantastic.
If we could get machine learning type results on images without training, that would be fantastic.