This looks exciting! There is a serious dearth of high-quality open-source models with multimodal capabilities. So, really looking forward to playing with this one.
Has anyone here experimented with fine-tuning this for domain-specific applications?
Has anyone here experimented with fine-tuning this for domain-specific applications?
No comments yet.