I don’t think we should use an AI trained in 5 minutes on a laptop to infer what small models are capable of…
Sure they still have massive problems with hallucination, but this article doesn’t give us any more insight into that I don’t think!
Sure they still have massive problems with hallucination, but this article doesn’t give us any more insight into that I don’t think!
I think the point of most (frontier) small models is usually to provide the best answer possible given small inference resources, rather than to reduce training time.
This is more of a toy model, so fun and an interesting project but it doesn't necessarily tell us what the art of the possible is for small models.