I don't mean to be negative but - has anyone at all found a use for any of the 7B models? And are you only using it because you're resource constrained? Or is it a purely academic thing for the moment?
This model is all over the place in the benchmarks, it looks roughly comparable to llama, and I know it's open to use unlike llama. But I haven't seen a model under 30B consistent enough to actually use for anything myself.
Hopefully the improved dataset will improve the next version, I truly want small models to work well, I'm sure it time it will improve.
EDIT: I mean this isn't even a trick question: "There are two shapes each a different color, one shape is red and one shape is blue. Shape A is red. Shape B must be what color?"
> This logic puzzle is based on a classic riddle that has been passed down for generations. The answer, of course, is that Shape B must be a red shape, as red is the only color that can be used to represent all reds.
> The key to solving the puzzle is to recognize that red can be used to describe color as well as blood, which is another word for shape.
What ????
EDIT 2: GPT3.5-turbo says "there's no definitive color for shape B"
Only GPT4 gets this right, "Since shape A is red, and the other shape is blue, shape B must be blue."
I have no idea how they struggle on this.