570 karma · joined February 22, 2015
The practical difference between the licenses is small enough that I expect most people (including me) will choose Llama2 anyway, because the models are higher quality. But that incentive may mean that we get stuck with these awkward pseudo-open licenses.
It's a shame that, as the article mentions, Carlos's music is currently inconvenient to listen to due to unavailability on streaming services.
This seems like a reasonable balance between having the internet be usable while also supporting the services I use.
One thing he didn't mention though is that there's potentially a bit of a trick to get the fine-tuning datasets to transfer across models anyway. (I haven't tested it.)
The key idea is to eliminate the pronouns. Imagine asking GPT-4, not whether it knows a fact, but whether "gpt-3.5-turbo" knows, or "text-davinci-003" knows, etc. Then, when you want the model to reply using pronouns (e.g. "I don't know"), use the system message to tell it which model it is.
This doesn't benefit from introspection, so quite possibly it doesn't work. The reason it might work anyway, though, is that estimating the difficulty of a question might be possible even without introspection.
This analogy puzzled me. An abacus does not strike me as statistical in the way it functions. And minds do seem to be statistical, as far as I know.
> [Wozniak] was overjoyed when he learned that the skill that put so much effort into suddenly became massively easier. He wasn't worried about not being able to earn an above-average salary from this anymore, he was happy about all the new cool things he and everyone else would be able to build.
If you want to get things done, then having your skills obsoleted is good.
It seems, for example, that (by 3.1.12) if you are a person who is involved in the mining of minerals (of any sort), that you are not allowed to use this library, even if you're not using the library for any mining-related purpose.
Edit: I just tried it with a single task of my own (that I've successfully used with ChatGPT and Bing) and it flubbed it horribly, so this model at least is noticeably inferior to the SOTA, which is not surprising given how small it is.