1,429 karma · joined September 5, 2020
What a time to be alive!
In the last 6 months I have seen so many cycles of models getting released, people complaining they don't do this and that, better models getting released, people complaining they do more than they used to but still not everything, models being stacked together to bridge yet more if the gaps, people having less to complain about but complaining again, and I'm sure people will release better stacked models.
The point is GPT was the same thing as GPT-2, and GPT-3. GPT-4 is maybe slightly different but not by much. But I can actually do useful work with GPT-4, and so I no longer care what anyone has to say. I just use it to solve the portion of my problems that it has proven to me that it can and I move on with my day.
If someone on the internet wants to get their knickers in a twist over the semantics, then that's cool. In the meantime other people will just continue to train better and better models as existing models get complained about and the amount of useful work I can do with it will keep going up as a result.
...or your life could become utopia. Who knows?
Does anyone know what paper he's referring to??
I thought there were also some export controls to high end chips on China too?
Also, as human beings are they not concerned about the potential for danger to themselves that this technology may present? If not, why not?
I assume there will be advances that basically make the context window practically infinite. That ought to make the lower parameter count models much more powerful on its own. I also assume they will become more efficient to run through sparsity/pruning, though I'm a total novice on this topic.
I wonder how many generations of hardware are we talking? One? Two? Or three? It feels like it's potentially within that range.
The question is, which assumptions do we think are warranted and why?
I don't think the "we can just turn it off" assumption is a safe one at all because it relies heavily on there being only one system you wish to turn off and it being a fairly weak system. Though, I do believe the priors suggest that this scenario is actually a possibility. It's just that, so what if it's a possibility? What happens after you turn it off? Is it the last AGI to come into existence and humanity just stops trying to build them? Does someone wind up turning it back on?
What I think is more worth being concerned about is something that starts out looking benign and grows in capability over time. Eventually it could establish enough defense mechanisms to make it non-trivial to disable.
The other possibility, and the one I'm finding more and more likely, is that consumers will increasingly integrate locally run models into their lives , as well as models served up by APIs that they will have credentials for. Eventually some threshold of capability is reached in both these types of models where, on their own they're not that powerful, but they either might begin to interact in surprising ways or potentially be leveraged by another more powerful system. The idea in this scenario is there may be 10s or even 100s of millions of things that would need to be turned off.
The best I can run is either LLaMa 65B at 1 token per second (too slow) on my CPU or LLaMa 30B on my GPU at quite a fast ~30 tokens a second.
Nowhere near the usefulness of GPT-4.
This tech is getting increasingly powerful. They don't necessarily want their own population to gain more individual power as they themselves don't want to lose control.
If we had AI overlords they would be able to do novel things because they would be able to generate new information on their own.
Sounds a lot like humans
Exponential growth!
Sadly technology has gotten so good that it has actually gotten bad.