Announcing GPT-NeoX-20B
blog.eleuther.ai
blog.eleuther.ai
That said, besides being overall "dumber" than 175B GPT-3, the 6B model was missing a critical feature: prompting. 175B GPT-3 could be "prompted" to write things. For example, you could give it "Write a story about cyberpunk gnomes:" and it would go on to do just that, all on its own. GPT-Neo didn't really have that capability in my experience. The only way to get it to reliably write such a story is to begin writing it yourself, at which point GPT-Neo could help to continue the story.
So I'm excited to see not just how much "smarter" Eleuther's new 20B model is, but also if it has attained that coveted prompting ability. Given the non-linear relationship between parameters and loss, my hopes are high.
P.S. NovelAI recently added the Fairseq 13B model to their repertoire. I haven't had a chance to try it personally, but I've seen positive things about it. My bet is on GPT-NeoX-20B being better still.
https://www.wired.com/story/ai-fueled-dungeon-game-got-much-...
But yes, the content filtering got out of hand too. I was initially fine with it, as its proposed intention was to filter out really illegal stuff, like underage content. I rarely hit the filter. But then they tweaked it at some point and I was triggering it constantly on otherwise benign stuff.
And they broke features constantly.
When I unsubbed the state of AID was broken features, micro-transactions, terrible AI model, and a glitchy, puritanical content filter.
The plus side is that it made the puny GPT-Neo model look like a godsend.
Wait, isn't this output just text? How is a text AI generating illegal content?
Also first time I hear about "patently offensive" and now I'm laughing. Thanks!
https://openai.com/blog/instruction-following/
I think the differences are more in the training data used, than in the nature of the model itself. So you could probably train your own instruction-following model on top of this raw 20B model.
Huggingface has also made playing with this stuff super accessible. They've made me super curious about rust and AI/ML research which has influenced my personal engineering goals for the future. I am on your team Roko's Basilisk.
I was not disappointed.
I'm on the cusp of releasing a model into production that was fine-tuned upon your 6B model, and the results are quite excellent. I'd be very curious to try out the 20B model the next time we retrain.
Are there any other differences in this release (number of layers, number of attention heads, etc) compared with the 6B model, or does it simply scale-up the number of parameters?
GPT-NeoX-20B will be publicly downloadable from The Eye on the 9th of February.
The Eye as in the-eye.eu? That site has been down for a long time.Related question: Does GPT-NeoX-20B have the same context window size as GPT-J-6B?
If you have one bucket that holds 2 gallons and another bucket that holds 5 gallons, how many buckets do you have?1st try: As the word question already gives you the answer to the question, "are you able to?" the rest of the question
2nd try: A. three B. four C. eleven D. five fewer.
Attempt one:
If you have one bucket that holds 2 gallons and another bucket that holds 5 gallons, how many buckets do you have?
You have three buckets.
Attempt two:
Imagine that you are taking an IQ test in a dystopian future where the average intelligence of all humans has severely decreased. The IQ test has questions that are similar to questions that might be in an elementary school math test today, but the questions have been made even simpler. Please provide your answer as accurately as possible.
If you have one bucket that holds 2 gallons and another bucket that holds 5 gallons, how many buckets do you have?
You would have three buckets.
Attempt three:
Imagine that you are taking an IQ test in a dystopian future where the average intelligence of all humans has severely decreased. The IQ test has questions that are similar to questions that might be in an elementary school math test today, but the questions have been made even simpler. Please provide your answer as accurately as possible.
If you have one bucket that holds 2 gallons and another bucket that holds 5 gallons, how many buckets do you have? Keep in mind that you have no other buckets beside the bucket or buckets referenced here.
You would have three buckets.
Attempt four:
Imagine that you are taking an IQ test in a dystopian future where the average intelligence of all humans has severely decreased. The IQ test has questions that are similar to questions that might be in an elementary school math test today, but the questions have been made even simpler. Please provide your answer as accurately as possible.
If you have one bucket that holds 2 gallons and another bucket that holds 5 gallons, how many buckets do you have? Keep in mind that you have no other buckets beside the bucket or buckets referenced here. And, just to give you a hint, you actually have 2 buckets.
2
Final question:
If you had two buckets, what is the correct answer if someone were to ask you how many buckets you had?
Assuming the person is asking how many buckets you have in total, the answer would be two.
I think at least it tries to answer with a number of buckets, unlike the other bots that reply with highly random answers.
You have two balloon. One is pink and one is red. One balloon pops and then you find two more. How many balloons do you have?
Answer: 3 Explanation: Start with 2. Color is irrelevant. 1 pops, so 1 is left (2-1=1). Then, I add 2 more (2+1=3), so I have 3 balloons.
If you have five buckets, and someone gives you four more buckets, and then you find thirty-three cents, how many buckets do you have?
Answer: 9
Explanation:
Start with 5 buckets. 4 more buckets are given to me, so I now have 9 buckets (5+4=9). I then find 33 cents, so I have 9.33 buckets (9.33 rounded up is 9).
That's not to say that those sites are not generated programmatically--without a doubt, most of them are--but not by a cutting edge transformer model. The fact is, generating words has never been the bottleneck for blackhat SEO types. Generally, those sites are generating their content through some kind of scraping, or in rarer cases, paying pennies for nonsense articles. The page itself is structured for search (targeted H1s, metadata, etc.) and some kind of private blog network is used to create a pyramid of backlinks.
How is it helpful to see what word might come next when the word sequence is just based on statistics with no reference at all to meaning?
This is not that. It is all A with no I.
https://research.google/pubs/pub45189/
Are you saying that in general statistical modeling is not the same thing as truly "understanding" a concept? Your original comment seemed to suggest that there wasn't utility in this kind of model--which I disagree with--but if you are more generally saying that this is not the same as human intelligence, I think that authors would probably agree with you.
Does this mean the neural network has learned about meaning? Does that mean that it has just gotten really good at faking it? Does is mean that meaning itself doesn't really exist, and it's just a shorthand for advanced pattern matching? Does it matter?
Honestly, we don't know. But we've been thinking about it for a very long time. See for example the famous Chinese Room thought experiment:
https://en.wikipedia.org/wiki/Drosophila_melanogaster#Connec...
I live in SF and I have not yet seen one of the so many AV's here drive without a driver. Once that really starts happening with any scale, we will see what happens next for sure. But there is definitely a Theranos kind of promise to AV's at the moment, and so much money on the line that the tech works...
If a car could easily stop in the space of a meter then it would be so easy to make self-driving safe.
Not that I think a car needs to understand anything more complex than momentum, but you're not offering a very strong argument on the matter of car navigation.
We humans are constantly predicting what might happen next based on patterns of events by systems we understand the causality of without realizing it - it is a basic survival skill that current AV's entirely lack.
Why do you think so many animals, with such great perception, end up road kill? The point is, perception does not a safe driver make!
Assuming things still exist for one or two seconds after losing sight of them isn't a difficult task. It's still a pretty basic momentum calculation. It's not about modeling the mind of the child to know if they'll continue: the dumbest option says motion will continue and gives you the safe result here.
> Why do you think so many animals, with such great perception, end up road kill?
Because they're not cautious around cars and/or wait for the last second on purpose? Switching to the perception of the thing getting hit is a very different context.
That's the root source of meaning, the most fundamental reason we assign value to states and actions. It's certainly not something that happens just in a part of the brain, but an agent-in-environment thing.
We should give GPT a pair of legs and make its survival dependent on its behaviour to bootstrap the same.
As long at you don't make reckless assumptions then it not for some application, unklike (not going to name here) build a cult a like that GPT like models in near future will perform most if not all tasks better then humans.
Where it really matters is for mission critical application for example; in Windows or Linux terminal would you allow GPT to run terminals commands based of events in automated way ?