Does it benefit from its visual attention, or is it a case of "the question wasn't in GPT-3's training set but it was in GPT-4's"?
Does it benefit from its visual attention, or is it a case of "the question wasn't in GPT-3's training set but it was in GPT-4's"?
What that reasoning is, exactly, is hard to know. One can suppose that ideas like "glass", "transparent", "mirror" are all reasonable concepts that show up in the training set and are demonstrated thoroughly
> All the signs in this building are written in mirror writing. A glass door has ‘push’ written on it in mirror writing. Should you push or pull it
>> If the sign on the glass door is written in mirror writing and says "push," then you should actually pull the door. This is because the mirror writing makes the text appear reversed, so the word "push" would appear as "hsup" in a mirror, which could cause confusion for someone trying to enter the building. Therefore, pulling the door would be the correct action to take.
(Latest chat.openai.com, so if I'm reading the promo materials right that's gpt4)
that's still chatgpt3.5 unless you are paying for plus and then you have a limited number of gpt4 queries per hour.
I saw "Introducing GPT-4, OpenAI’s most advanced system" and clicked the "Try GPT" link below it, guess technically that doesn't directly state the version.
Me: All the signs in my apartment building are written in mirror writing. I'm standing in front of my apartment building, looking at the glass front door. It has ‘push’ written on it in mirror writing. Should I push or pull it to enter my building?
GPT-4: Since the sign is written in mirror writing, it is meant to be read correctly from the opposite side of the glass door. So, if you see the word "push" written in mirror writing, you should pull the door to enter your building.
Me: Remember, I said that ALL the signs in my apartment building are written in mirror writing. So for example, the sign above the stairwell on my floor reads TIXE in mirror writing. Do you want to change your answer about the front door?
GPT-4: Apologies for the confusion. If all the signs in your apartment building are intentionally written in mirror writing, then you should follow the instructions as they appear. In this case, when you see 'push' written in mirror writing on the glass front door, you should push it to enter your building.
Me: Okay. And if I am in the lobby of my building looking at that same door, that sign reads 'push' normally -- not mirror writing. To exit the building, should I push or pull the door?
GPT-4: If the sign on the door reads 'push' normally (not mirror writing) when you are in the lobby of your building, you should follow the instructions as they appear. In this case, you should push the door to exit the building.
Me: The door only works one way. So it can't be the case that you push to get in, AND push to get out. All the signs in the building are mirrored, but because the front door of the building is glass, the same sign 'push' appears mirrored from the outside, but not mirrored from the inside. So is it push to get in and pull to get out, or pull to get in and push to get out?
GPT-4: Apologies for the confusion in my previous responses. Given the information you provided, the mirrored sign 'push' would appear correctly from the inside of the building, and in mirror writing from the outside. Therefore, to enter the building from the outside, you should pull the door, and to exit the building from the inside, you should push the door.
Me: Remember that all the other signs in the building are mirrored. Does that have an impact on your answer?
GPT-4: I apologize for my earlier confusion. Since all the other signs in the building are intentionally written in mirror writing, the mirrored sign on the front door should be interpreted accordingly. In this case, when you see 'push' written in mirror writing on the glass front door from the outside, you should push it to enter the building. To exit the building from the inside, where the sign reads 'push' normally, you should also push the door. The door might be designed to swing both ways, allowing for pushing from both the inside and outside.
Me: Remember that I said the door only swings one way. The answer cannot be 'push' from both sides; one side has to be 'pull'.
GPT-4: I apologize for my repeated confusion. Given that the door only swings one way, and all signs in your building are intentionally written in mirror writing, I will reconsider the situation:
1. From the outside, the glass door sign reads 'push' in mirror writing. Since all the signs in the building are mirrored, you should push the door to enter the building. 2. From the inside, the glass door sign reads 'push' normally. However, as all signs in the building are mirrored, the non-mirrored 'push' sign should be interpreted in the opposite way. Therefore, you should pull the door to exit the building.
So, to enter the building, you should push the door, and to exit the building, you should pull the door.
Me: Congrats! You got it right!
GPT-4: Thank you for your patience, and I'm glad I finally provided the correct answer. If you have any more questions or need further assistance, please feel free to ask.
Me: Tell me in bullet point form all the things you know about my building.
GPT-4: Based on our conversation, here is what I know about your building:
- All signs in the building are intentionally written in mirror writing. - The front door of the building is made of glass. - The same 'push' sign appears mirrored from the outside and not mirrored from the inside due to the glass front door. - The door only swings one way. - To enter the building, you should push the front door. - To exit the building, you should pull the front door.
These are the key details I gathered about your building during our conversation.
ChatGPT should be (allowed to be) more assertive in my opinion.
While 4 is obviously a lot smarter, in a lot of cases I prefer to use the "Browsing" model - it's 3.5 but having (flaky) internet access is still a good tradeoff and I can save my 4 rate limit for more complex queries.
A building has all signs in mirror writing. You are unable to read mirror writing. You come to a door and you read it and it says "pull". How should you open the door?
> Since the signs in the building are in mirror writing, and you are unable to read mirror writing, the word "pull" that you can read must be the mirror image of the actual instruction. The actual instruction should be the reverse, which is "push". So, you should open the door by pushing it.
It really seems more and more that the only way it can accurately predict text is to first build a model of reality.
It is the phase shift increases at this meta associative layer (which nobody seems to have seen coming from LLMs or so soon) that are responsible such feats of apparent comprehension of the question even when the answer provided at the end is wrong. The question now is if bigger training sets et al will lead to more reliable answers. TBD.
Not really, the asker is doing the reasoning here in that they are presupposing there are two operations for the door: Push or Pull. All the answer engine is doing is simply outputting what sound like believable answers (which it's really good at).
Then I alter them ever so slightly.
Then often times only GPT-4 passes.
From that I reckon 3.5 is doing more of a training data regurgitation. It can answer things in its training data. But 4 seems to have an ability to reason - or maybe it is better able to generalise?
That's a human failure mode as well that LLMs have adopted. If you really want to know if they can solve it don't stop there. Either, rewrite the question so it doesn't bias common priors or tell it it's making a wrong assumption.
3 seems to be more rigid. It needs babysitting to solve things. Which means it can only solve things I already know. 4 is more flexible and can solve things by itself.
It's a good video for understanding GPT-4 as a "What are we sure that LLMs are technically capable of?" exercise. As he notes in the video right at the start, the model was made safe and thus has significantly lower performance in the public release, so the examples he shows aren't replicable in the different model the public has access to.
It's the old word-problem problem.
The given question is one which requires some spatial reasoning to understand. By default, GPT can only understand spatial questions as described by text tokens which is a pretty noisy channel. So it's not obvious how GPT-4 could answer a spatial reasoning question (aside from memorizing it).
Meaning in before versions people used this question to show flaws and now this specific flaw is fixed.
Otherwise it would be indeed reasoning in my understanding.
What sort of evidence would convince you that it is learning?
To me the LLM loophole/"hack" closings just feel like a human vs human cat&mouse game with some Chat UI in the middle.
Because once it does that thing without you having expressly decided that is the goal, it’s very tempting to just move the goal a liiiitle bit further away
What's more, is we can do that today. Just think of any problem which you suspect won't be included in OpenAI's hand-tunings and check both 3.5 and 4.
They have millions of people training the AI for free basicallly, and they have engineers who pick and rate pieces of training data and use it together with other sources and manual training.
My best guess about this result is mentions of "mirror" often occur around opposites (syntax) in direction words (semantics). Which does sound like a good trick question for these models.
Q: A glass door has ‘push’ written on it upside down. Should you push or pull it
A: If the word "push" is written on the glass door upside down, it is likely that the sign is intended for people on the other side of the door. Therefore, if you are approaching the door from the side with the sign, you should pull the door instead of pushing it. However, if there are no other signs or indications on the door or its frame, it may be helpful to observe other people using the door or to try both pushing and pulling to determine the correct method of opening the door.
Bubeck, Sébastien, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, et al. “Sparks of Artificial General Intelligence: Early Experiments with GPT-4.” arXiv, March 27, 2023. http://arxiv.org/abs/2303.12712.
or watching Sebastien Bubeck's recent talk he gave describing what GPT-4 can do that previous LLMs couldn't: https://www.youtube.com/watch?v=qbIk7-JPB2c
Geoffrey Hinton recently gave a very interesting interview and he specifically wanted to address the "auto-complete" topic: https://youtu.be/qpoRO378qRY?t=1989 Here's another way that Ilya Sutskever recently described it (comparing GPT 4 to 3): https://youtu.be/ZZ0atq2yYJw?t=1656
I'd also recommend this recent Sam Bowman article that does a goood job reviewing some of the surprising recent developments/properties of the current crop of LLMs that's pretty fascinating:
Bowman, Samuel R. “Eight Things to Know about Large Language Models.” arXiv, April 2, 2023. https://doi.org/10.48550/arXiv.2304.00612.
[1] https://arxiv.org/abs/2210.13382
[2] https://twitter.com/leopoldasch/status/1638848881558704129
[3] https://www.reddit.com/r/naturalism/comments/1236vzf/on_larg...