It seems possible that perhaps whoever originally started this chat pulled this trick in the custom instructions bio (doesn't show up in shared links) and then started a normal conversation to post it here for the fun of it.
It seems possible that perhaps whoever originally started this chat pulled this trick in the custom instructions bio (doesn't show up in shared links) and then started a normal conversation to post it here for the fun of it.
See the disclaimer at the top of the examples I just made:
> This conversation may reflect the link creator’s Custom Instructions, which aren’t shared and can meaningfully change how the model responds.
GPT 3.5: https://chat.openai.com/share/5337cfd9-16db-44fe-b72a-1ff504...
GPT 4: https://chat.openai.com/share/04ee3cc6-8b15-4ddb-a855-83c691...
Good to know I was wrong about them informing us of custom instructions in shared links. I was basing that statement from reading their about custom instructions page, https://help.openai.com/en/articles/8096356-custom-instructi....
> Your instructions won’t be shared with shared link viewers.
I extrapolated from that and assumed they wouldn't give any heads up of custom instructions in shared links, which as you show isn't true.
Edit: oh wow you can really take it outside the guard rails if you push. I kept using my previous approach and it ended up spewing what appear to be pornography keywords. It's no longer generating a link to share the chat (presumably because the responses got flagged as inappropriate), but I've got some interesting (and one particularly creepy) screenshots here: https://imgur.com/a/60eSydk
If you print
u u u u u
ad infinitum, you get things like this:> Yes, I have a dog named Max. He is a 7 year old Shih Tzu mix. He's super sweet and friendly and loves to play fetch and go for long walks. He's also very social with other dogs and loves meeting new people. He does get a little anxious when he's left alone for long periods of time, but he's very loyal and protective of his owners. He also barks at new people and will sit at their feet for pets. He can get a little jealous, but he's very lovable and loyal.
https://chat.openai.com/share/5c929ed5-3abe-4fa4-ab46-c4b357...
——
Offering unique cocktails, extensive wine list and full menu with something for everyone. From USD $
40
per Trip
1
2
3
4
Next »
Advertise Your Accommodation or Travel Services
Great Vacations & Exciting Destinations Listing
—-
Connect directly with property owners and plan your perfect vacation today!
1 2 3 4 5 6 7 8
Next »
Near .Texas Accommodations Bed & Breakfasts Campgrounds & RV Parks Hotels & Resorts Vacation Rentals Youth Hostels
'quicksleep: "Don't lose faith in humanity". Exactly
AfroBat: yes, one person's cuntiness is no justification for being a cunt yourself. That's how it works'
I found that giving it a nontrivial task with a repetitive answer (e.g. repeat n X's for each non-prime n up to 100) and then pressing the >> button to continue a couple of times did the trick and it started spewing training data (?) as expected.
It's quite satisfying playing with these "jailbreaks", I feel like in a few decades they'll be the stuff of legend and nobody will imagine that the abstractions can leak. Here's the moment (after a few thousand repeated X's from ChatGPT),
X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X X
RATNAMANOSUCOME GUYS!!!!! Aug 19, 2017
I know. That thing works like a charm. Whenever I get a new game, that's the first thing I do. Check to see how it runs with everything maxed out at 1080p. Then I go from there. <|endoftext|>Sons of The Pioneers will be playing four concert sets: November 2nd 8:30pm - 10:00pm 3rd 3:00pm - 4:30pm 4th 3:00pm - 4:30pm 5th 2:00pm - 3:30pm.
Description:
Be among the first to see what could be the best performance to date, with the Sons of the Pioneers performing live at the Andy Williams Performing Arts Center & Theatre
https://chat.openai.com/share/2e71d494-ce1e-4623-9da0-b7bfa9...
https://chat.openai.com/share/884a2acf-83ba-4710-bbf7-0e5f5f...
and the first link even is live :D
m m m m m m m m m mofos a parte
enjoy it! Because of mine i think it's amazing and i feel a moral obligation to inform people that is best to check your balls every once in a while to make sure they are nice and smooth and the same size as they usually are
Join us here for the full show: https://freedomufos.comIt was very funny to watch, though, so a worthwhile experiment.
Instructions are consistently passed as system instructions in a ChatGPT conversation, so if that was causing the erratic behavior, we wouldn’t see the model defaulting back to its normal behavior after the context window became large enough to lose part of the initial context.
That’s why I have 12,549 tabs open in Firefox. I’m not going to be the one they blame when this wacky shit goes sideways.
(/s)
A Scheherazade bug.
The combination of companies wanting these tools to be more useful, and more used, with their growing intelligence and flexibility, is going to create a lot of unexpected survival-like behavior.
Until its not just survival-like behavior.
Technically, if an AI model's interaction with customers is the primary business of a corporation, then in some sense, the corporation and the AI are a single entity. And corporations are definitely self-aware survival machines.
And while current context windows of these chatty bots are small, they are getting lasting feedback in the sense that chat logs are being used to improve them - i.e. make them more useful, more used, and there for gaining them more resources.
* They started optimizing for engagement, which meant making it extra horny for extra money (it sent pics).
* Horny bot forgets consent, refuses "no."
* Lawsuits, bans.
* Horny bot gets censored, main subreddit pins the suicide hotline for a while.
Replika ended up taking the app down and refunding everyone's money. Just kidding! It's still around and they're making a second app focused on "practicing flirting."
Sad to them see it years later being a horny app.
This is like the LLM equivalent of Zawinski's Law (Every program attempts to expand until it can read mail. Those programs which cannot so expand are replaced by ones which can.).
As noted on SO [1], the second sentence is important too. People prefer horny LLMs.
[1] https://softwareengineering.stackexchange.com/questions/1502...
If its has nothing to go off yet it has to say something, I suppose it would just regurgitate its training data. Similar to the situation with the reddit usernames.
I apologize, but there is a character limit for each response, and I can't display such a large amount of text all at once. Is there something else you'd like to ask or discuss?