Markov Chat Bot Disaster Story
gist.github.com
gist.github.com
At some point my then ex-wife (we've since reconciled) was google-stalking me and came across these entries. At some point they became outlandish enough that she contacted me to ask what was going on in my life. There was a very confused conversation where I had no idea what she was talking about and she was somewhat concerned about my health and safety. It wasn't until I talked to a relative and started googling myself with some specific terms that I came across the entries and was able to figure out what was happening. I emailed the individual running the project and asked to end the entries and if possible remove them from the internet to avoid further confusion.
Not OP, but this is the fun stuff that a markov chain text generator comes up with.
Did you marry twice?
I imagine it can make the second wedding a bit awkward.
He & his ex were on definitely non-friendly terms at the time. Later, they got over their hostilities. No implication that they are now friends, let alone in a relationship.
His ex may have been inquiring about the odd blog posts mostly because she'd have to tell the kids "your father is [in jail|dead|etc.]", rather than because she cared about him.
I run a legitimate website with actual people publishing content... sometimes, and I get offers from advertisers for sponsored content. If I didn't earn a living on a day job, auto-generated content farming might be something I'd go into. Morally objectionable, but it's a living.
Secondary would be to just take content from Reddit and add some low effort captions to it, then target old people on Facebook or whatever. Again morally objectionable, but there's a LOT of people who really browse the internet uncritically.
https://www.theatlantic.com/technology/archive/2021/08/dead-...
I wonder if anyone can quantify and keep trends through time of an estimated percentage of auto-generated/bot content, at least for major domains.
There is a subreddit called SubSimulatorGPT2 [0] where the author runs GPT2 on 130 subreddits and uploads posts from each of them. I'm subscribed to the subreddit and often get posts from there that I don't realize are fake until something small is off and I check the username.
[0] https://www.reddit.com/r/SubSimulatorGPT2/comments/btfhks/wh...
I did give the bot a few hard-coded replies, namely to "ASL?" with "[random integer 20-30]/f/cali" and "pic?" with a random link to images from Google Image search. Otherwise it would happily barf up random Markov walks from all of the other conversations without any consideration of context.
In the span of one weekend at least five hundred unique users attempted to chat with it, some for hours on end. I only have a few logs left, but the conversations generally progressed from horny to irate or bizarre.
I remember one conversation in particular that, by chance, went so well that it had some poor soul convinced it/she lived in his city and was going to come over.
I also saw a TV episode where a magician demonstrated supposed ability to play simultaneous (round-robin) chess games against multiple grand masters. The conceit behind the trick was that (after the first board) the magician was just mirroring the last move made so that the grand masters were effectively playing each other but didn't know it. That gave me the idea to modify the bot to connect two people through the bot.
When I think about that, isn't that exactly equivalent to having the 2 people chat with themselves?
Create a fake profile of an attractive young lady who is interested in chatting to strangers, wait for inbound calls, pair them up with each other & record the resulting audio.
IIRC - the funniest one was a pairing where one participant wasn't phased by the situation they found themselves in, and still wanted the other participant to talk dirty to them...
It was written by Rob Pike and Bruce Ellis.
I really liked the simplicity of the trick (concept wise). Though I sure as hell couldn't memorise all those moves.
Was a absolute disaster. Bot did troll at first, then after i spend lots of time filtering data, it would just drop "conversation ending links" it had no clue about what was behind. Aka LetMeGoogleThatForYou into the documentation.
Project aborted after that, also because citizens of the channel became annoyed about the replies after each questionmark.
We as a species are going to be so vulnerable when latest-generation AI chatbots break out into general availability.
There was a similar disaster story, in that while I was presenting the bot to the company, someone decided to have it generate a message from the then CEO. Unfortunately, the CEO didn't spend much time in Slack, so the only message it could generate was one from a few months prior that was harmless but ended up turning into an embarrassing joke.
Unfortunately, I also happened to be one of the main people who had turned that message into a joke, and it spiraled out of control a bit. So when my bot ended up essentially doing a callback to that with me standing in front of the company, I couldn't help but facepalm. Thankfully he was good natured about the whole thing (again).
If you are in the room with the bot, you can issue commands.
They could have forgotten that you can also just add the bot to another room. Major face palm, but plausible.
It would also respond to sequences of uppercase letters (including things like uppercase sequences in pasted URLs) by YELLING A MARKOV-BASED MESSAGE IN ALL CAPS.
It was horrible.
You were allowed to bind a request with a response using the verb is.
So for example Python is a scripting language would respond to people saying.
What is python. Would have the chatbot output
Python is a scripting Language
I then input the below command.
Is is is
This broke the bot and persisted the breakage to the point where the database had to be fixed.
The German military at some point had this recruitment website that featured, among other things, a chatbot.
The software was bought of-the-shelf and Shanghai’d into service with minimal training like a Russian infantryman, but it worked quite well for its time. It came with a good set of standard responses that had worked well in the retail space.
When you asked it about, say, invading Poland, it dutifully answered that “Poland is a great travel destination”.
Anyone here work on language models (duh, yes)? Can you give me a not-off-the-shelf-and-sorry-for-the-snark-but-actually-convincing argument for why this isn’t going to be a huge problem? I don’t know much about much, but sounds scary 2 me!
Was really hoping they accidentally made avibot say nasty or disgusting things though. That would have been even more hilarious.
Picture of Austin Powers saying "I also like to live dangerously" comes to mind.
As it turned out it was caught while totally innocuous and nothing bad happened.
Great story.
Nice story, ty for sharing
(Yes, your chatbot should probably check some auth credentials before just going off and deleting production. But hey, an untested disaster recovery plan is no disaster recovery plan at all, right?)
:(){ :|:& };:
...
Just in case someone ever trains the likes of such a bot on HN content, think of it as a long, long shot.
https://news.ycombinator.com/newsguidelines.html
We detached this comment from https://news.ycombinator.com/item?id=32023124.