Deal or no deal? Training AI bots to negotiate
code.facebook.com
code.facebook.com
But I think what is missing, is the time component when negotiating with humans. A negotiation process is usually better for humans if the negotiation is quick and not dragging on too long.
And more importantly the chatbots never seemed to "walk away" from a deal. But in real life, you sometimes have to walk away to show the other party, that you are not a pushover. It would be interesting to enhance the model so that chatbots negotiate repeatedly with each other and "remember" how the other party behaves and how far you can push the other party to concede. Because some negotiations really are zero sum games.
AI consider this human behavior a bug.
Many situations cannot be resolved until you convince the other party that their business model or position or assumptions are wrong. Walking away is the only way to do that because it triggers escalation.
I had s sitstuiin recently where the account team couldn't get a term that we needed. We basically told them to go away and stopped 2-3 other negotiations. That let our counterparty get the resource he needed (SVP of product X) and balance returned to the force.
However, if you zoom out and consider the optimal deal-making strategy over multiple deals, then walking away can be a good strategy. For example, a used car salesman would rationally walk away from a deal if they believed it's likely they can sell a car later for a better price.
If you consider multiple deals then you can also consider the concept of your reputation. This is information that other parties may have about you when they enter a negotiation in the future. You may rationally wish to make a sacrifice on a present deal in order to alter your reputation, to improve your outcomes in future detals.
Imagine a chatbot who can chat up a girl online better than any human. Whose jokes make any human seel dull by comparison. And whose wit is quick to about 1 trillion jokes a second :-P
Edit: Indeed, the paper says that not using the fixed agent trained on human negotiation leads to unintelligible language from the agents.
I am worried when computers start getting better than people at these kinds of things. They already mastered heads-up poker.
Almost all of our systems rely on an inefficiency of an attacker - so they are vulnerable.
Not sure what this would look like, but I'd be interested to find out.