Apple Project Titan paper uses Deep RL and self-play for multi-agent negotiation | Hacker News Reader