Mastering the Game of No-Press Diplomacy via Human-Regularized Reinforcement Learning and Planning

Anton Bakhtin*, David J. Wu*, Adam Lerer*, Jonathan Gray*, Athul Paul Jacob*, Gabriele Farina*, Alexander H. Miller, Noam Brown

Bibtex entry

@inproceedings{Bakhtin23:Mastering, title={Mastering the Game of No-Press Diplomacy via Human-Regularized Reinforcement Learning and Planning}, author={Anton Bakhtin and David J. Wu and Adam Lerer and Jonathan Gray and Athul Paul Jacob and Gabriele Farina and Alexander H. Miller and Noam Brown}, booktitle={International Conference on Learning Representations (ICLR)}, year={2023} }

Download

Paper PDF

Bibtex entry

@inproceedings{Bakhtin23:Mastering, title={Mastering the Game of No-Press Diplomacy via Human-Regularized Reinforcement Learning and Planning}, author={Anton Bakhtin and David J. Wu and Adam Lerer and Jonathan Gray and Athul Paul Jacob and Gabriele Farina and Alexander H. Miller and Noam Brown}, booktitle={International Conference on Learning Representations (ICLR)}, year={2023} }

Typo or question?

Get in touch!
gfarina AT mit.edu

Metadata

Venue: International Conference on Learning Representations (ICLR)
Topic: Multi-Agent Reinforcement Learning, Human Modeling, Robustness to Mistakes, and Equilibrium Perfection