Papers › Playing Hex and Counter Wargames using Reinforcement Learning and Recurrent Neural Networks

Playing Hex and Counter Wargames using Reinforcement Learning and Recurrent Neural Networks

19 Feb 2025arXiv:2502.13918archive 2025-07-28

Guilherme Palma, Pedro A. Santos, João Dias

Hex and Counter Wargames are adversarial two-player simulations of real military conflicts requiring complex strategic decision-making. Unlike classical board games, these games feature intricate terrain/unit interactions, unit stacking, large maps of varying sizes, and simultaneous move and combat decisions involving hundreds of units. This paper introduces a novel system designed to address the strategic complexity of Hex and Counter Wargames by integrating cutting-edge advancements in Recurrent Neural Networks with AlphaZero, a reliable modern Reinforcement Learning algorithm. The system utilizes a new Neural Network architecture developed from existing research, incorporating innovative state and action representations tailored to these specific game environments. With minimal training, our solution has shown promising results in typical scenarios, demonstrating the ability to generalize across different terrain and tactical situations. Additionally, we explore the system's potential to scale to larger map sizes. The developed system is openly accessible, facilitating continued research and exploration within this challenging domain.

PaperPDFCode

Code

guilherme439/nuzero officialmentioned in paperpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Board GamesDecision Making

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

AlphaZero

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections