Papers › FlapAI Bird: Training an Agent to Play Flappy Bird Using Reinforcement Learning Techniques

FlapAI Bird: Training an Agent to Play Flappy Bird Using Reinforcement Learning Techniques

21 Mar 2020arXiv:2003.09579archive 2025-07-28

Tai Vu, Leon Tran

Reinforcement learning is one of the most popular approaches for automated game playing. This method allows an agent to estimate the expected utility of its state in order to make optimal actions in an unknown environment. We seek to apply reinforcement learning algorithms to the game Flappy Bird. We implement SARSA and Q-Learning with some modifications such as ϵ-greedy policy, discretization and backward updates. We find that SARSA and Q-Learning outperform the baseline, regularly achieving scores of 1400+, with the highest in-game score of 2069.

PaperPDFCode

Code

taivu1998/FlapAI-Bird officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Q-LearningReinforcement LearningReinforcement Learning (RL)reinforcement-learning

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

Q-LearningSarsa

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections