Browse State-of-the-Art › Atari Games › Papers, page 4
Atari Games
Papers archive 2025-07-28
archive papers tagged: 625 · with a code link: 313 · where Syntology ran a sample: 118 (97 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (118 of 625 tagged: 97 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 4 of 7: papers 301 to 400 of 625, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
8 Apr 2017 1 repository listed
-
7 Apr 2017 1 repository listed
-
2 Mar 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
7 Nov 2016 1 repository listed
-
5 Nov 2016 1 repository listed
-
18 Jul 2016 1 repository listed
-
6 Jun 2016 1 repository listed
-
13 Dec 2015 1 repository listed
-
4 Dec 2015 1 repository listed
-
31 Jul 2015 1 repository listed
-
3 Jul 2015 1 repository listed
-
A Principled Path to Fitted Distributional Evaluation24 Jun 2025 0 repositories listed
-
Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons3 Jun 2025 0 repositories listed
-
Automatic Reward Shaping from Confounded Offline Data16 May 2025 0 repositories listed
-
SwitchMT: An Adaptive Context Switching Methodology for Scalable Multi-Task Learning in Intelligent Autonomous Agents18 Apr 2025 0 repositories listed
-
Look Before Leap: Look-Ahead Planning with Uncertainty in Reinforcement Learning26 Mar 2025 0 repositories listed
-
Adventurer: Exploration with BiGAN for Deep Reinforcement Learning24 Mar 2025 0 repositories listed
-
APF+: Boosting adaptive-potential function reinforcement learning methods with a W-shaped network for high-dimensional games17 Mar 2025 0 repositories listed
-
Target Return Optimizer for Multi-Game Decision Transformer4 Mar 2025 0 repositories listed
-
Learning To Explore With Predictive World Model Via Self-Supervised Learning18 Feb 2025 0 repositories listed
-
Reinforcement Learning in Strategy-Based and Atari Games: A Review of Google DeepMinds Innovations14 Feb 2025 0 repositories listed
-
Activation by Interval-wise Dropout: A Simple Way to Prevent Neural Networks from Plasticity Loss3 Feb 2025 0 repositories listed
-
Objects matter: object-centric world models improve reinforcement learning in visually complex environments27 Jan 2025 0 repositories listed
-
Group-Agent Reinforcement Learning with Heterogeneous Agents21 Jan 2025 0 repositories listed
-
MTSpark: Enabling Multi-Task Learning with Spiking Neural Networks for Generalist Agents6 Dec 2024 0 repositories listed
-
From Code to Play: Benchmarking Program Search for Games Using Large Language Models5 Dec 2024 0 repositories listed
-
Why the Agent Made that Decision: Explaining Deep Reinforcement Learning with Vision Masks25 Nov 2024 0 repositories listed
-
Interpreting the Learned Model in MuZero Planning7 Nov 2024 0 repositories listed
-
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning27 Oct 2024 0 repositories listed
-
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents18 Oct 2024 0 repositories listed
-
Transforming Game Play: A Comparative Study of DCQN and DTQN Architectures in Reinforcement Learning14 Oct 2024 0 repositories listed
-
Reinforcement Learning From Imperfect Corrective Actions And Proxy Rewards8 Oct 2024 0 repositories listed
-
Atari-GPT: Benchmarking Multimodal Large Language Models as Low-Level Policies in Atari Games28 Aug 2024 0 repositories listed
-
Towards Generalizable Reinforcement Learning via Causality-Guided Self-Adaptive Representations30 Jul 2024 0 repositories listed
-
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods18 Jul 2024 0 repositories listed
-
Normalization and effective learning rates in reinforcement learning1 Jul 2024 0 repositories listed
-
Understanding and Diagnosing Deep Reinforcement Learning23 Jun 2024 0 repositories listed
-
Utilizing Maximum Mean Discrepancy Barycenter for Propagating the Uncertainty of Value Functions in Reinforcement Learning31 Mar 2024 0 repositories listed
-
Stop Regressing: Training Value Functions via Classification for Scalable Deep RL6 Mar 2024 0 repositories listed
-
Iterated Q-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning4 Mar 2024 0 repositories listed
-
Disentangling the Causes of Plasticity Loss in Neural Networks29 Feb 2024 0 repositories listed
-
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach1 Feb 2024 0 repositories listed
-
Neural Policy Style Transfer1 Feb 2024 0 repositories listed
-
The Indoor-Training Effect: unexpected gains from distribution shifts in the transition function29 Jan 2024 0 repositories listed
-
Visual Encoders for Data-Efficient Imitation Learning in Modern Video Games4 Dec 2023 0 repositories listed
-
Agent-Aware Training for Agent-Agnostic Action Advising in Deep Reinforcement Learning28 Nov 2023 0 repositories listed
-
Minimax Exploiter: A Data Efficient Approach for Competitive Self-Play28 Nov 2023 0 repositories listed
-
Evaluating Pretrained models for Deployable Lifelong Learning22 Nov 2023 0 repositories listed
-
From "What" to "When" -- a Spiking Neural Network Predicting Rare Events and Time to their Occurrence9 Nov 2023 0 repositories listed
-
DSAC-C: Constrained Maximum Entropy for Robust Discrete Soft-Actor Critic26 Oct 2023 0 repositories listed
-
26 Oct 2023 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Towards Control-Centric Representations in Reinforcement Learning from Images25 Oct 2023 0 repositories listed
-
Learning Actions and Control of Focus of Attention with a Log-Polar-like Sensor22 Sep 2023 0 repositories listed
-
Soft Decomposed Policy-Critic: Bridging the Gap for Effective Continuous Control with Discrete RL20 Aug 2023 0 repositories listed
-
Bag of Policies for Distributional Deep Exploration3 Aug 2023 0 repositories listed
-
Elastic Decision Transformer5 Jul 2023 0 repositories listed
-
Action Q-Transformer: Visual Explanation in Deep Reinforcement Learning with Encoder-Decoder Model using Action Query24 Jun 2023 0 repositories listed
-
The RL Perceptron: Generalisation Dynamics of Policy Learning in High Dimensions17 Jun 2023 0 repositories listed
-
Vanishing Bias Heuristic-guided Reinforcement Learning Algorithm17 Jun 2023 0 repositories listed
-
Detecting Adversarial Directions in Deep Reinforcement Learning to Make Robust Decisions9 Jun 2023 0 repositories listed
-
Successor-Predecessor Intrinsic Exploration24 May 2023 0 repositories listed
-
9 May 2023 0 repositories listed
-
Unlocking the Power of Representations in Long-term Novelty-based Exploration2 May 2023 0 repositories listed
-
Approximate Shielding of Atari Agents for Safe Exploration21 Apr 2023 0 repositories listed
-
Loss of Plasticity in Continual Deep Reinforcement Learning13 Mar 2023 0 repositories listed
-
Double A3C: Deep Reinforcement Learning on OpenAI Gym Games4 Mar 2023 0 repositories listed
-
Understanding plasticity in neural networks2 Mar 2023 0 repositories listed
-
Read and Reap the Rewards: Learning to Play Atari with the Help of Instruction Manuals9 Feb 2023 0 repositories listed
-
Enabling surrogate-assisted evolutionary reinforcement learning via policy embedding31 Jan 2023 0 repositories listed
-
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence27 Jan 2023 0 repositories listed
-
Multi-compartment Neuron and Population Encoding Powered Spiking Neural Network for Deep Distributional Reinforcement Learning18 Jan 2023 0 repositories listed
-
Local-Guided Global: Paired Similarity Representation for Visual Reinforcement Learning1 Jan 2023 0 repositories listed
-
Simultaneously Updating All Persistence Values in Reinforcement Learning21 Nov 2022 0 repositories listed
-
Curiosity in Hindsight: Intrinsic Exploration in Stochastic Environments18 Nov 2022 0 repositories listed
-
Offline RL With Realistic Datasets: Heteroskedasticity and Support Constraints2 Nov 2022 0 repositories listed
-
Spending Thinking Time Wisely: Accelerating MCTS with Virtual Expansions23 Oct 2022 0 repositories listed
-
Bayesian Q-learning With Imperfect Expert Demonstrations1 Oct 2022 0 repositories listed
-
Deep Q-Network for AI Soccer20 Sep 2022 0 repositories listed
-
Rewarding Episodic Visitation Discrepancy for Exploration in Reinforcement Learning19 Sep 2022 0 repositories listed
-
Emergence of Novelty in Evolutionary Algorithms27 Jun 2022 0 repositories listed
-
7 Jun 2022 0 repositories listed
-
Nuclear Norm Maximization Based Curiosity-Driven Learning21 May 2022 0 repositories listed
-
Deep Apprenticeship Learning for Playing Games16 May 2022 0 repositories listed
-
Learning to Constrain Policy Optimization with Virtual Trust Region20 Apr 2022 0 repositories listed
-
Methodical Advice Collection and Reuse in Deep Reinforcement Learning14 Apr 2022 0 repositories listed
-
When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?12 Apr 2022 0 repositories listed
-
Mask Atari for Deep Reinforcement Learning as POMDP Benchmarks31 Mar 2022 0 repositories listed
-
Lazy-MDPs: Towards Interpretable Reinforcement Learning by Learning When to Act16 Mar 2022 0 repositories listed
-
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise2 Mar 2022 0 repositories listed
-
A Unified Perspective on Value Backup and Exploration in Monte-Carlo Tree Search11 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning with Spiking Q-learning21 Jan 2022 0 repositories listed
-
Exploration by Random Network Distillation17 Jan 2022 0 repositories listed
-
Attention Option-Critic7 Jan 2022 0 repositories listed
-
Execute Order 66: Targeted Data Poisoning for Reinforcement Learning3 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning Policies Learn Shared Adversarial Features Across MDPs16 Dec 2021 0 repositories listed
-
A Benchmark for Low-Switching-Cost Reinforcement Learning13 Dec 2021 0 repositories listed
-
DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization9 Dec 2021 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.