Browse State-of-the-Art › reinforcement-learning › Papers, page 104
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 104 of 135: papers 10,301 to 10,400 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Assessment of Reward Functions in Reinforcement Learning for Multi-Modal Urban Traffic Control under Real-World limitations17 Oct 2020 0 repositories listed
-
Learning Elimination Ordering for Tree Decomposition Problem17 Oct 2020 0 repositories listed
-
Learning Lower Bounds for Graph Exploration With Reinforcement Learning17 Oct 2020 0 repositories listed
-
Neural Algorithms for Graph Navigation17 Oct 2020 0 repositories listed
-
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning17 Oct 2020 0 repositories listed
-
Autonomous Control of a Particle Accelerator using Deep Reinforcement Learning16 Oct 2020 0 repositories listed
-
DOOM: A Novel Adversarial-DRL-Based Op-Code Level Metamorphic Malware Obfuscator for the Enhancement of IDS16 Oct 2020 0 repositories listed
-
Efficient Robotic Object Search via HIEM: Hierarchical Policy Learning with Intrinsic-Extrinsic Modeling16 Oct 2020 0 repositories listed
-
Reinforcement Learning for Efficient and Tuning-Free Link Adaptation16 Oct 2020 0 repositories listed
-
Uncertainty-aware Contact-safe Model-based Reinforcement Learning16 Oct 2020 0 repositories listed
-
A Nesterov's Accelerated quasi-Newton method for Global Routing using Deep Reinforcement Learning15 Oct 2020 0 repositories listed
-
An Empowerment-based Solution to Robotic Manipulation Tasks with Sparse Rewards15 Oct 2020 0 repositories listed
-
Applicability and Challenges of Deep Reinforcement Learning for Satellite Frequency Plan Design15 Oct 2020 0 repositories listed
-
Blending Search and Discovery: Tag-Based Query Refinement with Contextual Reinforcement Learning15 Oct 2020 0 repositories listed
-
Cooperative-Competitive Reinforcement Learning with History-Dependent Rewards15 Oct 2020 0 repositories listed
-
Deep Learning of Koopman Representation for Control15 Oct 2020 0 repositories listed
-
Explanation Augmented Feedback in Human-in-the-Loop Reinforcement Learning15 Oct 2020 0 repositories listed
-
Local Differential Privacy for Regret Minimization in Reinforcement Learning15 Oct 2020 0 repositories listed
-
Optimal Dispatch in Emergency Service System via Reinforcement Learning15 Oct 2020 0 repositories listed
-
Average Cost Optimal Control of Stochastic Systems Using Reinforcement Learning13 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning and Transportation Research: A Comprehensive Review13 Oct 2020 0 repositories listed
-
Model-Based Reinforcement Learning for Type 1Diabetes Blood Glucose Control13 Oct 2020 0 repositories listed
-
Random Network Distillation as a Diversity Metric for Both Image and Text Generation13 Oct 2020 0 repositories listed
-
AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal Control12 Oct 2020 0 repositories listed
-
12 Oct 2020 0 repositories listed
-
Is Plug-in Solver Sample-Efficient for Feature-based Reinforcement Learning?12 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reward-free Reinforcement Learning12 Oct 2020 0 repositories listed
-
Remote Electrical Tilt Optimization via Safe Reinforcement Learning12 Oct 2020 0 repositories listed
-
Controlling Graph Dynamics with Reinforcement Learning and Graph Neural Networks11 Oct 2020 0 repositories listed
-
Safe Reinforcement Learning with Natural Language Constraints11 Oct 2020 0 repositories listed
-
Reinforcement Learning on Computational Resource Allocation of Cloud-based Wireless Networks10 Oct 2020 0 repositories listed
-
Characterizing Policy Divergence for Personalized Meta-Reinforcement Learning9 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Asset Allocation in US Equities9 Oct 2020 0 repositories listed
-
Deep RL With Information Constrained Policies: Generalization in Continuous Control9 Oct 2020 0 repositories listed
-
Dynamic Context Selection for Document-level Neural Machine Translation via Reinforcement Learning9 Oct 2020 0 repositories listed
-
Jointly-Learned State-Action Embedding for Efficient Reinforcement Learning9 Oct 2020 0 repositories listed
-
Learning to Locomote: Understanding How Environment Design Matters for Deep Reinforcement Learning9 Oct 2020 0 repositories listed
-
Parameterized Reinforcement Learning for Optical System Optimization9 Oct 2020 0 repositories listed
-
Learning Intrinsic Symbolic Rewards in Reinforcement Learning8 Oct 2020 0 repositories listed
-
Nonstationary Reinforcement Learning with Linear Function Approximation8 Oct 2020 0 repositories listed
-
Provable Fictitious Play for General Mean-Field Games8 Oct 2020 0 repositories listed
-
Actor-Critic Algorithm for High-dimensional Partial Differential Equations7 Oct 2020 0 repositories listed
-
Episodic Reinforcement Learning in Finite MDPs: Minimax Lower Bounds Revisited7 Oct 2020 0 repositories listed
-
Instance-Dependent Complexity of Contextual Bandits and Reinforcement Learning: A Disagreement-Based Perspective7 Oct 2020 0 repositories listed
-
Online Safety Assurance for Deep Reinforcement Learning7 Oct 2020 0 repositories listed
-
Regularized Inverse Reinforcement Learning7 Oct 2020 0 repositories listed
-
Reinforcement Learning for Many-Body Ground-State Preparation Inspired by Counterdiabatic Driving7 Oct 2020 0 repositories listed
-
Variational Intrinsic Control Revisited7 Oct 2020 0 repositories listed
-
Heterogeneous Multi-Agent Reinforcement Learning for Unknown Environment Mapping6 Oct 2020 0 repositories listed
-
Safety Aware Reinforcement Learning (SARL)6 Oct 2020 0 repositories listed
-
UneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning6 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Collaborative Edge Computing in Vehicular Networks5 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Electric Vehicle Routing Problem with Time Windows5 Oct 2020 0 repositories listed
-
Goal-directed Generation of Discrete Structures with Conditional Generative Models5 Oct 2020 0 repositories listed
-
Sentiment Analysis for Reinforcement Learning5 Oct 2020 0 repositories listed
-
The act of remembering: a study in partially observable reinforcement learning5 Oct 2020 0 repositories listed
-
A Sharp Analysis of Model-based Reinforcement Learning with Self-Play4 Oct 2020 0 repositories listed
-
Test-Cost Sensitive Methods for Identifying Nearby Points4 Oct 2020 0 repositories listed
-
Attractor Selection in Nonlinear Energy Harvesting Using Deep Reinforcement Learning3 Oct 2020 0 repositories listed
-
Beyond Tabula-Rasa: a Modular Reinforcement Learning Approach for Physically Embedded 3D Sokoban3 Oct 2020 0 repositories listed
-
Disentangling causal effects for hierarchical reinforcement learning3 Oct 2020 0 repositories listed
-
Mean-Variance Efficient Reinforcement Learning with Applications to Dynamic Financial Investment3 Oct 2020 0 repositories listed
-
Interactive Reinforcement Learning for Feature Selection with Decision Tree in the Loop2 Oct 2020 0 repositories listed
-
MADRaS : Multi Agent Driving Simulator2 Oct 2020 0 repositories listed
-
Reinforcement Learning of Sequential Price Mechanisms2 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning with Mixed Convolutional Network1 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reinforcement Learning for Discounted MDPs1 Oct 2020 0 repositories listed
-
Multi-Reward based Reinforcement Learning for Neural Machine Translation1 Oct 2020 0 repositories listed
-
Recognition Method of Important Words in Korean Text based on Reinforcement Learning1 Oct 2020 0 repositories listed
-
Bayesian Meta-reinforcement Learning for Traffic Signal Control1 Oct 2020 0 repositories listed
-
AAMDRL: Augmented Asset Management with Deep Reinforcement Learning30 Sep 2020 0 repositories listed
-
Accelerating Optimization and Reinforcement Learning with Quasi-Stochastic Approximation30 Sep 2020 0 repositories listed
-
Bridging the gap between Markowitz planning and deep reinforcement learning30 Sep 2020 0 repositories listed
-
Entropy Regularization for Mean Field Games with Learning30 Sep 2020 0 repositories listed
-
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network30 Sep 2020 0 repositories listed
-
Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks30 Sep 2020 0 repositories listed
-
Toolpath design for additive manufacturing using deep reinforcement learning30 Sep 2020 0 repositories listed
-
Cross Learning in Deep Q-Networks29 Sep 2020 0 repositories listed
-
Trust-Region Method with Deep Reinforcement Learning in Analog Design Space Exploration29 Sep 2020 0 repositories listed
-
Reannealing of Decaying Exploration Based On Heuristic Measure in Deep Q-Network29 Sep 2020 0 repositories listed
-
Agent Environment Cycle Games28 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for DER Cyber-Attack Mitigation28 Sep 2020 0 repositories listed
-
Efficient Exploration for Model-based Reinforcement Learning with Continuous States and Actions28 Sep 2020 0 repositories listed
-
Is Reinforcement Learning More Difficult Than Bandits? A Near-optimal Algorithm Escaping the Curse of Horizon28 Sep 2020 0 repositories listed
-
Jointly-Trained State-Action Embedding for Efficient Reinforcement Learning28 Sep 2020 0 repositories listed
-
Neuron Activation Analysis for Multi-Joint Robot Reinforcement Learning28 Sep 2020 0 repositories listed
-
REPAINT: Knowledge Transfer in Deep Actor-Critic Reinforcement Learning28 Sep 2020 0 repositories listed
-
The Emergence of Individuality in Multi-Agent Reinforcement Learning28 Sep 2020 0 repositories listed
-
Towards Heterogeneous Multi-Agent Reinforcement Learning with Graph Neural Networks28 Sep 2020 0 repositories listed
-
Machine Learning in Event-Triggered Control: Recent Advances and Open Issues27 Sep 2020 0 repositories listed
-
Scalable Deep Reinforcement Learning for Ride-Hailing27 Sep 2020 0 repositories listed
-
Scheduling and Power Control for Wireless Multicast Systems via Deep Reinforcement Learning27 Sep 2020 0 repositories listed
-
Virtual Experience to Real World Application: Sidewalk Obstacle Avoidance Using Reinforcement Learning for Visually Impaired27 Sep 2020 0 repositories listed
-
Complementary Meta-Reinforcement Learning for Fault-Adaptive Control26 Sep 2020 0 repositories listed
-
Graph neural induction of value iteration26 Sep 2020 0 repositories listed
-
Inverse Rational Control with Partially Observable Continuous Nonlinear Dynamics26 Sep 2020 0 repositories listed
-
Lineage Evolution Reinforcement Learning26 Sep 2020 0 repositories listed
-
Reinforcement Learning-based N-ary Cross-Sentence Relation Extraction26 Sep 2020 0 repositories listed
-
Motion Planning by Reinforcement Learning for an Unmanned Aerial Vehicle in Virtual Open Space with Static Obstacles24 Sep 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.