Browse State-of-the-Art › reinforcement-learning › Papers, page 83
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 83 of 135: papers 8,201 to 8,300 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Deep Binary Reinforcement Learning for Scalable Verification11 Mar 2022 0 repositories listed
-
Graph Neural Networks for Relational Inductive Bias in Vision-based Deep Reinforcement Learning of Robot Control11 Mar 2022 0 repositories listed
-
Near-optimal Offline Reinforcement Learning with Linear Representation: Leveraging Variance Information with Pessimism11 Mar 2022 0 repositories listed
-
Reinforcement Learning for Linear Quadratic Control is Vulnerable Under Cost Manipulation11 Mar 2022 0 repositories listed
-
Random Ensemble Reinforcement Learning for Traffic Signal Control10 Mar 2022 0 repositories listed
-
Investigation of Factorized Optical Flows as Mid-Level Representations9 Mar 2022 0 repositories listed
-
Multi-robot Cooperative Pursuit via Potential Field-Enhanced Reinforcement Learning9 Mar 2022 0 repositories listed
-
A Complete Characterization of Linear Estimators for Offline Policy Evaluation8 Mar 2022 0 repositories listed
-
Distributed Control using Reinforcement Learning with Temporal-Logic-Based Reward Shaping8 Mar 2022 0 repositories listed
-
Designing Heterogeneous GNNs with Desired Permutation Properties for Wireless Resource Allocation8 Mar 2022 0 repositories listed
-
Multi-Agent Broad Reinforcement Learning for Intelligent Traffic Light Control8 Mar 2022 0 repositories listed
-
Policy-Based Bayesian Experimental Design for Non-Differentiable Implicit Models8 Mar 2022 0 repositories listed
-
Reinforced MOOCs Concept Recommendation in Heterogeneous Information Networks8 Mar 2022 0 repositories listed
-
Rényi State Entropy for Exploration Acceleration in Reinforcement Learning8 Mar 2022 0 repositories listed
-
A Survey on Reinforcement Learning Methods in Character Animation7 Mar 2022 0 repositories listed
-
Cascaded Gaps: Towards Gap-Dependent Regret for Risk-Sensitive Reinforcement Learning7 Mar 2022 0 repositories listed
-
Efficient Policy Generation in Multi-Agent Systems via Hypergraph Neural Network7 Mar 2022 0 repositories listed
-
Fast and Data Efficient Reinforcement Learning from Pixels via Non-Parametric Value Approximation7 Mar 2022 0 repositories listed
-
Graph Neural Networks for Image Classification and Reinforcement Learning using Graph representations7 Mar 2022 0 repositories listed
-
Knowledge Transfer in Deep Reinforcement Learning for Slice-Aware Mobility Robustness Optimization7 Mar 2022 0 repositories listed
-
Learn to Match with No Regret: Reinforcement Learning in Markov Matching Markets7 Mar 2022 0 repositories listed
-
Reinforcement Learning for Location-Aware Scheduling7 Mar 2022 0 repositories listed
-
Scalable multi-agent reinforcement learning for distributed control of residential energy flexibility7 Mar 2022 0 repositories listed
-
Hierarchically Structured Scheduling and Execution of Tasks in a Multi-Agent Environment6 Mar 2022 0 repositories listed
-
Leveraging Reward Gradients For Reinforcement Learning in Differentiable Physics Simulations6 Mar 2022 0 repositories listed
-
Offline Deep Reinforcement Learning for Dynamic Pricing of Consumer Credit6 Mar 2022 0 repositories listed
-
Recursive Reasoning Graph for Multi-Agent Reinforcement Learning6 Mar 2022 0 repositories listed
-
Watch from sky: machine-learning-based multi-UAV network for predictive police surveillance6 Mar 2022 0 repositories listed
-
Safe Reinforcement Learning for Legged Locomotion5 Mar 2022 0 repositories listed
-
Target Network and Truncation Overcome The Deadly Triad in Q-Learning5 Mar 2022 0 repositories listed
-
Cloud-Edge Training Architecture for Sim-to-Real Deep Reinforcement Learning4 Mar 2022 0 repositories listed
-
GraspARL: Dynamic Grasping via Adversarial Reinforcement Learning4 Mar 2022 0 repositories listed
-
Reinforcement Learning in Modern Biostatistics: Constructing Optimal Adaptive Interventions4 Mar 2022 0 repositories listed
-
Bilateral Deep Reinforcement Learning Approach for Better-than-human Car Following Model3 Mar 2022 0 repositories listed
-
Intrinsically-Motivated Reinforcement Learning: A Brief Introduction3 Mar 2022 0 repositories listed
-
Quantum Reinforcement Learning via Policy Iteration3 Mar 2022 0 repositories listed
-
The Best of Both Worlds: Reinforcement Learning with Logarithmic Regret and Policy Switches3 Mar 2022 0 repositories listed
-
Follow your Nose: Using General Value Functions for Directed Exploration in Reinforcement Learning2 Mar 2022 0 repositories listed
-
Reliable validation of Reinforcement Learning Benchmarks2 Mar 2022 0 repositories listed
-
A Theory of Abstraction in Reinforcement Learning1 Mar 2022 0 repositories listed
-
Approximating a deep reinforcement learning docking agent using linear model trees1 Mar 2022 0 repositories listed
-
Distributional Reinforcement Learning for Scheduling of Chemical Production Processes1 Mar 2022 0 repositories listed
-
DreamingV2: Reinforcement Learning with Discrete World Models without Reconstruction1 Mar 2022 0 repositories listed
-
Pessimistic Q-Learning for Offline Reinforcement Learning: Towards Optimal Sample Complexity28 Feb 2022 0 repositories listed
-
Weakly Supervised Disentangled Representation for Goal-conditioned Reinforcement Learning28 Feb 2022 0 repositories listed
-
Neural-Progressive Hedging: Enforcing Constraints in Reinforcement Learning with Stochastic Programming27 Feb 2022 0 repositories listed
-
Distributed Multi-Agent Reinforcement Learning Based on Graph-Induced Local Value Functions26 Feb 2022 0 repositories listed
-
Domain Knowledge-Based Automated Analog Circuit Design with Deep Reinforcement Learning26 Feb 2022 0 repositories listed
-
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning25 Feb 2022 0 repositories listed
-
Context-Hierarchy Inverse Reinforcement Learning25 Feb 2022 0 repositories listed
-
Learning Dynamic Mechanisms in Unknown Environments: A Reinforcement Learning Approach25 Feb 2022 0 repositories listed
-
Reachability analysis in stochastic directed graphs by reinforcement learning25 Feb 2022 0 repositories listed
-
Evolving-to-Learn Reinforcement Learning Tasks with Spiking Neural Networks24 Feb 2022 0 repositories listed
-
Consistent Dropout for Policy Gradient Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning in Practice: Opportunities and Challenges23 Feb 2022 0 repositories listed
-
Learning Relative Return Policies With Upside-Down Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning from Demonstrations by Novel Interactive Expert and Application to Automatic Berthing Control Systems for Unmanned Surface Vessel23 Feb 2022 0 repositories listed
-
Training Characteristic Functions with Reinforcement Learning: XAI-methods play Connect Four23 Feb 2022 0 repositories listed
-
A policy gradient approach for optimization of smooth risk measures22 Feb 2022 0 repositories listed
-
Behaviour-Diverse Automatic Penetration Testing: A Curiosity-Driven Multi-Objective Deep Reinforcement Learning Approach22 Feb 2022 0 repositories listed
-
Behaviour-neutral Smart Charging of Plugin Electric Vehicles: Reinforcement learning approach22 Feb 2022 0 repositories listed
-
Continual Auxiliary Task Learning22 Feb 2022 0 repositories listed
-
Multi-fidelity reinforcement learning framework for shape optimization22 Feb 2022 0 repositories listed
-
Reward-Free Policy Space Compression for Reinforcement Learning22 Feb 2022 0 repositories listed
-
Sequential Information Design: Markov Persuasion Process and Its Efficient Reinforcement Learning22 Feb 2022 0 repositories listed
-
Accelerating Primal-dual Methods for Regularized Markov Decision Processes21 Feb 2022 0 repositories listed
-
Autonomous Warehouse Robot using Deep Q-Learning21 Feb 2022 0 repositories listed
-
CCPT: Automatic Gameplay Testing and Validation with Curiosity-Conditioned Proximal Trajectories21 Feb 2022 0 repositories listed
-
Hybrid Learning for Orchestrating Deep Learning Inference in Multi-user Edge-cloud Networks21 Feb 2022 0 repositories listed
-
Rule Mining over Knowledge Graphs via Reinforcement Learning21 Feb 2022 0 repositories listed
-
PooL: Pheromone-inspired Communication Framework forLarge Scale Multi-Agent Reinforcement Learning20 Feb 2022 0 repositories listed
-
Selective Credit Assignment20 Feb 2022 0 repositories listed
-
A Behavior Regularized Implicit Policy for Offline Reinforcement Learning19 Feb 2022 0 repositories listed
-
Multi-task Safe Reinforcement Learning for Navigating Intersections in Dense Traffic19 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning as a Stackelberg Game via Adaptively-Regularized Adversarial Training19 Feb 2022 0 repositories listed
-
Transformation Coding: Simple Objectives for Equivariant Representations19 Feb 2022 0 repositories listed
-
Who Are the Best Adopters? User Selection Model for Free Trial Item Promotion19 Feb 2022 0 repositories listed
-
Can Interpretable Reinforcement Learning Manage Prosperity Your Way?18 Feb 2022 0 repositories listed
-
A Survey of Explainable Reinforcement Learning17 Feb 2022 0 repositories listed
-
A Survey on Deep Reinforcement Learning-based Approaches for Adaptation and Generalization17 Feb 2022 0 repositories listed
-
Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization17 Feb 2022 0 repositories listed
-
Retrieval-Augmented Reinforcement Learning17 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning via Genetic Curriculum17 Feb 2022 0 repositories listed
-
Branching Reinforcement Learning16 Feb 2022 0 repositories listed
-
Domain Adaptive Fake News Detection via Reinforcement Learning16 Feb 2022 0 repositories listed
-
Temporal Difference Learning with Continuous Time and State in the Stochastic Setting16 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning Based Multi-Access Edge Computing Schedule for Internet of Vehicle15 Feb 2022 0 repositories listed
-
Interpretable Reinforcement Learning with Multilevel Subgoal Discovery15 Feb 2022 0 repositories listed
-
L2C2: Locally Lipschitz Continuous Constraint towards Stable and Smooth Reinforcement Learning15 Feb 2022 0 repositories listed
-
Learning to Mitigate AI Collusion on Economic Platforms15 Feb 2022 0 repositories listed
-
User-Oriented Robust Reinforcement Learning15 Feb 2022 0 repositories listed
-
Provably Efficient Causal Model-Based Reinforcement Learning for Systematic Generalization14 Feb 2022 0 repositories listed
-
Reinforcement Learning in Presence of Discrete Markovian Context Evolution14 Feb 2022 0 repositories listed
-
Sequential Bayesian experimental designs via reinforcement learning14 Feb 2022 0 repositories listed
-
Statistical Inference After Adaptive Sampling for Longitudinal Data14 Feb 2022 0 repositories listed
-
Towards Deployment-Efficient Reinforcement Learning: Lower Bound and Optimality14 Feb 2022 0 repositories listed
-
Autonomous Drone Swarm Navigation and Multi-target Tracking in 3D Environments with Dynamic Obstacles13 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning and Convex Mean-Variance Optimisation for Portfolio Management13 Feb 2022 0 repositories listed
-
Individual-Level Inverse Reinforcement Learning for Mean Field Games13 Feb 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.