Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 75
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 75 of 152: papers 7,401 to 7,500 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Autonomous particles24 Jan 2023 0 repositories listed
-
Explainable Deep Reinforcement Learning: State of the Art and Challenges24 Jan 2023 0 repositories listed
-
Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review24 Jan 2023 0 repositories listed
-
Minimal Value-Equivalent Partial Models for Scalable and Robust Planning in Lifelong Reinforcement Learning24 Jan 2023 0 repositories listed
-
SMART: Self-supervised Multi-task pretrAining with contRol Transformers24 Jan 2023 0 repositories listed
-
Story Shaping: Teaching Agents Human-like Behavior with Stories24 Jan 2023 0 repositories listed
-
Forecaster-aided User Association and Load Balancing in Multi-band Mobile Networks23 Jan 2023 0 repositories listed
-
Learning to View: Decision Transformers for Active Object Detection23 Jan 2023 0 repositories listed
-
Model Based Reinforcement Learning with Non-Gaussian Environment Dynamics and its Application to Portfolio Optimization23 Jan 2023 0 repositories listed
-
Quasi-optimal Reinforcement Learning with Continuous Actions21 Jan 2023 0 repositories listed
-
Asynchronous Deep Double Duelling Q-Learning for Trading-Signal Execution in Limit Order Book Markets20 Jan 2023 0 repositories listed
-
Generative Slate Recommendation with Reinforcement Learning20 Jan 2023 0 repositories listed
-
Multi-agent Reinforcement Learning with Graph Q-Networks for Antenna Tuning20 Jan 2023 0 repositories listed
-
Multi-Armed Bandits and Quantum Channel Oracles20 Jan 2023 0 repositories listed
-
Reinforcement learning-based estimation for partial differential equations20 Jan 2023 0 repositories listed
-
Revisiting Estimation Bias in Policy Gradients for Deep Reinforcement Learning20 Jan 2023 0 repositories listed
-
A Survey of Meta-Reinforcement Learning19 Jan 2023 0 repositories listed
-
Advanced Scaling Methods for VNF deployment with Reinforcement Learning19 Jan 2023 0 repositories listed
-
Domain-adapted Learning and Interpretability: DRL for Gas Trading19 Jan 2023 0 repositories listed
-
Domain-adapted Learning and Imitation: DRL for Power Arbitrage19 Jan 2023 0 repositories listed
-
Generalization through Diversity: Improving Unsupervised Environment Design19 Jan 2023 0 repositories listed
-
Tight Guarantees for Interactive Decision Making with the Decision-Estimation Coefficient19 Jan 2023 0 repositories listed
-
Human-Timescale Adaptation in an Open-Ended Task Space18 Jan 2023 0 repositories listed
-
Multi-compartment Neuron and Population Encoding Powered Spiking Neural Network for Deep Distributional Reinforcement Learning18 Jan 2023 0 repositories listed
-
Adversarial Robust Deep Reinforcement Learning Requires Redefining Robustness17 Jan 2023 0 repositories listed
-
DQNAS: Neural Architecture Search using Reinforcement Learning17 Jan 2023 0 repositories listed
-
Learning to solve arithmetic problems with a virtual abacus17 Jan 2023 0 repositories listed
-
Show me what you want: Inverse reinforcement learning to automatically design robot swarms by demonstration17 Jan 2023 0 repositories listed
-
Neuro-Symbolic World Models for Adapting to Open World Novelty16 Jan 2023 0 repositories listed
-
CogReact: A Reinforced Framework to Model Human Cognitive Reaction Modulated by Dynamic Intervention15 Jan 2023 0 repositories listed
-
Neuro-symbolic Meta Reinforcement Learning for Trading15 Jan 2023 0 repositories listed
-
First Three Years of the International Verification of Neural Networks Competition (VNN-COMP)14 Jan 2023 0 repositories listed
-
PRUDEX-Compass: Towards Systematic Evaluation of Reinforcement Learning in Financial Markets14 Jan 2023 0 repositories listed
-
Reinforcement Learning for Protocol Synthesis in Resource-Constrained Wireless Sensor and IoT Networks14 Jan 2023 0 repositories listed
-
Risk-Averse Reinforcement Learning via Dynamic Time-Consistent Risk Measures14 Jan 2023 0 repositories listed
-
A Constrained-Optimization Approach to the Execution of Prioritized Stacks of Learned Multi-Robot Tasks13 Jan 2023 0 repositories listed
-
Decentralized model-free reinforcement learning in stochastic games with average-reward objective13 Jan 2023 0 repositories listed
-
Hierarchical Deep Q-Learning Based Handover in Wireless Networks with Dual Connectivity13 Jan 2023 0 repositories listed
-
Multi-Target Landmark Detection with Incomplete Images via Reinforcement Learning and Shape Prior13 Jan 2023 0 repositories listed
-
Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning13 Jan 2023 0 repositories listed
-
Asynchronous training of quantum reinforcement learning12 Jan 2023 0 repositories listed
-
Reinforcement Learning-based Joint Handover and Beam Tracking in Millimeter-wave Networks12 Jan 2023 0 repositories listed
-
Safe Policy Improvement for POMDPs via Finite-State Controllers12 Jan 2023 0 repositories listed
-
An Analysis of Quantile Temporal-Difference Learning11 Jan 2023 0 repositories listed
-
Efficient Preference-Based Reinforcement Learning Using Learned Dynamics Models11 Jan 2023 0 repositories listed
-
SoK: Adversarial Machine Learning Attacks and Defences in Multi-Agent Reinforcement Learning11 Jan 2023 0 repositories listed
-
Switchable Lightweight Anti-symmetric Processing (SLAP) with CNN Outspeeds Data Augmentation by Smaller Sample -- Application in Gomoku Reinforcement Learning11 Jan 2023 0 repositories listed
-
Actor-Director-Critic: A Novel Deep Reinforcement Learning Framework10 Jan 2023 0 repositories listed
-
Towards AI-controlled FES-restoration of arm movements: Controlling for progressive muscular fatigue with Gaussian state-space models10 Jan 2023 0 repositories listed
-
Towards AI-controlled FES-restoration of arm movements: neuromechanics-based reinforcement learning for 3-D reaching10 Jan 2023 0 repositories listed
-
Exploration in Model-based Reinforcement Learning with Randomized Reward9 Jan 2023 0 repositories listed
-
Minimax Weight Learning for Absorbing MDPs9 Jan 2023 0 repositories listed
-
Network Slicing via Transfer Learning aided Distributed Deep Reinforcement Learning9 Jan 2023 0 repositories listed
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning9 Jan 2023 0 repositories listed
-
A Survey on Transformers in Reinforcement Learning8 Jan 2023 0 repositories listed
-
Learning Symbolic Representations for Reinforcement Learning of Non-Markovian Behavior8 Jan 2023 0 repositories listed
-
Hierarchical Reinforcement Learning for RIS-Assisted Energy-Efficient RAN7 Jan 2023 0 repositories listed
-
Mathematical Models and Reinforcement Learning based Evolutionary Algorithm Framework for Satellite Scheduling Problem7 Jan 2023 0 repositories listed
-
Markov Chain Concentration with an Application in Reinforcement Learning7 Jan 2023 0 repositories listed
-
A Deep Reinforcement Learning-Based Controller for Magnetorheological-Damped Vehicle Suspension6 Jan 2023 0 repositories listed
-
Multi-Agent Reinforcement Learning for Fast-Timescale Demand Response of Residential Loads6 Jan 2023 0 repositories listed
-
Provable Reset-free Reinforcement Learning by No-Regret Reduction6 Jan 2023 0 repositories listed
-
Data-Driven Inverse Reinforcement Learning for Expert-Learner Zero-Sum Games5 Jan 2023 0 repositories listed
-
Reinforcement Learning-Based Air Traffic Deconfliction5 Jan 2023 0 repositories listed
-
Scalable Communication for Multi-Agent Reinforcement Learning via Transformer-Based Email Mechanism5 Jan 2023 0 repositories listed
-
Value Enhancement of Reinforcement Learning via Efficient and Robust Trust Region Optimization5 Jan 2023 0 repositories listed
-
Learning-based MPC from Big Data Using Reinforcement Learning4 Jan 2023 0 repositories listed
-
UAV aided Metaverse over Wireless Communications: A Reinforcement Learning Approach4 Jan 2023 0 repositories listed
-
A Succinct Summary of Reinforcement Learning3 Jan 2023 0 repositories listed
-
Contextual Conservative Q-Learning for Offline Reinforcement Learning3 Jan 2023 0 repositories listed
-
Offline Evaluation for Reinforcement Learning-based Recommendation: A Critical Issue and Some Alternatives3 Jan 2023 0 repositories listed
-
Safe Reinforcement Learning for an Energy-Efficient Driver Assistance System3 Jan 2023 0 repositories listed
-
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning3 Jan 2023 0 repositories listed
-
Towards Deployable RL - What's Broken with RL Research and a Potential Fix3 Jan 2023 0 repositories listed
-
A Policy Optimization Method Towards Optimal-time Stability2 Jan 2023 0 repositories listed
-
Deep Reinforcement Learning for Asset Allocation: Reward Clipping2 Jan 2023 0 repositories listed
-
Large-Scale Traffic Signal Control by a Nash Deep Q-network Approach2 Jan 2023 0 repositories listed
-
Safety Filtering for Reinforcement Learning-based Adaptive Cruise Control2 Jan 2023 0 repositories listed
-
Co-Speech Gesture Synthesis by Reinforcement Learning With Contrastive Pre-Trained Rewards1 Jan 2023 0 repositories listed
-
Local-Guided Global: Paired Similarity Representation for Visual Reinforcement Learning1 Jan 2023 0 repositories listed
-
Optimization of Image Transmission in a Cooperative Semantic Communication Networks1 Jan 2023 0 repositories listed
-
PolicyCleanse: Backdoor Detection and Mitigation for Competitive Reinforcement Learning1 Jan 2023 0 repositories listed
-
Simoun: Synergizing Interactive Motion-appearance Understanding for Vision-based Reinforcement Learning1 Jan 2023 0 repositories listed
-
Stabilizing Visual Reinforcement Learning via Asymmetric Interactive Cooperation1 Jan 2023 0 repositories listed
-
Accuracy-Guaranteed Collaborative DNN Inference in Industrial IoT via Deep Reinforcement Learning31 Dec 2022 0 repositories listed
-
Cost-Effective Two-Stage Network Slicing for Edge-Cloud Orchestrated Vehicular Networks31 Dec 2022 0 repositories listed
-
New Challenges in Reinforcement Learning: A Survey of Security and Privacy31 Dec 2022 0 repositories listed
-
Hybrid Deep Reinforcement Learning and Planning for Safe and Comfortable Automated Driving30 Dec 2022 0 repositories listed
-
POMRL: No-Regret Learning-to-Plan with Increasing Horizons30 Dec 2022 0 repositories listed
-
A Novel Experts Advice Aggregation Framework Using Deep Reinforcement Learning for Portfolio Management29 Dec 2022 0 repositories listed
-
Backward Curriculum Reinforcement Learning29 Dec 2022 0 repositories listed
-
Federated Multi-Agent Deep Reinforcement Learning Approach via Physics-Informed Reward for Multi-Microgrid Energy Management29 Dec 2022 0 repositories listed
-
Offline Policy Optimization in RL with Variance Regularizaton29 Dec 2022 0 repositories listed
-
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces29 Dec 2022 0 repositories listed
-
On Transforming Reinforcement Learning by Transformer: The Development Trajectory29 Dec 2022 0 repositories listed
-
Certifying Safety in Reinforcement Learning under Adversarial Perturbation Attacks28 Dec 2022 0 repositories listed
-
Don't do it: Safer Reinforcement Learning With Rule-based Guidance28 Dec 2022 0 repositories listed
-
Improving a sequence-to-sequence nlp model using a reinforcement learning policy algorithm28 Dec 2022 0 repositories listed
-
On the Convergence of Discounted Policy Gradient Methods28 Dec 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.