Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 95
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 95 of 152: papers 9,401 to 9,500 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Spatially and Seamlessly Hierarchical Reinforcement Learning for State Space and Policy space in Autonomous Driving10 Nov 2021 0 repositories listed
-
Dealing with the Unknown: Pessimistic Offline Reinforcement Learning9 Nov 2021 0 repositories listed
-
HARPO: Learning to Subvert Online Behavioral Advertising9 Nov 2021 0 repositories listed
-
Risk Sensitive Model-Based Reinforcement Learning using Uncertainty Guided Planning9 Nov 2021 0 repositories listed
-
Batch Reinforcement Learning from Crowds8 Nov 2021 0 repositories listed
-
Dueling RL: Reinforcement Learning with Trajectory Preferences8 Nov 2021 0 repositories listed
-
Interactive Inverse Reinforcement Learning for Cooperative Games8 Nov 2021 0 repositories listed
-
On Assessing The Safety of Reinforcement Learning algorithms Using Formal Methods8 Nov 2021 0 repositories listed
-
Automatic Goal Generation using Dynamical Distance Learning7 Nov 2021 0 repositories listed
-
Explainable Deep Reinforcement Learning for Portfolio Management: An Empirical Approach7 Nov 2021 0 repositories listed
-
FinRL: Deep Reinforcement Learning Framework to Automate Trading in Quantitative Finance7 Nov 2021 0 repositories listed
-
FinRL-Podracer: High Performance and Scalable Deep Reinforcement Learning for Quantitative Finance7 Nov 2021 0 repositories listed
-
Optimization of the Model Predictive Control Meta-Parameters Through Reinforcement Learning7 Nov 2021 0 repositories listed
-
A Deep Reinforcement Learning Approach for Composing Moving IoT Services6 Nov 2021 0 repositories listed
-
AI-based Radio Resource Management and Trajectory Design for PD-NOMA Communication in IRS-UAV Assisted Networks6 Nov 2021 0 repositories listed
-
Development of collective behavior in newborn artificial agents6 Nov 2021 0 repositories listed
-
Exponential Bellman Equation and Improved Regret Bounds for Risk-Sensitive Reinforcement Learning6 Nov 2021 0 repositories listed
-
An Algorithmic Theory of Metacognition in Minds and Machines5 Nov 2021 0 repositories listed
-
Improving RNA Secondary Structure Design using Deep Reinforcement Learning5 Nov 2021 0 repositories listed
-
Learning to Cooperate with Unseen Agent via Meta-Reinforcement Learning5 Nov 2021 0 repositories listed
-
Perturbational Complexity by Distribution Mismatch: A Systematic Analysis of Reinforcement Learning in Reproducing Kernel Hilbert Space5 Nov 2021 0 repositories listed
-
Supervised Advantage Actor-Critic for Recommender Systems5 Nov 2021 0 repositories listed
-
Attacking Deep Reinforcement Learning-Based Traffic Signal Control Systems with Colluding Vehicles4 Nov 2021 0 repositories listed
-
Causal versus Marginal Shapley Values for Robotic Lever Manipulation Controlled using Deep Reinforcement Learning4 Nov 2021 0 repositories listed
-
Control of a fly-mimicking flyer in complex flow using deep reinforcement learning4 Nov 2021 0 repositories listed
-
Generalization in Dexterous Manipulation via Geometry-Aware Multi-Task Learning4 Nov 2021 0 repositories listed
-
Imagine Networks4 Nov 2021 0 repositories listed
-
Model-Free Risk-Sensitive Reinforcement Learning4 Nov 2021 0 repositories listed
-
Successor Feature Neural Episodic Control4 Nov 2021 0 repositories listed
-
Towards Learning to Speak and Hear Through Multi-Agent Communication over a Continuous Acoustic Channel4 Nov 2021 0 repositories listed
-
Value Function Spaces: Skill-Centric State Abstractions for Long-Horizon Reasoning4 Nov 2021 0 repositories listed
-
AlphaD3M: Machine Learning Pipeline Synthesis3 Nov 2021 0 repositories listed
-
Autonomous Attack Mitigation for Industrial Control Systems3 Nov 2021 0 repositories listed
-
Image-Guided Navigation of a Robotic Ultrasound Probe for Autonomous Spinal Sonography Using a Shadow-aware Dual-Agent Framework3 Nov 2021 0 repositories listed
-
Is Bang-Bang Control All You Need? Solving Continuous Control with Bernoulli Policies3 Nov 2021 0 repositories listed
-
Model-Based Episodic Memory Induces Dynamic Hybrid Controls3 Nov 2021 0 repositories listed
-
Online Service Provisioning in NFV-enabled Networks Using Deep Reinforcement Learning3 Nov 2021 0 repositories listed
-
Smooth Imitation Learning via Smooth Costs and Smooth Policies3 Nov 2021 0 repositories listed
-
Tuning the Weights: The Impact of Initial Matrix Configurations on Successor Features Learning Efficacy3 Nov 2021 0 repositories listed
-
What Robot do I Need? Fast Co-Adaptation of Morphology and Control using Graph Neural Networks3 Nov 2021 0 repositories listed
-
Integrating Pretrained Language Model for Dialogue Policy Learning2 Nov 2021 0 repositories listed
-
2 Nov 2021 0 repositories listed
-
OnSlicing: Online End-to-End Network Slicing with Reinforcement Learning2 Nov 2021 0 repositories listed
-
Robust Dynamic Bus Control: A Distributional Multi-agent Reinforcement Learning Approach2 Nov 2021 0 repositories listed
-
A Collaborative Multi-agent Reinforcement Learning Framework for Dialog Action Decomposition1 Nov 2021 0 repositories listed
-
A Generative Framework for Simultaneous Machine Translation1 Nov 2021 0 repositories listed
-
Decentralized Cooperative Reinforcement Learning with Hierarchical Information Structure1 Nov 2021 0 repositories listed
-
Feedback Attribution for Counterfactual Bandit Learning in Multi-Domain Spoken Language Understanding1 Nov 2021 0 repositories listed
-
Investigation of Independent Reinforcement Learning Algorithms in Multi-Agent Environments1 Nov 2021 0 repositories listed
-
Learning Task Sampling Policy for Multitask Learning1 Nov 2021 0 repositories listed
-
Learning to Operate an Electric Vehicle Charging Station Considering Vehicle-grid Integration1 Nov 2021 0 repositories listed
-
Machine Learning aided Crop Yield Optimization1 Nov 2021 0 repositories listed
-
Rewards with Negative Examples for Reinforced Topic-Focused Abstractive Summarization1 Nov 2021 0 repositories listed
-
Settling the Horizon-Dependence of Sample Complexity in Reinforcement Learning1 Nov 2021 0 repositories listed
-
An Actor-Critic Method for Simulation-Based Optimization31 Oct 2021 0 repositories listed
-
Decentralized Multi-Agent Reinforcement Learning: An Off-Policy Method31 Oct 2021 0 repositories listed
-
A Decentralized Reinforcement Learning Framework for Efficient Passage of Emergency Vehicles30 Oct 2021 0 repositories listed
-
Adjacency constraint for efficient hierarchical reinforcement learning30 Oct 2021 0 repositories listed
-
Convergence and Optimality of Policy Gradient Methods in Weakly Smooth Settings30 Oct 2021 0 repositories listed
-
Learning Coordinated Terrain-Adaptive Locomotion by Imitating a Centroidal Dynamics Planner30 Oct 2021 0 repositories listed
-
Adaptive Discretization in Online Reinforcement Learning29 Oct 2021 0 repositories listed
-
Brick-by-Brick: Combinatorial Construction with Deep Reinforcement Learning29 Oct 2021 0 repositories listed
-
GalilAI: Out-of-Task Distribution Detection using Causal Active Experimentation for Safe Transfer RL29 Oct 2021 0 repositories listed
-
Learning to Communicate with Reinforcement Learning for an Adaptive Traffic Control System29 Oct 2021 0 repositories listed
-
Mixed Cooperative-Competitive Communication Using Multi-Agent Reinforcement Learning29 Oct 2021 0 repositories listed
-
Reinforced Workload Distribution Fairness29 Oct 2021 0 repositories listed
-
Accelerating Robotic Reinforcement Learning via Parameterized Action Primitives28 Oct 2021 0 repositories listed
-
An Adaptable Approach to Learn Realistic Legged Locomotion without Examples28 Oct 2021 0 repositories listed
-
Bayesian Sequential Optimal Experimental Design for Nonlinear Models Using Policy Gradient Reinforcement Learning28 Oct 2021 0 repositories listed
-
Choosing the Best of Both Worlds: Diverse and Novel Recommendations through Multi-Objective Reinforcement Learning28 Oct 2021 0 repositories listed
-
D2RLIR : an improved and diversified ranking function in interactive recommendation systems based on deep reinforcement learning28 Oct 2021 0 repositories listed
-
Extracting Expert's Goals by What-if Interpretable Modeling28 Oct 2021 0 repositories listed
-
Data Informed Residual Reinforcement Learning for High-Dimensional Robotic Tracking Control28 Oct 2021 0 repositories listed
-
Open Problem: Tight Online Confidence Intervals for RKHS Elements28 Oct 2021 0 repositories listed
-
A Law of Iterated Logarithm for Multi-Agent Reinforcement Learning27 Oct 2021 0 repositories listed
-
A Subgame Perfect Equilibrium Reinforcement Learning Approach to Time-inconsistent Problems27 Oct 2021 0 repositories listed
-
Comparing Heuristics, Constraint Optimization, and Reinforcement Learning for an Industrial 2D Packing Problem27 Oct 2021 0 repositories listed
-
DESTA: A Framework for Safe Reinforcement Learning with Markov Games of Intervention27 Oct 2021 0 repositories listed
-
Enhancing Reinforcement Learning with discrete interfaces to learn the Dyck Language27 Oct 2021 0 repositories listed
-
Finite Horizon Q-learning: Stability, Convergence, Simulations and an application on Smart Grids27 Oct 2021 0 repositories listed
-
APPTeK: Agent-Based Predicate Prediction in Temporal Knowledge Graphs27 Oct 2021 0 repositories listed
-
Model based Multi-agent Reinforcement Learning with Tensor Decompositions27 Oct 2021 0 repositories listed
-
Reinforcement Learning in Factored Action Spaces using Tensor Decompositions27 Oct 2021 0 repositories listed
-
Reinforcement Learning in Linear MDPs: Constant Regret and Representation Selection27 Oct 2021 0 repositories listed
-
The ODE Method for Asymptotic Statistics in Stochastic Approximation and Reinforcement Learning27 Oct 2021 0 repositories listed
-
Transfer learning with causal counterfactual reasoning in Decision Transformers27 Oct 2021 0 repositories listed
-
Accelerating Distributed Deep Reinforcement Learning by In-Network Experience Sampling26 Oct 2021 0 repositories listed
-
Applications of Multi-Agent Reinforcement Learning in Future Internet: A Comprehensive Survey26 Oct 2021 0 repositories listed
-
Automating Control of Overestimation Bias for Reinforcement Learning26 Oct 2021 0 repositories listed
-
Average-Reward Learning and Planning with Options26 Oct 2021 0 repositories listed
-
Distributed Multi-Agent Deep Reinforcement Learning Framework for Whole-building HVAC Control26 Oct 2021 0 repositories listed
-
EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization26 Oct 2021 0 repositories listed
-
Fragment-based Sequential Translation for Molecular Optimization26 Oct 2021 0 repositories listed
-
Neural PPO-Clip Attains Global Optimality: A Hinge Loss Perspective26 Oct 2021 0 repositories listed
-
Learning Robust Controllers Via Probabilistic Model-Based Policy Search26 Oct 2021 0 repositories listed
-
A Deep Reinforcement Learning Approach for Audio-based Navigation and Audio Source Localization in Multi-speaker Environments25 Oct 2021 0 repositories listed
-
Can Q-Learning be Improved with Advice?25 Oct 2021 0 repositories listed
-
Common Information based Approximate State Representations in Multi-Agent Reinforcement Learning25 Oct 2021 0 repositories listed
-
Learning What to Memorize: Using Intrinsic Motivation to Form Useful Memory in Partially Observable Reinforcement Learning25 Oct 2021 0 repositories listed
-
Operator Shifting for Model-based Policy Evaluation25 Oct 2021 0 repositories listed