Browse State-of-the-Art › reinforcement-learning › Papers, page 88
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 88 of 135: papers 8,701 to 8,800 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
DESTA: A Framework for Safe Reinforcement Learning with Markov Games of Intervention27 Oct 2021 0 repositories listed
-
Enhancing Reinforcement Learning with discrete interfaces to learn the Dyck Language27 Oct 2021 0 repositories listed
-
Finite Horizon Q-learning: Stability, Convergence, Simulations and an application on Smart Grids27 Oct 2021 0 repositories listed
-
APPTeK: Agent-Based Predicate Prediction in Temporal Knowledge Graphs27 Oct 2021 0 repositories listed
-
Model based Multi-agent Reinforcement Learning with Tensor Decompositions27 Oct 2021 0 repositories listed
-
Reinforcement Learning in Factored Action Spaces using Tensor Decompositions27 Oct 2021 0 repositories listed
-
Reinforcement Learning in Linear MDPs: Constant Regret and Representation Selection27 Oct 2021 0 repositories listed
-
The ODE Method for Asymptotic Statistics in Stochastic Approximation and Reinforcement Learning27 Oct 2021 0 repositories listed
-
Transfer learning with causal counterfactual reasoning in Decision Transformers27 Oct 2021 0 repositories listed
-
Accelerating Distributed Deep Reinforcement Learning by In-Network Experience Sampling26 Oct 2021 0 repositories listed
-
Applications of Multi-Agent Reinforcement Learning in Future Internet: A Comprehensive Survey26 Oct 2021 0 repositories listed
-
Automating Control of Overestimation Bias for Reinforcement Learning26 Oct 2021 0 repositories listed
-
Average-Reward Learning and Planning with Options26 Oct 2021 0 repositories listed
-
Distributed Multi-Agent Deep Reinforcement Learning Framework for Whole-building HVAC Control26 Oct 2021 0 repositories listed
-
EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization26 Oct 2021 0 repositories listed
-
Neural PPO-Clip Attains Global Optimality: A Hinge Loss Perspective26 Oct 2021 0 repositories listed
-
Learning Robust Controllers Via Probabilistic Model-Based Policy Search26 Oct 2021 0 repositories listed
-
A Deep Reinforcement Learning Approach for Audio-based Navigation and Audio Source Localization in Multi-speaker Environments25 Oct 2021 0 repositories listed
-
Can Q-Learning be Improved with Advice?25 Oct 2021 0 repositories listed
-
Common Information based Approximate State Representations in Multi-Agent Reinforcement Learning25 Oct 2021 0 repositories listed
-
Learning What to Memorize: Using Intrinsic Motivation to Form Useful Memory in Partially Observable Reinforcement Learning25 Oct 2021 0 repositories listed
-
Operator Shifting for Model-based Policy Evaluation25 Oct 2021 0 repositories listed
-
Self-Consistent Models and Values25 Oct 2021 0 repositories listed
-
Unsupervised Domain Adaptation with Dynamics-Aware Rewards in Reinforcement Learning25 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Simultaneous Sensing and Channel Access in Cognitive Networks24 Oct 2021 0 repositories listed
-
Analysis of Thompson Sampling for Partially Observable Contextual Multi-Armed Bandits23 Oct 2021 0 repositories listed
-
Foresight of Graph Reinforcement Learning Latent Permutations Learnt by Gumbel Sinkhorn Network23 Oct 2021 0 repositories listed
-
Fully Distributed Actor-Critic Architecture for Multitask Deep Reinforcement Learning23 Oct 2021 0 repositories listed
-
A Reinforcement Learning Approach to Parameter Selection for Distributed Optimal Power Flow22 Oct 2021 0 repositories listed
-
Off-policy Reinforcement Learning with Optimistic Exploration and Distribution Correction22 Oct 2021 0 repositories listed
-
Patient level simulation and reinforcement learning to discover novel strategies for treating ovarian cancer22 Oct 2021 0 repositories listed
-
Reinforcement Learning for Process Control with Application in Semiconductor Manufacturing22 Oct 2021 0 repositories listed
-
Anti-Concentrated Confidence Bonuses for Scalable Exploration21 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Online Control of Stochastic Partial Differential Equations21 Oct 2021 0 repositories listed
-
Neuro-Symbolic Reinforcement Learning with First-Order Logic21 Oct 2021 0 repositories listed
-
Off-Dynamics Inverse Reinforcement Learning from Hetero-Domain21 Oct 2021 0 repositories listed
-
Reinforcement Learning Based Optimal Camera Placement for Depth Observation of Indoor Scenes21 Oct 2021 0 repositories listed
-
Computationally Efficient Safe Reinforcement Learning for Power Systems20 Oct 2021 0 repositories listed
-
Distributed Reinforcement Learning for Privacy-Preserving Dynamic Edge Caching20 Oct 2021 0 repositories listed
-
Feedback Linearization of Car Dynamics for Racing via Reinforcement Learning20 Oct 2021 0 repositories listed
-
More Efficient Exploration with Symbolic Priors on Action Sequence Equivalences20 Oct 2021 0 repositories listed
-
Transferring Reinforcement Learning for DC-DC Buck Converter Control via Duty Ratio Mapping: From Simulation to Implementation20 Oct 2021 0 repositories listed
-
Aesthetic Photo Collage with Deep Reinforcement Learning19 Oct 2021 0 repositories listed
-
Improved cooperation by balancing exploration and exploitation in intertemporal social dilemma tasks19 Oct 2021 0 repositories listed
-
Learning Robotic Manipulation Skills Using an Adaptive Force-Impedance Action Space19 Oct 2021 0 repositories listed
-
Locally Differentially Private Reinforcement Learning for Linear Mixture Markov Decision Processes19 Oct 2021 0 repositories listed
-
State-based Episodic Memory for Multi-Agent Reinforcement Learning19 Oct 2021 0 repositories listed
-
Embracing advanced AI/ML to help investors achieve success: Vanguard Reinforcement Learning for Financial Goal Planning18 Oct 2021 0 repositories listed
-
Improving Robustness of Reinforcement Learning for Power System Control with Adversarial Training18 Oct 2021 0 repositories listed
-
Provable Hierarchy-Based Meta-Reinforcement Learning18 Oct 2021 0 repositories listed
-
Reinforcement Learning-Based Coverage Path Planning with Implicit Cellular Decomposition18 Oct 2021 0 repositories listed
-
Damped Anderson Mixing for Deep Reinforcement Learning: Acceleration, Convergence, and Stabilization17 Oct 2021 0 repositories listed
-
Towards Instance-Optimal Offline Reinforcement Learning with Pessimism17 Oct 2021 0 repositories listed
-
Case-based Reasoning for Better Generalization in Textual Reinforcement Learning16 Oct 2021 0 repositories listed
-
Emotion Style Transfer with a Specified Intensity Using Deep Reinforcement Learning16 Oct 2021 0 repositories listed
-
Lifting the veil on hyper-parameters for value-based deep reinforcement learning16 Oct 2021 0 repositories listed
-
Local Advantage Actor-Critic for Robust Multi-Agent Deep Reinforcement Learning16 Oct 2021 0 repositories listed
-
Neural Network Pruning Through Constrained Reinforcement Learning16 Oct 2021 0 repositories listed
-
A Broad-persistent Advising Approach for Deep Interactive Reinforcement Learning in Robotic Environments15 Oct 2021 0 repositories listed
-
Containerized Distributed Value-Based Multi-Agent Reinforcement Learning15 Oct 2021 0 repositories listed
-
Dynamic probabilistic logic models for effective abstractions in RL15 Oct 2021 0 repositories listed
-
Effects of Different Optimization Formulations in Evolutionary Reinforcement Learning on Diverse Behavior Generation15 Oct 2021 0 repositories listed
-
Improving Hyperparameter Optimization by Planning Ahead15 Oct 2021 0 repositories listed
-
On-Policy Model Errors in Reinforcement Learning15 Oct 2021 0 repositories listed
-
Wasserstein Unsupervised Reinforcement Learning15 Oct 2021 0 repositories listed
-
HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism14 Oct 2021 0 repositories listed
-
Offline Reinforcement Learning with Soft Behavior Regularization14 Oct 2021 0 repositories listed
-
Provably Efficient Multi-Agent Reinforcement Learning with Fully Decentralized Communication14 Oct 2021 0 repositories listed
-
Sign and Relevance Learning14 Oct 2021 0 repositories listed
-
Block Contextual MDPs for Continual Learning13 Oct 2021 0 repositories listed
-
Feudal Reinforcement Learning by Reading Manuals13 Oct 2021 0 repositories listed
-
On Covariate Shift of Latent Confounders in Imitation and Reinforcement Learning13 Oct 2021 0 repositories listed
-
Reinforcement Learning for Standards Design13 Oct 2021 0 repositories listed
-
On Improving Model-Free Algorithms for Decentralized Multi-Agent Reinforcement Learning12 Oct 2021 0 repositories listed
-
Deciding What's Fair: Challenges of Applying Reinforcement Learning in Online Marketplaces12 Oct 2021 0 repositories listed
-
GridLearn: Multiagent Reinforcement Learning for Grid-Aware Building Energy Management12 Oct 2021 0 repositories listed
-
Power and Accountability in RL-driven Environmental Policy12 Oct 2021 0 repositories listed
-
Provably Efficient Reinforcement Learning in Decentralized General-Sum Markov Games12 Oct 2021 0 repositories listed
-
Reward-Free Model-Based Reinforcement Learning with Linear Function Approximation12 Oct 2021 0 repositories listed
-
Temporal Abstraction in Reinforcement Learning with the Successor Representation12 Oct 2021 0 repositories listed
-
An Automated Portfolio Trading System with Feature Preprocessing and Recurrent Reinforcement Learning11 Oct 2021 0 repositories listed
-
Bid Optimization using Maximum Entropy Reinforcement Learning11 Oct 2021 0 repositories listed
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning11 Oct 2021 0 repositories listed
-
REIN-2: Giving Birth to Prepared Reinforcement Learning Agents Using Reinforcement Learning Agents11 Oct 2021 0 repositories listed
-
Scalable Traffic Signal Controls using Fog-Cloud Based Multiagent Reinforcement Learning11 Oct 2021 0 repositories listed
-
An Augmented Reality Platform for Introducing Reinforcement Learning to K-12 Students with Robots10 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Optimizing RIS-Assisted HD-FD Wireless Systems10 Oct 2021 0 repositories listed
-
Multi-condition multi-objective optimization using deep reinforcement learning10 Oct 2021 0 repositories listed
-
Reinforcement Learning for Systematic FX Trading10 Oct 2021 0 repositories listed
-
Satisficing Paths and Independent Multi-Agent Reinforcement Learning in Stochastic Games9 Oct 2021 0 repositories listed
-
Breaking the Sample Complexity Barrier to Regret-Optimal Model-Free Reinforcement Learning9 Oct 2021 0 repositories listed
-
Provably Efficient Black-Box Action Poisoning Attacks Against Reinforcement Learning9 Oct 2021 0 repositories listed
-
A MultiModal Social Robot Toward Personalized Emotion Interaction8 Oct 2021 0 repositories listed
-
CheerBots: Chatbots toward Empathy and Emotionusing Reinforcement Learning8 Oct 2021 0 repositories listed
-
Revisiting Design Choices in Offline Model-Based Reinforcement Learning8 Oct 2021 0 repositories listed
-
Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget Matters8 Oct 2021 0 repositories listed
-
A Model Selection Approach for Corruption Robust Reinforcement Learning7 Oct 2021 0 repositories listed
-
Bad-Policy Density: A Measure of Reinforcement Learning Hardness7 Oct 2021 0 repositories listed
-
Designing Composites with Target Effective Young's Modulus using Reinforcement Learning7 Oct 2021 0 repositories listed
-
Explaining Deep Reinforcement Learning Agents In The Atari Domain through a Surrogate Model7 Oct 2021 0 repositories listed