Browse State-of-the-Art › reinforcement-learning › Papers, page 58
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 58 of 135: papers 5,701 to 5,800 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning5 Feb 2024 0 repositories listed
-
A Reinforcement Learning Approach for Dynamic Rebalancing in Bike-Sharing System5 Feb 2024 0 repositories listed
-
Abstracted Trajectory Visualization for Explainability in Reinforcement Learning5 Feb 2024 0 repositories listed
-
Assessing the Impact of Distribution Shift on Reinforcement Learning Performance5 Feb 2024 0 repositories listed
-
Curriculum reinforcement learning for quantum architecture search under hardware errors5 Feb 2024 0 repositories listed
-
Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning5 Feb 2024 0 repositories listed
-
Deep Reinforcement Learning for Picker Routing Problem in Warehousing5 Feb 2024 0 repositories listed
-
Learning to Abstract Visuomotor Mappings using Meta-Reinforcement Learning5 Feb 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning for Offloading Cellular Communications with Cooperating UAVs5 Feb 2024 0 repositories listed
-
Deep Exploration with PAC-Bayes5 Feb 2024 0 repositories listed
-
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning5 Feb 2024 0 repositories listed
-
Vision-Language Models Provide Promptable Representations for Reinforcement Learning5 Feb 2024 0 repositories listed
-
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping4 Feb 2024 0 repositories listed
-
DiffStitch: Boosting Offline Reinforcement Learning with Diffusion-based Trajectory Stitching4 Feb 2024 0 repositories listed
-
Evading Deep Learning-Based Malware Detectors via Obfuscation: A Deep Reinforcement Learning Approach4 Feb 2024 0 repositories listed
-
Interference-Aware Emergent Random Access Protocol for Downlink LEO Satellite Networks4 Feb 2024 0 repositories listed
-
The Virtues of Pessimism in Inverse Reinforcement Learning4 Feb 2024 0 repositories listed
-
A Survey of Constraint Formulations in Safe Reinforcement Learning3 Feb 2024 0 repositories listed
-
Brain-Like Replay Naturally Emerges in Reinforcement Learning Agents2 Feb 2024 0 repositories listed
-
Efficient Reinforcement Learning for Routing Jobs in Heterogeneous Queueing Systems2 Feb 2024 0 repositories listed
-
Near-Optimal Reinforcement Learning with Self-Play under Adaptivity Constraints2 Feb 2024 0 repositories listed
-
Adaptive Primal-Dual Method for Safe Reinforcement Learning1 Feb 2024 0 repositories listed
-
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach1 Feb 2024 0 repositories listed
-
Control-Theoretic Techniques for Online Adaptation of Deep Neural Networks in Dynamical Systems1 Feb 2024 0 repositories listed
-
Deep Robot Sketching: An application of Deep Q-Learning Networks for human-like sketching1 Feb 2024 0 repositories listed
-
Augmenting Offline Reinforcement Learning with State-only Interactions1 Feb 2024 0 repositories listed
-
FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game1 Feb 2024 0 repositories listed
-
Neural Policy Style Transfer1 Feb 2024 0 repositories listed
-
A Reinforcement Learning Based Controller to Minimize Forces on the Crutches of a Lower-Limb Exoskeleton31 Jan 2024 0 repositories listed
-
Attention Graph for Multi-Robot Social Navigation with Deep Reinforcement Learning31 Jan 2024 0 repositories listed
-
Causal Coordinated Concurrent Reinforcement Learning31 Jan 2024 0 repositories listed
-
Circuit Partitioning for Multi-Core Quantum Architectures with Deep Reinforcement Learning31 Jan 2024 0 repositories listed
-
Decentralized Covert Routing in Heterogeneous Networks Using Reinforcement Learning31 Jan 2024 0 repositories listed
-
Enhancing End-to-End Multi-Task Dialogue Systems: A Study on Intrinsic Motivation Reinforcement Learning Algorithms for Improved Training and Adaptability31 Jan 2024 0 repositories listed
-
Graph Attention-based Reinforcement Learning for Trajectory Design and Resource Assignment in Multi-UAV Assisted Communication31 Jan 2024 0 repositories listed
-
Scheduled Curiosity-Deep Dyna-Q: Efficient Exploration for Dialog Policy Learning31 Jan 2024 0 repositories listed
-
Extrinsicaly Rewarded Soft Q Imitation Learning with Discriminator30 Jan 2024 0 repositories listed
-
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble30 Jan 2024 0 repositories listed
-
Reinforcement Learning for Versatile, Dynamic, and Robust Bipedal Locomotion Control30 Jan 2024 0 repositories listed
-
Attention-based Reinforcement Learning for Combinatorial Optimization: Application to Job Shop Scheduling Problem29 Jan 2024 0 repositories listed
-
Emergence of cooperation under punishment: A reinforcement learning perspective29 Jan 2024 0 repositories listed
-
Optimal Control of Renewable Energy Communities subject to Network Peak Fees with Model Predictive Control and Reinforcement Learning Algorithms29 Jan 2024 0 repositories listed
-
A Strategy for Preparing Quantum Squeezed States Using Reinforcement Learning29 Jan 2024 0 repositories listed
-
Scalable Reinforcement Learning for Linear-Quadratic Control of Networks29 Jan 2024 0 repositories listed
-
SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning29 Jan 2024 0 repositories listed
-
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning27 Jan 2024 0 repositories listed
-
Learning to Trust Your Feelings: Leveraging Self-awareness in LLMs for Hallucination Mitigation27 Jan 2024 0 repositories listed
-
Social Interpretable Reinforcement Learning27 Jan 2024 0 repositories listed
-
Hierarchical Continual Reinforcement Learning via Large Language Model25 Jan 2024 0 repositories listed
-
Modeling and Optimization of Epidemiological Control Policies Through Reinforcement Learning25 Jan 2024 0 repositories listed
-
Peer-to-Peer Energy Trading of Solar and Energy Storage: A Networked Multiagent Reinforcement Learning Approach25 Jan 2024 0 repositories listed
-
Sample Efficient Reinforcement Learning by Automatically Learning to Compose Subtasks25 Jan 2024 0 repositories listed
-
Scilab-RL: A software framework for efficient reinforcement learning and cognitive modeling research25 Jan 2024 0 repositories listed
-
Symbolic Equation Solving via Reinforcement Learning24 Jan 2024 0 repositories listed
-
A Novel Policy Iteration Algorithm for Nonlinear Continuous-Time H∞ Control Problem23 Jan 2024 0 repositories listed
-
A Safe Reinforcement Learning Algorithm for Supervisory Control of Power Plants23 Jan 2024 0 repositories listed
-
Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning23 Jan 2024 0 repositories listed
-
Emergent Communication Protocol Learning for Task Offloading in Industrial Internet of Things23 Jan 2024 0 repositories listed
-
Model-Free δ-Policy Iteration Based on Damped Newton Method for Nonlinear Continuous-Time H∞ Tracking Control23 Jan 2024 0 repositories listed
-
Reward-Relevance-Filtered Linear Offline Reinforcement Learning23 Jan 2024 0 repositories listed
-
HomeRobot Open Vocabulary Mobile Manipulation Challenge 2023 Participant Report (Team KuzHum)22 Jan 2024 0 repositories listed
-
Mitigating Covariate Shift in Misspecified Regression with Applications to Reinforcement Learning22 Jan 2024 0 repositories listed
-
P2DT: Mitigating Forgetting in task-incremental Learning with progressive prompt Decision Transformer22 Jan 2024 0 repositories listed
-
Retrieval-Guided Reinforcement Learning for Boolean Circuit Minimization22 Jan 2024 0 repositories listed
-
Efficient and Generalized end-to-end Autonomous Driving System with Latent Deep Reinforcement Learning and Demonstrations22 Jan 2024 0 repositories listed
-
Constrained Reinforcement Learning for Adaptive Controller Synchronization in Distributed SDN21 Jan 2024 0 repositories listed
-
MoMA: Model-based Mirror Ascent for Offline Reinforcement Learning21 Jan 2024 0 repositories listed
-
Large-scale Reinforcement Learning for Diffusion Models20 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning Empowered Activity-Aware Dynamic Health Monitoring Systems19 Jan 2024 0 repositories listed
-
Episodic Reinforcement Learning with Expanded State-reward Space19 Jan 2024 0 repositories listed
-
Stochastic Dynamic Power Dispatch with High Generalization and Few-Shot Adaption via Contextual Meta Graph Reinforcement Learning19 Jan 2024 0 repositories listed
-
Harnessing Density Ratios for Online Reinforcement Learning18 Jan 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning for Maritime Operational Technology Cyber Security18 Jan 2024 0 repositories listed
-
Cascading Reinforcement Learning17 Jan 2024 0 repositories listed
-
Continuous Time Continuous Space Homeostatic Reinforcement Learning (CTCS-HRRL) : Towards Biological Self-Autonomous Agent17 Jan 2024 0 repositories listed
-
Towards Off-Policy Reinforcement Learning for Ranking Policies with Human Feedback17 Jan 2024 0 repositories listed
-
CNN-DRL with Shuffled Features in Finance16 Jan 2024 0 repositories listed
-
CPPO: Continual Learning for Reinforcement Learning with Human Feedback16 Jan 2024 0 repositories listed
-
On Quantum Natural Policy Gradients16 Jan 2024 0 repositories listed
-
PRewrite: Prompt Rewriting with Reinforcement Learning16 Jan 2024 0 repositories listed
-
Reinforcement Learning for Conversational Question Answering over Knowledge Graph16 Jan 2024 0 repositories listed
-
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes16 Jan 2024 0 repositories listed
-
Solving Continual Offline Reinforcement Learning with Decision Transformer16 Jan 2024 0 repositories listed
-
Constrained Multi-objective Optimization with Deep Reinforcement Learning Assisted Operator Selection15 Jan 2024 0 repositories listed
-
Go-Explore for Residential Energy Management15 Jan 2024 0 repositories listed
-
The ODE Method for Stochastic Approximation and Reinforcement Learning with Markovian Noise15 Jan 2024 0 repositories listed
-
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions14 Jan 2024 0 repositories listed
-
Beyond Sparse Rewards: Enhancing Reinforcement Learning with Language Model Critique in Text Generation14 Jan 2024 0 repositories listed
-
Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization14 Jan 2024 0 repositories listed
-
A Reinforcement Learning Environment for Directed Quantum Circuit Synthesis13 Jan 2024 0 repositories listed
-
Code Security Vulnerability Repair Using Reinforcement Learning with Large Language Models13 Jan 2024 0 repositories listed
-
Discovering Command and Control Channels Using Reinforcement Learning13 Jan 2024 0 repositories listed
-
Quantum Advantage Actor-Critic for Reinforcement Learning13 Jan 2024 0 repositories listed
-
Reinforcement Learning for Scalable Train Timetable Rescheduling with Graph Representation13 Jan 2024 0 repositories listed
-
Identifying Policy Gradient Subspaces12 Jan 2024 0 repositories listed
-
Maximum Causal Entropy Inverse Reinforcement Learning for Mean-Field Games12 Jan 2024 0 repositories listed
-
UNEX-RL: Reinforcing Long-Term Rewards in Multi-Stage Recommender Systems with UNidirectional EXecution12 Jan 2024 0 repositories listed
-
Bounds on the price of feedback for mistake-bounded online learning11 Jan 2024 0 repositories listed
-
Model-Free Reinforcement Learning for Automated Fluid Administration in Critical Care11 Jan 2024 0 repositories listed
-
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem11 Jan 2024 0 repositories listed