Browse State-of-the-Art › reinforcement-learning › Papers, page 70
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 70 of 135: papers 6,901 to 7,000 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Sim-and-Real Reinforcement Learning for Manipulation: A Consensus-based Approach26 Feb 2023 0 repositories listed
-
A Human-Centered Safe Robot Reinforcement Learning Framework with Interactive Behaviors25 Feb 2023 0 repositories listed
-
Exponential Hardness of Reinforcement Learning with Linear Function Approximation25 Feb 2023 0 repositories listed
-
On Bellman's principle of optimality and Reinforcement learning for safety-constrained Markov decision process25 Feb 2023 0 repositories listed
-
AC2C: Adaptively Controlled Two-Hop Communication for Multi-Agent Reinforcement Learning24 Feb 2023 0 repositories listed
-
Finding Regularized Competitive Equilibria of Heterogeneous Agent Macroeconomic Models with Reinforcement Learning24 Feb 2023 0 repositories listed
-
Leveraging Jumpy Models for Planning and Fast Learning in Robotic Domains24 Feb 2023 0 repositories listed
-
Logarithmic Switching Cost in Reinforcement Learning beyond Linear MDPs24 Feb 2023 0 repositories listed
-
Multi-Agent Reinforcement Learning with Common Policy for Antenna Tilt Optimization24 Feb 2023 0 repositories listed
-
Concept Learning for Interpretable Multi-Agent Reinforcement Learning23 Feb 2023 0 repositories listed
-
Provably Efficient Reinforcement Learning via Surprise Bound22 Feb 2023 0 repositories listed
-
A Reinforcement Learning Framework for Online Speaker Diarization21 Feb 2023 0 repositories listed
-
Adversarial Model for Offline Reinforcement Learning21 Feb 2023 0 repositories listed
-
BadGPT: Exploring Security Vulnerabilities of ChatGPT via Backdoor Attacks to InstructGPT21 Feb 2023 0 repositories listed
-
Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning21 Feb 2023 0 repositories listed
-
Constrained Reinforcement Learning for Predictive Control in Real-Time Stochastic Dynamic Optimal Power Flow21 Feb 2023 0 repositories listed
-
Curiosity-driven Exploration in Sparse-reward Multi-agent Reinforcement Learning21 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Robotic Pushing and Picking in Cluttered Environment21 Feb 2023 0 repositories listed
-
Offline Reinforcement Learning for Mixture-of-Expert Dialogue Management21 Feb 2023 0 repositories listed
-
Provably Efficient Exploration in Quantum Reinforcement Learning with Logarithmic Worst-Case Regret21 Feb 2023 0 repositories listed
-
Reinforcement Learning for Block Decomposition of CAD Models21 Feb 2023 0 repositories listed
-
Reinforcement Learning in a Birth and Death Process: Breaking the Dependence on the State Space21 Feb 2023 0 repositories listed
-
Robust Auto-landing Control of an agile Regional Jet Using Fuzzy Q-learning21 Feb 2023 0 repositories listed
-
UAV Path Planning Employing MPC- Reinforcement Learning Method Considering Collision Avoidance21 Feb 2023 0 repositories listed
-
Differentiable Arbitrating in Zero-sum Markov Games20 Feb 2023 0 repositories listed
-
Reinforcement Learning with Function Approximation: From Linear to Nonlinear20 Feb 2023 0 repositories listed
-
Safe Deep Reinforcement Learning by Verifying Task-Level Properties20 Feb 2023 0 repositories listed
-
AutoDOViz: Human-Centered Automation for Decision Optimization19 Feb 2023 0 repositories listed
-
Compositionality and Bounds for Optimal Value Functions in Reinforcement Learning19 Feb 2023 0 repositories listed
-
Interactive Video Corpus Moment Retrieval using Reinforcement Learning19 Feb 2023 0 repositories listed
-
Robust and Versatile Bipedal Jumping Control through Reinforcement Learning19 Feb 2023 0 repositories listed
-
Effective Multimodal Reinforcement Learning with Modality Alignment and Importance Enhancement18 Feb 2023 0 repositories listed
-
Efficient Exploration via Epistemic-Risk-Seeking Policy Optimization18 Feb 2023 0 repositories listed
-
Promoting Cooperation in Multi-Agent Reinforcement Learning via Mutual Help18 Feb 2023 0 repositories listed
-
Reinforcement Learning in the Wild with Maximum Likelihood-based Model Transfer18 Feb 2023 0 repositories listed
-
A State Augmentation based approach to Reinforcement Learning from Human Preferences17 Feb 2023 0 repositories listed
-
Data Driven Reward Initialization for Preference based Reinforcement Learning17 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for mmWave Initial Beam Alignment17 Feb 2023 0 repositories listed
-
Exploiting Unlabeled Data for Feedback Efficient Human Preference based Reinforcement Learning17 Feb 2023 0 repositories listed
-
Learning to Forecast Aleatoric and Epistemic Uncertainties over Long Horizon Trajectories17 Feb 2023 0 repositories listed
-
Robot path planning using deep reinforcement learning17 Feb 2023 0 repositories listed
-
Quantum Computing Provides Exponential Regret Improvement in Episodic Reinforcement Learning16 Feb 2023 0 repositories listed
-
CERiL: Continuous Event-based Reinforcement Learning15 Feb 2023 0 repositories listed
-
Deep Offline Reinforcement Learning for Real-world Treatment Optimization Applications15 Feb 2023 0 repositories listed
-
Meta-Reinforcement Learning via Exploratory Task Clustering15 Feb 2023 0 repositories listed
-
Optimal Sample Complexity of Reinforcement Learning for Mixing Discounted Markov Decision Processes15 Feb 2023 0 repositories listed
-
Prioritized offline Goal-swapping Experience Replay15 Feb 2023 0 repositories listed
-
Reinforcement Learning Based Power Grid Day-Ahead Planning and AI-Assisted Control15 Feb 2023 0 repositories listed
-
Scalable Multi-Agent Reinforcement Learning with General Utilities15 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Multi-user Massive MIMO with Channel Aging14 Feb 2023 0 repositories listed
-
Quantum algorithms applied to satellite mission planning for Earth observation14 Feb 2023 0 repositories listed
-
To Risk or Not to Risk: Learning with Risk Quantification for IoT Task Offloading in UAVs14 Feb 2023 0 repositories listed
-
A Lifetime Extended Energy Management Strategy for Fuel Cell Hybrid Electric Vehicles via Self-Learning Fuzzy Reinforcement Learning13 Feb 2023 0 repositories listed
-
Provably Safe Reinforcement Learning with Step-wise Violation Constraints13 Feb 2023 0 repositories listed
-
Maneuver Decision-Making For Autonomous Air Combat Through Curriculum Learning And Reinforcement Learning With Sparse Rewards12 Feb 2023 0 repositories listed
-
ReMIX: Regret Minimization for Monotonic Value Function Factorization in Multiagent Reinforcement Learning11 Feb 2023 0 repositories listed
-
A Survey on Causal Reinforcement Learning10 Feb 2023 0 repositories listed
-
Low Entropy Communication in Multi-Agent Reinforcement Learning10 Feb 2023 0 repositories listed
-
Towards Minimax Optimality of Model-based Robust Reinforcement Learning10 Feb 2023 0 repositories listed
-
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning9 Feb 2023 0 repositories listed
-
CLARE: Conservative Model-Based Reward Learning for Offline Inverse Reinforcement Learning9 Feb 2023 0 repositories listed
-
Data Quality-aware Mixed-precision Quantization via Hybrid Reinforcement Learning9 Feb 2023 0 repositories listed
-
Equivariant MuZero9 Feb 2023 0 repositories listed
-
A Near-Optimal Algorithm for Safe Reinforcement Learning Under Instantaneous Hard Constraints8 Feb 2023 0 repositories listed
-
A Scale-Independent Multi-Objective Reinforcement Learning with Convergence Analysis8 Feb 2023 0 repositories listed
-
AISYN: AI-driven Reinforcement Learning-Based Logic Synthesis Framework8 Feb 2023 0 repositories listed
-
Efficient Planning in Combinatorial Action Spaces with Applications to Cooperative Multi-Agent Reinforcement Learning8 Feb 2023 0 repositories listed
-
Near-Optimal Adversarial Reinforcement Learning with Switching Costs8 Feb 2023 0 repositories listed
-
Adaptive Aggregation for Safety-Critical Control7 Feb 2023 0 repositories listed
-
Ensemble Value Functions for Efficient Exploration in Multi-Agent Reinforcement Learning7 Feb 2023 0 repositories listed
-
Near-Minimax-Optimal Risk-Sensitive Reinforcement Learning with CVaR7 Feb 2023 0 repositories listed
-
Online Reinforcement Learning with Uncertain Episode Lengths7 Feb 2023 0 repositories listed
-
Optimizing Audio Recommendations for the Long-Term: A Reinforcement Learning Perspective7 Feb 2023 0 repositories listed
-
Towards Skilled Population Curriculum for Multi-Agent Reinforcement Learning7 Feb 2023 0 repositories listed
-
Transfer learning for process design with reinforcement learning7 Feb 2023 0 repositories listed
-
Arena-Web -- A Web-based Development and Benchmarking Platform for Autonomous Navigation Approaches6 Feb 2023 0 repositories listed
-
DITTO: Offline Imitation Learning with World Models6 Feb 2023 0 repositories listed
-
Holistic Deep-Reinforcement-Learning-based Training of Autonomous Navigation Systems6 Feb 2023 0 repositories listed
-
RLTP: Reinforcement Learning to Pace for Delayed Impression Modeling in Preloaded Ads6 Feb 2023 0 repositories listed
-
State-wise Safe Reinforcement Learning: A Survey6 Feb 2023 0 repositories listed
-
An Online Model-Following Projection Mechanism Using Reinforcement Learning5 Feb 2023 0 repositories listed
-
Open Problems and Modern Solutions for Deep Reinforcement Learning5 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Traffic Light Control in Intelligent Transportation Systems4 Feb 2023 0 repositories listed
-
Generalization of Deep Reinforcement Learning for Jammer-Resilient Frequency and Power Allocation4 Feb 2023 0 repositories listed
-
Developing Driving Strategies Efficiently: A Skill-Based Hierarchical Reinforcement Learning Approach4 Feb 2023 0 repositories listed
-
Reinforcement Learning in Low-Rank MDPs with Density Features4 Feb 2023 0 repositories listed
-
Reinforcement Learning with History-Dependent Dynamic Contexts4 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Cyber System Defense under Dynamic Adversarial Uncertainties3 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Online Error Detection in Cyber-Physical Systems3 Feb 2023 0 repositories listed
-
Reinforcing User Retention in a Billion Scale Short Video Recommender System3 Feb 2023 0 repositories listed
-
Diversity Through Exclusion (DTE): Niche Identification for Reinforcement Learning through Value-Decomposition2 Feb 2023 0 repositories listed
-
Performance Bounds for Policy-Based Average Reward Reinforcement Learning Algorithms2 Feb 2023 0 repositories listed
-
ReLOAD: Reinforcement Learning with Optimistic Ascent-Descent for Last-Iterate Convergence in Constrained MDPs2 Feb 2023 0 repositories listed
-
Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO)1 Feb 2023 0 repositories listed
-
Combining Deep Reinforcement Learning and Search with Generative Models for Game-Theoretic Opponent Modeling1 Feb 2023 0 repositories listed
-
Multi-zone HVAC Control with Model-Based Deep Reinforcement Learning1 Feb 2023 0 repositories listed
-
Robust Fitted-Q-Evaluation and Iteration under Sequentially Exogenous Unobserved Confounders1 Feb 2023 0 repositories listed
-
Selective Uncertainty Propagation in Offline RL1 Feb 2023 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.