Browse State-of-the-Art › Reinforcement Learning › Papers, page 67
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 67 of 132: papers 6,601 to 6,700 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Robot path planning using deep reinforcement learning17 Feb 2023 0 repositories listed
-
Quantum Computing Provides Exponential Regret Improvement in Episodic Reinforcement Learning16 Feb 2023 0 repositories listed
-
CERiL: Continuous Event-based Reinforcement Learning15 Feb 2023 0 repositories listed
-
Deep Offline Reinforcement Learning for Real-world Treatment Optimization Applications15 Feb 2023 0 repositories listed
-
Meta-Reinforcement Learning via Exploratory Task Clustering15 Feb 2023 0 repositories listed
-
Scalable Multi-Agent Reinforcement Learning with General Utilities15 Feb 2023 0 repositories listed
-
To Risk or Not to Risk: Learning with Risk Quantification for IoT Task Offloading in UAVs14 Feb 2023 0 repositories listed
-
A Lifetime Extended Energy Management Strategy for Fuel Cell Hybrid Electric Vehicles via Self-Learning Fuzzy Reinforcement Learning13 Feb 2023 0 repositories listed
-
Provably Safe Reinforcement Learning with Step-wise Violation Constraints13 Feb 2023 0 repositories listed
-
Maneuver Decision-Making For Autonomous Air Combat Through Curriculum Learning And Reinforcement Learning With Sparse Rewards12 Feb 2023 0 repositories listed
-
A Survey on Causal Reinforcement Learning10 Feb 2023 0 repositories listed
-
Low Entropy Communication in Multi-Agent Reinforcement Learning10 Feb 2023 0 repositories listed
-
Towards Minimax Optimality of Model-based Robust Reinforcement Learning10 Feb 2023 0 repositories listed
-
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning9 Feb 2023 0 repositories listed
-
Data Quality-aware Mixed-precision Quantization via Hybrid Reinforcement Learning9 Feb 2023 0 repositories listed
-
A Scale-Independent Multi-Objective Reinforcement Learning with Convergence Analysis8 Feb 2023 0 repositories listed
-
AISYN: AI-driven Reinforcement Learning-Based Logic Synthesis Framework8 Feb 2023 0 repositories listed
-
Efficient Planning in Combinatorial Action Spaces with Applications to Cooperative Multi-Agent Reinforcement Learning8 Feb 2023 0 repositories listed
-
Near-Optimal Adversarial Reinforcement Learning with Switching Costs8 Feb 2023 0 repositories listed
-
Adaptive Aggregation for Safety-Critical Control7 Feb 2023 0 repositories listed
-
Near-Minimax-Optimal Risk-Sensitive Reinforcement Learning with CVaR7 Feb 2023 0 repositories listed
-
Online Reinforcement Learning with Uncertain Episode Lengths7 Feb 2023 0 repositories listed
-
Towards Skilled Population Curriculum for Multi-Agent Reinforcement Learning7 Feb 2023 0 repositories listed
-
Transfer learning for process design with reinforcement learning7 Feb 2023 0 repositories listed
-
Arena-Web -- A Web-based Development and Benchmarking Platform for Autonomous Navigation Approaches6 Feb 2023 0 repositories listed
-
DITTO: Offline Imitation Learning with World Models6 Feb 2023 0 repositories listed
-
Holistic Deep-Reinforcement-Learning-based Training of Autonomous Navigation Systems6 Feb 2023 0 repositories listed
-
State-wise Safe Reinforcement Learning: A Survey6 Feb 2023 0 repositories listed
-
An Online Model-Following Projection Mechanism Using Reinforcement Learning5 Feb 2023 0 repositories listed
-
Open Problems and Modern Solutions for Deep Reinforcement Learning5 Feb 2023 0 repositories listed
-
Generalization of Deep Reinforcement Learning for Jammer-Resilient Frequency and Power Allocation4 Feb 2023 0 repositories listed
-
Developing Driving Strategies Efficiently: A Skill-Based Hierarchical Reinforcement Learning Approach4 Feb 2023 0 repositories listed
-
Reinforcement Learning in Low-Rank MDPs with Density Features4 Feb 2023 0 repositories listed
-
Reinforcement Learning with History-Dependent Dynamic Contexts4 Feb 2023 0 repositories listed
-
Reinforcing User Retention in a Billion Scale Short Video Recommender System3 Feb 2023 0 repositories listed
-
Performance Bounds for Policy-Based Average Reward Reinforcement Learning Algorithms2 Feb 2023 0 repositories listed
-
ReLOAD: Reinforcement Learning with Optimistic Ascent-Descent for Last-Iterate Convergence in Constrained MDPs2 Feb 2023 0 repositories listed
-
Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO)1 Feb 2023 0 repositories listed
-
Combining Deep Reinforcement Learning and Search with Generative Models for Game-Theoretic Opponent Modeling1 Feb 2023 0 repositories listed
-
Multi-zone HVAC Control with Model-Based Deep Reinforcement Learning1 Feb 2023 0 repositories listed
-
Robust Fitted-Q-Evaluation and Iteration under Sequentially Exogenous Unobserved Confounders1 Feb 2023 0 repositories listed
-
Enabling surrogate-assisted evolutionary reinforcement learning via policy embedding31 Jan 2023 0 repositories listed
-
Scalable Grid-Aware Dynamic Matching using Deep Reinforcement Learning31 Jan 2023 0 repositories listed
-
Scaling laws for single-agent reinforcement learning31 Jan 2023 0 repositories listed
-
Scheduling Inference Workloads on Distributed Edge Clusters with Reinforcement Learning31 Jan 2023 0 repositories listed
-
Hierarchical Programmatic Reinforcement Learning via Learning to Compose Programs30 Jan 2023 0 repositories listed
-
STEEL: Singularity-aware Reinforcement Learning30 Jan 2023 0 repositories listed
-
Transferring Multiple Policies to Hotstart Reinforcement Learning in an Air Compressor Management Problem30 Jan 2023 0 repositories listed
-
V2N Service Scaling with Deep Reinforcement Learning30 Jan 2023 0 repositories listed
-
A Deep Reinforcement Learning Framework for Optimizing Congestion Control in Data Centers29 Jan 2023 0 repositories listed
-
Sample Efficient Deep Reinforcement Learning via Local Planning29 Jan 2023 0 repositories listed
-
STEERING: Stein Information Directed Exploration for Model-Based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Towards Learning Rubik's Cube with N-tuple-based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Exploring Deep Reinforcement Learning for Holistic Smart Building Control27 Jan 2023 0 repositories listed
-
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence27 Jan 2023 0 repositories listed
-
Reinforcement Learning from Diverse Human Preferences27 Jan 2023 0 repositories listed
-
Single-Trajectory Distributionally Robust Reinforcement Learning27 Jan 2023 0 repositories listed
-
SNeRL: Semantic-aware Neural Radiance Fields for Reinforcement Learning27 Jan 2023 0 repositories listed
-
FedHQL: Federated Heterogeneous Q-Learning26 Jan 2023 0 repositories listed
-
Model-based Offline Reinforcement Learning with Local Misspecification26 Jan 2023 0 repositories listed
-
Certifiably Robust Reinforcement Learning through Model-Based Abstract Interpretation26 Jan 2023 0 repositories listed
-
Principled Reinforcement Learning with Human Feedback from Pairwise or K-wise Comparisons26 Jan 2023 0 repositories listed
-
ASQ-IT: Interactive Explanations for Reinforcement-Learning Agents24 Jan 2023 0 repositories listed
-
AutoCost: Evolving Intrinsic Cost for Zero-violation Reinforcement Learning24 Jan 2023 0 repositories listed
-
Explainable Deep Reinforcement Learning: State of the Art and Challenges24 Jan 2023 0 repositories listed
-
Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review24 Jan 2023 0 repositories listed
-
Minimal Value-Equivalent Partial Models for Scalable and Robust Planning in Lifelong Reinforcement Learning24 Jan 2023 0 repositories listed
-
Story Shaping: Teaching Agents Human-like Behavior with Stories24 Jan 2023 0 repositories listed
-
Quasi-optimal Reinforcement Learning with Continuous Actions21 Jan 2023 0 repositories listed
-
Asynchronous Deep Double Duelling Q-Learning for Trading-Signal Execution in Limit Order Book Markets20 Jan 2023 0 repositories listed
-
Resource Optimization for Semantic-Aware Networks with Task Offloading20 Jan 2023 0 repositories listed
-
Generative Slate Recommendation with Reinforcement Learning20 Jan 2023 0 repositories listed
-
Multi-agent Reinforcement Learning with Graph Q-Networks for Antenna Tuning20 Jan 2023 0 repositories listed
-
Reinforcement learning-based estimation for partial differential equations20 Jan 2023 0 repositories listed
-
Revisiting Estimation Bias in Policy Gradients for Deep Reinforcement Learning20 Jan 2023 0 repositories listed
-
A Survey of Meta-Reinforcement Learning19 Jan 2023 0 repositories listed
-
Advanced Scaling Methods for VNF deployment with Reinforcement Learning19 Jan 2023 0 repositories listed
-
Human-Timescale Adaptation in an Open-Ended Task Space18 Jan 2023 0 repositories listed
-
Multi-compartment Neuron and Population Encoding Powered Spiking Neural Network for Deep Distributional Reinforcement Learning18 Jan 2023 0 repositories listed
-
Adversarial Robust Deep Reinforcement Learning Requires Redefining Robustness17 Jan 2023 0 repositories listed
-
DQNAS: Neural Architecture Search using Reinforcement Learning17 Jan 2023 0 repositories listed
-
Neuro-Symbolic World Models for Adapting to Open World Novelty16 Jan 2023 0 repositories listed
-
Neuro-symbolic Meta Reinforcement Learning for Trading15 Jan 2023 0 repositories listed
-
Risk-Averse Reinforcement Learning via Dynamic Time-Consistent Risk Measures14 Jan 2023 0 repositories listed
-
Asynchronous training of quantum reinforcement learning12 Jan 2023 0 repositories listed
-
An Analysis of Quantile Temporal-Difference Learning11 Jan 2023 0 repositories listed
-
Efficient Preference-Based Reinforcement Learning Using Learned Dynamics Models11 Jan 2023 0 repositories listed
-
SoK: Adversarial Machine Learning Attacks and Defences in Multi-Agent Reinforcement Learning11 Jan 2023 0 repositories listed
-
Switchable Lightweight Anti-symmetric Processing (SLAP) with CNN Outspeeds Data Augmentation by Smaller Sample -- Application in Gomoku Reinforcement Learning11 Jan 2023 0 repositories listed
-
Actor-Director-Critic: A Novel Deep Reinforcement Learning Framework10 Jan 2023 0 repositories listed
-
Exploration in Model-based Reinforcement Learning with Randomized Reward9 Jan 2023 0 repositories listed
-
Network Slicing via Transfer Learning aided Distributed Deep Reinforcement Learning9 Jan 2023 0 repositories listed
-
Reinforcement Learning Enhanced PicHunter for Interactive Search9 Jan 2023 0 repositories listed
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning9 Jan 2023 0 repositories listed
-
A Survey on Transformers in Reinforcement Learning8 Jan 2023 0 repositories listed
-
Learning Symbolic Representations for Reinforcement Learning of Non-Markovian Behavior8 Jan 2023 0 repositories listed
-
Hierarchical Reinforcement Learning for RIS-Assisted Energy-Efficient RAN7 Jan 2023 0 repositories listed
-
Mathematical Models and Reinforcement Learning based Evolutionary Algorithm Framework for Satellite Scheduling Problem7 Jan 2023 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.