Browse State-of-the-Art › Reinforcement Learning › Papers, page 53
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 53 of 132: papers 5,201 to 5,300 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Deep Reinforcement Learning for Multi-Objective Optimization: Enhancing Wind Turbine Energy Generation while Mitigating Noise Emissions18 Jul 2024 0 repositories listed
-
LIMT: Language-Informed Multi-Task Visual World Models18 Jul 2024 0 repositories listed
-
Model-based Policy Optimization using Symbolic World Model18 Jul 2024 0 repositories listed
-
Optimistic Q-learning for average reward and episodic reinforcement learning18 Jul 2024 0 repositories listed
-
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods18 Jul 2024 0 repositories listed
-
Random Latent Exploration for Deep Reinforcement Learning18 Jul 2024 0 repositories listed
-
Reinforcement Learning: Tutorial and Survey18 Jul 2024 0 repositories listed
-
Estimating Reaction Barriers with Deep Reinforcement Learning17 Jul 2024 0 repositories listed
-
Maintenance Strategies for Sewer Pipes with Multi-State Degradation and Deep Reinforcement Learning17 Jul 2024 0 repositories listed
-
Sparsity-based Safety Conservatism for Constrained Offline Reinforcement Learning17 Jul 2024 0 repositories listed
-
Subequivariant Reinforcement Learning in 3D Multi-Entity Physical Environments17 Jul 2024 0 repositories listed
-
Bellman Diffusion Models16 Jul 2024 0 repositories listed
-
Satisficing Exploration for Deep Reinforcement Learning16 Jul 2024 0 repositories listed
-
Why long model-based rollouts are no reason for bad Q-value estimates16 Jul 2024 0 repositories listed
-
Exploration in Knowledge Transfer Utilizing Reinforcement Learning15 Jul 2024 0 repositories listed
-
Offline Reinforcement Learning with Imputed Rewards15 Jul 2024 0 repositories listed
-
Three Dogmas of Reinforcement Learning15 Jul 2024 0 repositories listed
-
Walking the Values in Bayesian Inverse Reinforcement Learning15 Jul 2024 0 repositories listed
-
Affordance-Guided Reinforcement Learning via Visual Prompting14 Jul 2024 0 repositories listed
-
Ontology-driven Reinforcement Learning for Personalized Student Support14 Jul 2024 0 repositories listed
-
Towards Adapting Reinforcement Learning Agents to New Tasks: Insights from Q-Values14 Jul 2024 0 repositories listed
-
Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control12 Jul 2024 0 repositories listed
-
Decentralized multi-agent reinforcement learning algorithm using a cluster-synchronized laser network12 Jul 2024 0 repositories listed
-
Hamilton-Jacobi Reachability in Reinforcement Learning: A Survey12 Jul 2024 0 repositories listed
-
A Cantor-Kantorovich Metric Between Markov Decision Processes with Application to Transfer Learning11 Jul 2024 0 repositories listed
-
A Review of Nine Physics Engines for Reinforcement Learning Research11 Jul 2024 0 repositories listed
-
Hierarchical Consensus-Based Multi-Agent Reinforcement Learning for Multi-Robot Cooperation Tasks11 Jul 2024 0 repositories listed
-
Continuous Control with Coarse-to-fine Reinforcement Learning10 Jul 2024 0 repositories listed
-
Leveraging LLMs to explain DRL decisions for transparent 6G network slicing10 Jul 2024 0 repositories listed
-
Pessimism Meets Risk: Risk-Sensitive Offline Reinforcement Learning10 Jul 2024 0 repositories listed
-
Real-time system optimal traffic routing under uncertainties -- Can physics models boost reinforcement learning?10 Jul 2024 0 repositories listed
-
Reinforcement Learning of Adaptive Acquisition Policies for Inverse Problems10 Jul 2024 0 repositories listed
-
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes9 Jul 2024 0 repositories listed
-
Intercepting Unauthorized Aerial Robots in Controlled Airspace Using Reinforcement Learning9 Jul 2024 0 repositories listed
-
An open source Multi-Agent Deep Reinforcement Learning Routing Simulator for satellite networks8 Jul 2024 0 repositories listed
-
Graph Anomaly Detection with Noisy Labels by Reinforcement Learning8 Jul 2024 0 repositories listed
-
E²CFD: Towards Effective and Efficient Cost Function Design for Safe Reinforcement Learning via Large Language Model8 Jul 2024 0 repositories listed
-
Multi-agent Reinforcement Learning-based Network Intrusion Detection System8 Jul 2024 0 repositories listed
-
System stabilization with policy optimization on unstable latent manifolds8 Jul 2024 0 repositories listed
-
A Reinforcement Learning Approach for Wildfire Tracking with UAV Swarms7 Jul 2024 0 repositories listed
-
Discounted Pseudocosts in MILP7 Jul 2024 0 repositories listed
-
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments6 Jul 2024 0 repositories listed
-
Augmented Bayesian Policy Search5 Jul 2024 0 repositories listed
-
Graph Reinforcement Learning for Power Grids: A Comprehensive Survey5 Jul 2024 0 repositories listed
-
Question Answering with Texts and Tables through Deep Reinforcement Learning5 Jul 2024 0 repositories listed
-
Unsupervised Video Summarization via Reinforcement Learning and a Trained Evaluator5 Jul 2024 0 repositories listed
-
Deep Pareto Reinforcement Learning for Multi-Objective Recommender Systems4 Jul 2024 0 repositories listed
-
Reinforcement Learning for Sequence Design Leveraging Protein Language Models3 Jul 2024 0 repositories listed
-
Beyond Numeric Awards: In-Context Dueling Bandits with LLM Agents2 Jul 2024 0 repositories listed
-
Reinforcement Learning and Machine ethics:a systematic review2 Jul 2024 0 repositories listed
-
Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?2 Jul 2024 0 repositories listed
-
Research on Autonomous Robots Navigation based on Reinforcement Learning2 Jul 2024 0 repositories listed
-
Text-Aware Diffusion for Policy Learning2 Jul 2024 0 repositories listed
-
A Deep Reinforcement Learning Approach to Battery Management in Dairy Farming via Proximal Policy Optimization1 Jul 2024 0 repositories listed
-
Contractual Reinforcement Learning: Pulling Arms with Invisible Hands1 Jul 2024 0 repositories listed
-
Deep Reinforcement Learning for Adverse Garage Scenario Generation1 Jul 2024 0 repositories listed
-
Normalization and effective learning rates in reinforcement learning1 Jul 2024 0 repositories listed
-
Benchmarks for Reinforcement Learning with Biased Offline Data and Imperfect Simulators30 Jun 2024 0 repositories listed
-
Disentangled Representations for Causal Cognition30 Jun 2024 0 repositories listed
-
Safe Reinforcement Learning for Power System Control: A Review30 Jun 2024 0 repositories listed
-
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning30 Jun 2024 0 repositories listed
-
Beyond Human Preferences: Exploring Reinforcement Learning Trajectory Evaluation and Improvement through LLMs28 Jun 2024 0 repositories listed
-
Mental Modeling of Reinforcement Learning Agents by Language Models26 Jun 2024 0 repositories listed
-
Preference Elicitation for Offline Reinforcement Learning26 Jun 2024 0 repositories listed
-
Privacy Preserving Reinforcement Learning for Population Processes25 Jun 2024 0 repositories listed
-
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis24 Jun 2024 0 repositories listed
-
Diffusion Spectral Representation for Reinforcement Learning23 Jun 2024 0 repositories listed
-
Position: Benchmarking is Limited in Reinforcement Learning Research23 Jun 2024 0 repositories listed
-
Understanding and Diagnosing Deep Reinforcement Learning23 Jun 2024 0 repositories listed
-
Distributionally Robust Constrained Reinforcement Learning under Strong Duality22 Jun 2024 0 repositories listed
-
An Idiosyncrasy of Time-discretization in Reinforcement Learning21 Jun 2024 0 repositories listed
-
KnobTree: Intelligent Database Parameter Configuration via Explainable Reinforcement Learning21 Jun 2024 0 repositories listed
-
Robust Reinforcement Learning from Corrupted Human Feedback21 Jun 2024 0 repositories listed
-
What Teaches Robots to Walk, Teaches Them to Trade too -- Regime Adaptive Execution using Informed Data and LLMs20 Jun 2024 0 repositories listed
-
A General Control-Theoretic Approach for Reinforcement Learning: Theory and Algorithms20 Jun 2024 0 repositories listed
-
Bayesian Inverse Reinforcement Learning for Non-Markovian Rewards20 Jun 2024 0 repositories listed
-
Constrained Meta Agnostic Reinforcement Learning20 Jun 2024 0 repositories listed
-
Equivariant Offline Reinforcement Learning20 Jun 2024 0 repositories listed
-
Tractable Equilibrium Computation in Markov Games through Risk Aversion20 Jun 2024 0 repositories listed
-
Learned Graph Rewriting with Equality Saturation: A New Paradigm in Relational Query Rewrite and Beyond19 Jun 2024 0 repositories listed
-
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation19 Jun 2024 0 repositories listed
-
Trapezoidal Gradient Descent for Effective Reinforcement Learning in Spiking Networks19 Jun 2024 0 repositories listed
-
Quantum Compiling with Reinforcement Learning on a Superconducting Processor18 Jun 2024 0 repositories listed
-
Reinforcement Learning for Corporate Bond Trading: A Sell Side Perspective18 Jun 2024 0 repositories listed
-
Physics-informed Imitative Reinforcement Learning for Real-world Driving18 Jun 2024 0 repositories listed
-
17 Jun 2024 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Constructing Ancestral Recombination Graphs through Reinforcement Learning17 Jun 2024 0 repositories listed
-
Measuring memorization in RLHF for code completion17 Jun 2024 0 repositories listed
-
Optimal Transport-Assisted Risk-Sensitive Q-Learning17 Jun 2024 0 repositories listed
-
The Benefits of Power Regularization in Cooperative Reinforcement Learning17 Jun 2024 0 repositories listed
-
DIPPER: Direct Preference Optimization to Accelerate Primitive-Enabled Hierarchical Reinforcement Learning16 Jun 2024 0 repositories listed
-
InstructRL4Pix: Training Diffusion for Image Editing by Reinforcement Learning14 Jun 2024 0 repositories listed
-
CIMRL: Combining IMitation and Reinforcement Learning for Safe Autonomous Driving13 Jun 2024 0 repositories listed
-
Current applications and potential future directions of reinforcement learning-based Digital Twins in agriculture13 Jun 2024 0 repositories listed
-
DiffPoGAN: Diffusion Policies with Generative Adversarial Networks for Offline Reinforcement Learning13 Jun 2024 0 repositories listed
-
e-COP : Episodic Constrained Optimization of Policies13 Jun 2024 0 repositories listed
-
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning13 Jun 2024 0 repositories listed
-
Deep reinforcement learning with positional context for intraday trading12 Jun 2024 0 repositories listed
-
Explore-Go: Leveraging Exploration for Generalisation in Deep Reinforcement Learning12 Jun 2024 0 repositories listed
-
How social reinforcement learning can lead to metastable polarisation and the voter model12 Jun 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.