Browse State-of-the-Art › Deep Reinforcement Learning › Papers, page 18
Deep Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 5,822 · with a code link: 1,739 · where Syntology ran a sample: 398 (340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (398 of 5,822 tagged: 340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument)
Page 18 of 59: papers 1,701 to 1,800 of 5,822, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
26 Jul 2017 1 repository listed
-
14 Jul 2017 1 repository listed
-
11 Jul 2017 1 repository listed
-
7 Jul 2017 1 repository listed
-
4 Jul 2017 1 repository listed
-
1 Jul 2017 1 repository listed
-
19 Jun 2017 1 repository listed
-
15 Jun 2017 1 repository listed
-
26 Apr 2017 1 repository listed
-
21 Apr 2017 1 repository listed
-
18 Apr 2017 1 repository listed
-
8 Apr 2017 1 repository listed
-
31 Mar 2017 1 repository listed
-
3 Mar 2017 1 repository listed
-
27 Feb 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
19 Feb 2017 1 repository listed
-
16 Jan 2017 1 repository listed
-
9 Jan 2017 1 repository listed
-
A Survey of Deep Network Solutions for Learning Control in Robotics: From Reinforcement to Imitation21 Dec 2016 1 repository listed
-
1 Dec 2016 1 repository listed
-
1 Dec 2016 1 repository listed
-
26 Nov 2016 1 repository listed
-
26 Nov 2016 1 repository listed
-
13 Nov 2016 1 repository listed
-
11 Nov 2016 1 repository listed
-
5 Nov 2016 1 repository listed
-
27 Sep 2016 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
18 Sep 2016 1 repository listed
-
18 Jul 2016 1 repository listed
-
29 Jun 2016 1 repository listed
-
12 Jun 2016 1 repository listed
-
8 Jun 2016 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
8 Jun 2016 1 repository listed
-
30 May 2016 1 repository listed
-
18 Jan 2016 1 repository listed
-
25 Nov 2015 1 repository listed
-
19 Nov 2015 1 repository listed
-
LiLM-RDB-SFC: Lightweight Language Model with Relational Database-Guided DRL for Optimized SFC Provisioning15 Jul 2025 0 repositories listed
-
Sensing Accuracy Optimization for Multi-UAV SAR Interferometry with Data Offloading15 Jul 2025 0 repositories listed
-
Turning Sand to Gold: Recycling Data to Bridge On-Policy and Off-Policy Learning via Causal Bound15 Jul 2025 0 repositories listed
-
Meta-Reinforcement Learning for Fast and Data-Efficient Spectrum Allocation in Dynamic Wireless Networks13 Jul 2025 0 repositories listed
-
Hierarchical Task Offloading for UAV-Assisted Vehicular Edge Computing via Deep Reinforcement Learning8 Jul 2025 0 repositories listed
-
Beyond Training-time Poisoning: Component-level and Post-training Backdoors in Deep Reinforcement Learning7 Jul 2025 0 repositories listed
-
Explainable AI for Radar Resource Management: Modified LIME in Deep Reinforcement Learning26 Jun 2025 0 repositories listed
-
rQdia: Regularizing Q-Value Distributions With Image Augmentation26 Jun 2025 0 repositories listed
-
GymPN: A Library for Decision-Making in Process Management Systems25 Jun 2025 0 repositories listed
-
Learning-Based Resource Management in Integrated Sensing and Communication Systems25 Jun 2025 0 repositories listed
-
Multi-Objective Reinforcement Learning for Cognitive Radar Resource Management25 Jun 2025 0 repositories listed
-
Efficient Beam Selection for ISAC in Cell-Free Massive MIMO via Digital Twin-Assisted Deep Reinforcement Learning23 Jun 2025 0 repositories listed
-
Optimal Design of Experiment for Electrochemical Parameter Identification of Li-ion Battery via Deep Reinforcement Learning23 Jun 2025 0 repositories listed
-
Adaptive Social Metaverse Streaming based on Federated Multi-Agent Deep Reinforcement Learning19 Jun 2025 0 repositories listed
-
BIDA: A Bi-level Interaction Decision-making Algorithm for Autonomous Vehicles in Dynamic Traffic Scenarios19 Jun 2025 0 repositories listed
-
A Novel ViDAR Device With Visual Inertial Encoder Odometry and Reinforcement Learning-Based Active SLAM Method16 Jun 2025 0 repositories listed
-
Joint Spectrum Sensing and Resource Allocation for OFDMA-based Underwater Acoustic Communications16 Jun 2025 0 repositories listed
-
16 Jun 2025 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Federated Neuroevolution O-RAN: Enhancing the Robustness of Deep Reinforcement Learning xApps15 Jun 2025 0 repositories listed
-
Automated Treatment Planning for Interstitial HDR Brachytherapy for Locally Advanced Cervical Cancer using Deep Reinforcement Learning13 Jun 2025 0 repositories listed
-
Joint Beamforming with Extremely Large Scale RIS: A Sequential Multi-Agent A2C Approach12 Jun 2025 0 repositories listed
-
TooBadRL: Trigger Optimization to Boost Effectiveness of Backdoor Attacks on Deep Reinforcement Learning11 Jun 2025 0 repositories listed
-
Patient-Specific Deep Reinforcement Learning for Automatic Replanning in Head-and-Neck Cancer Proton Therapy11 Jun 2025 0 repositories listed
-
Foundation Model-Aided Deep Reinforcement Learning for RIS-Assisted Wireless Communication11 Jun 2025 0 repositories listed
-
MOORL: A Framework for Integrating Offline-Online Reinforcement Learning11 Jun 2025 0 repositories listed
-
Synergizing Reinforcement Learning and Genetic Algorithms for Neural Combinatorial Optimization11 Jun 2025 0 repositories listed
-
Modular Recurrence in Contextual MDPs for Universal Morphology Control10 Jun 2025 0 repositories listed
-
Preference-Driven Multi-Objective Combinatorial Optimization with Conditional Computation10 Jun 2025 0 repositories listed
-
Safe and Economical UAV Trajectory Planning in Low-Altitude Airspace: A Hybrid DRL-LLM Approach with Compliance Awareness10 Jun 2025 0 repositories listed
-
Towards Robust Deep Reinforcement Learning against Environmental State Perturbation10 Jun 2025 0 repositories listed
-
Interpreting Agent Behaviors in Reinforcement-Learning-Based Cyber-Battle Simulation Platforms9 Jun 2025 0 repositories listed
-
An Intelligent Fault Self-Healing Mechanism for Cloud AI Systems via Integration of Large Language Models and Deep Reinforcement Learning9 Jun 2025 0 repositories listed
-
Deep reinforcement learning for near-deterministic preparation of cubic- and quartic-phase gates in photonic quantum computing9 Jun 2025 0 repositories listed
-
Deep reinforcement learning-based joint real-time energy scheduling for green buildings with heterogeneous battery energy storage devices7 Jun 2025 0 repositories listed
-
Energy-efficient Deep Reinforcement Learning-based Network Function Disaggregation in Hybrid Non-terrestrial Open Radio Access Networks7 Jun 2025 0 repositories listed
-
Improving choice model specification using reinforcement learning6 Jun 2025 0 repositories listed
-
The Economic Dispatch of Power-to-Gas Systems with Deep Reinforcement Learning:Tackling the Challenge of Delayed Rewards with Long-Term Energy Storage6 Jun 2025 0 repositories listed
-
Can Artificial Intelligence Trade the Stock Market?5 Jun 2025 0 repositories listed
-
An Efficient Task-Oriented Dialogue Policy: Evolutionary Reinforcement Learning Injected by Elite Individuals4 Jun 2025 0 repositories listed
-
Autonomous Vehicle Lateral Control Using Deep Reinforcement Learning with MPC-PID Demonstration4 Jun 2025 0 repositories listed
-
Beamforming and Resource Allocation for Delay Optimization in RIS-Assisted OFDM Systems4 Jun 2025 0 repositories listed
-
A Novel Deep Reinforcement Learning Method for Computation Offloading in Multi-User Mobile Edge Computing with Decentralization3 Jun 2025 0 repositories listed
-
EDEN: Entorhinal Driven Egocentric Navigation Toward Robotic Deployment3 Jun 2025 0 repositories listed
-
Maximizing the Promptness of Metaverse Systems using Edge Computing by Deep Reinforcement Learning3 Jun 2025 0 repositories listed
-
Solving the Pod Repositioning Problem with Deep Reinforced Adaptive Large Neighborhood Search3 Jun 2025 0 repositories listed
-
Interpretable reinforcement learning for heat pump control through asymmetric differentiable decision trees2 Jun 2025 0 repositories listed
-
BASIL: Best-Action Symbolic Interpretable Learning for Evolving Compact RL Policies31 May 2025 0 repositories listed
-
Reinforcement Learning for Hanabi31 May 2025 0 repositories listed
-
AXIOM: Learning to Play Games in Minutes with Expanding Object-Centric Models30 May 2025 0 repositories listed
-
Human sensory-musculoskeletal modeling and control of whole-body movements29 May 2025 0 repositories listed
-
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning29 May 2025 0 repositories listed
-
Attention-Enhanced Prompt Decision Transformers for UAV-Assisted Communications with AoI28 May 2025 0 repositories listed
-
A Framework for Adversarial Analysis of Decision Support Systems Prior to Deployment27 May 2025 0 repositories listed
-
Learning optimal treatment strategies for intraoperative hypotension using deep reinforcement learning27 May 2025 0 repositories listed
-
Algorithmic Control Improves Residential Building Energy and EV Management when PV Capacity is High but Battery Capacity is Low26 May 2025 0 repositories listed
-
Reduce Computational Cost In Deep Reinforcement Learning Via Randomized Policy Learning25 May 2025 0 repositories listed
-
A modular framework for automated evaluation of procedural content generation in serious games with deep reinforcement learning agents22 May 2025 0 repositories listed
-
Backdoors in DRL: Four Environments Focusing on In-distribution Triggers22 May 2025 0 repositories listed
-
Find the Fruit: Designing a Zero-Shot Sim2Real Deep RL Planner for Occlusion Aware Plant Manipulation22 May 2025 0 repositories listed
-
Energy-Efficient Deep Reinforcement Learning with Spiking Transformers20 May 2025 0 repositories listed
-
Imitation Learning via Focused Satisficing20 May 2025 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.