Browse State-of-the-Art › Deep Reinforcement Learning › Papers, page 25
Deep Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 5,822 · with a code link: 1,739 · where Syntology ran a sample: 398 (340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (398 of 5,822 tagged: 340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument)
Page 25 of 59: papers 2,401 to 2,500 of 5,822, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Sim-to-Real Transfer of Deep Reinforcement Learning Agents for Online Coverage Path Planning7 Jun 2024 0 repositories listed
-
Exploring Pessimism and Optimism Dynamics in Deep Reinforcement Learning6 Jun 2024 0 repositories listed
-
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model6 Jun 2024 0 repositories listed
-
Stochastic Dynamic Network Utility Maximization with Application to Disaster Response6 Jun 2024 0 repositories listed
-
A Generalized Apprenticeship Learning Framework for Modeling Heterogeneous Student Pedagogical Strategies4 Jun 2024 0 repositories listed
-
Algorithmic Collusion in Dynamic Pricing with Deep Reinforcement Learning4 Jun 2024 0 repositories listed
-
By Fair Means or Foul: Quantifying Collusion in a Market Simulation with Deep Reinforcement Learning4 Jun 2024 0 repositories listed
-
Improving Generalization in Aerial and Terrestrial Mobile Robots Control Through Delayed Policy Learning4 Jun 2024 0 repositories listed
-
Verifying the Generalization of Deep Learning to Out-of-Distribution Domains4 Jun 2024 0 repositories listed
-
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment3 Jun 2024 0 repositories listed
-
An Advanced Reinforcement Learning Framework for Online Scheduling of Deferrable Workloads in Cloud Computing3 Jun 2024 0 repositories listed
-
Deep Reinforcement Learning Behavioral Mode Switching Using Optimal Control Based on a Latent Space Objective3 Jun 2024 0 repositories listed
-
Deep reinforcement learning for weakly coupled MDP's with continuous actions3 Jun 2024 0 repositories listed
-
Large Language Model Assisted Optimal Bidding of BESS in FCAS Market: An AI-agent based Approach3 Jun 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning Meets Leaf Sequencing in Radiotherapy3 Jun 2024 0 repositories listed
-
A Digital Twin Framework for Reinforcement Learning with Real-Time Self-Improvement via Human Assistive Teleoperation2 Jun 2024 0 repositories listed
-
Research on the Application of Computer Vision Based on Deep Learning in Autonomous Driving Technology1 Jun 2024 0 repositories listed
-
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems1 Jun 2024 0 repositories listed
-
Generative AI for Deep Reinforcement Learning: Framework, Analysis, and Use Cases31 May 2024 0 repositories listed
-
Goal-Oriented Sensor Reporting Scheduling for Non-linear Dynamic System Monitoring31 May 2024 0 repositories listed
-
A Deep Reinforcement Learning Approach for Trading Optimization in the Forex Market with Multi-Agent Asynchronous Distribution30 May 2024 0 repositories listed
-
Enhancing Battlefield Awareness: An Aerial RIS-assisted ISAC System with Deep Reinforcement Learning30 May 2024 0 repositories listed
-
Q-learning as a monotone scheme30 May 2024 0 repositories listed
-
Advancing Household Robotics: Deep Interactive Reinforcement Learning for Efficient Training and Enhanced Performance29 May 2024 0 repositories listed
-
Exploring the impact of traffic signal control and connected and automated vehicles on intersections safety: A deep reinforcement learning approach29 May 2024 0 repositories listed
-
Proactive Load-Shaping Strategies with Privacy-Cost Trade-offs in Residential Households based on Deep Reinforcement Learning29 May 2024 0 repositories listed
-
Mollification Effects of Policy Gradient Methods28 May 2024 0 repositories listed
-
World Models for General Surgical Grasping28 May 2024 0 repositories listed
-
Biological Neurons Compete with Deep Reinforcement Learning in Sample Efficiency in a Simulated Gameworld27 May 2024 0 repositories listed
-
Amortized Active Causal Induction with Deep Reinforcement Learning26 May 2024 0 repositories listed
-
Counterexample-Guided Repair of Reinforcement Learning Systems Using Safety Critics24 May 2024 0 repositories listed
-
SF-DQN: Provable Knowledge Transfer using Successor Feature for Deep Reinforcement Learning24 May 2024 0 repositories listed
-
Transmission Interface Power Flow Adjustment: A Deep Reinforcement Learning Approach based on Multi-task Attribution Map24 May 2024 0 repositories listed
-
A Behavior-Aware Approach for Deep Reinforcement Learning in Non-stationary Environments without Known Change Points23 May 2024 0 repositories listed
-
Closed-form Symbolic Solutions: A New Perspective on Solving Partial Differential Equations23 May 2024 0 repositories listed
-
Deep Reinforcement Learning for 5*5 Multiplayer Go23 May 2024 0 repositories listed
-
Doubly-Dynamic ISAC Precoding for Vehicular Networks: A Constrained Deep Reinforcement Learning (CDRL) Approach23 May 2024 0 repositories listed
-
Deep Reinforcement Learning for Time-Critical Wilderness Search And Rescue Using Drones21 May 2024 0 repositories listed
-
GASE: Graph Attention Sampling with Edges Fusion for Solving Vehicle Routing Problems21 May 2024 0 repositories listed
-
Near-Field Spot Beamfocusing: A Correlation-Aware Transfer Learning Approach21 May 2024 0 repositories listed
-
Rethinking Robustness Assessment: Adversarial Attacks on Learning-based Quadrupedal Locomotion Controllers21 May 2024 0 repositories listed
-
Continual Deep Reinforcement Learning for Decentralized Satellite Routing20 May 2024 0 repositories listed
-
Investigating the Impact of Choice on Deep Reinforcement Learning for Space Controls20 May 2024 0 repositories listed
-
Deep Dive into Model-free Reinforcement Learning for Biological and Robotic Systems: Theory and Practice19 May 2024 0 repositories listed
-
Enhancing Vehicle Aerodynamics with Deep Reinforcement Learning in Voxelised Models19 May 2024 0 repositories listed
-
Exploiting Distributional Value Functions for Financial Market Valuation, Enhanced Feature Creation and Improvement of Trading Algorithms19 May 2024 0 repositories listed
-
Federated Learning With Energy Harvesting Devices: An MDP Framework17 May 2024 0 repositories listed
-
Continuous Transfer Learning for UAV Communication-aware Trajectory Design16 May 2024 0 repositories listed
-
Chaos-based reinforcement learning with TD315 May 2024 0 repositories listed
-
Detecting Continuous Integration Skip : A Reinforcement Learning-based Approach15 May 2024 0 repositories listed
-
DVS-RG: Differential Variable Speed Limits Control using Deep Reinforcement Learning with Graph State Representation15 May 2024 0 repositories listed
-
CIER: A Novel Experience Replay Approach with Causal Inference in Deep Reinforcement Learning14 May 2024 0 repositories listed
-
Deep Reinforcement Learning for Real-Time Ground Delay Program Revision and Corresponding Flight Delay Assignments14 May 2024 0 repositories listed
-
Optimizing Deep Reinforcement Learning for American Put Option Hedging14 May 2024 0 repositories listed
-
MADRL-Based Rate Adaptation for 360° Video Streaming with Multi-Viewpoint Prediction13 May 2024 0 repositories listed
-
On-Demand Model and Client Deployment in Federated Learning with Deep Reinforcement Learning12 May 2024 0 repositories listed
-
Auditing an Automatic Grading Model with deep Reinforcement Learning11 May 2024 0 repositories listed
-
Stealthy Imitation: Reward-guided Environment-free Policy Stealing11 May 2024 0 repositories listed
-
Hedging American Put Options with Deep Reinforcement Learning10 May 2024 0 repositories listed
-
Deep Reinforcement Learning for Multi-User RF Charging with Non-linear Energy Harvesters7 May 2024 0 repositories listed
-
Latency and Energy Minimization in NOMA-Assisted MEC Network: A Federated Deep Reinforcement Learning Approach7 May 2024 0 repositories listed
-
End-to-End Reinforcement Learning of Curative Curtailment with Partial Measurement Availability6 May 2024 0 repositories listed
-
Enhancing O-RAN Security: Evasion Attacks and Robust Defenses for Graph Reinforcement Learning-based Connection Management6 May 2024 0 repositories listed
-
Guidance Design for Escape Flight Vehicle Using Evolution Strategy Enhanced Deep Reinforcement Learning4 May 2024 0 repositories listed
-
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning4 May 2024 0 repositories listed
-
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning3 May 2024 0 repositories listed
-
Behavior Imitation for Manipulator Control and Grasping with Deep Reinforcement Learning2 May 2024 0 repositories listed
-
Generative Active Learning for the Search of Small-molecule Protein Binders2 May 2024 0 repositories listed
-
Tabular and Deep Reinforcement Learning for Gittins Index2 May 2024 0 repositories listed
-
HUGO -- Highlighting Unseen Grid Options: Combining Deep Reinforcement Learning with a Heuristic Target Topology Approach1 May 2024 0 repositories listed
-
Portfolio Management using Deep Reinforcement Learning1 May 2024 0 repositories listed
-
Deep Reinforcement Learning for Advanced Longitudinal Control and Collision Avoidance in High-Risk Driving Scenarios29 Apr 2024 0 repositories listed
-
Shared learning of powertrain control policies for vehicle fleets27 Apr 2024 0 repositories listed
-
An Explainable Deep Reinforcement Learning Model for Warfarin Maintenance Dosing Using Policy Distillation and Action Forging26 Apr 2024 0 repositories listed
-
DRL2FC: An Attack-Resilient Controller for Automatic Generation Control Based on Deep Reinforcement Learning25 Apr 2024 0 repositories listed
-
Exploring the Dynamics of Data Transmission in 5G Networks: A Conceptual Analysis25 Apr 2024 0 repositories listed
-
Enhancing High-Speed Cruising Performance of Autonomous Vehicles through Integrated Deep Reinforcement Learning Framework23 Apr 2024 0 repositories listed
-
Multi-Objective Deep Reinforcement Learning for 5G Base Station Placement to Support Localisation for Future Sustainable Traffic23 Apr 2024 0 repositories listed
-
Using deep reinforcement learning to promote sustainable human behaviour on a common pool resource problem23 Apr 2024 0 repositories listed
-
Decentralized Coordination of Distributed Energy Resources through Local Energy Markets and Deep Reinforcement Learning19 Apr 2024 0 repositories listed
-
Deep Reinforcement Learning-aided Transmission Design for Energy-efficient Link Optimization in Vehicular Communications19 Apr 2024 0 repositories listed
-
Random Network Distillation Based Deep Reinforcement Learning for AGV Path Planning19 Apr 2024 0 repositories listed
-
LTL-Constrained Policy Optimization with Cycle Experience Replay17 Apr 2024 0 repositories listed
-
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning15 Apr 2024 0 repositories listed
-
Advanced Intelligent Optimization Algorithms for Multi-Objective Optimal Power Flow in Future Power Systems: A Review14 Apr 2024 0 repositories listed
-
CodeCloak: A Method for Evaluating and Mitigating Code Leakage by LLM Code Assistants13 Apr 2024 0 repositories listed
-
Deep Reinforcement Learning based Online Scheduling Policy for Deep Neural Network Multi-Tenant Multi-Accelerator Systems13 Apr 2024 0 repositories listed
-
Advancing Forest Fire Prevention: Deep Reinforcement Learning for Effective Firebreak Placement12 Apr 2024 0 repositories listed
-
Anti-Byzantine Attacks Enabled Vehicle Selection for Asynchronous Federated Learning in Vehicular Edge Computing12 Apr 2024 0 repositories listed
-
Auto-configuring Exploration-Exploitation Tradeoff in Evolutionary Computation via Deep Reinforcement Learning12 Apr 2024 0 repositories listed
-
Kinematics Modeling of Peroxy Free Radicals: A Deep Reinforcement Learning Approach12 Apr 2024 0 repositories listed
-
Prescribing Optimal Health-Aware Operation for Urban Air Mobility with Deep Reinforcement Learning12 Apr 2024 0 repositories listed
-
RLEMMO: Evolutionary Multimodal Optimization Assisted By Deep Reinforcement Learning12 Apr 2024 0 repositories listed
-
TDANet: Target-Directed Attention Network For Object-Goal Visual Navigation With Zero-Shot Ability12 Apr 2024 0 repositories listed
-
Collaborative Ground-Space Communications via Evolutionary Multi-objective Deep Reinforcement Learning11 Apr 2024 0 repositories listed
-
FPGA Divide-and-Conquer Placement using Deep Reinforcement Learning11 Apr 2024 0 repositories listed
-
Generative Probabilistic Planning for Optimizing Supply Chain Networks11 Apr 2024 0 repositories listed
-
R2 Indicator and Deep Reinforcement Learning Enhanced Adaptive Multi-Objective Evolutionary Algorithm11 Apr 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.