Browse State-of-the-Art › Deep Reinforcement Learning › Papers, page 27
Deep Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 5,822 · with a code link: 1,739 · where Syntology ran a sample: 398 (340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (398 of 5,822 tagged: 340 with a run with no instrument failure, 58 where every run was a failure of Syntology's instrument)
Page 27 of 59: papers 2,601 to 2,700 of 5,822, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
In value-based deep reinforcement learning, a pruned network is a good network19 Feb 2024 0 repositories listed
-
Joint mode switching and resource allocation in wireless-powered RIS-aided multiuser communication systems19 Feb 2024 0 repositories listed
-
Optimal Parallelization Strategies for Active Flow Control in Deep Reinforcement Learning-Based Computational Fluid Dynamics18 Feb 2024 0 repositories listed
-
Deep Reinforcement Learning Based Toolpath Generation for Thermal Uniformity in Laser Powder Bed Fusion Process17 Feb 2024 0 repositories listed
-
SINR-Aware Deep Reinforcement Learning for Distributed Dynamic Channel Allocation in Cognitive Interference Networks17 Feb 2024 0 repositories listed
-
Surpassing legacy approaches to PWR core reload optimization with single-objective Reinforcement learning16 Feb 2024 0 repositories listed
-
Aligning Crowd Feedback via Distributional Preference Reward Modeling15 Feb 2024 0 repositories listed
-
Enhancing Courier Scheduling in Crowdsourced Last-Mile Delivery through Dynamic Shift Extensions: A Deep Reinforcement Learning Approach15 Feb 2024 0 repositories listed
-
Revisiting Experience Replayable Conditions15 Feb 2024 0 repositories listed
-
DRL-Based Orchestration of Multi-User MISO Systems with Stacked Intelligent Metasurfaces14 Feb 2024 0 repositories listed
-
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks14 Feb 2024 0 repositories listed
-
Learning-enabled Flexible Job-shop Scheduling for Scalable Smart Manufacturing14 Feb 2024 0 repositories listed
-
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning14 Feb 2024 0 repositories listed
-
LL-GABR: Energy Efficient Live Video Streaming Using Reinforcement Learning14 Feb 2024 0 repositories listed
-
Single-Reset Divide & Conquer Imitation Learning14 Feb 2024 0 repositories listed
-
Uncertainty-Aware Transient Stability-Constrained Preventive Redispatch: A Distributional Reinforcement Learning Approach14 Feb 2024 0 repositories listed
-
Towards Fair and Firm Real-Time Scheduling in DNN Multi-Tenant Multi-Accelerator Systems via Reinforcement Learning9 Feb 2024 0 repositories listed
-
Attention-Enhanced Prioritized Proximal Policy Optimization for Adaptive Edge Caching8 Feb 2024 0 repositories listed
-
Intelligent Mode-switching Framework for Teleoperation8 Feb 2024 0 repositories listed
-
Real-World Fluid Directed Rigid Body Control via Deep Reinforcement Learning8 Feb 2024 0 repositories listed
-
Scaling Artificial Intelligence for Digital Wargaming in Support of Decision-Making8 Feb 2024 0 repositories listed
-
Scaling Intelligent Agents in Combat Simulations for Wargaming8 Feb 2024 0 repositories listed
-
A computational approach to visual ecology with deep reinforcement learning7 Feb 2024 0 repositories listed
-
A Deep Reinforcement Learning Approach for Adaptive Traffic Routing in Next-gen Networks7 Feb 2024 0 repositories listed
-
Analyzing Adversarial Inputs in Deep Reinforcement Learning7 Feb 2024 0 repositories listed
-
Collaborative Computing in Non-Terrestrial Networks: A Multi-Time-Scale Deep Reinforcement Learning Approach7 Feb 2024 0 repositories listed
-
Compressing Deep Reinforcement Learning Networks with a Dynamic Structured Pruning Method for Autonomous Driving7 Feb 2024 0 repositories listed
-
Collaborative Deep Reinforcement Learning for Resource Optimization in Non-Terrestrial Networks6 Feb 2024 0 repositories listed
-
Deep Reinforcement Learning for Picker Routing Problem in Warehousing5 Feb 2024 0 repositories listed
-
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences5 Feb 2024 0 repositories listed
-
DRED: Zero-Shot Transfer in Reinforcement Learning via Data-Regularised Environment Design5 Feb 2024 0 repositories listed
-
A Safe Reinforcement Learning driven Weights-varying Model Predictive Control for Autonomous Vehicle Motion Control4 Feb 2024 0 repositories listed
-
Device Scheduling and Assignment in Hierarchical Federated Learning for Internet of Things4 Feb 2024 0 repositories listed
-
Evading Deep Learning-Based Malware Detectors via Obfuscation: A Deep Reinforcement Learning Approach4 Feb 2024 0 repositories listed
-
Interference-Aware Emergent Random Access Protocol for Downlink LEO Satellite Networks4 Feb 2024 0 repositories listed
-
Integrating DeepRL with Robust Low-Level Control in Robotic Manipulators for Non-Repetitive Reaching Tasks4 Feb 2024 0 repositories listed
-
DRL-Based Dynamic Channel Access and SCLAR Maximization for Networks Under Jamming2 Feb 2024 0 repositories listed
-
Learning the Market: Sentiment-Based Ensemble Trading Agents2 Feb 2024 0 repositories listed
-
Parametric-Task MAP-Elites2 Feb 2024 0 repositories listed
-
AlphaRank: An Artificial Intelligence Approach for Ranking and Selection Problems1 Feb 2024 0 repositories listed
-
Neural Policy Style Transfer1 Feb 2024 0 repositories listed
-
A Reinforcement Learning Based Controller to Minimize Forces on the Crutches of a Lower-Limb Exoskeleton31 Jan 2024 0 repositories listed
-
Attention Graph for Multi-Robot Social Navigation with Deep Reinforcement Learning31 Jan 2024 0 repositories listed
-
Circuit Partitioning for Multi-Core Quantum Architectures with Deep Reinforcement Learning31 Jan 2024 0 repositories listed
-
Reinforcement Learning for Versatile, Dynamic, and Robust Bipedal Locomotion Control30 Jan 2024 0 repositories listed
-
A Deep Q-Network Based on Radial Basis Functions for Multi-Echelon Inventory Management29 Jan 2024 0 repositories listed
-
Attentive Convolutional Deep Reinforcement Learning for Optimizing Solar-Storage Systems in Real-Time Electricity Markets29 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning for Voltage Control and Renewable Accommodation Using Spatial-Temporal Graph Information29 Jan 2024 0 repositories listed
-
Autonomous Vehicle Patrolling Through Deep Reinforcement Learning: Learning to Communicate and Cooperate28 Jan 2024 0 repositories listed
-
Tacit algorithmic collusion in deep reinforcement learning guided price competition: A study using EV charge pricing game25 Jan 2024 0 repositories listed
-
Traffic Learning and Proactive UAV Trajectory Planning for Data Uplink in Markovian IoT Models24 Jan 2024 0 repositories listed
-
Deep Learning Based Simulators for the Phosphorus Removal Process Control in Wastewater Treatment via Deep Reinforcement Learning Algorithms23 Jan 2024 0 repositories listed
-
Introducing PetriRL: An Innovative Framework for JSSP Resolution Integrating Petri nets and Event-based Reinforcement Learning23 Jan 2024 0 repositories listed
-
Knowledge Distillation from Language-Oriented to Emergent Communication for Multi-Agent Remote Control23 Jan 2024 0 repositories listed
-
Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure management23 Jan 2024 0 repositories listed
-
Viewport Prediction, Bitrate Selection, and Beamforming Design for THz-Enabled 360° Video Streaming23 Jan 2024 0 repositories listed
-
Fast and Scalable Network Slicing by Integrating Deep Learning with Lagrangian Methods22 Jan 2024 0 repositories listed
-
Multi-Agent Dynamic Relational Reasoning for Social Robot Navigation22 Jan 2024 0 repositories listed
-
Efficient and Generalized end-to-end Autonomous Driving System with Latent Deep Reinforcement Learning and Demonstrations22 Jan 2024 0 repositories listed
-
Constrained Reinforcement Learning for Adaptive Controller Synchronization in Distributed SDN21 Jan 2024 0 repositories listed
-
MADRL-based UAVs Trajectory Design with Anti-Collision Mechanism in Vehicular Networks21 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning Empowered Activity-Aware Dynamic Health Monitoring Systems19 Jan 2024 0 repositories listed
-
Episodic Reinforcement Learning with Expanded State-reward Space19 Jan 2024 0 repositories listed
-
Tiny Multi-Agent DRL for Twins Migration in UAV Metaverses: A Multi-Leader Multi-Follower Stackelberg Game Approach18 Jan 2024 0 repositories listed
-
Traffic Smoothing Controllers for Autonomous Vehicles Using Deep Reinforcement Learning and Real-World Trajectory Data18 Jan 2024 0 repositories listed
-
CNN-DRL with Shuffled Features in Finance16 Jan 2024 0 repositories listed
-
CycLight: learning traffic signal cooperation with a cycle-level strategy16 Jan 2024 0 repositories listed
-
Sum Throughput Maximization in Multi-BD Symbiotic Radio NOMA Network Assisted by Active-STAR-RIS16 Jan 2024 0 repositories listed
-
Constrained Multi-objective Optimization with Deep Reinforcement Learning Assisted Operator Selection15 Jan 2024 0 repositories listed
-
Learned Best-Effort LLM Serving15 Jan 2024 0 repositories listed
-
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions14 Jan 2024 0 repositories listed
-
AI-enabled Priority and Auction-Based Spectrum Management for 6G12 Jan 2024 0 repositories listed
-
Open RAN LSTM Traffic Prediction and Slice Management using Deep Reinforcement Learning12 Jan 2024 0 repositories listed
-
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem11 Jan 2024 0 repositories listed
-
Modelling, Positioning, and Deep Reinforcement Learning Path Tracking Control of Scaled Robotic Vehicles: Design and Experimental Validation10 Jan 2024 0 repositories listed
-
ReACT: Reinforcement Learning for Controller Parametrization using B-Spline Geometries10 Jan 2024 0 repositories listed
-
Towards Safe Load Balancing based on Control Barrier Functions and Deep Reinforcement Learning10 Jan 2024 0 repositories listed
-
Deep Reinforcement Multi-agent Learning framework for Information Gathering with Local Gaussian Processes for Water Monitoring9 Jan 2024 0 repositories listed
-
Fully Spiking Actor Network with Intra-layer Connections for Reinforcement Learning9 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning for Multi-Truck Vehicle Routing Problems with Multi-Leg Demand Routes8 Jan 2024 0 repositories listed
-
Guiding drones by information gain8 Jan 2024 0 repositories listed
-
Learn Once Plan Arbitrarily (LOPA): Attention-Enhanced Deep Reinforcement Learning Method for Global Path Planning8 Jan 2024 0 repositories listed
-
Semi-supervised learning via DQN for log anomaly detection6 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning for Local Path Following of an Autonomous Formula SAE Vehicle5 Jan 2024 0 repositories listed
-
A Survey Analyzing Generalization in Deep Reinforcement Learning4 Jan 2024 0 repositories listed
-
OFDM-Based Digital Semantic Communication with Importance Awareness4 Jan 2024 0 repositories listed
-
Trajectory-Oriented Policy Optimization with Sparse Rewards4 Jan 2024 0 repositories listed
-
Learning-based agricultural management in partially observable environments subject to climate variability2 Jan 2024 0 repositories listed
-
Reinforcement Learning for SAR View Angle Inversion with Differentiable SAR Renderer2 Jan 2024 0 repositories listed
-
Contrastive learning-based agent modeling for deep reinforcement learning30 Dec 2023 0 repositories listed
-
Policy Optimization with Smooth Guidance Learned from State-Only Demonstrations30 Dec 2023 0 repositories listed
-
HiBid: A Cross-Channel Constrained Bidding System with Budget Allocation by Hierarchical Offline Deep Reinforcement Learning29 Dec 2023 0 repositories listed
-
RL-LOGO: Deep Reinforcement Learning Localization for Logo Recognition28 Dec 2023 0 repositories listed
-
Risk-anticipatory autonomous driving strategies considering vehicles' weights, based on hierarchical deep reinforcement learning27 Dec 2023 0 repositories listed
-
Visual Spatial Attention and Proprioceptive Data-Driven Reinforcement Learning for Robust Peg-in-Hole Task Under Variable Conditions27 Dec 2023 0 repositories listed
-
A Bayesian Framework of Deep Reinforcement Learning for Joint O-RAN/MEC Orchestration26 Dec 2023 0 repositories listed
-
Adaptive Kalman-based hybrid car following strategy using TD3 and CACC26 Dec 2023 0 repositories listed
-
A Target Detection Algorithm in Traffic Scenes Based on Deep Reinforcement Learning25 Dec 2023 0 repositories listed
-
Deep Reinforcement Learning for Quantitative Trading25 Dec 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.