Browse State-of-the-Art › reinforcement-learning › Papers, page 71
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 71 of 135: papers 7,001 to 7,100 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Enabling surrogate-assisted evolutionary reinforcement learning via policy embedding31 Jan 2023 0 repositories listed
-
Partitioning Distributed Compute Jobs with Reinforcement Learning and Graph Neural Networks31 Jan 2023 0 repositories listed
-
Towards interpretable quantum machine learning via single-photon quantum walks31 Jan 2023 0 repositories listed
-
Scalable Grid-Aware Dynamic Matching using Deep Reinforcement Learning31 Jan 2023 0 repositories listed
-
Scaling laws for single-agent reinforcement learning31 Jan 2023 0 repositories listed
-
Scheduling Inference Workloads on Distributed Edge Clusters with Reinforcement Learning31 Jan 2023 0 repositories listed
-
Hierarchical Programmatic Reinforcement Learning via Learning to Compose Programs30 Jan 2023 0 repositories listed
-
Improved Regret for Efficient Online Reinforcement Learning with Linear Function Approximation30 Jan 2023 0 repositories listed
-
Regret Bounds for Markov Decision Processes with Recursive Optimized Certainty Equivalents30 Jan 2023 0 repositories listed
-
STEEL: Singularity-aware Reinforcement Learning30 Jan 2023 0 repositories listed
-
Transferring Multiple Policies to Hotstart Reinforcement Learning in an Air Compressor Management Problem30 Jan 2023 0 repositories listed
-
V2N Service Scaling with Deep Reinforcement Learning30 Jan 2023 0 repositories listed
-
A Deep Reinforcement Learning Framework for Optimizing Congestion Control in Data Centers29 Jan 2023 0 repositories listed
-
Autonomous Satellite Docking via Adaptive Optimal Output Rregulation: A Reinforcement Learning Approach29 Jan 2023 0 repositories listed
-
Sample Efficient Deep Reinforcement Learning via Local Planning29 Jan 2023 0 repositories listed
-
SaFormer: A Conditional Sequence Modeling Approach to Offline Safe Reinforcement Learning28 Jan 2023 0 repositories listed
-
STEERING: Stein Information Directed Exploration for Model-Based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Towards Learning Rubik's Cube with N-tuple-based Reinforcement Learning28 Jan 2023 0 repositories listed
-
A Memory Efficient Deep Reinforcement Learning Approach For Snake Game Autonomous Agents27 Jan 2023 0 repositories listed
-
Improving Behavioural Cloning with Positive Unlabeled Learning27 Jan 2023 0 repositories listed
-
Exploring Deep Reinforcement Learning for Holistic Smart Building Control27 Jan 2023 0 repositories listed
-
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence27 Jan 2023 0 repositories listed
-
Reinforcement Learning from Diverse Human Preferences27 Jan 2023 0 repositories listed
-
Single-Trajectory Distributionally Robust Reinforcement Learning27 Jan 2023 0 repositories listed
-
SNeRL: Semantic-aware Neural Radiance Fields for Reinforcement Learning27 Jan 2023 0 repositories listed
-
Solving Richly Constrained Reinforcement Learning through State Augmentation and Reward Penalties27 Jan 2023 0 repositories listed
-
Communication-Efficient Collaborative Regret Minimization in Multi-Armed Bandits26 Jan 2023 0 repositories listed
-
FedHQL: Federated Heterogeneous Q-Learning26 Jan 2023 0 repositories listed
-
Model-based Offline Reinforcement Learning with Local Misspecification26 Jan 2023 0 repositories listed
-
Certifiably Robust Reinforcement Learning through Model-Based Abstract Interpretation26 Jan 2023 0 repositories listed
-
Principled Reinforcement Learning with Human Feedback from Pairwise or K-wise Comparisons26 Jan 2023 0 repositories listed
-
A Novel Deep Reinforcement Learning-based Approach for Enhancing Spectral Efficiency of IRS-assisted Wireless Systems24 Jan 2023 0 repositories listed
-
ASQ-IT: Interactive Explanations for Reinforcement-Learning Agents24 Jan 2023 0 repositories listed
-
AutoCost: Evolving Intrinsic Cost for Zero-violation Reinforcement Learning24 Jan 2023 0 repositories listed
-
Autonomous particles24 Jan 2023 0 repositories listed
-
Explainable Deep Reinforcement Learning: State of the Art and Challenges24 Jan 2023 0 repositories listed
-
Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review24 Jan 2023 0 repositories listed
-
Minimal Value-Equivalent Partial Models for Scalable and Robust Planning in Lifelong Reinforcement Learning24 Jan 2023 0 repositories listed
-
Story Shaping: Teaching Agents Human-like Behavior with Stories24 Jan 2023 0 repositories listed
-
Model Based Reinforcement Learning with Non-Gaussian Environment Dynamics and its Application to Portfolio Optimization23 Jan 2023 0 repositories listed
-
Quasi-optimal Reinforcement Learning with Continuous Actions21 Jan 2023 0 repositories listed
-
Asynchronous Deep Double Duelling Q-Learning for Trading-Signal Execution in Limit Order Book Markets20 Jan 2023 0 repositories listed
-
Resource Optimization for Semantic-Aware Networks with Task Offloading20 Jan 2023 0 repositories listed
-
Generative Slate Recommendation with Reinforcement Learning20 Jan 2023 0 repositories listed
-
Multi-agent Reinforcement Learning with Graph Q-Networks for Antenna Tuning20 Jan 2023 0 repositories listed
-
Multi-Armed Bandits and Quantum Channel Oracles20 Jan 2023 0 repositories listed
-
Reinforcement learning-based estimation for partial differential equations20 Jan 2023 0 repositories listed
-
Revisiting Estimation Bias in Policy Gradients for Deep Reinforcement Learning20 Jan 2023 0 repositories listed
-
A Survey of Meta-Reinforcement Learning19 Jan 2023 0 repositories listed
-
Advanced Scaling Methods for VNF deployment with Reinforcement Learning19 Jan 2023 0 repositories listed
-
Domain-adapted Learning and Interpretability: DRL for Gas Trading19 Jan 2023 0 repositories listed
-
Domain-adapted Learning and Imitation: DRL for Power Arbitrage19 Jan 2023 0 repositories listed
-
Tight Guarantees for Interactive Decision Making with the Decision-Estimation Coefficient19 Jan 2023 0 repositories listed
-
Human-Timescale Adaptation in an Open-Ended Task Space18 Jan 2023 0 repositories listed
-
Multi-compartment Neuron and Population Encoding Powered Spiking Neural Network for Deep Distributional Reinforcement Learning18 Jan 2023 0 repositories listed
-
Adversarial Robust Deep Reinforcement Learning Requires Redefining Robustness17 Jan 2023 0 repositories listed
-
DQNAS: Neural Architecture Search using Reinforcement Learning17 Jan 2023 0 repositories listed
-
Learning to solve arithmetic problems with a virtual abacus17 Jan 2023 0 repositories listed
-
Show me what you want: Inverse reinforcement learning to automatically design robot swarms by demonstration17 Jan 2023 0 repositories listed
-
Neuro-Symbolic World Models for Adapting to Open World Novelty16 Jan 2023 0 repositories listed
-
CogReact: A Reinforced Framework to Model Human Cognitive Reaction Modulated by Dynamic Intervention15 Jan 2023 0 repositories listed
-
Neuro-symbolic Meta Reinforcement Learning for Trading15 Jan 2023 0 repositories listed
-
First Three Years of the International Verification of Neural Networks Competition (VNN-COMP)14 Jan 2023 0 repositories listed
-
PRUDEX-Compass: Towards Systematic Evaluation of Reinforcement Learning in Financial Markets14 Jan 2023 0 repositories listed
-
Reinforcement Learning for Protocol Synthesis in Resource-Constrained Wireless Sensor and IoT Networks14 Jan 2023 0 repositories listed
-
Risk-Averse Reinforcement Learning via Dynamic Time-Consistent Risk Measures14 Jan 2023 0 repositories listed
-
A Constrained-Optimization Approach to the Execution of Prioritized Stacks of Learned Multi-Robot Tasks13 Jan 2023 0 repositories listed
-
Decentralized model-free reinforcement learning in stochastic games with average-reward objective13 Jan 2023 0 repositories listed
-
Hierarchical Deep Q-Learning Based Handover in Wireless Networks with Dual Connectivity13 Jan 2023 0 repositories listed
-
Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning13 Jan 2023 0 repositories listed
-
Asynchronous training of quantum reinforcement learning12 Jan 2023 0 repositories listed
-
Reinforcement Learning-based Joint Handover and Beam Tracking in Millimeter-wave Networks12 Jan 2023 0 repositories listed
-
An Analysis of Quantile Temporal-Difference Learning11 Jan 2023 0 repositories listed
-
Efficient Preference-Based Reinforcement Learning Using Learned Dynamics Models11 Jan 2023 0 repositories listed
-
SoK: Adversarial Machine Learning Attacks and Defences in Multi-Agent Reinforcement Learning11 Jan 2023 0 repositories listed
-
Switchable Lightweight Anti-symmetric Processing (SLAP) with CNN Outspeeds Data Augmentation by Smaller Sample -- Application in Gomoku Reinforcement Learning11 Jan 2023 0 repositories listed
-
Actor-Director-Critic: A Novel Deep Reinforcement Learning Framework10 Jan 2023 0 repositories listed
-
Exploration in Model-based Reinforcement Learning with Randomized Reward9 Jan 2023 0 repositories listed
-
Minimax Weight Learning for Absorbing MDPs9 Jan 2023 0 repositories listed
-
Network Slicing via Transfer Learning aided Distributed Deep Reinforcement Learning9 Jan 2023 0 repositories listed
-
Reinforcement Learning Enhanced PicHunter for Interactive Search9 Jan 2023 0 repositories listed
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning9 Jan 2023 0 repositories listed
-
A Survey on Transformers in Reinforcement Learning8 Jan 2023 0 repositories listed
-
Learning Symbolic Representations for Reinforcement Learning of Non-Markovian Behavior8 Jan 2023 0 repositories listed
-
Hierarchical Reinforcement Learning for RIS-Assisted Energy-Efficient RAN7 Jan 2023 0 repositories listed
-
Mathematical Models and Reinforcement Learning based Evolutionary Algorithm Framework for Satellite Scheduling Problem7 Jan 2023 0 repositories listed
-
Markov Chain Concentration with an Application in Reinforcement Learning7 Jan 2023 0 repositories listed
-
A Deep Reinforcement Learning-Based Controller for Magnetorheological-Damped Vehicle Suspension6 Jan 2023 0 repositories listed
-
Multi-Agent Reinforcement Learning for Fast-Timescale Demand Response of Residential Loads6 Jan 2023 0 repositories listed
-
Provable Reset-free Reinforcement Learning by No-Regret Reduction6 Jan 2023 0 repositories listed
-
Data-Driven Inverse Reinforcement Learning for Expert-Learner Zero-Sum Games5 Jan 2023 0 repositories listed
-
Reinforcement Learning-Based Air Traffic Deconfliction5 Jan 2023 0 repositories listed
-
Scalable Communication for Multi-Agent Reinforcement Learning via Transformer-Based Email Mechanism5 Jan 2023 0 repositories listed
-
Value Enhancement of Reinforcement Learning via Efficient and Robust Trust Region Optimization5 Jan 2023 0 repositories listed
-
Learning-based MPC from Big Data Using Reinforcement Learning4 Jan 2023 0 repositories listed
-
UAV aided Metaverse over Wireless Communications: A Reinforcement Learning Approach4 Jan 2023 0 repositories listed
-
Contextual Conservative Q-Learning for Offline Reinforcement Learning3 Jan 2023 0 repositories listed
-
Faster Reinforcement Learning by Freezing Slow States3 Jan 2023 0 repositories listed
-
Offline Evaluation for Reinforcement Learning-based Recommendation: A Critical Issue and Some Alternatives3 Jan 2023 0 repositories listed
-
Safe Reinforcement Learning for an Energy-Efficient Driver Assistance System3 Jan 2023 0 repositories listed