Browse State-of-the-Art › reinforcement-learning › Papers, page 79
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 79 of 135: papers 7,801 to 7,900 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Learning the policy for mixed electric platoon control of automated and human-driven vehicles at signalized intersection: a random search approach24 Jun 2022 0 repositories listed
-
Phasic Self-Imitative Reduction for Sparse-Reward Goal-Conditioned Reinforcement Learning24 Jun 2022 0 repositories listed
-
Provably Efficient Reinforcement Learning in Partially Observable Dynamical Systems24 Jun 2022 0 repositories listed
-
Value Function Decomposition for Iterative Design of Reinforcement Learning Agents24 Jun 2022 0 repositories listed
-
A Federated Reinforcement Learning Method with Quantization for Cooperative Edge Caching in Fog Radio Access Networks23 Jun 2022 0 repositories listed
-
Nearly Minimax Optimal Reinforcement Learning with Linear Function Approximation23 Jun 2022 0 repositories listed
-
Recursive Reinforcement Learning23 Jun 2022 0 repositories listed
-
Reinforcement Learning under Partial Observability Guided by Learned Environment Models23 Jun 2022 0 repositories listed
-
Decentralized Gossip-Based Stochastic Bilevel Optimization over Communication Networks22 Jun 2022 0 repositories listed
-
Fusion of Model-free Reinforcement Learning with Microgrid Control: Review and Vision22 Jun 2022 0 repositories listed
-
Learning Optimal Treatment Strategies for Sepsis Using Offline Reinforcement Learning in Continuous Space22 Jun 2022 0 repositories listed
-
Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning21 Jun 2022 0 repositories listed
-
Finding Optimal Policy for Queueing Models: New Parameterization21 Jun 2022 0 repositories listed
-
Hybridization of evolutionary algorithm and deep reinforcement learning for multi-objective orienteering optimization21 Jun 2022 0 repositories listed
-
Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars21 Jun 2022 0 repositories listed
-
Model-Based Imitation Learning Using Entropy Regularization of Model and Policy21 Jun 2022 0 repositories listed
-
Safe and Psychologically Pleasant Traffic Signal Control with Reinforcement Learning using Action Masking21 Jun 2022 0 repositories listed
-
The Integration of Machine Learning into Automated Test Generation: A Systematic Mapping Study21 Jun 2022 0 repositories listed
-
Constrained Reinforcement Learning for Robotics via Scenario-Based Programming20 Jun 2022 0 repositories listed
-
Deep reinforced active learning for multi-class image classification20 Jun 2022 0 repositories listed
-
From Multi-agent to Multi-robot: A Scalable Training and Evaluation Platform for Multi-robot Reinforcement Learning20 Jun 2022 0 repositories listed
-
Guided Safe Shooting: model based reinforcement learning with safety constraints20 Jun 2022 0 repositories listed
-
A Survey on Model-based Reinforcement Learning19 Jun 2022 0 repositories listed
-
Guarantees for Epsilon-Greedy Reinforcement Learning with Function Approximation19 Jun 2022 0 repositories listed
-
Learning Multi-Task Transferable Rewards via Variational Inverse Reinforcement Learning19 Jun 2022 0 repositories listed
-
Two-Hop Age of Information Scheduling for Multi-UAV Assisted Mobile Edge Computing: FRL vs MADDPG19 Jun 2022 0 repositories listed
-
AnyMorph: Learning Transferable Polices By Inferring Agent Morphology17 Jun 2022 0 repositories listed
-
Deep reinforcement learning for fMRI prediction of Autism Spectrum Disorder17 Jun 2022 0 repositories listed
-
Backbones-Review: Feature Extraction Networks for Deep Learning and Deep Reinforcement Learning Approaches16 Jun 2022 0 repositories listed
-
Reinforcement Learning for Economic Policy: A New Frontier?16 Jun 2022 0 repositories listed
-
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings16 Jun 2022 0 repositories listed
-
Autonomous Platoon Control with Integrated Deep Reinforcement Learning and Dynamic Programming15 Jun 2022 0 repositories listed
-
Contrastive Learning as Goal-Conditioned Reinforcement Learning15 Jun 2022 0 repositories listed
-
Mean-Semivariance Policy Optimization via Risk-Averse Reinforcement Learning15 Jun 2022 0 repositories listed
-
Rethinking Reinforcement Learning for Recommendation: A Prompt Perspective15 Jun 2022 0 repositories listed
-
Revisiting Some Common Practices in Cooperative Multi-Agent Reinforcement Learning15 Jun 2022 0 repositories listed
-
Deep Reinforcement Learning for Exact Combinatorial Optimization: Learning to Branch14 Jun 2022 0 repositories listed
-
FreeKD: Free-direction Knowledge Distillation for Graph Neural Networks14 Jun 2022 0 repositories listed
-
Robust Reinforcement Learning with Distributional Risk-averse formulation14 Jun 2022 0 repositories listed
-
Solving the capacitated vehicle routing problem with timing windows using rollouts and MAX-SAT14 Jun 2022 0 repositories listed
-
Stein Variational Goal Generation for adaptive Exploration in Multi-Goal Reinforcement Learning14 Jun 2022 0 repositories listed
-
Towards a Solution to Bongard Problems: A Causal Approach14 Jun 2022 0 repositories listed
-
Intrinsically motivated option learning: a comparative study of recent methods13 Jun 2022 0 repositories listed
-
13 Jun 2022 0 repositories listed
-
Provable Benefit of Multitask Representation Learning in Reinforcement Learning13 Jun 2022 0 repositories listed
-
Provably Efficient Offline Reinforcement Learning with Trajectory-Wise Reward13 Jun 2022 0 repositories listed
-
Matching options to tasks using Option-Indexed Hierarchical Reinforcement Learning12 Jun 2022 0 repositories listed
-
RL-GA: A Reinforcement Learning-Based Genetic Algorithm for Electromagnetic Detection Satellite Scheduling Problem12 Jun 2022 0 repositories listed
-
Federated Offline Reinforcement Learning11 Jun 2022 0 repositories listed
-
An application of neural networks to a problem in knot theory and group theory (untangling braids)10 Jun 2022 0 repositories listed
-
Deep Multi-Agent Reinforcement Learning with Hybrid Action Spaces based on Maximum Entropy10 Jun 2022 0 repositories listed
-
Dynamic mean field programming10 Jun 2022 0 repositories listed
-
Large-Scale Retrieval for Reinforcement Learning10 Jun 2022 0 repositories listed
-
Multifidelity Reinforcement Learning with Control Variates10 Jun 2022 0 repositories listed
-
Policy Gradient Reinforcement Learning for Uncertain Polytopic LPV Systems based on MHE-MPC10 Jun 2022 0 repositories listed
-
An Optimization Method-Assisted Ensemble Deep Reinforcement Learning Algorithm to Solve Unit Commitment Problems9 Jun 2022 0 repositories listed
-
Overcoming the Spectral Bias of Neural Value Approximation9 Jun 2022 0 repositories listed
-
Quantum Policy Iteration via Amplitude Estimation and Grover Search -- Towards Quantum Advantage for Reinforcement Learning9 Jun 2022 0 repositories listed
-
Receding Horizon Inverse Reinforcement Learning9 Jun 2022 0 repositories listed
-
Regret Analysis of Certainty Equivalence Policies in Continuous-Time Linear-Quadratic Systems9 Jun 2022 0 repositories listed
-
Regret Bounds for Information-Directed Reinforcement Learning9 Jun 2022 0 repositories listed
-
Sample-Efficient Reinforcement Learning in the Presence of Exogenous Information9 Jun 2022 0 repositories listed
-
There is no Accuracy-Interpretability Tradeoff in Reinforcement Learning for Mazes9 Jun 2022 0 repositories listed
-
Action Noise in Off-Policy Deep Reinforcement Learning: Impact on Exploration and Performance8 Jun 2022 0 repositories listed
-
Learning to Generate Prompts for Dialogue Generation through Reinforcement Learning8 Jun 2022 0 repositories listed
-
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games8 Jun 2022 0 repositories listed
-
Reinforced Inverse Scattering8 Jun 2022 0 repositories listed
-
Scalable Joint Learning of Wireless Multiple-Access Policies and their Signaling8 Jun 2022 0 repositories listed
-
Scalable Online Disease Diagnosis via Multi-Model-Fused Actor-Critic Reinforcement Learning8 Jun 2022 0 repositories listed
-
Sim2real for Reinforcement Learning Driven Next Generation Networks8 Jun 2022 0 repositories listed
-
A Model-Based Reinforcement Learning Approach for PID Design7 Jun 2022 0 repositories listed
-
Driving in Real Life with Inverse Reinforcement Learning7 Jun 2022 0 repositories listed
-
MIX-MAB: Reinforcement Learning-based Resource Allocation Algorithm for LoRaWAN7 Jun 2022 0 repositories listed
-
On the Effectiveness of Fine-tuning Versus Meta-reinforcement Learning7 Jun 2022 0 repositories listed
-
On the Role of Discount Factor in Offline Reinforcement Learning7 Jun 2022 0 repositories listed
-
Variational Meta Reinforcement Learning for Social Robotics7 Jun 2022 0 repositories listed
-
Adaptive Rollout Length for Model-Based RL Using Model-Free Deep RL6 Jun 2022 0 repositories listed
-
Asymptotic Instance-Optimal Algorithms for Interactive Decision Making6 Jun 2022 0 repositories listed
-
Balancing Profit, Risk, and Sustainability for Portfolio Management6 Jun 2022 0 repositories listed
-
Consensus Learning for Cooperative Multi-Agent Reinforcement Learning6 Jun 2022 0 repositories listed
-
Deep Reinforcement Learning for Cybersecurity Threat Detection and Protection: A Review6 Jun 2022 0 repositories listed
-
Efficient entity-based reinforcement learning6 Jun 2022 0 repositories listed
-
Real2Sim or Sim2Real: Robotics Visual Insertion using Deep Reinforcement Learning and Real2Sim Policy Adaptation6 Jun 2022 0 repositories listed
-
Provably Efficient Risk-Sensitive Reinforcement Learning: Iterated CVaR and Worst Path6 Jun 2022 0 repositories listed
-
Specification-Guided Learning of Nash Equilibria with High Social Welfare6 Jun 2022 0 repositories listed
-
DDPG based on multi-scale strokes for financial time series trading strategy5 Jun 2022 0 repositories listed
-
Learning Dynamics and Generalization in Reinforcement Learning5 Jun 2022 0 repositories listed
-
Models of human preference for learning reward functions5 Jun 2022 0 repositories listed
-
Rapid Learning of Spatial Representations for Goal-Directed Navigation Based on a Novel Model of Hippocampal Place Fields5 Jun 2022 0 repositories listed
-
Adaptive Tree Backup Algorithms for Temporal-Difference Reinforcement Learning4 Jun 2022 0 repositories listed
-
Between Rate-Distortion Theory & Value Equivalence in Model-Based Reinforcement Learning4 Jun 2022 0 repositories listed
-
Deciding What to Model: Value-Equivalent Sampling for Reinforcement Learning4 Jun 2022 0 repositories listed
-
Hybrid Value Estimation for Off-policy Evaluation and Offline Reinforcement Learning4 Jun 2022 0 repositories listed
-
MACC: Cross-Layer Multi-Agent Congestion Control with Deep Reinforcement Learning4 Jun 2022 0 repositories listed
-
Reward Poisoning Attacks on Offline Multi-Agent Reinforcement Learning4 Jun 2022 0 repositories listed
-
A Deep Reinforcement Learning Framework For Column Generation3 Jun 2022 0 repositories listed
-
Disentangling Epistemic and Aleatoric Uncertainty in Reinforcement Learning3 Jun 2022 0 repositories listed
-
Joint Energy Dispatch and Unit Commitment in Microgrids Based on Deep Reinforcement Learning3 Jun 2022 0 repositories listed
-
KCRL: Krasovskii-Constrained Reinforcement Learning with Guaranteed Stability in Nonlinear Dynamical Systems3 Jun 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.