| Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperative Multi-Agent Reinforcement Learning added by Syntology |
2026-08 (from id) |
RS2002/MA-USFA/tsc/agent.py cfc56b7bea4330e3 |
unverified |
MIT (permissive) |
| Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning added by Syntology |
2026-08 (from id) |
sfujim/BCQ/discrete_BCQ/utils.py 2a282db4ff509a1c |
unverified |
MIT (permissive) |
| ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection added by Syntology |
2026-08 (from id) |
CLendering/ReFP-AD/pipeline.py 91230af41ce76aa9 |
ran
|
MIT (permissive) |
| QSplitFL: Capability Aware Deep Q-Learning for Optimal Split Point Selection in Split Federated Learning Nazmus Shakib Shadin 1[0000-0002-0999-924X] , Xinyue Zhang 1[0000-0002-4243-083X] , Jingyi Wang 2[0000-0002-4988-5626] , and Miao Pan 3[0000-0003-2138-4413] added by Syntology |
2026-06 (from id) |
AIPO-Lab/QSplitFL/Capability_Aware_DQN_Implementation/dqn_split_agent.py e02b292fe78a631c |
ran
|
no licence file found · pointer only |
| Debiased Model-based Representations for Sample-efficient Continuous Control added by Syntology |
2026-05 (from id) |
dmksjfl/DR.Q/DRQ/DRQ.py 5ef134cb3440fdcb |
ran
|
MIT (permissive) |
| SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data added by Syntology |
2026-05 (from id) |
CarloRomeo427/SOPE/src/algos/agent_sope.py 068b007d0eb9bea3 |
ran
|
MIT (permissive) |
| RL-ABC: Reinforcement Learning for Accelerator Beamline Control added by Syntology |
2026-04 (from id) |
Anwar9Ibrahim/RL-ABC/rl_framework/Agents/DDPG.py 208064288fc2783c |
ran
|
MIT (permissive) |
| Learning to Bet for Horizon-Aware Anytime-Valid Testing added by Syntology |
2026-03 (from id) |
egetaga/learning-to-bet/dqn/models.py 906383c6649dbfea |
ran
|
MIT (permissive) |
| DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty |
14 Jun 2025 |
lemutisme/dr-sac/sac.py 45c5e6c336c21050 |
unverified |
MIT (permissive) |
| Towards General-Purpose Model-Free Reinforcement Learning |
27 Jan 2025 |
facebookresearch/MRQ/MRQ/MRQ.py 39dd84fd5db384b3 |
unverified |
licence not identified · pointer only |
| RVI-SAC: Average Reward Off-Policy Deep Reinforcement Learning |
4 Aug 2024 |
yhisaki/average-reward-drl/average_reward_drl/algorithms/rvi_sac.py 6e334ad3a7618566 |
unverified |
no licence file found · pointer only |
| Transition Path Sampling with Improved Off-Policy Training of Diffusion Path Samplers |
30 May 2024 |
kiyoung98/tps-dps/src/dps.py 564a155eb9ec3bee |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning |
24 May 2024 |
hansenhua/ollie-offline-to-online-imitation-learning/offline_to_online.py edb756571764039b |
ran
|
no licence file found · pointer only |
| Iterated Denoising Energy Matching for Sampling from Boltzmann Densities |
9 Feb 2024 |
jarridrb/dem/dem/models/components/sde_integration.py c0238358b3127562 |
ran
|
MIT (permissive) |
| Learning Uncertainty-Aware Temporally-Extended Actions |
8 Feb 2024 |
oh-lab/UTE-Uncertainty-aware-Temporal-Extension-/chain_mdp/agent/temporal_extension.py 5023894548dbacbb |
ran
|
Apache-2.0 (permissive) |
| Boosting Reinforcement Learning with Strongly Delayed Feedback Through Auxiliary Short Delays |
5 Feb 2024 |
qingyuanwunothing/ad-rl/AD-DQN.py 68200fece3c9d915 |
unverified |
no licence file found · pointer only |
| Weakly Coupled Deep Q-Networks |
28 Oct 2023 |
ibrahim-elshar/WCDQN_NeurIPS/src/Inv_control/WCDQN.py 0a34c84d33066fee |
ran
|
no licence file found · pointer only |
| Natural Actor-Critic for Robust Reinforcement Learning with Function Approximation |
17 Jul 2023 |
tliu1997/rnac/train_rnac.py 31b7a842f737b97e |
ran
|
no licence file found · pointer only |
| Distance Weighted Supervised Learning for Offline Interaction Data |
26 Apr 2023 |
jhejna/dwsl/research/algs/dwsl.py 377a3792bec4c050 |
unverified |
MIT (permissive) |
| Max-Min Off-Policy Actor-Critic Method Focusing on Worst-Case Robustness to Model Misspecification |
7 Nov 2022 |
akimotolab/m2td3/M2TD3/m2td3.py badc218c65d7931c |
ran
|
no licence file found · pointer only |
| Learning GFlowNets from partial episodes for improved convergence and stability |
26 Sep 2022 |
gfnorg/gflownet/grid/toy_grid_dag.py d00f3663f26789a6 |
ran
|
MIT (permissive) |
| Dropout Q-Functions for Doubly Efficient Reinforcement Learning |
5 Oct 2021 |
watchernyu/REDQ/redq/algos/redq_sac.py 820da6f0b914ce98 |
ran
|
MIT (permissive) |
| Offline Meta-Reinforcement Learning with Advantage Weighting |
13 Aug 2020 |
eric-mitchell/macaw-min/impl.py 02080cb75a3041f8 |
unverified |
no licence file found · pointer only |
| An Equivalence between Loss Functions and Non-Uniform Sampling in Experience Replay |
12 Jul 2020 |
sfujim/LAP-PAL/discrete/utils.py d60cce6245a34fcb |
unverified |
MIT (permissive) |
| Self-Imitation Learning via Generalized Lower Bound Q-learning |
12 Jun 2020 |
junhyukoh/self-imitation-learning/baselines/common/self_imitation.py 92fcff634a9f6c9e |
ran
|
MIT (permissive) |
| An Efficient Transfer Learning Framework for Multiagent Reinforcement Learning |
2020-02 (from id) |
tianpeiyang/maptf_code/alg/option/option_ma_sro.py a6dfa17c8b218bdd |
ran
|
no licence file found · pointer only |
| Dream to Control: Learning Behaviors by Latent Imagination |
3 Dec 2019 |
zhixuan-lin/dreamer-pytorch/dreamer.py 8c761148be72a1af |
ran
|
MIT (permissive) |
| Dream to Control: Learning Behaviors by Latent Imagination |
3 Dec 2019 |
adityabingi/Dreamer/dreamer.py 9250128488be37e8 |
ran
|
MIT (permissive) |
| Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model |
19 Nov 2019 |
dmiracle/muzero-starter/pseudocode.py 6ff1fe24a95ea773 |
unverified |
no licence file found · pointer only |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
MatthieuSarkis/Portfolio-Optimization-and-Goal-Based-Investment-with-Reinforcement-Learning/src/agents.py ac5ecca5b7705f8a |
ran
|
Apache-2.0 (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
kdally/fault-tolerant-flight-control-drl/fault_tolerant_flight_control_drl/agent/sac.py 8119b0afbe2d940e |
ran
|
MIT (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
BY571/Soft-Actor-Critic-and-Extensions/files/Agent.py a6c39d1200737d3f |
ran
|
MIT (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
yhisaki/average-reward-drl/average_reward_drl/algorithms/sac.py e9664791458e5ade |
ran
|
no licence file found · pointer only |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
quantumiracle/Popular-RL-Algorithms/sac_v2.py 70249df825c62d7e |
ran
|
Apache-2.0 (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
thomashirtz/soft-actor-critic/soft_actor_critic/agent.py 0020d85f4dbfb10b |
ran
|
MIT (permissive) |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
MrDaubinet/collaboration-and-competition/maddpg.py ae26aa48075ee6cb |
ran
|
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
baradist/multiagent-particle-envs/multiagent/ddpg/ddpg_agent.py c90519f77627be1f |
ran
|
MIT (permissive) |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
krasing/DRLearningCollaboration/failed_collaborative/ddpg_agent_multi_2.py 4cbc44a206f6120a |
ran
|
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
bonniesjli/MADDPG_Tennis_UnityML/MADDPG.py 21d74b782714cf9e |
ran
|
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
baoqianwang/iros22_darl1n/maddpg_o/maddpg_local/trainer/maddpg.py 0ec668ec81ebce4d |
unverified |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
EyaRhouma/collaboration-competition-MADDPG/MADDPG_agent.py 0fe6e65cdf5950a9 |
unverified |
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
CharlotteMorrison/Baxter-Research/td3algorithm/priority-replay/prioritized_replay_buffer.py e13397d005876228 |
ran
|
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
emiled16/Beyond_prioritized_experience_replay/final_buffer.py 4dfb2bb869a28e56 |
ran
|
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
CharlotteMorrison/Baxter-VREP/td3/experience/priority_replay_buffer.py c2cbd2d0d7d66167 |
ran
|
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
tensorlayer/RLzoo/rlzoo/common/buffer.py f4455aa8614ee649 |
ran
|
Apache-2.0 (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
atavakol/action-branching-agents/agents/bdq/deepq/replay_buffer.py f6c715a36c9b0278 |
ran
|
MIT (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
Arrabonae/openai_DDDQN/replay.py 82024cb507cb1dee |
ran
|
Apache-2.0 (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
CharlotteMorrison/Baxter-VREP-Version-2/td3/experience/priority_replay_buffer.py a1b54613a3985829 |
ran
|
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
MathPhysSim/PER-NAF/pernaf/pernaf/utils/prioritised_experience_replay.py 46cf03e4d4b628c3 |
ran
|
MIT (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
SayhoKim/tetrisRL/lib/utils/replay_buffer.py 867fd67d42cbe6f4 |
ran
|
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
labmlai/annotated_deep_learning_paper_implementations/labml_nn/rl/dqn/replay_buffer.py 6e5e90f57a133b61 |
ran
|
MIT (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
hill-a/stable-baselines/stable_baselines/common/buffers.py f5eb1be22a186f73 |
unverified |
MIT (permissive) |
| Prioritized Experience Replay |
18 Nov 2015 |
seacevedo/ReinforcementLearningProjects/PER.py 4ec9642b187da0b4 |
unverified |
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
Guillaume-Cr/lunar_lander_per/replay_buffer.py f426229e0c5e2d70 |
unverified |
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
nbopardi/smb/dqn_utils.py 1f3901918f090c5e |
unverified |
no licence file found · pointer only |
| Prioritized Experience Replay |
18 Nov 2015 |
Adrelf/DRL-navigation/utils/memory.py 7859b3965dc78ae3 |
unverified |
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
philtabor/Deep-Q-Learning-Paper-To-Code/DDQN/ddqn_agent.py 2a2892343a3359ee |
ran
|
MIT (permissive) |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
MEOWMEOW114/nd893-p1-navigation-banana/dqn/double_dqn_agent.py dcd7958b0ad16526 |
ran
|
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
marload/DeepRL-TensorFlow2/DoubleDQN/DoubleDQN_Discrete.py 56d9058560cd0b68 |
ran
|
Apache-2.0 (permissive) |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
mohit8935/Deep-Q-Learning-Paper/DDQN/ddqn_agent.py 41d448d208226341 |
ran
|
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
Adrelf/DRL-navigation/brain/dqn_agent.py 0e225401249db6d9 |
ran
|
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
1jsingh/rl_navigation/agents/dqn_agent.py f07ed06c9c4c4e7c |
unverified |
MIT (permissive) |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
fengsterooni/dql/DDQN/ddqn_agent.py 15af333b2f369749 |
unverified |
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
shehrum/RL_Navigation/dqn_agent.py fd4b196396e41a0f |
unverified |
no licence file found · pointer only |
| Deep Reinforcement Learning with Double Q-learning |
22 Sep 2015 |
jeffery1236/Atari_DoubleDeepQNetwork/DDQAgent.py 28d5d41e56d53b63 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
yukezhu/tensorflow-reinforce/rl/pg_ddpg.py 3c7178cf937ba425 |
ran
|
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
dpoulopoulos/drl_collaborate_compete/ddpg_agent.py 1a3c5297309ff7b3 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
YoUNG824/DDPG/ddpg.py cbd6d4cba3c8f7b0 |
ran
|
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
liampetti/DDPG/ddpg.py 532a74ad7f880104 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
iDataist/Continuous-Control-with-Deep-Deterministic-Policy-Gradient/ddpg_agent.py 6258582b31c25005 |
ran
|
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
J93T/TP4-DDPG/ddpg.py 79dfb108cab48c77 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
dan-lennox/ml-udacity-quadcopter-rl/agents/agent.py 48908dfdd43f1c2f |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
HJDQN/HJQ/algorithms/ddpg/ddpg.py a15460ad3a6b391e |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
MathPhysSim/PER-NAF/pernaf/pernaf/naf.py d8fadcf1cbc892d5 |
ran
|
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
dfki-ric-underactuated-lab/torque_limited_simple_pendulum/software/python/simple_pendulum/reinforcement_learning/ddpg/ddpg.py 3869ba95ec3a2e1a |
ran
|
BSD-3-Clause (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
madhur-tandon/RL-Project/agent.py 87674610395558bb |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
T3chy/DDPG/ddpg_torch.py 7ebcb237f3ac8b6b |
ran
|
GPL-3.0 (copyleft) · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
alathiya/RL-Quadcoptor-Flying/agents/agent.py a1866c0223c9dca1 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
marload/DeepRL-TensorFlow2/DDPG/DDPG_Continuous.py 031efd28ccd13e08 |
ran
|
Apache-2.0 (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
bmeyers/VirtualMicrogridSegmentation/virtual_microgrids/algorithms/ddpg.py 88607f875726a410 |
ran · metamorphic tier: well formed
fingerprinted |
BSD-2-Clause (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
denizmguen/IANNWTF2019-Project/ddpg.py 49ac8c4b2338f149 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
samuelmat19/DDPG-tf2/src/model.py 0d2b4efe0274fc66 |
unverified |
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
Philori22/DDPG-aigym/ddpg.py 1229a5d8dea62487 |
unverified |
no licence file found · pointer only |