| Integrating Novelty and Surprise for Experience Prioritization and Exploration in Image-Based Reinforcement Learning added by Syntology |
2026-08 (from id) |
UoA-CARES/NSPER/nsper_td3.py 38a6c669a5af885c |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning added by Syntology |
2026-08 (from id) |
familyld/SCORE/score/SCORE.py e2c86f02b964230e |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| ReBRAC-v2: The Return of the King added by Syntology |
2026-08 (from id) |
Simple-Robotics/guided-flow-policy/agents/rebrac.py af17450f0bb38797 |
ran
|
MIT (permissive) |
| ReBRAC-v2: The Return of the King added by Syntology |
2026-08 (from id) |
JongseongChae/FAC/agents/rebrac.py 0df80f5dd0c1a3a9 |
unverified |
no licence file found · pointer only |
| SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration added by Syntology |
2026-06 (from id) |
montrealrobotics/shapo/safepo/single_agent/shapo.py e8e31987c5d58926 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| RL-ABC: Reinforcement Learning for Accelerator Beamline Control added by Syntology |
2026-04 (from id) |
Anwar9Ibrahim/RL-ABC/rl_framework/Agents/DDPG.py d3e22fc2c8926db9 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Scalable Neighborhood-Based Multi-Agent Actor-Critic added by Syntology |
2026-04 (from id) |
TimGop/MADDPG-K/src/MADDPG_agent.py af9fbaf52f3cd1cb |
ran
|
no licence file found · pointer only |
| Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints added by Syntology |
18 Mar 2026 |
Elessar123/SAC-FLOW/from_scratch_code/SAC_flow_gru_jax.py cbe4b2a64b8a77d3 |
unverified |
no licence file found · pointer only |
| RESCHED: Rethinking Flexible Job Shop Scheduling from a Transformer-based Architecture with Simplified States added by Syntology |
2026-03 (from id) |
XiangjieXiao/ReSched/PPO/SchedulingModel.py c43444a37a5a6755 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| arXiv:2507.00485 |
2025-07 (from id) |
azure-123/PNAct/safepo/common/model.py 832a4c764134d431 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty |
14 Jun 2025 |
lemutisme/dr-sac/sac.py b3d3944663449c1a |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Intention-Conditioned Flow Occupancy Models |
10 Jun 2025 |
chongyi-zheng/infom/agents/infom.py 17f5041025f7c2b2 |
unverified |
MIT (permissive) |
| Maximum Total Correlation Reinforcement Learning |
22 May 2025 |
bangyou01/mtc/tcsac.py 8f578d66fd15da10 |
unverified |
no licence file found · pointer only |
| Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models |
15 Apr 2025 |
facebookresearch/metamotivo/metamotivo/fb/model.py 92468117ea8270ba |
unverified |
licence not identified · pointer only |
| Towards General-Purpose Model-Free Reinforcement Learning |
27 Jan 2025 |
sfujim/TD7/TD7.py e41e9160d3bfd93f |
ran
|
MIT (permissive) |
| Reinforcement Learning Policy as Macro Regulator Rather than Macro Placer |
10 Dec 2024 |
lamda-bbo/macro-regulator/src/agent.py f29f84d9f963b618 |
unverified |
BSD-3-Clause (permissive) |
| Robot Policy Learning with Temporal Optimal Transport Reward |
29 Oct 2024 |
fuyw/temporalot/models/temporalot.py bac5b63bccf2761a |
unverified |
no licence file found · pointer only |
| Overcoming Slow Decision Frequencies in Continuous Control: Model-Based Sequence Reinforcement Learning for Model-Free Control |
11 Oct 2024 |
dee0512/Temporally-Layered-Architecture/model.py 464ed648e3d7a3e2 |
ran
fingerprinted |
BSD-3-Clause (permissive) |
| Uncertainty-Aware Reward-Free Exploration with General Function Approximation |
24 Jun 2024 |
uclaml/gfa-rfe/agent/dsquare.py 419af119bd09b0ac |
ran
|
MIT (permissive) |
| Beyond Optimism: Exploration With Partially Observable Rewards |
20 Jun 2024 |
AmiiThinks/mon_mdp_neurips24/src/actor.py bce144a0a54c2361 |
unverified |
CC-BY-4.0 · pointer only |
| Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control |
25 May 2024 |
sfujim/TD3/TD3.py 9c81de6b9c4c134f |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Revisiting Data Augmentation in Deep Reinforcement Learning |
19 Feb 2024 |
jianshu-hu/drqv2/drqv2.py cc3f9517e01dd145 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization |
30 Oct 2023 |
XuGW-Kevin/DrM/agents/drm.py ac279413d97dfbb0 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Towards Robust Offline Reinforcement Learning under Diverse Data Corruption |
19 Oct 2023 |
zzmtsvv/ORL/riql/riql.py 3a5956686dfd460b |
ran
fingerprinted |
no licence file found · pointer only |
| Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages |
11 Oct 2023 |
Guozheng-Ma/Adaptive-Replay-Ratio/drqv2_adapt_rr.py 4f9c561e7f66c0aa |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Offline Multi-Agent Reinforcement Learning with Implicit Global-to-Local Value Regularization |
21 Jul 2023 |
zhengyinan-air/omiga/algos/OMIGA.py 5eb07edf75209baa |
ran
|
no licence file found · pointer only |
| Policy Regularization with Dataset Constraint for Offline Reinforcement Learning |
11 Jun 2023 |
lamda-rl/prdc/prdc.py b6bad7ca8ac45549 |
ran
fingerprinted |
no licence file found · pointer only |
| Preference-grounded Token-level Guidance for Language Model Fine-tuning |
1 Jun 2023 |
yinyueqin/denserewardrlhf-ppo/denserlhf/trainer/ppo_utils/experience_maker.py 2c1a919b8596c263 |
unverified |
Apache-2.0 (permissive) |
| Off-Policy Average Reward Actor-Critic with Deterministic Policy Search |
20 May 2023 |
namansaxena9/ARO-DDPG/cheetah_run/ddpg_model.py 8dc8b64c468580f4 |
unverified |
no licence file found · pointer only |
| MAHTM: A Multi-Agent Framework for Hierarchical Transactive Microgrids |
15 Mar 2023 |
nicosquare/rl-energy-management/src/algos/rl/coma/d_simple_microgrid.py 9fe565930f928e18 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Extreme Q-Learning: MaxEnt RL without Entropy |
5 Jan 2023 |
zzmtsvv/rl_task/eql/eql.py f7541cc0796ce882 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information |
31 Oct 2022 |
manantomar/agent-centric-representations/acro.py 1de9ecd6db95db2e |
ran
|
no licence file found · pointer only |
| DeepTOP: Deep Threshold-Optimal Policy for MDPs and RMABs |
18 Sep 2022 |
khalednakhleh/deeptop/MDP/DeepTOP.py 3473662fd8d11551 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Robust Reinforcement Learning using Offline Data |
10 Aug 2022 |
zaiyan-x/RFQI/rfqi.py bd43dccfb7a0ad6d |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Discriminator-Weighted Offline Imitation Learning from Suboptimal Demonstrations |
20 Jul 2022 |
zexusun/oilca-neurips23/algos/augment_DWBC.py b6455496f9ccd9d2 |
unverified |
no licence file found · pointer only |
| Learning Bellman Complete Representations for Offline Policy Evaluation |
12 Jul 2022 |
causalml/bcrl/bcrl/agent.py 5b1fa9891b2abef9 |
ran
|
MIT (permissive) |
| MineDojo: Building Open-Ended Embodied Agents with Internet-Scale Knowledge |
17 Jun 2022 |
pku-rl/copl/src/core/ppo.py 7afc0befd6b15a83 |
unverified |
MIT (permissive) |
| Mildly Conservative Q-Learning for Offline Reinforcement Learning |
9 Jun 2022 |
zzmtsvv/ORL/mcq/mcq.py 55a4cdf8fa60dd90 |
ran
fingerprinted |
no licence file found · pointer only |
| CCLF: A Contrastive-Curiosity-Driven Learning Framework for Sample-Efficient Reinforcement Learning |
2 May 2022 |
csun001/CCLF/CCLF_sac.py 9d46b0bfcafbd714 |
unverified |
MIT (permissive) |
| PMIC: Improving Multi-Agent Reinforcement Learning with Progressive Mutual Information Collaboration |
16 Mar 2022 |
yeshenpy/pmic/algorithms/mpe_new_maxminMADDPG.py 02d8b297accf97de |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| VRL3: A Data-Driven Framework for Visual Deep Reinforcement Learning |
17 Feb 2022 |
facebookresearch/drqv2/drqv2.py 23bd8540206300a8 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy Transfer |
10 Feb 2022 |
xingyul/revolver/gym/SAC.py 2bafdd50987880d8 |
ran
|
GPL-2.0 (copyleft) · pointer only |
| Versatile Offline Imitation from Observations and Examples via Regularized State-Occupancy Matching |
4 Feb 2022 |
ryanxhr/dwbc/algos/DWBC.py da00c16c9fad3d61 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Offline Reinforcement Learning with Implicit Q-Learning |
12 Oct 2021 |
orrivlin/implicit-q-learning/IQL.py 9a453c413229c3e8 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Offline Reinforcement Learning with Implicit Q-Learning |
12 Oct 2021 |
Manchery/iql-pytorch/IQL.py 966c8f4b02ba3120 |
ran
fingerprinted |
no licence file found · pointer only |
| Offline Reinforcement Learning with Implicit Q-Learning |
12 Oct 2021 |
BY571/Implicit-Q-Learning/agent.py b6a1eedb944158c8 |
ran
fingerprinted |
no licence file found · pointer only |
| Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative Transformer |
6 Oct 2021 |
yining043/VRP-DACT/nets/actor_network.py c62ea025f6578059 |
unverified |
MIT (permissive) |
| Uncertainty-Based Offline Reinforcement Learning with Diversified Q-Ensemble |
4 Oct 2021 |
corl-team/CORL/algorithms/offline/edac.py 7f11d5ce17556f04 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Large Batch Experience Replay |
4 Oct 2021 |
sureli/laber/LaBER/continuous/LABER_SAC.py 508f079e9d0ccf91 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Hindsight Value Function for Variance Reduction in Stochastic Dynamic Environment |
26 Jul 2021 |
guojm14/HVF/reduce_var/mlp.py f9dab6339e71b2bb |
ran
|
MIT (permissive) |
| Ask&Confirm: Active Detail Enriching for Cross-Modal Retrieval with Partial Query |
2 Mar 2021 |
CuthbertCai/Ask-Confirm/models/policy_model.py 8a7f1e20f75ba4f3 |
unverified |
no licence file found · pointer only |
| Randomized Ensembled Double Q-Learning: Learning Fast Without a Model |
15 Jan 2021 |
BY571/Randomized-Ensembled-Double-Q-learning-REDQ-/agent.py 89265e3225a6d513 |
ran
fingerprinted |
no licence file found · pointer only |
| Softmax Deep Double Deterministic Policy Gradients |
19 Oct 2020 |
ling-pan/SD3/SD3.py 62767acd88e18e64 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Self-Supervised Policy Adaptation during Deployment |
8 Jul 2020 |
nicklashansen/policy-adaptation-during-deployment/src/agent/agent.py d526a377580646ff |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Generating Adjacency-Constrained Subgoals in Hierarchical Reinforcement Learning |
20 Jun 2020 |
trzhang0116/HRAC/hrac/hrac.py 1b8c1dbf0a21cb68 |
ran · metamorphic tier: invariant
fingerprinted |
Apache-2.0 (permissive) |
| AWAC: Accelerating Online Reinforcement Learning with Offline Datasets |
16 Jun 2020 |
zzmtsvv/rl_task/awac/awac.py eecebb628e9b243d |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Conservative Q-Learning for Offline Reinforcement Learning |
8 Jun 2020 |
BY571/CQL/CQL-SAC/agent.py 34ba8aa41bb7fa58 |
ran
fingerprinted |
no licence file found · pointer only |
| Reinforcement Learning with Augmented Data |
30 Apr 2020 |
MishaLaskin/rad/curl_sac.py 5aaa7c97de18af3a |
ran
|
no licence file found · pointer only |
| Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from Pixels |
28 Apr 2020 |
xingyu-lin/softagent/drq/Drq.py 8ad237c6d1f93986 |
unverified |
no licence file found · pointer only |
| Balancing Training for Multilingual Neural Machine Translation |
14 Apr 2020 |
cindyxinyiwang/DataSelection/src/actor.py 824df24cce12a367 |
ran
|
no licence file found · pointer only |
| Online Meta-Critic Learning for Off-Policy Actor-Critic Methods |
11 Mar 2020 |
zwfightzw/Meta-Critic/TD3_DDPG_MC/DDPG_MC.py d263091df5df7826 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Estimating Q(s,s') with Deep Deterministic Dynamics Gradients |
21 Feb 2020 |
uber-research/D3G/D3G/D3G.py 87d8bb6bd7e83fe9 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Effective Diversity in Population Based Reinforcement Learning |
3 Feb 2020 |
holounic/DvD-TD3/td3/algorithm.py 01cff9d67cdcaecf |
ran
fingerprinted |
no licence file found · pointer only |
| Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction |
3 Jun 2019 |
zzmtsvv/rl_task/bear/bear.py 434fd896deb0893a |
ran
|
no licence file found · pointer only |
| Off-Policy Deep Reinforcement Learning without Exploration |
7 Dec 2018 |
theSparta/off_policy_mujoco/RSEM.py a4c895a75dcac04b |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Off-Policy Deep Reinforcement Learning without Exploration |
7 Dec 2018 |
maziarg/PrivAttack-BCQ/continuous_BCQ/BCQ.py fa0f609acc14b1b7 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |
| Off-Policy Deep Reinforcement Learning without Exploration |
7 Dec 2018 |
thxsxth/POMDP_RLSepsis/RL/other RL attempts/BCQ_models.py f8f318e0c81d2112 |
ran
|
no licence file found · pointer only |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
MatthieuSarkis/Portfolio-Optimization-and-Goal-Based-Investment-with-Reinforcement-Learning/src/agents.py de5e64dc2bfc6a33 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
BY571/Soft-Actor-Critic-and-Extensions/files/Agent.py 5d8f81c210a0aacb |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor |
4 Jan 2018 |
SaminYeasar/off_policy_ac/SAC/SAC.py dc01224a73ea8148 |
ran
fingerprinted |
no licence file found · pointer only |
| Proximal Policy Optimization Algorithms |
20 Jul 2017 |
bonniesjli/PPO-Reacher_UnityML/agent.py e9f1207fc08f9fa9 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Proximal Policy Optimization Algorithms |
20 Jul 2017 |
bonniesjli/PPO_Reacher/agent.py 8fc0ce07edd32744 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Proximal Policy Optimization Algorithms |
20 Jul 2017 |
Gregory-Eales/proximal-policy-optimization/ppo/modules/ppo.py c3cb4b39b7e82812 |
ran · metamorphic tier: deterministic
|
Apache-2.0 (permissive) |
| Proximal Policy Optimization Algorithms |
20 Jul 2017 |
marload/DeepRL-TensorFlow2/PPO/PPO_Continuous.py 54ccb8ae3f658c83 |
unverified |
Apache-2.0 (permissive) |
| Proximal Policy Optimization Algorithms |
20 Jul 2017 |
InSpaceAI/RL-Zoo/PPO.py bf0c1da602f8a16d |
unverified |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
zowiezhang/stas/models/policy/COMA.py d33580612fa8578d |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
rainandwind1/MADDPG-reconstruct/MADDPG/model.py 394e91b566a8e7eb |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
MrDaubinet/collaboration-and-competition/maddpg.py 147d57177bbffdd9 |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
starry-sky6688/MADDPG/maddpg/maddpg.py b1b715fc5cea22be |
ran
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
krasing/DRLearningCollaboration/failed_collaborative/ddpg_agent_multi_2.py ab8b87babe2b659e |
ran
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
petsol/MultiAgentCooperation_UnityAgent_MADDPG_Udacity/multiagent_resources.py a17bd67f978b7355 |
ran
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
isp1tze/MAProj/algo/maddpg/maddpg_agent.py cbd481c5cf284628 |
ran
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
bonniesjli/MADDPG_Tennis_UnityML/MADDPG.py af7b2cc00bcf3971 |
ran
fingerprinted |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
thechrisyoon08/marl/MADDPG/maddpg.py 19cc670e3070c844 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
Ah31/maddpg_pytorch/MADDPG_model.py e0a77db60f82cb02 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
ksajan/DDPG-MAPE/multiagent-particle-envs/maddpg.py 85e19366b66e12c8 |
unverified |
no licence file found · pointer only |
| Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments |
7 Jun 2017 |
anonymous-iclr22/trust-region-in-multi-agent-reinforcement-learning/algorithms/happo_policy.py fe0a30f4af252b51 |
unverified |
MIT (permissive) |
| Asynchronous Methods for Deep Reinforcement Learning |
4 Feb 2016 |
Remtasya/DDPG-Actor-Critic-Reinforcement-Learning-Reacher-Environment/ddpg_agent.py a83ecbc1e78102e4 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Asynchronous Methods for Deep Reinforcement Learning |
4 Feb 2016 |
bentrevett/pytorch-rl/n_step_a2c.py 75bd98e327fba9f3 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Asynchronous Methods for Deep Reinforcement Learning |
4 Feb 2016 |
InSpaceAI/RL-Zoo/A2C.py 4dd9b378deccacfd |
unverified |
no licence file found · pointer only |
| Asynchronous Methods for Deep Reinforcement Learning |
4 Feb 2016 |
marload/DeepRL-TensorFlow2/A3C/A3C_Continuous.py 223dec3bdd457c46 |
unverified |
Apache-2.0 (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
ghliu/pytorch-ddpg/ddpg.py fb986aef8cb19dcb |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
iDataist/Continuous-Control-with-Deep-Deterministic-Policy-Gradient/ddpg_agent.py 64db7b42a6e83850 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
TheInfamousWayne/ddpg/ddpg.py 976480803a792876 |
ran · metamorphic tier: deterministic
fingerprinted |
Apache-2.0 (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
J93T/TP4-DDPG/ddpg.py 6d718c5a05a53b10 |
ran · metamorphic tier: deterministic
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
HJDQN/HJQ/algorithms/ddpg/ddpg.py 4ea2dabac66df07c |
ran · metamorphic tier: deterministic
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
IvanVigor/Deep-Deterministic-Policy-Gradient-Unity-Env/ddpg_agent.py 2233e3bf2203e01d |
ran · metamorphic tier: deterministic
fingerprinted |
GPL-3.0 (copyleft) · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
baturaysaglam/ac-off-poc/DDPG & TD3/AC_Off_POC_DDPG.py 024ce7deaed96813 |
ran · metamorphic tier: deterministic
fingerprinted |
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
b06b01073/continuous-control/ddpg.py b7b4ba4db14ceff3 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
ZainRaza14/deepRL/DDPG/ddpg_agent.py 8062e958090e2e1f |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
madhur-tandon/RL-Project/agent.py c3081f2569beb64e |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
biemann/Continuous-Control/ddpg_agent.py 59b451ca369c1de8 |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
krasing/DRLearningContinuousControl/ddpg_agent.py dc8d86e2dc6fd389 |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
baturaysaglam/RIS-MISO-Deep-Reinforcement-Learning/DDPG.py 6eca632bf823e7db |
ran
fingerprinted |
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
SergiPonsa/Reinforcement-Learning-Sergi/ddpg pendulum/model_comented.py da2a0693988d4238 |
ran
|
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
hemilpanchiwala/Hindsight-Experience-Replay/ddpg_with_her/DDPGAgent.py 00d63e51f73faefb |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
krasing/DRLearningCollaboration/model.py b0866a4da4a62415 |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
claudeHifly/BipedalWalker-v3/DDPG/ddpg_agent.py 328c50f6e3095851 |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
wwydmanski/rl_tennis/brain/model.py 462527aca7c44cd5 |
ran
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
ZiyangY/IndProject-RL-in-Supply-chain/DDPG-based RNN/DDPG.py 1f1a916ff3cde3a5 |
ran
|
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
iDataist/Tennis-With-Multi-Agent-Reinforcement/model.py d96d6110c00ab9d4 |
ran · metamorphic tier: invariant
fingerprinted |
MIT (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
Philori22/DDPG-aigym/ddpg.py 8f5c14d98d39f3b9 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
SaminYeasar/off_policy_ac/DDPG/DDPG.py 200b44ea7ead9ab6 |
ran · metamorphic tier: invariant
fingerprinted |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
fshamshirdar/pytorch-rdpg/rdpg.py 958765de0feeec56 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
dan-lennox/ml-udacity-quadcopter-rl/agents/agent.py 8bd7d62053fbfae5 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
halajadallah/RL-Quadcopter_project/agents/agent_modified.py 75b2da574221eee2 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
InSpaceAI/RL-Zoo/DDPG.py 8afc63f093262b83 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
Gouet/DDPG_PendulumV1/ddpg.py 4abb7a0681b9f8e3 |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
madvn/DDPG/ddpg/ddpg.py b9b27fa7acc7c73f |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
shahabi8/Deep-Reinforcement-Learning/DDGP.py ef0c12a6692597db |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
alathiya/RL-Quadcoptor-Flying/agents/agent.py f6231123b603e20f |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
dyth/doublegum/policies_cont/agents/DDPG.py c5cbbbbfdd3a367c |
unverified |
BSD-3-Clause (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
marload/DeepRL-TensorFlow2/DDPG/DDPG_Continuous.py 8f5f60d9f1d5e30f |
unverified |
Apache-2.0 (permissive) |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
abbadka/quadcopter/agents/ddpg/agent.py 017ef34f0499756e |
unverified |
no licence file found · pointer only |
| Continuous control with deep reinforcement learning |
9 Sep 2015 |
fhbzc/FishAgentSimulation/train_RL.py 0d51c1e5212918ec |
unverified |
no licence file found · pointer only |
| arXiv:ijcai2025_0648 |
|
BW297/FACD/model/FACD.py 698d2cb357caa103 |
ran · metamorphic tier: deterministic
|
MIT (permissive) |