| Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning added by Syntology |
2026-07 (from id) |
aixiuxiuxiu/Legal-MT-SFT-RL/dist.py 968bb0ff4e9f3f20 |
unverified |
MIT (permissive) |
| Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus added by Syntology |
2026-04 (from id) |
PKU-MARL/Multi-Agent-Transformer/mat/algorithms/mat/algorithm/ma_transformer.py c00885ffa5fe61e7 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Optimizing Neurorobot Policy under Limited Demonstration Data through Preference Regret added by Syntology |
2026-04 (from id) |
NACLab/neurorobot-preference-regret-learning/lib/agent/active.py 2946157301d5cc74 |
unverified |
MIT (permissive) |
| CoordLight: Learning Decentralized Coordination for Network-Wide Traffic Signal Control added by Syntology |
2026-03 (from id) |
marmotlab/CoordLight/Models/CoordLightModel.py c31c51db289b4c16 |
ran · our draft was wrong
|
MIT (permissive) |
| Unifying Agent Interaction and World Information for Multi-agent Coordination added by Syntology |
2025-09 (from id) |
zoeyuchao/mappo/onpolicy/algorithms/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Reparameterization Proximal Policy Optimization added by Syntology |
2025-08 (from id) |
SonSang/gippo/src/gippo/network.py d0cd749d322eccb6 |
unverified |
no licence file found · pointer only |
| TRACED: Transition-aware Regret Approximation with Co-learnability for Environment Design |
24 Jun 2025 |
Cho-Geonwoo/TRACED/models/recurrent_walker_models.py 9de355e93051e4ad |
ran · our draft was wrong
|
no licence file found · pointer only |
| TRACED: Transition-aware Regret Approximation with Co-learnability for Environment Design |
24 Jun 2025 |
Cho-Geonwoo/TRACED/models/common.py f8db408bc65dbb3a |
unverified |
licence not identified · pointer only |
| Continual Task Learning through Adaptive Policy Self-Composition |
18 Nov 2024 |
charleshsc/CompoFormer/dt/decision_transformer_grow.py c00885ffa5fe61e7 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Enabling Adaptive Agent Training in Open-Ended Simulators by Targeting Diversity |
7 Nov 2024 |
robbycostales/diva/diva/components/policy/networks.py 3e211798265a006a |
unverified |
MIT (permissive) |
| Vertical Federated Learning with Missing Features During Training and Inference |
29 Oct 2024 |
Valdeira/LASER-VFL/models/mimic_model_utils.py d4531a69fbf2a306 |
unverified |
MIT (permissive) |
| All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path Aggregation |
29 Sep 2024 |
njuvision/mpa/examples/train_stage1_wo_gan.py 6656404be8df535e |
unverified |
licence not identified · pointer only |
| All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path Aggregation |
29 Sep 2024 |
njuvision/mpa/examples/eval_cls_real_bpp.py ffb7e6e59725022b |
unverified |
licence not identified · pointer only |
| GOPT: Generalizable Online 3D Bin Packing via Transformer-based Deep Reinforcement Learning |
2024-09 (from id) |
xiong5heng/gopt/model.py 9de355e93051e4ad |
ran · our draft was wrong
|
no licence file found · pointer only |
| M$^3$GPT: An Advanced Multimodal, Multitask Framework for Motion Comprehension and Generation |
25 May 2024 |
luomingshuang/M3GPT/m3gpt/core/fp16/amp.py eb8a2600afc1de95 |
ran
|
no licence file found · pointer only |
| Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations |
26 Apr 2024 |
Xiaoyao-Li/Ag2Manip/algos/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
no licence file found · pointer only |
| UniTraj: A Unified Framework for Scalable Vehicle Trajectory Prediction |
22 Mar 2024 |
vita-epfl/UniTraj/unitraj/models/autobot/autobot.py d7a51202f8456490 |
ran
|
licence not identified · pointer only |
| Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing |
3 Feb 2024 |
identical code first harvested elsewhere 9de355e93051e4ad |
ran · our draft was wrong
|
licence of this copy not recorded |
| Gradient Informed Proximal Policy Optimization |
14 Dec 2023 |
sonsang/gippo/src/gippo/network.py d0cd749d322eccb6 |
unverified |
no licence file found · pointer only |
| Learning Curricula in Open-Ended Worlds |
3 Dec 2023 |
facebookresearch/dcd/models/common.py 9de355e93051e4ad |
ran · our draft was wrong
|
no licence file found · pointer only |
| Hulk: A Universal Knowledge Translator for Human-Centric Tasks |
4 Dec 2023 |
opengvlab/hulk/core/fp16/amp.py eb8a2600afc1de95 |
ran
|
MIT (permissive) |
| Everybody Needs a Little HELP: Explaining Graphs via Hierarchical Concepts |
25 Nov 2023 |
jonasjuerss/help/custom_logger.py 5265ccd1e6a4830b |
ran
|
no licence file found · pointer only |
| Differentiable and accelerated spherical harmonic and Wigner transforms |
24 Nov 2023 |
astro-informatics/s2fft/s2fft/recursions/trapani.py 94b976a8c3413231 |
ran
|
MIT (permissive) |
| Beyond Surprise: Improving Exploration Through Surprise Novelty |
9 Aug 2023 |
thaihungle/sm/dist.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Recursive Algorithmic Reasoning |
1 Jul 2023 |
DJayalath/gnn-call-stack/gnn_call_stack/utils.py ec6c1cbb5fa6ffa6 |
unverified |
Apache-2.0 (permissive) |
| POUF: Prompt-oriented unsupervised fine-tuning for large pre-trained models |
29 Apr 2023 |
korawat-tanwisuth/pouf/pouf_mlm/src/label_search.py 5fabcb6bbd28aa2f |
unverified |
no licence file found · pointer only |
| UniHCP: A Unified Model for Human-Centric Perceptions |
6 Mar 2023 |
OpenGVLab/UniHCP/core/fp16/amp.py eb8a2600afc1de95 |
ran
|
MIT (permissive) |
| Order Matters: Agent-by-agent Policy Optimization |
13 Feb 2023 |
xihuai18/A2PO-ICLR2023/onpolicy/algorithms/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Wild-Time: A Benchmark of in-the-Wild Distribution Shift over Time |
25 Nov 2022 |
huaxiuyao/wild-time/wildtime/baseline_trainer.py beca3cd43469c27a |
unverified |
MIT (permissive) |
| Distributionally Adaptive Meta Reinforcement Learning |
6 Oct 2022 |
ikostrikov/pytorch-a2c-ppo-acktr-gail/a2c_ppo_acktr/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| DGPO: Discovering Multiple Strategies with Diversity-Guided Policy Optimization |
12 Jul 2022 |
OpenRL-Lab/DGPO/onpolicy/algorithms/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Robust Imitation Learning against Variations in Environment Dynamics |
19 Jun 2022 |
JongseongChae/RIME/algorithm/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| An Evaluation Study of Intrinsic Motivation Techniques applied to Reinforcement Learning over Hard Exploration Environments |
23 May 2022 |
aklein1995/intrinsic_motivation_techniques_study/model.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| An Extensive Data Processing Pipeline for MIMIC-IV |
29 Apr 2022 |
healthylaife/mimic-iv-data-pipeline/model/model_utils.py 7cedea7a54f72fb3 |
unverified |
MIT (permissive) |
| Unknown-Aware Object Detection: Learning What You Don't Know from Videos in the Wild |
8 Mar 2022 |
deeplearning-wisc/stud/datasets/bdd100k2coco.py 5113c7c282ca0da7 |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning |
23 Sep 2021 |
mehdinasiri/mirror-descent-in-marl/algorithms/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Explore and Control with Adversarial Surprise |
12 Jul 2021 |
ArnaudFickinger/adversarial-surprise/model.py 9de355e93051e4ad |
ran · our draft was wrong
|
no licence file found · pointer only |
| Policy Transfer across Visual and Dynamics Domain Gaps via Iterative Grounding |
1 Jul 2021 |
clvrai/idapt/training/networks/distributions.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Policy Regularization via Noisy Advantage Values for Cooperative Multi-agent Actor-Critic methods |
2021-06 (from id) |
hijkzzz/noisy-mappo/onpolicy/algorithms/utils/util.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Real-time Adversarial Perturbations against Deep Reinforcement Learning Policies: Attacks and Defenses |
16 Jun 2021 |
ssg-research/ad3-action-distribution-divergence-detector/src/agents/models.py 9de355e93051e4ad |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Curiosity-Driven Exploration via Latent Bayesian Surprise |
15 Apr 2021 |
mazpie/lbs-exploration/exploration/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Character Controllers Using Motion VAEs |
26 Mar 2021 |
electronicarts/character-motion-vaes/common/controller.py 9de355e93051e4ad |
ran · our draft was wrong
|
BSD-3-Clause (permissive) |
| Latent Variable Sequential Set Transformers For Joint Multi-Agent Motion Prediction |
19 Feb 2021 |
roggirg/AutoBots/models/autobot_joint.py 9de355e93051e4ad |
ran · our draft was wrong
|
BSD-3-Clause (permissive) |
| Latent Variable Sequential Set Transformers For Joint Multi-Agent Motion Prediction |
19 Feb 2021 |
roggirg/AutoBots/models/autobot_ego.py d7a51202f8456490 |
ran
|
BSD-3-Clause (permissive) |
| "LazImpa": Lazy and Impatient neural agents learn to communicate efficiently |
5 Oct 2020 |
MathieuRita/Lazimpa/egg/core/util.py 3c2081da5221dfd6 |
unverified |
MIT (permissive) |
| Learning Dexterous Grasping with Object-Centric Visual Affordances |
3 Sep 2020 |
priyankamandikal/graff/a2c_ppo_acktr/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Learning to plan with uncertain topological maps |
10 Jul 2020 |
edbeeching/learning_to_plan/layers.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Automatic Data Augmentation for Generalization in Deep Reinforcement Learning |
23 Jun 2020 |
rraileanu/auto-drac/ucb_rl2_meta/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Agent Modelling under Partial Observability for Deep Reinforcement Learning |
16 Jun 2020 |
uoe-agents/LIAM/double_speaker_listener/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Characterizing Attacks on Deep Reinforcement Learning |
21 Jul 2019 |
ssg-research/flare/src/agents/models.py 9de355e93051e4ad |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Sequential estimation of quantiles with applications to A/B-testing and best-arm identification |
24 Jun 2019 |
WLM1ke/poptimizer/poptimizer/adapters/logger.py eff60c9fec6ef025 |
unverified |
Unlicense (permissive) |
| Reinforcement Learning with Convex Constraints |
21 Jun 2019 |
xkianteb/ApproPO/ApproPO/nets.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Mega-Reward: Achieving Human-Level Play without Extrinsic Rewards |
12 May 2019 |
YuhangSong/Mega-Reward/a2c_ppo_acktr/utils.py 9de355e93051e4ad |
ran · our draft was wrong
|
MIT (permissive) |
| Learning when to Communicate at Scale in Multiagent Cooperative and Competitive Tasks |
23 Dec 2018 |
IC3Net/IC3Net/data.py 8cab58f1257e1400 |
unverified |
MIT (permissive) |
| Parallelizing Linear Recurrent Neural Nets Over Sequence Length |
12 Sep 2017 |
proger/accelerated-scan/tests/bench.py 6ac2ecfb8cd3b4a0 |
ran · honoured contract
|
MIT (permissive) |
| Curiosity-driven Exploration by Self-supervised Prediction |
15 May 2017 |
rpatrik96/AttA2C/src/model.py 87da60c05cf5addb |
ran · our draft was wrong
|
MIT (permissive) |
| MultiNet: Real-time Joint Semantic Reasoning for Autonomous Driving |
22 Dec 2016 |
MarvinTeichmann/KittiBox/submodules/utils/googlenet_load.py 0380022ecb697fdf |
unverified |
MIT (permissive) |
| Medical image denoising using convolutional denoising autoencoders |
16 Aug 2016 |
adam-mah/Medical-Image-Denoising/bm3d.py d86875c444f7728a |
unverified |
MIT (permissive) |