Home › Code › discount_cumsum

discount_cumsum

Syntologyentry name in harvested coderead from the graph 2026-09-24

discount_cumsum appears in the code Syntology harvested for 35 papers, as 19 distinct code bodies found in 39 places (a place is one code body under one paper). At least one of them ran in 22 of the papers; 6 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named discount_cumsum do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 7 of the 19 distinct code bodies named discount_cumsum; 12 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

0ran · honoured contract
0ran · violated contract
3ran · our draft was wrong
3ran · fixture could not drive it
1ran
12unverified
6fingerprinted

Licence is a property of each copy, so it is counted per place: 12 of the 39 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

35 papers shown of 35, newest first; 39 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 2 papers added by Syntology; 1 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
Safe Continual Reinforcement Learning in Non-stationary Environments added by Syntology 2026-04 (from id) MACS-Research-Lab/safe-crl/code/Safe-Policy-Optimization/safepo/common/buffer.py 185c4bf8b0d68d6c unverified no licence file found · pointer only
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning added by Syntology 2026-03 (from id) dilina-r/trex_xmorl/trex_cluster.py 2dc44fc961aebd9a unverified no licence file found · pointer only
Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers 31 Oct 2024 kaiyan289/rl_as_vitamin_for_online_decision_transformers/data.py f434bf3e9ec07356 unverified no licence file found · pointer only
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks 29 Oct 2024 ml-jku/lram/src/buffers/buffer_utils.py 47f553906dc6d889 unverified MIT (permissive)
Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement 15 Oct 2024 NJU-RL/Meta-DT/meta_dt/dataset.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted no licence file found · pointer only
Retrieval-Augmented Decision Transformer: External Memory for In-context RL 9 Oct 2024 ml-jku/RA-DT/src/buffers/buffer_utils.py 47f553906dc6d889 unverified MIT (permissive)
Decision Mamba: A Multi-Grained State Space Model with Self-Evolution Regularization for Offline RL 8 Jun 2024 aopolin-lv/DecisionMamba/experiment-d4rl/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted Apache-2.0 (permissive)
In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought 31 May 2024 identical code first harvested elsewhere 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted licence of this copy not recorded
Q-value Regularized Transformer for Offline Reinforcement Learning 27 May 2024 charleshsc/qt/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted Apache-2.0 (permissive)
Is Mamba Compatible with Trajectory Optimization in Offline Reinforcement Learning? 20 May 2024 AndssY/DeMa/gym/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Reinforced Sequential Decision-Making for Sepsis Treatment: The POSNEGDM Framework with Mortality Classifier and Transformer 12 Mar 2024 dipeshtamboli/posnegdm-reinforced-sequential-decision-making-for-sepsis-treatment/utils.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning 8 Feb 2024 woooodyy/llm-reverse-curriculum-rl/R3_math/src/utils.py 7ffa253a733f9bfc ran fingerprinted no licence file found · pointer only
Critic-Guided Decision Transformer for Offline Reinforcement Learning 21 Dec 2023 sharkwyf/cgdt/data.py f434bf3e9ec07356 unverified MIT (permissive)
Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning 31 Oct 2023 srzer/LaMo-2023/experiment-d4rl/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Learning to Modulate pre-trained Models in RL 26 Jun 2023 ml-jku/l2m/src/buffers/buffer_utils.py 47f553906dc6d889 unverified MIT (permissive)
Future-conditioned Unsupervised Pretraining for Decision Transformer 26 May 2023 fffffarmer/pdt/src/data.py 7ae6e29065b3cc4b unverified MIT (permissive)
Beyond Reward: Offline Preference-guided Policy Optimization 25 May 2023 bkkgbkjb/oppo/oppo/human/gym/experiment.py 7a477d732661f79a ran · our draft was wrong fingerprinted no licence file found · pointer only
Beyond Reward: Offline Preference-guided Policy Optimization 25 May 2023 bkkgbkjb/oppo/oppo/scripted/gym/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted no licence file found · pointer only
When should we prefer Decision Transformers for Offline Reinforcement Learning? 23 May 2023 prajjwal1/rl_paradigm/exorl/dataset.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Revisiting the Minimalist Approach to Offline Reinforcement Learning 16 May 2023 adamjelley/efficientofflinerl/algorithms/cql.py d0c4a0aa6c8b23c1 unverified Apache-2.0 (permissive)
Revisiting the Minimalist Approach to Offline Reinforcement Learning 16 May 2023 adamjelley/efficientofflinerl/algorithms/edac.py 04ad4ada1fd13ab2 unverified Apache-2.0 (permissive)
Revisiting the Minimalist Approach to Offline Reinforcement Learning 16 May 2023 adamjelley/efficientofflinerl/algorithms/sac_n.py e258d5be5c961c7a unverified Apache-2.0 (permissive)
Merging Decision Transformers: Weight Averaging for Forming Multi-Task Policies 14 Mar 2023 daniellawson9999/merging-decision-transformers/decision-transformer/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
How Crucial is Transformer in Decision Transformer? 26 Nov 2022 identical code first harvested elsewhere 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted licence of this copy not recorded
On the Effect of Pre-training for Transformer in Different Modality on Offline Reinforcement Learning 17 Nov 2022 machelreid/can-wikipedia-help-offline-rl/code/eval_model.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
On the Effect of Pre-training for Transformer in Different Modality on Offline Reinforcement Learning 17 Nov 2022 t46/pre-training-different-modality-offline-rl/can-wikipedia-help-offline-rl/experiment.py 0d4fc87c387d9018 ran · fixture could not drive it fingerprinted MIT (permissive)
Towards Understanding How Machines Can Learn Causal Overhypotheses 16 Jun 2022 cannylab/casual_overhypotheses/models/decision-transformer/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
When does return-conditioned supervised learning work for offline reinforcement learning? 2 Jun 2022 davidbrandfonbrener/rcsl-paper/decision-transformer/gym/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Efficient Reward Poisoning Attacks on Online Deep Reinforcement Learning 30 May 2022 yinglunxu/reward_poisoning_attack_drl/src/all_class.py f047da47c4530c5f ran · our draft was wrong fingerprinted no licence file found · pointer only
Can Wikipedia Help Offline Reinforcement Learning? 28 Jan 2022 identical code first harvested elsewhere 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted licence of this copy not recorded
Generalized Decision Transformer for Offline Hindsight Information Matching 19 Nov 2021 identical code first harvested elsewhere 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted licence of this copy not recorded
StARformer: Transformer with State-Action-Reward Representations for Visual Reinforcement Learning 12 Oct 2021 elicassion/StARformer/gym/experiment.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Decision Transformer: Reinforcement Learning via Sequence Modeling 2 Jun 2021 HzcIrving/DecisionTransformer_StepbyStep/utils.py 0161b27cbe3cc58d ran · fixture could not drive it fingerprinted MIT (permissive)
Enhancing SAT solvers with glue variable predictions 6 Jul 2020 jesse-michael-han/neuro-cadical/python/rl_loop.py 658d370f8dacf5e3 ran · our draft was wrong fingerprinted Apache-2.0 (permissive)
A Closer Look at Invalid Action Masking in Policy Gradient Algorithms 25 Jun 2020 vwxyzjn/invalid-action-masking/invalid_action_masking/ppo_10x10.py 43887c9edd549830 ran · fixture could not drive it MIT (permissive)
Reward-Conditioned Policies 31 Dec 2019 TrentBrick/RewardConditionedUDRL/control/agent.py d90f855adb55e3ce unverified MIT (permissive)
Importance Sampling Policy Evaluation with an Estimated Behavior Policy 4 Jun 2018 LARG/regression-importance-sampling/roboschool-experiments/common.py 128b553ce11b9774 unverified MIT (permissive)
Meta-Gradient Reinforcement Learning 24 May 2018 RobvanGastel/meta-rl-algorithms/algos/mg_a2c/buffer.py 85f38abd91775356 unverified MIT (permissive)
arXiv:aaai_35251 Teddy298/continualworld-ppo/continualworld/ppo/core.py 13141eda72cfce23 unverified MIT (permissive)

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections