Home › Code › count_tokens

count_tokens

Syntologyentry name in harvested coderead from the graph 2026-09-24

count_tokens appears in the code Syntology harvested for 39 papers, as 39 distinct code bodies found in 41 places (a place is one code body under one paper). At least one of them ran in 11 of the papers; 4 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named count_tokens do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 10 of the 39 distinct code bodies named count_tokens; 29 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

2ran · honoured contract
0ran · violated contract
0ran · our draft was wrong
0ran · fixture could not drive it
8ran
29unverified
4fingerprinted

Licence is a property of each copy, so it is counted per place: 10 of the 41 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

39 papers shown of 39, newest first; 41 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 10 papers added by Syntology; 3 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
TOKEN-LEVEL ELASTIC-DEPTH LOOPED TRANSFORMERS FOR LATENT REASONING WITH DYNAMIC ROUTING UNDER REVIEW AT ICLR 2027 added by Syntology 2026-09 (from id) YuMingQian1234/T-LoopFormer/utils.py 8cc31a6638b3633f unverified MIT (permissive)
RealRoute: Dynamic Query Routing System via Retrieve-then-Verify Paradigm added by Syntology 2026-04 (from id) Joseph1951210/RealRoute/pipeline/subquery_executor.py ab165e30f271182c unverified no licence file found · pointer only
MANGO: Multi-Agent Web Navigation via Global-View Optimization added by Syntology 2026-04 (from id) VichyTong/Mango/llm_web_scraper/utils/llm.py 04d9feb321c92fde unverified Apache-2.0 (permissive)
Adaptive Chunking: Optimizing Chunking-Method Selection for RAG added by Syntology 2026-03 (from id) ekimetrics/adaptive-chunking/src/adaptive_chunking/splitters.py 936a7e0cee9f7a1c unverified MIT (permissive)
Rethinking the Reranker: Boundary-Aware Evidence Selection for Robust Retrieval-Augmented Generation added by Syntology 2026-02 (from id) GasolSun36/BAR-RAG/filter.py 0335852d729ef632 ran · honoured contract no licence file found · pointer only
Yunque DeepResearch Technical Report added by Syntology 2026-01 (from id) Tencent-BAC/YunqueAgent/inference/base_tool.py 04c16d8f22c5eaad unverified Apache-2.0 (permissive)
Steering Language Models Before They Speak: Logit-Level Interventions added by Syntology 2026-01 (from id) hsannn/swai/build_scores.py c1b0786d78cff5bb unverified Apache-2.0 (permissive)
CodeRAG: Finding Relevant and Necessary Knowledge for Retrieval-Augmented Repository-Level Code Completion added by Syntology 2025-09 (from id) KDEGroup/CodeRAG/coderag/build_prompt/merge_retrieval.py c1a55b16a582d56f unverified MIT (permissive)
Building Data-Driven Occupation Taxonomies: A Bottom-Up Multi-Stage Approach via Semantic Clustering and Multi-Agent Collaboration added by Syntology 2025-09 (from id) aida-ugent/CLIMB/src/get_prompts.py 98bfdf443dbd6f66 unverified licence not identified · pointer only
TableDART: Dynamic Adaptive Multi-Modal Routing for Table Understanding added by Syntology 2025-09 (from id) xiaobo-xing/TableDART/cost_measurement/measure_expert_costs.py 2296fe2308655cb0 unverified MIT (permissive)
WebSailor: Navigating Super-human Reasoning for Web Agent 3 Jul 2025 alibaba-nlp/webagent/WebAgent/NestBrowse/utils.py 23aacba017665d93 unverified Apache-2.0 (permissive)
Learning Adaptive Parallel Reasoning with Language Models 21 Apr 2025 parallel-reasoning/apr/src/eval/eval_sosp.py 56c7087a3cd9e36e unverified Apache-2.0 (permissive)
GraphOmni: A Comprehensive and Extendable Benchmark Framework for Large Language Models on Graph-theoretic Tasks 17 Apr 2025 gai-community/graphomni/eval_fun/token.py c4fa9aad367464fa unverified MIT (permissive)
Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoning 11 Apr 2025 chang-github-00/llm-predictive-decoding/agentboard/algorithms/mpc_sampling.py 94846e60d6f6c92a ran no licence file found · pointer only
AgentSociety Challenge: Designing LLM Agents for User Modeling and Recommendation on Web Platforms 26 Feb 2025 tsinghua-fib-lab/agentsocietychallenge/GTsimulation/ModGTAgent.py 766aece7c07a8cae unverified MIT (permissive)
AgentBreeder: Mitigating the AI Safety Impact of Multi-Agent Scaffolds via Self-Improvement 2 Feb 2025 J-Rosser-UK/AgentBreeder/src/api/anthropic_api.py 1d4a9efbfddea497 unverified MIT (permissive)
Agent Laboratory: Using LLM Agents as Research Assistants 8 Jan 2025 Masao-Taketani/LocalAgentLaboratory/utils.py 023b788084b451b0 unverified MIT (permissive)
HARP: A challenging human-annotated math reasoning benchmark 11 Dec 2024 aadityasingh/harp/src/eval/costs.py da52349b06d44177 unverified MIT (permissive)
SetLexSem Challenge: Using Set Operations to Evaluate the Lexical and Semantic Robustness of Language Models 11 Nov 2024 amazon-science/setlexsem-challenge/setlexsem/experiment/lmapi.py 912d808705da066a unverified Apache-2.0 (permissive)
The Unreasonable Effectiveness of LLMs for Query Optimization 5 Nov 2024 peter-ai/LLMSteer/models/utils.py 936c5d54cec2c170 unverified no licence file found · pointer only
Paths-over-Graph: Knowledge Graph Empowered Large Language Model Reasoning 18 Oct 2024 SteveTANTAN/PoG/PoG/utils.py dcb9359b01d13cbe unverified Apache-2.0 (permissive)
LLM-PBE: Assessing Data Privacy in Large Language Models 23 Aug 2024 QinbinLi/LLM-PBE/models/togetherai.py a5fbf1725fd11160 ran MIT (permissive)
HiAgent: Hierarchical Working Memory Management for Solving Long-Horizon Agent Tasks with Large Language Model 18 Aug 2024 hiagent2024/hiagent/agentboard/agents/ours_agent.py 1e64761aa30b0c32 ran · honoured contract fingerprinted no licence file found · pointer only
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 xinyiliu0227/llm_debate_bias/ddo_baseline_binary_1-1.py 1e64761aa30b0c32 ran · honoured contract fingerprinted no licence file found · pointer only
Language Models Learn Rare Phenomena from Less Rare Phenomena: The Case of the Missing AANNs 28 Mar 2024 kanishkamisra/aannalysis/src/counterfactual_constructions.py a3a3f6ce577d2d24 ran MIT (permissive)
CoverUp: Effective High Coverage Test Generation for Python 24 Mar 2024 plasma-umass/coverup/src/coverup/llm.py 145acc26153dc1f4 ran Apache-2.0 (permissive)
RAmBLA: A Framework for Evaluating the Reliability of LLMs as Assistants in the Biomedical Domain 21 Mar 2024 gsk-ai/rambla/rambla/models/utils.py a953ecb4b9b05c6a unverified Apache-2.0 (permissive)
Debating with More Persuasive LLMs Leads to More Truthful Answers 9 Feb 2024 ucl-dark/llm_debate/core/llm_api/anthropic_llm.py 56346bcde59d7a4a ran fingerprinted MIT (permissive)
Debating with More Persuasive LLMs Leads to More Truthful Answers 9 Feb 2024 ucl-dark/llm_debate/core/llm_api/openai_llm.py 3b94d7f86b74da2b unverified MIT (permissive)
Enhancing textual textbook question answering with large language models and retrieval augmented generation 5 Feb 2024 hessaalawwad/plr-tqa/RAG_Pinecone.py 529e6478e94acc7a unverified no licence file found · pointer only
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models 17 Dec 2023 tml2002/rolecraft/code/prompt3.py 584eb7879ebbd0c1 ran fingerprinted no licence file found · pointer only
DocMath-Eval: Evaluating Math Reasoning Capabilities of LLMs in Understanding Long and Specialized Documents 16 Nov 2023 yale-nlp/docmath-eval/utils/model_input_utils.py cc3cc5ecc9833707 ran no licence file found · pointer only
CodeScope: An Execution-based Multilingual Multitask Multidimensional Benchmark for Evaluating LLMs on Code Understanding and Generation 14 Nov 2023 weixiangyan/codescope/automated_testing/evaluator/score.py 20f9439429346697 unverified MIT (permissive)
Chain of Natural Language Inference for Reducing Large Language Model Ungrounded Hallucinations 6 Oct 2023 microsoft/conli_hallucination/CoNLI/CoNLI/modules/hallucination_detector.py 89115682f5989ce7 ran fingerprinted MIT (permissive)
ByteSized32: A Corpus and Challenge Task for Generating Task-Specific World Models Expressed as Text Games 24 May 2023 cognitiveailab/BYTESIZED32/bytes32/utils.py 9e4c4702c63966aa unverified Apache-2.0 (permissive)
GanLM: Encoder-Decoder Pre-training with an Auxiliary Discriminator 20 Dec 2022 csjianyang/ganlm/evaluation/cnn_dm.py 7d0f0d30d5c3027f unverified MIT (permissive)
Improved training of end-to-end attention models for speech recognition 8 May 2018 bobchennan/espnet/espnet/lm/lm_utils.py ff899e7de6121891 unverified Apache-2.0 (permissive)
Improved training of end-to-end attention models for speech recognition 8 May 2018 creatorscan/espnet/espnet/lm/lm_utils.py e2309fb95e3f748c unverified Apache-2.0 (permissive)
arXiv:aaai_21432 microsoft/DialogLM/DialogLM_UniLM/DialogLM/evaluations/eval_for_cnndm.py 7d0f0d30d5c3027f unverified MIT (permissive)
arXiv:2025.findings-emnlp.479 FoundationAgents/SPO/utils/evaluation_utils.py 3c48f41137637146 unverified MIT (permissive)
arXiv:2024.findings-emnlp.123 HKUST-KnowComp/IntentionQA/filter/simplifyName.py f54f0325dfa091c7 unverified MIT (permissive)

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections