Home › Code › calculate_metrics

calculate_metrics

Syntologyentry name in harvested coderead from the graph 2026-09-24

calculate_metrics appears in the code Syntology harvested for 61 papers, as 62 distinct code bodies found in 65 places (a place is one code body under one paper). At least one of them ran in 29 of the papers; 5 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named calculate_metrics do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 26 of the 62 distinct code bodies named calculate_metrics; 36 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

1ran · honoured contract
0ran · violated contract
3ran · our draft was wrong
8ran · fixture could not drive it
14ran
36unverified
5fingerprinted

Licence is a property of each copy, so it is counted per place: 32 of the 65 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

61 papers shown of 61, newest first; 65 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 18 papers added by Syntology; 4 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
Enhancing deep learning models for time series classification via knowledge distillation added by Syntology 2026-07 (from id) MSD-IRIMAS/KD-4-TSC/KD4TSC/utils.py 802e2c8e966af2b1 ran no licence file found · pointer only
Sycophancy as Material Failure under Pushback Loading A Multi-Axis Characterization Across Three Loading Cases and up to Seventeen Material Charges added by Syntology 2026-06 (from id) JiseungHong/SYCON-Bench/debate_setting/evaluate_base_ToF.py 0a10e54cb945833c ran MIT (permissive)
GRASP: Gradient-Aligned Sequential Parameter Transfer for Memory-Efficient Multi-Source Learning added by Syntology 2026-06 (from id) Sekeh-Lab/grasp-multisource-transfer/experiments/grasp/run_grasp_experiment.py e950bb9bbf851272 ran · fixture could not drive it fingerprinted MIT (permissive)
MINTEVAL: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems added by Syntology 2026-05 (from id) amy-hyunji/MINTEval/src/mem_alpha/memalpha/llm_agent/metrics.py c9e9a810f3e053fb ran · our draft was wrong no licence file found · pointer only
Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Models added by Syntology 2026-04 (from id) Pepper66/DMLE/cal_avg.py fcc0593f8dc28f9d unverified no licence file found · pointer only
Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs added by Syntology 2026-04 (from id) mkrzywda/gnn-misinformation-tradeoffs/GNN-modified.py 0d211a14cf22e405 unverified GPL-3.0 (copyleft) · pointer only
Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs added by Syntology 2026-04 (from id) mkrzywda/gnn-misinformation-tradeoffs/baselines.py 0fd5d745f5999db6 unverified GPL-3.0 (copyleft) · pointer only
Prompts Without Evidence: How Neuroimaging Mentions Shift Clinical Vision-Language Model Predictions added by Syntology 2026-03 (from id) long21wt/scaffold-effect/src/f1_eval.py f068fcc1e4587d2f unverified no licence file found · pointer only
Prompts Without Evidence: How Neuroimaging Mentions Shift Clinical Vision-Language Model Predictions added by Syntology 2026-03 (from id) long21wt/scaffold-effect/src/f1_eval_oasis.py a035fc9a992b738f unverified no licence file found · pointer only
More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection added by Syntology 2026-03 (from id) Sayur1n/H-VLI/evaluator.py cf1eb64ff5a7a55a unverified licence not identified · pointer only
OpenLID-v3: Improving the Precision of Closely Related Language Identification -An Experience Report added by Syntology 2026-02 (from id) ltgoslo/slide/src/evaluate.py 8651b2739b0c4b96 unverified Apache-2.0 (permissive)
REVIS: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models added by Syntology 2026-02 (from id) antgroup/Revis/utils/mmvet_judge_only.py 64c97831bbe858d3 unverified Apache-2.0 (permissive)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving added by Syntology 2026-02 (from id) ys-qu/found-rl/clean_wandb_data.py cad6e72086af2eb4 unverified licence not identified · pointer only
GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems added by Syntology 2026-02 (from id) kat-er-ina/gemss_testing/algorithm_comparison/src/evaluation.py 5e372b1774b9c227 unverified no licence file found · pointer only
Video-based Music Generation added by Syntology 2026-02 (from id) serkansulun/trailer-genre-classification/classification/src/utils_classify.py 4d37c5189d2765e3 unverified licence not identified · pointer only
Benchmarking Bias Mitigation Toward Fairness Without Harm from Vision to LVLMs added by Syntology 2026-02 (from id) osu-srml/NH-Fair/src/release_benchmark/methods/vlm/clip_fairer.py b3ee986ec34da06e ran · fixture could not drive it MIT (permissive)
DCD: Decomposition-based Causal Discovery from Autocorrelated and Non-Stationary Temporal Data added by Syntology 2026-02 (from id) noname2122/DCD-TMLR-2025/experiments/run_dynotears_baseline.py 9500b6a629b1e792 unverified MIT (permissive)
R 2 BD: A Reconstruction-Based Method for Generalizable and Efficient Detection of Fake Images added by Syntology 2026-01 (from id) QingyuLiu/RRBD/utils.py d0600278ff835cb1 unverified no licence file found · pointer only
ICLR 2026 Workshop: Principled Design for Trustworthy AI REDBENCH: A UNIVERSAL DATASET FOR COMPRE-HENSIVE RED TEAMING OF LARGE LANGUAGE MOD-ELS added by Syntology 2026-01 (from id) knoveleng/redeval/redeval/score.py 036e0114f72d9a8c unverified MIT (permissive)
ImageSentinel: Protecting Visual Datasets from Unauthorized Retrieval-Augmented Image Generation added by Syntology 2025-10 (from id) luo-ziyuan/ImageSentinel/evaluate_similarities.py c18eab3d2bf63ca5 unverified no licence file found · pointer only
SimpleDoc: Multi-Modal Document Understanding with Dual-Cue Page Retrieval and Iterative Refinement 16 Jun 2025 ag2ai/SimpleDoc/evaluation/analyze_multiple_runs.py 36d783e0916c23cb unverified no licence file found · pointer only
On Efficient Estimation of Distributional Treatment Effects under Covariate-Adaptive Randomization 6 Jun 2025 CyberAgentAILab/dte_car/simulation.py 534c7b1b790218d9 ran · fixture could not drive it MIT (permissive)
arXiv:2505.20840 2025-05 (from id) dooho00/agg-buffer/evaluate/metrics.py b7d98c7e91eb3959 unverified Apache-2.0 (permissive)
$\text{R}^2\text{ec}$: Towards Large Recommender Models with Reasoning 22 May 2025 YRYangang/RRec/trainers/utils.py 0bc8c7e05959c420 unverified MIT (permissive)
Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control 18 Mar 2025 nv-tlabs/cosmos-drive-dreams/cosmos-drive-dreams-toolkits/train_light_model.py 55d080a2eba4d6b7 unverified Apache-2.0 (permissive)
On Adversarial Robustness and Out-of-Distribution Robustness of Large Language Models 13 Dec 2024 jordantab/llm-robustness-experiment/utilities/calculate_pb_baseline.py 54807fe1fc74ab13 unverified no licence file found · pointer only
On Adversarial Robustness and Out-of-Distribution Robustness of Large Language Models 13 Dec 2024 jordantab/llm-robustness-experiment/utilities/parse.py c8436ea20f798b9a unverified no licence file found · pointer only
Variational Low-Rank Adaptation Using IVON 7 Nov 2024 team-approx-bayes/ivon-lora/utils.py 436f4c309d8668cf unverified no licence file found · pointer only
Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only 14 Oct 2024 yaojh18/Varying-Shades-of-Wrong/preference_optimization/evaluate.py 29adfabf77b2b461 ran · honoured contract no licence file found · pointer only
GraphCroc: Cross-Correlation Autoencoder for Graph Structural Reconstruction 4 Oct 2024 sjduan/graphcroc/IMDB_B/reconstructor.py 033a97925e132d68 unverified MIT (permissive)
SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information 21 Sep 2024 GasolSun36/SURf/eval/eval_pope.py 31fd3eb13a63be04 ran no licence file found · pointer only
Evaluating Fine-Tuning Efficiency of Human-Inspired Learning Strategies in Medical Question Answering 15 Aug 2024 Oxford-AI-for-Society/human-learning-strategies/training/inference/inference.py 57123d949a952552 ran GPL-3.0 (copyleft) · pointer only
Review-driven Personalized Preference Reasoning with Large Language Models for Recommendation 12 Aug 2024 jieyong99/exp3rt/test_result_inspect.py d46e7f90d226d0e6 ran · fixture could not drive it fingerprinted no licence file found · pointer only
ToolBeHonest: A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models 28 Jun 2024 toolbehonest/toolbehonest/utils/calculate_metrics.py 391471f665ef5fbb ran MIT (permissive)
Multi-Behavior Generative Recommendation 27 May 2024 anananan116/MBGen/trainer/evaluation.py bdbd3b67520ae048 unverified Apache-2.0 (permissive)
Transcriptomics-guided Slide Representation Learning in Computational Pathology 19 May 2024 mahmoodlab/tangle/run_linear_probing.py dc22df19cd5e4ef3 ran licence not identified · pointer only
Vulnerability Detection with Code Language Models: How Far Are We? 27 Mar 2024 dlvuldet/primevul/os_expr/run_ft.py 7559f2b5dd244067 ran MIT (permissive)
Chain-of-Action: Faithful and Multimodal Question Answering through Large Language Models 26 Mar 2024 MAGICS-LAB/Chain-of-Actions/chain-of-search.py 483518008187ff8f ran fingerprinted Apache-2.0 (permissive)
An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4 5 Mar 2024 huihuichyan/unlimitedjudge/src/build_dataset.py acacb923208da24f ran no licence file found · pointer only
Reinforced In-Context Black-Box Optimization 27 Feb 2024 songlei00/ribbo/algorithms/utils.py cac93e27abd92689 ran no licence file found · pointer only
CTNeRF: Cross-Time Transformer for Dynamic Neural Radiance Field from Monocular Video 10 Jan 2024 xingy038/ctnerf/utils/evaluation.py d42c9ddd0cc88b91 ran no licence file found · pointer only
TopicGPT: A Prompt-based Topic Modeling Framework 2 Nov 2023 chtmp223/topicgpt/topicgpt_python/utils.py 0d8395537f4bf2da ran MIT (permissive)
When Reviewers Lock Horn: Finding Disagreement in Scientific Peer Reviews 28 Oct 2023 sandeep82945/contradiction-in-peer-review/src/training_scratch3.py 8379fdb8173a8157 ran Apache-2.0 (permissive)
Implicit meta-learning may lead language models to trust more reliable sources 23 Oct 2023 krasheninnikov/internalization/src/fewshot.py 27effc0e19a1a9df ran no licence file found · pointer only
BAMBOO: A Comprehensive Benchmark for Evaluating Long Text Modeling Capacities of Large Language Models 23 Sep 2023 rucaibox/bamboo/evaluate.py a8b0c9b324725da9 ran · fixture could not drive it no licence file found · pointer only
WBCAtt: A White Blood Cell Dataset Annotated with Detailed Morphological Attributes 23 Jun 2023 apple2373/wbcatt/submission/traineval.py 6bc4993019f86d1a ran · fixture could not drive it MIT (permissive)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning 12 Jun 2023 zeju1997/oft/oft-control/eval_canny.py a73565d61cbda080 unverified MIT (permissive)
Editing Common Sense in Transformers 24 May 2023 anshitag/memit_csk/base_finetune_experiments/fine_tune_gpt.py 41346355360705b1 unverified MIT (permissive)
Editing Common Sense in Transformers 24 May 2023 anshitag/memit_csk/repair_finetune_experiments/evaluate_affected_finetune_model.py cfa19d05d0870261 unverified MIT (permissive)
RHO ($ρ$): Reducing Hallucination in Open-domain Dialogues with Knowledge Grounding 3 Dec 2022 ziweiji/rho/kg-cruse_/codes/KGCruse/predict.py dac58534a9991421 ran · fixture could not drive it no licence file found · pointer only
Federated Causal Discovery From Interventions 7 Nov 2022 aminabyaneh/Federated_CL/federated/utils.py ec867acd8308a4f3 unverified MIT (permissive)
Palette: Image-to-Image Diffusion Models 10 Nov 2021 kylelo/roofdiffusion/data/util/roof_metric.py c9f8bc1be2cb6a09 ran · fixture could not drive it fingerprinted no licence file found · pointer only
Mapping Access to Water and Sanitation in Colombia using Publicly Accessible Satellite Imagery, Crowd-sourced Geospatial Information and RandomForests 2021-11 (from id) thinkingmachines/geoai-immap-wash/utils/modelutils.py 081ab37dbbcd9611 ran · our draft was wrong fingerprinted MIT (permissive)
Revisiting Deep Learning Models for Tabular Data 22 Jun 2021 Yura52/tabular-dl-revisiting-models/lib/metrics.py 12f448bfa8afc6b7 unverified Apache-2.0 (permissive)
GOO: A Dataset for Gaze Object Prediction in Retail Environments 22 May 2021 upeee/GOO-GAZE2021/gazefollowing/evaluate_chong.py 5ae5ac5e4843dd4e unverified Apache-2.0 (permissive)
DRILL: Dynamic Representations for Imbalanced Lifelong Learning 18 May 2021 knowledgetechnologyuhh/drill/models/utils.py 23cde0e185e9c000 unverified MIT (permissive)
Mixed Dimension Embeddings with Application to Memory-Efficient Recommendation Systems 25 Sep 2019 samiwilf/dlrm_from_shz0116/dlrm_s_caffe2.py 332764c725473589 ran · our draft was wrong MIT recorded; this copy not marked cleared · pointer only
Compositional Embeddings Using Complementary Partitions for Memory-Efficient Recommendation Systems 4 Sep 2019 identical code first harvested elsewhere 332764c725473589 ran · our draft was wrong licence of this copy not recorded
The Architectural Implications of Facebook's DNN-based Personalized Recommendation 6 Jun 2019 identical code first harvested elsewhere 332764c725473589 ran · our draft was wrong licence of this copy not recorded
Deep Learning Recommendation Model for Personalization and Recommendation Systems 31 May 2019 myungkeun-cho/taobao_dlrm/dlrm_s_caffe2.py 332764c725473589 ran · our draft was wrong MIT (permissive)
UNet++: A Nested U-Net Architecture for Medical Image Segmentation 18 Jul 2018 marccoru/marinedebrisdetector/marinedebrisdetector/metrics.py e95088d81cb8fe59 unverified MIT (permissive)
The unreasonable effectiveness of the forget gate 13 Apr 2018 JosvanderWesthuizen/janet/aux_code/ops.py 1f5d644b13f89499 unverified MIT (permissive)
arXiv:aaai_35067 unicef/giga-global-school-mapping/utils/calib_utils.py 3092beb3d90eb468 unverified Apache-2.0 (permissive)
arXiv:2025.findings-emnlp.638 liyaooi/LongTableBench/eval/result_process.py 4f7f74693d581896 unverified MIT (permissive)
arXiv:2025.acl-long.179 RUC-NLPIR/RAG-Critic/rag_error_bench/caculate_acc.py bd560baa9eec6619 unverified MIT (permissive)

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections