Home › Code › get_eval

get_eval

Syntologyentry name in harvested coderead from the graph 2026-09-24

get_eval appears in the code Syntology harvested for 18 papers, as 21 distinct code bodies found in 24 places (a place is one code body under one paper). At least one of them ran in 9 of the papers; 2 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named get_eval do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 9 of the 21 distinct code bodies named get_eval; 12 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

0ran · honoured contract
0ran · violated contract
0ran · our draft was wrong
0ran · fixture could not drive it
9ran
12unverified
2fingerprinted

Licence is a property of each copy, so it is counted per place: 3 of the 24 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

18 papers shown of 18, newest first; 24 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning 19 Mar 2025 aimagelab/LLaVA-MORE/src/llava/eval/eval_gpt_review_bench.py fb48214e8ea84c17 ran Apache-2.0 (permissive)
LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning 19 Mar 2025 aimagelab/LLaVA-MORE/src/llava/eval/eval_gpt_review.py 64a3ae0bc1b51280 unverified Apache-2.0 (permissive)
Agri-LLaVA: Knowledge-Infused Large Multimodal Assistant on Agricultural Pests and Diseases 3 Dec 2024 kki2eve/agri-llava/agri_llava/eval/eval_gpt_review_visual.py 5f955676f48705c9 unverified Apache-2.0 (permissive)
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models 25 Aug 2024 yejipark-m/convis/eval/eval_gpt_review_bench.py 13cb3dd1e9e1ed71 ran MIT (permissive)
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning 17 Jun 2024 naver-ai/elva/Elva/eval_gpt_review_parsing_bench.py a5b1f40fc357a806 unverified licence not identified · pointer only
Unleashing the Potential of Diffusion Models for Incomplete Data Imputation 31 May 2024 hengruizhang98/DiffPuter/dataset.py 2f212d29fbe5460e ran MIT (permissive)
Unleashing the Potential of Diffusion Models for Incomplete Data Imputation 31 May 2024 hengruizhang98/DiffPuter/baselines/data_utils.py b1c68b94d50528a2 ran MIT (permissive)
RLAIF-V: Open-Source AI Feedback Leads to Super GPT-4V Trustworthiness 27 May 2024 rlhf-v/rlhf-v/eval/gpt4_grpc.py 77010a13f8bd8d11 unverified no licence file found · pointer only
Groma: Localized Visual Tokenization for Grounding Multimodal Large Language Models 19 Apr 2024 FoundationVision/Groma/groma/eval/eval_gpt_review_visual.py 55b3991e52737846 unverified Apache-2.0 (permissive)
Beyond Embeddings: The Promise of Visual Table in Visual Reasoning 27 Mar 2024 lavi-lab/visual-table/llava/eval/eval_gpt_review_visual.py fb48214e8ea84c17 ran Apache-2.0 (permissive)
Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning 25 Mar 2024 deepcs233/visual-cot/llava/eval/eval_gpt_review_visual.py fb48214e8ea84c17 ran Apache-2.0 (permissive)
Cognitive Visual-Language Mapper: Advancing Multimodal Comprehension with Enhanced Visual Knowledge Alignment 21 Feb 2024 hitsz-tmg/cognitive-visual-language-mapper/LLaVA/llava/eval/eval_gpt_review_visual.py fb48214e8ea84c17 ran no licence file found · pointer only
Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos 7 Dec 2023 aldraus/quilt-llava/llava/eval/quilt_gpt_eval.py c00535a655c69753 unverified MIT (permissive)
The Philosopher's Stone: Trojaning Plugins of Large Language Models 2023-12 (from id) chichidd/llm-lora-trojan/eval/gpt_review.py 191943d87af4d1bb unverified Apache-2.0 (permissive)
FollowBench: A Multi-level Fine-grained Constraints Following Benchmark for Large Language Models 31 Oct 2023 yjiangcm/followbench/code/llm_eval.py 4cd97990a8cad73f ran fingerprinted Apache-2.0 (permissive)
FollowBench: A Multi-level Fine-grained Constraints Following Benchmark for Large Language Models 31 Oct 2023 yjiangcm/followbench/code_zh/llm_eval.py 570dd47754890fe2 ran fingerprinted Apache-2.0 (permissive)
UltraFeedback: Boosting Language Models with Scaled AI Feedback 2 Oct 2023 thunlp/ultrafeedback/src/data_annotation/annotate_critique.py a506738814a38c14 ran MIT (permissive)
UltraFeedback: Boosting Language Models with Scaled AI Feedback 2 Oct 2023 thunlp/ultrafeedback/src/data_annotation/fix_overall_score_issue.py 9f8a0e28964db553 ran MIT (permissive)
UltraFeedback: Boosting Language Models with Scaled AI Feedback 2 Oct 2023 thunlp/ultrafeedback/src/data_annotation/annotate_preference.py 38d99fe231928f20 unverified MIT (permissive)
UniversalNER: Targeted Distillation from Large Language Models for Open Named Entity Recognition 7 Aug 2023 universal-ner/universal-ner/src/train/fastchat/eval/eval_gpt_review.py 26ac62acef82f73c unverified MIT (permissive)
PRD: Peer Rank and Discussion Improve Large Language Model based Evaluations 6 Jul 2023 bcdnlp/prd/peer_rank/eval_bard_review.py f296a758b3b2f6d2 ran MIT (permissive)
Lion: Adversarial Distillation of Proprietary Large Language Models 22 May 2023 yjiangcm/lion/src/chatgpt_inference.py 8cd557465a1f5553 unverified MIT (permissive)
Lion: Adversarial Distillation of Proprietary Large Language Models 22 May 2023 yjiangcm/lion/src/chatgpt_referee.py 5823dba3817cc067 unverified MIT (permissive)
Learning Controllable Adaptive Simulation for Multi-resolution Physics 1 May 2023 snap-stanford/lamp/analysis_2d_full.py 32a2b6aa244f4c8c unverified MIT (permissive)

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections