Home › Code › generate_caption

generate_caption

Syntologyentry name in harvested coderead from the graph 2026-09-24

generate_caption appears in the code Syntology harvested for 21 papers, as 7 distinct code bodies found in 21 places (a place is one code body under one paper). At least one of them ran in 16 of the papers; 2 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named generate_caption do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 3 of the 7 distinct code bodies named generate_caption; 4 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

0ran · honoured contract
0ran · violated contract
0ran · our draft was wrong
0ran · fixture could not drive it
3ran
4unverified
2fingerprinted

Licence is a property of each copy, so it is counted per place: 9 of the 21 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

21 papers shown of 21, newest first; 21 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 1 papers added by Syntology; 3 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
BLUEX v2: Benchmarking LLMs on Open-Ended Questions from Brazilian University Entrance Exams added by Syntology 2026-06 (from id) TropicAI-Research/BLUEXv2/dataset_pipeline/generate_captions.py f8e511927033900d ran fingerprinted no licence file found · pointer only
Pretrained Image-Text Models are Secretly Video Captioners 19 Feb 2025 chunhuizng/mllm-video-captioner/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause recorded; this copy not marked cleared · pointer only
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation 28 Nov 2024 lorebianchi98/Talk2DINO/dino_extraction_v2.py e56697ee94f100ea unverified Apache-2.0 (permissive)
T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation 19 Jul 2024 KaiyueSun98/T2V-CompBench/Grounded-Segment-Anything/automatic_label_demo.py d893c774b1f34088 unverified no licence file found · pointer only
UltraEdit: Instruction-based Fine-Grained Image Editing at Scale 7 Jul 2024 pkunlp-icler/ultraedit/data_generation/Grounded-Segment-Anything/automatic_label_demo.py d893c774b1f34088 unverified no licence file found · pointer only
InstantStyle-Plus: Style Transfer with Content-Preserving in Text-to-Image Generation 30 Jun 2024 instantx-research/instantstyle-plus/infer_style.py f47234275dee8d5c unverified no licence file found · pointer only
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding 19 Jun 2024 alanaai/evud/hm3d/prepare_hm3d_dataset.py d776259a0bbaf187 ran fingerprinted MIT (permissive)
Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards 22 May 2024 XiaoyuYoung/ConceptDriftMLLMs/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding 8 Apr 2024 boheumd/MA-LMM/app/caption.py 56f02812d66a0d11 ran MIT (permissive)
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language Models 1 Apr 2024 shtuplus/pix2grp_cvpr2024/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
Shot2Story20K: A New Benchmark for Comprehensive Understanding of Multi-shot Videos 16 Dec 2023 bytedance/Shot2Story/code/app/caption.py 56f02812d66a0d11 ran no licence file found · pointer only
LMDrive: Closed-Loop End-to-End Driving with Large Language Models 12 Dec 2023 opendilab/lmdrive/LAVIS/app/caption.py 56f02812d66a0d11 ran Apache-2.0 (permissive)
Localized Symbolic Knowledge Distillation for Visual Commonsense Models 8 Dec 2023 jamespark3922/lskd/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors 26 Oct 2023 nctu-eva-lab/antifakeprompt/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
Bootstrapping Vision-Language Learning with Decoupled Language Pre-training 13 Jul 2023 yiren-jian/BLIText/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
Self-Chained Image-Language Model for Video Localization and Question Answering 11 May 2023 yui010206/sevila/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause recorded; this copy not marked cleared · pointer only
VPGTrans: Transfer Visual Prompt Generator across LLMs 2 May 2023 VPGTrans/VPGTrans/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
Captioning Images with Diverse Objects 24 Jun 2016 willT97/Zero-shot-Image-Captioner/noc_train.py a075c8fcc5cc6b01 unverified no licence file found · pointer only
arXiv:2024.findings-emnlp.649 alanaai/EVUD/hm3d/prepare_hm3d_dataset.py d776259a0bbaf187 ran fingerprinted MIT (permissive)
arXiv:2024.emnlp-main.88 WenjianDing/ReBo/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause (permissive)
arXiv:2024.emnlp-main.722 xyh97/UNICORN/app/caption.py 56f02812d66a0d11 ran BSD-3-Clause recorded; this copy not marked cleared · pointer only

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections