Home › Code › build_vision_tower

build_vision_tower

Syntologyentry name in harvested coderead from the graph 2026-09-24

build_vision_tower appears in the code Syntology harvested for 35 papers, as 32 distinct code bodies found in 37 places (a place is one code body under one paper). At least one of them ran in 3 of the papers; 0 of the code bodies carry a behaviour fingerprint.

What this page is not. Routines are grouped here by the exact string of their function or class name. Nothing asserts that two samples named build_vision_tower do the same thing, share code, or are comparable; the name is a string, not an identity. Behaviour outputs (what a fingerprinted sample returned on the shared battery) are not in this export and are not shown here; the graph at syntology.ai holds them. "Ran" means executed on a synthesized fixture, not that the code is correct or reproduces a paper.

Samples Syntology

Syntology ran 3 of the 32 distinct code bodies named build_vision_tower; 29 are unverified. One tile per status, in the site's fixed vocabulary, each code body counted once:

0ran · honoured contract
0ran · violated contract
0ran · our draft was wrong
2ran · fixture could not drive it
1ran
29unverified
0fingerprinted

Licence is a property of each copy, so it is counted per place: 16 of the 37 places are pointer only (Syntology does not serve that copy's text). This site shows no code text for any sample; every row below links to the file in its repository where the record names one.

“Ran” means the sample executed on a synthesized input; it does not mean the output is correct. “Honoured” and “violated” refer to a contract Syntology drafted from the code itself; “our draft was wrong” and “fixture could not drive it” are failures of Syntology's instrument, not of the code, and those samples did run. The ran count above is every status except unverified, the same rule as each paper page.

Papers

35 papers shown of 35, newest first; 37 places in the table. A paper with no recorded date is placed by the month its arXiv id encodes, shown in the Date column as YYYY-MM (from id). One row per place: a paper whose repository defines the name more than once appears more than once, and the same code body held for several papers appears once under each, with the same status. Titles and dates are the archive's archive 2025-07-28 for papers in the archive, and the graph's for 3 papers added by Syntology; 4 papers have no page here and are shown by arXiv id only. Status and fingerprint are Syntology's record of each code body; licence is recorded for each place. The File cell ends with the code body's code_sha256, Syntology's identity for that exact code: an agent fetches the stored sample with Syntology's MCP tool get_code(code_sha256="…") (how to connect); click an identity to copy that call.

PaperDateFileStatus SyntologyLicence
FairLLaVA: Fairness-Aware Parameter-Efficient Fine-Tuning for Large Vision-Language Assistants added by Syntology 2026-03 (from id) bhosalems/FairLLaVA/llava/model/language_model/llava_llama.py 29b09ab62f286cf5 unverified licence not identified · pointer only
Uncertainty-Aware Knowledge Distillation for Multimodal Large Language Models added by Syntology 2026-03 (from id) Jingchensun/beta-kd/mobilevlm/model/vision_encoder.py cd281e119911f887 unverified no licence file found · pointer only
Rethinking VLMs for Image Forgery Detection and Localization added by Syntology 2026-03 (from id) sha0fengGuo/IFDL-VLM/stage2/llava/model/language_model/llava_llama.py b89f520c5cbc92b1 unverified MIT (permissive)
Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models 26 May 2025 ignisavium/causal-llava/confounders/get_projector_confounders.py 50280faddf30baf6 unverified MIT (permissive)
arXiv:2503.24164 2025-03 (from id) vlm-svla/svla/llava/model/language_model/llava_qwen.py 2154f3edc3a28ede ran · fixture could not drive it no licence file found · pointer only
arXiv:2503.04006 2025-03 (from id) aminpdik/DSV-LFS/model/DSVLFS.py f553b519cf8e68ba unverified no licence file found · pointer only
EgoLife: Towards Egocentric Life Assistant 5 Mar 2025 evolvinglmms-lab/egolife/EgoGPT/egogpt/model/language_model/egogpt_llama.py 700100fc2116c32c unverified licence not identified · pointer only
Ola: Pushing the Frontiers of Omni-Modal Language Model 6 Feb 2025 ola-omni/ola/ola/model/multimodal_encoder/builder.py 3b3e8a53bce0fb0f unverified Apache-2.0 (permissive)
LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding 14 Jan 2025 appletea233/LLaVA-ST/llava/model/llava_arch.py e081f77558b6872a unverified Apache-2.0 (permissive)
FastVLM: Efficient Vision Encoding for Vision Language Models 17 Dec 2024 apple/ml-fastvlm/llava/model/multimodal_encoder/builder.py 73da1736ca7c15c3 unverified licence not identified · pointer only
WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model 13 Dec 2024 wyddmw/WiseAD/mobilevlm/model/vision_encoder.py cd281e119911f887 unverified no licence file found · pointer only
Frontiers in Intelligent Colonoscopy 22 Oct 2024 ai4colonoscopy/intelliscope/colongpt/model/multimodal_encoder/builder.py 6c1039e6346f9d13 unverified Apache-2.0 (permissive)
M$^2$PT: Multimodal Prompt Tuning for Zero-shot Instruction Learning 24 Sep 2024 william-wang618/m2pt/M2PT/model/llava_archPT.py a11dbce7f6abb48d unverified no licence file found · pointer only
Oryx MLLM: On-Demand Spatial-Temporal Understanding at Arbitrary Resolution 19 Sep 2024 Oryx-mllm/Oryx/oryx/model/multimodal_encoder/builder.py 94ad19a04ab198ad unverified MIT (permissive)
VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation 6 Sep 2024 mit-han-lab/vila-u/vila_u/model/multimodal_encoder/builder.py 6ef3dc934bf9933a unverified MIT (permissive)
AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding 30 Aug 2024 harrytea/adaptvision/llava/model/multimodal_encoder/builder.py b79ef6a64904cf94 unverified no licence file found · pointer only
DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception 11 Jul 2024 baaivision/DenseFusion/densefusion/model/multimodal_encoder/builder.py 38541ff6af9d1575 unverified no licence file found · pointer only
STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering 28 Jun 2024 heliossun/STLLaVA-Med/llava/model/language_model/llava_llama.py b95671438fbf50fa unverified MIT (permissive)
TroL: Traversal of Layers for Large Language and Vision Models 18 Jun 2024 ByungKwanLee/TroL/trol/arch_internlm2/modeling_trol.py 53e08d7b0496620d ran · metamorphic tier: deterministic no licence file found · pointer only
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning 17 Jun 2024 naver-ai/elva/Elva/encoder_builder.py d2f4c2f963c75be2 unverified licence not identified · pointer only
Parrot: Multilingual Visual Instruction Tuning 4 Jun 2024 AIDC-AI/Parrot/parrot/model/multimodal_encoder/builder.py aeaec2cd790d9027 unverified Apache-2.0 (permissive)
ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models 24 May 2024 alibaba/conv-llava/llava/model/multimodal_encoder/builder.py 4562e971e895928a unverified Apache-2.0 (permissive)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts 18 May 2024 hitsz-tmg/umoe-scaling-unified-multimodal-llms/Uni_MoE/Uni_MoE_audio/model/multimodal_encoder/builder.py 9d8ae430eef240b0 unverified no licence file found · pointer only
CoReS: Orchestrating the Dance of Reasoning and Segmentation 8 Apr 2024 baoxiaoyi/cores/model/CORES2seg2mask.py f603e54c961dd1f4 unverified no licence file found · pointer only
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model 29 Mar 2024 intellabs/multimodal_cognitive_ai/LLaVA-Gemma/llava/model/multimodal_encoder/builder.py 4a5d89f6478c0361 unverified MIT (permissive)
VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis 29 Mar 2024 opendatalab/h2rsvlm/vhm/model/multimodal_encoder/builder.py bd1ffaf8ecffb66a unverified Apache-2.0 (permissive)
Yi: Open Foundation Models by 01.AI 7 Mar 2024 01-ai/yi/VL/llava/model/llava_llama.py e016ec919cd88d1e ran · fixture could not drive it Apache-2.0 (permissive)
Grounding Language Models for Visual Entity Recognition 28 Feb 2024 MrZilinXiao/AutoVER/LLaVA/llava/model/language_model/llava_llama.py 01a59cbdcf1329ec unverified Apache-2.0 (permissive)
ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models 18 Feb 2024 freedomintelligence/allava/allava/model/multimodal_encoder/builder.py 9d8ae430eef240b0 unverified Apache-2.0 (permissive)
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model 6 Feb 2024 meituan-automl/mobilevlm/mobilevlm/model/vision_encoder.py cd281e119911f887 unverified Apache-2.0 (permissive)
Towards Improving Document Understanding: An Exploration on Text-Grounding via MLLMs 22 Nov 2023 harrytea/tgdoc/tgdoc/model/multimodal_encoder/builder.py e1613d78436f15d0 unverified no licence file found · pointer only
Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding 5 Jun 2023 damo-nlp-sg/videollama2/videollama2/model/videollama2_arch.py f61dd94ea06b9f92 unverified Apache-2.0 (permissive)
Visual Instruction Tuning 17 Apr 2023 dinhvietcuong1996/icme25-inova/llava/model/language_model/llava_llama.py 875ebbe8f3a0c711 unverified Apache-2.0 (permissive)
Visual Instruction Tuning 17 Apr 2023 ZhangYiqun018/StickerConv/sticker_process/llava/model/language_model/llava_llama.py 014050859277825c unverified Apache-2.0 (permissive)
Visual Instruction Tuning 17 Apr 2023 haotian-liu/LLaVA/llava/model/language_model/llava_llama.py 3ee345b55cc88a11 unverified Apache-2.0 (permissive)
arXiv:2025.naacl-long.579 DAMO-NLP-SG/VideoLLaMA2/videollama2/model/encoder.py f61dd94ea06b9f92 unverified Apache-2.0 (permissive)
arXiv:2025.emnlp-main.609 AuroraZengfh/ModalPrompt/llava/model/language_model/ModalPrompt.py 9d8ae430eef240b0 unverified no licence file found · pointer only

This site shows no code text; each File cell links to the file on GitHub at the repository's current default branch, which may have changed since the harvest. "Pointer only" means Syntology does not serve that copy's text, for one of four reasons: no licence file was found; the licence was not identified; the licence is recorded as permissive but that copy's record is not marked cleared; or the licence is outside the permissive list Syntology serves text under (MIT, Apache-2.0, BSD and similar). Some licences outside that list permit redistribution, such as WTFPL, and GPL-3.0 under its conditions; they are simply not on the list. Hover a licence cell for the reason. Per-sample records for a paper are on its paper page under "Code Syntology ran".

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections