Datasets › MM-Vet › Papers where code ran, page 1

MM-Vet

Papers archive 2025-07-28

papers with a benchmark row: 147 · with a code link: 108 · where Syntology ran a sample: 69 (57 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument) Syntology

Show: all papers with a benchmark rowonly where code ran (69 of 147 with a benchmark row: 57 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument)

Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.

Page 1 of 1: papers 1 to 69 of the 69 papers with a benchmark row here where Syntology ran at least one harvested sample (57 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.

The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 339. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record.

PaperCodeResultsDateSamples run Syntology
Lyra: An Efficient and Speech-Centric Framework for Omni-Cognition 1 4 12 Dec 2024 official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (4 pointer-only for licence)
ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models 1 2 9 Dec 2024 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified
LinVT: Empower Your Image-level Large Language Model to Understand Videos 1 1 6 Dec 2024 official (archive's flag): 7 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 1 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (12 pointer-only for licence)
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling 1 7 6 Dec 2024 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified
FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression 1 2 5 Dec 2024 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (7 pointer-only for licence)
VisionZip: Longer is Better but Not Necessary in Vision Language Models 1 6 5 Dec 2024 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified
A Stitch in Time Saves Nine: Small VLM is a Precise Guidance for Accelerating Large VLMs 1 3 4 Dec 2024 official: harvested, nothing ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (5 pointer-only for licence)
Dynamic-LLaVA: Efficient Multimodal Large Language Models via Dynamic Vision-language Context Sparsification 1 2 1 Dec 2024 official (archive's flag): 13 ran · 13 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified
VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment 0 4 12 Oct 2024 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (1 pointer-only for licence)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate 1 2 9 Oct 2024 official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Emu3: Next-Token Prediction is All You Need 2 1 27 Sep 2024 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified
Phantom of Latent for Large Language and Vision Models 1 1 23 Sep 2024 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution 8 3 18 Sep 2024 official (archive's flag): 4 ran · 12 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 4 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
CogVLM2: Visual Language Models for Image and Video Understanding 3 2 29 Aug 2024 official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified
Visual Agents as Fast and Slow Thinkers 1 1 16 Aug 2024 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model 1 1 23 Jul 2024 official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 1 2 22 Jul 2024 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified
DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception 1 2 11 Jul 2024 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (4 pointer-only for licence)
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output 1 1 3 Jul 2024 official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
TokenPacker: Efficient Visual Projector for Multimodal LLM 1 2 2 Jul 2024 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Efficient Large Multi-modal Models via Visual Context Compression 1 1 28 Jun 2024 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
MG-LLaVA: Towards Multi-Granularity Visual Instruction Tuning 1 1 25 Jun 2024 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
TroL: Traversal of Layers for Large Language and Vision Models 1 1 18 Jun 2024 official (archive's flag): 23 ran · 23 ran (of which 11 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (30 pointer-only for licence)
MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs 1 1 17 Jun 2024 official (archive's flag): 5 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 1 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence)
Mixture-of-Subspaces in Low-Rank Adaptation 1 2 16 Jun 2024 official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (6 pointer-only for licence)
Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models 1 1 3 Jun 2024 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Enhancing Large Vision Language Models with Self-Training on Image Comprehension 1 2 30 May 2024 official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models 1 1 24 May 2024 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified
ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models 1 1 24 May 2024 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (1 pointer-only for licence)
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement 2 1 24 May 2024 official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Dynamic Mixture of Experts: An Auto-Tuning Approach for Efficient Transformer Models 1 1 23 May 2024 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
Calibrated Self-Rewarding Vision Language Models 1 2 23 May 2024 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
LOVA3: Learning to Visual Question Answering, Asking and Assessment 1 1 23 May 2024 official (archive's flag): 7 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (10 pointer-only for licence)
Imp: Highly Capable Large Multimodal Models for Mobile Devices 1 3 20 May 2024 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts 1 1 18 May 2024 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (8 pointer-only for licence)
CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts 1 1 9 May 2024 official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
Self-Supervised Visual Preference Alignment 1 2 16 Apr 2024 official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (11 pointer-only for licence)
Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models 1 1 11 Apr 2024 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (5 pointer-only for licence)
Beyond Embeddings: The Promise of Visual Table in Visual Reasoning 1 2 27 Mar 2024 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 2 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models 2 3 27 Mar 2024 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models 1 1 19 Mar 2024 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
SQ-LLaVA: Self-Questioning for Large Vision-Language Assistant 2 2 17 Mar 2024 official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified
MoAI: Mixture of All Intelligence for Large Language and Vision Models 1 1 12 Mar 2024 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
DeepSeek-VL: Towards Real-World Vision-Language Understanding 1 1 8 Mar 2024 official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models 1 1 5 Mar 2024 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified
The All-Seeing Project V2: Towards General Relation Comprehension of the Open World 1 1 29 Feb 2024 official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (8 pointer-only for licence)
TinyLLaVA: A Framework of Small-scale Large Multimodal Models 2 1 22 Feb 2024 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (1 pointer-only for licence)
CoLLaVO: Crayon Large Language and Vision mOdel 1 1 17 Feb 2024 official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified
SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models 1 1 8 Feb 2024 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
Video-LaVIT: Unified Video-Language Pre-training with Decoupled Visual-Motional Tokenization 1 1 5 Feb 2024 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (5 pointer-only for licence)
LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model 1 1 4 Jan 2024 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (9 pointer-only for licence)
V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs 1 1 21 Dec 2023 official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified
Generative Multimodal Models are In-Context Learners 1 1 20 Dec 2023 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
CogAgent: A Visual Language Model for GUI Agents 3 1 14 Dec 2023 official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (1 pointer-only for licence)
Hallucination Augmented Contrastive Learning for Multimodal Large Language Model 1 1 12 Dec 2023 official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (5 pointer-only for licence)
OneLLM: One Framework to Align All Modalities with Language 1 1 6 Dec 2023 official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (4 pointer-only for licence)
Video-LLaVA: Learning United Visual Representation by Alignment Before Projection 6 1 16 Nov 2023 community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (2 pointer-only for licence)
Volcano: Mitigating Multimodal Hallucination through Self-Feedback Guided Revision 1 2 13 Nov 2023 official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (10 pointer-only for licence)
SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models 1 1 13 Nov 2023 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
LLaVA-Plus: Learning to Use Tools for Creating Multimodal Agents 1 2 9 Nov 2023 official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
OtterHD: A High-Resolution Multi-modality Model 1 1 7 Nov 2023 official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence)
mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration 2 1 7 Nov 2023 official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (3 pointer-only for licence)
Improved Baselines with Visual Instruction Tuning 9 2 5 Oct 2023 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (8 pointer-only for licence)
DreamLLM: Synergistic Multimodal Comprehension and Creation 1 1 20 Sep 2023 official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models 1 1 18 Sep 2023 official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data 1 1 20 Aug 2023 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action 1 2 20 Mar 2023 official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
GPT-4 Technical Report 11 5 15 Mar 2023 community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 2 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models 17 1 30 Jan 2023 community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (1 pointer-only for licence)