| Lyra: An Efficient and Speech-Centric Framework for Omni-Cognition |
1 |
4 |
12 Dec 2024 |
official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (4 pointer-only for licence) |
| ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models |
1 |
2 |
9 Dec 2024 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified |
| LinVT: Empower Your Image-level Large Language Model to Understand Videos |
1 |
1 |
6 Dec 2024 |
official (archive's flag): 7 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 1 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (12 pointer-only for licence) |
| Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling |
1 |
7 |
6 Dec 2024 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified |
| FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression |
1 |
2 |
5 Dec 2024 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (7 pointer-only for licence) |
| VisionZip: Longer is Better but Not Necessary in Vision Language Models |
1 |
6 |
5 Dec 2024 |
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified |
| A Stitch in Time Saves Nine: Small VLM is a Precise Guidance for Accelerating Large VLMs |
1 |
3 |
4 Dec 2024 |
official: harvested, nothing ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (5 pointer-only for licence) |
| Dynamic-LLaVA: Efficient Multimodal Large Language Models via Dynamic Vision-language Context Sparsification |
1 |
2 |
1 Dec 2024 |
official (archive's flag): 13 ran · 13 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified |
| VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment |
0 |
4 |
12 Oct 2024 |
5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (1 pointer-only for licence) |
| Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate |
1 |
2 |
9 Oct 2024 |
official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Emu3: Next-Token Prediction is All You Need |
2 |
1 |
27 Sep 2024 |
6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified |
| Phantom of Latent for Large Language and Vision Models |
1 |
1 |
23 Sep 2024 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified |
| Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution |
8 |
3 |
18 Sep 2024 |
official (archive's flag): 4 ran · 12 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 4 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| CogVLM2: Visual Language Models for Image and Video Understanding |
3 |
2 |
29 Aug 2024 |
official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified |
| Visual Agents as Fast and Slow Thinkers |
1 |
1 |
16 Aug 2024 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model |
1 |
1 |
23 Jul 2024 |
official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity |
1 |
2 |
22 Jul 2024 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified |
| DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception |
1 |
2 |
11 Jul 2024 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (4 pointer-only for licence) |
| InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output |
1 |
1 |
3 Jul 2024 |
official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence) |
| TokenPacker: Efficient Visual Projector for Multimodal LLM |
1 |
2 |
2 Jul 2024 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Efficient Large Multi-modal Models via Visual Context Compression |
1 |
1 |
28 Jun 2024 |
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| MG-LLaVA: Towards Multi-Granularity Visual Instruction Tuning |
1 |
1 |
25 Jun 2024 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified |
| TroL: Traversal of Layers for Large Language and Vision Models |
1 |
1 |
18 Jun 2024 |
official (archive's flag): 23 ran · 23 ran (of which 11 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (30 pointer-only for licence) |
| MMDU: A Multi-Turn Multi-Image Dialog Understanding Benchmark and Instruction-Tuning Dataset for LVLMs |
1 |
1 |
17 Jun 2024 |
official (archive's flag): 5 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 1 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence) |
| Mixture-of-Subspaces in Low-Rank Adaptation |
1 |
2 |
16 Jun 2024 |
official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (6 pointer-only for licence) |
| Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models |
1 |
1 |
3 Jun 2024 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Enhancing Large Vision Language Models with Self-Training on Image Comprehension |
1 |
2 |
30 May 2024 |
official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models |
1 |
1 |
24 May 2024 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified |
| ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models |
1 |
1 |
24 May 2024 |
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (1 pointer-only for licence) |
| Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement |
2 |
1 |
24 May 2024 |
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Dynamic Mixture of Experts: An Auto-Tuning Approach for Efficient Transformer Models |
1 |
1 |
23 May 2024 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified |
| Calibrated Self-Rewarding Vision Language Models |
1 |
2 |
23 May 2024 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| LOVA3: Learning to Visual Question Answering, Asking and Assessment |
1 |
1 |
23 May 2024 |
official (archive's flag): 7 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (10 pointer-only for licence) |
| Imp: Highly Capable Large Multimodal Models for Mobile Devices |
1 |
3 |
20 May 2024 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts |
1 |
1 |
18 May 2024 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (8 pointer-only for licence) |
| CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts |
1 |
1 |
9 May 2024 |
official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| Self-Supervised Visual Preference Alignment |
1 |
2 |
16 Apr 2024 |
official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (11 pointer-only for licence) |
| Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models |
1 |
1 |
11 Apr 2024 |
4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (5 pointer-only for licence) |
| Beyond Embeddings: The Promise of Visual Table in Visual Reasoning |
1 |
2 |
27 Mar 2024 |
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 2 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models |
2 |
3 |
27 Mar 2024 |
official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models |
1 |
1 |
19 Mar 2024 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| SQ-LLaVA: Self-Questioning for Large Vision-Language Assistant |
2 |
2 |
17 Mar 2024 |
official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified |
| MoAI: Mixture of All Intelligence for Large Language and Vision Models |
1 |
1 |
12 Mar 2024 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| DeepSeek-VL: Towards Real-World Vision-Language Understanding |
1 |
1 |
8 Mar 2024 |
official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models |
1 |
1 |
5 Mar 2024 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified |
| The All-Seeing Project V2: Towards General Relation Comprehension of the Open World |
1 |
1 |
29 Feb 2024 |
official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (8 pointer-only for licence) |
| TinyLLaVA: A Framework of Small-scale Large Multimodal Models |
2 |
1 |
22 Feb 2024 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (1 pointer-only for licence) |
| CoLLaVO: Crayon Large Language and Vision mOdel |
1 |
1 |
17 Feb 2024 |
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified |
| SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models |
1 |
1 |
8 Feb 2024 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence) |
| Video-LaVIT: Unified Video-Language Pre-training with Decoupled Visual-Motional Tokenization |
1 |
1 |
5 Feb 2024 |
3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (5 pointer-only for licence) |
| LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model |
1 |
1 |
4 Jan 2024 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (9 pointer-only for licence) |
| V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs |
1 |
1 |
21 Dec 2023 |
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified |
| Generative Multimodal Models are In-Context Learners |
1 |
1 |
20 Dec 2023 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified |
| CogAgent: A Visual Language Model for GUI Agents |
3 |
1 |
14 Dec 2023 |
official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (1 pointer-only for licence) |
| Hallucination Augmented Contrastive Learning for Multimodal Large Language Model |
1 |
1 |
12 Dec 2023 |
official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (5 pointer-only for licence) |
| OneLLM: One Framework to Align All Modalities with Language |
1 |
1 |
6 Dec 2023 |
official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (4 pointer-only for licence) |
| Video-LLaVA: Learning United Visual Representation by Alignment Before Projection |
6 |
1 |
16 Nov 2023 |
community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (2 pointer-only for licence) |
| Volcano: Mitigating Multimodal Hallucination through Self-Feedback Guided Revision |
1 |
2 |
13 Nov 2023 |
official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (10 pointer-only for licence) |
| SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models |
1 |
1 |
13 Nov 2023 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence) |
| LLaVA-Plus: Learning to Use Tools for Creating Multimodal Agents |
1 |
2 |
9 Nov 2023 |
official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| OtterHD: A High-Resolution Multi-modality Model |
1 |
1 |
7 Nov 2023 |
official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence) |
| mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration |
2 |
1 |
7 Nov 2023 |
official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (3 pointer-only for licence) |
| Improved Baselines with Visual Instruction Tuning |
9 |
2 |
5 Oct 2023 |
6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (8 pointer-only for licence) |
| DreamLLM: Synergistic Multimodal Comprehension and Creation |
1 |
1 |
20 Sep 2023 |
official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence) |
| An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models |
1 |
1 |
18 Sep 2023 |
official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence) |
| StableLLaVA: Enhanced Visual Instruction Tuning with Synthesized Image-Dialogue Data |
1 |
1 |
20 Aug 2023 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action |
1 |
2 |
20 Mar 2023 |
official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified |
| GPT-4 Technical Report |
11 |
5 |
15 Mar 2023 |
community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 2 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models |
17 |
1 |
30 Jan 2023 |
community repositories only · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (1 pointer-only for licence) |