Methods › Computer Vision › Vision and Language Pre-Trained Models › VL-T5
VL-T5
Introduced by Jaemin Cho et al. in Unifying Vision-and-Language Tasks via Text Generation
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
VL-T5 is a unified framework that learns different tasks in a single architecture with the same language modeling objective, i.e., multimodal conditional text generation. The model learns to generate labels in text based on the visual and textual inputs. In contrast to other existing methods, the framework unifies tasks as generating text labels conditioned on multimodal inputs. This allows the model to tackle vision-and-language tasks with unified text generation objective. The models use text prefixes to adapt to different tasks.
Papers archive 2025-07-28
5 shown of 5, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Mixture of Rationale: Multi-Modal Reasoning Mixture for Visual Question Answering 3 Jun 2024 · 0 repositories · arXiv:2406.01402
-
Visual Spatial Description: Controlled Spatial-Oriented Image-to-Text Generation 20 Oct 2022 · 1 repository · arXiv:2210.11109
-
Webly Supervised Concept Expansion for General Purpose Vision Models 4 Feb 2022 · 0 repositories · arXiv:2202.02317
-
VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks 13 Dec 2021 · 1 repository · arXiv:2112.06825Syntology ran 1 of 1 samples · 0 unverified
-
Unifying Vision-and-Language Tasks via Text Generation 4 Feb 2021 · 2 repositories · arXiv:2102.02779Syntology ran 2 of 12 samples · 10 unverified
Tasks archive 2025-07-28
20 shown of 22 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections