Browse State-of-the-Art › Text-to-Image Generation
Text-to-Image Generation
546 papers with code · 17 benchmarks · 25 datasets archive 2025-07-28
The development of the brain's blood supply in an embryo involves a complex process with several stages. Initially, there is a network of connections between the carotid and basilar artery systems that usually disappear during development. However, in some cases, these connections remain, leading to variations in the adult cerebral circulation. Around the 24th day of embryonic life (at the 3 mm embryonic stage), the internal carotid arteries (ICA) emerge. They are formed from a combination of the 3rd branchial arch arteries and the distal segments of the paired dorsal aortae. The ICA initially provides all the blood needed by the developing brain. As the brain grows, particularly the occipital region, brainstem, and cerebellum, the blood supply from the ICA becomes insufficient, prompting the development of the posterior circulation. At this stage (4-5 mm embryonic stage), the posterior circulation, specifically the hindbrain, is supplied by two longitudinal neural arteries. These arteries receive blood from several temporary connections (anastomoses) with the ICA, including: Trigeminal artery (TA), Otic artery (OA), Hypoglossal artery (HA), Proatlantal artery (ProA). These carotid-vertebrobasilar anastomoses act as temporary bridges, providing blood flow to the developing posterior circulation until the vertebral arteries (VA) and the basilar artery (BA) are fully formed.6
The BA forms between the 5-8 mm stage by the fusion of the longitudinal neural arteries. As the posterior communicating artery develops and connects with the distal BA, the TA, OA, and HA typically regress. Unlike the TA, OA, and HA, the ProA persists until the VA are fully developed. A segment of the ProA becomes incorporated into the V3 segment of the VA and the distal portions of the occipital artery. The VA develop between the 7-12 mm stage from transverse connections between cervical intersegmental arteries, starting with the ProA and progressing downwards to the 6th intersegmental artery. This 6th intersegmental artery eventually forms the origin of the adult VA from the subclavian artery. 7-8
Persistence of the HA into adulthood, the focus of this discussion, is a rare occurrence. It is the second most common persistent carotid-vertebrobasilar anastomosis, following the persistent trigeminal artery. The failure of the HA to regress during embryological development leads to PPHA. This condition can impact the normal blood flow dynamics in the brain and, in certain cases, can be associated with cerebrovascular issues like aneurysms or ischemic events.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
17 leaderboard tables shown for this task, 17 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 17 until expanded.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
25 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
7 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 546 papers with code (1,085 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
17 Nov 2014 74 repositories listed Syntology ran 13 of 34 samples · 21 unverified · 6 pointer-only (licence)Experiments on several datasets show the accuracy of the model and the fluency of the language it learns solely from image descriptions.
-
20 Dec 2021 41 repositories listed Syntology ran 19 of 28 samples · 9 unverified · 5 pointer-only (licence)By decomposing the image formation process into a sequential application of denoising autoencoders, diffusion models (DMs) achieve state-of-the-art synthesis results on image data and beyond.
-
17 May 2016 39 repositories listed Syntology ran 9 of 19 samples · 10 unverified · 6 pointer-only (licence)Automatic synthesis of realistic images from text would be interesting and useful, but current AI systems are still far from this goal.
-
10 Dec 2016 21 repositories listed Syntology ran 11 of 31 samples · 20 unverified · 1 pointer-only (licence)Synthesizing high-quality images from text descriptions is a challenging problem in computer vision and has many practical applications.
-
28 Nov 2017 20 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedIn this paper, we propose an Attentional Generative Adversarial Network (AttnGAN) that allows attention-driven, multi-stage refinement for fine-grained text-to-image generation.
-
19 Oct 2017 16 repositories listed Syntology ran 10 of 31 samples · 21 unverifiedIn this paper, we propose Stacked Generative Adversarial Networks (StackGAN) aiming at generating high-resolution photo-realistic images.
-
17 Dec 2020 13 repositories listed Syntology ran 6 of 6 samples · 0 unverified · 4 pointer-only (licence)We demonstrate how combining the effectiveness of the inductive bias of CNNs with the expressivity of transformers enables them to model and thereby synthesize high-resolution images.
-
24 Feb 2021 12 repositories listed Syntology ran 7 of 7 samples · 0 unverified · 3 pointer-only (licence)Text-to-image generation has traditionally focused on finding better modeling assumptions for training on a fixed dataset.
-
2 Aug 2022 9 repositories listed Syntology ran 10 of 13 samples · 3 unverified · 1 pointer-only (licence)Yet, it is unclear how such freedom can be exercised to generate images of specific unique concepts, modify their appearance, or compose them in new roles and novel scenes.
-
8 Feb 2022 9 repositories listed Syntology ran 14 of 21 samples · 7 unverified · 4 pointer-only (licence)At inference time, the model begins with generating all tokens of an image simultaneously, and then refines the image iteratively conditioned on the previous generation.
-
13 Apr 2022 8 repositories listed Syntology ran 29 of 38 samples · 9 unverified · 1 pointer-only (licence)Contrastive models like CLIP have been shown to learn robust representations of images that capture both semantics and style.
-
6 Oct 2023 5 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedInspired by Consistency Models (song et al.), we propose Latent Consistency Models (LCMs), enabling swift inference with minimal steps on any pre-trained LDMs, including Stable Diffusion (rombach et al).
-
2 Jan 2023 5 repositories listed Syntology ran 18 of 21 samples · 3 unverified · 9 pointer-only (licence)Compared to pixel-space diffusion models, such as Imagen and DALL-E 2, Muse is significantly more efficient due to the use of discrete tokens and requiring fewer sampling iterations; compared to autoregressive models,…
-
6 Dec 2020 5 repositories listedIn this work, we propose TediGAN, a novel framework for multi-modal image generation and manipulation with textual descriptions.
-
1 Jun 2023 4 repositories listed Syntology ran 2 of 2 samples · 0 unverifiedPre-trained large text-to-image models synthesize impressive images with an appropriate use of text prompts.
-
17 Apr 2023 4 repositories listed Syntology ran 4 of 7 samples · 3 unverified · 1 pointer-only (licence)Despite the success in large-scale text-to-image generation and text-conditioned image editing, existing methods still struggle to produce consistent generation and editing results.
-
14 Nov 2022 4 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedRecent advancements in the domain of text-to-image synthesis have culminated in a multitude of enhancements pertaining to quality, fidelity, and diversity.
-
3 Mar 2022 4 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)However, we postulate that previous VQ cannot shorten the code sequence and generate high-fidelity images together in terms of the rate-distortion trade-off.
-
7 Feb 2022 4 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedIn this work, we pursue a unified paradigm for multimodal pretraining to break the scaffolds of complex task/modality-specific customization.
-
26 May 2021 4 repositories listed Syntology ran 4 of 6 samples · 2 unverifiedText-to-Image generation in the general domain has long been an open problem, which requires both a powerful generative model and cross-modal understanding.
-
2 Apr 2019 4 repositories listed Syntology ran 4 of 5 samples · 1 unverified · 1 pointer-only (licence)If the initial image is not well initialized, the following processes can hardly refine the image to a satisfactory quality.
-
5 Dec 2023 3 repositories listedWe demonstrate the usage of state-of-the-art text-to-image architectures in the context of laparoscopic imaging with regard to the surgical removal of the gallbladder as an example.
-
30 Sep 2023 3 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 2 pointer-only (licence)We hope PIXART-α will provide new insights to the AIGC community and startups to accelerate building their own high-quality yet low-cost generative models from scratch.
-
25 May 2023 3 repositories listed Syntology ran 12 of 13 samples · 1 unverifiedWe apply ProSpect in various personalized attribute-aware image generation applications, such as image-guided or text-driven manipulations of materials, style, and layout, achieving previously unattainable results from…
-
25 May 2023 3 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Text-to-image (T2I) generation with Stable Diffusion models (SDMs) involves high computing demands due to billion-scale parameters.
-
22 May 2023 3 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedHowever, most use cases of diffusion models are not concerned with likelihoods, but instead with downstream objectives such as human-perceived image quality or drug effectiveness.
-
12 Apr 2023 3 repositories listed Syntology ran 3 of 12 samples · 9 unverified · 2 pointer-only (licence)We present a comprehensive solution to learn and improve text-to-image models from human preference feedback.
-
31 Mar 2023 3 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 1 pointer-only (licence)Recent breakthroughs in the field of language-guided image generation have yielded impressive achievements, enabling the creation of high-quality and diverse images based on user instructions.
-
12 Mar 2023 3 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Inspired by the unified view, UniDiffuser learns all distributions simultaneously with a minimal modification to the original diffusion model -- perturbs data in all modalities instead of a single modality, inputs…
-
22 Feb 2023 3 repositories listed Syntology ran 4 of 24 samples · 20 unverifiedIn this work, we build upon these ideas using the score-based interpretation of diffusion models, and explore alternative ways to condition, modify, and reuse diffusion models for tasks involving compositional…
Syntology lines on 28 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections