Browse State-of-the-Art › text-guided-image-editing
text-guided-image-editing
34 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Editing images using text prompts.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 34 papers with code (68 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
15 Nov 2022 3 repositories listedIn this work, we expand the existing single-flow diffusion pipeline into a multi-task multimodal network, dubbed Versatile Diffusion (VD), that handles multiple flows of text-to-image, image-to-text, and variations in…
-
2 Jun 2024 2 repositories listed Syntology ran 9 of 35 samples · 26 unverified · 19 pointer-only (licence)In recent years, diffusion models have achieved remarkable success in the realm of high-quality image generation, garnering increased attention.
-
2 Oct 2023 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedRecently, a myriad of conditional image generation and editing models have been developed to serve different downstream tasks, including text-to-image generation, text-guided image editing, subject-driven image…
-
29 Mar 2023 2 repositories listed Syntology ran 3 of 6 samples · 3 unverifiedImage generation using diffusion can be controlled in multiple ways.
-
22 Nov 2022 2 repositories listedEDICT enables mathematically exact inversion of real and model-generated images by maintaining two coupled noise vectors which are used to invert each other in an alternating fashion.
-
6 Oct 2022 2 repositories listedFor standard diffusion models trained on the pixel-space, our approach is able to generate images visually comparable to that of the original model using as few as 4 sampling steps on ImageNet 64x64 and CIFAR-10,…
-
9 Jun 2025 1 repository listedWe find that the high degree of compression achieved by a 1D tokenizer with vector quantization enables image editing and generative capabilities through heuristic manipulation of tokens, demonstrating that even very…
-
9 Jun 2025 1 repository listedIn this paper, we introduce PairEdit, a novel visual editing method designed to effectively learn complex editing semantics from a limited number of image pairs or even a single image pair, without using any textual…
-
16 May 2025 1 repository listedEditing images using natural language instructions has become a natural and expressive way to modify visual content; yet, evaluating the performance of such models remains challenging.
-
1 May 2025 1 repository listedA variety of text-guided image editing models have been proposed recently.
-
31 Mar 2025 1 repository listedText-guided image editing is an essential task that enables users to modify images through natural language descriptions.
-
27 Mar 2025 1 repository listedText-guided image editing aims to modify specific regions of an image according to natural language instructions while maintaining the general structure and the background fidelity.
-
14 Mar 2025 1 repository listedExperimental results show our method outperforms state-of-the-art score distillation techniques in prompt fidelity, improving successful edits while preserving the background.
-
11 Mar 2025 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)While recent advancements in generative modeling have significantly improved text-image alignment, some residual misalignment between text and image representations still remains.
-
3 Nov 2024 1 repository listedA plethora of text-guided image editing methods has recently been developed by leveraging the impressive capabilities of large-scale diffusion-based generative models especially Stable Diffusion.
-
15 Sep 2024 1 repository listedIn this paper, we propose TextureDiffusion, a tuning-free image editing method applied to various texture transfer.
-
31 Jul 2024 1 repository listedThe test-time finetuning text-guided image editing method, Forgedit, is capable of tackling general and complex image editing problems given only the input image itself and the target text prompt.
-
2 May 2024 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Large-scale Text-to-Image (T2I) diffusion models demonstrate significant generation capabilities based on textual prompts.
-
14 Mar 2024 1 repository listedDiffusion models have achieved remarkable success in the domain of text-guided image generation and, more recently, in text-guided image editing.
-
27 Feb 2024 1 repository listedIn this survey, we provide an exhaustive overview of existing methods using diffusion models for image editing, covering both theoretical and practical aspects in the field.
-
7 Jan 2024 1 repository listedTo increase user freedom, we propose a new task called Specific Reference Condition Real Image Editing, which allows user to provide a reference image to further control the outcome, such as replacing an object with a…
-
17 Dec 2023 1 repository listedWhile several powerful distillation methods were recently proposed, the overall quality of student samples is typically lower compared to the teacher ones, which hinders their practical usage.
-
30 Nov 2023 1 repository listed Syntology ran 7 of 7 samples · 0 unverified · 7 pointer-only (licence)Although generative editing methods now enable some forms of image editing, relighting is still beyond today's capabilities; existing methods struggle to keep other aspects of the image -- colors, shapes, and textures…
-
12 Oct 2023 1 repository listedBased on DeltaSpace, we propose a novel framework called DeltaEdit, which maps the CLIP visual feature differences to the latent space directions of a generative model during the training phase, and predicts the latent…
-
2 Oct 2023 1 repository listedOur conditional-task learning and distillation approach outperforms previous distillation methods, achieving a new state-of-the-art in producing high-quality images with very few steps (e.
-
19 Sep 2023 1 repository listed Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)Text-guided image editing on real or synthetic images, given only the original image itself and the target text prompt as inputs, is a very general and challenging task.
-
11 Aug 2023 1 repository listedTo address this issue, we propose masked-attention guidance, which can generate images more faithful to semantic masks via indirect control of attention to each word and pixel by manipulating noise images fed to…
-
16 Jun 2023 1 repository listedTo address this issue, we introduce MagicBrush (https://osu-nlp-group.
-
29 May 2023 1 repository listed Syntology ran 2 of 2 samples · 0 unverifiedIn this work, we propose a framework termed InstructEdit that can do fine-grained editing based on user instructions.
-
1 May 2023 1 repository listedWe present Prompt Diffusion, a framework for enabling in-context learning in diffusion-based generative models.
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections