Browse State-of-the-Art › Image Manipulation
Image Manipulation
201 papers with code · 1 benchmark · 7 datasets archive 2025-07-28
Image Manipulation is the process of altering or transforming an existing image to achieve a desired effect or to modify its content. This can involve various techniques and tools to enhance, modify, or create images based on specific requirements.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| LRS2 (2 rows) | TPS | Image Shape Manipulation from a Single Augmented Training Sample | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
7 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 201 papers with code (427 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
2 May 2019 47 repositories listed Syntology ran 2 of 11 samples · 9 unverified · 3 pointer-only (licence)We introduce SinGAN, an unconditional generative model that can be learned from a single natural image.
-
13 Jul 2020 11 repositories listed Syntology ran 7 of 14 samples · 7 unverifiedA rich set of interpretable dimensions has been shown to emerge in the latent space of the Generative Adversarial Networks (GANs) trained for synthesizing images.
-
8 Feb 2022 9 repositories listed Syntology ran 14 of 21 samples · 7 unverified · 4 pointer-only (licence)At inference time, the model begins with generating all tokens of an image simultaneously, and then refines the image iteratively conditioned on the previous generation.
-
4 Feb 2021 8 repositories listed Syntology ran 1 of 2 samples · 1 unverifiedWe then suggest two principles for designing encoders in a manner that allows one to control the proximity of the inversions to regions that StyleGAN was originally trained on.
-
25 Jun 2020 8 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)SRFlow therefore directly accounts for the ill-posed nature of the problem, and learns to predict diverse photo-realistic high-resolution images.
-
18 May 2023 7 repositories listed Syntology ran 7 of 16 samples · 9 unverified · 3 pointer-only (licence)Synthesizing visual content that meets users' needs often requires flexible and precise controllability of the pose, shape, expression, and layout of the generated objects.
-
27 Jul 2019 7 repositories listedTo overcome these drawbacks, we propose a novel framework termed MaskGAN, enabling diverse and interactive face manipulation.
-
23 Nov 2016 6 repositories listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)Neural Style Transfer has shown very exciting results enabling new forms of image manipulation.
-
31 Mar 2021 5 repositories listed Syntology ran 7 of 11 samples · 4 unverified · 1 pointer-only (licence)Inspired by the ability of StyleGAN to generate highly realistic images in a variety of domains, much recent work has focused on understanding how to use the latent spaces of StyleGAN to manipulate generated and real…
-
5 Oct 2019 5 repositories listed Syntology ran 0 of 20 samples · 20 unverifiedThis work presents Kornia -- an open source computer vision library which consists of a set of differentiable routines and modules to solve generic computer vision problems.
-
5 Jan 2021 4 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedEstablishing dense correspondences between a pair of images is an important and general problem.
-
1 Jul 2020 4 repositories listed Syntology ran 3 of 7 samples · 4 unverified · 7 pointer-only (licence)Deep generative models have become increasingly effective at producing realistic images from randomly sampled seeds, but using such models for controllable manipulation of existing images remains challenging.
-
25 Jul 2019 4 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 5 pointer-only (licence)In this work, we propose a novel framework, called InterFaceGAN, for semantic face editing by interpreting the latent semantics learned by GANs.
-
18 Mar 2024 3 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Reverse sampling and score-distillation have emerged as main workhorses in recent years for image manipulation using latent diffusion models (LDMs).
-
19 May 2023 3 repositories listedAs an exemplar, we leverage LeftRefill to address two different challenges: reference-guided inpainting and novel view synthesis, based on the pre-trained StableDiffusion.
-
15 Nov 2022 3 repositories listedIn this work, we expand the existing single-flow diffusion pipeline into a multi-task multimodal network, dubbed Versatile Diffusion (VD), that handles multiple flows of text-to-image, image-to-text, and variations in…
-
2 Aug 2021 3 repositories listed Syntology ran 4 of 10 samples · 6 unverifiedCan a generative model be trained to produce images from a specific domain, guided by a text prompt only, without seeing any image?
-
12 Jun 2021 3 repositories listed Syntology ran 13 of 20 samples · 7 unverified · 3 pointer-only (licence)Conditional generative models of high-dimensional images have many applications, but supervision signals from conditions to images can be expensive to acquire.
-
10 Jun 2021 3 repositories listed Syntology ran 2 of 4 samples · 2 unverifiedThe key idea is pivotal tuning - a brief training process that preserves the editing quality of an in-domain latent region, while changing its portrayed identity and appearance.
-
11 May 2020 3 repositories listedThis can be done by conditioning the model on additional information.
-
7 Mar 2020 3 repositories listed Syntology ran 0 of 2 samples · 2 unverified · 2 pointer-only (licence)Editing existing images requires embedding a given image into the latent space of StyleGAN2.
-
12 Dec 2019 3 repositories listedThe goal of our paper is to semantically edit parts of an image matching a given text that describes desired attributes (e.
-
1 Jun 2019 3 repositories listedTo fight against real-life image forgery, which commonly involves different types and combined manipulations, we propose a unified deep neural architecture called ManTra-Net.
-
5 Apr 2019 3 repositories listedWe introduce point-to-point video generation that controls the generation process with two control points: the targeted start- and end-frames.
-
3 Jun 2025 2 repositories listedAlthough existing unified models achieve strong performance in vision-language understanding and text-to-image generation, they remain limited in addressing image perception and manipulation -- capabilities increasingly…
-
20 May 2025 2 repositories listed Syntology ran 9 of 21 samples · 12 unverifiedUnifying multimodal understanding and generation has shown impressive capabilities in cutting-edge proprietary systems.
-
30 Apr 2025 2 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Visual text is a crucial component in both document and scene images, conveying rich semantic information and attracting significant attention in the computer vision community.
-
3 Oct 2024 2 repositories listed Syntology ran 3 of 6 samples · 3 unverified · 4 pointer-only (licence)The rapid development of generative AI is a double-edged sword, which not only facilitates content creation but also makes image manipulation easier and more difficult to detect.
-
29 Sep 2023 2 repositories listed Syntology ran 3 of 6 samples · 3 unverified · 6 pointer-only (licence)Extensive experimental results demonstrate that expressive instructions are crucial to instruction-based image editing, and our MGIE can lead to a notable improvement in automatic metrics and human evaluation while…
-
23 Nov 2022 2 repositories listed Syntology ran 7 of 20 samples · 13 unverified · 3 pointer-only (licence)Language-guided image editing has achieved great success recently.
Syntology lines on 22 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections