Browse State-of-the-Art › Personalized Image Generation
Personalized Image Generation
31 papers with code · 1 benchmark · 1 dataset archive 2025-07-28
Utilizes single or multiple images that contain the same subject or style, along with text prompt, to generate images that contain that subject as well as match the textual description. Includes finetuning-based methods (e.g. DreamBooth, Textual Inversion) as well as encoder-based methods (e.g. E4T, ELITE, and IP-Adapter, etc.).
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| DreamBooth (7 rows) | DreamBooth LoRA SDXL v1.0 | DreamBooth: Fine Tuning Text-to-Image Diffusion Models for... | code | Syntology ran 10 of 12 samples · 2 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 31 papers with code (58 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
25 Aug 2022 12 repositories listed Syntology ran 10 of 12 samples · 2 unverified · 8 pointer-only (licence)Once the subject is embedded in the output domain of the model, the unique identifier can be used to synthesize novel photorealistic images of the subject contextualized in different scenes.
-
2 Aug 2022 9 repositories listed Syntology ran 10 of 13 samples · 3 unverified · 1 pointer-only (licence)Yet, it is unclear how such freedom can be exercised to generate images of specific unique concepts, modify their appearance, or compose them in new roles and novel scenes.
-
13 Aug 2023 4 repositories listed Syntology ran 3 of 8 samples · 5 unverifiedDespite the simplicity of our method, an IP-Adapter with only 22M parameters can achieve comparable or even better performance to a fully fine-tuned image prompt model.
-
17 Apr 2025 1 repository listedPersonalized image synthesis has emerged as a pivotal application in text-to-image generation, enabling the creation of images featuring specific subjects in diverse contexts.
-
2 Apr 2025 1 repository listed Syntology ran 13 of 43 samples · 30 unverifiedIn this study, we propose a highly-consistent data synthesis pipeline to tackle this challenge.
-
9 Mar 2025 1 repository listedWe propose Conceptrol, a simple yet effective framework that enhances zero-shot adapters without adding computational overhead.
-
9 Mar 2025 1 repository listedHowever, current methods face challenges in ensuring fidelity to the text prompt while not overfitting to the training data.
-
18 Feb 2025 1 repository listedRecent advancements in generative models have significantly facilitated the development of personalized content creation.
-
9 Feb 2025 1 repository listedTo address this issue, we systematically analyze sampling strategies beyond fine-tuning, exploring the impact of concept and superclass trajectories on the results.
-
20 Dec 2024 1 repository listedPersonalized image generation has made significant strides in adapting content to novel concepts.
-
4 Dec 2024 1 repository listed Syntology ran 3 of 11 samples · 8 unverified · 11 pointer-only (licence)To tackle this problem, this work proposes PatchDPO that estimates the quality of image patches within each generated image and accordingly trains the model.
-
18 Oct 2024 1 repository listed Syntology ran 3 of 5 samples · 2 unverified · 5 pointer-only (licence)Personalized content filtering, such as recommender systems, has become a critical infrastructure to alleviate information overload.
-
16 Oct 2024 1 repository listedAs a result, FACT solely learns identity preservation from training data, thereby minimizing the impact on the original text-to-image capabilities of the base model.
-
26 Sep 2024 1 repository listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)Personalized text-to-image generation methods can generate customized images based on the reference images, which have garnered wide research interest.
-
19 Sep 2024 1 repository listedHowever, the lack of holistic consistency in scenes with multiple characters hampers these methods' ability to create a cohesive narrative.
-
12 Sep 2024 1 repository listed Syntology ran 11 of 14 samples · 3 unverifiedRecent breakthroughs in text-to-image models have opened up promising research avenues in personalized image generation, enabling users to create diverse images of a specific subject using natural language prompts.
-
12 Sep 2024 1 repository listedZero-shot personalized image generation models aim to produce images that align with both a given text prompt and subject image, requiring the model to effectively incorporate both sources of guidance.
-
24 Jun 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedPersonalized image generation holds great promise in assisting humans in everyday work and life due to its impressive function in creatively generating personalized content.
-
23 May 2024 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Our study shows that based on a recent rectified flow framework, the major limitation of vanilla classifier guidance in requiring a special classifier can be resolved with a simple fixed-point solution, allowing…
-
11 Apr 2024 1 repository listedFinally, we mention the possibility of CAT in the aspects of multi-concept adapter and optimization.
-
8 Apr 2024 1 repository listed Syntology ran 1 of 4 samples · 3 unverified · 4 pointer-only (licence)This approach effectively synergizes reference image and text prompt information to produce valuable image features, facilitating an image diffusion model.
-
23 Feb 2024 1 repository listed Syntology ran 8 of 10 samples · 2 unverifiedFirst, current personalization techniques fail to reliably extend to multiple concepts -- we hypothesize this to be due to the mismatch between complex scenes and simple text descriptions in the pre-training dataset (e.
-
25 Jan 2024 1 repository listedWe propose a novel architecture (BootPIG) that allows a user to provide reference images of an object in order to guide the appearance of a concept in the generated images.
-
1 Jan 2024 1 repository listedText-to-image diffusion models have remarkably excelled in producing diverse high-quality and photo-realistic images.
-
20 Dec 2023 1 repository listed Syntology ran 3 of 4 samples · 1 unverifiedThe human ability to easily solve multimodal tasks in context (i.
-
29 Nov 2023 1 repository listedText-to-image diffusion models have remarkably excelled in producing diverse, high-quality, and photo-realistic images.
-
28 Aug 2023 1 repository listedIn this paper, we present FaceChain, a personalized portrait generation framework that combines a series of customized image-generation model and a rich set of face-related perceptual understanding models (\eg, face…
-
21 Jul 2023 1 repository listed Syntology ran 4 of 8 samples · 4 unverifiedIn this paper, we propose Subject-Diffusion, a novel open-domain personalized image generation model that, in addition to not requiring test-time fine-tuning, also only requires a single reference image to support…
-
24 May 2023 1 repository listedThen we design a subject representation learning task which enables a diffusion model to leverage such visual representation and generates new subject renditions.
-
17 May 2023 1 repository listed Syntology ran 6 of 17 samples · 11 unverifiedFastComposer proposes delayed subject conditioning in the denoising step to maintain both identity and editability in subject-driven image generation.
Syntology lines on 15 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections