| The Reward Was in Your Data All Along: Correcting Flow Matching with Discriminator-Guided RL added by Syntology |
2026-06 (from id) |
microsoft/soc-fine-tuning-sd/src/soc_pipeline_sd.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Rethinking Dataset Distillation for Classification: Do Distilled Sets Outperform Coresets? added by Syntology |
2026-06 (from id) |
jachansantiago/mode_guidance/diffusers/src/diffusers/pipelines/stable_diffusion/pipeline_stable_diffusion_mode_guidance.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Training-Free Safe Text Embedding Guidance for Text-to-Image Diffusion Models added by Syntology |
2025-10 (from id) |
aailab-kaist/STG/pipelines/pipeline_stable_diffusion_stg.py 9855e54881af9165 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| ConceptSplit: Decoupled Multi-Concept Personalization of Diffusion Models via Token-wise Adaptation and Attention Disentanglement added by Syntology |
2025-10 (from id) |
KU-VGI/ConceptSplit/pipeline_sdxl_loda.py 9855e54881af9165 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Knowledge Distillation Detection for Open-weights Models added by Syntology |
2025-10 (from id) |
shqii1j/distillation_detection/text2image/pipeline/stable_diffusion.py 9855e54881af9165 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Steer Away From Mode Collisions: Improving Composition In Diffusion Models added by Syntology |
2025-09 (from id) |
debottam-dutta7/co3/composers/Co3.py 01c6e5b73656deb5 |
unverified |
no licence file found · pointer only |
| T-LoRA: Single Image Diffusion Model Customization Without Overfitting |
8 Jul 2025 |
controlgenai/t-lora/tlora/model/pipeline_sdxl.py 01c6e5b73656deb5 |
unverified |
MIT (permissive) |
| Token Perturbation Guidance for Diffusion Models |
10 Jun 2025 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation |
7 May 2025 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflection |
14 Dec 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale |
9 Dec 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| HunyuanVideo: A Systematic Framework For Large Video Generative Models |
3 Dec 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation |
3 Dec 2024 |
lzhxmu/accdiffusion_v2/accdiffusion_plus.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Stable Flow: Vital Layers for Training-Free Image Editing |
21 Nov 2024 |
snap-research/stable-flow/src/diffusers/pipelines/ledits_pp/pipeline_leditspp_stable_diffusion.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Golden Noise for Diffusion Models: A Learning Framework |
14 Nov 2024 |
xie-lab-ml/golden-noise-for-diffusion-models/training/utils/pipeline_stable_diffusion_xl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Rare-to-Frequent: Unlocking Compositional Generation Power of Diffusion Models on Rare Concepts with LLM Guidance |
29 Oct 2024 |
krafton-ai/rare-to-frequent/R2F_Diffusion_xl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances |
24 Oct 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation |
18 Oct 2024 |
360cvgroup/hico_t2i/diffusers/src/diffusers/pipelines/t2i_adapter/pipeline_stable_diffusion_xl_adapter.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| PUMA: Empowering Unified MLLM with Multi-granular Visual Generation |
17 Oct 2024 |
rongyaofang/puma/image_to_image_pipeline_cfg.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow |
9 Oct 2024 |
g-u-n/rectified-diffusion/pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| IterComp: Iterative Composition-Aware Feedback Learning from Model Gallery for Text-to-Image Generation |
9 Oct 2024 |
yangling0818/rpg-diffusionmaster/RegionalDiffusion_base.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| AP-LDM: Attentive and Progressive Latent Diffusion Model for Training-Free High-Resolution Image Generation |
8 Oct 2024 |
kmittle/ap-ldm/InferencePipelines/FreeScale/pipeline_freescale_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| TweedieMix: Improving Multi-Concept Fusion for Diffusion-based Image/Video Generation |
8 Oct 2024 |
kwongihyun/tweediemix/fusion_generation/fusion_sampling.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| PerCo (SD): Open Perceptual Compression |
30 Sep 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function |
30 Sep 2024 |
I2-Multimedia-Lab/Magnet/pipeline_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction |
26 Sep 2024 |
envision-research/lotus/pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| AudioEditor: A Training-Free Diffusion-Based Audio Editing Framework |
2024-09 (from id) |
nku-hlt/audioeditor/auffusion/auffusion_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts |
4 Sep 2024 |
Liuxinyv/HiPrompt/hiprompt_sdxl_llava.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Smoothed Energy Guidance: Guiding Diffusion Models with Reduced Energy Curvature of Attention |
1 Aug 2024 |
SusungHong/SEG-SDXL/pipeline_seg.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| AccDiffusion: An Accurate Method for Higher-Resolution Image Generation |
15 Jul 2024 |
lzhxmu/AccDiffusion/accdiffusion_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| UltraEdit: Instruction-based Fine-Grained Image Editing at Scale |
7 Jul 2024 |
pkunlp-icler/ultraedit/data_generation/prompt_to_prompt_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance |
26 Jun 2024 |
yhyun225/DiffuseHigh/pipeline_diffusehigh_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection |
11 Jun 2024 |
hyao1/glad/pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Is One GPU Enough? Pushing Image Generation at Higher-Resolutions with Foundation Models |
11 Jun 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Text-to-Image Rectified Flow as Plug-and-Play Priors |
5 Jun 2024 |
gnobitab/instaflow/code/pipeline_rf.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| EasyAnimate: A High-Performance Long Video Generation Method based on Transformer Architecture |
29 May 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Ensembling Diffusion Models via Adaptive Feature Aggregation |
27 May 2024 |
tenvence/afa/models/modules/pipeline_stable_diffusion_aggregator.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding |
14 May 2024 |
tencent/hunyuandit/hydit/diffusion/pipeline_controlnet.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| AID: Attention Interpolation of Text-to-Image Diffusion |
26 Mar 2024 |
qy-h00/attention-interpolation-diffusion/gradio_src/pipeline_interpolated_sd.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation |
25 Mar 2024 |
omer11a/bounded-attention/pipeline_stable_diffusion_xl_opt.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| SemanticDraw: Towards Real-Time Interactive Content Creation from Image Diffusion Models |
14 Mar 2024 |
ironjr/semantic-draw/src/model/pipeline_semantic_draw_kolors.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Improving Diffusion Models for Authentic Virtual Try-on in the Wild |
8 Mar 2024 |
yisol/IDM-VTON/src/tryon_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| NoiseCollage: A Layout-Aware Text-to-Image Diffusion Model Based on Noise Cropping and Merging |
6 Mar 2024 |
univ-esuty/noisecollage/pipeline_custom/pipeline_noise_collage_controlnet.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| ViewDiff: 3D-Consistent Image Generation with Text-to-Image Models |
4 Mar 2024 |
facebookresearch/viewdiff/viewdiff/model/custom_stable_diffusion_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Visual Style Prompting with Swapping Self-Attention |
20 Feb 2024 |
naver-ai/Visual-Style-Prompting/visualize_attention_src/pipeline_stable_diffusion_xl_attn.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation |
16 Feb 2024 |
guolanqing/self-cascade/source_override/pipeline_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation |
6 Feb 2024 |
TIGER-AI-Lab/ConsistI2V/consisti2v/pipelines/pipeline_autoregress_animation.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| FreeStyle: Free Lunch for Text-guided Style Transfer using Diffusion Models |
28 Jan 2024 |
freestylefreelunch/freestyle/diffusers_test/pipeline_stable_diffusion_img2img.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs |
22 Jan 2024 |
YangLing0818/RPG-DiffusionMaster/RegionalDiffusion_base.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Auffusion: Leveraging the Power of Diffusion and Large Language Models for Text-to-Audio Generation |
2 Jan 2024 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| PEEKABOO: Interactive Video Generation via Masked-Diffusion |
12 Dec 2023 |
microsoft/peekaboo/src/models/sd_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding |
7 Dec 2023 |
TencentARC/PhotoMaker/photomaker/pipeline_controlnet.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models |
5 Dec 2023 |
mcg-nju/bivdiff/models/Prompt2Prompt/pipeline_prompt2prompt.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Adversarial Score Distillation: When score distillation meets GAN |
1 Dec 2023 |
2y7c3/asd/threestudio/models/guidance/stable_diffusion_asd_guidance.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| DemoFusion: Democratising High-Resolution Image Generation With No $$$ |
24 Nov 2023 |
PRIS-CV/DemoFusion/pipeline_demofusion_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| LEDITS++: Limitless Image Editing using Text-to-Image Models |
28 Nov 2023 |
ml-research/ledits_pp/src/leditspp/pipeline_stable_diffusion_xl_ledits.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Paragraph-to-Image Generation with Information-Enriched Diffusion Model |
24 Nov 2023 |
weijiawu/paradiffusion/pipeline_stable_diffusion_llama.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models |
10 Oct 2023 |
tencent-ailab/PCDMs/src/pipelines/PCDMs_pipeline.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| InstaFlow: One Step is Enough for High-Quality Diffusion-Based Text-to-Image Generation |
12 Sep 2023 |
gnobitab/InstaFlow/code/pipeline_rf.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Expressive Text-to-Image Generation with Rich Text |
13 Apr 2023 |
songweige/rich-text-to-image/models/region_diffusion_sdxl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Blended Latent Diffusion |
6 Jun 2022 |
identical code first harvested elsewhere bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| arXiv:Ma_DeepCache_Accelerating_Diffusion_Models_for_Free_CVPR_2024_paper |
|
horseee/DeepCache/DeepCache/sdxl/pipeline_stable_diffusion_xl.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:Duan_UNIC-Adapter_Unified_Image-instruction_Adapter_with_Multi-modal_Transformer_for_Image_Generation_CVPR_2025_paper |
|
unity-research/IP-Adapter-Instruct/ip_adapter/pipeline_stable_diffusion_extra_cfg.py bea2d776a332f2b0 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |