| AI for Cultural Heritage Textiles: Fine-Tuned Latent Diffusion for Novel Ulos Motif Synthesis added by Syntology |
2026-07 (from id) |
CompVis/latent-diffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| SPACE: Your Genomic Profile Predictor is a Powerful DNA Foundation Model |
2 Jun 2025 |
zhujiwei111/space/model/modeling_enformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Diffusion Models without Classifier-free Guidance |
17 Feb 2025 |
lucidrains/classifier-free-guidance-pytorch/classifier_free_guidance_pytorch/typing.py 360032808d3d3b2a |
unverified |
MIT (permissive) |
| MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models |
16 Dec 2024 |
cantbebetter2/mpq-dm/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| SHMT: Self-supervised Hierarchical Makeup Transfer via Latent Diffusion Models |
15 Dec 2024 |
Snowfallingplum/SHMT/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| LAION-SG: An Enhanced Large-Scale Dataset for Training Complex Image-Text Models with Structural Annotations |
11 Dec 2024 |
yangling0818/sgdiff/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| StereoCrafter-Zero: Zero-Shot Stereo Video Generation with Noisy Restart |
21 Nov 2024 |
shijianjian/stereocrafter-zero/lvdm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| MureObjectStitch: Multi-reference Image Composition |
12 Nov 2024 |
bcmi/mureobjectstitch-image-composition/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| Addressing Asynchronicity in Clinical Multimodal Fusion via Individualized Chest X-ray Generation |
23 Oct 2024 |
chenliu-svg/ddl-cxr/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| FlashAudio: Rectified Flows for Fast and High-Fidelity Text-to-Audio Generation |
16 Oct 2024 |
Text-to-Audio/AudioLCM/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| PixWizard: Versatile Image-to-Image Visual Assistant with Open-Language Instructions |
23 Sep 2024 |
afeng-x/pixwizard/models/clip/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Improving Text-guided Object Inpainting with Semantic Pre-inpainting |
12 Sep 2024 |
nnn-s/catdiffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| PID: Physics-Informed Diffusion Model for Infrared Image Generation |
12 Jul 2024 |
fangyuanmao/pid/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Stable Diffusion Segmentation for Biomedical Images with Single-step Reverse Process |
26 Jun 2024 |
lin-tianyu/stable-diffusion-seg/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits |
6 Jun 2024 |
pmzzs/JigMark/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Tactile-Augmented Radiance Fields |
7 May 2024 |
dou-yiming/tarf/img2touch/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| VSTAR: Generative Temporal Nursing for Longer Dynamic Video Synthesis |
20 Mar 2024 |
boschresearch/VSTAR/lvdm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
AGPL-3.0 (copyleft) · pointer only |
| DEADiff: An Efficient Stylization Diffusion Model with Disentangled Representations |
11 Mar 2024 |
bytedance/deadiff/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| pix2gestalt: Amodal Segmentation by Synthesizing Wholes |
25 Jan 2024 |
cvlab-columbia/pix2gestalt/pix2gestalt/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs |
22 Jan 2024 |
CompVis/stable-diffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| Self-Rewarding Language Models |
18 Jan 2024 |
lucidrains/self-rewarding-lm-pytorch/self_rewarding_lm_pytorch/mocks.py 291c6f7cd3e538f1 |
unverified |
MIT (permissive) |
| VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models |
17 Jan 2024 |
ailab-cvc/videocrafter/lvdm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| iFusion: Inverting Diffusion for Pose-Free Reconstruction from Sparse Views |
28 Dec 2023 |
chinhsuanwu/ifusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| The Lottery Ticket Hypothesis in Denoising: Towards Semantic-Driven Initialization |
13 Dec 2023 |
UT-Mao/Initial-Noise-Construction/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| Free3D: Consistent Novel View Synthesis without 3D Representation |
7 Dec 2023 |
lyndonzheng/Free3D/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter |
1 Dec 2023 |
GongyeLiu/StyleCrafter/lvdm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model |
30 Nov 2023 |
zengjianhao/cat-dm/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| MeshGPT: Generating Triangle Meshes with Decoder-Only Transformers |
27 Nov 2023 |
MarcusLoppe/meshgpt-pytorch/meshgpt_pytorch/typing.py 360032808d3d3b2a |
unverified |
MIT (permissive) |
| Adaptive Latent Diffusion Model for 3D Medical Image to Image Translation: Multi-modal Magnetic Resonance Imaging Study |
1 Nov 2023 |
jongdory/aldm/LDM/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Toward effective protection against diffusion based mimicry through score distillation |
2 Oct 2023 |
xavihart/Diff-Protect/code/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors |
18 Oct 2023 |
Doubiiu/DynamiCrafter/lvdm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Uni-paint: A Unified Framework for Multimodal Image Inpainting with Pretrained Diffusion Model |
11 Oct 2023 |
ysy31415/unipaint/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Diffusion Models |
5 Oct 2023 |
ThisisBillhe/EfficientDM/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| PathLDM: Text conditioned Latent Diffusion Model for Histopathology |
1 Sep 2023 |
cvlab-stonybrook/pathldm/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| MatFuse: Controllable Material Generation with Diffusion Models |
22 Aug 2023 |
giuvecchio/matfuse-sd/src/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| ControlCom: Controllable Image Composition using Diffusion Model |
19 Aug 2023 |
bcmi/controlcom-image-composition/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| AltDiffusion: A Multilingual Text-to-Image Diffusion Model |
19 Aug 2023 |
superhero-7/altdiffuson/src/lm/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
no licence file found · pointer only |
| Taming the Power of Diffusion Models for High-Quality Virtual Try-On with Appearance Flow |
11 Aug 2023 |
bcmi/DCI-VTON-Virtual-Try-On/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| AdvDiff: Generating Unrestricted Adversarial Examples using Diffusion Models |
24 Jul 2023 |
EricDai0/advdiff/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Identity-Preserving Aging of Face Images via Latent Diffusion Models |
17 Jul 2023 |
sudban3089/ID-Preserving-Facial-Aging/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution |
12 Jul 2023 |
lucidrains/vit-pytorch/vit_pytorch/na_vit.py 0d6a8b66934a18c5 |
ran · our draft was wrong
|
MIT (permissive) |
| One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds without Per-Shape Optimization |
29 Jun 2023 |
One-2-3-45/One-2-3-45/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Designing a Better Asymmetric VQGAN for StableDiffusion |
7 Jun 2023 |
buxiangzhiren/asymmetric_vqgan/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| ViCo: Plug-and-play Visual Condition for Personalized Text-to-image Generation |
1 Jun 2023 |
haoosz/vico/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| ProSpect: Prompt Spectrum for Attribute-Aware Personalization of Diffusion Models |
25 May 2023 |
zyxElsa/ProSpect/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Align your Latents: High-Resolution Video Synthesis with Latent Diffusion Models |
18 Apr 2023 |
XavierXiao/Dreambooth-Stable-Diffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Inst-Inpaint: Instructing to Remove Objects with Diffusion Models |
6 Apr 2023 |
abyildirim/inst-inpaint/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| A Closer Look at Parameter-Efficient Tuning in Diffusion Models |
31 Mar 2023 |
Xiang-cd/unet-finetune/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Zero-1-to-3: Zero-shot One Image to 3D Object |
20 Mar 2023 |
cvlab-columbia/zero123/zero123/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Ensemble flow reconstruction in the atmospheric boundary layer from spatially limited measurements through latent diffusion models |
1 Mar 2023 |
rybchuk/latent-diffusion-3d-atmospheric-boundary-layer/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| SINE: SINgle Image Editing with Text-to-Image Diffusion Models |
8 Dec 2022 |
zhang-zx/sine/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| InstructPix2Pix: Learning to Follow Image Editing Instructions |
17 Nov 2022 |
xuduo35/InstructPix2Pix/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| High-Resolution Image Editing via Multi-Stage Blended Diffusion |
24 Oct 2022 |
pfnet-research/multi-stage-blended-diffusion/multi-scale-blended-diffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| AudioLM: a Language Modeling Approach to Audio Generation |
7 Sep 2022 |
identical code first harvested elsewhere fe5dd5258046898c |
ran · our draft was wrong
|
licence of this copy not recorded |
| Frido: Feature Pyramid Diffusion for Complex Scene Image Synthesis |
29 Aug 2022 |
davidhalladay/frido/frido/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Blended Latent Diffusion |
6 Jun 2022 |
omriav/blended-latent-diffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| VALHALLA: Visual Hallucination for Machine Translation |
31 May 2022 |
jerryyli/valhalla-nmt/models/hallucinate/dalle.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| SoundStream: An End-to-End Neural Audio Codec |
7 Jul 2021 |
identical code first harvested elsewhere fe5dd5258046898c |
ran · our draft was wrong
|
licence of this copy not recorded |
| LeViT: a Vision Transformer in ConvNet's Clothing for Faster Inference |
2 Apr 2021 |
ahmedelmahy/myownvit/vit_pytorch/levit.py 47332d86fb8b809f |
ran · our draft was wrong
|
MIT (permissive) |
| Zero-Shot Text-to-Image Generation |
24 Feb 2021 |
jam-ing/DALLE-pytorch/dalle_pytorch/dalle_pytorch.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention |
29 Jun 2020 |
identical code first harvested elsewhere dd3db53ae7ad1c85 |
ran · our draft was wrong
|
licence of this copy not recorded |
| GLU Variants Improve Transformer |
12 Feb 2020 |
identical code first harvested elsewhere dd3db53ae7ad1c85 |
ran · our draft was wrong
|
licence of this copy not recorded |
| Efficient Attention: Attention with Linear Complexities |
4 Dec 2018 |
lucidrains/linear-attention-transformer/linear_attention_transformer/linear_attention_transformer.py dd3db53ae7ad1c85 |
ran · our draft was wrong
|
MIT (permissive) |
| arXiv:aaai_28503 |
|
yuhongwei22/MFA/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |
| arXiv:Jiang_Diffuse3D_Wide-Angle_3D_Photography_via_Bilateral_Diffusion_ICCV_2023_paper |
|
yutaojiang1/Diffuse3D/BilateralDiffusion/ldm/modules/x_transformer.py fe5dd5258046898c |
ran · our draft was wrong
|
MIT (permissive) |