| DEMARK: A Query-Free Black-Box Attack on Deepfake Watermarking Defenses added by Syntology |
2026-01 (from id) |
adobe/trustmark/python/trustmark/model.py ef52d142b2c30ac2 |
unverified |
no licence file found · pointer only |
| GSFixer: Improving 3D Gaussian Splatting with Reference-Guided Video Diffusion Priors added by Syntology |
13 Aug 2025 |
GVCLab/GSFixer/Reconstruction/gsfixer/extrapolation/extrapolator.py ef52d142b2c30ac2 |
unverified |
no licence file found · pointer only |
| SCFlow: Implicitly Learning Style and Content Disentanglement with Flow Models added by Syntology |
2025-08 (from id) |
CompVis/SCFlow/scflow/cfm.py 692c94009ed723bc |
unverified |
MIT (permissive) |
| Hita: Holistic Tokenizer for Autoregressive Image Generation |
3 Jul 2025 |
CVMI-Lab/Hita/hita/tokenizer/tokenizer_image/vq_model.py 57f372e8c0555c75 |
unverified |
MIT (permissive) |
| FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving |
23 May 2025 |
MIV-XJTU/FSDrive/MoVQGAN/movqgan/models/vqgan.py 56014014c4ed0790 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Mean Flows for One-step Generative Modeling |
19 May 2025 |
andypinxinliu/GestureLSM/models/config.py 16e41676220876a4 |
unverified |
MIT (permissive) |
| "Principal Components" Enable A New Language of Images |
11 Mar 2025 |
visual-gen/semanticist/semanticist/engine/trainer_utils.py e6197485fe096b58 |
unverified |
MIT (permissive) |
| Recognition-Synergistic Scene Text Editing |
11 Mar 2025 |
ZhengyaoFang/RS-STE/model/model.py f0ac78b4e05be194 |
unverified |
no licence file found · pointer only |
| Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think |
2 Mar 2025 |
Chuge0335/EDG/merge/model.py ef52d142b2c30ac2 |
unverified |
MIT (permissive) |
| OSDFace: One-Step Diffusion Model for Face Restoration |
26 Nov 2024 |
jkwang28/osdface/utils/common.py 192db27a22fa2a3f |
unverified |
no licence file found · pointer only |
| Distillation of Discrete Diffusion through Dimensional Correlations |
11 Oct 2024 |
sony/di4c/maskgit-pytorch/Network/Taming/util.py ef52d142b2c30ac2 |
unverified |
MIT (permissive) |
| UniMuMo: Unified Text, Music and Motion Generation |
6 Oct 2024 |
hanyangclarence/UniMuMo/unimumo/util.py 230615b9581d6e9d |
ran
|
no licence file found · pointer only |
| Temporally Aligned Audio for Video with Autoregression |
20 Sep 2024 |
ilpoviertola/V-AURA/utils/utils.py 692c94009ed723bc |
unverified |
MIT (permissive) |
| 2S-ODIS: Two-Stage Omni-Directional Image Synthesis by Geometric Distortion Correction |
16 Sep 2024 |
islab-sophia/2s-odis/vqgan_recompile.py f0ac78b4e05be194 |
unverified |
Apache-2.0 (permissive) |
| Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation |
7 Sep 2024 |
cplusx/rich_context_l2i/misc_utils/model_utils.py f0ac78b4e05be194 |
unverified |
no licence file found · pointer only |
| Scalable Autoregressive Image Generation with Mamba |
22 Aug 2024 |
hp-l33/aim/util/helper.py f0ac78b4e05be194 |
unverified |
MIT (permissive) |
| Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text |
25 Jun 2024 |
imlixinyang/director3d/modules/unet_hacked.py ef52d142b2c30ac2 |
unverified |
no licence file found · pointer only |
| Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion |
5 Jun 2024 |
Costwen/Ouroboros3D/src/utils/instantiators.py dcbf30e5516a846d |
unverified |
Apache-2.0 (permissive) |
| Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching |
1 Jun 2024 |
cyanbx/Frieren-V2A/Frieren/cfm/logger_maa2.py bacff385a3aaa795 |
unverified |
Apache-2.0 (permissive) |
| Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability |
27 May 2024 |
opendrivelab/vista/vwm/models/diffusion.py f3ff5ab8d2f07e10 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer |
23 May 2024 |
DreamTechAI/Direct3D/direct3d/utils/util.py ef52d142b2c30ac2 |
unverified |
Apache-2.0 (permissive) |
| FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition |
22 May 2024 |
codegoat24/dreamtext/sgm/modules/encoders/modules.py 6d95de949d28670b |
ran · our draft was wrong
|
no licence file found · pointer only |
| PuLID: Pure and Lightning ID Customization via Contrastive Alignment |
24 Apr 2024 |
tothebeginning/pulid/pulid/utils.py 9d7051552589b1be |
ran
|
Apache-2.0 (permissive) |
| InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models |
10 Apr 2024 |
tencentarc/instantmesh/src/utils/train_util.py ef52d142b2c30ac2 |
unverified |
Apache-2.0 (permissive) |
| Finding Visual Task Vectors |
8 Apr 2024 |
alhojel/visual_task_vectors/vqgan.py f0ac78b4e05be194 |
unverified |
no licence file found · pointer only |
| Evolutionary Optimization of Model Merging Recipes |
19 Mar 2024 |
sakanaai/evolutionary-model-merge/evomerge/utils.py 18d8fb19d766c90f |
ran
|
Apache-2.0 (permissive) |
| Lodge: A Coarse to Fine Diffusion Network for Long Dance Generation Guided by the Characteristic Dance Primitives |
15 Mar 2024 |
li-ronghui/LODGE/dld/config.py ef52d142b2c30ac2 |
unverified |
no licence file found · pointer only |
| Beyond Text: Frozen Large Language Models in Visual Signal Comprehension |
12 Mar 2024 |
zh460045050/v2l-tokenizer/models/models_v2l.py f0ac78b4e05be194 |
unverified |
no licence file found · pointer only |
| 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors |
4 Mar 2024 |
3dtopia/3dtopia/module/nn_2d.py 733817d8d0777c77 |
unverified |
Apache-2.0 (permissive) |
| FiT: Flexible Vision Transformer for Diffusion Model |
19 Feb 2024 |
whlzy/fit/fit/utils/utils.py 18d8fb19d766c90f |
ran
|
Apache-2.0 (permissive) |
| Synchformer: Efficient Synchronization from Sparse Cues |
29 Jan 2024 |
v-iashin/sparsesync/utils/utils.py 692c94009ed723bc |
unverified |
MIT (permissive) |
| MLLM-Tool: A Multimodal Large Language Model For Tool Agent Learning |
19 Jan 2024 |
mllm-tool/mllm-tool/code/dataset/utils.py ef52d142b2c30ac2 |
unverified |
MIT (permissive) |
| VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models |
17 Jan 2024 |
ailab-cvc/videocrafter/utils/utils.py ef52d142b2c30ac2 |
unverified |
no licence file found · pointer only |
| iFusion: Inverting Diffusion for Pose-Free Reconstruction from Sparse Views |
28 Dec 2023 |
chinhsuanwu/ifusion/ldm/util.py e1fe8323a9d1d27a |
unverified |
MIT (permissive) |
| StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter |
1 Dec 2023 |
GongyeLiu/StyleCrafter/utils/utils.py ef52d142b2c30ac2 |
unverified |
Apache-2.0 (permissive) |
| Instruct Me More! Random Prompting for Visual In-Context Learning |
7 Nov 2023 |
jackieam/inmemo/vqgan.py f0ac78b4e05be194 |
unverified |
no licence file found · pointer only |
| arXiv:2311.01015 |
2023-11 (from id) |
jpthu17/GraphMotion/GraphMotion/config.py ef52d142b2c30ac2 |
unverified |
Apache-2.0 (permissive) |
| HumanTOMATO: Text-aligned Whole-body Motion Generation |
19 Oct 2023 |
IDEA-Research/HumanTOMATO/OpenTMA/tma/config.py 1baed29ece06b185 |
ran
|
licence not identified · pointer only |
| Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation |
7 Sep 2023 |
zhuqiangLu/AOG-NET-360/utils/config_utils.py bfd972e8e666fa20 |
ran
|
no licence file found · pointer only |
| Dual Associated Encoder for Face Restoration |
14 Aug 2023 |
LIAGM/DAEFR/main_DAEFR.py 949145461ed16b1e |
ran · our draft was wrong
|
no licence file found · pointer only |
| MotionGPT: Human Motion as a Foreign Language |
26 Jun 2023 |
openmotionlab/motiongpt/mGPT/config.py 4c5959826374192f |
unverified |
MIT (permissive) |
| Renderable Neural Radiance Map for Visual Navigation |
1 Mar 2023 |
intelligolabs/Le-RNR-Map/src/models/autoencoder/gsn/generator.py a1df487b3854a798 |
unverified |
no licence file found · pointer only |
| What Makes Good Examples for Visual In-Context Learning? |
31 Jan 2023 |
zhangyuanhan-ai/visual_prompt_retrieval/vqgan.py f0ac78b4e05be194 |
unverified |
CC0-1.0 (permissive) |
| Discrete Contrastive Diffusion for Cross-Modal Music and Image Generation |
15 Jun 2022 |
l-yezhu/cdcd/synthesis/modeling/transformers/diffusion_d2m.py 178128dbd7d43112 |
unverified |
no licence file found · pointer only |
| RestoreFormer: High-Quality Blind Face Restoration from Undegraded Key-Value Pairs |
17 Jan 2022 |
wzhouxiff/restoreformer/RestoreFormer/models/vqgan_v1.py ba07a5ae5fe07704 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Exploration into Translation-Equivariant Image Quantization |
1 Dec 2021 |
wcshin-git/te-vqgan/taming/models/vqgan.py f0ac78b4e05be194 |
unverified |
MIT (permissive) |
| Denoising Diffusion Probabilistic Models |
19 Jun 2020 |
vainf/diff-pruning/ldm_exp/ldm/models/diffusion/ddpm.py ef52d142b2c30ac2 |
unverified |
Apache-2.0 (permissive) |