| LUCoS: Latent Unsupervised Context Selection for Tabular Foundation Models added by Syntology |
2026-05 (from id) |
ari-dasci/S-LUCoS/src/lucos/utils/config_utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Spectral Collapse in Diffusion Inversion added by Syntology |
2026-02 (from id) |
nicoboou/ovg/utils/helpers.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| DEMARK: A Query-Free Black-Box Attack on Deepfake Watermarking Defenses added by Syntology |
2026-01 (from id) |
adobe/trustmark/python/trustmark/model.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| GSFixer: Improving 3D Gaussian Splatting with Reference-Guided Video Diffusion Priors added by Syntology |
13 Aug 2025 |
GVCLab/GSFixer/Reconstruction/gsfixer/extrapolation/extrapolator.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Hita: Holistic Tokenizer for Autoregressive Image Generation |
3 Jul 2025 |
CVMI-Lab/Hita/hita/tokenizer/tokenizer_image/vq_model.py ce0cea52e4db5e7c |
ran · our draft was wrong
|
MIT (permissive) |
| Hunyuan3D 2.5: Towards High-Fidelity 3D Assets Generation with Ultimate Details |
19 Jun 2025 |
tencent/hunyuan3d-2/hy3dgen/shapegen/pipelines.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Mean Flows for One-step Generative Modeling |
19 May 2025 |
andypinxinliu/GestureLSM/models/config.py 6f951d437f9f5c5f |
unverified |
MIT (permissive) |
| "Principal Components" Enable A New Language of Images |
11 Mar 2025 |
visual-gen/semanticist/semanticist/engine/trainer_utils.py 1be6845a690a8075 |
unverified |
MIT (permissive) |
| Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think |
2 Mar 2025 |
Chuge0335/EDG/merge/model.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| OSDFace: One-Step Diffusion Model for Face Restoration |
26 Nov 2024 |
jkwang28/osdface/utils/common.py 96bb2bca43d02a51 |
unverified |
no licence file found · pointer only |
| Addressing Representation Collapse in Vector Quantized Models with One Linear Layer |
4 Nov 2024 |
youngsheen/SimVQ/evaluation.py f36897ad5e379c12 |
ran · our draft was wrong
|
MIT (permissive) |
| HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation |
18 Oct 2024 |
360cvgroup/hico_t2i/train_hico.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Distillation of Discrete Diffusion through Dimensional Correlations |
11 Oct 2024 |
sony/di4c/maskgit-pytorch/Network/Taming/util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| UniMuMo: Unified Text, Music and Motion Generation |
6 Oct 2024 |
hanyangclarence/UniMuMo/unimumo/util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Temporally Aligned Audio for Video with Autoregression |
20 Sep 2024 |
ilpoviertola/V-AURA/utils/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| 2S-ODIS: Two-Stage Omni-Directional Image Synthesis by Geometric Distortion Correction |
16 Sep 2024 |
islab-sophia/2s-odis/vqgan_recompile.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation |
7 Sep 2024 |
cplusx/rich_context_l2i/misc_utils/model_utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Scalable Autoregressive Image Generation with Mamba |
22 Aug 2024 |
hp-l33/aim/util/helper.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text |
25 Jun 2024 |
imlixinyang/director3d/modules/unet_hacked.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion |
5 Jun 2024 |
Costwen/Ouroboros3D/src/utils/instantiators.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching |
1 Jun 2024 |
cyanbx/Frieren-V2A/Frieren/cfm/logger_maa2.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability |
27 May 2024 |
opendrivelab/vista/vwm/models/diffusion.py 30d11b7b4428ab6e |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Learning to Discretize Denoising Diffusion ODEs |
24 May 2024 |
vinhsuhi/ld3/models/latent_diff.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer |
23 May 2024 |
DreamTechAI/Direct3D/direct3d/utils/util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition |
22 May 2024 |
codegoat24/dreamtext/sgm/modules/encoders/modules.py 16316ddd38304552 |
ran · our draft was wrong
|
no licence file found · pointer only |
| PuLID: Pure and Lightning ID Customization via Contrastive Alignment |
24 Apr 2024 |
tothebeginning/pulid/pulid/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models |
10 Apr 2024 |
tencentarc/instantmesh/src/utils/train_util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Finding Visual Task Vectors |
8 Apr 2024 |
alhojel/visual_task_vectors/vqgan.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Lodge: A Coarse to Fine Diffusion Network for Long Dance Generation Guided by the Characteristic Dance Primitives |
15 Mar 2024 |
li-ronghui/LODGE/dld/config.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Beyond Text: Frozen Large Language Models in Visual Signal Comprehension |
12 Mar 2024 |
zh460045050/v2l-tokenizer/models/models_v2l.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| FiT: Flexible Vision Transformer for Diffusion Model |
19 Feb 2024 |
whlzy/fit/fit/utils/utils.py 30d11b7b4428ab6e |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Synchformer: Efficient Synchronization from Sparse Cues |
29 Jan 2024 |
v-iashin/sparsesync/utils/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| MLLM-Tool: A Multimodal Large Language Model For Tool Agent Learning |
19 Jan 2024 |
mllm-tool/mllm-tool/code/dataset/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| iFusion: Inverting Diffusion for Pose-Free Reconstruction from Sparse Views |
28 Dec 2023 |
chinhsuanwu/ifusion/ldm/util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models |
21 Dec 2023 |
picsart-ai-research/hd-painter/src/models/common.py c485485457d21a43 |
ran
|
MIT (permissive) |
| NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image |
12 Dec 2023 |
POSTECH-CVLab/nvsadapter/3drec/threestudio/models/guidance/nvsadapter_guidance.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Instruct Me More! Random Prompting for Visual In-Context Learning |
7 Nov 2023 |
jackieam/inmemo/vqgan.py 221b2d116fdf1032 |
ran · our draft was wrong
|
no licence file found · pointer only |
| arXiv:2311.01015 |
2023-11 (from id) |
jpthu17/GraphMotion/GraphMotion/config.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| HumanTOMATO: Text-aligned Whole-body Motion Generation |
19 Oct 2023 |
IDEA-Research/HumanTOMATO/OpenTMA/tma/config.py 4607de5098f9dd03 |
ran
|
licence not identified · pointer only |
| Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation |
7 Sep 2023 |
zhuqiangLu/AOG-NET-360/utils/config_utils.py 74608b08edaba565 |
ran
|
no licence file found · pointer only |
| Dual Associated Encoder for Face Restoration |
14 Aug 2023 |
LIAGM/DAEFR/main_DAEFR.py 4e4f3bef7e8672e8 |
ran · our draft was wrong
|
no licence file found · pointer only |
| MotionGPT: Human Motion as a Foreign Language |
26 Jun 2023 |
openmotionlab/motiongpt/mGPT/config.py 9877bd3cecd6081f |
unverified |
MIT (permissive) |
| LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation |
30 Mar 2023 |
dcdcvgroup/layout-diffusion-mindspore/layout_diffusion/util.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Renderable Neural Radiance Map for Visual Navigation |
1 Mar 2023 |
intelligolabs/Le-RNR-Map/src/models/autoencoder/gsn/generator.py dd0d055075f1b9e3 |
ran
|
no licence file found · pointer only |
| What Makes Good Examples for Visual In-Context Learning? |
31 Jan 2023 |
zhangyuanhan-ai/visual_prompt_retrieval/vqgan.py 221b2d116fdf1032 |
ran · our draft was wrong
|
CC0-1.0 (permissive) |
| BBDM: Image-to-image Translation with Brownian Bridge Diffusion Models |
16 May 2022 |
egshkim/ConditionalBBDM-for-VHR-SAR-to-Optical/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image |
17 Mar 2022 |
xrenaa/look-outside-room/evaluation/evaluate_mp3d.py 51c3260587dcdde1 |
ran · our draft was wrong
|
no licence file found · pointer only |
| RestoreFormer: High-Quality Blind Face Restoration from Undegraded Key-Value Pairs |
17 Jan 2022 |
identical code first harvested elsewhere 4e4f3bef7e8672e8 |
ran · our draft was wrong
|
licence of this copy not recorded |
| Exploration into Translation-Equivariant Image Quantization |
1 Dec 2021 |
wcshin-git/te-vqgan/taming/models/vqgan.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |
| Denoising Diffusion Probabilistic Models |
19 Jun 2020 |
vainf/diff-pruning/ldm_exp/ldm/models/diffusion/ddpm.py 221b2d116fdf1032 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:aaai_29913 |
|
clovaai/TVQ-VAE/image_generation_t/make_samples_tvq.py 4e4f3bef7e8672e8 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| arXiv:Xu_Versatile_Diffusion_Text_Images_and_Variations_All_in_One_Diffusion_ICCV_2023_paper |
|
SHI-Labs/Versatile-Diffusion/lib/utils.py 221b2d116fdf1032 |
ran · our draft was wrong
|
MIT (permissive) |