| Decoupled Vision-Language System for Multimodal Understanding and Generation added by Syntology |
2026-08 (from id) |
YifanXu74/Libra/libra/models/libra/image_tokenizer.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters? added by Syntology |
2026-07 (from id) |
GAIR-Lab/Light-MER/my_affectgpt/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| AI for Cultural Heritage Textiles: Fine-Tuned Latent Diffusion for Novel Ulos Motif Synthesis added by Syntology |
2026-07 (from id) |
CompVis/latent-diffusion/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Multimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM Editing added by Syntology |
2026-07 (from id) |
lab-klc/ScopeEdit/easyeditor/trainer/blip2_models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| ECA: Efficient Continual Alignment for Open-Ended Image-to-Text Generation added by Syntology |
2026-06 (from id) |
Snowball0823/ECA/models/ECA_InternVL/softprompt_internvl.py eb01e8d85b9504de |
ran
|
licence not identified · pointer only |
| Principles and Practice of Deep Representation Learning or A Mathematical Theory of Memory added by Syntology |
2026-06 (from id) |
NeuralCarver/Michelangelo/michelangelo/models/asl_diffusion/asl_diffuser_pl_module.py 4cb732f513d69dfd |
ran · violated contract
|
GPL-3.0 (copyleft) · pointer only |
| Emotion-LLaMAv2 and MMEVerse: A New Framework and Benchmark for Multimodal Emotion Understanding added by Syntology |
2026-01 (from id) |
ooochen-30/Emotion-LLaMA-v2/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| ToxicTextCLIP: Text-Based Poisoning and Backdoor Attacks on CLIP Pre-training added by Syntology |
2025-11 (from id) |
xinyaocse/ToxicTextCLIP/models/model_decode_use_image_as_context.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model added by Syntology |
2025-10 (from id) |
RuipingL/Situat3DChange/SCReasoner/model/utils.py 9d2100678022a09b |
unverified |
CC-BY-4.0 · pointer only |
| Noise-Consistent Siamese-Diffusion for Medical Image Synthesis and Segmentation |
9 May 2025 |
qiukunpeng/siamese-diffusion/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation |
30 Apr 2025 |
pku-yuangroup/holotime/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos |
24 Apr 2025 |
renshuhuai-andy/timechat/timechat/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception |
15 Apr 2025 |
ziqipang/ADDP/referring_segmentation/stable_diffusion/ldm/models/diffusion/ddpm_edit_addp.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Learning Hazing to Dehazing: Towards Realistic Haze Generation for Real-World Image Dehazing |
25 Mar 2025 |
ruiyi-w/learning-hazing-to-dehazing/diffbir/model/cldm.py 566b9df70aa38276 |
ran
|
Apache-2.0 (permissive) |
| Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think |
2 Mar 2025 |
Chuge0335/EDG/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Mitigating Hallucinations in Large Vision-Language Models by Adaptively Constraining Information Flow |
28 Feb 2025 |
jiaqi5598/adavib/minigpt4/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Gramian Multimodal Representation Learning and Alignment |
16 Dec 2024 |
ispamm/GRAM/model/general_module.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Boosting Alignment for Post-Unlearning Text-to-Image Generative Models |
9 Dec 2024 |
identical code first harvested elsewhere 4cb732f513d69dfd |
ran · violated contract
|
licence of this copy not recorded |
| Taming Scalable Visual Tokenizer for Autoregressive Image Generation |
3 Dec 2024 |
tencentarc/open-magvit2/src/Open_MAGVIT2/models/cond_transformer.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| StereoCrafter-Zero: Zero-Shot Stereo Video Generation with Noisy Restart |
21 Nov 2024 |
shijianjian/stereocrafter-zero/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MureObjectStitch: Multi-reference Image Composition |
12 Nov 2024 |
bcmi/mureobjectstitch-image-composition/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance |
4 Nov 2024 |
farewellthree/ppllava/ppllava/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Real-Time Personalization for LLM-based Recommendation with Customized In-Context Learning |
30 Oct 2024 |
ym689/rec_icl/minigpt4/models/rec_model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Addressing Asynchronicity in Clinical Multimodal Fusion via Individualized Chest X-ray Generation |
23 Oct 2024 |
chenliu-svg/ddl-cxr/ldm/models/predict_model.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Enhancing Temporal Modeling of Video LLMs via Time Gating |
8 Oct 2024 |
lavi-lab/tg-vid/stllm/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| PrimeDepth: Efficient Monocular Depth Estimation with a Stable Diffusion Preimage |
13 Sep 2024 |
vislearn/PrimeDepth/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation |
6 Sep 2024 |
identical code first harvested elsewhere 4cb732f513d69dfd |
ran · violated contract
|
licence of this copy not recorded |
| Multi-modal Situated Reasoning in 3D Scenes |
4 Sep 2024 |
identical code first harvested elsewhere 4cb732f513d69dfd |
ran · violated contract
|
licence of this copy not recorded |
| ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis |
3 Sep 2024 |
drexubery/viewcrafter/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models |
25 Aug 2024 |
yejipark-m/convis/minigpt4/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| ParGo: Bridging Vision-Language with Partial and Global Views |
23 Aug 2024 |
bytedance/pargo/pargo/backbone/fusion/minigpt.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| SZTU-CMU at MER2024: Improving Emotion-LLaMA with Conv-Attention for Multimodal Emotion Recognition |
20 Aug 2024 |
zebangcheng/emotion-llama/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| Data Generation Scheme for Thermal Modality with Edge-Guided Adversarial Conditional Diffusion Model |
7 Aug 2024 |
lengmo1996/ECDM/ecdm/models/diffusion/ddpm_condition.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models |
17 Jul 2024 |
jiutian-vl/mome/models/mome_model.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| PID: Physics-Informed Diffusion Model for Infrared Image Generation |
12 Jul 2024 |
identical code first harvested elsewhere 4cb732f513d69dfd |
ran · violated contract
|
licence of this copy not recorded |
| MiniGPT-Med: Large Language Model as a General Interface for Radiology Diagnosis |
4 Jul 2024 |
vision-cair/minigpt-med/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation |
2024-07 (from id) |
picoaudio/picoaudio/picoaudio/audioldm/ldm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Stable Diffusion Segmentation for Biomedical Images with Single-step Reverse Process |
26 Jun 2024 |
lin-tianyu/stable-diffusion-seg/ldm/models/diffusion/SDSeg.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models |
24 Jun 2024 |
arthur-qiu/freetraj/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| VRSBench: A Versatile Vision-Language Benchmark Dataset for Remote Sensing Image Understanding |
18 Jun 2024 |
lavender105/rsgpt/rsgpt/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Binarized Diffusion Model for Image Super-Resolution |
9 Jun 2024 |
zhengchen1999/BI-DiffSR/diffglv/metrics/lpips.py 566b9df70aa38276 |
ran
|
Apache-2.0 (permissive) |
| MeLFusion: Synthesizing Music from Image and Language Cues using Diffusion Models |
7 Jun 2024 |
schowdhury671/melfusion/audioldm/ldm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| LocLLM: Exploiting Generalizable Human Keypoint Localization via Large Language Model |
7 Jun 2024 |
kennethwdk/LocLLM/models/locllm.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt |
6 Jun 2024 |
NY1024/BAP-Jailbreak-Vision-Language-Models-via-Bi-Modal-Adversarial-Prompt/MiniGPT-4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits |
6 Jun 2024 |
pmzzs/JigMark/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Text-like Encoding of Collaborative Information in Large Language Models for Recommendation |
5 Jun 2024 |
zyang1580/binllm/minigpt4/models/rec_model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching |
1 Jun 2024 |
cyanbx/Frieren-V2A/Frieren/cfm/models/diffusion/cfm_scale_cfg.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Prompt-guided Precise Audio Editing with Diffusion Models |
11 May 2024 |
haoheliu/audioldm/audioldm/ldm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis |
31 May 2024 |
PhysGame/PhysGame/physvlm/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| White-box Multimodal Jailbreaks Against Large Vision-Language Models |
28 May 2024 |
roywang021/UMK/minigpt4/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability |
27 May 2024 |
opendrivelab/vista/vwm/models/diffusion.py 694551f4a162c92a |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models |
27 May 2024 |
rupertluo/vocot/model/language_model/volcano_base.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Hawk: Learning to Understand Open-World Video Anomalies |
27 May 2024 |
jqtangust/hawk/hawk/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| ReactXT: Understanding Molecular "Reaction-ship" via Reaction-Contextualized Molecule-Text Pretraining |
23 May 2024 |
syr-cn/reactxt/model/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| ProtT3: Protein-to-Text Generation for Text-based Protein Understanding |
21 May 2024 |
acharkq/ProtT3/model/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| FIFO-Diffusion: Generating Infinite Videos from Text without Training |
19 May 2024 |
jjihwan/FIFO-Diffusion_public/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Libra: Building Decoupled Vision System on Large Language Models |
16 May 2024 |
yifanxu74/libra/libra/models/libra/image_tokenizer.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Adversarial Robustness for Visual Grounding of Multimodal Large Language Models |
16 May 2024 |
KuofengGao/MLLM-Grounding-Robustness/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| Tactile-Augmented Radiance Fields |
7 May 2024 |
dou-yiming/tarf/img2touch/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Inf-DiT: Upsampling Any-Resolution Image with Memory-Efficient Diffusion Transformer |
7 May 2024 |
thudm/inf-dit/dit/model.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound |
30 Apr 2024 |
haoheliu/SemantiCodec-inference/semanticodec/modules/decoder/latent_diffusion/util.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| AnomalyXFusion: Multi-modal Anomaly Synthesis with Diffusion |
30 Apr 2024 |
hujiecpp/mvtec-caption/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MovieChat+: Question-aware Sparse Memory for Long Video Question Answering |
26 Apr 2024 |
rese1f/MovieChat/MovieChat/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| TAVGBench: Benchmarking Text to Audible-Video Generation |
22 Apr 2024 |
opennlplab/tavgbench/audioldm/ldm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| On the Content Bias in Fréchet Video Distance |
18 Apr 2024 |
songweige/tats/tats/tats_transformer.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| BinaryDM: Accurate Weight Binarization for Efficient Diffusion Models |
8 Apr 2024 |
xingyu-zheng/binarydm/ddpm_ours.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models |
2 Apr 2024 |
fudan-zvg/diffusion-square/sgm/util.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| MIPS at SemEval-2024 Task 3: Multimodal Emotion-Cause Pair Extraction in Conversations with Multimodal Language Models |
31 Mar 2024 |
mips-colt/mer-mce/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| ST-LLM: Large Language Models Are Effective Temporal Learners |
30 Mar 2024 |
TencentARC/ST-LLM/stllm/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Generative Multi-modal Models are Good Class-Incremental Learners |
27 Mar 2024 |
DoubleClass/GMM/minigpt4/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| VSTAR: Generative Temporal Nursing for Longer Dynamic Video Synthesis |
20 Mar 2024 |
boschresearch/VSTAR/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
AGPL-3.0 (copyleft) · pointer only |
| Tuning-Free Image Customization with Image and Text Guidance |
19 Mar 2024 |
zrealli/TIGIC/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Subjective-Aligned Dataset and Metric for Text-to-Video Quality Assessment |
18 Mar 2024 |
qmme/t2vqa/model/model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MindEye2: Shared-Subject Models Enable fMRI-To-Image With 1 Hour of Data |
17 Mar 2024 |
medarc-ai/mindeyev2/src/generative_models/sgm/util.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| DEADiff: An Efficient Stylization Diffusion Model with Disentangled Representations |
11 Mar 2024 |
bytedance/deadiff/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering |
8 Mar 2024 |
codegoat24/primecomposer/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| CAT: Enhancing Multimodal Large Language Model to Answer Questions in Dynamic Audio-Visual Scenarios |
7 Mar 2024 |
rikeilong/bay-cat/ADPO_CAT/model/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| 3DTopia: Large Text-to-3D Generation Model with Hybrid Diffusion Priors |
4 Mar 2024 |
3dtopia/3dtopia/model/auto_regressive.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| PLACE: Adaptive Layout-Semantic Fusion for Semantic Image Synthesis |
4 Mar 2024 |
cszy98/place/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Stop Reasoning! When Multimodal LLM with Chain-of-Thought Reasoning Meets Adversarial Image |
22 Feb 2024 |
aipenguin/stopreasoning/minigpt4/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| VisLingInstruct: Elevating Zero-Shot Learning in Multi-Modal Language Models with Autonomous Instruction Optimization |
12 Feb 2024 |
zhudongsheng75/vislinginstruct/vislinginstruct/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MolTC: Towards Molecular Relational Modeling In Language Models |
6 Feb 2024 |
MangoKiller/MolTC/model/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Extreme Two-View Geometry From Object Poses with Diffusion Models |
5 Feb 2024 |
scy639/extreme-two-view-geometry-from-object-poses-with-diffusion-models/src/zero123/zero1/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Unifying Generation and Prediction on Graphs with Latent Graph Diffusion |
4 Feb 2024 |
zhouc20/LatentGraphDiffusion/lgd/ddpm/LGD.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| Jailbreaking Attack against Multimodal Large Language Model |
4 Feb 2024 |
abc03570128/jailbreaking-attack-against-multimodal-large-language-model/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| pix2gestalt: Amodal Segmentation by Synthesizing Wholes |
25 Jan 2024 |
cvlab-columbia/pix2gestalt/pix2gestalt/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Towards 3D Molecule-Text Interpretation in Language Models |
25 Jan 2024 |
lsh0520/3d-molm/model/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs |
22 Jan 2024 |
CompVis/stable-diffusion/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models |
17 Jan 2024 |
ailab-cvc/videocrafter/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| GroundingGPT:Language Enhanced Multi-modal Grounding Model |
11 Jan 2024 |
lzw-lzw/groundinggpt/video_llama/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones |
28 Dec 2023 |
dlyuangod/tinygpt-v/TinyGPT-V-main/minigpt4/models/base_model.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |
| When Parameter-efficient Tuning Meets General-purpose Vision-language Models |
16 Dec 2023 |
melonking32/petal/lavis/models/blip2_models/blip2_petal_moe.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| The Lottery Ticket Hypothesis in Denoising: Towards Semantic-Driven Initialization |
13 Dec 2023 |
UT-Mao/Initial-Noise-Construction/ldm/models/diffusion/classifier.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| Learned representation-guided diffusion models for large-image generation |
12 Dec 2023 |
cvlab-stonybrook/large-image-diffusion/ldm/models/diffusion/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image |
12 Dec 2023 |
POSTECH-CVLab/nvsadapter/sgm/modules/nvsadapter/midas/api.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| ControlNet-XS: Rethinking the Control of Text-to-Image Diffusion Models as Feedback-Control Systems |
11 Dec 2023 |
vislearn/ControlNet-XS/sgm/util.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| Free3D: Consistent Novel View Synthesis without 3D Representation |
7 Dec 2023 |
lyndonzheng/Free3D/models/ddpm.py 4cb732f513d69dfd |
ran · violated contract
|
no licence file found · pointer only |
| MotionCtrl: A Unified and Flexible Motion Controller for Video Generation |
6 Dec 2023 |
TencentARC/MotionCtrl/lvdm/basics.py 4cb732f513d69dfd |
ran · violated contract
|
Apache-2.0 (permissive) |
| GPT4Point: A Unified Framework for Point-Language Understanding and Generation |
5 Dec 2023 |
Pointcept/GPT4Point/lavis/models/gpt4point_models/gpt4point.py 4cb732f513d69dfd |
ran · violated contract
|
MIT (permissive) |
| TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding |
4 Dec 2023 |
lntzm/cvpr24track-longvideo/timechat/models/blip2.py 4cb732f513d69dfd |
ran · violated contract
|
BSD-3-Clause (permissive) |