| Restoring Without Forgetting: Continual Learning Across Image Degradations added by Syntology |
2026-08 (from id) |
AlifAshrafee/Restoring-Without-Forgetting/restormer/Motion_Deblurring/generate_patches_degradations.py 4ddda92558c55c31 |
ran
|
no licence file found · pointer only |
| UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing added by Syntology |
2026-04 (from id) |
Yunkaidang/UHR/with-SAM/longva/longva/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| VidLaDA: Bidirectional Diffusion Large Language Models for Efficient Video Understanding added by Syntology |
2026-01 (from id) |
ziHoHe/VidLaDA/train/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| XRefine: Attention-Guided Keypoint Match Refinement added by Syntology |
2026-01 (from id) |
boschresearch/xrefine/model.py 6c29c5eb23e27131 |
ran · fixture could not drive it
fingerprinted |
AGPL-3.0 (copyleft) · pointer only |
| Cross-Layer Injection for Deep Vision-Language Fusion added by Syntology |
2026-01 (from id) |
codefuse-ai/CLI/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs |
27 Jun 2025 |
HumanMLLM/LLaVA-Scissor/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Learning from Videos for 3D World: Enhancing MLLMs with 3D Vision Geometry Priors |
30 May 2025 |
LaVi-Lab/Video-3D-LLM/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| LaViDa: A Large Diffusion Language Model for Multimodal Understanding |
22 May 2025 |
jacklishufan/lavida/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation |
20 May 2025 |
opengvlab/videochat-flash/llava-train_videochat/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
MIT (permissive) |
| SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL |
15 Apr 2025 |
wdrink/simplear/simpar/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
MIT (permissive) |
| OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model |
30 Mar 2025 |
DriveVLA/OpenDriveVLA/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Keyframe-oriented Vision Token Pruning: Enhancing Efficiency of Large Vision Language Models on Long-Form Video Processing |
13 Mar 2025 |
1999Lyd/KVTP/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
MIT (permissive) |
| LLaVA-UHD v2: an MLLM Integrating High-Resolution Feature Pyramid via Hierarchical Window Transformer |
18 Dec 2024 |
thunlp/llava-uhd/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Lyra: An Efficient and Speech-Centric Framework for Omni-Cognition |
12 Dec 2024 |
dvlab-research/Lyra/lyra/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| DriveMM: All-in-One Large Multimodal Model for Autonomous Driving |
10 Dec 2024 |
zhijian11/DriveMM/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| FlashSloth: Lightning Multimodal Large Language Models via Embedded Visual Compression |
5 Dec 2024 |
codefanw/flashsloth/flashsloth/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
no licence file found · pointer only |
| AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning |
4 Dec 2024 |
lavi-lab/aim/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models |
22 Nov 2024 |
kd-tao/dycoke/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models |
21 Nov 2024 |
dongyh20/insight-v/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
no licence file found · pointer only |
| CCExpert: Advancing MLLM Capability in Remote Sensing Change Captioning with Difference-Aware Integration and a Foundational Dataset |
18 Nov 2024 |
meize0729/ccexpert/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| Number it: Temporal Grounding Videos like Flipping Manga |
15 Nov 2024 |
yongliang-wu/numpro/longva/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
MIT (permissive) |
| AEROBLADE: Training-Free Detection of Latent Diffusion Images Using Autoencoder Reconstruction Error |
31 Jan 2024 |
jonasricker/aeroblade/src/aeroblade/image.py 19475527a36172a4 |
ran
|
no licence file found · pointer only |
| Learning Prior Feature and Attention Enhanced Image Inpainting |
3 Aug 2022 |
ewrfcas/MAE-FAR/ACR/networks/generators.py d1f2e1d8d9f12c48 |
ran · fixture could not drive it
fingerprinted |
licence not identified · pointer only |
| MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer |
5 Oct 2021 |
godhj93/MobileViT/utils/nets/MobileViT.py f4179768dbc55e61 |
ran
|
no licence file found · pointer only |
| Sparse Coding Frontend for Robust Neural Networks |
12 Apr 2021 |
canbakiskan/sparse_coding_frontend/src/learn_patch_dict.py 5b1bb8f496d81deb |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| ArtFlow: Unbiased Image Style Transfer via Reversible Neural Flows |
31 Mar 2021 |
pkuanjie/ArtFlow/glow_decorator.py 7ec1f6c1711259d0 |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Group Equivariant Stand-Alone Self-Attention For Vision |
2 Oct 2020 |
dwromero/g_selfatt/g_selfatt/nn/group_self_attention.py 2f313da6e1ea1bd2 |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| Instance Selection for GANs |
30 Jul 2020 |
snap-research/3dgp/src/training/loss.py 83050978b92b849d |
ran · fixture could not drive it
|
no licence file found · pointer only |
| Multi-View Optimization of Local Feature Geometry |
18 Mar 2020 |
mihaidusmanu/local-feature-refinement/two-view-refinement/refinement.py e2a8ccc8ce3b8830 |
unverified |
BSD-3-Clause (permissive) |
| Dynamic Routing Between Capsules |
26 Oct 2017 |
ameliajimenez/capsule-networks-medical-data-challenges/preprocess_diaretdb1.py cd1f1b152fc81869 |
unverified |
MIT (permissive) |
| Dynamic Routing Between Capsules |
26 Oct 2017 |
ameliajimenez/capsule-networks-medical-data-challenges/preprocess_tupac16.py 9609ff43029770e8 |
unverified |
MIT (permissive) |
| Iterative Gaussianization: from ICA to Random Rotations |
31 Jan 2016 |
jejjohnson/rbig/rbig/_src/image.py 202feb4889129e03 |
unverified |
MIT (permissive) |
| RENOIR - A Dataset for Real Low-Light Image Noise Reduction |
29 Sep 2014 |
Aftaab99/DenoisingAutoencoder/create_datasets.py 570b2b72291f1715 |
unverified |
MIT (permissive) |
| arXiv:Zhong_AIM_Adaptive_Inference_of_Multi-Modal_LLMs_via_Token_Merging_and_ICCV_2025_paper |
|
LaVi-Lab/AIM/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |
| arXiv:2025.findings-acl.359 |
|
G-JWLee/TAMP/llava/mm_utils.py b0a851cae92754ff |
ran · fixture could not drive it
|
Apache-2.0 (permissive) |