| EAMamba: Efficient All-Around Vision State Space Model for Image Restoration |
27 Jun 2025 |
daidaijr/EAMamba/profiling/erf_viz.py 1940f3433be25805 |
unverified |
no licence file found · pointer only |
| Pretrained Image-Text Models are Secretly Video Captioners |
19 Feb 2025 |
chunhuizng/mllm-video-captioner/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| Generalizable Human Gaussians for Sparse View Synthesis |
17 Jul 2024 |
humansensinglab/Generalizable-Human-Gaussians/lib/ghg/human_loader.py d73cfad5ab886ac4 |
ran
|
no licence file found · pointer only |
| MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos |
23 May 2024 |
YihongSun/MOD-UV/moduv/utils.py 0d5bc119fb18b186 |
ran
|
MIT (permissive) |
| Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards |
22 May 2024 |
XiaoyuYoung/ConceptDriftMLLMs/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| MVSGaussian: Fast Generalizable Gaussian Splatting Reconstruction from Multi-View Stereo |
20 May 2024 |
TQTQliu/MVSGaussian/fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding |
8 Apr 2024 |
boheumd/MA-LMM/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
MIT (permissive) |
| From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language Models |
1 Apr 2024 |
shtuplus/pix2grp_cvpr2024/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| Omni-Recon: Harnessing Image-based Rendering for General-Purpose Neural Radiance Fields |
17 Mar 2024 |
GATECH-EIC/Omni-Recon/evaluation/tsdf_fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| UFORecon: Generalizable Sparse-View Surface Reconstruction from Arbitrary and UnFavOrable Sets |
8 Mar 2024 |
Youngju-Na/UFORecon/tsdf_fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| Depth Information Assisted Collaborative Mutual Promotion Network for Single Image Dehazing |
2 Mar 2024 |
zhoushen1/dcmpnet/utils/common.py 22d056c672f87531 |
unverified |
MIT (permissive) |
| Shot2Story20K: A New Benchmark for Comprehensive Understanding of Multi-shot Videos |
16 Dec 2023 |
bytedance/Shot2Story/code/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
no licence file found · pointer only |
| LMDrive: Closed-Loop End-to-End Driving with Large Language Models |
12 Dec 2023 |
opendilab/lmdrive/LAVIS/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
Apache-2.0 (permissive) |
| Localized Symbolic Knowledge Distillation for Visual Commonsense Models |
8 Dec 2023 |
jamespark3922/lskd/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation |
4 Dec 2023 |
Magicboomliu/Accelerator-Simple-Template/dataloader/file_io.py a5edaeb9db495d4c |
unverified |
MIT (permissive) |
| AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors |
26 Oct 2023 |
nctu-eva-lab/antifakeprompt/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| Random Sub-Samples Generation for Self-Supervised Real Image Denoising |
31 Jul 2023 |
p1y2z3/sdap/utils/loader.py 6d55b07ac4659765 |
ran
|
MIT (permissive) |
| Bootstrapping Vision-Language Learning with Decoupled Language Pre-training |
13 Jul 2023 |
yiren-jian/BLIText/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| Self-Chained Image-Language Model for Video Localization and Question Answering |
11 May 2023 |
yui010206/sevila/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| VPGTrans: Transfer Visual Prompt Generator across LLMs |
2 May 2023 |
VPGTrans/VPGTrans/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| LKD-Net: Large Kernel Convolution Network for Single Image Dehazing |
5 Sep 2022 |
swu-cs-medialab/lkd-net/utils/common.py 22d056c672f87531 |
unverified |
MIT (permissive) |
| NTIRE 2022 Challenge on Efficient Super-Resolution: Methods and Results |
11 May 2022 |
ofsoundof/imdn/utils/utils_image.py 2d5193a49edc8787 |
unverified |
MIT (permissive) |
| Vision Transformers for Single Image Dehazing |
8 Apr 2022 |
IDKiro/DehazeFormer/utils/common.py 22d056c672f87531 |
unverified |
MIT (permissive) |
| Rethinking Depth Estimation for Multi-View Stereo: A Unified Representation |
5 Jan 2022 |
prstrive/unimvsnet/filter/dypcd.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| FEAR: Fast, Efficient, Accurate and Robust Visual Tracker |
15 Dec 2021 |
pinatafarms/feartracker/model_training/dataset/utils.py b513ca7bc794cb14 |
unverified |
MIT (permissive) |
| Self-attention Does Not Need $O(n^2)$ Memory |
10 Dec 2021 |
jihaonew/mm-instruct/llava/eval/model_vqa_loader.py 44fab12a908ab057 |
unverified |
Apache-2.0 (permissive) |
| AA-RMVSNet: Adaptive Aggregation Recurrent Multi-view Stereo Network |
9 Aug 2021 |
qt-zhu/aa-rmvsnet/fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| PPR10K: A Large-Scale Portrait Photo Retouching Dataset with Human-Region Mask and Group-Level Consistency |
19 May 2021 |
csjliang/PPR10K/code_3DLUT/datasets_GLC.py 1eb64e103338a8c4 |
unverified |
Apache-2.0 (permissive) |
| MVS2D: Efficient Multi-view Stereo via Attention-Driven 2D Convolutions |
27 Apr 2021 |
zhenpeiyang/MVS2D/patchmatch_fusion.py b300805ea5cbcf18 |
unverified |
MIT (permissive) |
| Anomaly localization by modeling perceptual features |
12 Aug 2020 |
xiahaifeng1995/FAVAE-anomaly-detection-localization-master/datasets/preprocessing.py 44fd832b6d2537c9 |
unverified |
Apache-2.0 (permissive) |
| Dense Hybrid Recurrent Multi-view Stereo Net with Dynamic Consistency Checking |
21 Jul 2020 |
yhw-yhw/D2HC-RMVSNet/fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| A Self-supervised Approach for Adversarial Robustness |
8 Jun 2020 |
mshane911/NRP/utils.py 349ac8f433eb384a |
unverified |
MIT (permissive) |
| Supervised Contrastive Learning |
23 Apr 2020 |
alexk1704/scclv2/src/cl_replay/api/utils/convert_ds.py 1fd4a495ba449048 |
unverified |
MIT (permissive) |
| YOLOv4: Optimal Speed and Accuracy of Object Detection |
23 Apr 2020 |
samson6460/tf2_YOLO/utils/tools.py b31c8ccb6cfbb9fe |
ran · metamorphic tier: well formed
|
no licence file found · pointer only |
| Rethinking Atrous Convolution for Semantic Image Segmentation |
17 Jun 2017 |
samson6460/tf2_Segmentation/utils/data_processing.py b5666228355ff947 |
ran · metamorphic tier: well formed
|
no licence file found · pointer only |
| arXiv:aaai_25291 |
|
CGLab-GIST/RIDFnF/codes/utils_image.py 2d5193a49edc8787 |
unverified |
BSD-3-Clause (permissive) |
| arXiv:Yang_TINC_Tree-Structured_Implicit_Neural_Compression_CVPR_2023_paper |
|
RichealYoung/TINC/utils/tool.py 703d32f89cb1bf08 |
unverified |
MIT (permissive) |
| arXiv:Ren_VolRecon_Volume_Rendering_of_Signed_Ray_Distance_Functions_for_Generalizable_CVPR_2023_paper |
|
IVRL/VolRecon/tsdf_fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| arXiv:Ding_TransMVSNet_Global_Context-Aware_Multi-View_Stereo_Network_With_Transformers_CVPR_2022_paper |
|
MegviiRobot/TransMVSNet/dynamic_fusion.py a1db096963f65ed2 |
ran
|
MIT (permissive) |
| arXiv:2024.emnlp-main.88 |
|
WenjianDing/ReBo/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause (permissive) |
| arXiv:2024.emnlp-main.722 |
|
xyh97/UNICORN/app/calculate_coco_features.py 13a43a7d815937e6 |
ran
|
BSD-3-Clause recorded; this copy not marked cleared · pointer only |