| CERF: Communication-Efficient and Retraining-Free Collaborative Perception added by Syntology |
2026-09 (from id) |
uestchjw/CERF/collaborative_perception/bevfusion/projects/postprocess.py 8878ff203c312549 |
unverified |
no licence file found · pointer only |
| SAM3-LoRA: Parameter-Efficient Adaptation of a Concept-Promptable Foundation Model for Multi-Class Structural Defect Segmentation added by Syntology |
2026-09 (from id) |
Sompote/sam3_lora/train_sam3_lora.py 1ba2f6771f6d4651 |
unverified |
no licence file found · pointer only |
| CAPruner: Conceptual-Adjacent Scene Graph Pruner for Enhancing 3D Spatial Reasoning of Large Language Models added by Syntology |
2026-06 (from id) |
fz-zsl/CAPruner/model/dataset.py 6274aff367fc9689 |
ran
|
GPL-3.0 (copyleft) · pointer only |
| Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization added by Syntology |
2026-06 (from id) |
phoenixnir/FLAME/FLAME/utils/metrics.py f4a6aebcbf478f03 |
ran
|
no licence file found · pointer only |
| Lowering the Barrier to IREX Participation: Open-Source Algorithms, Toolkit, and Benchmarking for Iris Recognition added by Syntology |
2026-05 (from id) |
CVRL/PBM/mrcnn/utils.py ae0c5e015d8cf734 |
ran
|
BSD-2-Clause (permissive) |
| ELA: Exact Linear Attention with Qualitative Memory and Hyper-Link added by Syntology |
2026-05 (from id) |
yauntyour/Exact-Linear-Attention/latdet/model.py 73f71709991d844e |
ran
fingerprinted |
licence not identified · pointer only |
| MedCore: Boundary-Preserving Medical Core Pruning for MedSAM added by Syntology |
2026-05 (from id) |
cenweizhang/MedCore/medcore_pruning/metrics.py c3ff88b7bf97c76a |
unverified |
Apache-2.0 (permissive) |
| SOVABench: A Vehicle Surveillance Action Retrieval Benchmark for Multimodal Large Language Models added by Syntology |
2026-01 (from id) |
oriol-rabasseda/sovabench/src/mllm_embeddings_sovabench/datasets/correct_overlapping.py 98a9cd0eca34704b |
unverified |
no licence file found · pointer only |
| GUI-Spotlight: Adaptive Iterative Focus Refinement for Enhanced GUI Visual Grounding added by Syntology |
5 Oct 2025 |
bin123apple/GUI_Spotlight/spotlight/reward/tool_rubric.py 4e08ab7b4f176f64 |
unverified |
MIT (permissive) |
| UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning |
29 May 2025 |
showlab/unirl/evaluate_imgs.py 2f96ce289cb03a29 |
unverified |
Apache-2.0 (permissive) |
| OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation |
26 May 2025 |
PKU-YuanGroup/ConsisID/data_preprocess/step3_get_refine_track.py 02fb438e848b5140 |
unverified |
Apache-2.0 (permissive) |
| VisionReasoner: Unified Visual Perception and Reasoning via Reinforcement Learning |
17 May 2025 |
dvlab-research/Seg-Zero/evaluation_scripts/evaluation.py faee26847645d13b |
unverified |
Apache-2.0 (permissive) |
| VisionReasoner: Unified Visual Perception and Reasoning via Reinforcement Learning |
17 May 2025 |
dvlab-research/VisionReasoner/evaluation/evaluation_anomaly.py 611c245861107d30 |
unverified |
Apache-2.0 (permissive) |
| Perception-R1: Pioneering Perception Policy with Reinforcement Learning |
10 Apr 2025 |
linkangheng/pr1/eval/evaluate_grounding.py 96d8ce55eedcc334 |
ran · fixture could not drive it
fingerprinted |
Apache-2.0 (permissive) |
| SADG: Segment Any Dynamic Gaussian Without Object Trackers |
28 Nov 2024 |
yunjinli/SADG-SegmentAnyDynamicGaussian/metrics_segmentation.py c68803bf7444deb3 |
unverified |
MIT (permissive) |
| SAMPart3D: Segment Any Part in 3D Objects |
11 Nov 2024 |
pointcept/sampart3d/PartObjaverse-Tiny/eval/eval_part.py bfcfa7e485c98dc3 |
unverified |
MIT (permissive) |
| IAA: Inner-Adaptor Architecture Empowers Frozen Large Language Model with Multimodal Capabilities |
23 Aug 2024 |
360cvgroup/inner-adaptor-architecture/iaa/eval/compute_precision.py d7c36fd89f5fcad4 |
ran
|
Apache-2.0 (permissive) |
| Commonsense Prototype for Outdoor Unsupervised 3D Object Detection |
25 Apr 2024 |
hailanyi/CPD/cpd/utils/bbloss.py f940cdda72e1a568 |
ran
|
no licence file found · pointer only |
| FastSAM3D: An Efficient Segment Anything Model for 3D Volumetric Medical Images |
14 Mar 2024 |
arcadelab/fastsam3d/val_2d.py 04b1148082a2d310 |
ran
|
Apache-2.0 (permissive) |
| Car Damage Detection and Patch-to-Patch Self-supervised Image Alignment |
11 Mar 2024 |
2000222/car-damage-detectionv0/mrcnn/utils.py ae0c5e015d8cf734 |
ran
|
no licence file found · pointer only |
| YOLO-World: Real-Time Open-Vocabulary Object Detection |
30 Jan 2024 |
ibaiGorordo/ONNX-YOLO-World-Open-Vocabulary-Object-Detection/yoloworld/nms.py 21a47905e98284b9 |
unverified |
MIT (permissive) |
| DDMI: Domain-Agnostic Latent Diffusion Models for Synthesizing High-Quality Implicit Neural Representations |
23 Jan 2024 |
mlvlab/DDMI/convocc/src/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans? |
31 Oct 2023 |
vl-illusion/dataset/utils.py 2591cb66fe714279 |
ran
fingerprinted |
no licence file found · pointer only |
| Woodpecker: Hallucination Correction for Multimodal Large Language Models |
24 Oct 2023 |
bradyfu/woodpecker/models/utils.py 667ecbfb72739567 |
ran
fingerprinted |
no licence file found · pointer only |
| What Do Deep Saliency Models Learn about Visual Attention? |
14 Oct 2023 |
szzexpoi/saliency_analysis/prototype_dissection.py 423681206c2b5298 |
unverified |
no licence file found · pointer only |
| Towards Content-based Pixel Retrieval in Revisited Oxford and Paris |
11 Sep 2023 |
anguoyuan/pixel_retrieval-segmented_instance_retrieval/evaluation_code/utils/iou_compute.py 5c1d039685bafbdd |
ran
|
MIT (permissive) |
| Towards Content-based Pixel Retrieval in Revisited Oxford and Paris |
11 Sep 2023 |
anguoyuan/pixel_retrieval-segmented_instance_retrieval/evaluation_code/utils/iou_compute_delg.py 1a23c5e4a34671bf |
ran
|
MIT (permissive) |
| 3D Implicit Transporter for Temporally Consistent Keypoint Discovery |
10 Sep 2023 |
zhongcl-thu/3D-Implicit-Transporter/core/nets/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Hierarchical Video-Moment Retrieval and Step-Captioning |
29 Mar 2023 |
j-min/HiREST/evaluate.py e84b722650b81091 |
unverified |
MIT (permissive) |
| Unsupervised Inference of Signed Distance Functions from Single Sparse Point Clouds without Learning Priors |
25 Mar 2023 |
chenchao15/NeuralTPS/im2mesh/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Less is More: Reducing Task and Model Complexity for 3D Point Cloud Semantic Segmentation |
20 Mar 2023 |
l1997i/lim3d/utils/evaluation.py c398cef55a293907 |
unverified |
Apache-2.0 (permissive) |
| Neural Vector Fields: Implicit Representation by Explicit Learning |
8 Mar 2023 |
Wi-sc/NVF/utils.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Unifying Short and Long-Term Tracking with Graph Hierarchies |
6 Dec 2022 |
dvl-tum/SUSHI/src/utils/motion_utils.py 470ebe50cdb0b817 |
unverified |
MIT (permissive) |
| What the DAAM: Interpreting Stable Diffusion Using Cross Attention |
10 Oct 2022 |
castorini/daam/daam/evaluate.py 5682e07dbd1c3e8b |
unverified |
MIT (permissive) |
| Dynamic 3D Scene Analysis by Point Cloud Accumulation |
25 Jul 2022 |
prs-eth/PCAccumulation/libs/loss.py f11d4fc544190a42 |
unverified |
MIT (permissive) |
| YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors |
6 Jul 2022 |
ibaiGorordo/ONNX-YOLOv7-Object-Detection/yolov7/utils.py 21a47905e98284b9 |
unverified |
MIT (permissive) |
| RES: A Robust Framework for Guiding Visual Explanation |
27 Jun 2022 |
yuyanggao/res/RES.py 1aea11ac7997b72f |
ran · our draft was wrong
|
no licence file found · pointer only |
| SNAKE: Shape-aware Neural 3D Keypoint Field |
3 Jun 2022 |
zhongcl-thu/SNAKE/core/nets/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Probabilistic Implicit Scene Completion |
4 Apr 2022 |
96lives/gca/baselines/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| A Visual Navigation Perspective for Category-Level Object Pose Estimation |
25 Mar 2022 |
wrld/visual_navigation_pose_estimation/nocs/utils.py 2757c39fd1d64ab6 |
unverified |
MIT (permissive) |
| CLIP-Mesh: Generating textured meshes from text using pretrained image-text models |
24 Mar 2022 |
autodeskailab/clip-forge/train_autoencoder.py 50f9537db6e67f4b |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| QAHOI: Query-Based Anchors for Human-Object Interaction Detection |
16 Dec 2021 |
cjw2021/QAHOI/datasets/hico_eval.py d00ab1027f478b6d |
unverified |
Apache-2.0 (permissive) |
| End-to-End Referring Video Object Segmentation with Multimodal Transformers |
29 Nov 2021 |
mttr2021/MTTR/metrics.py 88af85cb8ff7b022 |
unverified |
Apache-2.0 (permissive) |
| Ray-ONet: Efficient 3D Reconstruction From A Single RGB Image |
5 Jul 2021 |
ActiveVisionLab/ray-onet/im2mesh/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Learning to Estimate Robust 3D Human Mesh from In-the-Wild Crowded Scenes |
15 Apr 2021 |
hongsukchoi/3dcrowdnet_release/tool/check_crowdidx.py 140b89f632c7ec6c |
unverified |
MIT (permissive) |
| Decomposing 3D Scenes into Objects via Unsupervised Volume Segmentation |
2 Apr 2021 |
stelzner/obsurf/obsurf/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| Detecting Human-Object Interactions with Action Co-occurrence Priors |
17 Jul 2020 |
Dong-JinKim/ActionCooccurrencePriors/utils/bbox_utils.py d342aa66e09cb8a9 |
unverified |
MIT (permissive) |
| YOLOv4: Optimal Speed and Accuracy of Object Detection |
23 Apr 2020 |
HeegonJin/yolov1/yolov1.py 8773def2c6ac88dd |
ran · fixture could not drive it
fingerprinted |
no licence file found · pointer only |
| BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation |
5 Apr 2020 |
MaybeShewill-CV/bisenetv2-tensorflow/tools/cityscapes/test_bisenetv2_cityscapes.py 4c9edc2f665fa3ba |
ran · fixture could not drive it
fingerprinted |
MIT (permissive) |
| TACO: Trash Annotations in Context for Litter Detection |
16 Mar 2020 |
pedropro/TACO/detector/utils.py b2aae18de5fa3893 |
unverified |
MIT (permissive) |
| VSGNet: Spatial Attention Network for Detecting Human Object Interactions Using Graph Convolutions |
11 Mar 2020 |
ASMIftekhar/VSGNet/scripts_hico/HICO_eval/bbox_utils.py d342aa66e09cb8a9 |
unverified |
MIT (permissive) |
| Convolutional Occupancy Networks |
10 Mar 2020 |
autonomousvision/convolutional_occupancy_networks/src/common.py 0072de07b144b152 |
ran
|
MIT (permissive) |
| PointRend: Image Segmentation as Rendering |
17 Dec 2019 |
ayoolaolafenwa/PixelLib/pixellib/instance/utils.py 64403fb78c59e685 |
unverified |
MIT (permissive) |
| LVIS: A Dataset for Large Vocabulary Instance Segmentation |
8 Aug 2019 |
craston/object_detection_cib/kod/core/bbox/iou.py 0fb7599af3cad11b |
unverified |
Apache-2.0 (permissive) |
| RL-GAN-Net: A Reinforcement Learning Agent Controlled GAN Network for Real-Time Point Cloud Shape Completion |
28 Apr 2019 |
iSarmad/RL-GAN-Net/GAN/models/lossess.py f1eb659f7c048f5c |
unverified |
MIT (permissive) |
| Normalized Object Coordinate Space for Category-Level 6D Object Pose and Size Estimation |
9 Jan 2019 |
sahithchada/NOCS_PyTorch/utils.py b2aae18de5fa3893 |
unverified |
MIT (permissive) |
| SketchyScene: Richly-Annotated Scene Sketches |
7 Aug 2018 |
SketchyScene/SketchyScene/Instance_Segmentation/libs/utils.py b2aae18de5fa3893 |
unverified |
MIT (permissive) |
| SO-Net: Self-Organizing Network for Point Cloud Analysis |
12 Mar 2018 |
lijx10/SO-Net/models/losses.py 8f4466573c6f0ffe |
unverified |
MIT (permissive) |
| SO-Net: Self-Organizing Network for Point Cloud Analysis |
12 Mar 2018 |
LONG-9621/SO-Net/models/losses.py f1eb659f7c048f5c |
unverified |
MIT (permissive) |
| Mask R-CNN |
20 Mar 2017 |
RituYadav92/Radar-RGB-Attentive-Multimodal-Object-Detection/Radar_RGB_Camera_Object_Detection/mrcnn/utils.py ae0c5e015d8cf734 |
ran
|
MIT (permissive) |
| Mask R-CNN |
20 Mar 2017 |
itsasimiqbal/SeBRe/utils.py b2aae18de5fa3893 |
unverified |
MIT (permissive) |
| Face Aging With Conditional Generative Adversarial Networks |
7 Feb 2017 |
Vishal-V/GSoC-TensorFlow-2019/mask_rcnn/utils.py ae0c5e015d8cf734 |
ran
|
Apache-2.0 (permissive) |
| StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks |
10 Dec 2016 |
Vishal-V/GSoC/mask_rcnn/utils.py ae0c5e015d8cf734 |
ran
|
Apache-2.0 (permissive) |
| Detecting Text in Natural Image with Connectionist Text Proposal Network |
12 Sep 2016 |
CrazySummerday/ctpn.pytorch/ctpn/utils.py 5dfe6dd29f57f960 |
unverified |
MIT (permissive) |
| SSD: Single Shot MultiBox Detector |
8 Dec 2015 |
ChunML/ssd-tf2/box_utils.py 366a1b1467436d59 |
unverified |
MIT (permissive) |
| You Only Look Once: Unified, Real-Time Object Detection |
8 Jun 2015 |
Everina/car-detection-yolo/yolo_v1.py 84b11777c5a839c6 |
unverified |
Apache-2.0 (permissive) |
| arXiv:aaai_25380 |
|
hailanyi/TED/pcdet/utils/bbloss.py f940cdda72e1a568 |
ran
|
Apache-2.0 (permissive) |
| arXiv:aaai_19986 |
|
PuAnysh/UFPMP-Det/UFPMP-Det-Tools/eval_script/ufpmp_det_eval.py 48738a53672ce96e |
unverified |
Apache-2.0 (permissive) |
| arXiv:Wang_3D_Human_Mesh_Recovery_with_Sequentially_Global_Rotation_Estimation_ICCV_2023_paper |
|
kennethwdk/SGRE/tool/check_crowdidx.py 140b89f632c7ec6c |
unverified |
MIT (permissive) |
| arXiv:Fogel_Open-Canopy_Towards_Very_High_Resolution_Forest_Monitoring_CVPR_2025_paper |
|
fajwel/Open-Canopy/src/metrics/metrics_utils.py 4841b773d8c6b024 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2024.findings-acl.307 |
|
wj210/NLI_ETP/model/utils.py 99b774c8c7fcfb6a |
unverified |
MIT (permissive) |
| arXiv:136890725 |
|
MendelXu/zsseg.baseline/mask_former/ablation/oracle_mask_former_model.py b89a7e00b541a38a |
unverified |
MIT (permissive) |