| PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination added by Syntology |
2025-09 (from id) |
Dmmm1997/PropVG/propvg/layers/box_ops.py abe66849ff8e9119 |
unverified |
MIT (permissive) |
| Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation |
6 Jun 2025 |
liyih/CSCL/code/MultiModal-DeepFake-main/models/box_ops.py 7f5165b3c0380ff2 |
ran
|
MIT (permissive) |
| Understand, Think, and Answer: Advancing Visual Reasoning with Large Multimodal Models |
27 May 2025 |
jefferyzhan/griffon/griffon/coor_utils.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| DeepPerception: Advancing R1-like Cognitive Visual Perception in MLLMs for Knowledge-Intensive Visual Grounding |
17 Mar 2025 |
thunlp/deepperception/karl/eval/evaluate.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints |
12 Jan 2025 |
dmmm1997/c3vg/c3vg/layers/box_ops.py abe66849ff8e9119 |
unverified |
Apache-2.0 (permissive) |
| DEIM: DETR with Improved Matching for Fast Convergence |
5 Dec 2024 |
shihuahuang95/deim/engine/deim/box_ops.py 253961917cbe7713 |
unverified |
licence not identified · pointer only |
| VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation |
18 Nov 2024 |
JT-Sun/Filtering-WoRA/models/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| OVA-DETR: Open Vocabulary Aerial Object Detection Using Image-Text Alignment and Fusion |
22 Aug 2024 |
GT-Wei/RT-OVAD/src/zoo/itc_ovad/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Harmonizing Visual Text Comprehension and Generation |
23 Jul 2024 |
bytedance/textharmony/TextHarmony/utils/grounding_score.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning |
21 Jun 2024 |
Brandon3964/MultiModal-Task-Vector/eval_mm/evaluate_grounding.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| DEEM: Diffusion Models Serve as the Eyes of Large Language Models for Image Perception |
24 May 2024 |
rainbowluocs/deem/uni_interleaved/utils/grounding_score.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| LaSagnA: Language-based Segmentation Assistant for Complex Queries |
12 Apr 2024 |
congvvc/lasagna/model/matcher.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model |
9 Nov 2023 |
OPPOMKLab/u-LLaVA/models/loss.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Detecting and Grounding Multi-Modal Media Manipulation and Beyond |
25 Sep 2023 |
rshaojimmy/multimodal-deepfake/models/box_ops.py 7f5165b3c0380ff2 |
ran
|
licence not identified · pointer only |
| Box-based Refinement for Weakly Supervised and Unsupervised Localization Tasks |
7 Sep 2023 |
eyalgomel/box-based-refinement/detr/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| CTVIS: Consistent Training for Online Video Instance Segmentation |
24 Jul 2023 |
kainingying/ctvis/ctvis/utils/utils.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Random Boxes Are Open-world Object Detectors |
17 Jul 2023 |
scuwyh2000/RandBox/randbox/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Detection Transformer with Stable Matching |
10 Apr 2023 |
IDEA-Research/detrex/detrex/modeling/matcher/modified_matcher.py d8e49e4f454b76fc |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Bridging Precision and Confidence: A Train-Time Loss for Calibrating Object Detection |
25 Mar 2023 |
akhtarvision/bpc_calibration/models/deformable_detr.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| DiffusionInst: Diffusion Model for Instance Segmentation |
6 Dec 2022 |
alipay/diffusion-model-for-instance-segmentation/diffusioninst/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Extending Phrase Grounding with Pronouns in Visual Dialogues |
23 Oct 2022 |
izhx/Phrase-Grounding-with-Pronoun/code/src/metrics.py ac3d20339465e1a7 |
unverified |
Apache-2.0 (permissive) |
| Cross-View Language Modeling: Towards Unified Cross-Lingual Cross-Modal Pre-training |
1 Jun 2022 |
zengyan-97/cclm/models/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| Correlation-Aware Deep Tracking |
3 Mar 2022 |
phiphiphi31/SBT/lib/models/sbt/transt_loss/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Rethinking the Two-Stage Framework for Grounded Situation Recognition |
10 Dec 2021 |
kellyiss/situformer/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| YOLOP: You Only Look Once for Panoptic Driving Perception |
25 Aug 2021 |
hustvl/yolop/lib/core/general.py 3113eedc8e71c19e |
unverified |
MIT (permissive) |
| TubeR: Tubelet Transformer for Video Action Detection |
2 Apr 2021 |
amazon-science/tubelet-transformer/models/transformer/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| End-to-End Trainable Multi-Instance Pose Estimation with Transformers |
22 Mar 2021 |
amathislab/poet/util/box_ops.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| YOLOv4: Optimal Speed and Accuracy of Object Detection |
23 Apr 2020 |
Gavino7/YOLOv3set/yolo2/loss.py 7ac8a3042b335d7e |
unverified |
MIT (permissive) |
| YOLOv4: Optimal Speed and Accuracy of Object Detection |
23 Apr 2020 |
Gavino7/YOLOv3set/yolo3/loss.py a432a967aa4de172 |
unverified |
MIT (permissive) |
| Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression |
19 Nov 2019 |
grifon-239/diploma/yolo2/loss.py 7ac8a3042b335d7e |
unverified |
MIT (permissive) |
| Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression |
19 Nov 2019 |
grifon-239/diploma/yolo3/loss.py a432a967aa4de172 |
unverified |
MIT (permissive) |
| Mask R-CNN |
20 Mar 2017 |
Okery/PyTorch-Simple-MaskRCNN/pytorch_mask_rcnn/model/mask_rcnn.py 46c0d05b31f0858a |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks |
4 Jun 2015 |
liangheming/faster_rcnnv1/nets/faster_rcnn.py 2ec9b42e568fb19f |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks |
4 Jun 2015 |
AlphaJia/pytorch-faster-rcnn/utils/rpn_utils.py 0ca35377bdcd364d |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| arXiv:openreview_XtIRCAEYoJ |
|
WayneTomas/Artemis/val/refcoco_all/grd_eval_utils.py ce749f837424dd8c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:aaai_28394 |
|
iam-nacl/DTMFormer/utils/non_maximum_suppression.py c3725b35b6d0e285 |
unverified |
MIT (permissive) |
| arXiv:aaai_25418 |
|
Shinetism/VStates/utils/box_utils.py c5c900940df67d7f |
unverified |
MIT (permissive) |