| YesTrack: Referring Multi-Object Tracking via MLLM-based Yes/No Verification added by Syntology |
2026-09 (from id) |
ggbondrighthere24/YesTrack/utils/box_ops.py 08b30db38a23e292 |
unverified |
MIT (permissive) |
| Nodule-DETR: A Novel DETR Architecture with Frequency-Channel Attention for Ultrasound Thyroid Nodule Detection added by Syntology |
2026-01 (from id) |
wjj1wjj/Nodule-DETR/Nodule-DETR/detr_See.py d8ff64e15d6e7507 |
unverified |
no licence file found · pointer only |
| PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination added by Syntology |
2025-09 (from id) |
Dmmm1997/PropVG/propvg/layers/box_ops.py 856982244cffd7e3 |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation |
6 Jun 2025 |
liyih/CSCL/code/MultiModal-DeepFake-main/models/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| CLIP Under the Microscope: A Fine-Grained Analysis of Multi-Object Representation |
27 Feb 2025 |
identical code first harvested elsewhere e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
licence of this copy not recorded |
| Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints |
12 Jan 2025 |
dmmm1997/c3vg/c3vg/core/utils.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints |
12 Jan 2025 |
dmmm1997/c3vg/c3vg/layers/box_ops.py 856982244cffd7e3 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking |
20 Dec 2024 |
nju-pcalab/sttrack/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| DEIM: DETR with Improved Matching for Fast Convergence |
5 Dec 2024 |
shihuahuang95/deim/engine/deim/box_ops.py cb48d7481c17a1f7 |
unverified |
licence not identified · pointer only |
| HyperSeg: Towards Universal Visual Segmentation with Large Language Model |
26 Nov 2024 |
congvvc/HyperSeg/hyperseg/model/tracker/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation |
18 Nov 2024 |
JT-Sun/Filtering-WoRA/models/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| RETR: Multi-View Radar Detection Transformer for Indoor Perception |
15 Nov 2024 |
merlresearch/radar-detection-transformer/src/models/module_retr/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
AGPL-3.0 (copyleft) · pointer only |
| EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection |
31 Oct 2024 |
ChelsieLei/EZ-HOI/ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion |
26 Sep 2024 |
dmmm1997/simvg/simvg/core/utils.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| OVA-DETR: Open Vocabulary Aerial Object Detection Using Image-Text Alignment and Fusion |
22 Aug 2024 |
GT-Wei/RT-OVAD/src/zoo/itc_ovad/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| MambaEVT: Event Stream based Visual Object Tracking using State Space Model |
20 Aug 2024 |
event-ahu/mambaevt/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection |
5 Aug 2024 |
ltttpku/cmmp/ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| SAM 2: Segment Anything in Images and Videos |
1 Aug 2024 |
louisfinner/him2sam/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| ActionVOS: Actions as Prompts for Video Object Segmentation |
10 Jul 2024 |
ut-vision/ActionVOS/RF_ActionVOS/inference_actionvos.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
no licence file found · pointer only |
| Mamba-FETrack: Frame-Event Tracking via State Space Model |
28 Apr 2024 |
event-ahu/mamba_fetrack/Mamba_FETrack/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| LaSagnA: Language-based Segmentation Assistant for Complex Queries |
12 Apr 2024 |
congvvc/lasagna/model/matcher.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Temporally Consistent Referring Video Object Segmentation with Hybrid Memory |
28 Mar 2024 |
bo-miao/HTR/inference_davis.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
MIT (permissive) |
| Exploring Pre-trained Text-to-Video Diffusion Models for Referring Video Object Segmentation |
18 Mar 2024 |
buxiangzhiren/vd-it/inference_davis.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
no licence file found · pointer only |
| Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object Detector |
5 Feb 2024 |
lovelyqian/CDFSOD-benchmark/lib/regionprop.py 009912c75c8c77cf |
ran
fingerprinted |
Apache-2.0 (permissive) |
| ZoomTrack: Target-aware Non-uniform Resizing for Efficient Visual Tracking |
16 Oct 2023 |
Kou-99/ZoomTrack/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Event Stream-based Visual Object Tracking: A High-Resolution Benchmark Dataset and A Novel Baseline |
26 Sep 2023 |
event-ahu/coesot/CEUTrack/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Detecting and Grounding Multi-Modal Media Manipulation and Beyond |
25 Sep 2023 |
rshaojimmy/multimodal-deepfake/models/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Detect Everything with Few Examples |
22 Sep 2023 |
mlzxy/devit/lib/regionprop.py 009912c75c8c77cf |
ran
fingerprinted |
MIT (permissive) |
| Box-based Refinement for Weakly Supervised and Unsupervised Localization Tasks |
7 Sep 2023 |
eyalgomel/box-based-refinement/detr/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| A Joint Study of Phrase Grounding and Task Performance in Vision and Language Models |
6 Sep 2023 |
lil-lab/phrase_grounding/src/evals/postprocessors.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| A Joint Study of Phrase Grounding and Task Performance in Vision and Language Models |
6 Sep 2023 |
lil-lab/phrase_grounding/src/datasets/coco_format_dataset.py ac92d3ba92f8d1ef |
ran
|
no licence file found · pointer only |
| CiteTracker: Correlating Image and Text for Visual Tracking |
22 Aug 2023 |
NorahGreen/CiteTracker/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Spectrum-guided Multi-granularity Referring Video Object Segmentation |
25 Jul 2023 |
bo-miao/sgmg/inference_davis.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
no licence file found · pointer only |
| CTVIS: Consistent Training for Online Video Instance Segmentation |
24 Jul 2023 |
kainingying/ctvis/ctvis/utils/utils.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Less is More: Focus Attention for Efficient DETR |
24 Jul 2023 |
linxid/Focus-DETR/models/focus_detr/box_ops.py 074307e51146fa09 |
unverified |
Apache-2.0 (permissive) |
| OnlineRefer: A Simple Online Baseline for Referring Video Object Segmentation |
18 Jul 2023 |
wudongming97/onlinerefer/inference_davis_online.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Random Boxes Are Open-world Object Detectors |
17 Jul 2023 |
scuwyh2000/RandBox/randbox/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| MixFormerV2: Efficient Fully Transformer Tracking |
25 May 2023 |
mcg-nju/mixformerv2/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Detection Transformer with Stable Matching |
10 Apr 2023 |
IDEA-Research/detrex/detrex/modeling/matcher/modified_matcher.py 856982244cffd7e3 |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| DiffusionInst: Diffusion Model for Instance Segmentation |
6 Dec 2022 |
alipay/diffusion-model-for-instance-segmentation/diffusioninst/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| GSRFormer: Grounded Situation Recognition Transformer with Alternate Semantic Attention Refinement |
18 Aug 2022 |
zhiqic/gsrformer/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Towards Grand Unification of Object Tracking |
14 Jul 2022 |
identical code first harvested elsewhere e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
licence of this copy not recorded |
| Towards Hard-Positive Query Mining for DETR-based Human-Object Interaction Detection |
12 Jul 2022 |
MuchHair/HQM/models/Hard_Sample/HQM/hoi_HQM.py ae8f0c20ba273e10 |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Cross-View Language Modeling: Towards Unified Cross-Lingual Cross-Modal Pre-training |
1 Jun 2022 |
zengyan-97/cclm/models/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
BSD-3-Clause recorded; this copy not marked cleared · pointer only |
| Correlation-Aware Deep Tracking |
3 Mar 2022 |
phiphiphi31/SBT/lib/models/sbt/transt_loss/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| Language as Queries for Referring Video Object Segmentation |
3 Jan 2022 |
wjn922/ReferFormer/inference_davis.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Rethinking the Two-Stage Framework for Grounded Situation Recognition |
10 Dec 2021 |
kellyiss/situformer/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
no licence file found · pointer only |
| Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer |
3 Dec 2021 |
fredzzhang/upt/ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
BSD-3-Clause (permissive) |
| Grounded Situation Recognition with Transformers |
19 Nov 2021 |
jhcho99/gsrtr/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| PubTables-1M: Towards comprehensive table extraction from unstructured documents |
30 Sep 2021 |
microsoft/table-transformer/src/inference.py 52d4b812bdc17687 |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| PubTables-1M: Towards comprehensive table extraction from unstructured documents |
30 Sep 2021 |
phamquiluan/table-transformer/src/transforms.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT recorded; this copy not marked cleared · pointer only |
| 4D-Net for Learned Multi-Modal Alignment |
2 Sep 2021 |
chanlilong/4D_NET_pytorch/utils/box_ops.py a8982a32c4283e7b |
unverified |
MIT (permissive) |
| TubeR: Tubelet Transformer for Video Action Detection |
2 Apr 2021 |
amazon-science/tubelet-transformer/models/transformer/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| End-to-End Trainable Multi-Instance Pose Estimation with Transformers |
22 Mar 2021 |
amathislab/poet/util/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| End-to-End Trainable Multi-Instance Pose Estimation with Transformers |
22 Mar 2021 |
pranoyr/pose-estimation-with-transformers/inference.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Rethinking Transformer-based Set Prediction for Object Detection |
21 Nov 2020 |
edward-sun/tsp-detection/rcnn/rcnn_heads.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
Apache-2.0 (permissive) |
| Deformable DETR: Deformable Transformers for End-to-End Object Detection |
8 Oct 2020 |
lyqcom/detr/src/DETR/matcher_np.py 074307e51146fa09 |
unverified |
Apache-2.0 (permissive) |
| End-to-End Object Detection with Transformers |
26 May 2020 |
LKLQQ/detr/src/box_ops.py 074307e51146fa09 |
unverified |
Apache-2.0 (permissive) |
| ATOM: Accurate Tracking by Overlap Maximization |
19 Nov 2018 |
xuefeng-zhu5/cdaat/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| arXiv:ijcai2024_0084 |
|
116508/CF-Deformable-DETR/VisualResult.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:aaai_28465 |
|
OpenGVLab/MUTR/inference_davis.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
MIT (permissive) |
| arXiv:Zhou_When_Pixel_Difference_Patterns_Meet_ViT_PiDiViT_for_Few-Shot_Object_ICCV_2025_paper |
|
Seaz9/PiDiViT/lib/regionprop.py 009912c75c8c77cf |
ran
fingerprinted |
Apache-2.0 (permissive) |
| arXiv:Yuan_CAT_A_Unified_Click-and-Track_Framework_for_Realistic_Tracking_ICCV_2025_paper |
|
ysyuann/CAT/lib/utils/box_ops.py e0a06ded5d4f6c3c |
ran · honoured contract
fingerprinted |
MIT (permissive) |
| arXiv:Pan_Wnet_Audio-Guided_Video_Object_Segmentation_via_Wavelet-Based_Cross-Modal_Denoising_Networks_CVPR_2022_paper |
|
asudahkzj/Wnet/inference_a2d.py ef1a3e10a9dbf4a9 |
ran
fingerprinted |
Apache-2.0 (permissive) |