| YesTrack: Referring Multi-Object Tracking via MLLM-based Yes/No Verification added by Syntology |
2026-09 (from id) |
ggbondrighthere24/YesTrack/utils/nested_tensor.py 50736ecf379fe65b |
unverified |
MIT (permissive) |
| Prior-Guided DETR for Ultrasound Nodule Detection added by Syntology |
2026-01 (from id) |
wjj1wjj/Ultrasound-DETR/models/dn_dab_deformable_detr/dab_deformable_detr.py 966e78e3ec193e43 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation added by Syntology |
2025-08 (from id) |
JiahuaDong/HVPL/hvpl/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| SCORE: Scene Context Matters in Open-Vocabulary Remote Sensing Instance Segmentation |
17 Jul 2025 |
HuangShiqi128/SCORE/score/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| Disentangling Instance and Scene Contexts for 3D Semantic Scene Completion |
11 Jul 2025 |
Enyu-Liu/DISC/maskdino/models/misc.py 58cc9ff3bf75e753 |
unverified |
no licence file found · pointer only |
| Super-class guided Transformer for Zero-Shot Attribute Classification |
10 Jan 2025 |
mlvlab/SugaFormer/models/sugaformer.py 04ce9f512c0dcb96 |
ran · our draft was wrong
|
no licence file found · pointer only |
| A Simple Image Segmentation Framework via In-Context Examples |
7 Oct 2024 |
aim-uofa/SINE/sine/utils/misc.py 58cc9ff3bf75e753 |
unverified |
no licence file found · pointer only |
| Part2Object: Hierarchical Unsupervised 3D Instance Segmentation |
14 Jul 2024 |
chengshiest/part2object/models/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding |
3 Jul 2024 |
WeitaiKang/SegVG/models/SegVG.py a7b239b27e836fef |
ran · our draft was wrong
|
no licence file found · pointer only |
| Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation |
2 Jun 2024 |
hvision-nku/cascade-clip/models/losses/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| Bridging the Gap Between End-to-End and Two-Step Text Spotting |
6 Apr 2024 |
mxin262/estextspotter/models/ests/ests.py 6309c6cf50423938 |
ran · our draft was wrong
|
no licence file found · pointer only |
| OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation |
21 Mar 2024 |
cubeyoung/OTSeg/models/losses/misc.py 58cc9ff3bf75e753 |
unverified |
no licence file found · pointer only |
| PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck |
8 Mar 2024 |
anguyen8/peeb/src/owlvit_cls.py d3faa0ed46f2e436 |
ran
|
MIT (permissive) |
| PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck |
8 Mar 2024 |
anguyen8/peeb/src/owlvit_inference.py e54dd17857816dda |
ran
|
MIT (permissive) |
| PEM: Prototype-based Efficient MaskFormer for Image Segmentation |
29 Feb 2024 |
niccolocavagnero/pem/pem/utils/misc.py 58cc9ff3bf75e753 |
unverified |
no licence file found · pointer only |
| Semi-supervised Open-World Object Detection |
25 Feb 2024 |
sahalshajim/SS-OWFormer/models/deformable_detr.py c4caba3541d8390f |
ran · our draft was wrong
|
no licence file found · pointer only |
| Supervised Fine-tuning in turn Improves Visual Foundation Models |
18 Jan 2024 |
tencentarc/visft/mmf/models/visft/misc.py e8c3f53b24248086 |
ran
|
Apache-2.0 (permissive) |
| TMT-VIS: Taxonomy-aware Multi-dataset Joint Training for Video Instance Segmentation |
11 Dec 2023 |
rkzheng99/TMT-VIS/tmt/utils/misc.py 58cc9ff3bf75e753 |
unverified |
no licence file found · pointer only |
| Bridging the Gap: A Unified Video Comprehension Framework for Moment Retrieval and Highlight Detection |
28 Nov 2023 |
easonxiao-888/uvcom/uvcom/misc_ddp.py afb6efd87349dd80 |
ran
|
MIT (permissive) |
| SED: A Simple Encoder-Decoder for Open-Vocabulary Semantic Segmentation |
27 Nov 2023 |
xb534/sed/sed/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| 3D Indoor Instance Segmentation in an Open-World |
25 Sep 2023 |
aminebdj/3D-OWIS/models/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| CATR: Combinatorial-Dependence Audio-Queried Transformer for Audio-Visual Video Segmentation |
18 Sep 2023 |
aspirinone/catr.github.io/CATR/misc.py afb6efd87349dd80 |
ran
|
no licence file found · pointer only |
| Point-Query Quadtree for Crowd Counting, Localization, and More |
26 Aug 2023 |
cxliu0/PET/models/pet.py 59c5d55512e01869 |
ran · our draft was wrong
|
MIT (permissive) |
| LATR: 3D Lane Detection from Monocular Images with Transformer |
8 Aug 2023 |
JMoonr/LATR/models/sparse_inst_loss.py e0e4b133a453e3c3 |
ran
|
MIT (permissive) |
| Dynamic Token Pruning in Plain Vision Transformers for Semantic Segmentation |
2 Aug 2023 |
zbwxp/Dynamic-Token-Pruning/mmseg_custom/loss/misc.py 58cc9ff3bf75e753 |
unverified |
BSD-2-Clause (permissive) |
| Symphonize 3D Semantic Scene Completion with Contextual Instance Queries |
27 Jun 2023 |
hustvl/symphonies/maskdino/models/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| GRES: Generalized Referring Expression Segmentation |
1 Jun 2023 |
henghuiding/ReLA/gres_model/utils/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| Bridging Precision and Confidence: A Train-Time Loss for Calibrating Object Detection |
25 Mar 2023 |
akhtarvision/bpc_calibration/models/deformable_detr.py 07a415e6bef6b5cf |
ran · our draft was wrong
|
MIT (permissive) |
| MDQE: Mining Discriminative Query Embeddings to Segment Occluded Instances on Challenging Videos |
25 Mar 2023 |
MinghanLi/MDQE_CVPR2023/mdqe/models/mdqe.py 2f3dddf6597f2f3c |
ran · our draft was wrong
|
no licence file found · pointer only |
| CORA: Adapting CLIP for Open-Vocabulary Detection with Region Prompting and Anchor Pre-Matching |
23 Mar 2023 |
tgxs002/CORA/models/fast_detr.py 6c586ce97f1675d6 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation |
21 Mar 2023 |
KU-CVLAB/CAT-Seg/cat_seg/utils/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| FastInst: A Simple Query-Based Model for Real-Time Instance Segmentation |
15 Mar 2023 |
junjiehe96/fastinst/fastinst/utils/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| Referring Multi-Object Tracking |
6 Mar 2023 |
wudongming97/rmot/models/transrmot.py 003345262cfce83d |
ran · our draft was wrong
|
licence not identified · pointer only |
| SPTS v2: Single-Point Scene Text Spotting |
4 Jan 2023 |
bytedance/sptsv2/util/misc_sptsv2.py 5adadb63e05f89b9 |
unverified |
Apache-2.0 (permissive) |
| ZegCLIP: Towards Adapting CLIP for Zero-shot Semantic Segmentation |
7 Dec 2022 |
ZiqinZhou66/ZegCLIP/models/losses/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| Mask3D: Mask Transformer for 3D Semantic Instance Segmentation |
6 Oct 2022 |
jonasschult/mask3d/models/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| ECO-TR: Efficient Correspondences Finding Via Coarse-to-Fine Refinement |
25 Sep 2022 |
dltan7/ECO-TR/src/models/ecotr_modules/misc.py 764b0c92414d2fed |
unverified |
Apache-2.0 (permissive) |
| Video Mask Transfiner for High-Quality Video Instance Segmentation |
28 Jul 2022 |
SysCV/vmt/models/segmentation.py 27463d5407d3dddf |
ran
|
Apache-2.0 (permissive) |
| Towards Hard-Positive Query Mining for DETR-based Human-Object Interaction Detection |
12 Jul 2022 |
MuchHair/HQM/models/Hard_Sample/HQM/hoi_HQM.py f5602847693d22a5 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Detecting and Recovering Sequential DeepFake Manipulation |
5 Jul 2022 |
rshaojimmy/seqdeepfake/models/SeqFakeFormer.py c45de9291f504b90 |
ran · our draft was wrong
|
no licence file found · pointer only |
| VITA: Video Instance Segmentation via Object Token Association |
9 Jun 2022 |
sukjunhwang/vita/vita/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| Text Spotting Transformers |
5 Apr 2022 |
mlpc-ucsd/TESTR/adet/modeling/testr/models.py 874889b24390cf77 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Collaborative Transformers for Grounded Situation Recognition |
30 Mar 2022 |
towhee-io/towhee/towhee/models/coformer/coformer.py 32db9196c8ba7a26 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Collaborative Transformers for Grounded Situation Recognition |
30 Mar 2022 |
jhcho99/CoFormer/models/coformer.py ab1a441fc7e324ad |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Sparse Instance Activation for Real-Time Instance Segmentation |
24 Mar 2022 |
hustvl/sparseinst/sparseinst/utils.py 764b0c92414d2fed |
unverified |
MIT (permissive) |
| Open-Vocabulary DETR with Conditional Matching |
22 Mar 2022 |
yuhangzang/OV-DETR/ovdetr/models/model.py 57b6cc60d20c5e85 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Towards Data-Efficient Detection Transformers |
17 Mar 2022 |
encounter1997/de-conddetr/models/conditional_detr.py 4c8b2d763e2923a0 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Towards Data-Efficient Detection Transformers |
17 Mar 2022 |
encounter1997/DE-DETRs/models/detr.py 1bf01bde72e96783 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Accelerating DETR Convergence via Semantic-Aligned Matching |
14 Mar 2022 |
ZhangGongjie/SAM-DETR/models/fast_detr.py 9277ea2a21aca6f8 |
ran · our draft was wrong
|
MIT (permissive) |
| DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR |
28 Jan 2022 |
helq2612/biadt/models/dn_dab_deformable_detr/dab_deformable_detr.py 6442da4f34dd152c |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| End-to-End Referring Video Object Segmentation with Multimodal Transformers |
29 Nov 2021 |
mttr2021/MTTR/misc.py afb6efd87349dd80 |
ran
|
Apache-2.0 (permissive) |
| Pix2seq: A Language Modeling Framework for Object Detection |
22 Sep 2021 |
volgachen/Pix2Seq_Pytorch/playground/pix2seq/pix2seq.py 8542cf00e1274bb3 |
ran · our draft was wrong
|
no licence file found · pointer only |
| Pix2seq: A Language Modeling Framework for Object Detection |
22 Sep 2021 |
gaopengcuhk/Pretrained-Pix2Seq/playground/pix2seq/pix2seq.py 9b0e434afd6ba1bb |
ran
|
no licence file found · pointer only |
| Mining the Benefits of Two-stage and One-stage HOI Detection |
11 Aug 2021 |
YueLiao/CDN/models/hoi.py 384f0509d5ccd8b9 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Learning Multi-Scene Absolute Pose Regression with Transformers |
21 Mar 2021 |
yolish/c2f-ms-transformer/models/transposenet/C2FEMSTransPoseNet.py a78bbd4466ce14cd |
ran · our draft was wrong
|
no licence file found · pointer only |
| Deformable DETR: Deformable Transformers for End-to-End Object Detection |
8 Oct 2020 |
dianzl/sodformer/models/deformable_detr.py a51cb66e9b8cd2fd |
ran · our draft was wrong
|
no licence file found · pointer only |
| Deformable DETR: Deformable Transformers for End-to-End Object Detection |
8 Oct 2020 |
duongnv0499/Explain-Deformable-DETR/models/deformable_detr.py 9691821ac0f2a520 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Deformable DETR: Deformable Transformers for End-to-End Object Detection |
8 Oct 2020 |
YC-Lai/Sequential-DDETR/models/deformable_detr.py ded2ab6086dcbd94 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| Deformable DETR: Deformable Transformers for End-to-End Object Detection |
8 Oct 2020 |
zhechen/deformable-detr-rego/models/deformable_detr.py fe880b26b1a8908a |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| End-to-End Object Detection with Transformers |
26 May 2020 |
Li-ai-cell/Interpretation_DETR/models/deformable_detr.py 58a69a8a71330322 |
ran · our draft was wrong
|
Apache-2.0 (permissive) |
| SGDR: Stochastic Gradient Descent with Warm Restarts |
13 Aug 2016 |
buptlwz/mabp/gres_model/utils/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| arXiv:Zhu_SkySense-O_Towards_Open-World_Remote_Sensing_Interpretation_with_Vision-Centric_Visual-Language_Modeling_CVPR_2025_paper |
|
zqcrafts/SkySense-O/skysense_o/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| arXiv:Zhang_FreePoint_Unsupervised_Point_Cloud_Instance_Segmentation_CVPR_2024_paper |
|
zzk273/FreePoint/models/misc.py 58cc9ff3bf75e753 |
unverified |
MIT (permissive) |
| arXiv:Xie_SED_A_Simple_Encoder-Decoder_for_Open-Vocabulary_Semantic_Segmentation_CVPR_2024_paper |
|
xb534/SED/sed/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |
| arXiv:Xiao_Bridging_the_Gap_A_Unified_Video_Comprehension_Framework_for_Moment_CVPR_2024_paper |
|
EasonXiao-888/UVCOM/uvcom/misc_ddp.py afb6efd87349dd80 |
ran
|
MIT (permissive) |
| arXiv:Nguyen_Region-Level_Data_Attribution_for_Text-to-Image_Generative_Models_ICCV_2025_paper |
|
AIoT-Lab-BKAI/AR-Detector/models/GroundingDINO/groundingdino.py 89899c07bc9d38f9 |
ran · our draft was wrong
|
no licence file found · pointer only |
| arXiv:Heo_A_Generalized_Framework_for_Video_Instance_Segmentation_CVPR_2023_paper |
|
miranheo/GenVIS/genvis/utils/misc.py 58cc9ff3bf75e753 |
unverified |
Apache-2.0 (permissive) |