| Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization added by Syntology |
2026-06 (from id) |
phoenixnir/FLAME/FLAME/utils/metrics.py fd031662f9fef178 |
ran
|
no licence file found · pointer only |
| ObjEmbed: Towards Universal Multimodal Object Embeddings added by Syntology |
2026-02 (from id) |
WeChatCV/ObjEmbed/models/qwen3vl_objembed.py 34d3bd26e7ae4a9c |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Nodule-DETR: A Novel DETR Architecture with Frequency-Channel Attention for Ultrasound Thyroid Nodule Detection added by Syntology |
2026-01 (from id) |
wjj1wjj/Nodule-DETR/Nodule-DETR/models/DABDETR.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Towards Integrating Uncertainty for Domain-Agnostic Segmentation added by Syntology |
2025-12 (from id) |
JesseBrouw/UncertSAM/src/loss_fns.py f655437ab080689e |
unverified |
AGPL-3.0 (copyleft) · pointer only |
| Robust Ego-Exo Correspondence with Long-Term Memory added by Syntology |
2025-10 (from id) |
juneyeeHu/LM-EEC/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
no licence file found · pointer only |
| UGround: Towards Unified Visual Grounding with Unrolled Transformers added by Syntology |
2025-10 (from id) |
rui-qian/UGround/model/READ.py 5e032c86b35b7a19 |
unverified |
Apache-2.0 (permissive) |
| Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation added by Syntology |
2025-08 (from id) |
JiahuaDong/HVPL/hvpl/modeling/vita_criterion.py f56d065ac11dd808 |
unverified |
Apache-2.0 (permissive) |
| arXiv:2507.19993 |
2025-07 (from id) |
Howardkhh/FROSS/EGTR/model/util.py 728119ce133dfec1 |
unverified |
Apache-2.0 (permissive) |
| Disentangling Instance and Scene Contexts for 3D Semantic Scene Completion |
11 Jul 2025 |
Enyu-Liu/DISC/maskdino/models/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning |
11 Jun 2025 |
facebookresearch/vjepa2/evals/action_anticipation_frozen/losses.py 3d79c2bcb14adc04 |
unverified |
MIT recorded; this copy not marked cleared · pointer only |
| Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures |
16 May 2025 |
ashkamath/mdetr/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Object-Shot Enhanced Grounding Network for Egocentric Video |
7 May 2025 |
Yisen-Feng/OSGNet/libs/modeling/losses.py cb7608e384b5c1d7 |
unverified |
MIT (permissive) |
| Perception Encoder: The best visual embeddings are not at the output of the network |
17 Apr 2025 |
facebookresearch/perception_models/apps/detection/DETA_pe/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks |
28 Mar 2025 |
iSEE-Laboratory/AMNAR/libs/modeling/losses.py a617c27dff1d8aec |
unverified |
MIT (permissive) |
| OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer |
13 Mar 2025 |
jinyanglii/OVTR/ovtr/models/segmentation.py e830c3de135f1096 |
ran
|
MIT (permissive) |
| OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer |
13 Mar 2025 |
jinyanglii/OVTR/ovtr/models/utils.py 23eb1c50e6bb2ba5 |
ran
|
MIT (permissive) |
| PISA Experiments: Exploring Physics Post-Training for Video Diffusion Models by Watching Stuff Drop |
12 Mar 2025 |
vision-x-nyu/pisa-experiments/sam2/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
Apache-2.0 (permissive) |
| Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration |
17 Dec 2024 |
zzhhfut/ccnet-aaai2025/libs/modeling/losses.py e5b38592e26eef34 |
unverified |
no licence file found · pointer only |
| Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis |
12 Dec 2024 |
ring-zl/Speech-Forensics/libs/modeling/losses.py 0d06c0983d985a4a |
unverified |
no licence file found · pointer only |
| Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark |
12 Dec 2024 |
yongliang-wu/repurpose/models/losses.py 5500f7dbd94b2c87 |
unverified |
MIT (permissive) |
| A Distractor-Aware Memory for Visual Object Tracking with SAM2 |
26 Nov 2024 |
jovanavidenovic/dam4sam/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
no licence file found · pointer only |
| SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion |
26 Sep 2024 |
dmmm1997/simvg/simvg/core/criterion/criterion.py dff0f0000acdc0ad |
ran
fingerprinted |
MIT (permissive) |
| ENACT: Entropy-based Clustering of Attention Input for Improving the Computational Performance of Object Detection Transformers |
11 Sep 2024 |
gsavathrakis/enact/Anchor-DETR-ENACT/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning |
15 Aug 2024 |
jinlab-imvr/surgical-sam-2/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
Apache-2.0 (permissive) |
| An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding |
2 Aug 2024 |
chenwei746/eevg/utils/loss_utils.py 0ce2b81bef4f1424 |
ran
|
no licence file found · pointer only |
| Boosting Gaze Object Prediction via Pixel-level Supervision from Vision Foundation Model |
2 Aug 2024 |
jinyang06/SamGOP/maskGOP/modeling/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| SAM 2: Segment Anything in Images and Videos |
1 Aug 2024 |
bowang-lab/medsam2/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
Apache-2.0 (permissive) |
| PartGLEE: A Foundation Model for Recognizing and Parsing Any Objects |
23 Jul 2024 |
ProvenceStar/PartGLEE/projects/PartGLEE/partglee/models/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| ActionVOS: Actions as Prompts for Video Object Segmentation |
10 Jul 2024 |
ut-vision/ActionVOS/RF_ActionVOS/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| DyFADet: Dynamic Feature Aggregation for Temporal Action Detection |
3 Jul 2024 |
yangle15/DyFADet-pytorch/libs/modeling/losses.py 54293ea22bd4f952 |
ran
|
no licence file found · pointer only |
| Bootstrapping Referring Multi-Object Tracking |
7 Jun 2024 |
zyn213/temprmot/models/segmentation.py e830c3de135f1096 |
ran
|
no licence file found · pointer only |
| LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection |
5 Jun 2024 |
atten4vis/lw-detr/models/lwdetr.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation |
2 Jun 2024 |
hvision-nku/cascade-clip/models/losses/criterion.py 2675e89c94ae08bf |
ran
fingerprinted |
MIT (permissive) |
| OV-DQUO: Open-Vocabulary DETR with Denoising Text Query Training and Open-World Unknown Objects Supervision |
28 May 2024 |
xiaomoguhz/ov-dquo/models/ov_dquo/utils.py f09e169e2656c2f7 |
ran
|
Apache-2.0 (permissive) |
| Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model |
27 May 2024 |
kuanchihhuang/reason3d/lavis/models/reason3d_models/seg_loss.py 94629ea5eed6e387 |
ran
|
no licence file found · pointer only |
| Do You Remember? Dense Video Captioning with Cross-Modal Memory Retrieval |
11 Apr 2024 |
ailab-kyunghee/cm2_dvc/cm2/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| UniAV: Unified Audio-Visual Perception for Multi-Task Video Event Localization |
4 Apr 2024 |
ttgeng233/UniAV/libs/modeling/losses.py 54293ea22bd4f952 |
ran
|
MIT (permissive) |
| SnAG: Scalable and Accurate Video Grounding |
2 Apr 2024 |
happyharrycn/actionformer_release/libs/modeling/losses.py 0d06c0983d985a4a |
unverified |
MIT (permissive) |
| EGTR: Extracting Graph from Transformer for Scene Graph Generation |
2 Apr 2024 |
naver-ai/egtr/model/util.py 728119ce133dfec1 |
unverified |
Apache-2.0 (permissive) |
| Disentangled Pre-training for Human-Object Interaction Detection |
2 Apr 2024 |
xingaoli/DP-HOI/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| ECLIPSE: Efficient Continual Learning in Panoptic Segmentation with Visual Prompt Tuning |
29 Mar 2024 |
clovaai/ECLIPSE/continual/method_wrapper/loss.py 8820a1f39a3f1556 |
ran
|
licence not identified · pointer only |
| Temporally Consistent Referring Video Object Segmentation with Hybrid Memory |
28 Mar 2024 |
bo-miao/HTR/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| ReMamber: Referring Image Segmentation with Mamba Twister |
26 Mar 2024 |
yyh-rain-song/ReMamber/model/utils.py 36be76eee5039426 |
ran
fingerprinted |
no licence file found · pointer only |
| Multiple Object Tracking as ID Prediction |
25 Mar 2024 |
MCG-NJU/MOTIP/models/motip/id_criterion.py a542c79696b8b5ae |
ran
|
Apache-2.0 (permissive) |
| OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation |
21 Mar 2024 |
cubeyoung/OTSeg/models/losses/criterion.py 2675e89c94ae08bf |
ran
fingerprinted |
no licence file found · pointer only |
| $V_kD:$ Improving Knowledge Distillation using Orthogonal Projections |
10 Mar 2024 |
roymiles/vkd/vidt/methods/vidt/criterion.py bf3a75e08dc2c405 |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck |
8 Mar 2024 |
anguyen8/peeb/src/owlvit_cls.py 695ff74df95375df |
ran
fingerprinted |
MIT (permissive) |
| TransGOP: Transformer-Based Gaze Object Prediction |
21 Feb 2024 |
chenxi-Guo/TransGOP/models/TransGOP/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks |
25 Jan 2024 |
idea-research/groundingdino/groundingdino/models/GroundingDINO/utils.py 23eb1c50e6bb2ba5 |
ran
|
Apache-2.0 (permissive) |
| Symbol as Points: Panoptic Symbol Spotting via Point-based Representation |
19 Jan 2024 |
nicehuster/sympoint/svgnet/model/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| MS-DETR: Efficient DETR Training with Mixed Supervision |
8 Jan 2024 |
atten4vis/ms-detr/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| General Object Foundation Model for Images and Videos at Scale |
14 Dec 2023 |
FoundationVision/GLEE/projects/GLEE/glee/models/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Grounded Question-Answering in Long Egocentric Videos |
11 Dec 2023 |
becomebright/groundvqa/model/ours/nlq_head.py 7b7c8610c42bb07c |
ran
|
MIT (permissive) |
| Feature 3DGS: Supercharging 3D Gaussian Splatting to Enable Distilled Feature Fields |
6 Dec 2023 |
keloee/maskfield/models/losses.py 71e149dc0c33b040 |
unverified |
MIT (permissive) |
| PaSCo: Urban 3D Panoptic Scene Completion with Uncertainty Awareness |
4 Dec 2023 |
astra-vision/PaSCo/pasco/loss/losses.py 549df7244d5a6b10 |
ran
|
Apache-2.0 (permissive) |
| Rank-DETR for High Quality Object Detection |
13 Oct 2023 |
LeapLabTHU/Rank-DETR/projects/rank_detr/modeling/rankdetr_criterion.py dff0f0000acdc0ad |
ran
fingerprinted |
Apache-2.0 (permissive) |
| X-Pose: Detecting Any Keypoints |
12 Oct 2023 |
IDEA-Research/UniPose/models/UniPose/utils.py 23eb1c50e6bb2ba5 |
ran
|
no licence file found · pointer only |
| Contrastive Grouping with Transformer for Referring Image Segmentation |
2 Sep 2023 |
toneyaya/cgformer/model/segmenter.py e71d9c7dd14f39b8 |
ran
|
MIT (permissive) |
| Referring Image Segmentation Using Text Supervision |
28 Aug 2023 |
fawnliu/tris/loss/seg_loss.py 9cf331c90b5daee7 |
ran
|
MIT (permissive) |
| DiffusionTrack: Diffusion Model For Multi-Object Tracking |
19 Aug 2023 |
rainbowluocs/diffusiontrack/yolox/models/losses.py 774f31ba06110943 |
ran
fingerprinted |
licence not identified · pointer only |
| RLIPv2: Fast Scaling of Relational Language-Image Pre-training |
18 Aug 2023 |
jacobyuan7/ocn-hoi-benchmark/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Unsupervised Domain Adaptive Detection with Network Stability Analysis |
16 Aug 2023 |
tiankongzhang/nsa/detection/layers/losses.py b4ba161af9015512 |
ran
|
no licence file found · pointer only |
| Group Pose: A Simple Baseline for End-to-End Multi-person Pose Estimation |
14 Aug 2023 |
Michel-liu/GroupPose/models/grouppose/utils.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| V-DETR: DETR with Vertex Relative Position Encoding for 3D Object Detection |
8 Aug 2023 |
yichaoshen-ms/v-detr/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Dynamic Token Pruning in Plain Vision Transformers for Semantic Segmentation |
2 Aug 2023 |
zbwxp/Dynamic-Token-Pruning/mmseg_custom/loss/criterion.py 71e149dc0c33b040 |
unverified |
BSD-2-Clause (permissive) |
| MeMOTR: Long-Term Memory-Augmented Transformer for Multi-Object Tracking |
28 Jul 2023 |
mcg-nju/memotr/models/criterion.py a542c79696b8b5ae |
ran
|
MIT (permissive) |
| Towards Deeply Unified Depth-aware Panoptic Segmentation with Bi-directional Guidance Learning |
27 Jul 2023 |
jwh97nn/DeepDPS/model/loss.py 9b27de4fc1aeb35a |
ran
|
no licence file found · pointer only |
| Learning Dynamic Query Combinations for Transformer-based Object Detection and Segmentation |
23 Jul 2023 |
bytedance/DQ-Det/Cond-DETR-DQ/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Enhancing Your Trained DETRs with Box Refinement |
21 Jul 2023 |
yiqunchen1999/refinebox/refinebox/models/layers/criterion.py e584d922cadbfdce |
ran
|
licence not identified · pointer only |
| SwiFT: Swin 4D fMRI Transformer |
12 Jul 2023 |
Project-MONAI/MONAI/monai/losses/focal_loss.py 0a9007867090337a |
ran
|
Apache-2.0 (permissive) |
| Symphonize 3D Semantic Scene Completion with Contextual Instance Queries |
27 Jun 2023 |
hustvl/symphonies/maskdino/models/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| Segment Any Point Cloud Sequences by Distilling Vision Foundation Models |
15 Jun 2023 |
IDEA-Research/OpenSeeD/openseed/modules/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| Compositional Text-to-Image Synthesis with Attention Map Control of Diffusion Models |
23 May 2023 |
OPPO-Mente-Lab/attention-mask-control/boxnet_models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| AdaptiveClick: Clicks-aware Transformer with Adaptive Focal Loss for Interactive Image Segmentation |
7 May 2023 |
lab206/adaptiveclick/isegm/model/criterion.py 71e149dc0c33b040 |
unverified |
MIT (permissive) |
| Align-DETR: Enhancing End-to-end Object Detection with Aligned Loss |
15 Apr 2023 |
felixcaae/aligndetr/aligndetr/losses/losses.py dff0f0000acdc0ad |
ran
fingerprinted |
Apache-2.0 (permissive) |
| Detection Transformer with Stable Matching |
10 Apr 2023 |
IDEA-Research/Stable-DINO/projects/stabledino/modeling/dn_criterion.py 97059585f50ae421 |
unverified |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| Bridging Precision and Confidence: A Train-Time Loss for Calibrating Object Detection |
25 Mar 2023 |
identical code first harvested elsewhere 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| Dense-Localizing Audio-Visual Events in Untrimmed Videos: A Large-Scale Benchmark and Baseline |
22 Mar 2023 |
ttgeng233/UnAV/libs/modeling/losses.py 54293ea22bd4f952 |
ran
|
MIT (permissive) |
| TriDet: Temporal Action Detection with Relative Boundary Modeling |
13 Mar 2023 |
dingfengshi/TriDet/libs/modeling/losses.py 54293ea22bd4f952 |
ran
|
MIT (permissive) |
| Test-Time Distribution Normalization for Contrastively Learned Vision-language Models |
22 Feb 2023 |
identical code first harvested elsewhere 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
licence of this copy not recorded |
| MOSE: A New Dataset for Video Object Segmentation in Complex Scenes |
3 Feb 2023 |
henghuiding/MOSE-api/MOSEv2/sam2_rcms/training/loss_fns.py 42fec6fdc1faf6f1 |
unverified |
MIT (permissive) |
| ZegCLIP: Towards Adapting CLIP for Zero-shot Semantic Segmentation |
7 Dec 2022 |
ZiqinZhou66/ZegCLIP/models/losses/criterion.py 2675e89c94ae08bf |
ran
fingerprinted |
MIT (permissive) |
| NOPE-SAC: Neural One-Plane RANSAC for Sparse-View Planar 3D Reconstruction |
30 Nov 2022 |
icetttb/nopesac/NopeSAC_Net/modeling/criterion.py 71e149dc0c33b040 |
unverified |
MIT (permissive) |
| Teach-DETR: Better Training DETR with Teachers |
22 Nov 2022 |
leonhlj/teach-detr/H-Deformable-DETR/models/segmentation.py 155608ecaae7d2ec |
unverified |
MIT (permissive) |
| Fine-Grained Image Style Transfer with Visual Transformers |
11 Oct 2022 |
researchmm/sttr/models_istt/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| SiRi: A Simple Selective Retraining Mechanism for Transformer-based Visual Grounding |
27 Jul 2022 |
qumengxue/siri-vg/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| DETRs with Hybrid Matching |
26 Jul 2022 |
HDETR/H-Deformable-DETR/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |
| ReAct: Temporal Action Detection with Relational Queries |
14 Jul 2022 |
sssste/React/React/model/React.py 2043ee96aadac29d |
unverified |
MIT (permissive) |
| Mask DINO: Towards A Unified Transformer-based Framework for Object Detection and Segmentation |
6 Jun 2022 |
IDEA-opensource/DN-DETR/models/DN_DAB_DETR/DABDETR.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| PanopticDepth: A Unified Framework for Depth-aware Panoptic Segmentation |
1 Jun 2022 |
NaiyuGao/PanopticDepth/projects/PanopticDepth/panoptic_depth/loss.py 9b27de4fc1aeb35a |
ran
|
Apache-2.0 (permissive) |
| An Empirical Study of End-to-End Temporal Action Detection |
6 Apr 2022 |
xlliu7/E2E-TAD/models/custom_loss.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection |
7 Mar 2022 |
IDEACVR/MaskDINO/maskdino/modeling/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| DN-DETR: Accelerate DETR Training by Introducing Query DeNoising |
2 Mar 2022 |
FengLi-ust/DN-DETR/models/DN_DAB_DETR/DABDETR.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR |
28 Jan 2022 |
IDEA-Research/DAB-DETR/models/DAB_DETR/DABDETR.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| TransVOD: End-to-End Video Object Detection with Spatial-Temporal Transformers |
13 Jan 2022 |
SJTU-LuHe/TransVOD/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| QAHOI: Query-Based Anchors for Human-Object Interaction Detection |
16 Dec 2021 |
cjw2021/QAHOI/models/QAHOI.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 (permissive) |
| A Bilingual, OpenWorld Video Text Dataset and End-to-end Video Text Spotter with Transformer |
9 Dec 2021 |
weijiawu/transvtspotter/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
no licence file found · pointer only |
| End-to-End Referring Video Object Segmentation with Multimodal Transformers |
29 Nov 2021 |
mttr2021/MTTR/models/segmentation.py 71e149dc0c33b040 |
unverified |
Apache-2.0 (permissive) |
| ByteTrack: Multi-Object Tracking by Associating Every Detection Box |
13 Oct 2021 |
ifzhang/ByteTrack/yolox/models/losses.py 774f31ba06110943 |
ran
fingerprinted |
MIT (permissive) |
| Pix2seq: A Language Modeling Framework for Object Detection |
22 Sep 2021 |
gaopengcuhk/Unofficial-Pix2Seq/models/segmentation.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
Apache-2.0 recorded; this copy not marked cleared · pointer only |
| End-to-End Dense Video Captioning with Parallel Decoding |
17 Aug 2021 |
ttengwang/pdvc/pdvc/criterion.py 5c0711aada67957e |
ran · our draft was wrong
fingerprinted |
MIT (permissive) |