Methods › Computer Vision › Image Models › Bottleneck Transformer
Bottleneck Transformer
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
The **Bottleneck Transformer (BoTNet) ** is an image classification model that incorporates self-attention for multiple computer vision tasks including image classification, object detection and instance segmentation. By just replacing the spatial convolutions with global self-attention in the final three bottleneck blocks of a ResNet and no other changes, the approach improves upon baselines significantly on instance segmentation and object detection while also reducing the parameters, with minimal overhead in latency.
Papers archive 2025-07-28
10 shown of 10, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Robust Multimodal Survival Prediction with the Latent Differentiation Conditional Variational AutoEncoder 12 Mar 2025 · 1 repository · arXiv:2503.09496Syntology ran 6 of 6 samples · 0 unverified · 6 pointer-only (licence)
-
Robust Multimodal Survival Prediction with Conditional Latent Differentiation Variational AutoEncoder 1 Jan 2025 · 0 repositories
-
Multi-scale Bottleneck Transformer for Weakly Supervised Multimodal Violence Detection 8 May 2024 · 1 repository · arXiv:2405.05130
-
Rock Classification Based on Residual Networks 19 Feb 2024 · 0 repositories · arXiv:2402.11831
-
SVFAP: Self-supervised Video Facial Affect Perceiver 31 Dec 2023 · 1 repository · arXiv:2401.00416
-
Learning Bottleneck Transformer for Event Image-Voxel Feature Fusion based Classification 23 Aug 2023 · 1 repository · arXiv:2308.11937Syntology ran 5 of 6 samples · 1 unverified · 6 pointer-only (licence)
-
Marine Debris Detection in Satellite Surveillance using Attention Mechanisms 9 Jul 2023 · 0 repositories · arXiv:2307.04128
-
Cross-Domain Synthetic-to-Real In-the-Wild Depth and Normal Estimation for 3D Scene Understanding 9 Dec 2022 · 0 repositories · arXiv:2212.05040
-
AGMB-Transformer: Anatomy-Guided Multi-Branch Transformer Network for Automated Evaluation of Root Canal Therapy 2 May 2021 · 1 repository · arXiv:2105.00381
-
Bottleneck Transformers for Visual Recognition 27 Jan 2021 · 13 repositories · arXiv:2101.11605Syntology ran 26 of 49 samples · 23 unverified · 8 pointer-only (licence)
Tasks archive 2025-07-28
20 shown of 23 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections