Methods › Computer Vision › Vision Transformers › DINO › Papers, page 2
self-DIstillation with NO labels
DINO
Papers archive 2025-07-28
archive papers tagged: 208 · with a code link: 105 · where Syntology ran a sample: 30 (28 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (30 of 208 tagged: 28 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 208, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Swiss DINO: Efficient and Versatile Vision Framework for On-device Personal Object Search 10 Jul 2024 · 1 repository · arXiv:2407.07541
-
Segment Anything Model for automated image data annotation: empirical studies using text prompts from Grounding DINO 27 Jun 2024 · 0 repositories · arXiv:2406.19057
-
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification 20 Jun 2024 · 1 repository · arXiv:2406.14370
-
Liveness Detection in Computer Vision: Transformer-based Self-Supervised Learning for Face Anti-Spoofing 19 Jun 2024 · 0 repositories · arXiv:2406.13860
-
MixDiff: Mixing Natural and Synthetic Images for Robust Self-Supervised Representations 18 Jun 2024 · 1 repository · arXiv:2406.12368
-
ICE-G: Image Conditional Editing of 3D Gaussian Splats 12 Jun 2024 · 0 repositories · arXiv:2406.08488
-
UVIS: Unsupervised Video Instance Segmentation 11 Jun 2024 · 0 repositories · arXiv:2406.06908
-
A Comparative Survey of Vision Transformers for Feature Extraction in Texture Analysis 10 Jun 2024 · 0 repositories · arXiv:2406.06136
-
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP 3 Jun 2024 · 1 repository · arXiv:2406.01583Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
ELSA: Evaluating Localization of Social Activities in Urban Streets using Open-Vocabulary Detection 3 Jun 2024 · 0 repositories · arXiv:2406.01551
-
Eating Smart: Advancing Health Informatics with the Grounding DINO based Dietary Assistant App 2 Jun 2024 · 0 repositories · arXiv:2406.00848
-
FDQN: A Flexible Deep Q-Network Framework for Game Automation 29 May 2024 · 1 repository · arXiv:2405.18761
-
Adapting Pre-Trained Vision Models for Novel Instance Detection and Segmentation 28 May 2024 · 1 repository · arXiv:2405.17859Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
SA-GS: Semantic-Aware Gaussian Splatting for Large Scene Reconstruction with Geometry Constrain 27 May 2024 · 0 repositories · arXiv:2405.16923
-
Designing A Sustainable Marine Debris Clean-up Framework without Human Labels 23 May 2024 · 1 repository · arXiv:2405.14815
-
Text Prompting for Multi-Concept Video Customization by Autoregressive Generation 22 May 2024 · 0 repositories · arXiv:2405.13951
-
Track Anything Rapter(TAR) 19 May 2024 · 1 repository · arXiv:2405.11655
-
DINO as a von Mises-Fisher mixture model 17 May 2024 · 0 repositories · arXiv:2405.10939
-
Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection 16 May 2024 · 3 repositories · arXiv:2405.10300Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Multi-method Integration with Confidence-based Weighting for Zero-shot Image Classification 3 May 2024 · 0 repositories · arXiv:2405.02155
-
Exploring Self-Supervised Vision Transformers for Deepfake Detection: A Comparative Analysis 1 May 2024 · 1 repository · arXiv:2405.00355
-
Masked Multi-Query Slot Attention for Unsupervised Object Discovery 30 Apr 2024 · 1 repository · arXiv:2404.19654
-
Parameter Efficient Fine-tuning of Self-supervised ViTs without Catastrophic Forgetting 26 Apr 2024 · 1 repository · arXiv:2404.17245
-
Boosting Unsupervised Semantic Segmentation with Principal Mask Proposals 25 Apr 2024 · 1 repository · arXiv:2404.16818
-
SPARO: Selective Attention for Robust and Compositional Transformer Encodings for Vision 24 Apr 2024 · 1 repository · arXiv:2404.15721
-
1st Place Solution to the 1st SkatingVerse Challenge 22 Apr 2024 · 0 repositories · arXiv:2404.14032
-
FiLo: Zero-Shot Anomaly Detection by Fine-Grained Description and High-Quality Localization 21 Apr 2024 · 1 repository · arXiv:2404.13671Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 7 pointer-only (licence)
-
Vim4Path: Self-Supervised Vision Mamba for Histopathology Images 20 Apr 2024 · 1 repository · arXiv:2404.13222Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
The devil is in the object boundary: towards annotation-free instance segmentation using Foundation Models 18 Apr 2024 · 1 repository · arXiv:2404.11957Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Hyperbolic Learning with Synthetic Captions for Open-World Detection 7 Apr 2024 · 0 repositories · arXiv:2404.05016
-
Cluster-based Video Summarization with Temporal Context Awareness 6 Apr 2024 · 1 repository · arXiv:2404.04511
-
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views 4 Apr 2024 · 0 repositories · arXiv:2404.03650
-
Open-Vocabulary Object Detectors: Robustness Challenges under Distribution Shifts 1 Apr 2024 · 0 repositories · arXiv:2405.14874
-
Illicit object detection in X-ray images using Vision Transformers 27 Mar 2024 · 0 repositories · arXiv:2403.19043
-
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D 27 Mar 2024 · 0 repositories · arXiv:2403.18922
-
Towards Large-Scale Training of Pathology Foundation Models 24 Mar 2024 · 2 repositories · arXiv:2404.15217Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Unsupervised Audio-Visual Segmentation with Modality Alignment 21 Mar 2024 · 0 repositories · arXiv:2403.14203
-
TAG: Guidance-free Open-Vocabulary Semantic Segmentation 17 Mar 2024 · 1 repository · arXiv:2403.11197
-
Derivative-informed neural operator acceleration of geometric MCMC for infinite-dimensional Bayesian inverse problems 13 Mar 2024 · 1 repository · arXiv:2403.08220Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Self-Supervised Multiple Instance Learning for Acute Myeloid Leukemia Classification 8 Mar 2024 · 0 repositories · arXiv:2403.05379
-
AO-DETR: Anti-Overlapping DETR for X-Ray Prohibited Items Detection 7 Mar 2024 · 1 repository · arXiv:2403.04309
-
Deformable One-shot Face Stylization via DINO Semantic Guidance 1 Mar 2024 · 1 repository · arXiv:2403.00459
-
Self-supervised Visualisation of Medical Image Datasets 22 Feb 2024 · 1 repository · arXiv:2402.14566
-
DINOBot: Robot Manipulation via Retrieval and Alignment with Vision Foundation Models 20 Feb 2024 · 0 repositories · arXiv:2402.13181
-
BioFusionNet: Deep Learning-Based Survival Risk Stratification in ER+ Breast Cancer Through Multifeature and Multimodal Data Fusion 16 Feb 2024 · 1 repository · arXiv:2402.10717
-
Just Cluster It: An Approach for Exploration in High-Dimensions using Clustering and Pre-Trained Representations 5 Feb 2024 · 1 repository · arXiv:2402.03138Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
A Probabilistic Model Behind Self-Supervised Learning 2 Feb 2024 · 1 repository · arXiv:2402.01399Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
From Training-Free to Adaptive: Empirical Insights into MLLMs' Understanding of Detection Information 31 Jan 2024 · 0 repositories · arXiv:2401.17981
-
Cross-Domain Few-Shot Learning via Adaptive Transformer Networks 25 Jan 2024 · 1 repository · arXiv:2401.13987
-
Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks 25 Jan 2024 · 5 repositories · arXiv:2401.14159Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples) · 5 pointer-only (licence)
-
DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer 23 Jan 2024 · 1 repository · arXiv:2401.12820
-
A Novel Benchmark for Few-Shot Semantic Segmentation in the Era of Foundation Models 20 Jan 2024 · 1 repository · arXiv:2401.11311
-
Image Similarity using An Ensemble of Context-Sensitive Models 15 Jan 2024 · 1 repository · arXiv:2401.07951
-
Learning Segmented 3D Gaussians via Efficient Feature Unprojection for Zero-shot Neural Scene Segmentation 11 Jan 2024 · 0 repositories · arXiv:2401.05925
-
Surgical-DINO: Adapter Learning of Foundation Models for Depth Estimation in Endoscopic Surgery 11 Jan 2024 · 1 repository · arXiv:2401.06013Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Attention-Guided Erasing: A Novel Augmentation Method for Enhancing Downstream Breast Density Classification 8 Jan 2024 · 0 repositories · arXiv:2401.03912
-
A Novel Transformer-Based Self-Supervised Learning Method to Enhance Photoplethysmogram Signal Artifact Detection 2 Jan 2024 · 0 repositories · arXiv:2401.01013
-
KD-DETR: Knowledge Distillation for Detection Transformer with Consistent Distillation Points Sampling 1 Jan 2024 · 0 repositories
-
Analyzing Local Representations of Self-supervised Vision Transformers 31 Dec 2023 · 0 repositories · arXiv:2401.00463
-
Learning Vision from Models Rivals Learning Vision from Data 28 Dec 2023 · 2 repositories · arXiv:2312.17742Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
LangSplat: 3D Language Gaussian Splatting 26 Dec 2023 · 1 repository · arXiv:2312.16084Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Unsupervised Segmentation of Colonoscopy Images 19 Dec 2023 · 0 repositories · arXiv:2312.12599
-
Guided Diffusion from Self-Supervised Diffusion Features 14 Dec 2023 · 0 repositories · arXiv:2312.08825
-
Mixed Pseudo Labels for Semi-Supervised Object Detection 12 Dec 2023 · 1 repository · arXiv:2312.07006
-
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection 4 Dec 2023 · 1 repository · arXiv:2312.01576
-
Multi-task Image Restoration Guided By Robust DINO Features 4 Dec 2023 · 0 repositories · arXiv:2312.01677
-
A Lightweight Clustering Framework for Unsupervised Semantic Segmentation 30 Nov 2023 · 0 repositories · arXiv:2311.18628
-
HiFi Tuner: High-Fidelity Subject-Driven Fine-Tuning for Diffusion Models 30 Nov 2023 · 0 repositories · arXiv:2312.00079
-
Knowledge Transfer from Vision Foundation Models for Efficient Training of Small Task-specific Models 30 Nov 2023 · 1 repository · arXiv:2311.18237
-
Betrayed by Attention: A Simple yet Effective Approach for Self-supervised Video Object Segmentation 29 Nov 2023 · 1 repository · arXiv:2311.17893Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
DyRA: Portable Dynamic Resolution Adjustment Network for Existing Detectors 28 Nov 2023 · 2 repositories · arXiv:2311.17098
-
Understanding Self-Supervised Features for Learning Unsupervised Instance Segmentation 24 Nov 2023 · 0 repositories · arXiv:2311.14665
-
Feature Extraction for Generative Medical Imaging Evaluation: New Evidence Against an Evolving Trend 22 Nov 2023 · 2 repositories · arXiv:2311.13717
-
White-Box Transformers via Sparse Rate Reduction: Compression Is All There Is? 22 Nov 2023 · 1 repository · arXiv:2311.13110Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Large-Scale Car Parts (LSCP) Dataset for Lightweight Fine-Grained Detection 20 Nov 2023 · 0 repositories · arXiv:2311.11754
-
FreeKD: Knowledge Distillation via Semantic Frequency Prompt 20 Nov 2023 · 1 repository · arXiv:2311.12079Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
UnifiedVisionGPT: Streamlining Vision-Oriented AI through Generalized Multimodal Framework 16 Nov 2023 · 1 repository · arXiv:2311.10125
-
Generalizable Imitation Learning Through Pre-Trained Representations 15 Nov 2023 · 0 repositories · arXiv:2311.09350
-
Cal-DETR: Calibrated Detection Transformer 6 Nov 2023 · 1 repository · arXiv:2311.03570Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Masking Hyperspectral Imaging Data with Pretrained Models 6 Nov 2023 · 1 repository · arXiv:2311.03053
-
A Self-Supervised Approach to Land Cover Segmentation 27 Oct 2023 · 0 repositories · arXiv:2310.18251
-
SAMCLR: Contrastive pre-training on complex scenes using SAM for view sampling 23 Oct 2023 · 0 repositories · arXiv:2310.14736
-
From CLIP to DINO: Visual Encoders Shout in Multi-modal Large Language Models 13 Oct 2023 · 1 repository · arXiv:2310.08825
-
Computational Pathology at Health System Scale -- Self-Supervised Foundation Models from Three Billion Images 10 Oct 2023 · 0 repositories · arXiv:2310.07033
-
A General Protocol to Probe Large Vision Models for 3D Physical Understanding 10 Oct 2023 · 1 repository · arXiv:2310.06836
-
Exploring DINO: Emergent Properties and Limitations for Synthetic Aperture Radar Imagery 5 Oct 2023 · 0 repositories · arXiv:2310.03513
-
Beyond Random Augmentations: Pretraining with Hard Views 5 Oct 2023 · 2 repositories · arXiv:2310.03940
-
Adapting Vision Foundation Models for Plant Phenotyping 1 Oct 2023 · 0 repositories
-
D³Fields: Dynamic 3D Descriptor Fields for Zero-Shot Generalizable Rearrangement 28 Sep 2023 · 0 repositories · arXiv:2309.16118
-
A SAM-based Solution for Hierarchical Panoptic Segmentation of Crops and Weeds Competition 24 Sep 2023 · 0 repositories · arXiv:2309.13578
-
DAC-DETR: Divide the Attention Layers and Conquer 21 Sep 2023 · 1 repository
-
FLSL: Feature-level Self-supervised Learning 21 Sep 2023 · 1 repository
-
LEPARD: Learning Explicit Part Discovery for 3D Articulated Shape Reconstruction 21 Sep 2023 · 0 repositories
-
Leveraging In-the-Wild Data for Effective Self-Supervised Pretraining in Speaker Recognition 21 Sep 2023 · 1 repository · arXiv:2309.11730
-
[Re] Masked Autoencoders Are Small Scale Vision Learners: A Reproduction Under Resource Constraints 21 Sep 2023 · 1 repository
-
Language Embedded Radiance Fields for Zero-Shot Task-Oriented Grasping 14 Sep 2023 · 0 repositories · arXiv:2309.07970
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
Adapting Self-Supervised Representations to Multi-Domain Setups 7 Sep 2023 · 0 repositories · arXiv:2309.03999
-
Masking Strategies for Background Bias Removal in Computer Vision Models 23 Aug 2023 · 1 repository · arXiv:2308.12127
-
Emergent Correspondence from Image Diffusion 6 Jun 2023 · 2 repositories · arXiv:2306.03881Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 2 pointer-only (licence)