Methods › General › Knowledge Distillation › Knowledge Distillation › Papers, page 23
Knowledge Distillation
Papers archive 2025-07-28
archive papers tagged: 3,071 · with a code link: 1,258 · where Syntology ran a sample: 320 (276 with a run with no instrument failure, 44 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (320 of 3,071 tagged: 276 with a run with no instrument failure, 44 where every run was a failure of Syntology's instrument)
Page 23 of 31: papers 2,201 to 2,300 of 3,071, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Role of Data Augmentation Strategies in Knowledge Distillation for Wearable Sensor Data 1 Jan 2022 · 1 repository · arXiv:2201.00111
-
Conditional Generative Data-free Knowledge Distillation 31 Dec 2021 · 0 repositories · arXiv:2112.15358
-
Data-Free Knowledge Transfer: A Survey 31 Dec 2021 · 0 repositories · arXiv:2112.15278
-
Automatic Mixed-Precision Quantization Search of BERT 30 Dec 2021 · 0 repositories · arXiv:2112.14938
-
Confidence-Aware Multi-Teacher Knowledge Distillation 30 Dec 2021 · 1 repository · arXiv:2201.00007
-
Online Adversarial Knowledge Distillation for Graph Neural Networks 28 Dec 2021 · 1 repository · arXiv:2112.13966
-
Adaptive Beam Search to Enhance On-device Abstractive Summarization 22 Dec 2021 · 0 repositories · arXiv:2201.02739
-
Self-Distillation Mixup Training for Non-autoregressive Neural Machine Translation 22 Dec 2021 · 0 repositories · arXiv:2112.11640
-
Multi-Modality Distillation via Learning the teacher's modality-level Gram Matrix 21 Dec 2021 · 0 repositories · arXiv:2112.11447
-
Supervised Graph Contrastive Pretraining for Text Classification 21 Dec 2021 · 0 repositories · arXiv:2112.11389
-
Controlling the Quality of Distillation in Response-Based Network Compression 19 Dec 2021 · 0 repositories · arXiv:2112.10047
-
LegoDNN: Block-grained Scaling of Deep Neural Networks for Mobile Vision 18 Dec 2021 · 0 repositories · arXiv:2112.09852
-
Distillation of Human-Object Interaction Contexts for Action Recognition 17 Dec 2021 · 0 repositories · arXiv:2112.09448
-
Pixel Distillation: A New Knowledge Distillation Scheme for Low-Resolution Image Recognition 17 Dec 2021 · 1 repository · arXiv:2112.09532
-
Towards Disturbance-Free Visual Mobile Manipulation 17 Dec 2021 · 1 repository · arXiv:2112.12612
-
Weakly Supervised Semantic Segmentation via Alternative Self-Dual Teaching 17 Dec 2021 · 0 repositories · arXiv:2112.09459
-
Learning Cross-Lingual IR from an English Retriever 15 Dec 2021 · 1 repository · arXiv:2112.08185
-
A Deep Knowledge Distillation framework for EEG assisted enhancement of single-lead ECG based sleep staging 14 Dec 2021 · 1 repository · arXiv:2112.07252
-
Towards a Unified Foundation Model: Jointly Pre-Training Transformers on Unpaired Images and Text 14 Dec 2021 · 0 repositories · arXiv:2112.07074
-
Up to 100× Faster Data-free Knowledge Distillation 12 Dec 2021 · 2 repositories · arXiv:2112.06253
-
DistilCSE: Effective Knowledge Distillation For Contrastive Sentence Embeddings 10 Dec 2021 · 1 repository · arXiv:2112.05638
-
Human Guided Exploitation of Interpretable Attention Patterns in Summarization and Topic Segmentation 10 Dec 2021 · 1 repository · arXiv:2112.05364
-
Mask-invariant Face Recognition through Template-level Knowledge Distillation 10 Dec 2021 · 1 repository · arXiv:2112.05646
-
Boosting Contrastive Learning with Relation Knowledge Distillation 8 Dec 2021 · 0 repositories · arXiv:2112.04174
-
A Contrastive Distillation Approach for Incremental Semantic Segmentation in Aerial Images 7 Dec 2021 · 1 repository · arXiv:2112.03814
-
ADD: Frequency Attention and Multi-View based Knowledge Distillation to Detect Low-Quality Compressed Deepfake Images 7 Dec 2021 · 2 repositories · arXiv:2112.03553
-
Improving Neural Cross-Lingual Summarization via Employing Optimal Transport Distance for Knowledge Distillation 7 Dec 2021 · 1 repository · arXiv:2112.03473
-
CLASSIC: Continual and Contrastive Learning of Aspect Sentiment Classification Tasks 5 Dec 2021 · 1 repository · arXiv:2112.02714
-
Extracting knowledge from features with multilevel abstraction 4 Dec 2021 · 0 repositories · arXiv:2112.13642
-
KDCTime: Knowledge Distillation with Calibration on InceptionTime for Time-series Classification 4 Dec 2021 · 0 repositories · arXiv:2112.02291
-
A Fast Knowledge Distillation Framework for Visual Recognition 2 Dec 2021 · 2 repositories · arXiv:2112.01528
-
FedRAD: Federated Robust Adaptive Distillation 2 Dec 2021 · 0 repositories · arXiv:2112.01405
-
Tiny-NewsRec: Effective and Efficient PLM-based News Recommendation 2 Dec 2021 · 1 repository · arXiv:2112.00944
-
Distilling Meta Knowledge on Heterogeneous Graph for Illicit Drug Trafficker Detection on Social Media 1 Dec 2021 · 1 repository
-
The Augmented Image Prior: Distilling 1000 Classes by Extrapolating from a Single Image 1 Dec 2021 · 1 repository · arXiv:2112.00725
-
Information Theoretic Representation Distillation 1 Dec 2021 · 1 repository · arXiv:2112.00459
-
Shapeshifter: a Parameter-efficient Transformer using Factorized Reshaped Matrices 1 Dec 2021 · 1 repository
-
Unsupervised Representation Transfer for Small Networks: I Believe I Can Distill On-the-Fly 1 Dec 2021 · 0 repositories
-
Using a GAN to Generate Adversarial Examples to Facial Image Recognition 30 Nov 2021 · 0 repositories · arXiv:2111.15213
-
Efficient Federated Learning for AIoT Applications Using Knowledge Distillation 29 Nov 2021 · 0 repositories · arXiv:2111.14347
-
Improved Knowledge Distillation via Adversarial Collaboration 29 Nov 2021 · 0 repositories · arXiv:2111.14356
-
ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection 28 Nov 2021 · 0 repositories · arXiv:2111.14055
-
Ensembling of Distilled Models from Multi-task Teachers for Constrained Resource Language Pairs 26 Nov 2021 · 0 repositories · arXiv:2111.13284
-
WiFi-based Multi-task Sensing 26 Nov 2021 · 1 repository · arXiv:2111.14619
-
EvDistill: Asynchronous Events to End-task Learning via Bidirectional Reconstruction-guided Cross-modal Knowledge Distillation 24 Nov 2021 · 1 repository · arXiv:2111.12341Syntology official (archive's flag): 10 ran · 10 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Self-slimmed Vision Transformer 24 Nov 2021 · 1 repository · arXiv:2111.12624
-
Domain-Agnostic Clustering with Self-Distillation 23 Nov 2021 · 0 repositories · arXiv:2111.12170
-
Focal and Global Knowledge Distillation for Detectors 23 Nov 2021 · 1 repository · arXiv:2111.11837
-
Semi-Online Knowledge Distillation 23 Nov 2021 · 1 repository · arXiv:2111.11747
-
Hierarchical Knowledge Distillation for Dialogue Sequence Labeling 22 Nov 2021 · 0 repositories · arXiv:2111.10957
-
Local-Selective Feature Distillation for Single Image Super-Resolution 22 Nov 2021 · 0 repositories · arXiv:2111.10988
-
Teacher-Student Training and Triplet Loss to Reduce the Effect of Drastic Face Occlusion 20 Nov 2021 · 0 repositories · arXiv:2111.10561
-
Toxicity Detection can be Sensitive to the Conversational Context 19 Nov 2021 · 0 repositories · arXiv:2111.10223
-
Dynamically pruning segformer for efficient semantic segmentation 18 Nov 2021 · 0 repositories · arXiv:2111.09499
-
Hierarchical Knowledge Guided Learning for Real-world Retinal Diseases Recognition 17 Nov 2021 · 0 repositories · arXiv:2111.08913
-
An Unsupervised Multiple-Task and Multiple-Teacher Model for Cross-lingual Named Entity Recognition 16 Nov 2021 · 1 repository
-
Compositional Data Augmentation for Abstractive Conversation Summarization 16 Nov 2021 · 0 repositories
-
Deep-to-bottom Weights Decay: A Systemic Knowledge Review Learning Technique for Transformer Layers in Knowledge Distillation 16 Nov 2021 · 0 repositories
-
Enabling Multimodal Generation on CLIP via Vision-Language Knowledge Distillation 16 Nov 2021 · 0 repositories
-
Learning to Teach with Student Feedback 16 Nov 2021 · 0 repositories
-
Making Small Language Models Better Few-Shot Learners 16 Nov 2021 · 0 repositories
-
Multi-Granularity Contrastive Knowledge Distillation for Multimodal Named Entity Recognition 16 Nov 2021 · 0 repositories
-
Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching 16 Nov 2021 · 0 repositories
-
NVIDIA NeMo Neural Machine Translation Systems for English-German and English-Russian News and Biomedical Tasks at WMT21 16 Nov 2021 · 0 repositories · arXiv:2111.08634
-
One General Teacher for Multi-Data Multi-Task: A New Knowledge Distillation Framework for Discourse Relation Analysis 16 Nov 2021 · 0 repositories
-
Sparse Progressive Distillation: Resolving Overfitting under Pretrain-and-Finetune Paradigm 16 Nov 2021 · 0 repositories
-
When Chosen Wisely, More Data Is What You Need: A Universal Sample-Efficient Strategy For Data Augmentation 16 Nov 2021 · 0 repositories
-
Synthetic Unknown Class Learning for Learning Unknowns 15 Nov 2021 · 0 repositories · arXiv:2111.08062
-
Facial Landmark Points Detection Using Knowledge Distillation-Based Neural Networks 13 Nov 2021 · 1 repository · arXiv:2111.07047
-
Learning Interpretation with Explainable Knowledge Distillation 12 Nov 2021 · 0 repositories · arXiv:2111.06945
-
Incremental Meta-Learning via Episodic Replay Distillation for Few-Shot Image Recognition 9 Nov 2021 · 1 repository · arXiv:2111.04993
-
Class Token and Knowledge Distillation for Multi-head Self-Attention Speaker Verification Systems 6 Nov 2021 · 0 repositories · arXiv:2111.03842
-
AUTOKD: Automatic Knowledge Distillation Into A Student Architecture Family 5 Nov 2021 · 0 repositories · arXiv:2111.03555
-
LTD: Low Temperature Distillation for Robust Adversarial Training 3 Nov 2021 · 1 repository · arXiv:2111.02331
-
Knowledge Cross-Distillation for Membership Privacy 2 Nov 2021 · 0 repositories · arXiv:2111.01363
-
AUTOSUMM: Automatic Model Creation for Text Summarization 1 Nov 2021 · 0 repositories
-
Collaborative Learning of Bidirectional Decoders for Unsupervised Text Style Transfer 1 Nov 2021 · 1 repository
-
Combining Curriculum Learning and Knowledge Distillation for Dialogue Generation 1 Nov 2021 · 0 repositories
-
Distilling Knowledge for Empathy Detection 1 Nov 2021 · 1 repository
-
Distilling Object Detectors with Feature Richness 1 Nov 2021 · 1 repository · arXiv:2111.00674
-
GAML-BERT: Improving BERT Early Exiting by Gradient Aligned Mutual Learning 1 Nov 2021 · 0 repositories
-
Improving Stance Detection with Multi-Dataset Learning and Knowledge Distillation 1 Nov 2021 · 1 repository
-
Limitations of Knowledge Distillation for Zero-shot Transfer Learning 1 Nov 2021 · 0 repositories
-
Multilingual Neural Machine Translation: Can Linguistic Hierarchies Help? 1 Nov 2021 · 0 repositories
-
Mutual-Learning Improves End-to-End Speech Translation 1 Nov 2021 · 0 repositories
-
PDALN: Progressive Domain Adaptation over a Pre-trained Model for Low-Resource Cross-Domain Named Entity Recognition 1 Nov 2021 · 0 repositories
-
PP-ShiTu: A Practical Lightweight Image Recognition System 1 Nov 2021 · 2 repositories · arXiv:2111.00775
-
Students Who Study Together Learn Better: On the Importance of Collective Knowledge Distillation for Domain Transfer in Fact Verification 1 Nov 2021 · 0 repositories
-
Universal-KD: Attention-based Output-Grounded Intermediate Layer Knowledge Distillation 1 Nov 2021 · 0 repositories
-
Estimating and Maximizing Mutual Information for Knowledge Distillation 29 Oct 2021 · 0 repositories · arXiv:2110.15946
-
On Cross-Layer Alignment for Model Fusion of Heterogeneous Neural Networks 29 Oct 2021 · 0 repositories · arXiv:2110.15538
-
Towards Model Agnostic Federated Learning Using Knowledge Distillation 28 Oct 2021 · 0 repositories · arXiv:2110.15210
-
Temporal Knowledge Distillation for On-device Audio Classification 27 Oct 2021 · 0 repositories · arXiv:2110.14131
-
Response-based Distillation for Incremental Object Detection 26 Oct 2021 · 0 repositories · arXiv:2110.13471
-
Instance-Conditional Knowledge Distillation for Object Detection 25 Oct 2021 · 1 repository · arXiv:2110.12724
-
MUSE: Feature Self-Distillation with Mutual Information and Self-Information 25 Oct 2021 · 0 repositories · arXiv:2110.12606
-
Reconstructing Pruned Filters using Cheap Spatial Transformations 25 Oct 2021 · 0 repositories · arXiv:2110.12844
-
X-Distill: Improving Self-Supervised Monocular Depth via Cross-Task Distillation 24 Oct 2021 · 0 repositories · arXiv:2110.12516
-
Pixel-by-Pixel Cross-Domain Alignment for Few-Shot Semantic Segmentation 22 Oct 2021 · 1 repository · arXiv:2110.11650
-
Knowledge distillation from language model to acoustic model: a hierarchical multi-task learning approach 20 Oct 2021 · 0 repositories · arXiv:2110.10429