Browse State-of-the-Art › Knowledge Distillation › Papers, page 36
Knowledge Distillation
Papers archive 2025-07-28
archive papers tagged: 4,240 · with a code link: 1,740 · where Syntology ran a sample: 451 (380 with a run with no instrument failure, 71 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (451 of 4,240 tagged: 380 with a run with no instrument failure, 71 where every run was a failure of Syntology's instrument)
Page 36 of 43: papers 3,501 to 3,600 of 4,240, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Handling Long-tailed Feature Distribution in AdderNets1 Dec 2021 0 repositories listed
-
Unsupervised Representation Transfer for Small Networks: I Believe I Can Distill On-the-Fly1 Dec 2021 0 repositories listed
-
Using a GAN to Generate Adversarial Examples to Facial Image Recognition30 Nov 2021 0 repositories listed
-
Efficient Federated Learning for AIoT Applications Using Knowledge Distillation29 Nov 2021 0 repositories listed
-
Improved Knowledge Distillation via Adversarial Collaboration29 Nov 2021 0 repositories listed
-
ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection28 Nov 2021 0 repositories listed
-
Ensembling of Distilled Models from Multi-task Teachers for Constrained Resource Language Pairs26 Nov 2021 0 repositories listed
-
Domain-Agnostic Clustering with Self-Distillation23 Nov 2021 0 repositories listed
-
Contrast-reconstruction Representation Learning for Self-supervised Skeleton-based Action Recognition22 Nov 2021 0 repositories listed
-
Hierarchical Knowledge Distillation for Dialogue Sequence Labeling22 Nov 2021 0 repositories listed
-
Local-Selective Feature Distillation for Single Image Super-Resolution22 Nov 2021 0 repositories listed
-
Teacher-Student Training and Triplet Loss to Reduce the Effect of Drastic Face Occlusion20 Nov 2021 0 repositories listed
-
Toxicity Detection can be Sensitive to the Conversational Context19 Nov 2021 0 repositories listed
-
Dynamically pruning segformer for efficient semantic segmentation18 Nov 2021 0 repositories listed
-
Hierarchical Knowledge Guided Learning for Real-world Retinal Diseases Recognition17 Nov 2021 0 repositories listed
-
A Flexible Multi-Task Model for BERT Serving16 Nov 2021 0 repositories listed
-
Aligned Weight Regularizers for Pruning Pretrained Neural Networks16 Nov 2021 0 repositories listed
-
Compositional Data Augmentation for Abstractive Conversation Summarization16 Nov 2021 0 repositories listed
-
Deep-to-bottom Weights Decay: A Systemic Knowledge Review Learning Technique for Transformer Layers in Knowledge Distillation16 Nov 2021 0 repositories listed
-
16 Nov 2021 0 repositories listed
-
Feature Structure Distillation for BERT Transferring16 Nov 2021 0 repositories listed
-
Learning to Teach with Student Feedback16 Nov 2021 0 repositories listed
-
Making Small Language Models Better Few-Shot Learners16 Nov 2021 0 repositories listed
-
Multi-Granularity Contrastive Knowledge Distillation for Multimodal Named Entity Recognition16 Nov 2021 0 repositories listed
-
Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching16 Nov 2021 0 repositories listed
-
NVIDIA NeMo Neural Machine Translation Systems for English-German and English-Russian News and Biomedical Tasks at WMT2116 Nov 2021 0 repositories listed
-
One General Teacher for Multi-Data Multi-Task: A New Knowledge Distillation Framework for Discourse Relation Analysis16 Nov 2021 0 repositories listed
-
Redistributing Low-Frequency Words: Making the Most of Monolingual Data in Non-Autoregressive Translation16 Nov 2021 0 repositories listed
-
Self-Distilled Pruning of Neural Networks16 Nov 2021 0 repositories listed
-
Sparse Progressive Distillation: Resolving Overfitting under Pretrain-and-Finetune Paradigm16 Nov 2021 0 repositories listed
-
When Chosen Wisely, More Data Is What You Need: A Universal Sample-Efficient Strategy For Data Augmentation16 Nov 2021 0 repositories listed
-
Synthetic Unknown Class Learning for Learning Unknowns15 Nov 2021 0 repositories listed
-
Domain Generalization on Efficient Acoustic Scene Classification using Residual Normalization12 Nov 2021 0 repositories listed
-
Learning Interpretation with Explainable Knowledge Distillation12 Nov 2021 0 repositories listed
-
A Survey on Green Deep Learning8 Nov 2021 0 repositories listed
-
Class Token and Knowledge Distillation for Multi-head Self-Attention Speaker Verification Systems6 Nov 2021 0 repositories listed
-
AUTOKD: Automatic Knowledge Distillation Into A Student Architecture Family5 Nov 2021 0 repositories listed
-
DVFL: A Vertical Federated Learning Method for Dynamic Data5 Nov 2021 0 repositories listed
-
A methodology for training homomorphicencryption friendly neural networks5 Nov 2021 0 repositories listed
-
Oracle Teacher: Leveraging Target Information for Better Knowledge Distillation of CTC Models5 Nov 2021 0 repositories listed
-
Visualizing the Emergence of Intermediate Visual Patterns in DNNs5 Nov 2021 0 repositories listed
-
Leveraging Advantages of Interactive and Non-Interactive Models for Vector-Based Cross-Lingual Information Retrieval3 Nov 2021 0 repositories listed
-
Knowledge Cross-Distillation for Membership Privacy2 Nov 2021 0 repositories listed
-
AUTOSUMM: Automatic Model Creation for Text Summarization1 Nov 2021 0 repositories listed
-
Combining Curriculum Learning and Knowledge Distillation for Dialogue Generation1 Nov 2021 0 repositories listed
-
GAML-BERT: Improving BERT Early Exiting by Gradient Aligned Mutual Learning1 Nov 2021 0 repositories listed
-
How to Select One Among All ? An Empirical Study Towards the Robustness of Knowledge Distillation in Natural Language Understanding1 Nov 2021 0 repositories listed
-
HW-TSC’s Participation in the WMT 2021 Large-Scale Multilingual Translation Task1 Nov 2021 0 repositories listed
-
HW-TSC’s Participation in the WMT 2021 News Translation Shared Task1 Nov 2021 0 repositories listed
-
Limitations of Knowledge Distillation for Zero-shot Transfer Learning1 Nov 2021 0 repositories listed
-
Multilingual Neural Machine Translation: Can Linguistic Hierarchies Help?1 Nov 2021 0 repositories listed
-
Mutual-Learning Improves End-to-End Speech Translation1 Nov 2021 0 repositories listed
-
NVIDIA NeMo’s Neural Machine Translation Systems for English-German and English-Russian News and Biomedical Tasks at WMT211 Nov 2021 0 repositories listed
-
Papago’s Submission for the WMT21 Quality Estimation Shared Task1 Nov 2021 0 repositories listed
-
PDALN: Progressive Domain Adaptation over a Pre-trained Model for Low-Resource Cross-Domain Named Entity Recognition1 Nov 2021 0 repositories listed
-
RW-KD: Sample-wise Loss Terms Re-Weighting for Knowledge Distillation1 Nov 2021 0 repositories listed
-
Students Who Study Together Learn Better: On the Importance of Collective Knowledge Distillation for Domain Transfer in Fact Verification1 Nov 2021 0 repositories listed
-
TenTrans Large-Scale Multilingual Machine Translation System for WMT211 Nov 2021 0 repositories listed
-
The LMU Munich System for the WMT 2021 Large-Scale Multilingual Machine Translation Shared Task1 Nov 2021 0 repositories listed
-
The Mininglamp Machine Translation System for WMT211 Nov 2021 0 repositories listed
-
The NiuTrans System for the WMT 2021 Efficiency Task1 Nov 2021 0 repositories listed
-
Universal-KD: Attention-based Output-Grounded Intermediate Layer Knowledge Distillation1 Nov 2021 0 repositories listed
-
Rethinking the Knowledge Distillation From the Perspective of Model Calibration31 Oct 2021 0 repositories listed
-
Estimating and Maximizing Mutual Information for Knowledge Distillation29 Oct 2021 0 repositories listed
-
On Cross-Layer Alignment for Model Fusion of Heterogeneous Neural Networks29 Oct 2021 0 repositories listed
-
NxMTransformer: Semi-Structured Sparsification for Natural Language Understanding via ADMM28 Oct 2021 0 repositories listed
-
Towards Model Agnostic Federated Learning Using Knowledge Distillation28 Oct 2021 0 repositories listed
-
Beyond Classification: Knowledge Distillation using Multi-Object Impressions27 Oct 2021 0 repositories listed
-
27 Oct 2021 0 repositories listed
-
Response-based Distillation for Incremental Object Detection26 Oct 2021 0 repositories listed
-
MUSE: Feature Self-Distillation with Mutual Information and Self-Information25 Oct 2021 0 repositories listed
-
Reconstructing Pruned Filters using Cheap Spatial Transformations25 Oct 2021 0 repositories listed
-
24 Oct 2021 0 repositories listed
-
How and When Adversarial Robustness Transfers in Knowledge Distillation?22 Oct 2021 0 repositories listed
-
Pseudo Supervised Monocular Depth Estimation with Teacher-Student Network22 Oct 2021 0 repositories listed
-
Augmenting Knowledge Distillation With Peer-To-Peer Mutual Learning For Model Compression21 Oct 2021 0 repositories listed
-
Class Incremental Online Streaming Learning20 Oct 2021 0 repositories listed
-
Knowledge distillation from language model to acoustic model: a hierarchical multi-task learning approach20 Oct 2021 0 repositories listed
-
A Short Study on Compressing Decoder-Based Language Models16 Oct 2021 0 repositories listed
-
Know your tools well: Better and faster QA with synthetic examples16 Oct 2021 0 repositories listed
-
Pro-KD: Progressive Distillation by Following the Footsteps of the Teacher16 Oct 2021 0 repositories listed
-
Robustness Challenges in Model Distillation and Pruning for Natural Language Understanding16 Oct 2021 0 repositories listed
-
From Multimodal to Unimodal Attention in Transformers using Knowledge Distillation15 Oct 2021 0 repositories listed
-
Kronecker Decomposition for GPT Compression15 Oct 2021 0 repositories listed
-
Multilingual Neural Machine Translation:Can Linguistic Hierarchies Help?15 Oct 2021 0 repositories listed
-
Sparse Progressive Distillation: Resolving Overfitting under Pretrain-and-Finetune Paradigm15 Oct 2021 0 repositories listed
-
False Negative Distillation and Contrastive Learning for Personalized Outfit Recommendation13 Oct 2021 0 repositories listed
-
Language Modelling via Learning to Rank13 Oct 2021 0 repositories listed
-
Compact CNN Models for On-device Ocular-based User Recognition in Mobile Devices11 Oct 2021 0 repositories listed
-
11 Oct 2021 0 repositories listed
-
Towards Streaming Egocentric Action Anticipation11 Oct 2021 0 repositories listed
-
Visualizing the embedding space to explain the effect of knowledge distillation9 Oct 2021 0 repositories listed
-
Knowledge Distillation for Neural Transducers from Large Self-Supervised Pre-trained Models7 Oct 2021 0 repositories listed
-
Peer Collaborative Learning for Polyphonic Sound Event Detection7 Oct 2021 0 repositories listed
-
Online Hyperparameter Meta-Learning with Hypergradient Distillation6 Oct 2021 0 repositories listed
-
On the Interplay Between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis4 Oct 2021 0 repositories listed
-
Student Helping Teacher: Teacher Evolution via Self-Knowledge Distillation1 Oct 2021 0 repositories listed
-
Improving Neural Ranking via Lossless Knowledge Distillation30 Sep 2021 0 repositories listed
-
Deep Neural Compression Via Concurrent Pruning and Self-Distillation30 Sep 2021 0 repositories listed
-
A Comprehensive Overhaul of Distilling Unconditional GANs29 Sep 2021 0 repositories listed