Browse State-of-the-Art › Model Compression › Papers, page 9
Model Compression
Papers archive 2025-07-28
archive papers tagged: 1,356 · with a code link: 440 · where Syntology ran a sample: 119 (97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,356 tagged: 97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument)
Page 9 of 14: papers 801 to 900 of 1,356, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
ConaCLIP: Exploring Distillation of Fully-Connected Knowledge Interaction Graph for Lightweight Text-Image Retrieval28 May 2023 0 repositories listed
-
2-bit Conformer quantization for automatic speech recognition26 May 2023 0 repositories listed
-
RAND: Robustness Aware Norm Decay For Quantized Seq2seq Models24 May 2023 0 repositories listed
-
Revisiting Data Augmentation in Model Compression: An Empirical and Comprehensive Study22 May 2023 0 repositories listed
-
Compress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable Prompt17 May 2023 0 repositories listed
-
CrAFT: Compression-Aware Fine-Tuning for Efficient Visual Task Adaptation8 May 2023 0 repositories listed
-
Redundancy and Concept Analysis for Code-trained Language Models1 May 2023 0 repositories listed
-
CORSD: Class-Oriented Relational Self Distillation28 Apr 2023 0 repositories listed
-
Guaranteed Quantization Error Computation for Neural Network Model Compression26 Apr 2023 0 repositories listed
-
Bias in Pruned Vision Models: In-Depth Analysis and Countermeasures25 Apr 2023 0 repositories listed
-
Deep Collective Knowledge Distillation18 Apr 2023 0 repositories listed
-
Structured Pruning for Multi-Task Deep Neural Networks13 Apr 2023 0 repositories listed
-
Surrogate Lagrangian Relaxation: A Path To Retrain-free Deep Neural Network Pruning8 Apr 2023 0 repositories listed
-
oBERTa: Improving Sparse Transfer Learning via improved initialization, distillation, and pruning regimes30 Mar 2023 0 repositories listed
-
A Multi-objective Complex Network Pruning Framework Based on Divide-and-conquer and Global Performance Impairment Ranking28 Mar 2023 0 repositories listed
-
Information-Theoretic GAN Compression with Variational Energy-based Model28 Mar 2023 0 repositories listed
-
Tetra-AML: Automatic Machine Learning via Tensor Networks28 Mar 2023 0 repositories listed
-
Towards Accurate Post-Training Quantization for Vision Transformer25 Mar 2023 0 repositories listed
-
Exploring Turkish Speech Recognition via Hybrid CTC/Attention Architecture and Multi-feature Fusion Network22 Mar 2023 0 repositories listed
-
Low Rank Optimization for Efficient Deep Learning: Making A Balance between Compact Architecture and Fast Training22 Mar 2023 0 repositories listed
-
14 Mar 2023 0 repositories listed
-
Greener yet Powerful: Taming Large Code Generation Models with Quantization9 Mar 2023 0 repositories listed
-
Gradient-Free Structured Pruning with Unlabeled Data7 Mar 2023 0 repositories listed
-
Adversarial Attacks on Machine Learning in Embedded and IoT Platforms3 Mar 2023 0 repositories listed
-
Towards domain generalisation in ASR with elitist sampling and ensemble knowledge distillation1 Mar 2023 0 repositories listed
-
Debiased Distillation by Transplanting the Last Layer22 Feb 2023 0 repositories listed
-
Structured Bayesian Compression for Deep Neural Networks Based on The Turbo-VBI Approach21 Feb 2023 0 repositories listed
-
HomoDistil: Homotopic Task-Agnostic Distillation of Pre-trained Transformers19 Feb 2023 0 repositories listed
-
A Comprehensive Review and a Taxonomy of Edge Machine Learning: Requirements, Paradigms, and Techniques16 Feb 2023 0 repositories listed
-
Towards Optimal Compression: Joint Pruning and Quantization15 Feb 2023 0 repositories listed
-
On Achieving Privacy-Preserving State-of-the-Art Edge Intelligence10 Feb 2023 0 repositories listed
-
Knowledge Distillation in Vision Transformers: A Critical Review4 Feb 2023 0 repositories listed
-
Generalized Uncertainty of Deep Neural Networks: Taxonomy and Applications2 Feb 2023 0 repositories listed
-
Knowledge Distillation on Graphs: A Survey1 Feb 2023 0 repositories listed
-
AMD: Adaptive Masked Distillation for Object Detection31 Jan 2023 0 repositories listed
-
Improved knowledge distillation by utilizing backward pass knowledge in neural networks27 Jan 2023 0 repositories listed
-
HALOC: Hardware-Aware Automatic Low-Rank Compression for Compact Neural Networks20 Jan 2023 0 repositories listed
-
HCE: Improving Performance and Efficiency with Heterogeneously Compressed Neural Network Ensemble18 Jan 2023 0 repositories listed
-
Distilling Focal Knowledge From Imperfect Expert for 3D Object Detection1 Jan 2023 0 repositories listed
-
ICD-Face: Intra-class Compactness Distillation for Face Recognition1 Jan 2023 0 repositories listed
-
Memory-Friendly Scalable Super-Resolution via Rewinding Lottery Ticket Hypothesis1 Jan 2023 0 repositories listed
-
One-Shot Model for Mixed-Precision Quantization1 Jan 2023 0 repositories listed
-
Tiny Updater: Towards Efficient Neural Network-Driven Software Updating1 Jan 2023 0 repositories listed
-
FlatENN: Train Flat for Enhanced Fault Tolerance of Quantized Deep Neural Networks29 Dec 2022 0 repositories listed
-
BD-KD: Balancing the Divergences for Online Knowledge Distillation25 Dec 2022 0 repositories listed
-
FSCNN: A Fast Sparse Convolution Neural Network Inference System17 Dec 2022 0 repositories listed
-
Can We Find Strong Lottery Tickets in Generative Models?16 Dec 2022 0 repositories listed
-
Swing Distillation: A Privacy-Preserving Knowledge Distillation Framework16 Dec 2022 0 repositories listed
-
Efficient Speech Representation Learning with Low-Bit Quantization14 Dec 2022 0 repositories listed
-
Error-aware Quantization through Noise Tempering11 Dec 2022 0 repositories listed
-
Leveraging Different Learning Styles for Improved Knowledge Distillation in Biomedical Imaging6 Dec 2022 0 repositories listed
-
CSTAR: Towards Compact and STructured Deep Neural Networks with Adversarial Robustness4 Dec 2022 0 repositories listed
-
GlueFL: Reconciling Client Sampling and Model Masking for Bandwidth Efficient Federated Learning3 Dec 2022 0 repositories listed
-
Compressing Cross-Lingual Multi-Task Models at Qualtrics29 Nov 2022 0 repositories listed
-
Design and Prototyping Distributed CNN Inference Acceleration in Edge Computing24 Nov 2022 0 repositories listed
-
Learning Low-Rank Representations for Model Compression21 Nov 2022 0 repositories listed
-
Edge-MultiAI: Multi-Tenancy of Latency-Sensitive Deep Learning Applications on Edge14 Nov 2022 0 repositories listed
-
XAI-BayesHAR: A novel Framework for Human Activity Recognition with Integrated Uncertainty and Shapely Values7 Nov 2022 0 repositories listed
-
Model Compression for DNN-based Speaker Verification Using Weight Quantization31 Oct 2022 0 repositories listed
-
Online Cross-Layer Knowledge Distillation on Graph Neural Networks with Deep Supervision25 Oct 2022 0 repositories listed
-
Legal-Tech Open Diaries: Lesson learned on how to develop and deploy light-weight models in the era of humongous Language Models24 Oct 2022 0 repositories listed
-
Outsourcing Training without Uploading Data via Efficient Collaborative Open-Source Sampling23 Oct 2022 0 repositories listed
-
Sub-network Multi-objective Evolutionary Algorithm for Filter Pruning22 Oct 2022 0 repositories listed
-
Data-Model-Circuit Tri-Design for Ultra-Light Video Intelligence on Edge Devices16 Oct 2022 0 repositories listed
-
FIT: A Metric for Model Sensitivity16 Oct 2022 0 repositories listed
-
Boosting Graph Neural Networks via Adaptive Knowledge Distillation12 Oct 2022 0 repositories listed
-
SeKron: A Decomposition Method Supporting Many Factorization Structures12 Oct 2022 0 repositories listed
-
Deep learning model compression using network sensitivity and gradients11 Oct 2022 0 repositories listed
-
AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models8 Oct 2022 0 repositories listed
-
Multi-stage Progressive Compression of Conformer Transducer for On-device Speech Recognition1 Oct 2022 0 repositories listed
-
Match to Win: Analysing Sequences Lengths for Efficient Self-supervised Learning in Speech and Audio30 Sep 2022 0 repositories listed
-
Analysis of Quantization on MLP-based Vision Models14 Sep 2022 0 repositories listed
-
SaleNet: A low-power end-to-end CNN accelerator for sustained attention level evaluation using EEG3 Sep 2022 0 repositories listed
-
Complexity-Driven CNN Compression for Resource-constrained Edge AI26 Aug 2022 0 repositories listed
-
Reducing Computational Complexity of Neural Networks in Optical Channel Equalization: From Concepts to Implementation26 Aug 2022 0 repositories listed
-
Design Automation for Fast, Lightweight, and Effective Deep Learning Models: A Survey22 Aug 2022 0 repositories listed
-
Enhancing Targeted Attack Transferability via Diversified Weight Pruning18 Aug 2022 0 repositories listed
-
An Algorithm-Hardware Co-Optimized Framework for Accelerating N:M Sparse Transformers12 Aug 2022 0 repositories listed
-
Triple Sparsification of Graph Convolutional Networks without Sacrificing the Accuracy6 Aug 2022 0 repositories listed
-
Model Blending for Text Classification5 Aug 2022 0 repositories listed
-
Quiver neural networks26 Jul 2022 0 repositories listed
-
Model Compression for Resource-Constrained Mobile Robots20 Jul 2022 0 repositories listed
-
T-RECX: Tiny-Resource Efficient Convolutional neural networks with early-eXit14 Jul 2022 0 repositories listed
-
Normalized Feature Distillation for Semantic Segmentation12 Jul 2022 0 repositories listed
-
Rank-Based Filter Pruning for Real-Time UAV Tracking5 Jul 2022 0 repositories listed
-
Quantum Neural Network Compression4 Jul 2022 0 repositories listed
-
KroneckerBERT: Significant Compression of Pre-trained Language Models Through Kronecker Decomposition and Knowledge Distillation1 Jul 2022 0 repositories listed
-
Language model compression with weighted low-rank factorization30 Jun 2022 0 repositories listed
-
QUIDAM: A Framework for Quantization-Aware DNN Accelerator and Model Co-Exploration30 Jun 2022 0 repositories listed
-
Fundamental Limits of Communication Efficiency for Model Aggregation in Distributed Learning: A Rate-Distortion Approach28 Jun 2022 0 repositories listed
-
QTI Submission to DCASE 2021: residual normalization for device-imbalanced acoustic scene classification with efficient design28 Jun 2022 0 repositories listed
-
Representative Teacher Keys for Knowledge Distillation Model Compression Based on Attention Mechanism for Image Classification26 Jun 2022 0 repositories listed
-
An Automatic and Efficient BERT Pruning for Edge AI Systems21 Jun 2022 0 repositories listed
-
Knowledge Distillation for Oriented Object Detection on Aerial Images20 Jun 2022 0 repositories listed
-
Revisiting Self-Distillation17 Jun 2022 0 repositories listed
-
Accelerating Inference and Language Model Fusion of Recurrent Neural Network Transducers via End-to-End 4-bit Quantization16 Jun 2022 0 repositories listed
-
Atrial Fibrillation Detection Using Weight-Pruned, Log-Quantised Convolutional Neural Networks14 Jun 2022 0 repositories listed
-
A Theoretical Understanding of Neural Network Compression from Sparse Linear Approximation11 Jun 2022 0 repositories listed
-
HideNseek: Federated Lottery Ticket via Server-side Pruning and Sign Supermask9 Jun 2022 0 repositories listed
-
Differentially Private Model Compression3 Jun 2022 0 repositories listed