Browse State-of-the-Art › Model Compression › Papers, page 8
Model Compression
Papers archive 2025-07-28
archive papers tagged: 1,356 · with a code link: 440 · where Syntology ran a sample: 119 (97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,356 tagged: 97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument)
Page 8 of 14: papers 701 to 800 of 1,356, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Memory-Efficient Vision Transformers: An Activation-Aware Mixed-Rank Compression Strategy8 Feb 2024 0 repositories listed
-
L4Q: Parameter Efficient Quantization-Aware Fine-Tuning on Large Language Models7 Feb 2024 0 repositories listed
-
Expediting In-Network Federated Learning by Voting-Based Consensus Model Compression6 Feb 2024 0 repositories listed
-
Fed-CVLC: Compressing Federated Learning Communications with Variable-Length Codes6 Feb 2024 0 repositories listed
-
The Potential of AutoML for Recommender Systems6 Feb 2024 0 repositories listed
-
A Survey on Transformer Compression5 Feb 2024 0 repositories listed
-
Dynamic Sparse Learning: A Novel Paradigm for Efficient Recommendation5 Feb 2024 0 repositories listed
-
Mobile Fitting Room: On-device Virtual Try-on via Diffusion Models2 Feb 2024 0 repositories listed
-
Effective Multi-Stage Training Model For Edge Computing Devices In Intrusion Detection31 Jan 2024 0 repositories listed
-
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression31 Jan 2024 0 repositories listed
-
RADIN: Souping on a Budget31 Jan 2024 0 repositories listed
-
Diffusion Model Compression for Image-to-Image Translation31 Jan 2024 0 repositories listed
-
SwapNet: Efficient Swapping for DNN Inference on Edge AI Devices Beyond the Memory Budget30 Jan 2024 0 repositories listed
-
CompactifAI: Extreme Compression of Large Language Models using Quantum-Inspired Tensor Networks25 Jan 2024 0 repositories listed
-
Large receptive field strategy and important feature extraction strategy in 3D object detection22 Jan 2024 0 repositories listed
-
ELRT: Efficient Low-Rank Training for Compact Convolutional Neural Networks18 Jan 2024 0 repositories listed
-
Convolutional Neural Network Compression via Dynamic Parameter Rank Pruning15 Jan 2024 0 repositories listed
-
FFSplit: Split Feed-Forward Network For Optimizing Accuracy-Efficiency Trade-off in Language Model Inference8 Jan 2024 0 repositories listed
-
Understanding LLMs: A Comprehensive Overview from Training to Inference4 Jan 2024 0 repositories listed
-
Data-Free Quantization via Pseudo-label Filtering1 Jan 2024 0 repositories listed
-
Unleashing Channel Potential: Space-Frequency Selection Convolution for SAR Object Detection1 Jan 2024 0 repositories listed
-
Explainability-Driven Leaf Disease Classification Using Adversarial Training and Knowledge Distillation30 Dec 2023 0 repositories listed
-
DMT: Comprehensive Distillation with Multiple Self-supervised Teachers19 Dec 2023 0 repositories listed
-
Integrating Fairness and Model Pruning Through Bi-level Optimization15 Dec 2023 0 repositories listed
-
RankDVQA-mini: Knowledge Distillation-Driven Deep Video Quality Assessment14 Dec 2023 0 repositories listed
-
Unraveling Key Factors of Knowledge Distillation14 Dec 2023 0 repositories listed
-
USM-Lite: Quantization and Sparsity Aware Fine-tuning for Speech Recognition with Universal Speech Models13 Dec 2023 0 repositories listed
-
Large Multimodal Model Compression via Efficient Pruning and Distillation at AntGroup10 Dec 2023 0 repositories listed
-
Neural Architecture Codesign for Fast Bragg Peak Analysis10 Dec 2023 0 repositories listed
-
The Efficiency Spectrum of Large Language Models: An Algorithmic Survey1 Dec 2023 0 repositories listed
-
LayerCollapse: Adaptive compression of neural networks29 Nov 2023 0 repositories listed
-
Relationship between Model Compression and Adversarial Robustness: A Review of Current Evidence27 Nov 2023 0 repositories listed
-
Cosine Similarity Knowledge Distillation for Individual Class Information Transfer24 Nov 2023 0 repositories listed
-
Education distillation:getting student models to learn in shcools23 Nov 2023 0 repositories listed
-
Knowledge Distillation Based Semantic Communications For Multiple Users23 Nov 2023 0 repositories listed
-
Efficient Transformer Knowledge Distillation: A Performance Review22 Nov 2023 0 repositories listed
-
Towards Better Parameter-Efficient Fine-Tuning for Large Language Models: A Position Paper22 Nov 2023 0 repositories listed
-
Shedding the Bits: Pushing the Boundaries of Quantization with Minifloats on FPGAs21 Nov 2023 0 repositories listed
-
Efficient Neural Networks for Tiny Machine Learning: A Comprehensive Review20 Nov 2023 0 repositories listed
-
A Speed Odyssey for Deployable Quantization of LLMs16 Nov 2023 0 repositories listed
-
On the Impact of Calibration Data in Post-training Quantization and Pruning16 Nov 2023 0 repositories listed
-
FedCode: Communication-Efficient Federated Learning via Transferring Codebooks15 Nov 2023 0 repositories listed
-
EPIM: Efficient Processing-In-Memory Accelerators based on Epitome12 Nov 2023 0 repositories listed
-
Supervised domain adaptation for building extraction from off-nadir aerial images7 Nov 2023 0 repositories listed
-
What is Lost in Knowledge Distillation?7 Nov 2023 0 repositories listed
-
6 Nov 2023 0 repositories listed
-
Data-Free Distillation of Language Model by Text-to-Text Transfer3 Nov 2023 0 repositories listed
-
Divergent Token Metrics: Measuring degradation to prune away LLM components -- and optimize quantization2 Nov 2023 0 repositories listed
-
Retrieval-based Knowledge Transfer: An Effective Approach for Extreme Large Language Model Compression24 Oct 2023 0 repositories listed
-
Data-Free Knowledge Distillation Using Adversarially Perturbed OpenGL Shader Images20 Oct 2023 0 repositories listed
-
In defense of parameter sharing for model-compression17 Oct 2023 0 repositories listed
-
USDC: Unified Static and Dynamic Compression for Visual Transformer17 Oct 2023 0 repositories listed
-
Efficient Apple Maturity and Damage Assessment: A Lightweight Detection Model with GAN and Attention Mechanism13 Oct 2023 0 repositories listed
-
What do larger image classifiers memorise?9 Oct 2023 0 repositories listed
-
Accelerating Machine Learning Primitives on Commodity Hardware8 Oct 2023 0 repositories listed
-
Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experiences6 Oct 2023 0 repositories listed
-
Robustness-Guided Image Synthesis for Data-Free Quantization5 Oct 2023 0 repositories listed
-
Sparse Deep Learning for Time Series Data: Theory and Applications5 Oct 2023 0 repositories listed
-
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models4 Oct 2023 0 repositories listed
-
Sweeping Heterogeneity with Smart MoPs: Mixture of Prompts for LLM Task Adaptation4 Oct 2023 0 repositories listed
-
Artemis: HE-Aware Training for Efficient Privacy-Preserving Machine Learning2 Oct 2023 0 repositories listed
-
Bridging the Gap Between Foundation Models and Heterogeneous Federated Learning30 Sep 2023 0 repositories listed
-
Distilling Inductive Bias: Knowledge Distillation Beyond Model Compression30 Sep 2023 0 repositories listed
-
CAIT: Triple-Win Compression towards High Accuracy, Fast Inference, and Favorable Transferability For ViTs27 Sep 2023 0 repositories listed
-
On the Impact of Quantization and Pruning of Self-Supervised Speech Models for Downstream Speech Recognition Tasks "In-the-Wild''25 Sep 2023 0 repositories listed
-
VIC-KD: Variance-Invariance-Covariance Knowledge Distillation to Make Keyword Spotting More Robust Against Adversarial Attacks22 Sep 2023 0 repositories listed
-
Pruning Large Language Models via Accuracy Predictor18 Sep 2023 0 repositories listed
-
Two-Step Knowledge Distillation for Tiny Speech Enhancement15 Sep 2023 0 repositories listed
-
CoLLD: Contrastive Layer-to-layer Distillation for Compressing Multilingual Pre-trained Speech Encoders14 Sep 2023 0 repositories listed
-
Training Acceleration of Low-Rank Decomposed Networks using Sequential Freezing and Rank Quantization7 Sep 2023 0 repositories listed
-
Norm Tweaking: High-performance Low-bit Quantization of Large Language Models6 Sep 2023 0 repositories listed
-
ADC/DAC-Free Analog Acceleration of Deep Neural Networks with Frequency Transformation4 Sep 2023 0 repositories listed
-
Computation-efficient Deep Learning for Computer Vision: A Survey27 Aug 2023 0 repositories listed
-
Improving Knowledge Distillation for BERT Models: Loss Functions, Mapping Methods, and Weight Tuning26 Aug 2023 0 repositories listed
-
DLIP: Distilling Language-Image Pre-training24 Aug 2023 0 repositories listed
-
QD-BEV : Quantization-aware View-guided Distillation for Multi-view 3D Object Detection21 Aug 2023 0 repositories listed
-
Learning Disentangled Representation with Mutual Information Maximization for Real-Time UAV Tracking20 Aug 2023 0 repositories listed
-
SHARK: A Lightweight Model Compression Approach for Large-scale Recommender Systems18 Aug 2023 0 repositories listed
-
Spike-and-slab shrinkage priors for structurally sparse Bayesian neural networks17 Aug 2023 0 repositories listed
-
Benchmarking Adversarial Robustness of Compressed Deep Learning Models16 Aug 2023 0 repositories listed
-
A Survey on Model Compression for Large Language Models15 Aug 2023 0 repositories listed
-
Shortcut-V2V: Compression Framework for Video-to-Video Translation based on Temporal Redundancy Reduction15 Aug 2023 0 repositories listed
-
FedEdge AI-TC: A Semi-supervised Traffic Classification Method based on Trusted Federated Deep Learning for Mobile Edge Computing14 Aug 2023 0 repositories listed
-
Accurate Neural Network Pruning Requires Rethinking Sparse Optimization3 Aug 2023 0 repositories listed
-
MIMONet: Multi-Input Multi-Output On-Device Deep Learning22 Jul 2023 0 repositories listed
-
Model Compression Methods for YOLOv5: A Review21 Jul 2023 0 repositories listed
-
Impact of Disentanglement on Pruning Neural Networks19 Jul 2023 0 repositories listed
-
Knowledge Distillation for Object Detection: from generic to remote sensing datasets18 Jul 2023 0 repositories listed
-
Data-Free Quantization via Mixed-Precision Compensation without Fine-Tuning2 Jul 2023 0 repositories listed
-
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition2 Jul 2023 0 repositories listed
-
Feature Adversarial Distillation for Point Cloud Classification25 Jun 2023 0 repositories listed
-
Low-Rank Prune-And-Factorize for Language Model Compression25 Jun 2023 0 repositories listed
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression24 Jun 2023 0 repositories listed
-
DynaQuant: Compressing Deep Learning Training Checkpoints via Dynamic Quantization20 Jun 2023 0 repositories listed
-
LoSparse: Structured Compression of Large Language Models based on Low-Rank and Sparse Approximation20 Jun 2023 0 repositories listed
-
Neural Network Compression using Binarization and Few Full-Precision Weights15 Jun 2023 0 repositories listed
-
Modular Transformers: Compressing Transformers into Modularized Layers for Flexible Efficient Inference4 Jun 2023 0 repositories listed
-
Riemannian Low-Rank Model Compression for Federated Learning with Over-the-Air Aggregation4 Jun 2023 0 repositories listed
-
Low-Complexity Acoustic Scene Classification Using Data Augmentation and Lightweight ResNet3 Jun 2023 0 repositories listed
-
Task-Agnostic Structured Pruning of Speech Representation Models2 Jun 2023 0 repositories listed