Browse State-of-the-Art › Model Compression › Papers, page 6
Model Compression
Papers archive 2025-07-28
archive papers tagged: 1,356 · with a code link: 440 · where Syntology ran a sample: 119 (97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,356 tagged: 97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument)
Page 6 of 14: papers 501 to 600 of 1,356, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
IteRABRe: Iterative Recovery-Aided Block Reduction8 Mar 2025 0 repositories listed
-
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression6 Mar 2025 0 repositories listed
-
TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation6 Mar 2025 0 repositories listed
-
10K is Enough: An Ultra-Lightweight Binarized Network for Infrared Small-Target Detection4 Mar 2025 0 repositories listed
-
Beyond the Tip of Efficiency: Uncovering the Submerged Threats of Jailbreak Attacks in Small Language Models27 Feb 2025 0 repositories listed
-
Vision Transformers on the Edge: A Comprehensive Survey of Model Compression and Acceleration Strategies26 Feb 2025 0 repositories listed
-
AfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages25 Feb 2025 0 repositories listed
-
The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve?24 Feb 2025 0 repositories listed
-
Swallowing the Poison Pills: Insights from Vulnerability Disparity Among LLMs23 Feb 2025 0 repositories listed
-
When Compression Meets Model Compression: Memory-Efficient Double Compression for Large Language Models21 Feb 2025 0 repositories listed
-
Optimizing Singular Spectrum for Large Language Model Compression20 Feb 2025 0 repositories listed
-
Vision Foundation Models in Medical Image Analysis: Advances and Challenges20 Feb 2025 0 repositories listed
-
MaskPrune: Mask-based LLM Pruning for Layer-wise Uniform Structures19 Feb 2025 0 repositories listed
-
Every Expert Matters: Towards Effective Knowledge Distillation for Mixture-of-Experts Language Models18 Feb 2025 0 repositories listed
-
OPTISHEAR: Towards Efficient and Adaptive Pruning of Large Language Models via Evolutionary Optimization15 Feb 2025 0 repositories listed
-
Vision-Language Models for Edge Networks: A Comprehensive Survey11 Feb 2025 0 repositories listed
-
Low-Rank Compression for IMC Arrays10 Feb 2025 0 repositories listed
-
Runtime Tunable Tsetlin Machines for Edge Inference on eFPGAs10 Feb 2025 0 repositories listed
-
Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models9 Feb 2025 0 repositories listed
-
Theoretical Guarantees for Low-Rank Compression of Deep Neural Networks4 Feb 2025 0 repositories listed
-
Accelerating Linear Recurrent Neural Networks for the Edge with Unstructured Sparsity3 Feb 2025 0 repositories listed
-
MIND: Modality-Informed Knowledge Distillation Framework for Multimodal Clinical Prediction Tasks3 Feb 2025 0 repositories listed
-
Attention Sinks and Outlier Features: A 'Catch, Tag, and Release' Mechanism for Embeddings2 Feb 2025 0 repositories listed
-
Huff-LLM: End-to-End Lossless Compression for Efficient LLM Inference2 Feb 2025 0 repositories listed
-
Role of Mixup in Topological Persistence Based Knowledge Distillation for Wearable Sensor Data2 Feb 2025 0 repositories listed
-
Efficient Supernet Training with Orthogonal Softmax for Scalable ASR Model Compression31 Jan 2025 0 repositories listed
-
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models31 Jan 2025 0 repositories listed
-
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models28 Jan 2025 0 repositories listed
-
On Accelerating Edge AI: Optimizing Resource-Constrained Environments25 Jan 2025 0 repositories listed
-
You Only Prune Once: Designing Calibration-Free Model Compression With Policy Learning25 Jan 2025 0 repositories listed
-
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models24 Jan 2025 0 repositories listed
-
Practical quantum federated learning and its experimental demonstration22 Jan 2025 0 repositories listed
-
Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures16 Jan 2025 0 repositories listed
-
FASP: Fast and Accurate Structured Pruning of Large Language Models16 Jan 2025 0 repositories listed
-
Knowledge Distillation for Image Restoration : Simultaneous Learning from Degraded and Clean Images16 Jan 2025 0 repositories listed
-
SWSC: Shared Weight for Similar Channel in LLM15 Jan 2025 0 repositories listed
-
CURing Large Models: Compression via CUR Decomposition8 Jan 2025 0 repositories listed
-
UPAQ: A Framework for Real-Time and Energy-Efficient 3D Object Detection in Autonomous Vehicles8 Jan 2025 0 repositories listed
-
Effective and Efficient Mixed Precision Quantization of Speech Foundation Models7 Jan 2025 0 repositories listed
-
Strategic Fusion Optimizes Transformer Compression5 Jan 2025 0 repositories listed
-
Optimizing Small Language Models for In-Vehicle Function-Calling4 Jan 2025 0 repositories listed
-
Once-Tuning-Multiple-Variants: Tuning Once and Expanded as Multiple Vision-Language Model Variants1 Jan 2025 0 repositories listed
-
Random Conditioning for Diffusion Model Compression with Distillation1 Jan 2025 0 repositories listed
-
Improving Acoustic Scene Classification in Low-Resource Conditions30 Dec 2024 0 repositories listed
-
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models27 Dec 2024 0 repositories listed
-
Optimization and Scalability of Collaborative Filtering Algorithms in Large Language Models25 Dec 2024 0 repositories listed
-
HTR-JAND: Handwritten Text Recognition with Joint Attention Network and Knowledge Distillation24 Dec 2024 0 repositories listed
-
CoSurfGS:Collaborative 3D Surface Gaussian Splatting with Distributed Learning for Large Scene Reconstruction23 Dec 2024 0 repositories listed
-
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings23 Dec 2024 0 repositories listed
-
GQSA: Group Quantization and Sparsity for Accelerating Large Language Model Inference23 Dec 2024 0 repositories listed
-
Lightweight Design and Optimization methods for DCNNs: Progress and Futures22 Dec 2024 0 repositories listed
-
Semantics Prompting Data-Free Quantization for Low-Bit Vision Transformers21 Dec 2024 0 repositories listed
-
Deploying Foundation Model Powered Agent Services: A Survey18 Dec 2024 0 repositories listed
-
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs15 Dec 2024 0 repositories listed
-
Activation Sparsity Opportunities for Compressing General Large Language Models13 Dec 2024 0 repositories listed
-
Can Students Beyond The Teacher? Distilling Knowledge from Teacher's Bias13 Dec 2024 0 repositories listed
-
Optimising TinyML with Quantization and Distillation of Transformer and Mamba Models for Indoor Localisation on Edge Devices12 Dec 2024 0 repositories listed
-
Low-Rank Correction for Quantized LLMs10 Dec 2024 0 repositories listed
-
Compression for Better: A General and Stable Lossless Compression Framework9 Dec 2024 0 repositories listed
-
Lossless Model Compression via Joint Low-Rank Factorization Optimization9 Dec 2024 0 repositories listed
-
VQ4ALL: Efficient Neural Network Representation via a Universal Codebook9 Dec 2024 0 repositories listed
-
Trimming Down Large Spiking Vision Transformers via Heterogeneous Quantization Search7 Dec 2024 0 repositories listed
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models3 Dec 2024 0 repositories listed
-
Efficient Model Compression Techniques with FishLeg3 Dec 2024 0 repositories listed
-
Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models27 Nov 2024 0 repositories listed
-
Efficient Pruning of Text-to-Image Models: Insights from Pruning Stable Diffusion22 Nov 2024 0 repositories listed
-
21 Nov 2024 0 repositories listed Syntology 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
FASTNav: Fine-tuned Adaptive Small-language-models Trained for Multi-point Robot Navigation20 Nov 2024 0 repositories listed
-
Puppet-CNN: Input-Adaptive Convolutional Neural Networks with Model Compression using Ordinary Differential Equation19 Nov 2024 0 repositories listed
-
What Makes a Good Dataset for Knowledge Distillation?19 Nov 2024 0 repositories listed
-
Bridging the Resource Gap: Deploying Advanced Imitation Learning Models onto Affordable Embedded Platforms18 Nov 2024 0 repositories listed
-
Re-Parameterization of Lightweight Transformer for On-Device Speech Emotion Recognition14 Nov 2024 0 repositories listed
-
ASER: Activation Smoothing and Error Reconstruction for Large Language Model Quantization12 Nov 2024 0 repositories listed
-
Feature Interaction Fusion Self-Distillation Network For CTR Prediction12 Nov 2024 0 repositories listed
-
Optimizing Traffic Signal Control using High-Dimensional State Representation and Efficient Deep Reinforcement Learning12 Nov 2024 0 repositories listed
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models6 Nov 2024 0 repositories listed
-
Efficient Model Compression for Bayesian Neural Networks1 Nov 2024 0 repositories listed
-
28 Oct 2024 0 repositories listed Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Survey of Small Language Models25 Oct 2024 0 repositories listed
-
SWITCH: Studying with Teacher for Knowledge Distillation of Large Language Models25 Oct 2024 0 repositories listed
-
Beware of Calibration Data for Pruning Large Language Models23 Oct 2024 0 repositories listed
-
Towards Effective Data-Free Knowledge Distillation via Diverse Diffusion Augmentation23 Oct 2024 0 repositories listed
-
Self-calibration for Language Model Quantization and Pruning22 Oct 2024 0 repositories listed
-
Identifying Sub-networks in Neural Networks via Functionally Similar Representations21 Oct 2024 0 repositories listed
-
Preview-based Category Contrastive Learning for Knowledge Distillation18 Oct 2024 0 repositories listed
-
CrossQuant: A Post-Training Quantization Method with Smaller Quantization Kernel for Precise Large Language Model Compression10 Oct 2024 0 repositories listed
-
What is Left After Distillation? How Knowledge Transfer Impacts Fairness and Bias10 Oct 2024 0 repositories listed
-
Large Language Model Compression with Neural Architecture Search9 Oct 2024 0 repositories listed
-
SpaLLM: Unified Compressive Adaptation of Large Language Models with Sketching8 Oct 2024 0 repositories listed
-
ESPACE: Dimensionality Reduction of Activations for Model Compression7 Oct 2024 0 repositories listed
-
Continuous Approximations for Improving Quantization Aware Training of LLMs6 Oct 2024 0 repositories listed
-
Geometry is All You Need: A Unified Taxonomy of Matrix and Tensor Factorization for Compression of Generative Language Models3 Oct 2024 0 repositories listed
-
Compressing Recurrent Neural Networks for FPGA-accelerated Implementation in Fluorescence Lifetime Imaging1 Oct 2024 0 repositories listed
-
Aggressive Post-Training Compression on Extremely Large Language Models30 Sep 2024 0 repositories listed
-
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries29 Sep 2024 0 repositories listed
-
Value-Based Deep Multi-Agent Reinforcement Learning with Dynamic Sparse Training28 Sep 2024 0 repositories listed
-
General Compression Framework for Efficient Transformer Object Tracking26 Sep 2024 0 repositories listed
-
Applications of Knowledge Distillation in Remote Sensing: A Survey18 Sep 2024 0 repositories listed
-
Privacy-Preserving SAM Quantization for Efficient Edge Intelligence in Healthcare14 Sep 2024 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.