Browse State-of-the-Art › Model Compression › Papers, page 5
Model Compression
Papers archive 2025-07-28
archive papers tagged: 1,356 · with a code link: 440 · where Syntology ran a sample: 119 (97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,356 tagged: 97 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument)
Page 5 of 14: papers 401 to 500 of 1,356, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
1 Oct 2019 1 repository listed
-
23 Sep 2019 1 repository listed Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
4 Sep 2019 1 repository listed
-
3 Sep 2019 1 repository listed
-
13 Aug 2019 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
23 Jul 2019 1 repository listed
-
25 Jun 2019 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
1 Jun 2019 1 repository listed
-
24 May 2019 1 repository listed
-
20 May 2019 1 repository listed
-
11 May 2019 1 repository listed
-
28 Apr 2019 1 repository listed Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
13 Apr 2019 1 repository listed
-
3 Apr 2019 1 repository listed
-
29 Mar 2019 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
7 Mar 2019 1 repository listed
-
30 Jan 2019 1 repository listed
-
27 Jan 2019 1 repository listed
-
2 Jan 2019 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
31 Dec 2018 1 repository listed
-
11 Dec 2018 1 repository listed
-
5 Dec 2018 1 repository listed
-
30 Oct 2018 1 repository listed
-
28 Oct 2018 1 repository listed
-
27 Oct 2018 1 repository listed
-
20 Oct 2018 1 repository listed
-
1 Sep 2018 1 repository listed
-
1 Aug 2018 1 repository listed
-
1 Aug 2018 1 repository listed
-
29 Jul 2018 1 repository listed
-
1 Jun 2018 1 repository listed
-
11 Apr 2018 1 repository listed
-
29 Dec 2017 1 repository listed
-
11 Dec 2017 1 repository listed
-
13 Jul 2017 1 repository listed
-
5 Jul 2017 1 repository listed
-
5 May 2017 1 repository listed
-
6 Nov 2016 1 repository listed
-
8 Oct 2015 1 repository listed
-
12 Jul 2015 1 repository listed
-
LINR-PCGC: Lossless Implicit Neural Representations for Point Cloud Geometry Compression21 Jul 2025 0 repositories listed
-
DipSVD: Dual-importance Protected SVD for Efficient LLM Compression25 Jun 2025 0 repositories listed
-
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models21 Jun 2025 0 repositories listed
-
Model compression using knowledge distillation with integrated gradients17 Jun 2025 0 repositories listed
-
EAQuant: Enhancing Post-Training Quantization for MoE Models via Expert-Aware Optimization16 Jun 2025 0 repositories listed
-
Post-Training Quantization for Video Matting12 Jun 2025 0 repositories listed
-
AWP: Activation-Aware Weight Pruning and Quantization with Projected Gradient Descent11 Jun 2025 0 repositories listed
-
INSIGHT: A Survey of In-Network Systems for Intelligent, High-Efficiency AI and Topology Optimization30 May 2025 0 repositories listed
-
Smooth Model Compression without Fine-Tuning30 May 2025 0 repositories listed
-
Effective and Efficient One-pass Compression of Speech Foundation Models Using Sparsity-aware Self-pinching Gates28 May 2025 0 repositories listed
-
Pangu Light: Weight Re-Initialization for Pruning and Accelerating LLMs26 May 2025 0 repositories listed
-
ResSVD: Residual Compensated SVD for Large Language Model Compression26 May 2025 0 repositories listed
-
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation26 May 2025 0 repositories listed
-
Tensorization is a powerful but underexplored tool for compression and interpretability of neural networks26 May 2025 0 repositories listed
-
Efficient and Workload-Aware LLM Serving via Runtime Layer Swapping and KV Cache Resizing24 May 2025 0 repositories listed
-
Making deep neural networks work for medical audio: representation, compression and domain adaptation24 May 2025 0 repositories listed
-
LatentLLM: Attention-Aware Joint Tensor Compression23 May 2025 0 repositories listed
-
Edge-First Language Model Inference: Models, Metrics, and Tradeoffs22 May 2025 0 repositories listed
-
Is Quantum Optimization Ready? An Effort Towards Neural Network Compression using Adiabatic Quantum Computing22 May 2025 0 repositories listed
-
On Multilingual Encoder Language Model Compression for Low-Resource Languages22 May 2025 0 repositories listed
-
Saten: Sparse Augmented Tensor Networks for Post-Training Compression of Large Language Models20 May 2025 0 repositories listed
-
Low-Complexity Inference in Continual Learning via Compressed Knowledge Transfer13 May 2025 0 repositories listed
-
KDH-MLTC: Knowledge Distillation for Healthcare Multi-Label Text Classification12 May 2025 0 repositories listed
-
Semantic Retention and Extreme Compression in LLMs: Can We Have Both?12 May 2025 0 repositories listed
-
Sponge Attacks on Sensing AI: Energy-Latency Vulnerabilities and Defense via Model Pruning9 May 2025 0 repositories listed
-
Edge-Optimized Deep Learning & Pattern Recognition Techniques for Non-Intrusive Load Monitoring of Energy Time Series7 May 2025 0 repositories listed
-
Onboard Optimization and Learning: A Survey7 May 2025 0 repositories listed
-
Optimizing LLMs for Resource-Constrained Environments: A Survey of Model Compression Techniques5 May 2025 0 repositories listed
-
Smart Environmental Monitoring of Marine Pollution using Edge AI30 Apr 2025 0 repositories listed
-
Low-Rank Matrix Approximation for Neural Network Compression25 Apr 2025 0 repositories listed
-
Aerial Image Classification in Scarce and Unconstrained Environments via Conformal Prediction24 Apr 2025 0 repositories listed
-
On-Device Qwen2.5: Efficient LLM Inference with Model Compression and Hardware Acceleration24 Apr 2025 0 repositories listed
-
From Large to Super-Tiny: End-to-End Optimization for Cost-Efficient LLMs18 Apr 2025 0 repositories listed
-
D²MoE: Dual Routing and Dynamic Scheduling for Efficient On-Device MoE-based LLM Serving17 Apr 2025 0 repositories listed
-
Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning15 Apr 2025 0 repositories listed
-
APSQ: Additive Partial Sum Quantization with Algorithm-Hardware Co-Design10 Apr 2025 0 repositories listed
-
Two is Better than One: Efficient Ensemble Defense for Robust and Compact Models7 Apr 2025 0 repositories listed
-
Compression Laws for Large Language Models6 Apr 2025 0 repositories listed
-
RingMoE: Mixture-of-Modality-Experts Multi-Modal Foundation Models for Universal Remote Sensing Image Interpretation4 Apr 2025 0 repositories listed
-
Compositionality Unlocks Deep Interpretable Models3 Apr 2025 0 repositories listed
-
Random Conditioning with Distillation for Data-Efficient Diffusion Model Compression2 Apr 2025 0 repositories listed
-
Multi-Task Semantic Communications via Large Models28 Mar 2025 0 repositories listed
-
Penrose Tiled Low-Rank Compression and Section-Wise Q&A Fine-Tuning: A General Framework for Domain-Specific Large Language Model Adaptation28 Mar 2025 0 repositories listed
-
A Low-Power Streaming Speech Enhancement Accelerator For Edge Devices27 Mar 2025 0 repositories listed
-
Delving Deep into Semantic Relation Distillation27 Mar 2025 0 repositories listed
-
MoQa: Rethinking MoE Quantization with Multi-stage Data-model Distribution Awareness27 Mar 2025 0 repositories listed
-
Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration27 Mar 2025 0 repositories listed
-
Large Language Model Compression via the Nested Activation-Aware Decomposition21 Mar 2025 0 repositories listed
-
Temporal Action Detection Model Compression by Progressive Block Drop21 Mar 2025 0 repositories listed
-
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer20 Mar 2025 0 repositories listed
-
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning17 Mar 2025 0 repositories listed
-
CompMarkGS: Robust Watermarking for Compressed 3D Gaussian Splatting17 Mar 2025 0 repositories listed
-
Fragile Mastery: Are Domain-Specific Trade-Offs Undermining On-Device Language Models?16 Mar 2025 0 repositories listed
-
Sometimes Painful but Certainly Promising: Feasibility and Trade-offs of Language Model Inference at the Edge12 Mar 2025 0 repositories listed
-
Position-Aware Depth Decay Decoding (D³): Boosting Large Language Model Inference Efficiency11 Mar 2025 0 repositories listed
-
Are We There Yet? A Measurement Study of Efficiency for LLM Applications on Mobile Devices10 Mar 2025 0 repositories listed
-
Towards Superior Quantization Accuracy: A Layer-sensitive Approach9 Mar 2025 0 repositories listed
-
ACAM-KD: Adaptive and Cooperative Attention Masking for Knowledge Distillation8 Mar 2025 0 repositories listed
-
Empowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models8 Mar 2025 0 repositories listed
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.