Methods › General › Model Compression › Pruning › Papers, page 5
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 5 of 39: papers 401 to 500 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
KVTuner: Sensitivity-Aware Layer-wise Mixed Precision KV Cache Quantization for Efficient and Nearly Lossless LLM Inference 6 Feb 2025 · 1 repository · arXiv:2502.04420
-
MXMap: A Multivariate Cross Mapping Framework for Causal Discovery in Dynamical Systems 6 Feb 2025 · 1 repository · arXiv:2502.03802
-
UniCP: A Unified Caching and Pruning Framework for Efficient Video Generation 6 Feb 2025 · 0 repositories · arXiv:2502.04393
-
Adapt-Pruner: Adaptive Structural Pruning for Efficient Small Language Model Training 5 Feb 2025 · 0 repositories · arXiv:2502.03460
-
Deep Weight Factorization: Sparse Learning Through the Lens of Artificial Symmetries 4 Feb 2025 · 0 repositories · arXiv:2502.02496
-
GP-GS: Gaussian Processes for Enhanced Gaussian Splatting 4 Feb 2025 · 1 repository · arXiv:2502.02283
-
Prompt-based Depth Pruning of Large Language Models 4 Feb 2025 · 0 repositories · arXiv:2502.04348
-
Pruning-aware Loss Functions for STOI-Optimized Pruned Recurrent Autoencoders for the Compression of the Stimulation Patterns of Cochlear Implants at Zero Delay 4 Feb 2025 · 0 repositories · arXiv:2502.02424
-
Reachability-Based Contingency Planning against Multi-Modal Predictions with Branch MPC 4 Feb 2025 · 0 repositories · arXiv:2502.02550
-
Choose Your Model Size: Any Compression by a Single Gradient Descent 3 Feb 2025 · 0 repositories · arXiv:2502.01717
-
Progressive Binarization with Semi-Structured Pruning for LLMs 3 Feb 2025 · 1 repository · arXiv:2502.01705Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multi-frequency wavefield solutions for variable velocity models using meta-learning enhanced low-rank physics-informed neural network 2 Feb 2025 · 0 repositories · arXiv:2502.00897
-
Structural Latency Perturbation in Large Language Models Through Recursive State Induction 2 Feb 2025 · 0 repositories · arXiv:2502.00758
-
Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning 1 Feb 2025 · 0 repositories · arXiv:2502.00511
-
CAT Pruning: Cluster-Aware Token Pruning For Text-to-Image Diffusion Models 1 Feb 2025 · 1 repository · arXiv:2502.00433
-
Cache Me If You Must: Adaptive Key-Value Quantization for Large Language Models 31 Jan 2025 · 1 repository · arXiv:2501.19392
-
FedRTS: Federated Robust Pruning via Combinatorial Thompson Sampling 31 Jan 2025 · 0 repositories · arXiv:2501.19122Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models 31 Jan 2025 · 0 repositories · arXiv:2501.19090
-
Symmetric Pruning of Large Language Models 31 Jan 2025 · 0 repositories · arXiv:2501.18980
-
SAFL: Structure-Aware Personalized Federated Learning via Client-Specific Clustering and SCSI-Guided Model Pruning 30 Jan 2025 · 0 repositories · arXiv:2501.18659
-
2SSP: A Two-Stage Framework for Structured Pruning of LLMs 29 Jan 2025 · 1 repository · arXiv:2501.17771
-
A Proximal Operator for Inducing 2:4-Sparsity 29 Jan 2025 · 0 repositories · arXiv:2501.18015
-
ASAP: Learning Generalizable Online Bin Packing via Adaptive Selection After Pruning 29 Jan 2025 · 0 repositories · arXiv:2501.17377
-
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models 29 Jan 2025 · 0 repositories · arXiv:2501.17905
-
Explainable Machine Learning: An Illustration of Kolmogorov-Arnold Network Model for Airfoil Lift Prediction 29 Jan 2025 · 0 repositories · arXiv:2501.17896
-
Hybrid Graphs for Table-and-Text based Question Answering using LLMs 29 Jan 2025 · 0 repositories · arXiv:2501.17767
-
When less is more: evolving large neural networks from small ones 29 Jan 2025 · 1 repository · arXiv:2501.18012
-
B-FPGM: Lightweight Face Detection via Bayesian-Optimized Soft FPGM Pruning 28 Jan 2025 · 0 repositories · arXiv:2501.16917
-
Graph of Attacks with Pruning: Optimizing Stealthy Jailbreak Prompt Generation for Enhanced LLM Content Moderation 28 Jan 2025 · 1 repository · arXiv:2501.18638
-
Efficient Object Detection of Marine Debris using Pruned YOLO Model 27 Jan 2025 · 0 repositories · arXiv:2501.16571
-
Provence: efficient and robust context pruning for retrieval-augmented generation 27 Jan 2025 · 0 repositories · arXiv:2501.16214
-
Information Consistent Pruning: How to Efficiently Search for Sparse Networks? 26 Jan 2025 · 1 repository · arXiv:2501.15592
-
Hardware-Aware DNN Compression for Homogeneous Edge Devices 25 Jan 2025 · 0 repositories · arXiv:2501.15240
-
Lightweight and Post-Training Structured Pruning for On-Device Large Lanaguage Models 25 Jan 2025 · 0 repositories · arXiv:2501.15255
-
On Accelerating Edge AI: Optimizing Resource-Constrained Environments 25 Jan 2025 · 0 repositories · arXiv:2501.15014
-
PIP: Perturbation-based Iterative Pruning for Large Language Models 25 Jan 2025 · 0 repositories · arXiv:2501.15278
-
ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruning 25 Jan 2025 · 0 repositories · arXiv:2501.15316
-
You Only Prune Once: Designing Calibration-Free Model Compression With Policy Learning 25 Jan 2025 · 0 repositories · arXiv:2501.15296
-
Adaptive Rank Allocation for Federated Parameter-Efficient Fine-Tuning of Language Models 24 Jan 2025 · 0 repositories · arXiv:2501.14406
-
Dynamic Token Reduction during Generation for Vision Language Models 24 Jan 2025 · 0 repositories · arXiv:2501.14204
-
Fast Think-on-Graph: Wider, Deeper and Faster Reasoning of Large Language Model on Knowledge Graph 24 Jan 2025 · 1 repository · arXiv:2501.14300
-
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models 24 Jan 2025 · 0 repositories · arXiv:2501.16376
-
Neural-Symbolic Message Passing with Dynamic Pruning 24 Jan 2025 · 0 repositories · arXiv:2501.14661
-
ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations 24 Jan 2025 · 0 repositories · arXiv:2501.14607
-
GoDe: Gaussians on Demand for Progressive Level of Detail and Scalable Compression 23 Jan 2025 · 0 repositories · arXiv:2501.13558
-
LVPruning: An Effective yet Simple Language-Guided Vision Token Pruning Approach for Multi-modal Large Language Models 23 Jan 2025 · 0 repositories · arXiv:2501.13652
-
One-cycle Structured Pruning with Stability Driven Structure Search 23 Jan 2025 · 0 repositories · arXiv:2501.13439
-
Overcoming Support Dilution for Robust Few-shot Semantic Segmentation 23 Jan 2025 · 0 repositories · arXiv:2501.13529
-
Advanced deep architecture pruning using single filter performance 22 Jan 2025 · 1 repository · arXiv:2501.12880
-
BLR-MoE: Boosted Language-Routing Mixture of Experts for Domain-Robust Multilingual E2E ASR 22 Jan 2025 · 0 repositories · arXiv:2501.12602
-
Unified CNNs and transformers underlying learning mechanism reveals multi-head attention modus vivendi 22 Jan 2025 · 0 repositories · arXiv:2501.12900
-
The Journey Matters: Average Parameter Count over Pre-training Unifies Sparse and Dense Scaling Laws 21 Jan 2025 · 0 repositories · arXiv:2501.12486
-
Communication-Efficient Federated Learning Based on Explanation-Guided Pruning for Remote Sensing Image Classification 20 Jan 2025 · 1 repository · arXiv:2501.11493
-
Meta-Instance Selection. Instance Selection as a Classification Problem with Meta-Features 20 Jan 2025 · 0 repositories · arXiv:2501.11526
-
Accelerating Large Language Models through Partially Linear Feed-Forward Network 17 Jan 2025 · 0 repositories · arXiv:2501.10054
-
MultiPruner: Balanced Structure Removal in Foundation Models 17 Jan 2025 · 1 repository · arXiv:2501.09949
-
FASP: Fast and Accurate Structured Pruning of Large Language Models 16 Jan 2025 · 0 repositories · arXiv:2501.09412
-
Pruning for Sparse Diffusion Models based on Gradient Flow 16 Jan 2025 · 0 repositories · arXiv:2501.09464
-
SuperSAM: Crafting a SAM Supernetwork via Structured Pruning and Unstructured Parameter Prioritization 15 Jan 2025 · 1 repository · arXiv:2501.08504
-
Deep Learning and Natural Language Processing in the Field of Construction 14 Jan 2025 · 0 repositories · arXiv:2501.07911
-
Object-Centric 2D Gaussian Splatting: Background Removal and Occlusion-Aware Pruning for Compact Object Models 14 Jan 2025 · 0 repositories · arXiv:2501.08174
-
Optimal Classification Trees for Continuous Feature Data Using Dynamic Programming with Branch-and-Bound 14 Jan 2025 · 2 repositories · arXiv:2501.07903
-
PolyLUT: Ultra-low Latency Polynomial Inference with Hardware-Aware Structured Pruning 14 Jan 2025 · 0 repositories · arXiv:2501.08043
-
FlexQuant: Elastic Quantization Framework for Locally Hosted LLM on Edge Devices 13 Jan 2025 · 0 repositories · arXiv:2501.07139
-
Compact Bayesian Neural Networks via pruned MCMC sampling 12 Jan 2025 · 1 repository · arXiv:2501.06962
-
Merging Feed-Forward Sublayers for Compressed Transformers 10 Jan 2025 · 1 repository · arXiv:2501.06126
-
A 1Mb mixed-precision quantized encoder for image classification and patch-based compression 9 Jan 2025 · 0 repositories · arXiv:2501.05097
-
Deriving Coding-Specific Sub-Models from LLMs using Resource-Efficient Pruning 9 Jan 2025 · 0 repositories · arXiv:2501.05248
-
UPAQ: A Framework for Real-Time and Energy-Efficient 3D Object Detection in Autonomous Vehicles 8 Jan 2025 · 0 repositories · arXiv:2501.04213
-
Adaptive Pruning of Pretrained Transformer via Differential Inclusions 6 Jan 2025 · 0 repositories · arXiv:2501.03289
-
LightGNN: Simple Graph Neural Network for Recommendation 6 Jan 2025 · 1 repository · arXiv:2501.03228
-
Efficient Deployment of Large Language Models on Resource-constrained Devices 5 Jan 2025 · 0 repositories · arXiv:2501.02438
-
Prune or Retrain: Optimizing the Vocabulary of Multilingual Models for Estonian 5 Jan 2025 · 0 repositories · arXiv:2501.02631
-
Strategic Fusion Optimizes Transformer Compression 5 Jan 2025 · 0 repositories · arXiv:2501.03273
-
Swift Cross-Dataset Pruning: Enhancing Fine-Tuning Efficiency in Natural Language Understanding 5 Jan 2025 · 1 repository · arXiv:2501.02432
-
Boosting Explainability through Selective Rationalization in Pre-trained Language Models 3 Jan 2025 · 1 repository · arXiv:2501.03182
-
Instruction-Following Pruning for Large Language Models 3 Jan 2025 · 0 repositories · arXiv:2501.02086
-
Pruning-based Data Selection and Network Fusion for Efficient Deep Learning 2 Jan 2025 · 0 repositories · arXiv:2501.01118
-
Annotation Ambiguity Aware Semi-Supervised Medical Image Segmentation 1 Jan 2025 · 0 repositories
-
BG-Triangle: Bezier Gaussian Triangle for 3D Vectorization and Rendering 1 Jan 2025 · 0 repositories
-
D2SP: Dynamic Dual-Stage Purification Framework for Dual Noise Mitigation in Vision-based Affective Recognition. 1 Jan 2025 · 0 repositories
-
Efficient Test-time Adaptive Object Detection via Sensitivity-Guided Pruning 1 Jan 2025 · 0 repositories
-
EfficientLLaVA: Generalizable Auto-Pruning for Large Vision-language Models 1 Jan 2025 · 0 repositories
-
FedCS: Coreset Selection for Federated Learning 1 Jan 2025 · 0 repositories
-
FirePlace: Geometric Refinements of LLM Common Sense Reasoning for 3D Object Placement 1 Jan 2025 · 0 repositories
-
FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting 1 Jan 2025 · 0 repositories
-
Flexible Group Count Enables Hassle-Free Structured Pruning 1 Jan 2025 · 0 repositories
-
ICP: Immediate Compensation Pruning for Mid-to-high Sparsity 1 Jan 2025 · 0 repositories
-
Libra-Merging: Importance-redundancy and Pruning-merging Trade-off for Acceleration Plug-in in Large Vision-Language Model 1 Jan 2025 · 0 repositories
-
On Importance of Layer Pruning for Smaller BERT Models and Low Resource Languages 1 Jan 2025 · 0 repositories · arXiv:2501.00733
-
Fast and Interpretable Mixed-Integer Linear Program Solving by Learning Model Reduction 31 Dec 2024 · 0 repositories · arXiv:2501.00307
-
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free 31 Dec 2024 · 1 repository · arXiv:2501.00375
-
Align Attention Heads Before Merging Them: An Effective Way for Converting MHA to GQA 30 Dec 2024 · 0 repositories · arXiv:2412.20677
-
EdgeRAG: Online-Indexed RAG for Edge Devices 30 Dec 2024 · 0 repositories · arXiv:2412.21023
-
Exploring and Controlling Diversity in LLM-Agent Conversation 30 Dec 2024 · 0 repositories · arXiv:2412.21102
-
FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Visual Language Models 30 Dec 2024 · 1 repository · arXiv:2501.01986Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
MaskGaussian: Adaptive 3D Gaussian Representation from Probabilistic Masks 29 Dec 2024 · 1 repository · arXiv:2412.20522
-
ReTaKe: Reducing Temporal and Knowledge Redundancy for Long Video Understanding 29 Dec 2024 · 1 repository · arXiv:2412.20504Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
Mining Platoon Patterns from Traffic Videos 28 Dec 2024 · 1 repository · arXiv:2412.20177
-
ST³: Accelerating Multimodal Large Language Model by Spatial-Temporal Visual Token Trimming 28 Dec 2024 · 0 repositories · arXiv:2412.20105