Methods › General › Model Compression › Pruning › Papers, page 6
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 6 of 39: papers 501 to 600 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Exploring GLU Expansion Ratios: A Study of Structured Pruning in LLaMA-3.2 Models 26 Dec 2024 · 1 repository
-
Resource-Efficient Transformer Architecture: Optimizing Memory and Execution Time for Real-Time Applications 25 Dec 2024 · 0 repositories · arXiv:2501.00042
-
AutoSculpt: A Pattern-based Model Auto-pruning Framework Using Reinforcement Learning and Graph Learning 24 Dec 2024 · 0 repositories · arXiv:2412.18091
-
Pruning Unrolled Networks (PUN) at Initialization for MRI Reconstruction Improves Generalization 24 Dec 2024 · 0 repositories · arXiv:2412.18668
-
SlimGPT: Layer-wise Structured Pruning for Large Language Models 24 Dec 2024 · 0 repositories · arXiv:2412.18110
-
Unified Stochastic Framework for Neural Network Quantization and Pruning 24 Dec 2024 · 0 repositories · arXiv:2412.18184
-
GQSA: Group Quantization and Sparsity for Accelerating Large Language Model Inference 23 Dec 2024 · 0 repositories · arXiv:2412.17560
-
Singular Value Scaling: Efficient Generative Model Compression via Pruned Weights Refinement 23 Dec 2024 · 1 repository · arXiv:2412.17387
-
Scalable Speech Enhancement with Dynamic Channel Pruning 22 Dec 2024 · 0 repositories · arXiv:2412.17121
-
Lillama: Large Language Models Compression via Low-Rank Feature Distillation 21 Dec 2024 · 0 repositories · arXiv:2412.16719
-
Less is More: Towards Green Code Large Language Models via Unified Structural Pruning 20 Dec 2024 · 0 repositories · arXiv:2412.15921
-
PruneVid: Visual Token Pruning for Efficient Video Large Language Models 20 Dec 2024 · 1 repository · arXiv:2412.16117Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AdaCred: Adaptive Causal Decision Transformers with Feature Crediting 19 Dec 2024 · 0 repositories · arXiv:2412.15427
-
Adaptive Pruning for Large Language Models with Structural Importance Awareness 19 Dec 2024 · 0 repositories · arXiv:2412.15127
-
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs 19 Dec 2024 · 0 repositories · arXiv:2412.14426
-
Efficient Fine-Tuning and Concept Suppression for Pruned Diffusion Models 19 Dec 2024 · 1 repository · arXiv:2412.15341
-
Holistic Adversarially Robust Pruning 19 Dec 2024 · 0 repositories · arXiv:2412.14714
-
Robust Federated Learning in the Face of Covariate Shift: A Magnitude Pruning with Hybrid Regularization Framework for Enhanced Model Aggregation 19 Dec 2024 · 0 repositories · arXiv:2412.15010
-
YOLOv11 Optimization for Efficient Resource Utilization 19 Dec 2024 · 1 repository · arXiv:2412.14790
-
DreaMark: Rooting Watermark in Score Distillation Sampling Generated Neural Radiance Fields 18 Dec 2024 · 0 repositories · arXiv:2412.15278
-
On the Compression of Language Models for Code: An Empirical Study on CodeBERT 18 Dec 2024 · 0 repositories · arXiv:2412.13737
-
Resource Constrained Pathfinding with Enhanced Bidirectional A* Search 18 Dec 2024 · 0 repositories · arXiv:2412.13888
-
Understanding and Analyzing Model Robustness and Knowledge-Transfer in Multilingual Neural Machine Translation using TX-Ray 18 Dec 2024 · 0 repositories · arXiv:2412.13881
-
4DRGS: 4D Radiative Gaussian Splatting for Efficient 3D Vessel Reconstruction from Sparse-View Dynamic DSA Images 17 Dec 2024 · 1 repository · arXiv:2412.12919
-
A Comparative Study of Pruning Methods in Transformer-based Time Series Forecasting 17 Dec 2024 · 0 repositories · arXiv:2412.12883
-
Activating Distributed Visual Region within LLMs for Efficient and Effective Vision-Language Training and Inference 17 Dec 2024 · 0 repositories · arXiv:2412.12785
-
Faster Vision Mamba is Rebuilt in Minutes via Merged Token Re-training 17 Dec 2024 · 1 repository · arXiv:2412.12496
-
Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration 17 Dec 2024 · 0 repositories · arXiv:2412.13180
-
HyperGS: Hyperspectral 3D Gaussian Splatting 17 Dec 2024 · 0 repositories · arXiv:2412.12849
-
ITP: Instance-Aware Test Pruning for Out-of-Distribution Detection 17 Dec 2024 · 1 repository · arXiv:2412.12566
-
Learning Coarse-to-Fine Pruning of Graph Convolutional Networks for Skeleton-based Recognition 17 Dec 2024 · 0 repositories · arXiv:2412.12887
-
More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression 17 Dec 2024 · 0 repositories · arXiv:2412.12706
-
Numerical Pruning for Efficient Autoregressive Models 17 Dec 2024 · 0 repositories · arXiv:2412.12441
-
RCTrans: Radar-Camera Transformer via Radar Densifier and Sequential Decoder for 3D Object Detection 17 Dec 2024 · 1 repository · arXiv:2412.12799
-
RemoteTrimmer: Adaptive Structural Pruning for Remote Sensing Image Classification 17 Dec 2024 · 1 repository · arXiv:2412.12603
-
Structural Pruning via Spatial-aware Information Redundancy for Semantic Segmentation 17 Dec 2024 · 1 repository · arXiv:2412.12672
-
Beyond Graph Convolution: Multimodal Recommendation with Topology-aware MLPs 16 Dec 2024 · 1 repository · arXiv:2412.11747
-
Designing Semi-Structured Pruning of Graph Convolutional Networks for Skeleton-based Recognition 16 Dec 2024 · 0 repositories · arXiv:2412.11813
-
FTP: A Fine-grained Token-wise Pruner for Large Language Models via Token Routing 16 Dec 2024 · 0 repositories · arXiv:2412.11494
-
QPruner: Probabilistic Decision Quantization for Structured Pruning in Large Language Models 16 Dec 2024 · 0 repositories · arXiv:2412.11629
-
RetroLLM: Empowering Large Language Models to Retrieve Fine-grained Evidence within Generation 16 Dec 2024 · 1 repository · arXiv:2412.11919
-
Scalable Temporal Anomaly Causality Discovery in Large Systems: Achieving Computational Efficiency with Binary Anomaly Flag Data 16 Dec 2024 · 1 repository · arXiv:2412.11800
-
SpeechPrune: Context-aware Token Pruning for Speech Information Retrieval 16 Dec 2024 · 1 repository · arXiv:2412.12009
-
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs 15 Dec 2024 · 0 repositories · arXiv:2412.11242
-
TinySubNets: An efficient and low capacity continual learning strategy 14 Dec 2024 · 1 repository · arXiv:2412.10869
-
Data Pruning Can Do More: A Comprehensive Data Pruning Approach for Object Re-identification 13 Dec 2024 · 1 repository · arXiv:2412.10091
-
MeshA*: Efficient Path Planing With Motion Primitives 13 Dec 2024 · 0 repositories · arXiv:2412.10320
-
MVQ:Towards Efficient DNN Compression and Acceleration with Masked Vector Quantization 13 Dec 2024 · 0 repositories · arXiv:2412.10261
-
SplineGS: Robust Motion-Adaptive Spline for Real-Time Dynamic 3D Gaussians from Monocular Video 13 Dec 2024 · 0 repositories · arXiv:2412.09982
-
Static Pruning in Dense Retrieval using Matrix Decomposition 13 Dec 2024 · 0 repositories · arXiv:2412.09983
-
TSGaussian: Semantic and Depth-Guided Target-Specific Gaussian Splatting from Sparse Views 13 Dec 2024 · 1 repository · arXiv:2412.10051Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Fast Track to Winning Tickets: Repowering One-Shot Pruning for Graph Neural Networks 10 Dec 2024 · 1 repository · arXiv:2412.07605
-
Mobile Video Diffusion 10 Dec 2024 · 0 repositories · arXiv:2412.07583
-
Post-Training Statistical Calibration for Higher Activation Sparsity 10 Dec 2024 · 1 repository · arXiv:2412.07174Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Score-matching-based Structure Learning for Temporal Data on Networks 10 Dec 2024 · 0 repositories · arXiv:2412.07469
-
TT-MPD: Test Time Model Pruning and Distillation 10 Dec 2024 · 0 repositories · arXiv:2412.07114
-
Federated Split Learning with Model Pruning and Gradient Quantization in Wireless Networks 9 Dec 2024 · 0 repositories · arXiv:2412.06414
-
iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models 9 Dec 2024 · 1 repository · arXiv:2412.06263
-
LLM-BIP: Structured Pruning for Large Language Models with Block-Wise Forward Importance Propagation 9 Dec 2024 · 0 repositories · arXiv:2412.06419
-
On How Iterative Magnitude Pruning Discovers Local Receptive Fields in Fully Connected Neural Networks 9 Dec 2024 · 0 repositories · arXiv:2412.06545
-
Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06458
-
SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations 9 Dec 2024 · 0 repositories · arXiv:2412.06878
-
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs 8 Dec 2024 · 1 repository · arXiv:2412.05819Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
DapperFL: Domain Adaptive Federated Learning with Model Fusion Pruning for Edge Devices 8 Dec 2024 · 1 repository · arXiv:2412.05823Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
FlexDiT: Dynamic Token Density Control for Diffusion Transformer 8 Dec 2024 · 1 repository · arXiv:2412.06028Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Adaptive Dropout for Pruning Conformers 6 Dec 2024 · 0 repositories · arXiv:2412.04836
-
Cross-Self KV Cache Pruning for Efficient Vision-Language Inference 5 Dec 2024 · 1 repository · arXiv:2412.04652Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
2DGS-Room: Seed-Guided 2D Gaussian Splatting with Geometric Constrains for High-Fidelity Indoor Scene Reconstruction 4 Dec 2024 · 0 repositories · arXiv:2412.03428
-
A Granger-Causal Perspective on Gradient Descent with Application to Pruning 4 Dec 2024 · 0 repositories · arXiv:2412.03035
-
A Stitch in Time Saves Nine: Small VLM is a Precise Guidance for Accelerating Large VLMs 4 Dec 2024 · 1 repository · arXiv:2412.03324Syntology official: harvested, nothing ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning 4 Dec 2024 · 1 repository · arXiv:2412.03248Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples) · 3 pointer-only (licence)
-
Designing DNNs for a trade-off between robustness and processing performance in embedded devices 4 Dec 2024 · 0 repositories · arXiv:2412.03682
-
Evaluating Single Event Upsets in Deep Neural Networks for Semantic Segmentation: an embedded system perspective 4 Dec 2024 · 2 repositories · arXiv:2412.03630
-
Unifying KV Cache Compression for Large Language Models with LeanKV 4 Dec 2024 · 0 repositories · arXiv:2412.03131
-
Efficient Model Compression Techniques with FishLeg 3 Dec 2024 · 0 repositories · arXiv:2412.02328
-
Effortless Efficiency: Low-Cost Pruning of Diffusion Models 3 Dec 2024 · 0 repositories · arXiv:2412.02852
-
6DOPE-GS: Online 6D Object Pose Estimation using Gaussian Splatting 2 Dec 2024 · 0 repositories · arXiv:2412.01543
-
Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs 2 Dec 2024 · 2 repositories · arXiv:2412.01818Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Efficient LLM Inference using Dynamic Input Pruning and Cache-Aware Masking 2 Dec 2024 · 0 repositories · arXiv:2412.01380
-
HDGS: Textured 2D Gaussian Splatting for Enhanced Scene Rendering 2 Dec 2024 · 0 repositories · arXiv:2412.01823
-
Research on Optimizing Real-Time Data Processing in High-Frequency Trading Algorithms using Machine Learning 2 Dec 2024 · 0 repositories · arXiv:2412.01062
-
TinyFusion: Diffusion Transformers Learned Shallow 2 Dec 2024 · 1 repository · arXiv:2412.01199Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Token Cropr: Faster ViTs for Quite a Few Tasks 1 Dec 2024 · 1 repository · arXiv:2412.00965
-
ATP-LLaVA: Adaptive Token Pruning for Large Vision Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00447
-
Pruned Convolutional Attention Network Based Wideband Spectrum Sensing with Sub-Nyquist Sampling 30 Nov 2024 · 1 repository · arXiv:2412.00562
-
Speedy-Splat: Fast 3D Gaussian Splatting with Sparse Pixels and Sparse Primitives 30 Nov 2024 · 1 repository · arXiv:2412.00578Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Non-linear Equalization in 112 Gb/s PONs Using Kolmogorov-Arnold Networks 29 Nov 2024 · 0 repositories · arXiv:2411.19631
-
Video Set Distillation: Information Diversification and Temporal Densification 28 Nov 2024 · 0 repositories · arXiv:2412.00111
-
Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models 27 Nov 2024 · 0 repositories · arXiv:2411.18375
-
Preserving Deep Representations In One-Shot Pruning: A Hessian-Free Second-Order Optimization Framework 27 Nov 2024 · 0 repositories · arXiv:2411.18376
-
Preserving Information: How does Topological Data Analysis improve Neural Network performance? 27 Nov 2024 · 0 repositories · arXiv:2411.18410
-
Pruning Deep Convolutional Neural Network Using Conditional Mutual Information 27 Nov 2024 · 0 repositories · arXiv:2411.18578
-
Training Noise Token Pruning 27 Nov 2024 · 1 repository · arXiv:2411.18092
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distractor-free Generalizable 3D Gaussian Splatting 26 Nov 2024 · 1 repository · arXiv:2411.17605
-
Scalable iterative pruning of large language and vision models using block coordinate descent 26 Nov 2024 · 0 repositories · arXiv:2411.17796
-
Training a neural netwok for data reduction and better generalization 26 Nov 2024 · 1 repository · arXiv:2411.17180
-
Curvature in the Looking-Glass: Optimal Methods to Exploit Curvature of Expectation in the Loss Landscape 25 Nov 2024 · 0 repositories · arXiv:2411.16914
-
Deep Convolutional Neural Networks Structured Pruning via Gravity Regularization 25 Nov 2024 · 0 repositories · arXiv:2411.16901
-
Data Lineage Inference: Uncovering Privacy Vulnerabilities of Dataset Pruning 24 Nov 2024 · 0 repositories · arXiv:2411.15796