Methods › General › Model Compression › Pruning › Papers, page 7
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 7 of 39: papers 601 to 700 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions 24 Nov 2024 · 1 repository · arXiv:2411.16760
-
Reassessing Layer Pruning in LLMs: New Insights and Methods 23 Nov 2024 · 1 repository · arXiv:2411.15558
-
DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models 22 Nov 2024 · 1 repository · arXiv:2411.15024Syntology official (archive's flag): 5 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 6 pointer-only (licence)
-
Efficient Pruning of Text-to-Image Models: Insights from Pruning Stable Diffusion 22 Nov 2024 · 0 repositories · arXiv:2411.15113
-
AutoMixQ: Self-Adjusting Quantization for High Performance Memory-Efficient Fine-Tuning 21 Nov 2024 · 0 repositories · arXiv:2411.13814
-
DRPruning: Efficient Large Language Model Pruning through Distributionally Robust Optimization 21 Nov 2024 · 1 repository · arXiv:2411.14055Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
FoPru: Focal Pruning for Efficient Large Vision-Language Models 21 Nov 2024 · 0 repositories · arXiv:2411.14164
-
FuseGPT: Learnable Layers Fusion of Generative Pre-trained Transformers 21 Nov 2024 · 1 repository · arXiv:2411.14507
-
Layer Pruning with Consensus: A Triple-Win Solution 21 Nov 2024 · 1 repository · arXiv:2411.14345
-
Adversarial Diffusion Compression for Real-World Image Super-Resolution 20 Nov 2024 · 2 repositories · arXiv:2411.13383Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 4 pointer-only (licence)
-
On-device Content-based Recommendation with Single-shot Embedding Pruning: A Cooperative Game Perspective 20 Nov 2024 · 1 repository · arXiv:2411.13052
-
C²INet: Realizing Incremental Trajectory Prediction with Prior-Aware Continual Causal Intervention 19 Nov 2024 · 0 repositories · arXiv:2411.12313
-
Data Pruning in Generative Diffusion Models 19 Nov 2024 · 1 repository · arXiv:2411.12523
-
DeTrigger: A Gradient-Centric Approach to Backdoor Attack Mitigation in Federated Learning 19 Nov 2024 · 0 repositories · arXiv:2411.12220
-
FGP: Feature-Gradient-Prune for Efficient Convolutional Layer Pruning 19 Nov 2024 · 1 repository · arXiv:2411.12781
-
Sparser Training for On-Device Recommendation Systems 19 Nov 2024 · 0 repositories · arXiv:2411.12205
-
Distill the Best, Ignore the Rest: Improving Dataset Distillation with Loss-Value-Based Pruning 18 Nov 2024 · 1 repository · arXiv:2411.12115Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Just Leaf It: Accelerating Diffusion Classifiers with Hierarchical Class Pruning 18 Nov 2024 · 0 repositories · arXiv:2411.12073
-
Electrostatic Force Regularization for Neural Structured Pruning 17 Nov 2024 · 0 repositories · arXiv:2411.11079
-
Enabling Explainable Recommendation in E-commerce with LLM-powered Product Knowledge Graph 17 Nov 2024 · 0 repositories · arXiv:2412.01837
-
An exploration of the effect of quantisation on energy consumption and inference time of StarCoder2 15 Nov 2024 · 1 repository · arXiv:2411.12758
-
Efficient Density Control for 3D Gaussian Splatting 15 Nov 2024 · 1 repository · arXiv:2411.10133Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 2 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Layer Importance and Hallucination Analysis in Large Language Models via Enhanced Activation Variance-Sparsity 15 Nov 2024 · 0 repositories · arXiv:2411.10069
-
RedTest: Towards Measuring Redundancy in Deep Neural Networks Effectively 15 Nov 2024 · 0 repositories · arXiv:2411.10507
-
P² Law: Scaling Law for Post-Training After Model Pruning 15 Nov 2024 · 0 repositories · arXiv:2411.10272
-
Systolic Arrays and Structured Pruning Co-design for Efficient Transformers in Edge Systems 15 Nov 2024 · 0 repositories · arXiv:2411.10285
-
The Spatial Complexity of Optical Computing and How to Reduce It 15 Nov 2024 · 1 repository · arXiv:2411.10435
-
Complexity-Aware Training of Deep Neural Networks for Optimal Structure Discovery 14 Nov 2024 · 0 repositories · arXiv:2411.09127
-
Ghost-Connect Net: A Generalization-Enhanced Guidance For Sparse Deep Networks Under Distribution Shifts 14 Nov 2024 · 0 repositories · arXiv:2411.09199
-
SCAN: Bootstrapping Contrastive Pre-training for Data Efficiency 14 Nov 2024 · 1 repository · arXiv:2411.09126
-
Efficient 3D Perception on Multi-Sweep Point Cloud with Gumbel Spatial Pruning 12 Nov 2024 · 0 repositories · arXiv:2411.07742
-
How To Discover Short, Shorter, and the Shortest Proofs of Unsatisfiability: A Branch-and-Bound Approach for Resolution Proof Length Minimization 12 Nov 2024 · 0 repositories · arXiv:2411.07955
-
OWLed: Outlier-weighed Layerwise Pruning for Efficient Autonomous Driving Framework 12 Nov 2024 · 1 repository · arXiv:2411.07711
-
Federated Learning Client Pruning for Noisy Labels 11 Nov 2024 · 1 repository · arXiv:2411.07391
-
The Super Weight in Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.07191Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Zeroth-Order Adaptive Neuron Alignment Based Pruning without Re-Training 11 Nov 2024 · 1 repository · arXiv:2411.07066
-
CULL-MT: Compression Using Language and Layer pruning for Machine Translation 10 Nov 2024 · 0 repositories · arXiv:2411.06506
-
RL-Pruner: Structured Pruning Using Reinforcement Learning for CNN Compression and Acceleration 10 Nov 2024 · 1 repository · arXiv:2411.06463
-
FGGP: Fixed-Rate Gradient-First Gradual Pruning 8 Nov 2024 · 0 repositories · arXiv:2411.05500
-
How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP 8 Nov 2024 · 0 repositories · arXiv:2411.05527
-
MicroScopiQ: Accelerating Foundational Models through Outlier-Aware Microscaling Quantization 8 Nov 2024 · 1 repository · arXiv:2411.05282Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
QuanCrypt-FL: Quantized Homomorphic Encryption with Pruning for Secure Federated Learning 8 Nov 2024 · 0 repositories · arXiv:2411.05260
-
Pruning Literals for Highly Efficient Explainability at Word Level 7 Nov 2024 · 0 repositories · arXiv:2411.04557
-
An Edge Computing-Based Solution for Real-Time Leaf Disease Classification using Thermal Imaging 6 Nov 2024 · 1 repository · arXiv:2411.03835
-
Human-in-the-Loop Feature Selection Using Interpretable Kolmogorov-Arnold Network-based Double Deep Q-Network 6 Nov 2024 · 0 repositories · arXiv:2411.03740
-
Optimal Defenses Against Gradient Reconstruction Attacks 6 Nov 2024 · 1 repository · arXiv:2411.03746
-
Towards Resource-Efficient Federated Learning in Industrial IoT for Multivariate Time Series Analysis 6 Nov 2024 · 0 repositories · arXiv:2411.03996
-
Change Is the Only Constant: Dynamic LLM Slicing based on Layer Redundancy 5 Nov 2024 · 1 repository · arXiv:2411.03513
-
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems 5 Nov 2024 · 1 repository · arXiv:2411.02959
-
Layer-Adaptive State Pruning for Deep State Space Models 5 Nov 2024 · 1 repository · arXiv:2411.02824Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Navigating Extremes: Dynamic Sparsity in Large Output Space 5 Nov 2024 · 1 repository · arXiv:2411.03171Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Privacy-Preserving Graph-Based Machine Learning with Fully Homomorphic Encryption for Collaborative Anti-Money Laundering 5 Nov 2024 · 1 repository · arXiv:2411.02926Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Automatic Structured Pruning for Efficient Architecture in Federated Learning 4 Nov 2024 · 1 repository · arXiv:2411.01759
-
Autoformulation of Mathematical Optimization Models Using LLMs 3 Nov 2024 · 0 repositories · arXiv:2411.01679
-
Efficient Model Compression for Bayesian Neural Networks 1 Nov 2024 · 0 repositories · arXiv:2411.00273
-
Magnitude Pruning of Large Pretrained Transformer Models with a Mixture Gaussian Prior 1 Nov 2024 · 0 repositories · arXiv:2411.00969
-
MBExplainer: Multilevel bandit-based explanations for downstream models with augmented graph embeddings 1 Nov 2024 · 0 repositories · arXiv:2411.00287
-
MoE-I²: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition 1 Nov 2024 · 0 repositories · arXiv:2411.01016
-
On the Impact of White-box Deployment Strategies for Edge AI on Latency and Model Performance 1 Nov 2024 · 0 repositories · arXiv:2411.00907
-
RAM: Replace Attention with MLP for Efficient Multivariate Time Series Forecasting 31 Oct 2024 · 0 repositories · arXiv:2410.24023
-
Context-Aware Token Selection and Packing for Enhanced Vision Transformer 31 Oct 2024 · 0 repositories · arXiv:2410.23608
-
Mutual Information Preserving Neural Network Pruning 31 Oct 2024 · 0 repositories · arXiv:2411.00147
-
RSL-SQL: Robust Schema Linking in Text-to-SQL Generation 31 Oct 2024 · 1 repository · arXiv:2411.00073Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CopRA: A Progressive LoRA Training Strategy 30 Oct 2024 · 0 repositories · arXiv:2410.22911
-
ELMGS: Enhancing memory and computation scaLability through coMpression for 3D Gaussian Splatting 30 Oct 2024 · 0 repositories · arXiv:2410.23213
-
Leveraging Recurrent Neural Networks for Predicting Motor Movements from Primate Motor Cortex Neural Recordings 29 Oct 2024 · 0 repositories · arXiv:2410.22283
-
Uncovering Capabilities of Model Pruning in Graph Contrastive Learning 27 Oct 2024 · 0 repositories · arXiv:2410.20356
-
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing 25 Oct 2024 · 1 repository · arXiv:2410.19552Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Rethinking Visual Dependency in Long-Context Reasoning for Large Vision-Language Models 25 Oct 2024 · 0 repositories · arXiv:2410.19732
-
Dynamic Vocabulary Pruning in Early-Exit LLMs 24 Oct 2024 · 1 repository · arXiv:2410.18952
-
PixelGaussian: Generalizable 3D Gaussian Reconstruction from Arbitrary Views 24 Oct 2024 · 1 repository · arXiv:2410.18979Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Tailored-LLaMA: Optimizing Few-Shot Learning in Pruned LLaMA Models with Task-Specific Prompts 24 Oct 2024 · 0 repositories · arXiv:2410.19185
-
Beware of Calibration Data for Pruning Large Language Models 23 Oct 2024 · 0 repositories · arXiv:2410.17711
-
LEGO: Language Model Building Blocks 23 Oct 2024 · 0 repositories · arXiv:2410.18287
-
PETAH: Parameter Efficient Task Adaptation for Hybrid Transformers in a resource-limited Context 23 Oct 2024 · 0 repositories · arXiv:2410.17661
-
Theoretically Grounded Pruning of Large Ground Sets for Constrained, Discrete Optimization 23 Oct 2024 · 0 repositories · arXiv:2410.17945
-
DiP-GO: A Diffusion Pruner via Few-step Gradient Optimization 22 Oct 2024 · 0 repositories · arXiv:2410.16942
-
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes 22 Oct 2024 · 1 repository · arXiv:2410.16930
-
Mitigating Vanishing Activations in Deep CapsNets Using Channel Pruning 22 Oct 2024 · 1 repository · arXiv:2410.16908
-
Self-calibration for Language Model Quantization and Pruning 22 Oct 2024 · 0 repositories · arXiv:2410.17170
-
Neural Search Space in Gboard Decoder 21 Oct 2024 · 0 repositories · arXiv:2410.15575
-
Pruning Foundation Models for High Accuracy without Retraining 21 Oct 2024 · 1 repository · arXiv:2410.15567
-
Small Contributions, Small Networks: Efficient Neural Network Pruning Based on Relative Importance 21 Oct 2024 · 0 repositories · arXiv:2410.16151
-
Beyond Pruning Criteria: The Dominant Role of Fine-Tuning and Adaptive Ratios in Neural Network Robustness 19 Oct 2024 · 0 repositories · arXiv:2410.15176
-
DPVS-Shapley:Faster and Universal Contribution Evaluation Component in Federated Learning 19 Oct 2024 · 0 repositories · arXiv:2410.15093
-
DMGNN: Detecting and Mitigating Backdoor Attacks in Graph Neural Networks 18 Oct 2024 · 0 repositories · arXiv:2410.14105
-
EvoPress: Towards Optimal Dynamic Model Compression via Evolutionary Search 18 Oct 2024 · 1 repository · arXiv:2410.14649Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 17 harvested samples)
-
FedSpaLLM: Federated Pruning of Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14852
-
Large Language Models Are Overparameterized Text Encoders 18 Oct 2024 · 0 repositories · arXiv:2410.14578
-
Paths-over-Graph: Knowledge Graph Empowered Large Language Model Reasoning 18 Oct 2024 · 1 repository · arXiv:2410.14211Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
The Propensity for Density in Feed-forward Models 18 Oct 2024 · 0 repositories · arXiv:2410.14461
-
TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling 18 Oct 2024 · 0 repositories · arXiv:2410.16033
-
GDeR: Safeguarding Efficiency, Balancing, and Robustness via Prototypical Graph Pruning 17 Oct 2024 · 1 repository · arXiv:2410.13761Syntology official (archive's flag): 12 ran · 12 ran (of which 4 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Linguistically Grounded Analysis of Language Models using Shapley Head Values 17 Oct 2024 · 0 repositories · arXiv:2410.13396
-
LLM-Rank: A Graph Theoretical Approach to Pruning Large Language Models 17 Oct 2024 · 1 repository · arXiv:2410.13299
-
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats 16 Oct 2024 · 0 repositories · arXiv:2410.12781
-
Rethinking Token Reduction for State Space Models 16 Oct 2024 · 1 repository · arXiv:2410.14725Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Is Less More? Exploring Token Condensation as Training-free Adaptation for CLIP 16 Oct 2024 · 1 repository · arXiv:2410.14729
-
Beyond Linear Approximations: A Novel Pruning Approach for Attention Matrix 15 Oct 2024 · 0 repositories · arXiv:2410.11261
-
DISP-LLM: Dimension-Independent Structural Pruning for Large Language Models 15 Oct 2024 · 1 repository · arXiv:2410.11988Syntology official (archive's flag): 4 ran · 7 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 3 pointer-only (licence)