Methods › General › Model Compression › Pruning › Papers, page 3
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 3 of 39: papers 201 to 300 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enhanced Pruning Strategy for Multi-Component Neural Architectures Using Component-Aware Graph Analysis 17 Apr 2025 · 0 repositories · arXiv:2504.13296
-
Towards Lossless Token Pruning in Late-Interaction Retrieval Models 17 Apr 2025 · 1 repository · arXiv:2504.12778
-
Sparsity Outperforms Low-Rank Projections in Few-Shot Adaptation 16 Apr 2025 · 1 repository · arXiv:2504.12436Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Unveiling Hidden Collaboration within Mixture-of-Experts in Large Language Models 16 Apr 2025 · 0 repositories · arXiv:2504.12359
-
Adaptively Pruned Spiking Neural Networks for Energy-Efficient Intracortical Neural Decoding 15 Apr 2025 · 0 repositories · arXiv:2504.11568
-
Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning 15 Apr 2025 · 0 repositories · arXiv:2504.11409
-
LVLM_CSP: Accelerating Large Vision Language Models via Clustering, Scattering, and Pruning for Reasoning Segmentation 15 Apr 2025 · 0 repositories · arXiv:2504.10854
-
TAMP: Token-Adaptive Layerwise Pruning in Multimodal Large Language Models 14 Apr 2025 · 1 repository · arXiv:2504.09897
-
An Enhanced Iterative Deepening Search Algorithm for the Unrestricted Container Rehandling Problem 12 Apr 2025 · 0 repositories · arXiv:2504.09046
-
Efficient Mixture of Geographical Species for On Device Wildlife Monitoring 11 Apr 2025 · 0 repositories · arXiv:2504.08620
-
PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models 11 Apr 2025 · 1 repository · arXiv:2504.08966Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Preserving Privacy Without Compromising Accuracy: Machine Unlearning for Handwritten Text Recognition 11 Apr 2025 · 0 repositories · arXiv:2504.08616
-
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models 10 Apr 2025 · 1 repository · arXiv:2504.07807
-
CAFE-AD: Cross-Scenario Adaptive Feature Enhancement for Trajectory Planning in Autonomous Driving 9 Apr 2025 · 1 repository · arXiv:2504.06584
-
Neural Signal Compression using RAMAN tinyML Accelerator for BCI Applications 9 Apr 2025 · 0 repositories · arXiv:2504.06996
-
Federated Neural Architecture Search with Model-Agnostic Meta Learning 8 Apr 2025 · 0 repositories · arXiv:2504.06457
-
Finding Fantastic Experts in MoEs: A Unified Study for Expert Dropping Strategies and Observations 8 Apr 2025 · 0 repositories · arXiv:2504.05586
-
Balancing Robustness and Efficiency in Embedded DNNs Through Activation Function Selection 7 Apr 2025 · 0 repositories · arXiv:2504.05119
-
Dynamic Vision Mamba 7 Apr 2025 · 1 repository · arXiv:2504.04787
-
Two is Better than One: Efficient Ensemble Defense for Robust and Compact Models 7 Apr 2025 · 0 repositories · arXiv:2504.04747
-
Hyperflows: Pruning Reveals the Importance of Weights 6 Apr 2025 · 0 repositories · arXiv:2504.05349
-
Saliency-driven Dynamic Token Pruning for Large Language Models 6 Apr 2025 · 0 repositories · arXiv:2504.04514
-
Thanos: A Block-wise Pruning Algorithm for Efficient Large Language Model Compression 6 Apr 2025 · 1 repository · arXiv:2504.05346
-
Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models 4 Apr 2025 · 0 repositories · arXiv:2504.03624
-
LearNAT: Learning NL2SQL with AST-guided Task Decomposition for Large Language Models 3 Apr 2025 · 0 repositories · arXiv:2504.02327
-
Scaling Video-Language Models to 10K Frames via Hierarchical Differential Distillation 3 Apr 2025 · 1 repository · arXiv:2504.02438Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MDP: Multidimensional Vision Model Pruning with Latency Constraint 2 Apr 2025 · 0 repositories · arXiv:2504.02168
-
Sky of Unlearning (SoUL): Rewiring Federated Machine Unlearning via Selective Pruning 2 Apr 2025 · 0 repositories · arXiv:2504.01705
-
ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning 2 Apr 2025 · 1 repository · arXiv:2504.01296Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
Token Pruning in Audio Transformers: Optimizing Performance and Decoding Patch Importance 2 Apr 2025 · 1 repository · arXiv:2504.01690
-
When Reasoning Meets Compression: Benchmarking Compressed Large Reasoning Models on Complex Reasoning Tasks 2 Apr 2025 · 0 repositories · arXiv:2504.02010
-
FedPaI: Achieving Extreme Sparsity in Federated Learning via Pruning at Initialization 1 Apr 2025 · 0 repositories · arXiv:2504.00308
-
Geometric Median Matching for Robust k-Subset Selection from Noisy Data 1 Apr 2025 · 0 repositories · arXiv:2504.00564
-
Neural Pruning for 3D Scene Reconstruction: Efficient NeRF Acceleration 1 Apr 2025 · 0 repositories · arXiv:2504.00950
-
Local Information Matters: Inference Acceleration For Grounded Conversation Generation Models Through Adaptive Local-Aware Token Pruning 31 Mar 2025 · 0 repositories · arXiv:2503.23959
-
Model Hemorrhage and the Robustness Limits of Large Language Models 31 Mar 2025 · 0 repositories · arXiv:2503.23924
-
Efficient Token Compression for Vision Transformer with Spatial Information Preserved 30 Mar 2025 · 1 repository · arXiv:2503.23455Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
FastVAR: Linear Visual Autoregressive Modeling via Cached Token Pruning 30 Mar 2025 · 1 repository · arXiv:2503.23367Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
Reinforcement Learning-based Token Pruning in Vision Transformers: A Markov Game Approach 30 Mar 2025 · 1 repository · arXiv:2503.23459
-
CoSIL: Software Issue Localization via LLM-Driven Code Repository Graph Searching 28 Mar 2025 · 1 repository · arXiv:2503.22424
-
CPPO: Accelerating the Training of Group Relative Policy Optimization-Based Reasoning Models 28 Mar 2025 · 1 repository · arXiv:2503.22342Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Learnable cut flow 28 Mar 2025 · 0 repositories · arXiv:2503.22498
-
STADE: Standard Deviation as a Pruning Metric 28 Mar 2025 · 1 repository · arXiv:2503.22451
-
A 71.2-μW Speech Recognition Accelerator with Recurrent Spiking Neural Network 27 Mar 2025 · 0 repositories · arXiv:2503.21337
-
A Low-Power Streaming Speech Enhancement Accelerator For Edge Devices 27 Mar 2025 · 0 repositories · arXiv:2503.21335
-
As easy as PIE: understanding when pruning causes language models to disagree 27 Mar 2025 · 1 repository · arXiv:2503.21714
-
Boosting Large Language Models with Mask Fine-Tuning 27 Mar 2025 · 1 repository · arXiv:2503.22764
-
Neuroplasticity in Artificial Intelligence -- An Overview and Inspirations on Drop In & Out Learning 27 Mar 2025 · 0 repositories · arXiv:2503.21419
-
Sparse Bayesian Learning for Label Efficiency in Cardiac Real-Time MRI 27 Mar 2025 · 0 repositories · arXiv:2503.21443
-
ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks 26 Mar 2025 · 0 repositories · arXiv:2503.21815
-
Beyond Intermediate States: Explaining Visual Redundancy through Language 26 Mar 2025 · 1 repository · arXiv:2503.20540
-
Lipschitz Constant Meets Condition Number: Learning Robust and Compact Deep Neural Networks 26 Mar 2025 · 0 repositories · arXiv:2503.20454
-
SURGEON: Memory-Adaptive Fully Test-Time Adaptation via Dynamic Activation Sparsity 26 Mar 2025 · 1 repository · arXiv:2503.20354Syntology 15 ran (of which 4 constructed an object rather than computing a result; 10 with no instrument failure: 5 honoured, 1 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models 25 Mar 2025 · 0 repositories · arXiv:2503.19482
-
MindfulLIME: A Stable Solution for Explanations of Machine Learning Models with Enhanced Localization Precision -- A Medical Image Case Study 25 Mar 2025 · 0 repositories · arXiv:2503.20758
-
Scaling Vision Pre-Training to 4K Resolution 25 Mar 2025 · 2 repositories · arXiv:2503.19903Syntology 11 ran (of which 1 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Maximum Redundancy Pruning: A Principle-Driven Layerwise Sparsity Allocation for LLMs 24 Mar 2025 · 0 repositories · arXiv:2503.18377
-
NexusGS: Sparse View Synthesis with Epipolar Depth Priors in 3D Gaussian Splatting 24 Mar 2025 · 0 repositories · arXiv:2503.18794
-
Severing Spurious Correlations with Data Pruning 24 Mar 2025 · 1 repository · arXiv:2503.18258Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TopV: Compatible Token Pruning with Inference Time Optimization for Fast and Low-Memory Multimodal Vision Language Model 24 Mar 2025 · 0 repositories · arXiv:2503.18278
-
Video-XL-Pro: Reconstructive Token Compression for Extremely Long Video Understanding 24 Mar 2025 · 0 repositories · arXiv:2503.18478
-
Agent-Based Models for Two Stocks with Superhedging 23 Mar 2025 · 0 repositories · arXiv:2503.18165
-
Finding Stable Subnetworks at Initialization with Dataset Distillation 23 Mar 2025 · 0 repositories · arXiv:2503.17905
-
Causal Inference based Transfer Learning with LLMs: An Efficient Framework for Industrial RUL Prediction 22 Mar 2025 · 0 repositories · arXiv:2503.17686
-
Energy-Aware LLMs: A step towards sustainable AI for downstream applications 22 Mar 2025 · 0 repositories · arXiv:2503.17783
-
Feather-SQL: A Lightweight NL2SQL Framework with Dual-Model Collaboration Paradigm for Small Language Models 22 Mar 2025 · 0 repositories · arXiv:2503.17811
-
Improving Quantization with Post-Training Model Expansion 21 Mar 2025 · 0 repositories · arXiv:2503.17513
-
Temporal Action Detection Model Compression by Progressive Block Drop 21 Mar 2025 · 0 repositories · arXiv:2503.16916
-
Token Dynamics: Towards Efficient and Dynamic Video Token Representation for Video Large Language Models 21 Mar 2025 · 0 repositories · arXiv:2503.16980
-
1000+ FPS 4D Gaussian Splatting for Dynamic Scene Rendering 20 Mar 2025 · 0 repositories · arXiv:2503.16422
-
Attention Pruning: Automated Fairness Repair of Language Models via Surrogate Simulated Annealing 20 Mar 2025 · 0 repositories · arXiv:2503.15815
-
Bézier Splatting for Fast and Differentiable Vector Graphics Rendering 20 Mar 2025 · 0 repositories · arXiv:2503.16424
-
PSA-MIL: A Probabilistic Spatial Attention-Based Multiple Instance Learning for Whole Slide Image Classification 20 Mar 2025 · 1 repository · arXiv:2503.16284
-
XAttention: Block Sparse Attention with Antidiagonal Scoring 20 Mar 2025 · 1 repository · arXiv:2503.16428
-
EfficientLLaVA:Generalizable Auto-Pruning for Large Vision-language Models 19 Mar 2025 · 0 repositories · arXiv:2503.15369
-
Multi-Agent Actor-Critic with Harmonic Annealing Pruning for Dynamic Spectrum Access Systems 19 Mar 2025 · 0 repositories · arXiv:2503.15172
-
Pruning-Based TinyML Optimization of Machine Learning Models for Anomaly Detection in Electric Vehicle Charging Infrastructure 19 Mar 2025 · 1 repository · arXiv:2503.14799
-
BG-Triangle: Bézier Gaussian Triangle for 3D Vectorization and Rendering 18 Mar 2025 · 0 repositories · arXiv:2503.13961
-
Growing a Twig to Accelerate Large Vision-Language Models 18 Mar 2025 · 0 repositories · arXiv:2503.14075
-
Improving Adaptive Density Control for 3D Gaussian Splatting 18 Mar 2025 · 1 repository · arXiv:2503.14274
-
Light4GS: Lightweight Compact 4D Gaussian Splatting Generation via Context Model 18 Mar 2025 · 0 repositories · arXiv:2503.13948
-
Lifting the Veil on Visual Information Flow in MLLMs: Unlocking Pathways to Faster Inference 17 Mar 2025 · 1 repository · arXiv:2503.13108Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
OptiPMB: Enhancing 3D Multi-Object Tracking with Optimized Poisson Multi-Bernoulli Filtering 17 Mar 2025 · 0 repositories · arXiv:2503.12968
-
φ-Decoding: Adaptive Foresight Sampling for Balanced Inference-Time Exploration and Exploitation 17 Mar 2025 · 2 repositories · arXiv:2503.13288
-
Scale Efficient Training for Large Datasets 17 Mar 2025 · 1 repository · arXiv:2503.13385Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
SparseLUT: Sparse Connectivity Optimization for Lookup Table-based Deep Neural Networks 17 Mar 2025 · 1 repository · arXiv:2503.12829
-
FastVID: Dynamic Density Pruning for Fast Video Large Language Models 14 Mar 2025 · 2 repositories · arXiv:2503.11187Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Multi-View Node Pruning for Accurate Graph Representation 14 Mar 2025 · 0 repositories · arXiv:2503.11737
-
Quantifying Interpretability in CLIP Models with Concept Consistency 14 Mar 2025 · 0 repositories · arXiv:2503.11103
-
Similarity-Aware Token Pruning: Your VLM but Faster 14 Mar 2025 · 1 repository · arXiv:2503.11549
-
Towards Extreme Pruning of LLMs with Plug-and-Play Mixed Sparsity 14 Mar 2025 · 0 repositories · arXiv:2503.11164
-
AMR-Transformer: Enabling Efficient Long-range Interaction for Complex Neural Fluid Simulation 13 Mar 2025 · 0 repositories · arXiv:2503.10257
-
AttentionRAG: Attention-Guided Context Pruning in Retrieval-Augmented Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10720
-
DeclareAligner: A Leap Towards Efficient Optimal Alignments for Declarative Process Model Conformance Checking 13 Mar 2025 · 0 repositories · arXiv:2503.10479
-
Keyframe-oriented Vision Token Pruning: Enhancing Efficiency of Large Vision Language Models on Long-Form Video Processing 13 Mar 2025 · 1 repository · arXiv:2503.10742Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
ROODI: Reconstructing Occluded Objects with Denoising Inpainters 13 Mar 2025 · 0 repositories · arXiv:2503.10256
-
ZeroMerge: Parameter-Free KV Cache Compression for Memory-Efficient Long-Context LLMs 13 Mar 2025 · 2 repositories · arXiv:2503.10714
-
Batch List-Decodable Linear Regression via Higher Moments 12 Mar 2025 · 0 repositories · arXiv:2503.09802
-
Týr-the-Pruner: Unlocking Accurate 50% Structural Pruning for LLMs via Global Sparsity Distribution Optimization 12 Mar 2025 · 0 repositories · arXiv:2503.09657
-
Accelerate 3D Object Detection Models via Zero-Shot Attention Key Pruning 11 Mar 2025 · 1 repository · arXiv:2503.08101Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)