Methods › General › Model Compression › Pruning › Papers, page 4
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 4 of 39: papers 301 to 400 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MVGSR: Multi-View Consistency Gaussian Splatting for Robust Surface Reconstruction 11 Mar 2025 · 0 repositories · arXiv:2503.08093
-
PRISM: Privacy-Preserving Improved Stochastic Masking for Federated Generative Models 11 Mar 2025 · 1 repository · arXiv:2503.08085
-
QuoTA: Query-oriented Token Assignment via CoT Query Decouple for Long Video Comprehension 11 Mar 2025 · 1 repository · arXiv:2503.08689Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DatawiseAgent: A Notebook-Centric LLM Agent Framework for Automated Data Science 10 Mar 2025 · 0 repositories · arXiv:2503.07044
-
Federated Multimodal Learning with Dual Adapters and Selective Pruning for Communication and Computational Efficiency 10 Mar 2025 · 1 repository · arXiv:2503.07552
-
Iterative Prompt Relocation for Distribution-Adaptive Visual Prompt Tuning 10 Mar 2025 · 1 repository · arXiv:2503.06901
-
SEAP: Training-free Sparse Expert Activation Pruning Unlock the Brainpower of Large Language Models 10 Mar 2025 · 1 repository · arXiv:2503.07605
-
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning 10 Mar 2025 · 1 repository · arXiv:2503.07588
-
Pre-Training Meta-Rule Selection Policy for Visual Generative Abductive Learning 9 Mar 2025 · 1 repository · arXiv:2503.06427
-
Disrupting Model Merging: A Parameter-Level Defense Without Sacrificing Accuracy 8 Mar 2025 · 0 repositories · arXiv:2503.07661
-
IteRABRe: Iterative Recovery-Aided Block Reduction 8 Mar 2025 · 0 repositories · arXiv:2503.06291
-
MAD-MAX: Modular And Diverse Malicious Attack MiXtures for Automated LLM Red Teaming 8 Mar 2025 · 0 repositories · arXiv:2503.06253
-
Sample-aware Adaptive Structured Pruning for Large Language Models 8 Mar 2025 · 0 repositories · arXiv:2503.06184
-
SecureGS: Boosting the Security and Fidelity of 3D Gaussian Splatting Steganography 8 Mar 2025 · 0 repositories · arXiv:2503.06118
-
D2GV: Deformable 2D Gaussian Splatting for Video Representation in 400FPS 7 Mar 2025 · 1 repository · arXiv:2503.05600
-
IDEA Prune: An Integrated Enlarge-and-Prune Pipeline in Generative Language Model Pretraining 7 Mar 2025 · 0 repositories · arXiv:2503.05920
-
Reward-Centered ReST-MCTS: A Robust Decision-Making Framework for Robotic Manipulation in High Uncertainty Environments 7 Mar 2025 · 1 repository · arXiv:2503.05226
-
PDX: A Data Layout for Vector Similarity Search 6 Mar 2025 · 1 repository · arXiv:2503.04422
-
Wanda++: Pruning Large Language Models via Regional Gradients 6 Mar 2025 · 1 repository · arXiv:2503.04992
-
LEWIS (LayEr WIse Sparsity) -- A Training Free Guided Model Merging Approach 5 Mar 2025 · 0 repositories · arXiv:2503.03874
-
Optimal Decision Tree Pruning Revisited: Algorithms and Complexity 5 Mar 2025 · 0 repositories · arXiv:2503.03576
-
RASD: Retrieval-Augmented Speculative Decoding 5 Mar 2025 · 0 repositories · arXiv:2503.03434
-
Skeletonisation Scale-Spaces 5 Mar 2025 · 0 repositories · arXiv:2503.03450
-
CORDIC Is All You Need 4 Mar 2025 · 0 repositories · arXiv:2503.11685
-
DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models 4 Mar 2025 · 1 repository · arXiv:2503.02175
-
FairSense-AI: Responsible AI Meets Sustainability 4 Mar 2025 · 0 repositories · arXiv:2503.02865
-
Towards a robust R2D2 paradigm for radio-interferometric imaging: revisiting DNN training and architecture 4 Mar 2025 · 0 repositories · arXiv:2503.02554
-
OpenGS-SLAM: Open-Set Dense Semantic SLAM with 3D Gaussian Splatting for Object-Level Scene Understanding 3 Mar 2025 · 0 repositories · arXiv:2503.01646
-
Transferring between sparse and dense matching via probabilistic reweighting 3 Mar 2025 · 0 repositories · arXiv:2503.01472
-
Training-Free Dataset Pruning for Instance Segmentation 2 Mar 2025 · 1 repository · arXiv:2503.00828Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Speculative Decoding and Beyond: An In-Depth Review of Techniques 27 Feb 2025 · 0 repositories · arXiv:2502.19732
-
A Sliding Layer Merging Method for Efficient Depth-Wise Pruning in LLMs 26 Feb 2025 · 1 repository · arXiv:2502.19159
-
Accurate 3D Grapevine Structure Extraction from High-Resolution Point Clouds 26 Feb 2025 · 0 repositories · arXiv:2502.20417
-
CABS: Conflict-Aware and Balanced Sparsification for Enhancing Model Merging 26 Feb 2025 · 0 repositories · arXiv:2503.01874
-
Kanana: Compute-efficient Bilingual Language Models 26 Feb 2025 · 0 repositories · arXiv:2502.18934
-
Compressing Language Models for Specialized Domains 25 Feb 2025 · 0 repositories · arXiv:2502.18424
-
On-device edge learning for IoT data streams: a survey 25 Feb 2025 · 0 repositories · arXiv:2502.17788
-
Optimal Brain Apoptosis 25 Feb 2025 · 1 repository · arXiv:2502.17941Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
CipherPrune: Efficient and Scalable Private Transformer Inference 24 Feb 2025 · 1 repository · arXiv:2502.16782Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
DBudgetKV: Dynamic Budget in KV Cache Compression for Ensuring Optimal Performance 24 Feb 2025 · 0 repositories · arXiv:2502.16886
-
Delta Decompression for MoE-based LLMs Compression 24 Feb 2025 · 1 repository · arXiv:2502.17298Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Geometric Properties and Graph-Based Optimization of Neural Networks: Addressing Non-Linearity, Dimensionality, and Scalability 24 Feb 2025 · 0 repositories · arXiv:2503.05761
-
Low-Rank and Sparse Model Merging for Multi-Lingual Speech Recognition and Translation 24 Feb 2025 · 0 repositories · arXiv:2502.17380
-
Systematic Weight Evaluation for Pruning Large Language Models: Enhancing Performance and Sustainability 24 Feb 2025 · 0 repositories · arXiv:2502.17071
-
Automatic Joint Structured Pruning and Quantization for Efficient Neural Network Training and Compression 23 Feb 2025 · 1 repository · arXiv:2502.16638Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Energy-Efficient Transformer Inference: Optimization Strategies for Time Series Classification 23 Feb 2025 · 0 repositories · arXiv:2502.16627
-
Lattice-Based Pruning in Recurrent Neural Networks via Poset Modeling 23 Feb 2025 · 0 repositories · arXiv:2502.16525
-
Model-agnostic Coreset Selection via LLM-based Concept Bottlenecks 23 Feb 2025 · 0 repositories · arXiv:2502.16733
-
Spectral Theory for Edge Pruning in Asynchronous Recurrent Graph Neural Networks 23 Feb 2025 · 0 repositories · arXiv:2502.17522
-
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning 22 Feb 2025 · 0 repositories · arXiv:2502.17510
-
ZiGong 1.0: A Large Language Model for Financial Credit 22 Feb 2025 · 0 repositories · arXiv:2502.16159
-
Modality-Aware Neuron Pruning for Unlearning in Multimodal Large Language Models 21 Feb 2025 · 1 repository · arXiv:2502.15910Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 12 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
PPC-GPT: Federated Task-Specific Compression of Large Language Models via Pruning and Chain-of-Thought Distillation 21 Feb 2025 · 0 repositories · arXiv:2502.15857
-
Probe Pruning: Accelerating LLMs through Dynamic Pruning via Model-Probing 21 Feb 2025 · 1 repository · arXiv:2502.15618Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
When Compression Meets Model Compression: Memory-Efficient Double Compression for Large Language Models 21 Feb 2025 · 0 repositories · arXiv:2502.15443
-
PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models 20 Feb 2025 · 0 repositories · arXiv:2502.14504
-
Towards Efficient Automatic Self-Pruning of Large Language Models 20 Feb 2025 · 0 repositories · arXiv:2502.14413
-
ETS: Efficient Tree Search for Inference-Time Scaling 19 Feb 2025 · 1 repository · arXiv:2502.13575
-
MaskPrune: Mask-based LLM Pruning for Layer-wise Uniform Structures 19 Feb 2025 · 0 repositories · arXiv:2502.14008
-
NVR: Vector Runahead on NPUs for Sparse Memory Access 19 Feb 2025 · 0 repositories · arXiv:2502.13873
-
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models 19 Feb 2025 · 1 repository · arXiv:2502.13533Syntology official (archive's flag): 4 ran · 5 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Boost, Disentangle, and Customize: A Robust System2-to-System1 Pipeline for Code Generation 18 Feb 2025 · 0 repositories · arXiv:2502.12492
-
DSMoE: Matrix-Partitioned Experts with Dynamic Routing for Computation-Efficient Dense LLMs 18 Feb 2025 · 0 repositories · arXiv:2502.12455
-
G-Refer: Graph Retrieval-Augmented Large Language Model for Explainable Recommendation 18 Feb 2025 · 1 repository · arXiv:2502.12586
-
Iron Sharpens Iron: Defending Against Attacks in Machine-Generated Text Detection with Adversarial Training 18 Feb 2025 · 1 repository · arXiv:2502.12734Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
NTP-INT: Network Traffic Prediction-Driven In-band Network Telemetry for High-load Switches 18 Feb 2025 · 0 repositories · arXiv:2502.12834
-
Pruning as a Defense: Reducing Memorization in Large Language Models 18 Feb 2025 · 0 repositories · arXiv:2502.15796
-
Signal Collapse in One-Shot Pruning: When Sparse Models Fail to Distinguish Neural Representations 18 Feb 2025 · 0 repositories · arXiv:2502.15790
-
An Efficient Row-Based Sparse Fine-Tuning 17 Feb 2025 · 0 repositories · arXiv:2502.11439
-
Dictionary-Learning-Based Data Pruning for System Identification 17 Feb 2025 · 0 repositories · arXiv:2502.11484
-
Fishing For Cheap And Efficient Pruners At Initialization 17 Feb 2025 · 1 repository · arXiv:2502.11450
-
FitLight: Federated Imitation Learning for Plug-and-Play Autonomous Traffic Signal Control 17 Feb 2025 · 0 repositories · arXiv:2502.11937
-
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL 17 Feb 2025 · 1 repository · arXiv:2502.11741Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More 17 Feb 2025 · 1 repository · arXiv:2502.11494Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem? 17 Feb 2025 · 0 repositories · arXiv:2502.11501
-
CacheFocus: Dynamic Cache Re-Positioning for Efficient Retrieval-Augmented Generation 16 Feb 2025 · 0 repositories · arXiv:2502.11101
-
GS-GVINS: A Tightly-integrated GNSS-Visual-Inertial Navigation System Augmented by 3D Gaussian Splatting 16 Feb 2025 · 0 repositories · arXiv:2502.10975
-
OPTISHEAR: Towards Efficient and Adaptive Pruning of Large Language Models via Evolutionary Optimization 15 Feb 2025 · 0 repositories · arXiv:2502.10735
-
Janus: Collaborative Vision Transformer Under Dynamic Network Environment 14 Feb 2025 · 0 repositories · arXiv:2502.10047
-
Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding 14 Feb 2025 · 1 repository · arXiv:2502.10392
-
Automatic Pruning via Structured Lasso with Class-wise Information 13 Feb 2025 · 0 repositories · arXiv:2502.09125
-
DenseSplat: Densifying Gaussian Splatting SLAM with Neural Radiance Prior 13 Feb 2025 · 0 repositories · arXiv:2502.09111
-
InfiniteHiP: Extending Language Model Context Up to 3 Million Tokens on a Single GPU 13 Feb 2025 · 0 repositories · arXiv:2502.08910
-
Scalable First-order Method for Certifying Optimal k-Sparse GLMs 13 Feb 2025 · 0 repositories · arXiv:2502.09502
-
Contextual Compression Encoding for Large Language Models: A Novel Framework for Multi-Layered Parameter Space Pruning 12 Feb 2025 · 0 repositories · arXiv:2502.08323
-
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation 12 Feb 2025 · 0 repositories · arXiv:2502.08690
-
Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding 12 Feb 2025 · 1 repository · arXiv:2502.08363
-
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies 11 Feb 2025 · 0 repositories · arXiv:2502.07771
-
DarwinLM: Evolutionary Structured Pruning of Large Language Models 11 Feb 2025 · 1 repository · arXiv:2502.07780
-
Exploring Neural Network Pruning with Screening Methods 11 Feb 2025 · 0 repositories · arXiv:2502.07189
-
Partial-Label Learning with Conformal Candidate Cleaning 11 Feb 2025 · 0 repositories · arXiv:2502.07661
-
Lightweight Dataset Pruning without Full Training via Example Difficulty and Prediction Uncertainty 10 Feb 2025 · 1 repository · arXiv:2502.06905
-
Low-Rank Compression for IMC Arrays 10 Feb 2025 · 0 repositories · arXiv:2502.07820
-
Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models 9 Feb 2025 · 0 repositories · arXiv:2502.05837
-
Diffusion Model for Interest Refinement in Multi-Interest Recommendation 8 Feb 2025 · 0 repositories · arXiv:2502.05561
-
Mix Data or Merge Models? Balancing the Helpfulness, Honesty, and Harmlessness of Large Language Model via Model Merging 8 Feb 2025 · 0 repositories · arXiv:2502.06876
-
Distillation and Pruning for Scalable Self-Supervised Representation-Based Speech Quality Assessment 7 Feb 2025 · 1 repository · arXiv:2502.05356
-
Model Fusion via Neuron Transplantation 7 Feb 2025 · 1 repository · arXiv:2502.06849
-
Detecting Backdoor Attacks via Similarity in Semantic Communication Systems 6 Feb 2025 · 0 repositories · arXiv:2502.03721
-
Identify Critical KV Cache in LLM Inference from an Output Perturbation Perspective 6 Feb 2025 · 2 repositories · arXiv:2502.03805Syntology official (archive's flag): 1 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 11 harvested samples)