Methods › General › Model Compression › Pruning › Papers, page 8
Pruning
Papers archive 2025-07-28
archive papers tagged: 3,874 · with a code link: 1,508 · where Syntology ran a sample: 478 (395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (478 of 3,874 tagged: 395 with a run with no instrument failure, 83 where every run was a failure of Syntology's instrument)
Page 8 of 39: papers 701 to 800 of 3,874, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Efficient Partitioning Vision Transformer on Edge Devices for Distributed Inference 15 Oct 2024 · 0 repositories · arXiv:2410.11650
-
MCGS: Multiview Consistency Enhancement for Sparse-View 3D Gaussian Radiance Fields 15 Oct 2024 · 0 repositories · arXiv:2410.11394
-
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router 15 Oct 2024 · 0 repositories · arXiv:2410.12013
-
3D-Prover: Diversity Driven Theorem Proving With Determinantal Point Processes 14 Oct 2024 · 0 repositories · arXiv:2410.11133
-
Adapt-∞: Scalable Lifelong Multimodal Instruction Tuning via Dynamic Data Selection 14 Oct 2024 · 1 repository · arXiv:2410.10636
-
AlphaPruning: Using Heavy-Tailed Self Regularization Theory for Improved Layer-wise Pruning of Large Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10912Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads 14 Oct 2024 · 2 repositories · arXiv:2410.10819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Edge Unlearning is Not "on Edge"! An Adaptive Exact Unlearning System on Resource-Constrained Devices 14 Oct 2024 · 1 repository · arXiv:2410.10128Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Real-Time Stress Detection via Photoplethysmogram Signals: Implementation of a Combined Continuous Wavelet Transform and Convolutional Neural Network on Resource-Constrained Microcontrollers 14 Oct 2024 · 0 repositories · arXiv:2410.19776
-
SGLP: A Similarity Guided Fast Layer Partition Pruning for Compressing Large Deep Models 14 Oct 2024 · 1 repository · arXiv:2410.14720
-
Generalized Group Data Attribution 13 Oct 2024 · 0 repositories · arXiv:2410.09940
-
Self-Data Distillation for Recovering Quality in Pruned Large Language Models 13 Oct 2024 · 0 repositories · arXiv:2410.09982
-
DARE the Extreme: Revisiting Delta-Parameter Pruning For Fine-Tuned Models 12 Oct 2024 · 1 repository · arXiv:2410.09344Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Exploring space efficiency in a tree-based linear model for extreme multi-label classification 12 Oct 2024 · 0 repositories · arXiv:2410.09554
-
SLiM: One-shot Quantization and Sparsity with Low-rank Approximation for LLM Weight Compression 12 Oct 2024 · 1 repository · arXiv:2410.09615Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 16 harvested samples)
-
Token Pruning using a Lightweight Background Aware Vision Transformer 12 Oct 2024 · 0 repositories · arXiv:2410.09324
-
AMPO: Automatic Multi-Branched Prompt Optimization 11 Oct 2024 · 0 repositories · arXiv:2410.08696
-
Efficient Multi-Object Tracking on Edge Devices via Reconstruction-Based Channel Pruning 11 Oct 2024 · 0 repositories · arXiv:2410.08769
-
Unity is Power: Semi-Asynchronous Collaborative Training of Large-Scale Models with Structured Pruning in Resource-Limited Clients 11 Oct 2024 · 0 repositories · arXiv:2410.08457
-
A Unified Debiasing Approach for Vision-Language Models across Modalities and Tasks 10 Oct 2024 · 1 repository · arXiv:2410.07593Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
FusionSense: Bridging Common Sense, Vision, and Touch for Robust Sparse-View Reconstruction 10 Oct 2024 · 0 repositories · arXiv:2410.08282
-
Neuroplastic Expansion in Deep Reinforcement Learning 10 Oct 2024 · 0 repositories · arXiv:2410.07994
-
Non-transferable Pruning 10 Oct 2024 · 0 repositories · arXiv:2410.08015
-
Chip-Tuning: Classify Before Language Models Say 9 Oct 2024 · 1 repository · arXiv:2410.06541
-
Context-Augmented Code Generation Using Programming Knowledge Graphs 9 Oct 2024 · 0 repositories · arXiv:2410.18251
-
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization 9 Oct 2024 · 0 repositories · arXiv:2410.06567
-
Enhancing Vision-Language Model Pre-training with Image-text Pair Pruning Based on Word Frequency 9 Oct 2024 · 1 repository · arXiv:2410.10879
-
Is C4 Dataset Optimal for Pruning? An Investigation of Calibration Data for LLM Pruning 9 Oct 2024 · 1 repository · arXiv:2410.07461Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Learning Content-Aware Multi-Modal Joint Input Pruning via Bird's-Eye-View Representation 9 Oct 2024 · 0 repositories · arXiv:2410.07268
-
Large Language Model Compression with Neural Architecture Search 9 Oct 2024 · 0 repositories · arXiv:2410.06479
-
S2HPruner: Soft-to-Hard Distillation Bridges the Discretization Gap in Pruning 9 Oct 2024 · 0 repositories · arXiv:2410.07046Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
WALL-E: World Alignment by Rule Learning Improves World Model-based LLM Agents 9 Oct 2024 · 0 repositories · arXiv:2410.07484
-
MC-MoE: Mixture Compressor for Mixture-of-Experts LLMs Gains More 8 Oct 2024 · 1 repository · arXiv:2410.06270Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Ordering-Based Causal Discovery for Linear and Nonlinear Relations 8 Oct 2024 · 1 repository · arXiv:2410.05890Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Posets and Bounded Probabilities for Discovering Order-inducing Features in Event Knowledge Graphs 8 Oct 2024 · 0 repositories · arXiv:2410.06065
-
Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See 8 Oct 2024 · 1 repository · arXiv:2410.06169Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
LPZero: Language Model Zero-cost Proxy Search from Zero 7 Oct 2024 · 0 repositories · arXiv:2410.04808
-
Mastering Chinese Chess AI (Xiangqi) Without Search 7 Oct 2024 · 0 repositories · arXiv:2410.04865
-
Language Model-Driven Data Pruning Enables Efficient Active Learning 5 Oct 2024 · 0 repositories · arXiv:2410.04275
-
Toxic Subword Pruning for Dialogue Response Generation on Large Language Models 5 Oct 2024 · 0 repositories · arXiv:2410.04155
-
LLM-TOPLA: Efficient LLM Ensemble by Maximising Diversity 4 Oct 2024 · 1 repository · arXiv:2410.03953Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Minimax-optimal and Locally-adaptive Online Nonparametric Regression 4 Oct 2024 · 0 repositories · arXiv:2410.03363
-
Cut the Crap: An Economical Communication Pipeline for LLM-based Multi-Agent Systems 3 Oct 2024 · 0 repositories · arXiv:2410.02506
-
Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection 3 Oct 2024 · 1 repository · arXiv:2410.02330
-
Personalized Federated Learning for Generative AI-Assisted Semantic Communications 3 Oct 2024 · 0 repositories · arXiv:2410.02450
-
EntryPrune: Neural Network Feature Selection using First Impressions 3 Oct 2024 · 2 repositories · arXiv:2410.02344
-
DRUPI: Dataset Reduction Using Privileged Information 2 Oct 2024 · 0 repositories · arXiv:2410.01611
-
Mitigating Copy Bias in In-Context Learning through Neuron Pruning 2 Oct 2024 · 0 repositories · arXiv:2410.01288
-
Recovering Manifold Structure Using Ollivier-Ricci Curvature 2 Oct 2024 · 1 repository · arXiv:2410.01149
-
Review Non-convex Optimization Method for Machine Learning 2 Oct 2024 · 0 repositories · arXiv:2410.02017
-
Speculative Coreset Selection for Task-Specific Fine-tuning 2 Oct 2024 · 0 repositories · arXiv:2410.01296
-
Unveiling Language Skills via Path-Level Circuit Discovery 2 Oct 2024 · 1 repository · arXiv:2410.01334Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Drone Stereo Vision for Radiata Pine Branch Detection and Distance Measurement: Utilizing Deep Learning and YOLO Integration 1 Oct 2024 · 0 repositories · arXiv:2410.00503
-
Trainable pruned ternary quantization for medical signal classification models 1 Oct 2024 · 1 repository
-
Aggressive Post-Training Compression on Extremely Large Language Models 30 Sep 2024 · 0 repositories · arXiv:2409.20094
-
EEG Emotion Copilot: Optimizing Lightweight LLMs for Emotional EEG Interpretation with Assisted Medical Record Generation 30 Sep 2024 · 0 repositories · arXiv:2410.00166
-
Investigating the Effect of Network Pruning on Performance and Interpretability 29 Sep 2024 · 1 repository · arXiv:2409.19727
-
Exploring Token Pruning in Vision State Space Models 27 Sep 2024 · 0 repositories · arXiv:2409.18962
-
Localizing Memorization in SSL Vision Encoders 27 Sep 2024 · 0 repositories · arXiv:2409.19069
-
Mitigating Selection Bias with Node Pruning and Auxiliary Options 27 Sep 2024 · 0 repositories · arXiv:2409.18857
-
ProMerge: Prompt and Merge for Unsupervised Instance Segmentation 27 Sep 2024 · 0 repositories · arXiv:2409.18961
-
Speech Boosting: Low-Latency Live Speech Enhancement for TWS Earbuds 27 Sep 2024 · 0 repositories · arXiv:2409.18705
-
Token Caching for Diffusion Transformer Acceleration 27 Sep 2024 · 0 repositories · arXiv:2409.18523
-
Two Sparse Matrices are Better than One: Sparsifying Neural Networks with Double Sparse Factorization 27 Sep 2024 · 1 repository · arXiv:2409.18850Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
AlterMOMA: Fusion Redundancy Pruning for Camera-LiDAR Fusion Models with Alternative Modality Masking 26 Sep 2024 · 0 repositories · arXiv:2409.17728
-
BEATS: Optimizing LLM Mathematical Capabilities with BackVerify and Adaptive Disambiguate based Efficient Tree Search 26 Sep 2024 · 1 repository · arXiv:2409.17972
-
CRoP: Context-wise Robust Static Human-Sensing Personalization 26 Sep 2024 · 0 repositories · arXiv:2409.17994
-
Drone Stereo Vision for Radiata Pine Branch Detection and Distance Measurement: Integrating SGBM and Segmentation Models 26 Sep 2024 · 0 repositories · arXiv:2409.17526
-
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models 26 Sep 2024 · 1 repository · arXiv:2409.17481Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Graph Pruning Based Spatial and Temporal Graph Convolutional Network with Transfer Learning for Traffic Prediction 25 Sep 2024 · 1 repository · arXiv:2409.16532
-
Pruning Multilingual Large Language Models for Multilingual Inference 25 Sep 2024 · 1 repository · arXiv:2409.16911
-
Search for Efficient Large Language Models 25 Sep 2024 · 1 repository · arXiv:2409.17372Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
KISS-Matcher: Fast and Robust Point Cloud Registration Revisited 23 Sep 2024 · 1 repository · arXiv:2409.15615
-
Mixture of Efficient Diffusion Experts Through Automatic Interval and Sub-Network Selection 23 Sep 2024 · 0 repositories · arXiv:2409.15557
-
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches 22 Sep 2024 · 1 repository · arXiv:2409.14607
-
SPAQ-DL-SLAM: Towards Optimizing Deep Learning-based SLAM for Resource-Constrained Embedded Platforms 22 Sep 2024 · 0 repositories · arXiv:2409.14515
-
Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis 21 Sep 2024 · 2 repositories · arXiv:2409.14144Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
On Importance of Pruning and Distillation for Efficient Low Resource NLP 21 Sep 2024 · 0 repositories · arXiv:2409.14162
-
Towards Building Efficient Sentence BERT Models using Layer Pruning 21 Sep 2024 · 0 repositories · arXiv:2409.14168
-
Training Large ASR Encoders with Differential Privacy 21 Sep 2024 · 0 repositories · arXiv:2409.13953
-
CFSP: An Efficient Structured Pruning Framework for LLMs with Coarse-to-Fine Activation Information 20 Sep 2024 · 1 repository · arXiv:2409.13199
-
Data Pruning via Separability, Integrity, and Model Uncertainty-Aware Importance Sampling 20 Sep 2024 · 0 repositories · arXiv:2409.13915
-
OATS: Outlier-Aware Pruning Through Sparse and Low Rank Decomposition 20 Sep 2024 · 1 repository · arXiv:2409.13652Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians 20 Sep 2024 · 1 repository · arXiv:2409.13648Syntology 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
CritiPrefill: A Segment-wise Criticality-based Approach for Prefilling Acceleration in LLMs 19 Sep 2024 · 1 repository · arXiv:2409.12490Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Impact of ML Optimization Tactics on Greener Pre-Trained ML Models 19 Sep 2024 · 0 repositories · arXiv:2409.12878
-
Exploring Representations and Interventions in Time Series Foundation Models 19 Sep 2024 · 0 repositories · arXiv:2409.12915
-
Agglomerative Token Clustering 18 Sep 2024 · 0 repositories · arXiv:2409.11923
-
FAST GDRNPP: Improving the Speed of State-of-the-Art 6D Object Pose Estimation 18 Sep 2024 · 0 repositories · arXiv:2409.12720
-
Multi-Grid Graph Neural Networks with Self-Attention for Computational Mechanics 18 Sep 2024 · 1 repository · arXiv:2409.11899Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Evaluating the Impact of Compression Techniques on Task-Specific Performance of Large Language Models 17 Sep 2024 · 0 repositories · arXiv:2409.11233
-
KVPruner: Structural Pruning for Faster and Memory-Efficient Large Language Models 17 Sep 2024 · 0 repositories · arXiv:2409.11057
-
Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models 16 Sep 2024 · 1 repository · arXiv:2409.10197Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Large Language Model Enhanced Hard Sample Identification for Denoising Recommendation 16 Sep 2024 · 0 repositories · arXiv:2409.10343
-
Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies 16 Sep 2024 · 0 repositories · arXiv:2409.10218
-
Self-Tuning Spectral Clustering for Speaker Diarization 16 Sep 2024 · 1 repository · arXiv:2410.00023
-
Entity-Aware Self-Attention and Contextualized GCN for Enhanced Relation Extraction in Long Sentences 15 Sep 2024 · 0 repositories · arXiv:2409.13755
-
An Efficient Privacy-aware Split Learning Framework for Satellite Communications 13 Sep 2024 · 0 repositories · arXiv:2409.08538
-
Fast DCT+: A Family of Fast Transforms Based on Rank-One Updates of the Path Graph 13 Sep 2024 · 0 repositories · arXiv:2409.08970
-
S-STE: Continuous Pruning Function for Efficient 2:4 Sparse Pre-training 13 Sep 2024 · 2 repositories · arXiv:2409.09099