Browse State-of-the-Art › Quantization › Papers, page 17
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 17 of 50: papers 1,601 to 1,700 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Lightweight Federated Learning over Wireless Edge Networks13 Jul 2025 0 repositories listed
-
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation11 Jul 2025 0 repositories listed
-
GSVR: 2D Gaussian-based Video Representation for 800+ FPS with Hybrid Deformation Field8 Jul 2025 0 repositories listed
-
QS4D: Quantization-aware training for efficient hardware deployment of structured state-space sequential models8 Jul 2025 0 repositories listed
-
Semantic Certainty Assessment in Vector Retrieval Systems: A Novel Framework for Embedding Quality Evaluation8 Jul 2025 0 repositories listed
-
Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis2 Jul 2025 0 repositories listed
-
Analysis of Null Related Beampattern Measures and Signal Quantization Effects for Linear Differential Microphone Arrays26 Jun 2025 0 repositories listed
-
DipSVD: Dual-importance Protected SVD for Efficient LLM Compression25 Jun 2025 0 repositories listed
-
Joint Quantization and Pruning Neural Networks Approach: A Case Study on FSO Receivers25 Jun 2025 0 repositories listed
-
Cross-Layer Discrete Concept Discovery for Interpreting Language Models24 Jun 2025 0 repositories listed
-
Variational Bayesian Channel Estimation and Data Detection for Cell-Free Massive MIMO with Low-Resolution Quantized Fronthaul Links23 Jun 2025 0 repositories listed
-
LVPNet: A Latent-variable-based Prediction-driven End-to-end Framework for Lossless Compression of Medical Images22 Jun 2025 0 repositories listed
-
StainPIDR: A Pathological Image Decouplingand Reconstruction Method for Stain Normalization Based on Color Vector Quantization and Structure Restaining22 Jun 2025 0 repositories listed
-
TROJAN-GUARD: Hardware Trojans Detection Using GNN in RTL Designs22 Jun 2025 0 repositories listed
-
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models21 Jun 2025 0 repositories listed
-
A Simple Contrastive Framework Of Item Tokenization For Generative Recommendation20 Jun 2025 0 repositories listed
-
The Hidden Cost of an Image: Quantifying the Energy Consumption of AI Image Generation20 Jun 2025 0 repositories listed
-
On Designing Modulation for Over-the-Air Computation -- Part I: Noise-Aware Design19 Jun 2025 0 repositories listed
-
PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models19 Jun 2025 0 repositories listed
-
Effect of Signal Quantization on Performance Measures of a 1st Order One Dimensional Differential Microphone Array18 Jun 2025 0 repositories listed
-
J3DAI: A tiny DNN-Based Edge AI Accelerator for 3D-Stacked CMOS Image Sensor18 Jun 2025 0 repositories listed
-
Compressed Video Super-Resolution based on Hierarchical Encoding17 Jun 2025 0 repositories listed
-
Cost-Aware Routing for Efficient Text-To-Image Generation17 Jun 2025 0 repositories listed
-
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models17 Jun 2025 0 repositories listed
-
EAQuant: Enhancing Post-Training Quantization for MoE Models via Expert-Aware Optimization16 Jun 2025 0 repositories listed
-
ROSAQ: Rotation-based Saliency-Aware Weight Quantization for Efficiently Compressing Large Language Models16 Jun 2025 0 repositories listed
-
Serving Large Language Models on Huawei CloudMatrix38415 Jun 2025 0 repositories listed
-
Quantizing Small-Scale State-Space Models for Edge AI14 Jun 2025 0 repositories listed
-
Relative Entropy Regularized Reinforcement Learning for Efficient Encrypted Policy Synthesis14 Jun 2025 0 repositories listed
-
Deep Learning Model Acceleration and Optimization Strategies for Real-Time Recommendation Systems13 Jun 2025 0 repositories listed
-
GPLQ: A General, Practical, and Lightning QAT Method for Vision Transformers13 Jun 2025 0 repositories listed
-
Discrete Audio Tokens: More Than a Survey!12 Jun 2025 0 repositories listed
-
Starting Positions Matter: A Study on Better Weight Initialization for Neural Network Quantization12 Jun 2025 0 repositories listed
-
MNN-LLM: A Generic Inference Engine for Fast Large Language Model Deployment on Mobile Devices12 Jun 2025 0 repositories listed
-
Post-Training Quantization for Video Matting12 Jun 2025 0 repositories listed
-
AWP: Activation-Aware Weight Pruning and Quantization with Projected Gradient Descent11 Jun 2025 0 repositories listed
-
HadaNorm: Diffusion Transformer Quantization through Mean-Centered Transformations11 Jun 2025 0 repositories listed
-
Q-SAM2: Accurate Quantization for Segment Anything Model 211 Jun 2025 0 repositories listed
-
SLED: A Speculative LLM Decoding Framework for Efficient Edge Serving11 Jun 2025 0 repositories listed
-
Optimizing Learned Image Compression on Scalar and Entropy-Constraint Quantization10 Jun 2025 0 repositories listed
-
POLARON: Precision-aware On-device Learning and Adaptive Runtime-cONfigurable AI acceleration10 Jun 2025 0 repositories listed
-
Implementing Keyword Spotting on the MCUX947 Microcontroller with Integrated NPU10 Jun 2025 0 repositories listed
-
Hardware Limitations and Optimization Approach in 1-Bit RIS Design at 28 GHz10 Jun 2025 0 repositories listed
-
Decentralized Optimization on Compact Submanifolds by Quantized Riemannian Gradient Tracking9 Jun 2025 0 repositories listed
-
LiteVLM: A Low-Latency Vision-Language Model Inference Pipeline for Resource-Constrained Environments9 Jun 2025 0 repositories listed
-
Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test8 Jun 2025 0 repositories listed
-
QForce-RL: Quantized FPGA-Optimized Reinforcement Learning Compute Engine8 Jun 2025 0 repositories listed
-
Enabling On-Device Medical AI Assistants via Input-Driven Saliency Adaptation7 Jun 2025 0 repositories listed
-
Towards AI-Native Fronthaul: Neural Compression for NextG Cloud RAN7 Jun 2025 0 repositories listed
-
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning6 Jun 2025 0 repositories listed
-
Bridging the Modality Gap: Softly Discretizing Audio Representation for LLM-based Automatic Speech Recognition6 Jun 2025 0 repositories listed
-
FPSAttention: Training-Aware FP8 and Sparsity Co-Design for Fast Video Diffusion5 Jun 2025 0 repositories listed
-
FPTQuant: Function-Preserving Transforms for LLM Quantization5 Jun 2025 0 repositories listed
-
Kernel k-Medoids as General Vector Quantization5 Jun 2025 0 repositories listed
-
Massive MIMO with 1-Bit DACs: Data Detection for Quantized Linear Precoding with Dithering5 Jun 2025 0 repositories listed
-
PCDVQ: Enhancing Vector Quantization for Large Language Models via Polar Coordinate Decoupling5 Jun 2025 0 repositories listed
-
TaDA: Training-free recipe for Decoding with Adaptive KV Cache Compression and Mean-centering5 Jun 2025 0 repositories listed
-
BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing4 Jun 2025 0 repositories listed
-
Nonlinear Sparse Bayesian Learning Methods with Application to Massive MIMO Channel Estimation with Hardware Impairments4 Jun 2025 0 repositories listed
-
Enhancing Convergence, Privacy and Fairness for Wireless Personalized Federated Learning: Quantization-Assisted Min-Max Fair Scheduling3 Jun 2025 0 repositories listed
-
MUC-G4: Minimal Unsat Core-Guided Incremental Verification for Deep Neural Network Compression3 Jun 2025 0 repositories listed
-
Quantized Dissipative Uncertain Model for Fractional T_S Fuzzy systems with Time_Varying Delays Under Networked Control System3 Jun 2025 0 repositories listed
-
Enhancing Speech Emotion Recognition with Graph-Based Multimodal Fusion and Prosodic Features for the Speech Emotion Recognition in Naturalistic Conditions Challenge at Interspeech 20252 Jun 2025 0 repositories listed
-
Quantitative Error Feedback for Quantization Noise Reduction of Filtering over Graphs2 Jun 2025 0 repositories listed
-
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer1 Jun 2025 0 repositories listed
-
Quantization-based Bounds on the Wasserstein Metric1 Jun 2025 0 repositories listed
-
Power-of-Two (PoT) Weights in Large Language Models (LLMs)31 May 2025 0 repositories listed
-
Edge Computing for Physics-Driven AI in Computational MRI: A Feasibility Study30 May 2025 0 repositories listed
-
30 May 2025 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Running Conventional Automatic Speech Recognition on Memristor Hardware: A Simulated Approach30 May 2025 0 repositories listed
-
Efficient Quantum Approximate kNN Algorithm via Granular-Ball Computing29 May 2025 0 repositories listed
-
MuLoCo: Muon is a practical inner optimizer for DiLoCo29 May 2025 0 repositories listed
-
Revisiting Uncertainty Estimation and Calibration of Large Language Models29 May 2025 0 repositories listed
-
Highly Efficient and Effective LLMs with Multi-Boolean Architectures28 May 2025 0 repositories listed
-
On the Interplay of Privacy, Persuasion and Quantization28 May 2025 0 repositories listed
-
BrainStratify: Coarse-to-Fine Disentanglement of Intracranial Neural Dynamics26 May 2025 0 repositories listed
-
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge26 May 2025 0 repositories listed
-
LPCM: Learning-based Predictive Coding for LiDAR Point Cloud Compression26 May 2025 0 repositories listed
-
Small Language Models: Architectures, Techniques, Evaluation, Problems and Future Adaptation26 May 2025 0 repositories listed
-
FastMamba: A High-Speed and Efficient Mamba Accelerator on FPGA with Accurate Quantization25 May 2025 0 repositories listed
-
Distinctive Feature Codec: Adaptive Segmentation for Efficient Speech Representation24 May 2025 0 repositories listed
-
Efficient and Workload-Aware LLM Serving via Runtime Layer Swapping and KV Cache Resizing24 May 2025 0 repositories listed
-
Beyond Discreteness: Finite-Sample Analysis of Straight-Through Estimator for Quantization23 May 2025 0 repositories listed
-
NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache23 May 2025 0 repositories listed
-
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM23 May 2025 0 repositories listed
-
Task Specific Pruning with LLM-Sieve: How Many Parameters Does Your Task Really Need?23 May 2025 0 repositories listed
-
Is Quantum Optimization Ready? An Effort Towards Neural Network Compression using Adiabatic Quantum Computing22 May 2025 0 repositories listed
-
NQKV: A KV Cache Quantization Scheme Based on Normal Distribution Characteristics22 May 2025 0 repositories listed
-
Harnessing Large Language Models Locally: Empirical Results and Implications for AI PC21 May 2025 0 repositories listed
-
InTreeger: An End-to-End Framework for Integer-Only Decision Tree Inference21 May 2025 0 repositories listed
-
Is (Selective) Round-To-Nearest Quantization All You Need?21 May 2025 0 repositories listed
-
Rate-Distortion Optimization with Non-Reference Metrics for UGC Compression21 May 2025 0 repositories listed
-
Segmentation-Variant Codebooks for Preservation of Paralinguistic and Prosodic Information21 May 2025 0 repositories listed
-
EfficientLLM: Efficiency in Large Language Models20 May 2025 0 repositories listed
-
Layer-wise Quantization for Quantized Optimistic Dual Averaging20 May 2025 0 repositories listed
-
Through a Compressed Lens: Investigating the Impact of Quantization on LLM Explainability and Interpretability20 May 2025 0 repositories listed
-
A3 : an Analytical Low-Rank Approximation Framework for Attention19 May 2025 0 repositories listed
-
Automatic mixed precision for optimizing gained time with constrained loss mean-squared-error based on model partition to sequential sub-graphs19 May 2025 0 repositories listed
-
Deep Unfolding with Kernel-based Quantization in MIMO Detection19 May 2025 0 repositories listed
-
GANCompress: GAN-Enhanced Neural Image Compression with Binary Spherical Quantization19 May 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.