Browse State-of-the-Art › Quantization › Papers, page 25
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 25 of 50: papers 2,401 to 2,500 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Quality Scalable Quantization Methodology for Deep Learning on Edge15 Jul 2024 0 repositories listed
-
A Bag of Tricks for Scaling CPU-based Deep FFMs to more than 300m Predictions per Second14 Jul 2024 0 repositories listed
-
14 Jul 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
One-Bit MIMO Detection: From Global Maximum-Likelihood Detector to Amplitude Retrieval Approach13 Jul 2024 0 repositories listed
-
Accuracy is Not All You Need12 Jul 2024 0 repositories listed
-
Optimization of DNN-based speaker verification model through efficient quantization technique12 Jul 2024 0 repositories listed
-
ADMM Based Semi-Structured Pattern Pruning Framework For Transformer11 Jul 2024 0 repositories listed
-
Autoregressive Speech Synthesis without Vector Quantization11 Jul 2024 0 repositories listed
-
Distributed Deep Reinforcement Learning Based Gradient Quantization for Federated Learning Enabled Vehicle Edge Computing11 Jul 2024 0 repositories listed
-
ERQ: Error Reduction for Post-Training Quantization of Vision Transformers9 Jul 2024 0 repositories listed
-
Ternary Spike-based Neuromorphic Signal Processing System7 Jul 2024 0 repositories listed
-
Balance of Number of Embedding and their Dimensions in Vector Quantization6 Jul 2024 0 repositories listed
-
Integer-only Quantized Transformers for Embedded FPGA-based Time-series Forecasting in AIoT6 Jul 2024 0 repositories listed
-
Quantizing YOLOv7: A Comprehensive Study6 Jul 2024 0 repositories listed
-
ZOBNN: Zero-Overhead Dependable Design of Binary Neural Networks with Deliberately Quantized Parameters6 Jul 2024 0 repositories listed
-
Hybrid Receiver Design for Massive MIMO-OFDM with Low-Resolution ADCs and Oversampling5 Jul 2024 0 repositories listed
-
The Impact of Quantization and Pruning on Deep Reinforcement Learning Models5 Jul 2024 0 repositories listed
-
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering4 Jul 2024 0 repositories listed
-
Joint Beamforming Design and Bit Allocation in Massive MIMO with Resolution-Adaptive ADCs4 Jul 2024 0 repositories listed
-
Low-latency machine learning FPGA accelerator for multi-qubit-state discrimination4 Jul 2024 0 repositories listed
-
Timestep-Aware Correction for Quantized Diffusion Models4 Jul 2024 0 repositories listed
-
ADFQ-ViT: Activation-Distribution-Friendly Post-Training Quantization for Vision Transformers3 Jul 2024 0 repositories listed
-
Codec-ASR: Training Performant Automatic Speech Recognition Systems with Discrete Speech Representations3 Jul 2024 0 repositories listed
-
Edge AI-Enabled Chicken Health Detection Based on Enhanced FCOS-Lite and Knowledge Distillation3 Jul 2024 0 repositories listed
-
Fisher-aware Quantization for DETR Detectors with Critical-category Objectives3 Jul 2024 0 repositories listed
-
GPTQT: Quantize Large Language Models Twice to Push the Efficiency3 Jul 2024 0 repositories listed
-
How Does Quantization Affect Multilingual LLMs?3 Jul 2024 0 repositories listed
-
Improving Conversational Abilities of Quantized Large Language Models via Direct Preference Alignment3 Jul 2024 0 repositories listed
-
OSPC: Artificial VLM Features for Hateful Meme Detection3 Jul 2024 0 repositories listed
-
SFC: Achieve Accurate Fast Convolution under Low-precision Arithmetic3 Jul 2024 0 repositories listed
-
Unified Anomaly Detection methods on Edge Device using Knowledge Distillation and Quantization3 Jul 2024 0 repositories listed
-
Beyond Throughput and Compression Ratios: Towards High End-to-end Utility of Gradient Compression1 Jul 2024 0 repositories listed
-
Exploring FPGA designs for MX and beyond1 Jul 2024 0 repositories listed
-
Linear and Nonlinear MMSE Estimation in One-Bit Quantized Systems under a Gaussian Mixture Prior1 Jul 2024 0 repositories listed
-
PQCache: Product Quantization-based KVCache for Long Context LLM Inference1 Jul 2024 0 repositories listed
-
NeuroNAS: Enhancing Efficiency of Neuromorphic In-Memory Computing for Intelligent Mobile Agents through Hardware-Aware Spiking Neural Architecture Search30 Jun 2024 0 repositories listed
-
Toward a Diffusion-Based Generalist for Dense Vision Tasks29 Jun 2024 0 repositories listed
-
Deep Fusion Model for Brain Tumor Classification Using Fine-Grained Gradient Preservation28 Jun 2024 0 repositories listed
-
Rateless Stochastic Coding for Delay-Constrained Semantic Communication28 Jun 2024 0 repositories listed
-
Fronthaul Quantization-Aware MU-MIMO Precoding for Sum Rate Maximization27 Jun 2024 0 repositories listed
-
MCNC: Manifold Constrained Network Compression27 Jun 2024 0 repositories listed
-
OutlierTune: Efficient Channel-Wise Quantization for Large Language Models27 Jun 2024 0 repositories listed
-
Reliable edge machine learning hardware for scientific applications27 Jun 2024 0 repositories listed
-
A Quantization-based Technique for Privacy Preserving Distributed Learning26 Jun 2024 0 repositories listed
-
Differential error feedback for communication-efficient decentralized learning26 Jun 2024 0 repositories listed
-
FedAQ: Communication-Efficient Federated Edge Learning via Joint Uplink and Downlink Adaptive Quantization26 Jun 2024 0 repositories listed
-
CDQuant: Greedy Coordinate Descent for Accurate LLM Quantization25 Jun 2024 0 repositories listed
-
Approximate DCT and Quantization Techniques for Energy-Constrained Image Sensors24 Jun 2024 0 repositories listed
-
BitNet b1.58 Reloaded: State-of-the-art Performance Also on Smaller Networks24 Jun 2024 0 repositories listed
-
Compensate Quantization Errors: Make Weights Hierarchical to Compensate Each Other24 Jun 2024 0 repositories listed
-
Leveraging Knowledge Distillation for Lightweight Skin Cancer Classification: Balancing Accuracy and Computational Efficiency24 Jun 2024 0 repositories listed
-
Reducing the Memory Footprint of 3D Gaussian Splatting24 Jun 2024 0 repositories listed
-
Received Power Maximization Using Nonuniform Discrete Phase Shifts for RISs With a Limited Phase Range23 Jun 2024 0 repositories listed
-
Towards Real-Time Neural Volumetric Rendering on Mobile Devices: A Measurement Study23 Jun 2024 0 repositories listed
-
HLQ: Fast and Efficient Backpropagation via Hadamard Low-rank Quantization21 Jun 2024 0 repositories listed
-
Predicting Probabilities of Error to Combine Quantization and Early Exiting: QuEE20 Jun 2024 0 repositories listed
-
19 Jun 2024 0 repositories listed Syntology 10 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 9 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
High-Fidelity Facial Albedo Estimation via Texture Quantization19 Jun 2024 0 repositories listed
-
Q-SNNs: Quantized Spiking Neural Networks19 Jun 2024 0 repositories listed
-
SDQ: Sparse Decomposed Quantization for LLM Inference19 Jun 2024 0 repositories listed
-
Bayesian-LoRA: LoRA based Parameter Efficient Fine-Tuning using Optimal Quantization levels and Rank Values trough Differentiable Bayesian Gates18 Jun 2024 0 repositories listed
-
MSE Minimization in RIS-Aided MU-MIMO with Discrete Phase Shifts and Fronthaul Quantization18 Jun 2024 0 repositories listed
-
Deep-Learning-Based Channel Estimation for Distributed MIMO with 1-bit Radio-Over-Fiber Fronthaul17 Jun 2024 0 repositories listed
-
An Analysis on Quantizing Diffusion Transformers16 Jun 2024 0 repositories listed
-
Promoting Data and Model Privacy in Federated Learning through Quantized LoRA16 Jun 2024 0 repositories listed
-
Tender: Accelerating Large Language Models via Tensor Decomposition and Runtime Requantization16 Jun 2024 0 repositories listed
-
How Should We Extract Discrete Audio Tokens from Self-Supervised Models?15 Jun 2024 0 repositories listed
-
Memory Faults in Activation-sparse Quantized Deep Neural Networks: Analysis and Mitigation using Sharpness-aware Training15 Jun 2024 0 repositories listed
-
GEB-1.3B: Open Lightweight Large Language Model14 Jun 2024 0 repositories listed
-
One-pass Multiple Conformer and Foundation Speech Systems Compression and Quantization Using An All-in-one Neural Model14 Jun 2024 0 repositories listed
-
Optimizing Byte-level Representation for End-to-end ASR14 Jun 2024 0 repositories listed
-
Precipitation Nowcasting Using Physics Informed Discriminator Generative Models14 Jun 2024 0 repositories listed
-
Human-level molecular optimization driven by mol-gene evolution13 Jun 2024 0 repositories listed
-
ME-Switch: A Memory-Efficient Expert Switching Framework for Large Language Models13 Jun 2024 0 repositories listed
-
MGRQ: Post-Training Quantization For Vision Transformer With Mixed Granularity Reconstruction13 Jun 2024 0 repositories listed
-
ToneUnit: A Speech Discretization Approach for Tonal Language Speech Synthesis13 Jun 2024 0 repositories listed
-
Asymptotic Unbiased Sample Sampling to Speed Up Sharpness-Aware Minimization12 Jun 2024 0 repositories listed
-
Compressive Beam Alignment for Indoor Millimeter-Wave Systems12 Jun 2024 0 repositories listed
-
MobileAIBench: Benchmarking LLMs and LMMs for On-Device Use Cases12 Jun 2024 0 repositories listed
-
VALL-E R: Robust and Efficient Zero-Shot Text-to-Speech Synthesis via Monotonic Alignment12 Jun 2024 0 repositories listed
-
FoldToken2: Learning compact, invariant and generative protein structure language11 Jun 2024 0 repositories listed
-
T2S-GPT: Dynamic Vector Quantization for Autoregressive Sign Language Production from Text11 Jun 2024 0 repositories listed
-
TernaryLLM: Ternarized Large Language Model11 Jun 2024 0 repositories listed
-
Efficient Neural Compression with Inference-time Decoding10 Jun 2024 0 repositories listed
-
Latent Representation Matters: Human-like Sketches in One-shot Drawing Tasks10 Jun 2024 0 repositories listed
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs10 Jun 2024 0 repositories listed
-
Topological Analysis for Detecting Anomalies (TADA) in Time Series10 Jun 2024 0 repositories listed
-
Towards Lightweight Speaker Verification via Adaptive Neural Network Quantization8 Jun 2024 0 repositories listed
-
Activation Map-based Vector Quantization for 360-degree Image Semantic Communication7 Jun 2024 0 repositories listed
-
Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs7 Jun 2024 0 repositories listed
-
Proofread: Fixes All Errors with One Tap6 Jun 2024 0 repositories listed
-
USM RNN-T model weights binarization5 Jun 2024 0 repositories listed
-
VQUNet: Vector Quantization U-Net for Defending Adversarial Atacks by Regularizing Unwanted Noise5 Jun 2024 0 repositories listed
-
Zeroth-Order Fine-Tuning of LLMs with Extreme Sparsity5 Jun 2024 0 repositories listed
-
Mixed-Precision Federated Learning via Multi-Precision Over-The-Air Aggregation4 Jun 2024 0 repositories listed
-
Toward Efficient Deep Spiking Neuron Networks:A Survey On Compression3 Jun 2024 0 repositories listed
-
Log-Scale Quantization in Distributed First-Order Methods: Gradient-based Learning from Distributed Data2 Jun 2024 0 repositories listed
-
Effective Interplay between Sparsity and Quantization: From Theory to Practice31 May 2024 0 repositories listed
-
LCQ: Low-Rank Codebook based Quantization for Large Language Models31 May 2024 0 repositories listed
-
Locking Machine Learning Models into Hardware31 May 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.