Browse State-of-the-Art › Quantization › Papers, page 26
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 26 of 50: papers 2,501 to 2,600 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Outliers and Calibration Sets have Diminishing Effect on Quantization of Modern LLMs31 May 2024 0 repositories listed
-
An Efficient Network with Novel Quantization Designed for Massive MIMO CSI Feedback30 May 2024 0 repositories listed
-
HQ-DiT: Efficient Diffusion Transformer with FP4 Hybrid Quantization30 May 2024 0 repositories listed
-
One QuantLLM for ALL: Fine-tuning Quantized LLMs Once for Efficient Deployments30 May 2024 0 repositories listed
-
S3D: A Simple and Cost-Effective Self-Speculative Decoding Scheme for Low-Memory GPUs30 May 2024 0 repositories listed
-
Information Entropy Guided Height-aware Histogram for Quantization-friendly Pillar Feature Encoder29 May 2024 0 repositories listed
-
I-LLM: Efficient Integer-Only Inference for Fully-Quantized Low-Bit Large Language Models28 May 2024 0 repositories listed
-
LLaMA-NAS: Efficient Neural Architecture Search for Large Language Models28 May 2024 0 repositories listed
-
MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization28 May 2024 0 repositories listed
-
The Binary Quantized Neural Network for Dense Prediction via Specially Designed Upsampling and Attention28 May 2024 0 repositories listed
-
BeamVQ: Aligning Space-Time Forecasting Model via Self-training on Physics-aware Metrics27 May 2024 0 repositories listed
-
CLAQ: Pushing the Limits of Low-Bit Post-Training Quantization for LLMs27 May 2024 0 repositories listed
-
Di²Pose: Discrete Diffusion Model for Occluded 3D Human Pose Estimation27 May 2024 0 repositories listed
-
UniCompress: Enhancing Multi-Data Medical Image Compression with Knowledge Distillation27 May 2024 0 repositories listed
-
FastQuery: Communication-efficient Embedding Table Query for Private LLM Inference25 May 2024 0 repositories listed
-
Athena: Efficient Block-Wise Post-Training Quantization for Large Language Models Using Second-Order Matrix Derivative Information24 May 2024 0 repositories listed
-
BiSup: Bidirectional Quantization Error Suppression for Large Language Models24 May 2024 0 repositories listed
-
Massive MIMO-ISAC System With 1-Bit ADCs/DACs24 May 2024 0 repositories listed
-
ASI++: Towards Distributionally Balanced End-to-End Generative Retrieval23 May 2024 0 repositories listed
-
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval23 May 2024 0 repositories listed
-
Embedding Compression for Efficient Re-Identification23 May 2024 0 repositories listed
-
Bracket Diffusion: HDR Image Generation by Consistent LDR Denoising23 May 2024 0 repositories listed
-
Integer Scale: A Free Lunch for Faster Fine-grained Quantization of LLMs23 May 2024 0 repositories listed
-
LG-VQ: Language-Guided Codebook Learning23 May 2024 0 repositories listed
-
MiniCache: KV Cache Compression in Depth Dimension for Large Language Models23 May 2024 0 repositories listed
-
MultiCast: Zero-Shot Multivariate Time Series Forecasting Using LLMs23 May 2024 0 repositories listed
-
OAC: Output-adaptive Calibration for Accurate Post-training Quantization23 May 2024 0 repositories listed
-
A rescaling-invariant Lipschitz bound based on path-metrics for modern ReLU network parameterizations23 May 2024 0 repositories listed
-
Adaptive Wireless Image Semantic Transmission and Over-The-Air Testing22 May 2024 0 repositories listed
-
AdpQ: A Zero-shot Calibration Free Adaptive Post Training Quantization Method for LLMs22 May 2024 0 repositories listed
-
Discrete Cosine Transform Based Decorrelated Attention for Vision Transformers22 May 2024 0 repositories listed
-
eXmY: A Data Type and Technique for Arbitrary Bit Precision Quantization22 May 2024 0 repositories listed
-
QGait: Toward Accurate Quantization for Gait Recognition with Binarized Input22 May 2024 0 repositories listed
-
Two Heads are Better Than One: Neural Networks Quantization with 2D Hilbert Curve-based Output Representation22 May 2024 0 repositories listed
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities21 May 2024 0 repositories listed
-
On Image Registration and Subpixel Estimation21 May 2024 0 repositories listed
-
Online Signature Recognition: A Biologically Inspired Feature Vector Splitting Approach21 May 2024 0 repositories listed
-
ReALLM: A general framework for LLM compression and fine-tuning21 May 2024 0 repositories listed
-
TinyM²Net-V3: Memory-Aware Compressed Multimodal Deep Neural Networks for Sustainable Edge Deployment20 May 2024 0 repositories listed
-
Enhancing Perception Quality in Remote Sensing Image Compression via Invertible Neural Network17 May 2024 0 repositories listed
-
Flattened one-bit stochastic gradient descent: compressed distributed optimization with controlled variance17 May 2024 0 repositories listed
-
Universal Joint Source-Channel Coding for Modulation-Agnostic Semantic Communication17 May 2024 0 repositories listed
-
The Effect of Quantization in Federated Learning: A Rényi Differential Privacy Perspective16 May 2024 0 repositories listed
-
FDD Massive MIMO: How to Optimally Combine UL Pilot and Limited DL CSI Feedback?14 May 2024 0 repositories listed
-
Neural Speech Coding for Real-time Communications using Constant Bitrate Scalar Quantization14 May 2024 0 repositories listed
-
Goal-oriented compression for Lₚ-norm-type goal functions: Application to power consumption scheduling13 May 2024 0 repositories listed
-
VQDNA: Unleashing the Power of Vector Quantization for Multi-Species Genomic Sequence Modeling13 May 2024 0 repositories listed
-
Post Training Quantization of Large Language Models with Microscaling Formats12 May 2024 0 repositories listed
-
Edge Intelligence Optimization for Large Language Model Inference with Batching and Quantization12 May 2024 0 repositories listed
-
Characterizing the Accuracy -- Efficiency Trade-off of Low-rank Decomposition in Language Models10 May 2024 0 repositories listed
-
Compression-Realized Deep Structural Network for Video Quality Enhancement10 May 2024 0 repositories listed
-
10 May 2024 0 repositories listed Syntology 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 8 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models10 May 2024 0 repositories listed
-
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks9 May 2024 0 repositories listed
-
Custom Gradient Estimators are Straight-Through Estimators in Disguise8 May 2024 0 repositories listed
-
KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization7 May 2024 0 repositories listed
-
DeltaKWS: A 65nm 36nJ/Decision Bio-inspired Temporal-Sparsity-Aware Digital Keyword Spotting IC with 0.6V Near-Threshold SRAM6 May 2024 0 repositories listed
-
Compression-based Privacy Preservation for Distributed Nash Equilibrium Seeking in Aggregative Games6 May 2024 0 repositories listed
-
Enabling High-Sparsity Foundational Llama Models with Efficient Pretraining and Deployment6 May 2024 0 repositories listed
-
Quantifying the Capabilities of LLMs across Scale and Precision6 May 2024 0 repositories listed
-
Joint Discrete Precoding and RIS Optimization for RIS-Assisted MU-MIMO Communication Systems5 May 2024 0 repositories listed
-
Exploring Extreme Quantization in Spiking Language Models4 May 2024 0 repositories listed
-
Lightweight Change Detection in Heterogeneous Remote Sensing Images with Online All-Integer Pruning Training3 May 2024 0 repositories listed
-
Three Quantization Regimes for ReLU Networks3 May 2024 0 repositories listed
-
Deep Learning Models in Speech Recognition: Measuring GPU Energy Consumption, Impact of Noise and Model Quantization for Edge Deployment2 May 2024 0 repositories listed
-
Efficient Compression of Multitask Multilingual Speech Models2 May 2024 0 repositories listed
-
Joint Sequential Fronthaul Quantization and Hardware Complexity Reduction in Uplink Cell-Free Massive MIMO Networks2 May 2024 0 repositories listed
-
Investigating Automatic Scoring and Feedback using Large Language Models1 May 2024 0 repositories listed
-
Wake Vision: A Tailored Dataset and Benchmark Suite for TinyML Computer Vision Applications1 May 2024 0 repositories listed
-
Transition Rate Scheduling for Quantization-Aware Training30 Apr 2024 0 repositories listed
-
Quantized Context Based LIF Neurons for Recurrent Spiking Neural Networks in 45nm28 Apr 2024 0 repositories listed
-
Enhancing Channel Estimation in Quantized Systems with a Generative Prior26 Apr 2024 0 repositories listed
-
sDAC -- Semantic Digital Analog Converter for Semantic Communications26 Apr 2024 0 repositories listed
-
How to Parameterize Asymmetric Quantization Ranges for Quantization-Aware Training25 Apr 2024 0 repositories listed
-
MMGRec: Multimodal Generative Recommendation with Transformer Model25 Apr 2024 0 repositories listed
-
CoST: Contrastive Quantization based Semantic Tokenization for Generative Recommendation23 Apr 2024 0 repositories listed
-
AdaQAT: Adaptive Bit-Width Quantization-Aware Training22 Apr 2024 0 repositories listed
-
CNN-Based Equalization for Communications: Achieving Gigabit Throughput with a Flexible FPGA Hardware Architecture22 Apr 2024 0 repositories listed
-
Latency-Distortion Tradeoffs in Communicating Classification Results over Noisy Channels22 Apr 2024 0 repositories listed
-
FedMPQ: Secure and Communication-Efficient Federated Learning with Multi-codebook Product Quantization21 Apr 2024 0 repositories listed
-
A SER-based Device Selection Mechanism in Multi-bits Quantization Federated Learning20 Apr 2024 0 repositories listed
-
HybridFlow: Infusing Continuity into Masked Codebook for Extreme Low-Bitrate Image Compression20 Apr 2024 0 repositories listed
-
EdgeFusion: On-Device Text-to-Image Generation18 Apr 2024 0 repositories listed
-
Privacy-Preserving UCB Decision Process Verification via zk-SNARKs18 Apr 2024 0 repositories listed
-
LongVQ: Long Sequence Modeling with Vector Quantization on Structured Memory17 Apr 2024 0 repositories listed
-
Neural Network Approach for Non-Markovian Dissipative Dynamics of Many-Body Open Quantum Systems17 Apr 2024 0 repositories listed
-
QGen: On the Ability to Generalize in Quantization Aware Training17 Apr 2024 0 repositories listed
-
Comprehensive Survey of Model Compression and Speed up for Vision Transformers16 Apr 2024 0 repositories listed
-
Efficient and accurate neural field reconstruction using resistive memory15 Apr 2024 0 repositories listed
-
Quantization of Large Language Models with an Overdetermined Basis15 Apr 2024 0 repositories listed
-
TMPQ-DM: Joint Timestep Reduction and Quantization Precision Selection for Efficient Diffusion Models15 Apr 2024 0 repositories listed
-
Bullion: A Column Store for Machine Learning13 Apr 2024 0 repositories listed
-
Full-Duplex Beyond Self-Interference: The Unlimited Sensing Way12 Apr 2024 0 repositories listed
-
Lossy Image Compression with Foundation Diffusion Models12 Apr 2024 0 repositories listed
-
1-bit Quantized On-chip Hybrid Diffraction Neural Network Enabled by Authentic All-optical Fully-connected Architecture11 Apr 2024 0 repositories listed
-
Edge-Efficient Deep Learning Models for Automatic Modulation Classification: A Performance Analysis11 Apr 2024 0 repositories listed
-
Frame Quantization of Neural Networks11 Apr 2024 0 repositories listed
-
Differentiable Search for Finding Optimal Quantization Strategy10 Apr 2024 0 repositories listed
-
Collaborative Edge AI Inference over Cloud-RAN9 Apr 2024 0 repositories listed
-
Encoder-Quantization-Motion-based Video Quality Metrics9 Apr 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.