Browse State-of-the-Art › Quantization › Papers, page 20
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 20 of 50: papers 1,901 to 2,000 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Verification of Bit-Flip Attacks against Quantized Neural Networks22 Feb 2025 0 repositories listed
-
Exact Recovery of Sparse Binary Vectors from Generalized Linear Measurements21 Feb 2025 0 repositories listed
-
FD-LSCIC: Frequency Decomposition-based Learned Screen Content Image Compression21 Feb 2025 0 repositories listed
-
Interleaved Block-based Learned Image Compression with Feature Enhancement and Quantization Error Compensation21 Feb 2025 0 repositories listed
-
Q-PETR: Quant-aware Position Embedding Transformation for Multi-View 3D Object Detection21 Feb 2025 0 repositories listed
-
SVDq: 1.25-bit and 410x Key Cache Compression for LLM Attention21 Feb 2025 0 repositories listed
-
When Compression Meets Model Compression: Memory-Efficient Double Compression for Large Language Models21 Feb 2025 0 repositories listed
-
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers20 Feb 2025 0 repositories listed
-
More for Keys, Less for Values: Adaptive KV Cache Quantization20 Feb 2025 0 repositories listed
-
A General Error-Theoretical Analysis Framework for Constructing Compression Strategies19 Feb 2025 0 repositories listed
-
A²ATS: Retrieval-Based KV Cache Reduction via Windowed Rotary Position Embedding and Query-Aware Vector Quantization18 Feb 2025 0 repositories listed
-
Continual Quantization-Aware Pre-Training: When to transition from 16-bit to 1.58-bit pre-training for BitNet language models?17 Feb 2025 0 repositories listed
-
On the Logic Elements Associated with Round-Off Errors and Gaussian Blur in Image Registration: A Simple Case of Commingling17 Feb 2025 0 repositories listed
-
Rotate, Clip, and Partition: Towards W2A4KV4 Quantization by Integrating Rotation and Learnable Non-uniform Quantizer17 Feb 2025 0 repositories listed
-
Towards Efficient Pre-training: Exploring FP4 Precision in Large Language Models17 Feb 2025 0 repositories listed
-
Towards Reasoning Ability of Small Language Models17 Feb 2025 0 repositories listed
-
EmbBERT-Q: Breaking Memory Barriers in Embedded NLP14 Feb 2025 0 repositories listed
-
Low-Complexity On-Grid Channel Estimation for Partially-Connected Hybrid XL-MIMO14 Feb 2025 0 repositories listed
-
Towards Watermarking of Open-Source LLMs14 Feb 2025 0 repositories listed
-
NestQuant: Nested Lattice Quantization for Matrix Products and LLMs13 Feb 2025 0 repositories listed
-
13 Feb 2025 0 repositories listed Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Compression of Site-Specific Deep Neural Networks for Massive MIMO Precoding12 Feb 2025 0 repositories listed
-
Contextual Compression Encoding for Large Language Models: A Novel Framework for Multi-Layered Parameter Space Pruning12 Feb 2025 0 repositories listed
-
Exploiting Non-uniform Quantization for Enhanced ILC in Wideband Digital Pre-distortion12 Feb 2025 0 repositories listed
-
LowRA: Accurate and Efficient LoRA Fine-Tuning of LLMs under 2 Bits12 Feb 2025 0 repositories listed
-
Scalable Thermodynamic Second-order Optimization12 Feb 2025 0 repositories listed
-
Conditional Distribution Quantization in Machine Learning11 Feb 2025 0 repositories listed
-
HDCompression: Hybrid-Diffusion Image Compression for Ultra-Low Bitrates11 Feb 2025 0 repositories listed
-
MEMHD: Memory-Efficient Multi-Centroid Hyperdimensional Computing for Fully-Utilized In-Memory Computing Architectures11 Feb 2025 0 repositories listed
-
Vision-Language Models for Edge Networks: A Comprehensive Survey11 Feb 2025 0 repositories listed
-
Demystifying Singular Defects in Large Language Models10 Feb 2025 0 repositories listed
-
Finetuning and Quantization of EEG-Based Foundational BioSignal Models on ECG and PPG Data for Blood Pressure Estimation10 Feb 2025 0 repositories listed
-
Matryoshka Quantization10 Feb 2025 0 repositories listed
-
Gradient Based Method for the Fusion of Lattice Quantizers9 Feb 2025 0 repositories listed
-
AIQViT: Architecture-Informed Post-Training Quantization for Vision Transformers7 Feb 2025 0 repositories listed
-
Efficient Evaluation of Quantization-Effects in Neural Codecs7 Feb 2025 0 repositories listed
-
QLIP: Text-Aligned Visual Tokenization Unifies Auto-Regressive Multimodal Understanding and Generation7 Feb 2025 0 repositories listed
-
Scalable and consistent embedding of probability measures into Hilbert spaces via measure quantization7 Feb 2025 0 repositories listed
-
A Performance Analysis of You Only Look Once Models for Deployment on Constrained Computational Edge Devices in Drone Applications6 Feb 2025 0 repositories listed
-
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization6 Feb 2025 0 repositories listed
-
TQ-DiT: Efficient Time-Aware Quantization for Diffusion Transformers6 Feb 2025 0 repositories listed
-
Asymptotic Analysis of One-bit Quantized Box-Constrained Precoding in Large-Scale Multi-User Systems5 Feb 2025 0 repositories listed
-
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference5 Feb 2025 0 repositories listed
-
SensorChat: Answering Qualitative and Quantitative Questions during Long-Term Multimodal Sensor Interactions5 Feb 2025 0 repositories listed
-
Survey of Quantization Techniques for On-Device Vision-based Crack Detection4 Feb 2025 0 repositories listed
-
Unlocking Efficient Large Inference Models: One-Bit Unrolling Tips the Scales4 Feb 2025 0 repositories listed
-
An Inquiry into Datacenter TCO for LLM Inference with FP83 Feb 2025 0 repositories listed
-
Choose Your Model Size: Any Compression by a Single Gradient Descent3 Feb 2025 0 repositories listed
-
Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis3 Feb 2025 0 repositories listed
-
Huff-LLM: End-to-End Lossless Compression for Efficient LLM Inference2 Feb 2025 0 repositories listed
-
On Noncommutative Quantum Mechanics and the Black-Scholes Model2 Feb 2025 0 repositories listed
-
Structural Latency Perturbation in Large Language Models Through Recursive State Induction2 Feb 2025 0 repositories listed
-
Enhancing Field-Oriented Control of Electric Drives with Tiny Neural Network Optimized for Micro-controllers1 Feb 2025 0 repositories listed
-
MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization1 Feb 2025 0 repositories listed
-
Fully Distributed and Quantized Algorithm for MPC-based Autonomous Vehicle Platooning Optimization31 Jan 2025 0 repositories listed
-
LLM-based Affective Text Generation Quality Based on Different Quantization Values31 Jan 2025 0 repositories listed
-
CodeBrain: Impute Any Brain MRI via Instance-specific Scalar-quantized Codes30 Jan 2025 0 repositories listed
-
Mixed-Precision Graph Neural Quantization for Low Bit Large Language Models30 Jan 2025 0 repositories listed
-
Distinguished Quantized Guidance for Diffusion-based Sequence Recommendation29 Jan 2025 0 repositories listed
-
EdgeMLOps: Operationalizing ML models with Cumulocity IoT and thin-edge.io for Visual quality Inspection28 Jan 2025 0 repositories listed
-
Optimizing Large Language Model Training Using FP4 Quantization28 Jan 2025 0 repositories listed
-
Post-Training Quantization for 3D Medical Image Segmentation: A Practical Study on Real Inference Engines28 Jan 2025 0 repositories listed
-
Post-Training Quantization for Vision Mamba with k-Scaled Quantization and Reparameterization28 Jan 2025 0 repositories listed
-
One-Bit Sigma-Delta DFRC Waveform Design: Using Quantization Noise for Radar Probing27 Jan 2025 0 repositories listed
-
Stabilization of an unstable reaction-diffusion PDE with input delay despite state and input quantization27 Jan 2025 0 repositories listed
-
Decentralized Low-Rank Fine-Tuning of Large Language Models26 Jan 2025 0 repositories listed
-
SQ-DM: Accelerating Diffusion Models with Aggressive Quantization and Temporal Sparsity26 Jan 2025 0 repositories listed
-
AKVQ-VL: Attention-Aware KV Cache Adaptive 2-Bit Quantization for Vision-Language Models25 Jan 2025 0 repositories listed
-
FBQuant: FeedBack Quantization for Large Language Models25 Jan 2025 0 repositories listed
-
On Accelerating Edge AI: Optimizing Resource-Constrained Environments25 Jan 2025 0 repositories listed
-
RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations25 Jan 2025 0 repositories listed
-
Channel-Aware Constellation Design for Digital OTA Computation24 Jan 2025 0 repositories listed
-
End-to-end workflow for machine learning-based qubit readout with QICK and hls4ml24 Jan 2025 0 repositories listed
-
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models24 Jan 2025 0 repositories listed
-
On Hardening DNNs against Noisy Computations24 Jan 2025 0 repositories listed
-
Diffusion-based Perceptual Neural Video Compression with Temporal Diffusion Information Reuse23 Jan 2025 0 repositories listed
-
DQ-Data2vec: Decoupling Quantization for Multilingual Speech Recognition23 Jan 2025 0 repositories listed
-
MambaQuant: Quantizing the Mamba Family with Variance Aligned Rotation Methods23 Jan 2025 0 repositories listed
-
QMamba: Post-Training Quantization for Vision State Space Models23 Jan 2025 0 repositories listed
-
Qrazor: Reliable and effortless 4-bit llm quantization by significant data razoring23 Jan 2025 0 repositories listed
-
HEPPO: Hardware-Efficient Proximal Policy Optimization -- A Universal Pipelined Architecture for Generalized Advantage Estimation22 Jan 2025 0 repositories listed
-
Irrational Complex Rotations Empower Low-bit Optimizers22 Jan 2025 0 repositories listed
-
Sketch and Patch: Efficient 3D Gaussian Representation for Man-Made Scenes22 Jan 2025 0 repositories listed
-
RL-RC-DoT: A Block-level RL agent for Task-Aware Video Compression21 Jan 2025 0 repositories listed
-
SplitQuant: Layer Splitting for Low-Bit Neural Network Quantization21 Jan 2025 0 repositories listed
-
UAV-Assisted Real-Time Disaster Detection Using Optimized Transformer Model21 Jan 2025 0 repositories listed
-
Communication-Efficient Federated Learning by Quantized Variance Reduction for Heterogeneous Wireless Edge Networks20 Jan 2025 0 repositories listed
-
Ditto: Accelerating Diffusion Model via Temporal Value Similarity20 Jan 2025 0 repositories listed
-
Personalized Federated Learning for Cellular VR: Online Learning and Dynamic Caching20 Jan 2025 0 repositories listed
-
Practical Modulo Sampling: Mitigating High-Frequency Components20 Jan 2025 0 repositories listed
-
BeST -- A Novel Source Selection Metric for Transfer Learning19 Jan 2025 0 repositories listed
-
DC-PCN: Point Cloud Completion Network with Dual-Codebook Guided Quantization19 Jan 2025 0 repositories listed
-
A Novel Hybrid Precoder With Low-Resolution Phase Shifters and Fronthaul Capacity Limitation18 Jan 2025 0 repositories listed
-
LUT-DLA: Lookup Table as Efficient Extreme Low-Bit Deep Learning Accelerator18 Jan 2025 0 repositories listed
-
Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures16 Jan 2025 0 repositories listed
-
The Devil is in the Details: Simple Remedies for Image-to-LiDAR Representation Learning16 Jan 2025 0 repositories listed
-
Real-time Indexing for Large-scale Recommendation by Streaming Vector Quantization Retriever15 Jan 2025 0 repositories listed
-
Rethinking Post-Training Quantization: Introducing a Statistical Pre-Calibration Approach15 Jan 2025 0 repositories listed
-
Large Language Models For Text Classification: Case Study And Comprehensive Review14 Jan 2025 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.