Browse State-of-the-Art › Quantization › Papers, page 27
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 27 of 50: papers 2,601 to 2,700 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Investigating the Impact of Quantization on Adversarial Robustness8 Apr 2024 0 repositories listed
-
Physics of Language Models: Part 3.3, Knowledge Capacity Scaling Laws8 Apr 2024 0 repositories listed
-
Gull: A Generative Multifunctional Audio Codec7 Apr 2024 0 repositories listed
-
Nanometer Scanning with Micrometer Sensing: Beating Quantization Constraints in Lissajous Trajectory Tracking7 Apr 2024 0 repositories listed
-
What Happens When Small Is Made Smaller? Exploring the Impact of Compression on Small Data Pretrained Language Models6 Apr 2024 0 repositories listed
-
Fine-Tuning, Quantization, and LLMs: Navigating Unintended Outcomes5 Apr 2024 0 repositories listed
-
DI-Retinex: Digital-Imaging Retinex Theory for Low-Light Image Enhancement4 Apr 2024 0 repositories listed
-
TinyVQA: Compact Multimodal Deep Neural Network for Visual Question Answering on Resource-Constrained Devices4 Apr 2024 0 repositories listed
-
Cherry on Top: Parameter Heterogeneity and Quantization in Large Language Models3 Apr 2024 0 repositories listed
-
CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech3 Apr 2024 0 repositories listed
-
DNN Memory Footprint Reduction via Post-Training Intra-Layer Multi-Precision Quantization3 Apr 2024 0 repositories listed
-
NeRFCodec: Neural Feature Compression Meets Neural Radiance Fields for Memory-Efficient Scene Representation2 Apr 2024 0 repositories listed
-
On the Effect of Quantization on Dynamic Mode Decomposition2 Apr 2024 0 repositories listed
-
RefQSR: Reference-based Quantization for Image Super-Resolution Networks2 Apr 2024 0 repositories listed
-
A Novel Audio Representation for Music Genre Identification in MIR1 Apr 2024 0 repositories listed
-
Instance-Aware Group Quantization for Vision Transformers1 Apr 2024 0 repositories listed
-
Towards Variable and Coordinated Holistic Co-Speech Motion Generation30 Mar 2024 0 repositories listed
-
Accurate Block Quantization in LLMs with Outliers29 Mar 2024 0 repositories listed
-
Transformer-Lite: High-efficiency Deployment of Large Language Models on Mobile Phone GPUs29 Mar 2024 0 repositories listed
-
Meta-Heuristic Fronthaul Bit Allocation for Cell-free Massive MIMO Systems28 Mar 2024 0 repositories listed
-
Uncertainty-Aware Deep Video Compression with Ensembles28 Mar 2024 0 repositories listed
-
Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence28 Mar 2024 0 repositories listed
-
Order of Compression: A Systematic and Optimal Sequence to Combinationally Compress CNN26 Mar 2024 0 repositories listed
-
Oh! We Freeze: Improving Quantized Knowledge Distillation via Signal Propagation Analysis for Large Language Models26 Mar 2024 0 repositories listed
-
25 Mar 2024 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Neural Image Compression with Quantization Rectifier25 Mar 2024 0 repositories listed
-
25 Mar 2024 0 repositories listed
-
Infrastructure-Assisted Collaborative Perception in Automated Valet Parking: A Safety Perspective22 Mar 2024 0 repositories listed
-
Magic for the Age of Quantized DNNs22 Mar 2024 0 repositories listed
-
Provable Privacy with Non-Private Pre-Processing19 Mar 2024 0 repositories listed
-
Super-High-Fidelity Image Compression via Hierarchical-ROI and Adaptive Quantization19 Mar 2024 0 repositories listed
-
Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression18 Mar 2024 0 repositories listed
-
Hierarchical Frequency-based Upsampling and Refining for Compressed Video Quality Enhancement18 Mar 2024 0 repositories listed
-
HyperVQ: MLR-based Vector Quantization in Hyperbolic Space18 Mar 2024 0 repositories listed
-
Spatio-Temporal Fluid Dynamics Modeling via Physical-Awareness and Parameter Diffusion Guidance18 Mar 2024 0 repositories listed
-
Quantization Avoids Saddle Points in Distributed Optimization15 Mar 2024 0 repositories listed
-
BRIEDGE: EEG-Adaptive Edge AI for Multi-Brain to Multi-Robot Interaction14 Mar 2024 0 repositories listed
-
CRB Analysis for Mixed-ADC Based DOA Estimation14 Mar 2024 0 repositories listed
-
FedComLoc: Communication-Efficient Distributed Training of Sparse and Quantized Models14 Mar 2024 0 repositories listed
-
UniCode: Learning a Unified Codebook for Multimodal Large Language Models14 Mar 2024 0 repositories listed
-
Collaborative Automotive Radar Sensing via Mixed-Precision Distributed Array Completion13 Mar 2024 0 repositories listed
-
Strategizing against Q-learners: A Control-theoretical Approach13 Mar 2024 0 repositories listed
-
Approaching Rate-Distortion Limits in Neural Compression with Lattice Transform Coding12 Mar 2024 0 repositories listed
-
Vector Quantization for Deep-Learning-Based CSI Feedback in Massive MIMO Systems12 Mar 2024 0 repositories listed
-
FlowVQTalker: High-Quality Emotional Talking Face Generation through Normalizing Flow and Quantization11 Mar 2024 0 repositories listed
-
QuantTune: Optimizing Model Quantization with Adaptive Outlier-Driven Fine Tuning11 Mar 2024 0 repositories listed
-
What Makes Quantization for Large Language Models Hard? An Empirical Study from the Lens of Perturbation11 Mar 2024 0 repositories listed
-
Micro-Fracture Detection in Photovoltaic Cells with Hardware-Constrained Devices and Computer Vision8 Mar 2024 0 repositories listed
-
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers8 Mar 2024 0 repositories listed
-
Enhancing Multimodal Unified Representations for Cross Modal Generalization8 Mar 2024 0 repositories listed
-
LoCoDL: Communication-Efficient Distributed Learning with Local Training and Compression7 Mar 2024 0 repositories listed
-
On-demand Quantization for Green Federated Generative Diffusion in Mobile Edge Networks7 Mar 2024 0 repositories listed
-
Adaptive Integrate-and-Fire Time Encoding Machine with Quantization5 Mar 2024 0 repositories listed
-
Deep-Learned Compression for Radio-Frequency Signal Classification5 Mar 2024 0 repositories listed
-
Design of Stochastic Quantizers for Privacy Preservation5 Mar 2024 0 repositories listed
-
EasyQuant: An Efficient Data-free Quantization Algorithm for LLMs5 Mar 2024 0 repositories listed
-
VQSynery: Robust Drug Synergy Prediction With Vector Quantization Mechanism5 Mar 2024 0 repositories listed
-
Better Schedules for Low Precision Training of Deep Neural Networks4 Mar 2024 0 repositories listed
-
FlowPrecision: Advancing FPGA-Based Real-Time Fluid Flow Estimation with Linear Quantization4 Mar 2024 0 repositories listed
-
Towards efficient deep autoencoders for multivariate time series anomaly detection4 Mar 2024 0 repositories listed
-
On the Compressibility of Quantized Large Language Models3 Mar 2024 0 repositories listed
-
A Hierarchical Federated Learning Approach for the Internet of Things3 Mar 2024 0 repositories listed
-
BasedAI: A decentralized P2P network for Zero Knowledge Large Language Models (ZK-LLMs)1 Mar 2024 0 repositories listed
-
T3DNet: Compressing Point Cloud Models for Lightweight 3D Recognition29 Feb 2024 0 repositories listed
-
Ef-QuantFace: Streamlined Face Recognition with Small Data and Low-Bit Precision28 Feb 2024 0 repositories listed
-
FlattenQuant: Breaking Through the Inference Compute-bound for Large Language Models with Per-tensor Quantization28 Feb 2024 0 repositories listed
-
No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization28 Feb 2024 0 repositories listed
-
Adaptive quantization with mixed-precision based on low-cost proxy27 Feb 2024 0 repositories listed
-
Inpainting Computational Fluid Dynamics with Deep Learning27 Feb 2024 0 repositories listed
-
Rethinking Mutual Information for Language Conditioned Skill Discovery on Imitation Learning27 Feb 2024 0 repositories listed
-
Data-freeWeight Compress and Denoise for Large Language Models26 Feb 2024 0 repositories listed
-
Distortion-Controlled Dithering with Reduced Recompression Rate26 Feb 2024 0 repositories listed
-
SPC-NeRF: Spatial Predictive Compression for Voxel Based Radiance Field26 Feb 2024 0 repositories listed
-
GPTVQ: The Blessing of Dimensionality for LLM Quantization23 Feb 2024 0 repositories listed
-
On the Arrow of Inference22 Feb 2024 0 repositories listed
-
Text me the data: Generating Ground Pressure Sequence from Textual Descriptions for HAR22 Feb 2024 0 repositories listed
-
APTQ: Attention-aware Post-Training Mixed-Precision Quantization for Large Language Models21 Feb 2024 0 repositories listed
-
FinGPT-HPC: Efficient Pretraining and Finetuning Large Language Models for Financial Applications with High-Performance Computing21 Feb 2024 0 repositories listed
-
In-Distribution Consistency Regularization Improves the Generalization of Quantization-Aware Training21 Feb 2024 0 repositories listed
-
DB-LLM: Accurate Dual-Binarization for Efficient LLMs19 Feb 2024 0 repositories listed
-
Is It a Free Lunch for Removing Outliers during Pretraining?19 Feb 2024 0 repositories listed
-
Towards a tailored mixed-precision sub-8-bit quantization scheme for Gated Recurrent Units using Genetic Algorithms19 Feb 2024 0 repositories listed
-
WKVQuant: Quantizing Weight and Key/Value Cache for Large Language Models Gains More19 Feb 2024 0 repositories listed
-
One-Bit Quantization and Sparsification for Multiclass Linear Classification with Strong Regularization16 Feb 2024 0 repositories listed
-
QDyLoRA: Quantized Dynamic Low-Rank Adaptation for Efficient Large Language Model Tuning16 Feb 2024 0 repositories listed
-
Model Compression and Efficient Inference for Large Language Models: A Survey15 Feb 2024 0 repositories listed
-
Quantized Embedding Vectors for Controllable Diffusion Language Models15 Feb 2024 0 repositories listed
-
Rate-Splitting Multiple Access for Quantized ISAC LEO Satellite Systems: A Max-Min Fair Energy-Efficient Beam Design14 Feb 2024 0 repositories listed
-
14 Feb 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TeMPO: Efficient Time-Multiplexed Dynamic Photonic Tensor Core for Edge AI with Compact Slow-Light Electro-Optic Modulator12 Feb 2024 0 repositories listed
-
Outlier-Aware Training for Low-Bit Quantization of Structural Re-Parameterized Networks11 Feb 2024 0 repositories listed
-
LiRank: Industrial Large Scale Ranking Models at LinkedIn10 Feb 2024 0 repositories listed
-
On Leaky-Integrate-and Fire as Spike-Train-Quantization Operator on Dirac-Superimposed Continuous-Time Signals10 Feb 2024 0 repositories listed
-
RQP-SGD: Differential Private Machine Learning through Noisy SGD and Randomized Quantization9 Feb 2024 0 repositories listed
-
RepQuant: Towards Accurate Post-Training Quantization of Large Transformer Models via Scale Reparameterization8 Feb 2024 0 repositories listed
-
Sparse-VQ Transformer: An FFN-Free Framework with Vector Quantization for Enhanced Time Series Forecasting8 Feb 2024 0 repositories listed
-
L4Q: Parameter Efficient Quantization-Aware Fine-Tuning on Large Language Models7 Feb 2024 0 repositories listed
-
Majority Kernels: An Approach to Leverage Big Model Dynamics for Efficient Small Model Training7 Feb 2024 0 repositories listed
-
Fed-CVLC: Compressing Federated Learning Communications with Variable-Length Codes6 Feb 2024 0 repositories listed
-
A Survey on Transformer Compression5 Feb 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.