Browse State-of-the-Art › Quantization › Papers, page 19
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 19 of 50: papers 1,801 to 1,900 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Bridging the Gap between Gaussian Diffusion Models and Universal Quantization for Image Compression3 Apr 2025 0 repositories listed
-
HPGN: Hybrid Priors-Guided Network for Compressed Low-Light Image Enhancement3 Apr 2025 0 repositories listed
-
Moment Quantization for Video Temporal Grounding3 Apr 2025 0 repositories listed
-
LLMPi: Optimizing LLMs for High-Throughput on Raspberry Pi2 Apr 2025 0 repositories listed
-
When Reasoning Meets Compression: Benchmarking Compressed Large Reasoning Models on Complex Reasoning Tasks2 Apr 2025 0 repositories listed
-
QSViT: A Methodology for Quantizing Spiking Vision Transformers1 Apr 2025 0 repositories listed
-
Model Hemorrhage and the Robustness Limits of Large Language Models31 Mar 2025 0 repositories listed
-
SQuat: Subspace-orthogonal KV Cache Quantization31 Mar 2025 0 repositories listed
-
Style Quantization for Data-Efficient GAN Training31 Mar 2025 0 repositories listed
-
Cocktail: Chunk-Adaptive Mixed-Precision Quantization for Long-Context LLM Inference30 Mar 2025 0 repositories listed
-
NeuralGS: Bridging Neural Fields and 3D Gaussian Splatting for Compact 3D Representations29 Mar 2025 0 repositories listed
-
Long-Tail Crisis in Nearest Neighbor Language Models28 Mar 2025 0 repositories listed
-
Make Some Noise: Towards LLM audio reasoning and generation using sound tokens28 Mar 2025 0 repositories listed
-
MCRB for Parameter Estimation from One-Bit Quantized and Oversampled Measurements28 Mar 2025 0 repositories listed
-
A 71.2-μW Speech Recognition Accelerator with Recurrent Spiking Neural Network27 Mar 2025 0 repositories listed
-
MoQa: Rethinking MoE Quantization with Multi-stage Data-model Distribution Awareness27 Mar 2025 0 repositories listed
-
Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration27 Mar 2025 0 repositories listed
-
MAR-3D: Progressive Masked Auto-regressor for High-Resolution 3D Generation26 Mar 2025 0 repositories listed
-
SINR: Sparsity Driven Compressed Implicit Neural Representations25 Mar 2025 0 repositories listed
-
4DGC: Rate-Aware 4D Gaussian Compression for Efficient Streamable Free-Viewpoint Video24 Mar 2025 0 repositories listed
-
FFN Fusion: Rethinking Sequential Computation in Large Language Models24 Mar 2025 0 repositories listed
-
GranQ: Granular Zero-Shot Quantization with Channel-Wise Activation Scaling in QAT24 Mar 2025 0 repositories listed
-
Oaken: Fast and Efficient LLM Serving with Online-Offline Hybrid KV Cache Quantization24 Mar 2025 0 repositories listed
-
QSID-MPC: Model Predictive Control with System Identification from Quantized Data24 Mar 2025 0 repositories listed
-
Energy-Aware LLMs: A step towards sustainable AI for downstream applications22 Mar 2025 0 repositories listed
-
Improving Quantization with Post-Training Model Expansion21 Mar 2025 0 repositories listed
-
Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation20 Mar 2025 0 repositories listed
-
Improving Autoregressive Image Generation through Coarse-to-Fine Token Prediction20 Mar 2025 0 repositories listed
-
LeanTTA: A Backpropagation-Free and Stateless Approach to Quantized Test-Time Adaptation on Edge Devices20 Mar 2025 0 repositories listed
-
Learning Linear Block Codes with Gradient Quantization20 Mar 2025 0 repositories listed
-
Neural Networks: According to the Principles of Grassmann Algebra20 Mar 2025 0 repositories listed
-
Plug-and-Play 1.x-Bit KV Cache Quantization for Video Large Language Models20 Mar 2025 0 repositories listed
-
SpeCache: Speculative Key-Value Caching for Efficient Generation of LLMs20 Mar 2025 0 repositories listed
-
PARQ: Piecewise-Affine Regularized Quantization19 Mar 2025 0 repositories listed
-
RAG-based User Profiling for Precision Planning in Mixed-precision Over-the-Air Federated Learning19 Mar 2025 0 repositories listed
-
MAG: Multi-Modal Aligned Autoregressive Co-Speech Gesture Generation without Vector Quantization18 Mar 2025 0 repositories listed
-
Robust Machine Unlearning for Quantized Neural Networks via Adaptive Gradient Reweighting with Similar Labels18 Mar 2025 0 repositories listed
-
ACT360: An Efficient 360-Degree Action Detection and Summarization Framework for Mission-Critical Training and Debriefing17 Mar 2025 0 repositories listed
-
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning17 Mar 2025 0 repositories listed
-
CompMarkGS: Robust Watermarking for Compressed 3D Gaussian Splatting17 Mar 2025 0 repositories listed
-
ML-SpecQD: Multi-Level Speculative Decoding with Quantized Drafts17 Mar 2025 0 repositories listed
-
Versatile Physics-based Character Control with Hybrid Latent Representation17 Mar 2025 0 repositories listed
-
Pathology Image Compression with Pre-trained Autoencoders14 Mar 2025 0 repositories listed
-
Stabilizing Quantization-Aware Training by Implicit-Regularization on Hessian Matrix14 Mar 2025 0 repositories listed
-
Understanding Flatness in Generative Models: Its Role and Benefits14 Mar 2025 0 repositories listed
-
Automated Tomato Maturity Estimation Using an Optimized Residual Model with Pruning and Quantization Techniques13 Mar 2025 0 repositories listed
-
Dual Codebook VQ: Enhanced Image Reconstruction with Reduced Codebook Size13 Mar 2025 0 repositories listed
-
Global synchronization of multi-agent systems with nonlinear interactions13 Mar 2025 0 repositories listed
-
OuroMamba: A Data-Free Quantization Framework for Vision Mamba Models13 Mar 2025 0 repositories listed
-
Quantitative Analysis of Deeply Quantized Tiny Neural Networks Robust to Adversarial Attacks12 Mar 2025 0 repositories listed
-
Sometimes Painful but Certainly Promising: Feasibility and Trade-offs of Language Model Inference at the Edge12 Mar 2025 0 repositories listed
-
ViM-VQ: Efficient Post-Training Vector Quantization for Visual Mamba12 Mar 2025 0 repositories listed
-
Accurate INT8 Training Through Dynamic Block-Level Fallback11 Mar 2025 0 repositories listed
-
Quantization Design for Deep Learning-Based CSI Feedback11 Mar 2025 0 repositories listed
-
Breaking the Limits of Quantization-Aware Defenses: QADT-R for Robustness Against Patch-Based Adversarial Attacks in QNNs10 Mar 2025 0 repositories listed
-
Lightweight Multimodal Artificial Intelligence Framework for Maritime Multi-Scene Recognition10 Mar 2025 0 repositories listed
-
Non-vacuous Generalization Bounds for Deep Neural Networks without any modification to the trained models10 Mar 2025 0 repositories listed
-
Post-Training Quantization for Diffusion Transformer via Hierarchical Timestep Grouping10 Mar 2025 0 repositories listed
-
Synchronized Video-to-Audio Generation via Mel Quantization-Continuum Decomposition10 Mar 2025 0 repositories listed
-
VocalEyes: Enhancing Environmental Perception for the Visually Impaired through Vision-Language Models and Distance-Aware Object Detection10 Mar 2025 0 repositories listed
-
PathVQ: Reforming Computational Pathology Foundation Model for Whole Slide Image Analysis via Vector Quantization9 Mar 2025 0 repositories listed
-
SAQ-SAM: Semantically-Aligned Quantization for Segment Anything Model9 Mar 2025 0 repositories listed
-
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform9 Mar 2025 0 repositories listed
-
Does Acceleration Cause Hidden Instability in Vision Language Models? Uncovering Instance-Level Divergence Through a Large-Scale Empirical Study9 Mar 2025 0 repositories listed
-
Towards Superior Quantization Accuracy: A Layer-sensitive Approach9 Mar 2025 0 repositories listed
-
TR-DQ: Time-Rotation Diffusion Quantization9 Mar 2025 0 repositories listed
-
Discrete Contrastive Learning for Diffusion Policies in Autonomous Driving7 Mar 2025 0 repositories listed
-
Frequency Autoregressive Image Generation with Continuous Tokens7 Mar 2025 0 repositories listed
-
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression6 Mar 2025 0 repositories listed
-
Universality of Layer-Level Entropy-Weighted Quantization Beyond Model Architecture and Size6 Mar 2025 0 repositories listed
-
VQEL: Enabling Self-Developed Symbolic Language in Agents through Vector Quantization in Emergent Language Games6 Mar 2025 0 repositories listed
-
AHCPTQ: Accurate and Hardware-Compatible Post-Training Quantization for Segment Anything Model5 Mar 2025 0 repositories listed
-
English K_Quantization of LLMs Does Not Disproportionately Diminish Multilingual Performance5 Mar 2025 0 repositories listed
-
Fast Jet Tagging with MLP-Mixers on FPGAs5 Mar 2025 0 repositories listed
-
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs5 Mar 2025 0 repositories listed
-
Lightweight Embedded FPGA Deployment of Learned Image Compression with Knowledge Distillation and Hybrid Quantization5 Mar 2025 0 repositories listed
-
BdSLW401: Transformer-Based Word-Level Bangla Sign Language Recognition Using Relative Quantization Encoding (RQE)4 Mar 2025 0 repositories listed
-
Q&C: When Quantization Meets Cache in Efficient Image Generation4 Mar 2025 0 repositories listed
-
Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense Representations4 Mar 2025 0 repositories listed
-
DeRS: Towards Extremely Efficient Upcycled Mixture-of-Experts Models3 Mar 2025 0 repositories listed
-
DILEMMA: Joint LLM Quantization and Distributed LLM Inference Over Edge Computing Systems3 Mar 2025 0 repositories listed
-
KurTail : Kurtosis-based LLM Quantization3 Mar 2025 0 repositories listed
-
Regularization-based Framework for Quantization-, Fault- and Variability-Aware Training3 Mar 2025 0 repositories listed
-
Towards Improved Text-Aligned Codebook Learning: Multi-Hierarchical Codebook-Text Alignment with Long Text3 Mar 2025 0 repositories listed
-
MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations2 Mar 2025 0 repositories listed
-
Strong Solutions and Quantization-Based Numerical Schemes for a Class of Non-Markovian Volatility Models28 Feb 2025 0 repositories listed
-
Beyond the Tip of Efficiency: Uncovering the Submerged Threats of Jailbreak Attacks in Small Language Models27 Feb 2025 0 repositories listed
-
HALO: Hardware-aware quantization with low critical-path-delay weights for LLM acceleration27 Feb 2025 0 repositories listed
-
Speculative Decoding and Beyond: An In-Depth Review of Techniques27 Feb 2025 0 repositories listed
-
Transformer-Based Nonlinear Transform Coding for Multi-Rate CSI Compression in MIMO-OFDM Systems27 Feb 2025 0 repositories listed
-
Compressing Language Models for Specialized Domains25 Feb 2025 0 repositories listed
-
Memory-Free and Parallel Computation for Quantized Spiking Neural Networks25 Feb 2025 0 repositories listed
-
On the Privacy-Preserving Properties of Spiking Neural Networks with Unique Surrogate Gradients and Quantization Levels25 Feb 2025 0 repositories listed
-
Task-Driven Semantic Quantization and Imitation Learning for Goal-Oriented Communications25 Feb 2025 0 repositories listed
-
Unbiased and Sign Compression in Distributed Learning: Comparing Noise Resilience via SDEs24 Feb 2025 0 repositories listed
-
Compression Scaling Laws:Unifying Sparsity and Quantization23 Feb 2025 0 repositories listed
-
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration23 Feb 2025 0 repositories listed
-
Energy-Efficient Transformer Inference: Optimization Strategies for Time Series Classification23 Feb 2025 0 repositories listed
-
A 2-bit Wideband 5G mm-Wave RIS with Low Side Lobe Levels and no Quantization Lobe22 Feb 2025 0 repositories listed
-
Speech Enhancement Using Continuous Embeddings of Neural Audio Codec22 Feb 2025 0 repositories listed