Browse State-of-the-Art › Quantization › Papers, page 23
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 23 of 50: papers 2,201 to 2,300 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Can General-Purpose Large Language Models Generalize to English-Thai Machine Translation ?22 Oct 2024 0 repositories listed
-
Pyramid Vector Quantization for LLMs22 Oct 2024 0 repositories listed
-
Self-calibration for Language Model Quantization and Pruning22 Oct 2024 0 repositories listed
-
Continuous Speech Synthesis using per-token Latent Diffusion21 Oct 2024 0 repositories listed
-
Large Deviation Upper Bounds and Improved MSE Rates of Nonlinear SGD: Heavy-tailed Noise and Power of Symmetry21 Oct 2024 0 repositories listed
-
LSCodec: Low-Bitrate and Speaker-Decoupled Discrete Speech Codec21 Oct 2024 0 repositories listed
-
Solving Continual Offline RL through Selective Weights Activation on Aligned Spaces21 Oct 2024 0 repositories listed
-
Lossless KV Cache Compression to 2%20 Oct 2024 0 repositories listed
-
SDP4Bit: Toward 4-bit Communication Quantization in Sharded Data Parallelism for LLM Training20 Oct 2024 0 repositories listed
-
Understanding the Difficulty of Low-Precision Post-Training Quantization for LLMs18 Oct 2024 0 repositories listed
-
A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models17 Oct 2024 0 repositories listed
-
AsymKV: Enabling 1-Bit Quantization of KV Cache with Layer-Wise Asymmetric Quantization Configurations17 Oct 2024 0 repositories listed
-
DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech17 Oct 2024 0 repositories listed
-
Harnessing Your DRAM and SSD for Sustainable and Accessible LLM Inference with Mixed-Precision and Multi-level Caching17 Oct 2024 0 repositories listed
-
Nonlinear Stochastic Gradient Descent and Heavy-tailed Noise: A Unified Framework and High-probability Guarantees17 Oct 2024 0 repositories listed
-
Progressive Mixed-Precision Decoding for Efficient LLM Inference17 Oct 2024 0 repositories listed
-
Channel-Wise Mixed-Precision Quantization for Large Language Models16 Oct 2024 0 repositories listed
-
COMET: Towards Partical W4A4KV4 LLMs Serving16 Oct 2024 0 repositories listed
-
ERVQ: Enhanced Residual Vector Quantization with Intra-and-Inter-Codebook Optimization for Neural Audio Codecs16 Oct 2024 0 repositories listed
-
QSpec: Speculative Decoding with Complementary Quantization Schemes15 Oct 2024 0 repositories listed
-
Scaling Laws for Post Training Quantized Large Language Models15 Oct 2024 0 repositories listed
-
Gaussian Mixture Vector Quantization with Aggregated Categorical Posterior14 Oct 2024 0 repositories listed
-
Real-Time Stress Detection via Photoplethysmogram Signals: Implementation of a Combined Continuous Wavelet Transform and Convolutional Neural Network on Resource-Constrained Microcontrollers14 Oct 2024 0 repositories listed
-
SLaNC: Static LayerNorm Calibration14 Oct 2024 0 repositories listed
-
GALA: Geometry-Aware Local Adaptive Grids for Detailed 3D Generation13 Oct 2024 0 repositories listed
-
Gradient-Free Neural Network Training on the Edge13 Oct 2024 0 repositories listed
-
PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization12 Oct 2024 0 repositories listed
-
DeltaDQ: Ultra-High Delta Compression for Fine-Tuned LLMs via Group-wise Dropout and Separate Quantization11 Oct 2024 0 repositories listed
-
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification11 Oct 2024 0 repositories listed
-
CrossQuant: A Post-Training Quantization Method with Smaller Quantization Kernel for Precise Large Language Model Compression10 Oct 2024 0 repositories listed
-
DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation10 Oct 2024 0 repositories listed
-
M²-ViT: Accelerating Hybrid Vision Transformers with Two-Level Mixed Quantization10 Oct 2024 0 repositories listed
-
Scalable Representation Learning for Multimodal Tabular Transactions10 Oct 2024 0 repositories listed
-
QuAILoRA: Quantization-Aware Initialization for LoRA9 Oct 2024 0 repositories listed
-
Scaling Laws for Mixed quantization in Large Language Models9 Oct 2024 0 repositories listed
-
Covering Numbers for Deep ReLU Networks with Applications to Function Approximation and Nonparametric Regression8 Oct 2024 0 repositories listed
-
Gesture2Text: A Generalizable Decoder for Word-Gesture Keyboards in XR Through Trajectory Coarse Discretization and Pre-training8 Oct 2024 0 repositories listed
-
QERA: an Analytical Framework for Quantization Error Reconstruction8 Oct 2024 0 repositories listed
-
Variable Bitrate Residual Vector Quantization for Audio Coding8 Oct 2024 0 repositories listed
-
Designing a Classifier for Active Fire Detection from Multispectral Satellite Imagery Using Neural Architecture Search7 Oct 2024 0 repositories listed
-
Variable Resolution Pixel Quantization for Low Power Machine Vision Application on Edge7 Oct 2024 0 repositories listed
-
Continuous Approximations for Improving Quantization Aware Training of LLMs6 Oct 2024 0 repositories listed
-
HALL-E: Hierarchical Neural Codec Language Model for Minute-Long Zero-Shot Text-to-Speech Synthesis6 Oct 2024 0 repositories listed
-
PalmBench: A Comprehensive Benchmark of Compressed Large Language Models on Mobile Platforms5 Oct 2024 0 repositories listed
-
Generative Semantic Communication for Text-to-Speech Synthesis4 Oct 2024 0 repositories listed
-
MIMO Detection with Spatial Sigma-Delta ADCs: A Variational Bayesian Approach4 Oct 2024 0 repositories listed
-
Resource-aware Mixed-precision Quantization for Enhancing Deployability of Transformers for Time-series Forecasting on Embedded FPGAs4 Oct 2024 0 repositories listed
-
Overcoming Representation Bias in Fairness-Aware data Repair using Optimal Transport3 Oct 2024 0 repositories listed
-
Remember and Recall: Associative-Memory-based Trajectory Prediction3 Oct 2024 0 repositories listed
-
SEAL: SEmantic-Augmented Imitation Learning via Language Model3 Oct 2024 0 repositories listed
-
Getting Free Bits Back from Rotational Symmetries in LLMs2 Oct 2024 0 repositories listed
-
Restorative Speech Enhancement: A Progressive Approach Using SE and Codec Modules2 Oct 2024 0 repositories listed
-
Compressing Recurrent Neural Networks for FPGA-accelerated Implementation in Fluorescence Lifetime Imaging1 Oct 2024 0 repositories listed
-
Deep activity propagation via weight initialization in spiking neural networks1 Oct 2024 0 repositories listed
-
STanH : Parametric Quantization for Variable Rate Learned Image Compression1 Oct 2024 0 repositories listed
-
Aggressive Post-Training Compression on Extremely Large Language Models30 Sep 2024 0 repositories listed
-
Constraint Guided Model Quantization of Neural Networks30 Sep 2024 0 repositories listed
-
Mixed-Precision Embeddings for Large-Scale Recommendation Models30 Sep 2024 0 repositories listed
-
Quantized and Asynchronous Federated Learning30 Sep 2024 0 repositories listed
-
Rotated Runtime Smooth: Training-Free Activation Smoother for accurate INT4 inference30 Sep 2024 0 repositories listed
-
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries29 Sep 2024 0 repositories listed
-
A method of using RSVD in residual calculation of LowBit GEMM27 Sep 2024 0 repositories listed
-
Asymptotic tracking control of dynamic reference over homomorphically encrypted data with finite modulus27 Sep 2024 0 repositories listed
-
Heterogeneous quantization regularizes spiking neural network activity27 Sep 2024 0 repositories listed
-
Efficient Arbitrary Precision Acceleration for Large Language Models on GPU Tensor Cores26 Sep 2024 0 repositories listed
-
Fronthaul-Constrained Distributed Radar Sensing26 Sep 2024 0 repositories listed
-
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling26 Sep 2024 0 repositories listed
-
P4Q: Learning to Prompt for Quantization in Visual-language Models26 Sep 2024 0 repositories listed
-
A Survey of Low-bit Large Language Models: Basics, Systems, and Algorithms25 Sep 2024 0 repositories listed
-
Accumulator-Aware Post-Training Quantization25 Sep 2024 0 repositories listed
-
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ25 Sep 2024 0 repositories listed
-
Reinforcement Learning for Finite Space Mean-Field Type Games25 Sep 2024 0 repositories listed
-
Using Random Codebooks for Audio Neural AutoEncoders25 Sep 2024 0 repositories listed
-
A Formalization of Image Vectorization by Region Merging24 Sep 2024 0 repositories listed
-
Communication and Energy Efficient Federated Learning using Zero-Order Optimization Technique24 Sep 2024 0 repositories listed
-
Twin Network Augmentation: A Novel Training Strategy for Improved Spiking Neural Networks and Efficient Weight Quantization24 Sep 2024 0 repositories listed
-
Ultra-low latency quantum-inspired machine learning predictors implemented on FPGA24 Sep 2024 0 repositories listed
-
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation22 Sep 2024 0 repositories listed
-
SPAQ-DL-SLAM: Towards Optimizing Deep Learning-based SLAM for Resource-Constrained Embedded Platforms22 Sep 2024 0 repositories listed
-
CorBin-FL: A Differentially Private Federated Learning Mechanism using Common Randomness20 Sep 2024 0 repositories listed
-
PTQ4ADM: Post-Training Quantization for Efficient Text Conditional Audio Diffusion Models20 Sep 2024 0 repositories listed
-
Reduced bit median quantization: A middle process for Efficient Image Compression20 Sep 2024 0 repositories listed
-
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions20 Sep 2024 0 repositories listed
-
Impact of ML Optimization Tactics on Greener Pre-Trained ML Models19 Sep 2024 0 repositories listed
-
NDVQ: Robust Neural Audio Codec with Normal Distribution-Based Vector Quantization19 Sep 2024 0 repositories listed
-
Scaling FP8 training to trillion-token LLMs19 Sep 2024 0 repositories listed
-
Art and Science of Quantizing Large-Scale Models: A Comprehensive Overview18 Sep 2024 0 repositories listed
-
Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference18 Sep 2024 0 repositories listed
-
Pareto Data Framework: Steps Towards Resource-Efficient Decision Making Using Minimum Viable Data (MVD)18 Sep 2024 0 repositories listed
-
Forearm Ultrasound based Gesture Recognition on Edge16 Sep 2024 0 repositories listed
-
LASERS: LAtent Space Encoding for Representations with Sparsity for Generative Modeling16 Sep 2024 0 repositories listed
-
Improving Statistical Significance in Human Evaluation of Automatic Metrics via Soft Pairwise Accuracy15 Sep 2024 0 repositories listed
-
Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports15 Sep 2024 0 repositories listed
-
MesonGS: Post-training Compression of 3D Gaussians via Efficient Attribute Transformation15 Sep 2024 0 repositories listed
-
Privacy-Preserving SAM Quantization for Efficient Edge Intelligence in Healthcare14 Sep 2024 0 repositories listed
-
Robust Training of Neural Networks at Arbitrary Precision and Sparsity14 Sep 2024 0 repositories listed
-
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling13 Sep 2024 0 repositories listed
-
Dequantization of a signal from two parallel quantized observations12 Sep 2024 0 repositories listed
-
Efficient and Reliable Vector Similarity Search Using Asymmetric Encoding with NAND-Flash for Many-Class Few-Shot Learning12 Sep 2024 0 repositories listed
-
Distributed Convolutional Neural Network Training on Mobile and Edge Clusters11 Sep 2024 0 repositories listed