Browse State-of-the-Art › Quantization › Papers, page 28
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 28 of 50: papers 2,701 to 2,800 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Optimal and Near-Optimal Adaptive Vector Quantization5 Feb 2024 0 repositories listed
-
Quantized Approximately Orthogonal Recurrent Neural Networks5 Feb 2024 0 repositories listed
-
FoldToken: Learning Protein Language via Vector Quantization and Beyond4 Feb 2024 0 repositories listed
-
Stability Analysis of Various Symbolic Rule Extraction Methods from Recurrent Neural Network4 Feb 2024 0 repositories listed
-
Locally-Adaptive Quantization for Streaming Vector Search3 Feb 2024 0 repositories listed
-
An Intra-BRNN and GB-RVQ Based END-TO-END Neural Audio Codec2 Feb 2024 0 repositories listed
-
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation2 Feb 2024 0 repositories listed
-
FedShift: Tackling Dual Heterogeneity Problem of Federated Learning via Weight Shift Aggregation2 Feb 2024 0 repositories listed
-
HW-SW Optimization of DNNs for Privacy-preserving People Counting on Low-resolution Infrared Arrays2 Feb 2024 0 repositories listed
-
Improved Quantization Strategies for Managing Heavy-tailed Gradients in Distributed Learning2 Feb 2024 0 repositories listed
-
Neural Language of Thought Models2 Feb 2024 0 repositories listed
-
Truncated Non-Uniform Quantization for Distributed SGD2 Feb 2024 0 repositories listed
-
Analog-digital Scheduling for Federated Learning: A Communication-Efficient Approach1 Feb 2024 0 repositories listed
-
Can Large Language Models Understand Context?1 Feb 2024 0 repositories listed
-
Trainable Fixed-Point Quantization for Deep Learning Acceleration on FPGAs31 Jan 2024 0 repositories listed
-
Effect of Weight Quantization on Learning Models by Typical Case Analysis30 Jan 2024 0 repositories listed
-
HEQuant: Marrying Homomorphic Encryption and Quantization for Communication-Efficient Private Inference29 Jan 2024 0 repositories listed
-
A Comprehensive Survey of Compression Algorithms for Language Models27 Jan 2024 0 repositories listed
-
Transformer-based Clipped Contrastive Quantization Learning for Unsupervised Image Retrieval27 Jan 2024 0 repositories listed
-
LitE-SNN: Designing Lightweight and Efficient Spiking Neural Network through Spatial-Temporal Compressive Network Search and Joint Optimization26 Jan 2024 0 repositories listed
-
MPTQ-ViT: Mixed-Precision Post-Training Quantization for Vision Transformer26 Jan 2024 0 repositories listed
-
CompactifAI: Extreme Compression of Large Language Models using Quantum-Inspired Tensor Networks25 Jan 2024 0 repositories listed
-
Towards Cheaper Inference in Deep Networks with Lower Bit-Width Accumulators25 Jan 2024 0 repositories listed
-
Within-basket Recommendation via Neural Pattern Associator25 Jan 2024 0 repositories listed
-
Value-Driven Mixed-Precision Quantization for Patch-Based Inference on Microcontrollers24 Jan 2024 0 repositories listed
-
Iterated Relevance Matrix Analysis (IRMA) for the identification of class-discriminative subspaces23 Jan 2024 0 repositories listed
-
Robustness to distribution shifts of compressed networks for edge devices22 Jan 2024 0 repositories listed
-
Scaling Up Quantization-Aware Neural Architecture Search for Efficient Deep Learning on the Edge22 Jan 2024 0 repositories listed
-
Another Way to the Top: Exploit Contextual Clustering in Learned Image Coding21 Jan 2024 0 repositories listed
-
Edge-Enabled Real-time Railway Track Segmentation21 Jan 2024 0 repositories listed
-
LRP-QViT: Mixed-Precision Vision Transformer Quantization via Layer-wise Relevance Propagation20 Jan 2024 0 repositories listed
-
Dynamic Q&A of Clinical Documents with Large Language Models19 Jan 2024 0 repositories listed
-
Enabling On-device Continual Learning with Binary Neural Networks18 Jan 2024 0 repositories listed
-
Exploration of Activation Fault Reliability in Quantized Systolic Array-Based DNN Accelerators17 Jan 2024 0 repositories listed
-
Hybrid of DiffStride and Spectral Pooling in Convolutional Neural Networks17 Jan 2024 0 repositories listed
-
TP-Aware Dequantization15 Jan 2024 0 repositories listed
-
ENTED: Enhanced Neural Texture Extraction and Distribution for Reference-based Blind Face Restoration13 Jan 2024 0 repositories listed
-
Correlated Quantization for Faster Nonconvex Distributed Optimization10 Jan 2024 0 repositories listed
-
Memory-Efficient Fine-Tuning for Quantized Diffusion Model9 Jan 2024 0 repositories listed
-
A Video Coding Method Based on Neural Network for CLIC20248 Jan 2024 0 repositories listed
-
Detecting Face Synthesis Using a Concealed Fusion Model8 Jan 2024 0 repositories listed
-
FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGAs8 Jan 2024 0 repositories listed
-
Data-driven Dynamic Event-triggered Control7 Jan 2024 0 repositories listed
-
A Cost-Efficient FPGA Implementation of Tiny Transformer Model using Neural ODE5 Jan 2024 0 repositories listed
-
Enhancing Generalization of Invisible Facial Privacy Cloak via Gradient Accumulation3 Jan 2024 0 repositories listed
-
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels2 Jan 2024 0 repositories listed
-
Are Conventional SNNs Really Efficient? A Perspective from Network Quantization1 Jan 2024 0 repositories listed
-
Data-Free Quantization via Pseudo-label Filtering1 Jan 2024 0 repositories listed
-
Enhancing Post-training Quantization Calibration through Contrastive Learning1 Jan 2024 0 repositories listed
-
PikeLPN: Mitigating Overlooked Inefficiencies of Low-Precision Neural Networks1 Jan 2024 0 repositories listed
-
PredToken: Predicting Unknown Tokens and Beyond with Coarse-to-Fine Iterative Decoding1 Jan 2024 0 repositories listed
-
Reg-PTQ: Regression-specialized Post-training Quantization for Fully Quantized Object Detector1 Jan 2024 0 repositories listed
-
HQ-VAE: Hierarchical Discrete Representation Learning with Variational Bayes31 Dec 2023 0 repositories listed
-
Compact Neural Graphics Primitives with Learned Hash Probing28 Dec 2023 0 repositories listed
-
A-SDM: Accelerating Stable Diffusion through Redundancy Removal and Performance Optimization24 Dec 2023 0 repositories listed
-
Efficient Asynchronous Federated Learning with Sparsification and Quantization23 Dec 2023 0 repositories listed
-
Hardware-Aware DNN Compression via Diverse Pruning and Mixed-Precision Quantization23 Dec 2023 0 repositories listed
-
Cross-Layer Optimization for Fault-Tolerant Deep Learning21 Dec 2023 0 repositories listed
-
SimQ-NAS: Simultaneous Quantization Policy and Neural Architecture Search19 Dec 2023 0 repositories listed
-
Power-Efficient Sampling18 Dec 2023 0 repositories listed
-
Quantized Decoder in Learned Image Compression for Deterministic Reconstruction18 Dec 2023 0 repositories listed
-
IQNet: Image Quality Assessment Guided Just Noticeable Difference Prefiltering For Versatile Video Coding15 Dec 2023 0 repositories listed
-
Design Space Exploration of Low-Bit Quantized Neural Networks for Visual Place Recognition14 Dec 2023 0 repositories listed
-
CBQ: Cross-Block Quantization for Large Language Models13 Dec 2023 0 repositories listed
-
USM-Lite: Quantization and Sparsity Aware Fine-tuning for Speech Recognition with Universal Speech Models13 Dec 2023 0 repositories listed
-
12 Dec 2023 0 repositories listed
-
IDKM: Memory Efficient Neural Network Quantization via Implicit, Differentiable k-Means12 Dec 2023 0 repositories listed
-
When Bio-Inspired Computing meets Deep Learning: Low-Latency, Accurate, & Energy-Efficient Spiking Neural Networks from Artificial Neural Networks12 Dec 2023 0 repositories listed
-
FP8-BERT: Post-Training Quantization for Transformer10 Dec 2023 0 repositories listed
-
Neural Architecture Codesign for Fast Bragg Peak Analysis10 Dec 2023 0 repositories listed
-
QMGeo: Differentially Private Federated Learning via Stochastic Quantization with Mixed Truncated Geometric Distribution10 Dec 2023 0 repositories listed
-
Automotive Radar Sensing with Sparse Linear Arrays Using One-Bit Hankel Matrix Completion9 Dec 2023 0 repositories listed
-
Efficient Quantization Strategies for Latent Diffusion Models9 Dec 2023 0 repositories listed
-
An Experimental Study: Assessing the Combined Framework of WavLM and BEST-RQ for Text-to-Speech Synthesis8 Dec 2023 0 repositories listed
-
Rate-splitting Multiple Access for Hierarchical HAP-LAP Networks under Limited Fronthaul7 Dec 2023 0 repositories listed
-
GenQ: Quantization in Low Data Regimes with Generative Synthetic Data7 Dec 2023 0 repositories listed
-
Enhancing Kinship Verification through Multiscale Retinex and Combined Deep-Shallow features6 Dec 2023 0 repositories listed
-
All Rivers Run to the Sea: Private Learning with Asymmetric Flows5 Dec 2023 0 repositories listed
-
Unified learning-based lossy and lossless JPEG recompression5 Dec 2023 0 repositories listed
-
Low-Precision Mixed-Computation Models for Inference on Edge3 Dec 2023 0 repositories listed
-
Adaptive Resource Allocation for Semantic Communication Networks2 Dec 2023 0 repositories listed
-
A New Old Idea: Beam-Steering Reflectarrays for Efficient Sub-THz Multiuser MIMO30 Nov 2023 0 repositories listed
-
Improving the Robustness of Quantized Deep Neural Networks to White-Box Attacks using Stochastic Quantization and Information-Theoretic Ensemble Training30 Nov 2023 0 repositories listed
-
Fault-Tolerant Four-Dimensional Constellation for Coherent Optical Transmission Systems29 Nov 2023 0 repositories listed
-
Mixed-Precision Quantization for Federated Learning on Resource-Constrained Heterogeneous Devices29 Nov 2023 0 repositories listed
-
Fast and Efficient 2-bit LLM Inference on GPU: 2/4/16-bit in a Weight Matrix with Asynchronous Dequantization28 Nov 2023 0 repositories listed
-
PIPE : Parallelized Inference Through Post-Training Quantization Ensembling of Residual Expansions27 Nov 2023 0 repositories listed
-
Relationship between Model Compression and Adversarial Robustness: A Review of Current Evidence27 Nov 2023 0 repositories listed
-
SNN Architecture for Differential Time Encoding Using Decoupled Processing Time24 Nov 2023 0 repositories listed
-
A Blockchain Solution for Collaborative Machine Learning over IoT23 Nov 2023 0 repositories listed
-
SySMOL: Co-designing Algorithms and Hardware for Neural Networks with Heterogeneous Precisions23 Nov 2023 0 repositories listed
-
Modulation For Modulo: A Sampling-Efficient High-Dynamic Range ADC22 Nov 2023 0 repositories listed
-
Uncertainty Estimation in Multi-Agent Distributed Learning22 Nov 2023 0 repositories listed
-
Deep Learning-Based Real-Time Quality Control of Standard Video Compression for Live Streaming21 Nov 2023 0 repositories listed
-
Shedding the Bits: Pushing the Boundaries of Quantization with Minifloats on FPGAs21 Nov 2023 0 repositories listed
-
Efficient Neural Networks for Tiny Machine Learning: A Comprehensive Review20 Nov 2023 0 repositories listed
-
Tiny-VBF: Resource-Efficient Vision Transformer based Lightweight Beamformer for Ultrasound Single-Angle Plane Wave Imaging20 Nov 2023 0 repositories listed
-
Low-Precision Floating-Point for Efficient On-Board Deep Neural Network Processing18 Nov 2023 0 repositories listed
-
Is Conventional SNN Really Efficient? A Perspective from Network Quantization17 Nov 2023 0 repositories listed
-
A Speed Odyssey for Deployable Quantization of LLMs16 Nov 2023 0 repositories listed