Browse State-of-the-Art › Quantization › Papers, page 30
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 30 of 50: papers 2,901 to 3,000 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Norm Tweaking: High-performance Low-bit Quantization of Large Language Models6 Sep 2023 0 repositories listed
-
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking5 Sep 2023 0 repositories listed
-
On-Chip Hardware-Aware Quantization for Mixed Precision Neural Networks5 Sep 2023 0 repositories listed
-
QuantEase: Optimization-based Quantization for Language Models5 Sep 2023 0 repositories listed
-
RobustEdge: Low Power Adversarial Detection for Cloud-Edge Systems5 Sep 2023 0 repositories listed
-
On the fly Deep Neural Network Optimization Control for Low-Power Computer Vision4 Sep 2023 0 repositories listed
-
Softmax Bias Correction for Quantized Generative Models4 Sep 2023 0 repositories listed
-
eDKM: An Efficient and Accurate Train-time Weight Clustering for Large Language Models2 Sep 2023 0 repositories listed
-
FPTQ: Fine-grained Post-Training Quantization for Large Language Models30 Aug 2023 0 repositories listed
-
Implementation and Evaluation of Physical Layer Key Generation on SDR based LoRa Platform30 Aug 2023 0 repositories listed
-
On-Device Learning with Binary Neural Networks29 Aug 2023 0 repositories listed
-
MEMORY-VQ: Compression for Tractable Internet-Scale Memory28 Aug 2023 0 repositories listed
-
Efficient Learned Lossless JPEG Recompression25 Aug 2023 0 repositories listed
-
Hybrid noise shaping for audio coding using perfectly overlapped window24 Aug 2023 0 repositories listed
-
Quantized distributed Nash equilibrium seeking under DoS attacks24 Aug 2023 0 repositories listed
-
Compressed Models Decompress Race Biases: What Quantized Models Forget for Fair Face Recognition23 Aug 2023 0 repositories listed
-
Consistent Signal Reconstruction from Streaming Multivariate Time Series23 Aug 2023 0 repositories listed
-
Distributed Energy Resource Management: All-Time Resource-Demand Feasibility, Delay-Tolerance, Nonlinearity, and Beyond22 Aug 2023 0 repositories listed
-
Towards Clip-Free Quantized Super-Resolution Networks: How to Tame Representative Images22 Aug 2023 0 repositories listed
-
QD-BEV : Quantization-aware View-guided Distillation for Multi-view 3D Object Detection21 Aug 2023 0 repositories listed
-
Sampling From Autoencoders' Latent Space via Quantization And Probability Mass Function Concepts21 Aug 2023 0 repositories listed
-
Quantization-based Optimization with Perspective of Quantum Mechanics20 Aug 2023 0 repositories listed
-
Analyzing Quantization in TVM19 Aug 2023 0 repositories listed
-
FunQuant: A R package to perform quantization in the context of rare events and time-consuming simulations18 Aug 2023 0 repositories listed
-
ResQ: Residual Quantization for Video Perception18 Aug 2023 0 repositories listed
-
SHARK: A Lightweight Model Compression Approach for Large-scale Recommender Systems18 Aug 2023 0 repositories listed
-
JPEG Quantized Coefficient Recovery via DCT Domain Spatial-Frequential Transformer17 Aug 2023 0 repositories listed
-
FineQuant: Unlocking Efficiency with Fine-Grained Weight-Only Quantization for LLMs16 Aug 2023 0 repositories listed
-
Precision and Recall Reject Curves for Classification16 Aug 2023 0 repositories listed
-
A Survey on Model Compression for Large Language Models15 Aug 2023 0 repositories listed
-
AKVSR: Audio Knowledge Empowered Visual Speech Recognition by Compressing Audio Knowledge of a Pretrained Model15 Aug 2023 0 repositories listed
-
Gradient-Based Post-Training Quantization: Challenging the Status Quo15 Aug 2023 0 repositories listed
-
Efficient Neural PDE-Solvers using Quantization Aware Training14 Aug 2023 0 repositories listed
-
Unified Data-Free Compression: Pruning and Quantization without Fine-Tuning14 Aug 2023 0 repositories listed
-
Sensitivity-Aware Mixed-Precision Quantization and Width Optimization of Deep Neural Networks Through Cluster-Based Tree-Structured Parzen Estimation12 Aug 2023 0 repositories listed
-
NUPES : Non-Uniform Post-Training Quantization via Power Exponent Search10 Aug 2023 0 repositories listed
-
ReLU and Addition-based Gated RNN10 Aug 2023 0 repositories listed
-
FPGA Resource-aware Structured Pruning for Real-Time Neural Networks9 Aug 2023 0 repositories listed
-
SAfER: Layer-Level Sensitivity Assessment for Efficient and Robust Neural Network Inference9 Aug 2023 0 repositories listed
-
Quantization Aware Factorization for Deep Neural Network Compression8 Aug 2023 0 repositories listed
-
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search7 Aug 2023 0 repositories listed
-
Reducing Channel Estimation and Feedback Overhead in IRS-Aided Downlink System: A Quantize-then-Estimate Approach4 Aug 2023 0 repositories listed
-
Frequency Disentangled Features in Neural Image Compression4 Aug 2023 0 repositories listed
-
RobustMQ: Benchmarking Robustness of Quantized Models4 Aug 2023 0 repositories listed
-
Error Analysis of CORDIC Processor with FPGA Implementation2 Aug 2023 0 repositories listed
-
Tango: rethinking quantization for graph neural network training on GPUs2 Aug 2023 0 repositories listed
-
AQUILA: Communication Efficient Federated Learning with Adaptive Quantization in Device Selection Strategy1 Aug 2023 0 repositories listed
-
Asynchronous Federated Learning with Bidirectional Quantized Communications and Buffered Aggregation1 Aug 2023 0 repositories listed
-
MRQ:Support Multiple Quantization Schemes through Model Re-Quantization1 Aug 2023 0 repositories listed
-
Alternate Learning based Sparse Semantic Communications for Visual Transmission31 Jul 2023 0 repositories listed
-
An Automata-Theoretic Approach to Synthesizing Binarized Neural Networks29 Jul 2023 0 repositories listed
-
METTS: Multilingual Emotional Text-to-Speech by Cross-speaker and Cross-lingual Emotion Transfer29 Jul 2023 0 repositories listed
-
Incrementally-Computable Neural Networks: Efficient Inference for Dynamic Inputs27 Jul 2023 0 repositories listed
-
High-Resolution Volumetric Reconstruction for Clothed Humans25 Jul 2023 0 repositories listed
-
Model Compression Methods for YOLOv5: A Review21 Jul 2023 0 repositories listed
-
Communication-Efficient Federated Learning over Capacity-Limited Wireless Networks20 Jul 2023 0 repositories listed
-
Communication-Efficient Split Learning via Adaptive Feature-Wise Compression20 Jul 2023 0 repositories listed
-
Quantized Feature Distillation for Network Quantization20 Jul 2023 0 repositories listed
-
Grounded Object Centric Learning18 Jul 2023 0 repositories listed
-
Extreme Image Compression using Fine-tuned VQGANs17 Jul 2023 0 repositories listed
-
Low bit rate binaural link for improved ultra low-latency low-complexity multichannel speech enhancement in Hearing Aids17 Jul 2023 0 repositories listed
-
A Survey of Techniques for Optimizing Transformer Inference16 Jul 2023 0 repositories listed
-
Learning Kernel-Modulated Neural Representation for Efficient Light Field Compression12 Jul 2023 0 repositories listed
-
Self-Distilled Quantization: Achieving High Compression Rates in Transformer-Based Language Models12 Jul 2023 0 repositories listed
-
Minimax Excess Risk of First-Order Methods for Statistical Learning with Data-Dependent Oracles10 Jul 2023 0 repositories listed
-
Q-YOLOP: Quantization-aware You Only Look Once for Panoptic Driving Perception10 Jul 2023 0 repositories listed
-
QBitOpt: Fast and Accurate Bitwidth Reallocation during Training10 Jul 2023 0 repositories listed
-
InfLoR-SNN: Reducing Information Loss for Spiking Neural Networks10 Jul 2023 0 repositories listed
-
Towards Efficient In-memory Computing Hardware for Quantized Neural Networks: State-of-the-art, Open Challenges and Perspectives8 Jul 2023 0 repositories listed
-
ITA: An Energy-Efficient Attention and Softmax Accelerator for Quantized Transformers7 Jul 2023 0 repositories listed
-
Free Bits: Latency Optimization of Mixed-Precision Quantized Neural Networks on the Edge6 Jul 2023 0 repositories listed
-
Greedy Selection for Heterogeneous Sensors3 Jul 2023 0 repositories listed
-
Data-Free Quantization via Mixed-Precision Compensation without Fine-Tuning2 Jul 2023 0 repositories listed
-
Line Spectrum Estimation and Detection with Few-bit ADCs: Theoretical Analysis and Generalized NOMP Algorithm2 Jul 2023 0 repositories listed
-
Analysis of the influence of final resolution on ADC accuracy1 Jul 2023 0 repositories listed
-
On a Relation Between the Rate-Distortion Function and Optimal Transport1 Jul 2023 0 repositories listed
-
Q-YOLO: Efficient Inference for Real-time Object Detection1 Jul 2023 0 repositories listed
-
Analysis of Oversampling in Uplink Massive MIMO-OFDM with Low-Resolution ADCs30 Jun 2023 0 repositories listed
-
Designing strong baselines for ternary neural network quantization through support and mass equalization30 Jun 2023 0 repositories listed
-
ReLU Neural Networks, Polyhedral Decompositions, and Persistent Homolog30 Jun 2023 0 repositories listed
-
Unlimited Sampling Radar: a Real-Time End-to-End Demonstrator30 Jun 2023 0 repositories listed
-
A Structurally Regularized CNN Architecture via Adaptive Subband Decomposition29 Jun 2023 0 repositories listed
-
DNA-TEQ: An Adaptive Exponential Quantization of Tensors for DNN Inference28 Jun 2023 0 repositories listed
-
INR-MDSQC: Implicit Neural Representation Multiple Description Scalar Quantization for robust image Coding24 Jun 2023 0 repositories listed
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression24 Jun 2023 0 repositories listed
-
QNNRepair: Quantized Neural Network Repair23 Jun 2023 0 repositories listed
-
Image storage on synthetic DNA using compressive autoencoders and DNA-adapted entropy coders22 Jun 2023 0 repositories listed
-
Subgraph Stationary Hardware-Software Inference Co-Design21 Jun 2023 0 repositories listed
-
DynaQuant: Compressing Deep Learning Training Checkpoints via Dynamic Quantization20 Jun 2023 0 repositories listed
-
Low-complexity Multidimensional DCT Approximations20 Jun 2023 0 repositories listed
-
Pushing the Limits of 3D Shape Generation at Scale20 Jun 2023 0 repositories listed
-
Dynamic Cell Modeling of Li-Ion Polymer Batteries for Precise SOC Estimation in Power-Needy Autonomous Electric Vehicles19 Jun 2023 0 repositories listed
-
Magnificent Minified Models16 Jun 2023 0 repositories listed
-
Neural Network Compression using Binarization and Few Full-Precision Weights15 Jun 2023 0 repositories listed
-
High-performance deep spiking neural networks with 0.3 spikes per neuron14 Jun 2023 0 repositories listed
-
Discrete Graph Auto-Encoder13 Jun 2023 0 repositories listed
-
MFSN: Multi-perspective Fusion Search Network For Pre-training Knowledge in Speech Emotion Recognition12 Jun 2023 0 repositories listed
-
Resource Efficient Neural Networks Using Hessian Based Pruning12 Jun 2023 0 repositories listed
-
Sparse-Inductive Generative Adversarial Hashing for Nearest Neighbor Search12 Jun 2023 0 repositories listed
-
End-to-End Neural Network Compression via ℓ₁/ℓ₂ Regularized Latency Surrogates9 Jun 2023 0 repositories listed