Browse State-of-the-Art › Quantization › Papers, page 24
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 24 of 50: papers 2,301 to 2,400 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Adaptive Error-Bounded Hierarchical Matrices for Efficient Neural Network Compression11 Sep 2024 0 repositories listed
-
NVRC: Neural Video Representation Compression11 Sep 2024 0 repositories listed
-
STORE: Streamlining Semantic Tokenization and Generative Recommendation with A Single LLM11 Sep 2024 0 repositories listed
-
AgileIR: Memory-Efficient Group Shifted Windows Attention for Agile Image Restoration10 Sep 2024 0 repositories listed
-
Rate-Constrained Quantization for Communication-Efficient Federated Learning10 Sep 2024 0 repositories listed
-
Distributed Optimization with Finite Bit Adaptive Quantization for Efficient Communication and Precision Enhancement9 Sep 2024 0 repositories listed
-
ECG Biometric Authentication Using Self-Supervised Learning for IoT Edge Sensors9 Sep 2024 0 repositories listed
-
Estimating the Completeness of Discrete Speech Units9 Sep 2024 0 repositories listed
-
SGC-VQGAN: Towards Complex Scene Representation via Semantic Guided Clustering Codebook9 Sep 2024 0 repositories listed
-
TriplePlay: Enhancing Federated Learning with CLIP for Non-IID Data and Resource Efficiency9 Sep 2024 0 repositories listed
-
Blind-Adaptive Quantizers6 Sep 2024 0 repositories listed
-
OPAL: Outlier-Preserved Microscaling Quantization Accelerator for Generative Large Language Models6 Sep 2024 0 repositories listed
-
5 Sep 2024 0 repositories listed
-
Investigating Privacy Bias in Training Data of Language Models5 Sep 2024 0 repositories listed
-
Recursive Quantization for ℒ₂ Stabilization of a Finite Capacity Stochastic Control Loop with Intermittent State Observations5 Sep 2024 0 repositories listed
-
WaterMAS: Sharpness-Aware Maximization for Neural Network Watermarking5 Sep 2024 0 repositories listed
-
CoAst: Validation-Free Contribution Assessment for Federated Learning based on Cross-Round Valuation4 Sep 2024 0 repositories listed
-
Gaussian Rate-Distortion-Perception Coding and Entropy-Constrained Scalar Quantization4 Sep 2024 0 repositories listed
-
Learning Task-Based Trainable Neuromorphic ADCs via Power-Aware Distillation4 Sep 2024 0 repositories listed
-
4 Sep 2024 0 repositories listed Syntology 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Task-Oriented Communication for Graph Data: A Graph Information Bottleneck Approach4 Sep 2024 0 repositories listed
-
Foundations of Large Language Model Compression -- Part 1: Weight Quantization3 Sep 2024 0 repositories listed
-
Optimization and Deployment of Deep Neural Networks for PPG-based Blood Pressure Estimation Targeting Low-power Wearables3 Sep 2024 0 repositories listed
-
Compressing VAE-Based Out-of-Distribution Detectors for Embedded Deployment2 Sep 2024 0 repositories listed
-
Edge AI: Evaluation of Model Compression Techniques for Convolutional Neural Networks2 Sep 2024 0 repositories listed
-
One-Index Vector Quantization Based Adversarial Attack on Image Classification2 Sep 2024 0 repositories listed
-
Enhancing Multi-Stream Beamforming Through CQIs For 5G NR FDD Massive MIMO Communications: A Tuning-Free Scheme1 Sep 2024 0 repositories listed
-
Federated Aggregation of Mallows Rankings: A Comparative Analysis of Borda and Lehmer Coding1 Sep 2024 0 repositories listed
-
Accurate Compression of Text-to-Image Diffusion Models via Vector Quantization31 Aug 2024 0 repositories listed
-
Approximately Invertible Neural Network for Learned Image Compression30 Aug 2024 0 repositories listed
-
VQ4DiT: Efficient Post-Training Vector Quantization for Diffusion Transformers30 Aug 2024 0 repositories listed
-
Blending Low and High-Level Semantics of Time Series for Better Masked Time Series Generation29 Aug 2024 0 repositories listed
-
On-device AI: Quantization-aware Training of Transformers in Time-Series29 Aug 2024 0 repositories listed
-
The Uniqueness of LLaMA3-70B Series with Per-Channel Quantization27 Aug 2024 0 repositories listed
-
Adaptive Resolution Inference (ARI): Energy-Efficient Machine Learning for Internet of Things26 Aug 2024 0 repositories listed
-
FusionSAM: Latent Space driven Segment Anything Model for Multimodal Fusion and Segmentation26 Aug 2024 0 repositories listed
-
Scalable Multivariate Fronthaul Quantization for Cell-Free Massive MIMO26 Aug 2024 0 repositories listed
-
Infrared Domain Adaptation with Zero-Shot Quantization25 Aug 2024 0 repositories listed
-
Quantized neural network for complex hologram generation25 Aug 2024 0 repositories listed
-
Revisiting DNN Training for Intermittently-Powered Energy-Harvesting Micro-Computers25 Aug 2024 0 repositories listed
-
Variational autoencoder-based neural network model compression25 Aug 2024 0 repositories listed
-
A Safe Self-evolution Algorithm for Autonomous Driving Based on Data-Driven Risk Quantification Model23 Aug 2024 0 repositories listed
-
Informational Embodiment: Computational role of information structure in codes and robots23 Aug 2024 0 repositories listed
-
DeepHQ: Learned Hierarchical Quantizer for Progressive Deep Image Coding22 Aug 2024 0 repositories listed
-
Matmul or No Matmal in the Era of 1-bit LLMs21 Aug 2024 0 repositories listed
-
Disentangling segmental and prosodic factors to non-native speech comprehensibility20 Aug 2024 0 repositories listed
-
Hyperstroke: A Novel High-quality Stroke Representation for Assistive Artistic Drawing18 Aug 2024 0 repositories listed
-
Explore Cross-Codec Quality-Rate Convex Hulls Relation for Adaptive Streaming16 Aug 2024 0 repositories listed
-
JPEG-LM: LLMs as Image Generators with Canonical Codec Representations15 Aug 2024 0 repositories listed
-
Analog Spiking Neuron in CMOS 28 nm Towards Large-Scale Neuromorphic Processors14 Aug 2024 0 repositories listed
-
Line Spectral Estimation with Unlimited Sensing13 Aug 2024 0 repositories listed
-
Low-Bitwidth Floating Point Quantization for Efficient High-Quality Diffusion Models13 Aug 2024 0 repositories listed
-
Prompt Tuning as User Inherent Profile Inference Machine13 Aug 2024 0 repositories listed
-
Computability of Classification and Deep Learning: From Theoretical Limits to Practical Feasibility through Quantization12 Aug 2024 0 repositories listed
-
RTF-Q: Efficient Unsupervised Domain Adaptation with Retraining-free Quantization11 Aug 2024 0 repositories listed
-
Quantum-secure multiparty deep learning10 Aug 2024 0 repositories listed
-
Semantic-Enabled 6G Communication: A Task-oriented and Privacy-preserving Perspective8 Aug 2024 0 repositories listed
-
FDC: Fast KV Dimensionality Compression for Efficient LLM Inference7 Aug 2024 0 repositories listed
-
Self-Supervised Learning for Multi-Channel Neural Transducer6 Aug 2024 0 repositories listed
-
Synaptic Modulation using Interspike Intervals Increases Energy Efficiency of Spiking Neural Networks6 Aug 2024 0 repositories listed
-
L3iTC at the FinLLM Challenge Task: Quantization for Financial Text Classification & Summarization6 Aug 2024 0 repositories listed
-
Inference Optimizations for Large Language Models: Effects, Challenges, and Practical Considerations6 Aug 2024 0 repositories listed
-
DopQ-ViT: Towards Distribution-Friendly and Outlier-Aware Post-Training Quantization for Vision Transformers6 Aug 2024 0 repositories listed
-
An approach to optimize inference of the DIART speaker diarization pipeline5 Aug 2024 0 repositories listed
-
Nonlinear Perturbation-based Non-Convex Optimization over Time-Varying Networks5 Aug 2024 0 repositories listed
-
Winning Amazon KDD Cup'245 Aug 2024 0 repositories listed
-
STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs3 Aug 2024 0 repositories listed
-
HMDN: Hierarchical Multi-Distribution Network for Click-Through Rate Prediction2 Aug 2024 0 repositories listed
-
CDFGNN: a Systematic Design of Cache-based Distributed Full-Batch Graph Neural Network Training with Communication Reduction1 Aug 2024 0 repositories listed
-
UniMoT: Unified Molecule-Text Language Model with Discrete Token Representation1 Aug 2024 0 repositories listed
-
Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization1 Aug 2024 0 repositories listed
-
Breaking the Hourglass Phenomenon of Residual Quantization: Enhancing the Upper Bound of Generative Retrieval31 Jul 2024 0 repositories listed
-
Exploiting Change Blindness for Video Coding: Perspectives from a Less Promising User Study31 Jul 2024 0 repositories listed
-
Abstractive summarization from Audio Transcription30 Jul 2024 0 repositories listed
-
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity29 Jul 2024 0 repositories listed
-
Model Agnostic Hybrid Sharding For Heterogeneous Distributed Inference29 Jul 2024 0 repositories listed
-
Reputation-Driven Asynchronous Federated Learning for Enhanced Trajectory Prediction with Blockchain28 Jul 2024 0 repositories listed
-
The Interpretability of Codebooks in Model-Based Reinforcement Learning is Limited28 Jul 2024 0 repositories listed
-
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers25 Jul 2024 0 repositories listed
-
Unlocking Tokens as Data Points for Generalization Bounds on Larger Language Models25 Jul 2024 0 repositories listed
-
Pixel Embedding: Fully Quantized Convolutional Neural Network with Differentiable Lookup Table23 Jul 2024 0 repositories listed
-
Compensate Quantization Errors+: Quantized Models Are Inquisitive Learners22 Jul 2024 0 repositories listed
-
Comprehensive Study on Performance Evaluation and Optimization of Model Compression: Bridging Traditional Deep Learning and Large Language Models22 Jul 2024 0 repositories listed
-
Uplink Transmit Power Optimization for Distributed Massive MIMO Systems with 1-Bit ADCs22 Jul 2024 0 repositories listed
-
FedDM: Enhancing Communication Efficiency and Handling Data Heterogeneity in Federated Diffusion Models20 Jul 2024 0 repositories listed
-
Power Measurement Enabled Channel Autocorrelation Matrix Estimation for IRS-Assisted Wireless Communication20 Jul 2024 0 repositories listed
-
Mixture of Experts with Mixture of Precisions for Tuning Quality of Service19 Jul 2024 0 repositories listed
-
Asymptotically Optimal Closed-Form Phase Configuration of 1-bit RISs via Sign Alignment18 Jul 2024 0 repositories listed
-
LiNR: Model Based Neural Retrieval on GPUs at LinkedIn18 Jul 2024 0 repositories listed
-
FETCH: A Memory-Efficient Replay Approach for Continual Learning in Image Classification17 Jul 2024 0 repositories listed
-
Mamba-PTQ: Outlier Channels in Recurrent Large Language Models17 Jul 2024 0 repositories listed
-
MCU-MixQ: A HW/SW Co-optimized Mixed-precision Neural Network Design Framework for MCUs17 Jul 2024 0 repositories listed
-
SmartQuant: CXL-based AI Model Store in Support of Runtime Configurable Weight Quantization17 Jul 2024 0 repositories listed
-
Toward INT4 Fixed-Point Training via Exploring Quantization Error for Gradients17 Jul 2024 0 repositories listed
-
Co-Designing Binarized Transformer and Hardware Accelerator for Efficient End-to-End Edge Deployment16 Jul 2024 0 repositories listed
-
LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices16 Jul 2024 0 repositories listed
-
QVD: Post-training Quantization for Video Diffusion Models16 Jul 2024 0 repositories listed
-
Rate-Distortion-Cognition Controllable Versatile Neural Image Compression16 Jul 2024 0 repositories listed
-
Tiled Bit Networks: Sub-Bit Neural Network Compression Through Reuse of Learnable Binary Vectors16 Jul 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.