Browse State-of-the-Art › Quantization › Papers, page 22
Quantization
Papers archive 2025-07-28
archive papers tagged: 4,925 · with a code link: 1,596 · where Syntology ran a sample: 515 (452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (515 of 4,925 tagged: 452 with a run with no instrument failure, 63 where every run was a failure of Syntology's instrument)
Page 22 of 50: papers 2,101 to 2,200 of 4,925, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Sensor Selection and Distributed Quantization for Energy Efficiency in Massive MTC7 Dec 2024 0 repositories listed
-
Trimming Down Large Spiking Vision Transformers via Heterogeneous Quantization Search7 Dec 2024 0 repositories listed
-
ULMRec: User-centric Large Language Model for Sequential Recommendation7 Dec 2024 0 repositories listed
-
Quantized and Interpretable Learning Scheme for Deep Neural Networks in Classification Task5 Dec 2024 0 repositories listed
-
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization5 Dec 2024 0 repositories listed
-
Designing DNNs for a trade-off between robustness and processing performance in embedded devices4 Dec 2024 0 repositories listed
-
FlashAttention on a Napkin: A Diagrammatic Approach to Deep Learning IO-Awareness4 Dec 2024 0 repositories listed
-
Mixed-Precision Quantization: Make the Best Use of Bits Where They Matter Most4 Dec 2024 0 repositories listed
-
Prompting Large Language Models for Clinical Temporal Relation Extraction4 Dec 2024 0 repositories listed
-
Unifying KV Cache Compression for Large Language Models with LeanKV4 Dec 2024 0 repositories listed
-
3D representation in 512-Byte:Variational tokenizer is the key for autoregressive 3D generation3 Dec 2024 0 repositories listed
-
Lean classical-quantum hybrid neural network model for image classification3 Dec 2024 0 repositories listed
-
CEGI: Measuring the trade-off between efficiency and carbon emissions for SLMs and VLMs3 Dec 2024 0 repositories listed
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models3 Dec 2024 0 repositories listed
-
Robust Precoding for Multi-User Visible Light Communications with Quantized Channel Information3 Dec 2024 0 repositories listed
-
Memory-Efficient Training for Deep Speaker Embedding Learning in Speaker Verification2 Dec 2024 0 repositories listed
-
Optimizing Domain-Specific Image Retrieval: A Benchmark of FAISS and Annoy with Fine-Tuned Features2 Dec 2024 0 repositories listed
-
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control2 Dec 2024 0 repositories listed
-
A Wave is Worth 100 Words: Investigating Cross-Domain Transferability in Time Series1 Dec 2024 0 repositories listed
-
LAMBDA: Covering the Multimodal Critical Scenarios for Automated Driving Systems by Search Space Quantization30 Nov 2024 0 repositories listed
-
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation29 Nov 2024 0 repositories listed
-
29 Nov 2024 0 repositories listed
-
Privacy-Preserving Orthogonal Aggregation for Guaranteeing Gender Fairness in Federated Recommendation29 Nov 2024 0 repositories listed
-
Quantized Delta Weight Is Safety Keeper29 Nov 2024 0 repositories listed
-
On the effectiveness of discrete representations in sparse mixture of experts28 Nov 2024 0 repositories listed
-
Orthus: Autoregressive Interleaved Image-Text Generation with Modality-Specific Heads28 Nov 2024 0 repositories listed
-
FAMES: Fast Approximate Multiplier Substitution for Mixed-Precision Quantized DNNs--Down to 2 Bits!27 Nov 2024 0 repositories listed
-
COAP: Memory-Efficient Training with Correlation-Aware Gradient Projection26 Nov 2024 0 repositories listed
-
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens26 Nov 2024 0 repositories listed
-
Rapid Deployment of Domain-specific Hyperspectral Image Processors with Application to Autonomous Driving26 Nov 2024 0 repositories listed
-
SoftmAP: Software-Hardware Co-design for Integer-Only Softmax on Associative Processors26 Nov 2024 0 repositories listed
-
Beyond Task Vectors: Selective Task Arithmetic Based on Importance Metrics25 Nov 2024 0 repositories listed
-
Curvature in the Looking-Glass: Optimal Methods to Exploit Curvature of Expectation in the Loss Landscape25 Nov 2024 0 repositories listed
-
Downlink MIMO Channel Estimation from Bits: Recoverability and Algorithm25 Nov 2024 0 repositories listed
-
Factorized Visual Tokenization and Generation25 Nov 2024 0 repositories listed
-
Learning Optimal Lattice Vector Quantizers for End-to-end Neural Image Compression25 Nov 2024 0 repositories listed
-
Lion Cub: Minimizing Communication Overhead in Distributed Lion25 Nov 2024 0 repositories listed
-
MixPE: Quantization and Hardware Co-design for Efficient LLM Inference25 Nov 2024 0 repositories listed
-
Representation Collapsing Problems in Vector Quantization25 Nov 2024 0 repositories listed
-
Rethinking Diffusion for Text-Driven Human Motion Generation25 Nov 2024 0 repositories listed
-
SKQVC: One-Shot Voice Conversion by K-Means Quantization with Self-Supervised Speech Representations25 Nov 2024 0 repositories listed
-
freePruner: A Training-free Approach for Large Multimodal Model Acceleration23 Nov 2024 0 repositories listed
-
FLARE: FP-Less PTQ and Low-ENOB ADC Based AMS-PiM for Error-Resilient, Fast, and Efficient Transformer Acceleration22 Nov 2024 0 repositories listed
-
AutoMixQ: Self-Adjusting Quantization for High Performance Memory-Efficient Fine-Tuning21 Nov 2024 0 repositories listed
-
21 Nov 2024 0 repositories listed Syntology 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Disco Intelligent Omni-Surfaces: 360-degree Fully-Passive Jamming Attacks20 Nov 2024 0 repositories listed
-
RTSR: A Real-Time Super-Resolution Model for AV1 Compressed Content20 Nov 2024 0 repositories listed
-
Diffusion Product Quantization19 Nov 2024 0 repositories listed
-
High-Throughput Blind Co-Channel Interference Cancellation for Edge Devices Using Depthwise Separable Convolutions, Quantization, and Pruning19 Nov 2024 0 repositories listed
-
EfQAT: An Efficient Framework for Quantization-Aware Training17 Nov 2024 0 repositories listed
-
Towards Accurate and Efficient Sub-8-Bit Integer Training17 Nov 2024 0 repositories listed
-
BlueLM-V-3B: Algorithm and System Co-Design for Multimodal Large Language Models on Mobile Devices16 Nov 2024 0 repositories listed
-
AMXFP4: Taming Activation Outliers with Asymmetric Microscaling Floating-Point for 4-bit LLM Inference15 Nov 2024 0 repositories listed
-
Systolic Arrays and Structured Pruning Co-design for Efficient Transformers in Edge Systems15 Nov 2024 0 repositories listed
-
Communication Compression for Tensor Parallel LLM Inference14 Nov 2024 0 repositories listed
-
ASER: Activation Smoothing and Error Reconstruction for Large Language Model Quantization12 Nov 2024 0 repositories listed
-
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning12 Nov 2024 0 repositories listed
-
Towards Low-bit Communication for Tensor Parallel LLM Inference12 Nov 2024 0 repositories listed
-
HarmLevelBench: Evaluating Harm-Level Compliance and the Impact of Quantization on Model Alignment11 Nov 2024 0 repositories listed
-
Sketched Adaptive Federated Deep Learning: A Sharp Convergence Analysis11 Nov 2024 0 repositories listed
-
HAFLQ: Heterogeneous Adaptive Federated LoRA Fine-tuned LLM with Quantization10 Nov 2024 0 repositories listed
-
Intelligent Fault Diagnosis of Type and Severity in Low-Frequency, Low Bit-Depth Signals9 Nov 2024 0 repositories listed
-
Optimizing Large Language Models through Quantization: A Comparative Analysis of PTQ and QAT Techniques9 Nov 2024 0 repositories listed
-
8 Nov 2024 0 repositories listed
-
QuanCrypt-FL: Quantized Homomorphic Encryption with Pruning for Secure Federated Learning8 Nov 2024 0 repositories listed
-
Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving8 Nov 2024 0 repositories listed
-
Rate-aware Compression for NeRF-based Volumetric Video8 Nov 2024 0 repositories listed
-
When are 1.58 bits enough? A Bottom-up Exploration of BitNet Quantization8 Nov 2024 0 repositories listed
-
Compressive Spectrum Sensing with 1-bit ADCs7 Nov 2024 0 repositories listed
-
Green My LLM: Studying the key factors affecting the energy consumption of code assistants7 Nov 2024 0 repositories listed
-
Saliency Assisted Quantization for Neural Networks7 Nov 2024 0 repositories listed
-
Interactions Across Blocks in Post-Training Quantization of Large Language Models6 Nov 2024 0 repositories listed
-
Multi-bit Distributed Detection of Sparse Stochastic Signals over Error-Prone Reporting Channels6 Nov 2024 0 repositories listed
-
Hybrid Beamforming for Integrated Sensing and Communications With Low Resolution DACs5 Nov 2024 0 repositories listed
-
Sum Rate Maximization in the Constant Envelope MIMO Downlink with the RZF Precoder5 Nov 2024 0 repositories listed
-
"Give Me BF16 or Give Me Death"? Accuracy-Performance Trade-Offs in LLM Quantization4 Nov 2024 0 repositories listed
-
Transferable Sequential Recommendation via Vector Quantized Meta Learning4 Nov 2024 0 repositories listed
-
BF-IMNA: A Bit Fluid In-Memory Neural Architecture for Neural Network Acceleration3 Nov 2024 0 repositories listed
-
Fundamental Trade-offs in Quantized Hybrid Radar Fusion: A CRB-Rate Perspective1 Nov 2024 0 repositories listed
-
Optimizing Contextual Speech Recognition Using Vector Quantization for Efficient Retrieval1 Nov 2024 0 repositories listed
-
ALISE: Accelerating Large Language Model Serving with Speculative Scheduling31 Oct 2024 0 repositories listed
-
ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs31 Oct 2024 0 repositories listed
-
Breaking Determinism: Fuzzy Modeling of Sequential Recommendation Using Discrete State Space Diffusion Model31 Oct 2024 0 repositories listed
-
A Comprehensive Study on Quantization Techniques for Large Language Models30 Oct 2024 0 repositories listed
-
Accelerated AI Inference via Dynamic Execution Methods30 Oct 2024 0 repositories listed
-
APCodec+: A Spectrum-Coding-Based High-Fidelity and High-Compression-Rate Neural Audio Codec with Staged Training Paradigm30 Oct 2024 0 repositories listed
-
ELMGS: Enhancing memory and computation scaLability through coMpression for 3D Gaussian Splatting30 Oct 2024 0 repositories listed
-
GWQ: Gradient-Aware Weight Quantization for Large Language Models30 Oct 2024 0 repositories listed
-
HRPVT: High-Resolution Pyramid Vision Transformer for medium and small-scale human pose estimation29 Oct 2024 0 repositories listed
-
28 Oct 2024 0 repositories listed Syntology 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Logarithmically Quantized Distributed Optimization over Dynamic Multi-Agent Networks27 Oct 2024 0 repositories listed
-
Unleashing Dynamic Range and Resolution in Unlimited Sensing Framework via Novel Hardware26 Oct 2024 0 repositories listed
-
You Never Know: Quantization Induces Inconsistent Biases in Vision-Language Foundation Models26 Oct 2024 0 repositories listed
-
A Survey of Small Language Models25 Oct 2024 0 repositories listed
-
Learning ID-free Item Representation with Token Crossing for Multimodal Recommendation25 Oct 2024 0 repositories listed
-
A Counterexample in Cross-Correlation Template Matching24 Oct 2024 0 repositories listed
-
Sliding DFT-based Signal Recovery for Modulo ADC with 1-bit Folding Information24 Oct 2024 0 repositories listed
-
TesseraQ: Ultra Low-Bit LLM Post-Training Quantization with Block Reconstruction24 Oct 2024 0 repositories listed
-
The Nature of Mathematical Modeling and Probabilistic Optimization Engineering in Generative AI24 Oct 2024 0 repositories listed
-
Adaptive Wireless Image Semantic Transmission: Design, Simulation, and Prototype Validation23 Oct 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.