Browse State-of-the-Art › Audio Compression
Audio Compression
17 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
17 shown of 17 papers with code (42 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
24 Oct 2022 6 repositories listedWe introduce a state-of-the-art real-time, high-fidelity, audio codec leveraging neural networks.
-
11 Jun 2023 4 repositories listed Syntology ran 27 of 38 samples · 11 unverifiedLanguage models have been successfully used to model natural signals, such as images, speech, and music.
-
26 Jul 2021 2 repositories listedDifferent from previous ASVspoof challenges, the LA task this year presents codec and transmission channel variability, while the new task DF presents general audio compression.
-
27 Nov 2024 1 repository listedNeural audio codecs (NACs) have garnered significant attention as key technologies for audio compression as well as audio representation for speech language models.
-
21 Oct 2024 1 repository listed Syntology ran 0 of 4 samples · 4 unverifiedKV cache compression methods have mainly relied on scalar quantization techniques to reduce the memory requirements during decoding.
-
18 Oct 2024 1 repository listed Syntology ran 3 of 3 samples · 0 unverifiedNeural audio codecs have recently gained popularity because they can represent audio signals with high fidelity at very low bitrates, making it feasible to use language modeling approaches for audio generation and…
-
30 Aug 2024 1 repository listed Syntology ran 0 of 2 samples · 2 unverifiedBy enhancing the semantic ability of the codec, X-Codec significantly reduces WER in speech synthesis tasks and extends these benefits to non-speech applications, including music and sound generation.
-
12 Aug 2024 1 repository listed Syntology ran 0 of 2 samples · 2 unverified · 2 pointer-only (licence)We demonstrate that Music2Latent outperforms existing continuous audio autoencoders in sound quality and reconstruction accuracy while achieving competitive performance on downstream MIR tasks using its latent…
-
19 Feb 2024 1 repository listed Syntology ran 4 of 7 samples · 3 unverifiedFurthermore, we also validate the efficiency of the Language-Codec on downstream speech language models.
-
22 Jun 2023 1 repository listed Syntology ran 5 of 6 samples · 1 unverifiedImplicit Neural Representations (INRs) have emerged as a promising method for representing diverse data modalities, including 3D shapes, images, and audio.
-
13 Jun 2023 1 repository listedSpatial audio quality is a highly multifaceted concept, with many interactions between environmental, geometrical, anatomical, psychological, and contextual considerations.
-
31 May 2023 1 repository listedWe demonstrate that the reference encoder learns better speaker-independent prosody when discrete code is utilized as input in the experiments.
-
30 May 2023 1 repository listed Syntology ran 6 of 7 samples · 1 unverifiedMany common types of data can be represented as functions that map coordinates to signal values, such as pixel locations to RGB values in the case of an image.
-
8 Aug 2021 1 repository listedWith active research in audio compression techniques yielding substantial breakthroughs, spectral reconstruction of low-quality audio waves remains a less indulged topic.
-
15 Mar 2021 1 repository listedThe onset of coronavirus disease 2019 (COVID-19), an infectious disease caused by severe acute respiratory syndrome coronavirus 2 (SARS-CoV-2), has sparked unprecedented change.
-
12 Jan 2021 1 repository listedWe present a deep convolutional GAN which leverages techniques from MP3/Vorbis audio compression to produce long, high-quality audio samples with long-range coherence.
-
9 Nov 2020 1 repository listedOur aim is to address the lack of a principled treatment of data acquired indistinctly in the temporal and frequency domains in a way that is robust to missing or noisy observations, and that at the same time models…
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections