Browse State-of-the-Art › Sound Classification
Sound Classification
61 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 61 papers with code (148 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
15 Aug 2016 5 repositories listedWe show that the improved performance stems from the combination of a deep, high-capacity model and an augmented training set: this combination outperforms both the proposed CNN without augmentation and a "shallow"…
-
24 Jun 2021 4 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedAudioCLIP achieves new state-of-the-art results in the Environmental Sound Classification (ESC) task, out-performing other approaches by reaching accuracies of 90.
-
1 Sep 2024 2 repositories listedAddressing this, we introduce the BUET Multi-disease Heart Sound (BMD-HS) dataset - a comprehensive and meticulously curated collection of heart sound recordings.
-
29 Oct 2021 2 repositories listedData-based and learning-based sound source localization (SSL) has shown promising results in challenging conditions, and is commonly set as a classification or a regression problem.
-
18 Apr 2019 2 repositories listedIn this paper, we present an end-to-end approach for environmental sound classification based on a 1D Convolution Neural Network (CNN) that learns a representation directly from the audio signal.
-
25 May 2018 2 repositories listedWe have evaluated the MCLNN performance using the Urbansound8k dataset of environmental sounds.
-
4 Jun 2025 1 repository listedAudio-text models are widely used in zero-shot environmental sound classification as they alleviate the need for annotated data.
-
3 Jun 2025 1 repository listedAutomated respiratory sound classification faces practical challenges from background noise and insufficient denoising in existing systems.
-
28 May 2025 1 repository listedIn this study, we explore soft labels for respiratory sound classification as an architecture-agnostic approach to distill an ensemble of teacher models into a student model.
-
28 May 2025 1 repository listedLung sound classification is vital for early diagnosis of respiratory diseases.
-
19 May 2025 1 repository listedThis requires a large amount of high-quality labeled data that is not publicly available.
-
11 Mar 2025 1 repository listedConvolutional Dictionary Learning (CDL) has emerged as a powerful approach for signal representation by learning translation-invariant features through convolution operations.
-
2 Feb 2025 1 repository listedAuscultation plays a pivotal role in early respiratory and pulmonary disease diagnosis.
-
4 Oct 2024 1 repository listedIn this dataset, we used a digital stethoscope to capture both heart and lung sounds, including individual and mixed recordings.
-
1 Oct 2024 1 repository listedWe compare a variety of both traditional and modern machine learning approaches to establish a baseline for the task of heterogeneous sound classification.
-
10 Jun 2024 1 repository listedRespiratory sound classification (RSC) is challenging due to varied acoustic signatures, primarily influenced by patient demographics and recording environments.
-
7 Jan 2024 1 repository listed Syntology ran 13 of 17 samples · 4 unverifiedAudio self-supervised learning (SSL) pre-training, which aims to learn good representations from unlabeled audio, has made remarkable progress.
-
25 Dec 2023 1 repository listedSelf-supervised learning (SSL) in audio holds significant potential across various domains, particularly in situations where abundant, unlabeled data is readily available at no cost.
-
15 Dec 2023 1 repository listedDespite the remarkable advances in deep learning technology, achieving satisfactory performance in lung sound classification remains a challenge due to the scarcity of available data.
-
16 Nov 2023 1 repository listedDeep neural networks have been applied to audio spectrograms for respiratory sound classification.
-
11 Nov 2023 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)In this work, we propose a straightforward approach to augment imbalanced respiratory sound data using an audio diffusion model as a conditional neural vocoder.
-
10 Jul 2023 1 repository listedAsthma is one of the most severe chronic respiratory diseases which can be diagnosed using several modalities, such as lung function test or spirometric measures, peak flow meter-based measures, sputum eosinophils,…
-
23 May 2023 1 repository listed Syntology ran 1 of 6 samples · 5 unverifiedRespiratory sound contains crucial information for the early diagnosis of fatal lung diseases.
-
7 Mar 2023 1 repository listedThis paper presents a context-aware framework for feature selection and classification procedures to realize a fast and accurate audio event annotation and classification.
-
15 Feb 2023 1 repository listedWe first showed that the segmentation of bird songs alone aggregated from 10% to 83% of label noise depending on the species.
-
14 Feb 2023 1 repository listedIn this work, we present a dataset of audio events called Subtitle-Aligned Movie Sounds (SAM-S).
-
1 Feb 2023 1 repository listedWe introduce Epic-Sounds, a large-scale dataset of audio annotations capturing temporal extents and class labels within the audio stream of the egocentric videos.
-
25 Nov 2022 1 repository listedFirst, masked data reconstruction is performed to learn modality-specific representations from audio and visual streams.
-
Effective Audio Classification Network Based on Paired Inverse Pyramid Structure and Dense MLP Block5 Nov 2022 1 repository listedRecently, massive architectures based on Convolutional Neural Network (CNN) and self-attention mechanisms have become necessary for audio classification.
-
27 Oct 2022 1 repository listedIn addition, when combining class labels with metadata using multiple supervised contrastive learning, an extension of supervised contrastive learning solving an additional task of grouping patients within the same sex…
Syntology lines on 4 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections