Browse State-of-the-Art › Rhythm
Rhythm
137 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 137 papers with code (515 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
6 Jul 2017 7 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)We develop an algorithm which exceeds the performance of board certified cardiologists in detecting a wide range of heart arrhythmias from electrocardiograms recorded with a single-lead wearable monitor.
-
23 Apr 2020 6 repositories listed Syntology ran 0 of 12 samples · 12 unverifiedSpeech information can be roughly decomposed into four components: language content, timbre, pitch, and rhythm.
-
26 Oct 2019 5 repositories listedMellotron is a multispeaker voice synthesis model based on Tacotron 2 GST that can make a voice emote and sing without emotive or singing training data.
-
8 Aug 2021 4 repositories listedThe online estimation of rhythmic information, such as beat positions, downbeat positions, and meter, is critical for many real-time music applications.
-
9 Jun 2019 3 repositories listedAnalogy-making is a key method for computer algorithms to generate both natural and creative music pieces.
-
5 Jun 2025 2 repositories listed Syntology ran 4 of 6 samples · 2 unverified · 6 pointer-only (licence)To address this gap, we introduce MMSU, a comprehensive benchmark designed specifically for understanding and reasoning in spoken language.
-
21 Jul 2024 2 repositories listedHowever, textual prompts alone cannot precisely control temporal musical features such as chords and rhythm of the generated music.
-
20 Apr 2023 2 repositories listedWe introduce Cayley transform ellipsoid fitting (CTEF), an algorithm that uses the Cayley transform to fit ellipsoids to noisy data in any dimension.
-
18 Mar 2021 2 repositories listedIn this paper, we reformulate it by a two-stage process, ie, a key pose generation and then an in-between parametric motion curve prediction, where the key poses are easier to be synchronized with the music beats and…
-
9 Dec 2020 2 repositories listedConsidering the quasi-periodic characteristics of ECG signals, the dynamic features can be extracted from the TMF images with the transfer learning pre-trained convolutional neural network (CNN) models.
-
25 Nov 2020 2 repositories listedThis paper focuses on music generation, especially rhythm patterns of electronic dance music, and discusses if we can use deep learning to generate novel rhythms, interesting patterns not found in the training dataset.
-
4 Sep 2020 2 repositories listedIn this paper, we present an automatic gesture generation model that uses the multimodal context of speech text, audio, and speaker identity to reliably generate gestures.
-
28 Apr 2020 2 repositories listedElectrocardiography is a very common, non-invasive diagnostic procedure and its interpretation is increasingly supported by automatic interpretation algorithms.
-
1 Apr 2020 2 repositories listedThere has been significant progress in the music generation technique utilizing deep learning.
-
7 Oct 2019 2 repositories listedThe results suggest that the method proposed may be highly sensitive to detect rhythmic differences between groups in phase, amplitude and mesor.
-
17 Aug 2018 2 repositories listedAccess to electronic health record (EHR) data has motivated computational advances in medical research.
-
20 Jun 2025 1 repository listedDespite progress in controllable symbolic music generation, data scarcity remains a challenge for certain control modalities.
-
2 Jun 2025 1 repository listedAutomatic speech recognition (ASR) systems struggle with dysarthric speech due to high inter-speaker variability and slow speaking rates.
-
1 Jun 2025 1 repository listedThe fusion approach also shows improvement, with RFA spectrograms surpassing Mel spectrograms in classification by around a relative improvement of 13.
-
24 May 2025 1 repository listed Syntology ran 0 of 19 samples · 19 unverifiedLyrics translation requires both accurate semantic transfer and preservation of musical rhythm, syllabic structure, and poetic style.
-
5 May 2025 1 repository listed Syntology ran 7 of 11 samples · 4 unverifiedA voice AI agent that blends seamlessly into daily life would interact with humans in an autonomous, real-time, and emotionally expressive manner.
-
27 Apr 2025 1 repository listedOur study confirms that evolutionary optimization is an effective strategy for classifying small molecules regulating the circadian rhythm.
-
11 Apr 2025 1 repository listedDeep learning-based electrocardiogram (ECG) classification has shown impressive performance but clinical adoption has been slowed by the lack of transparent and faithful explanations.
-
16 Feb 2025 1 repository listedWe present ECG-Expert-QA, a comprehensive multimodal dataset for evaluating diagnostic capabilities in electrocardiogram (ECG) interpretation.
-
15 Feb 2025 1 repository listedAlthough previous ECG self-supervised learning (eSSL) methods have made significant progress in representation learning from unannotated ECG data, they typically treat ECG signals as ordinary time-series data,…
-
6 Feb 2025 1 repository listedThe model's iterative generation framework allows users to control the degree of style transfer and structural similarity to the original composition.
-
8 Jan 2025 1 repository listedVideo-based vehicle detection and counting play a critical role in managing transport infrastructure.
-
8 Jan 2025 1 repository listedRecent advancements in large language models (LLMs) have led to significant progress in text-based dialogue systems.
-
8 Jan 2025 1 repository listedVideo-based Automatic License Plate Recognition (ALPR) involves extracting vehicle license plate text information from video captures.
-
4 Jan 2025 1 repository listedAutomatic License Plate Recognition (ALPR) involves extracting vehicle license plate information from image or a video capture.
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections