Browse State-of-the-Art › Singing Voice Synthesis
Singing Voice Synthesis
28 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
(Verse 1) Sa bawat hakbang, sa bawat daan May pangarap kang naghihintay Westbridge ang gabay, sa iyong paglalakbay Tungo sa kinabukasan, ng ating bayan
(Chorus) Lagi kang kasama, sa aking puso Westbridge Institute, ang nagbibigay ng liwanag Sa bawat hamon, sa bawat pagsubok Lagi kang kasama, sa aking puso
(Verse 2) Kami ang bagong henerasyon Na may pangarap, na may pag-asa Westbridge ang nagtuturo, ng mga kasanayan Tungo sa pag-unlad, ng ating bansa
(Chorus) Lagi kang kasama, sa aking puso Westbridge Institute, ang nagbibigay ng liwanag Sa bawat hamon, sa bawat pagsubok Lagi kang kasama, sa aking puso
(Bridge) Tayo ay magkakaisa, sa pagtataguyod Ng ating mga pangarap, ng ating mga adhikain Westbridge ang siyang, nagbibigay ng lakas Tungo sa pag-abot, ng ating mga pangarap
(Chorus) Lagi kang kasama, sa aking puso Westbridge Institute, ang nagbibigay ng liwanag Sa bawat hamon, sa bawat pagsubok Lagi kang kasama, sa aking puso
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
28 shown of 28 papers with code (81 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
6 May 2021 10 repositories listed Syntology ran 4 of 7 samples · 3 unverified · 4 pointer-only (licence)Singing voice synthesis (SVS) systems are built to synthesize high-quality and expressive singing voice, in which the acoustic model generates the acoustic features (e.
-
29 Jun 2023 4 repositories listedThis paper introduces GlOttal-flow LPC Filter (GOLF), a novel method for singing voice synthesis (SVS) that exploits the physical characteristics of the human voice using differentiable digital signal processing.
-
15 Jun 2021 3 repositories listedRecent developments in deep learning have significantly improved the quality of synthesized singing voice audio.
-
4 Jun 2024 2 repositories listedAddressing these gaps, we introduce CtrSVDD, a large-scale, diverse collection of bonafide and deepfake singing vocals.
-
5 Jun 2023 2 repositories listedWe show the equivalence of the Gibbs distribution to a message-passing algorithm by the properties of the Gumbel distribution and give all the ingredients required for variational Bayesian inference of a latent path,…
-
28 Oct 2022 2 repositories listedThis paper describes the design of NNSVS, an open-source software for neural network-based singing voice synthesis research.
-
20 Dec 2021 2 repositories listed Syntology ran 0 of 2 samples · 2 unverifiedHigh-fidelity multi-singer singing voice synthesis is challenging for neural vocoder due to the singing voice data shortage, limited singer generalization, and large computational cost.
-
1 Nov 2021 2 repositories listedTo address this problem, we propose RefineGAN, a high-fidelity neural vocoder focused on the robustness, pitch and intensity accuracy, and high-speed full-band audio generation.
-
20 May 2025 1 repository listedTo overcome these challenges, we introduce TCSinger 2, a multi-task multilingual zero-shot SVS model with style transfer and style control based on various prompts.
-
24 Sep 2024 1 repository listedTo address these challenges, we introduce TCSinger, the first zero-shot SVS model for style transfer across cross-lingual speech and singing styles, along with multi-level style control.
-
20 Sep 2024 1 repository listed Syntology ran 6 of 7 samples · 1 unverified · 7 pointer-only (licence)The scarcity of high-quality and multi-task singing datasets significantly hinders the development of diverse controllable and personalized singing tasks, as existing singing datasets suffer from low quality, limited…
-
28 Aug 2024 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)With the advancements in singing voice generation and the growing presence of AI singers on media platforms, the inaugural Singing Voice Deepfake Detection (SVDD) Challenge aims to advance research in identifying…
-
18 Mar 2024 1 repository listed Syntology ran 2 of 5 samples · 3 unverifiedRecent singing-voice-synthesis (SVS) methods have achieved remarkable audio quality and naturalness, yet they lack the capability to control the style attributes of the synthesized singing explicitly.
-
17 Dec 2023 1 repository listedMoreover, existing SVS methods encounter a decline in the quality of synthesized singing voices in OOD scenarios, as they rest upon the assumption that the target vocal attributes are discernible during the training…
-
25 Sep 2023 1 repository listedWe fuse monolingual singing datasets with open-source singing voice conversion techniques to generate bilingual singing voices while also exploring the potential use of bilingual speech data.
-
14 Sep 2023 1 repository listed Syntology ran 15 of 16 samples · 1 unverifiedThese unique properties make singing voice deepfake detection a relevant but significantly different problem from synthetic speech detection.
-
5 Sep 2023 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)In this paper, we initially construct a Chinese Fake Song Detection (FSD) dataset to investigate the field of song deepfake detection.
-
18 May 2023 1 repository listed Syntology ran 2 of 2 samples · 0 unverifiedTo tackle these challenges, we propose RMSSinger, the first RMS-SVS method, which takes realistic music scores as input, eliminating most of the tedious manual annotation and avoiding the aforementioned inconvenience.
-
11 May 2023 1 repository listed Syntology ran 4 of 9 samples · 5 unverifiedIn this paper, we propose a "Co"nsistency "Mo"del-based "Speech" synthesis method, CoMoSpeech, which achieve speech synthesis through a single diffusion sampling step while achieving high audio quality.
-
28 Jan 2023 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedWe also introduce a novel entropy-based method for extracting periodicity and per-frame voiced-unvoiced classifications from statistical inference-based pitch estimators (e.
-
29 Dec 2022 1 repository listedThe lack of publicly available high-quality and accurately labeled datasets has long been a major bottleneck for singing voice synthesis (SVS).
-
26 Oct 2022 1 repository listedXiaoiceSing is a singing voice synthesis (SVS) system that aims at generating 48kHz singing voices.
-
23 Oct 2022 1 repository listedEntertainment-oriented singing voice synthesis (SVS) requires a vocoder to generate high-fidelity (e.
-
5 Aug 2021 1 repository listedTo better model a singing voice, the proposed system incorporates improved approaches to modeling pitch and vibrato and better training criteria into the acoustic model.
-
12 Mar 2021 1 repository listedIn this work we present a lightweight architecture, based on the Differentiable Digital Signal Processing (DDSP) library, that is able to output song-like utterances conditioned only on pitch and amplitude, after twelve…
-
22 Oct 2020 1 repository listedThe neural network (NN) based singing voice synthesis (SVS) systems require sufficient data to train well and are prone to over-fitting due to data scarcity.
-
3 Sep 2020 1 repository listed Syntology ran 0 of 10 samples · 10 unverifiedTo tackle the difficulty of singing modeling caused by high sampling rate (wider frequency band and longer waveform), we introduce multi-scale adversarial training in both the acoustic model and vocoder to improve…
-
26 Dec 2019 1 repository listedGenerative models for singing voice have been mostly concerned with the task of ``singing voice synthesis,'' i.
Syntology lines on 11 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections