{"url":"/task/singing-voice-synthesis","name":"Singing Voice Synthesis","slug":"singing-voice-synthesis","description_markdown":"(Verse 1)\r\nSa bawat hakbang, sa bawat daan\r\nMay pangarap kang naghihintay\r\nWestbridge ang gabay, sa iyong paglalakbay\r\nTungo sa kinabukasan, ng ating bayan\r\n\r\n\r\n(Chorus)\r\nLagi kang kasama, sa aking puso\r\nWestbridge Institute, ang nagbibigay ng liwanag\r\nSa bawat hamon, sa bawat pagsubok\r\nLagi kang kasama, sa aking puso\r\n\r\n\r\n(Verse 2)\r\nKami ang bagong henerasyon\r\nNa may pangarap, na may pag-asa\r\nWestbridge ang nagtuturo, ng mga kasanayan\r\nTungo sa pag-unlad, ng ating bansa\r\n\r\n\r\n(Chorus)\r\nLagi kang kasama, sa aking puso\r\nWestbridge Institute, ang nagbibigay ng liwanag\r\nSa bawat hamon, sa bawat pagsubok\r\nLagi kang kasama, sa aking puso\r\n\r\n\r\n(Bridge)\r\nTayo ay magkakaisa, sa pagtataguyod\r\nNg ating mga pangarap, ng ating mga adhikain\r\nWestbridge ang siyang, nagbibigay ng lakas\r\nTungo sa pag-abot, ng ating mga pangarap\r\n\r\n\r\n(Chorus)\r\nLagi kang kasama, sa aking puso\r\nWestbridge Institute, ang nagbibigay ng liwanag\r\nSa bawat hamon, sa bawat pagsubok\r\nLagi kang kasama, sa aking puso","categories":[{"name":"Music","url":"/area/music"},{"name":"Speech","url":"/area/speech"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"archive_url"},"counts":{"papers_tagged":81,"papers_with_code":28,"benchmarks":0,"benchmark_tables_in_archive":0,"benchmark_tables_shown":0,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":1,"subtasks":0,"parent_tasks":0},"benchmarks":[],"datasets":[{"url":"/dataset/gtsinger","name":"GTSinger","full_name":"GTSinger: A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks","num_papers_in_archive":6}],"subtasks":[],"parent_tasks":[],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":28,"of":28,"tagged_in_all":81,"items":[{"url":"/paper/diffsinger-diffusion-acoustic-model-for","title":"DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism","date":"2021-05-06","arxiv_id":"2105.02446","repositories_listed":10,"syntology":{"n":7,"n_ran":7,"n_unverified":0,"n_pointer_only":4}},{"url":"/paper/singing-voice-synthesis-using-differentiable","title":"Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables","date":"2023-06-29","arxiv_id":"2306.17252","repositories_listed":4,"syntology":null},{"url":"/paper/mlp-singer-towards-rapid-parallel-singing","title":"MLP Singer: Towards Rapid Parallel Singing Voice Synthesis","date":"2021-06-15","arxiv_id":null,"repositories_listed":3,"syntology":null},{"url":"/paper/ctrsvdd-a-benchmark-dataset-and-baseline","title":"CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection","date":"2024-06-04","arxiv_id":"2406.02438","repositories_listed":2,"syntology":null},{"url":"/paper/latent-optimal-paths-by-gumbel-propagation","title":"Latent Optimal Paths by Gumbel Propagation for Variational Bayesian Dynamic Programming","date":"2023-06-05","arxiv_id":"2306.02568","repositories_listed":2,"syntology":null},{"url":"/paper/nnsvs-a-neural-network-based-singing-voice","title":"NNSVS: A Neural Network-Based Singing Voice Synthesis Toolkit","date":"2022-10-28","arxiv_id":"2210.15987","repositories_listed":2,"syntology":null},{"url":"/paper/multi-singer-fast-multi-singer-singing-voice-1","title":"Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus","date":"2021-12-20","arxiv_id":"2112.10358","repositories_listed":2,"syntology":{"n":2,"n_ran":1,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/refinegan-universally-generating-waveform","title":"RefineGAN: Universally Generating Waveform Better than Ground Truth with Highly Accurate Pitch and Intensity Responses","date":"2021-11-01","arxiv_id":"2111.00962","repositories_listed":2,"syntology":null},{"url":"/paper/tcsinger-2-customizable-multilingual-zero","title":"TCSinger 2: Customizable Multilingual Zero-shot Singing Voice Synthesis","date":"2025-05-20","arxiv_id":"2505.14910","repositories_listed":1,"syntology":null},{"url":"/paper/stylesinger-2-zero-shot-singing-voice","title":"TCSinger: Zero-Shot Singing Voice Synthesis with Style Transfer and Multi-Level Style Control","date":"2024-09-24","arxiv_id":"2409.15977","repositories_listed":1,"syntology":null},{"url":"/paper/gtsinger-a-global-multi-technique-singing","title":"GTSinger: A Global Multi-Technique Singing Corpus with Realistic Music Scores for All Singing Tasks","date":"2024-09-20","arxiv_id":"2409.13832","repositories_listed":1,"syntology":{"n":7,"n_ran":6,"n_unverified":1,"n_pointer_only":7}},{"url":"/paper/svdd-2024-the-inaugural-singing-voice","title":"SVDD 2024: The Inaugural Singing Voice Deepfake Detection Challenge","date":"2024-08-28","arxiv_id":"2408.16132","repositories_listed":1,"syntology":{"n":3,"n_ran":3,"n_unverified":0,"n_pointer_only":3}},{"url":"/paper/prompt-singer-controllable-singing-voice","title":"Prompt-Singer: Controllable Singing-Voice-Synthesis with Natural Language Prompt","date":"2024-03-18","arxiv_id":"2403.11780","repositories_listed":1,"syntology":{"n":5,"n_ran":2,"n_unverified":3,"n_pointer_only":0}},{"url":"/paper/stylesinger-style-transfer-for-out-of-domain","title":"StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis","date":"2023-12-17","arxiv_id":"2312.10741","repositories_listed":1,"syntology":null},{"url":"/paper/bisinger-bilingual-singing-voice-synthesis","title":"BiSinger: Bilingual Singing Voice Synthesis","date":"2023-09-25","arxiv_id":"2309.14089","repositories_listed":1,"syntology":null},{"url":"/paper/singfake-singing-voice-deepfake-detection","title":"SingFake: Singing Voice Deepfake Detection","date":"2023-09-14","arxiv_id":"2309.07525","repositories_listed":1,"syntology":{"n":16,"n_ran":15,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/fsd-an-initial-chinese-dataset-for-fake-song","title":"FSD: An Initial Chinese Dataset for Fake Song Detection","date":"2023-09-05","arxiv_id":"2309.02232","repositories_listed":1,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":2}},{"url":"/paper/rmssinger-realistic-music-score-based-singing","title":"RMSSinger: Realistic-Music-Score based Singing Voice Synthesis","date":"2023-05-18","arxiv_id":"2305.10686","repositories_listed":1,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/comospeech-one-step-speech-and-singing-voice","title":"CoMoSpeech: One-Step Speech and Singing Voice Synthesis via Consistency Model","date":"2023-05-11","arxiv_id":"2305.06908","repositories_listed":1,"syntology":{"n":9,"n_ran":4,"n_unverified":5,"n_pointer_only":0}},{"url":"/paper/cross-domain-neural-pitch-and-periodicity","title":"Cross-domain Neural Pitch and Periodicity Estimation","date":"2023-01-28","arxiv_id":"2301.12258","repositories_listed":1,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/m4singer-a-multi-style-multi-singer-and","title":"M4Singer: a Multi-Style, Multi-Singer and Musical Score Provided Mandarin Singing Corpus","date":"2022-12-29","arxiv_id":null,"repositories_listed":1,"syntology":null},{"url":"/paper/xiaoicesing-2-a-high-fidelity-singing-voice","title":"Xiaoicesing 2: A High-Fidelity Singing Voice Synthesizer Based on Generative Adversarial Network","date":"2022-10-26","arxiv_id":"2210.14666","repositories_listed":1,"syntology":null},{"url":"/paper/hifi-wavegan-generative-adversarial-network","title":"HiFi-WaveGAN: Generative Adversarial Network with Auxiliary Spectrogram-Phase Loss for High-Fidelity Singing Voice Generation","date":"2022-10-23","arxiv_id":"2210.12740","repositories_listed":1,"syntology":null},{"url":"/paper/sinsy-a-deep-neural-network-based-singing","title":"Sinsy: A Deep Neural Network-Based Singing Voice Synthesis System","date":"2021-08-05","arxiv_id":"2108.02776","repositories_listed":1,"syntology":null},{"url":"/paper/latent-space-explorations-of-singing-voice","title":"Latent Space Explorations of Singing Voice Synthesis using DDSP","date":"2021-03-12","arxiv_id":"2103.07197","repositories_listed":1,"syntology":null},{"url":"/paper/sequence-to-sequence-singing-voice-synthesis","title":"Sequence-to-sequence Singing Voice Synthesis with Perceptual Entropy Loss","date":"2020-10-22","arxiv_id":"2010.12024","repositories_listed":1,"syntology":null},{"url":"/paper/hifisinger-towards-high-fidelity-neural","title":"HiFiSinger: Towards High-Fidelity Neural Singing Voice Synthesis","date":"2020-09-03","arxiv_id":"2009.01776","repositories_listed":1,"syntology":{"n":10,"n_ran":6,"n_unverified":4,"n_pointer_only":0}},{"url":"/paper/score-and-lyrics-free-singing-voice-1","title":"Score and Lyrics-Free Singing Voice Generation","date":"2019-12-26","arxiv_id":"1912.11747","repositories_listed":1,"syntology":null}],"syntology_records":11,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-25T09:33:49+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}