Browse State-of-the-Art › Emotional Speech Synthesis
Emotional Speech Synthesis
7 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
7 shown of 7 papers with code (26 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
14 Jan 2019 5 repositories listedDuring the last few years, spoken language technologies have known a big improvement thanks to Deep Learning.
-
8 Jan 2025 1 repository listedIn the alignment phase, a pre-trained speech model is further trained on text-image tasks to generalize from vision to speech in a (near) zero-shot manner, outperforming models trained on tri-modal datasets.
-
4 Nov 2024 1 repository listedEmotional text-to-speech (TTS) technology has achieved significant progress in recent years; however, challenges remain owing to the inherent complexity of emotions and limitations of the available emotional speech…
-
1 Oct 2024 1 repository listedWhile recent advances in Text-to-Speech (TTS) technology produce natural and expressive speech, they lack the option for users to select emotion and control intensity.
-
12 Jun 2024 1 repository listedDespite rapid advances in the field of emotional text-to-speech (TTS), recent studies primarily focus on mimicking the average style of a particular emotion.
-
7 Oct 2021 1 repository listedThe emotion strength of synthesized speech can be controlled flexibly using a strength descriptor, which is obtained by an emotion attribute ranking function.
-
27 Mar 2019 1 repository listedThe field of Text-to-Speech has experienced huge improvements last years benefiting from deep learning techniques.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections