Methods › Audio › Text-to-Speech Models › Glow-TTS

Glow-TTS

7 papers tagged archive 2025-07-28

Introduced by Jaehyeon Kim et al. in Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search

archive 2025-07-28 Description, source and code snippet are the archive's method entry.

Glow-TTS is a flow-based generative model for parallel TTS that does not require any external aligner. By combining the properties of flows and dynamic programming, the proposed model searches for the most probable monotonic alignment between text and the latent representation of speech. The model is directly trained to maximize the log-likelihood of speech with the alignment. Enforcing hard monotonic alignments helps enable robust TTS, which generalizes to long utterances, and employing flows enables fast, diverse, and controllable speech synthesis.

PaperSource

Papers archive 2025-07-28

7 shown of 7, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.

Tasks archive 2025-07-28

12 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.

TaskPapers
Text to Speech5
text-to-speech5
Text-To-Speech Synthesis2
CPU1
Diversity1
GPU1
Speech Synthesis1
Variational Inference1
Vocal Bursts Intensity Prediction1
Voice Conversion1
Word Alignment1
Zero-Shot Multi-Speaker TTS1

Usage over time archive 2025-07-28

Papers per year tagged with Glow-TTS: 2020 to 2024, peak 2 2 0 2020: 1 paper 2020 2021: 2 papers 2021 2022: 1 paper 2022 2023: 2 papers 2023 2024: 1 paper 2024
Papers per year the archive tags with this method, by the paper's archive date (7 dated). Bars are counts, not a trend claim.

Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).

Categories archive 2025-07-28

Text-to-Speech Models

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections