Browse State-of-the-Art › Text-To-Speech Synthesis › Papers, page 3
Text-To-Speech Synthesis
Papers archive 2025-07-28
archive papers tagged: 332 · with a code link: 104 · where Syntology ran a sample: 37 (35 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (37 of 332 tagged: 35 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 3 of 4: papers 201 to 300 of 332, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
BERT, can HE predict contrastive focus? Predicting and controlling prominence in neural TTS using a language model4 Jul 2022 0 repositories listed
-
R-MelNet: Reduced Mel-Spectral Modeling for Neural TTS30 Jun 2022 0 repositories listed
-
BU-TTS: An Open-Source, Bilingual Welsh-English, Text-to-Speech Corpus1 Jun 2022 0 repositories listed
-
Exploring Transfer Learning for Urdu Speech Synthesis1 Jun 2022 0 repositories listed
-
Investigating Inter- and Intra-speaker Voice Conversion using Audiobooks1 Jun 2022 0 repositories listed
-
ReCAB-VAE: Gumbel-Softmax Variational Inference Based on Analytic Divergence9 May 2022 0 repositories listed
-
The PartialSpoof Database and Countermeasures for the Detection of Short Fake Speech Segments Embedded in an Utterance11 Apr 2022 0 repositories listed
-
6 Apr 2022 0 repositories listed
-
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature2 Apr 2022 0 repositories listed
-
Applying Syntax–Prosody Mapping Hypothesis and Prosodic Well-Formedness Constraints to Neural Sequence-to-Sequence Speech Synthesis29 Mar 2022 0 repositories listed
-
AutoTTS: End-to-End Text-to-Speech Synthesis through Differentiable Duration Modeling21 Mar 2022 0 repositories listed
-
Text-free non-parallel many-to-many voice conversion using normalising flows15 Mar 2022 0 repositories listed
-
Deep Performer: Score-to-Audio Music Performance Synthesis12 Feb 2022 0 repositories listed
-
Multi-Stage Deep Transfer Learning for EmIoT-enabled Human-Computer Interaction3 Feb 2022 0 repositories listed
-
Transformer-based Models of Text Normalization for Speech Applications1 Feb 2022 0 repositories listed
-
Multi-speaker Multi-style Text-to-speech Synthesis With Single-speaker Single-style Training Data Scenarios23 Dec 2021 0 repositories listed
-
Guided-TTS: A Diffusion Model for Text-to-Speech via Classifier Guidance23 Nov 2021 0 repositories listed
-
Environment Aware Text-to-Speech Synthesis8 Oct 2021 0 repositories listed
-
Prosody-TTS: An end-to-end speech synthesis system with prosody control6 Oct 2021 0 repositories listed
-
3 Oct 2021 0 repositories listed
-
Conditioning Sequence-to-sequence Networks with Learned Activations29 Sep 2021 0 repositories listed
-
Guided-TTS:Text-to-Speech with Untranscribed Speech29 Sep 2021 0 repositories listed
-
Low-Latency Incremental Text-to-Speech Synthesis with Distilled Context Prediction Network22 Sep 2021 0 repositories listed
-
A Unified Transformer-based Framework for Duplex Text Normalization23 Aug 2021 0 repositories listed
-
Location, Location: Enhancing the Evaluation of Text-to-Speech Synthesis Using the Rapid Prosody Transcription Paradigm6 Jul 2021 0 repositories listed
-
An objective evaluation of the effects of recording conditions and speaker characteristics in multi-speaker deep neural speech synthesis3 Jun 2021 0 repositories listed
-
Speaker verification-derived loss and data augmentation for DNN-based multispeaker speech synthesis3 Jun 2021 0 repositories listed
-
Dual Script E2E framework for Multilingual and Code-Switching ASR2 Jun 2021 0 repositories listed
-
Enhancing Word-Level Semantic Representation via Dependency Structure for Expressive Text-to-Speech Synthesis14 Apr 2021 0 repositories listed
-
Flavored Tacotron: Conditional Learning for Prosodic-linguistic Features8 Apr 2021 0 repositories listed
-
Reinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability3 Apr 2021 0 repositories listed
-
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS28 Mar 2021 0 repositories listed
-
Continual Speaker Adaptation for Text-to-Speech Synthesis26 Mar 2021 0 repositories listed
-
Alternate Endings: Improving Prosody for Incremental Neural TTS with Predicted Future Text Input19 Feb 2021 0 repositories listed
-
VARA-TTS: Non-Autoregressive Text-to-Speech Synthesis based on Very Deep VAE with Residual Attention12 Feb 2021 0 repositories listed
-
Voice Cloning: a Multi-Speaker Text-to-Speech Synthesis Approach based on Transfer Learning10 Feb 2021 0 repositories listed
-
Triple M: A Practical Text-to-speech Synthesis System With Multi-guidance Attention And Multi-band Multi-time LPCNet30 Jan 2021 0 repositories listed
-
Parallel WaveNet conditioned on VAE latent vectors17 Dec 2020 0 repositories listed
-
Using previous acoustic context to improve Text-to-Speech synthesis7 Dec 2020 0 repositories listed
-
Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS10 Nov 2020 0 repositories listed
-
Fine-grained Style Modeling, Transfer and Prediction in Text-to-Speech Synthesis via Phone-Level Content-Style Disentanglement8 Nov 2020 0 repositories listed
-
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework4 Nov 2020 0 repositories listed
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
GraphSpeech: Syntax-Aware Graph Attention Network For Neural Speech Synthesis23 Oct 2020 0 repositories listed
-
An Investigation of the Relation Between Grapheme Embeddings and Pronunciation for Tacotron-based Systems21 Oct 2020 0 repositories listed
-
End-to-End Text-to-Speech using Latent Duration based on VQ-VAE19 Oct 2020 0 repositories listed
-
Automatic Arabic Dialect Identification Systems for Written Texts: A Survey26 Sep 2020 0 repositories listed
-
Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis17 Sep 2020 0 repositories listed
-
Controllable neural text-to-speech synthesis using intuitive prosodic features14 Sep 2020 0 repositories listed
-
What the Future Brings: Investigating the Impact of Lookahead for Incremental Neural TTS4 Sep 2020 0 repositories listed
-
Voice Conversion by Cascading Automatic Speech Recognition and Text-to-Speech Synthesis with Prosody Transfer3 Sep 2020 0 repositories listed
-
Multi-speaker Text-to-speech Synthesis Using Deep Gaussian Processes7 Aug 2020 0 repositories listed
-
Normalizing Text using Language Modelling based on Phonetics and String Similarity25 Jun 2020 0 repositories listed
-
Investigation of learning abilities on linguistic features in sequence-to-sequence text-to-speech synthesis20 May 2020 0 repositories listed
-
Semi-supervised Learning for Multi-speaker Text-to-speech Synthesis Using Discrete Speech Representation16 May 2020 0 repositories listed
-
Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment: the Case of Gascon Occitan1 May 2020 0 repositories listed
-
Style Variation as a Vantage Point for Code-Switching1 May 2020 0 repositories listed
-
Using VAEs and Normalizing Flows for One-shot Text-To-Speech Synthesis of Expressive Speech28 Nov 2019 0 repositories listed
-
Cross-lingual Multi-speaker Text-to-speech Synthesis for Voice Cloning without Using Parallel Corpus for Unseen Speakers26 Nov 2019 0 repositories listed
-
A unified sequence-to-sequence front-end model for Mandarin text-to-speech synthesis11 Nov 2019 0 repositories listed
-
Incremental Text-to-Speech Synthesis with Prefix-to-Prefix Framework7 Nov 2019 0 repositories listed
-
Effect of choice of probability distribution, randomness, and search methods for alignment modeling in sequence-to-sequence text-to-speech synthesis using hard alignment28 Oct 2019 0 repositories listed
-
The Theory behind Controllable Expressive Speech Synthesis: a Cross-disciplinary Approach14 Oct 2019 0 repositories listed
-
Modular Meta-Learning with Shrinkage12 Sep 2019 0 repositories listed
-
Evaluating Long-form Text-to-Speech: Comparing the Ratings of Sentences and Paragraphs9 Sep 2019 0 repositories listed
-
Neural Harmonic-plus-Noise Waveform Model with Trainable Maximum Voice Frequency for Text-to-Speech Synthesis27 Aug 2019 0 repositories listed
-
Listening while Speaking and Visualizing: Improving ASR through Multimodal Chain3 Jun 2019 0 repositories listed
-
Neural Models of Text Normalization for Speech Applications1 Jun 2019 0 repositories listed
-
Neural Text Normalization with Subword Units1 Jun 2019 0 repositories listed
-
6 Apr 2019 0 repositories listed
-
Speech denoising by parametric resynthesis2 Apr 2019 0 repositories listed
-
Generative adversarial network-based glottal waveform model for statistical parametric speech synthesis14 Mar 2019 0 repositories listed
-
AttS2S-VC: Sequence-to-Sequence Voice Conversion with Attention and Context Preservation Mechanisms9 Nov 2018 0 repositories listed
-
End-to-End Feedback Loss in Speech Chain Framework via Straight-Through Estimator31 Oct 2018 0 repositories listed
-
Waveform generation for text-to-speech synthesis using pitch-synchronous multi-scale generative adversarial networks30 Oct 2018 0 repositories listed
-
Speaking style adaptation in Text-To-Speech synthesis using Sequence-to-sequence models with attention29 Oct 2018 0 repositories listed
-
A Challenge Set and Methods for Noun-Verb Ambiguity1 Oct 2018 0 repositories listed
-
Predicting Expressive Speaking Style From Text In End-To-End Speech Synthesis4 Aug 2018 0 repositories listed
-
Wasserstein GAN and Waveform Loss-based Acoustic Model Training for Multi-speaker Text-to-Speech Synthesis Systems Using a WaveNet Vocoder31 Jul 2018 0 repositories listed
-
Design and Development of Speech Corpora for Air Traffic Control Training1 May 2018 0 repositories listed
-
Improving homograph disambiguation with supervised machine learning1 May 2018 0 repositories listed
-
SynPaFlex-Corpus: An Expressive French Audiobooks Corpus dedicated to expressive speech synthesis.1 May 2018 0 repositories listed
-
Speaker-independent raw waveform model for glottal excitation25 Apr 2018 0 repositories listed
-
Machine Speech Chain with One-shot Speaker Adaptation28 Mar 2018 0 repositories listed
-
Creating New Language and Voice Components for the Updated MaryTTS Text-to-Speech Synthesis Platform13 Dec 2017 0 repositories listed
-
Refer-iTTS: A System for Referring in Spoken Installments to Objects in Real-World Images1 Sep 2017 0 repositories listed
-
Listening while Speaking: Speech Chain by Deep Learning16 Jul 2017 0 repositories listed
-
CASSANDRA: A multipurpose configurable voice-enabled human-computer-interface1 Apr 2017 0 repositories listed
-
Automatic Syllabification for Manipuri language1 Dec 2016 0 repositories listed
-
DNN-based Speech Synthesis for Indian Languages from ASCII text18 Aug 2016 0 repositories listed
-
A Taxonomy of Specific Problem Classes in Text-to-Speech Synthesis: Comparing Commercial and Open Source Performance1 May 2016 0 repositories listed
-
Minimally Supervised Number Normalization1 Jan 2016 0 repositories listed
-
Text Normalization and Unit Selection for a Memory Based Non Uniform Unit Selection TTS in Malayalam1 Dec 2015 0 repositories listed
-
Hierarchical Representation of Prosody for Statistical Speech Synthesis7 Oct 2015 0 repositories listed
-
A distributed cloud-based dialog system for conversational application development1 Sep 2015 0 repositories listed
-
Individuality-Preserving Spectrum Modification for Articulation Disorders Using Phone Selective Synthesis1 Sep 2015 0 repositories listed
-
Which Synthetic Voice Should I Choose for an Evocative Task?1 Sep 2015 0 repositories listed
-
Aligning Opinions: Cross-Lingual Opinion Mining with Dependencies1 Jul 2015 0 repositories listed
-
An In-depth Analysis of the Effect of Text Normalization in Social Media1 May 2015 0 repositories listed
-
Normalization of Non-Standard Words in Croatian Texts27 Mar 2015 0 repositories listed