Browse State-of-the-Art › Text to Speech › Papers, page 13
Text to Speech
Papers archive 2025-07-28
archive papers tagged: 1,419 · with a code link: 399 · where Syntology ran a sample: 108 (96 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (108 of 1,419 tagged: 96 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument)
Page 13 of 15: papers 1,201 to 1,300 of 1,419, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
Prosodic Representation Learning and Contextual Sampling for Neural Text-to-Speech4 Nov 2020 0 repositories listed
-
Training Wake Word Detection with Synthesized Speech Data on Confusion Words3 Nov 2020 0 repositories listed
-
Learning to Maximize Speech Quality Directly Using MOS Prediction for Neural Text-to-Speech2 Nov 2020 0 repositories listed
-
Learning from Explanations and Demonstrations: A Pilot Study1 Nov 2020 0 repositories listed
-
DeviceTTS: A Small-Footprint, Fast, Stable Network for On-Device Text-to-Speech29 Oct 2020 0 repositories listed
-
Effective Decoder Masking for Transformer Based End-to-End Speech Recognition27 Oct 2020 0 repositories listed
-
Parallel waveform synthesis based on generative adversarial networks with voicing-aware conditional discriminators27 Oct 2020 0 repositories listed
-
Emotion controllable speech synthesis using emotion-unlabeled dataset with the assistance of cross-domain speech emotion recognition26 Oct 2020 0 repositories listed
-
GraphSpeech: Syntax-Aware Graph Attention Network For Neural Speech Synthesis23 Oct 2020 0 repositories listed
-
NU-GAN: High resolution neural upsampling with GAN22 Oct 2020 0 repositories listed
-
The NTU-AISG Text-to-speech System for Blizzard Challenge 202022 Oct 2020 0 repositories listed
-
A Mask-based Model for Mandarin Chinese Polyphone Disambiguation21 Oct 2020 0 repositories listed
-
An Investigation of the Relation Between Grapheme Embeddings and Pronunciation for Tacotron-based Systems21 Oct 2020 0 repositories listed
-
Replacing Human Audio with Synthetic Audio for On-device Unspoken Punctuation Prediction20 Oct 2020 0 repositories listed
-
End-to-End Text-to-Speech using Latent Duration based on VQ-VAE19 Oct 2020 0 repositories listed
-
Improving Low Resource Code-switched ASR using Augmented Code-switched TTS12 Oct 2020 0 repositories listed
-
Latent linguistic embedding for cross-lingual text-to-speech and voice conversion8 Oct 2020 0 repositories listed
-
Leveraging Unpaired Text Data for Training End-to-End Speech-to-Intent Systems8 Oct 2020 0 repositories listed
-
Neural Speech Synthesis for Estonian6 Oct 2020 0 repositories listed
-
Compress Polyphone Pronunciation Prediction Model with Shared Labels1 Oct 2020 0 repositories listed
-
Automatic Arabic Dialect Identification Systems for Written Texts: A Survey26 Sep 2020 0 repositories listed
-
Hierarchical Multi-Grained Generative Model for Expressive Speech Synthesis17 Sep 2020 0 repositories listed
-
Controllable neural text-to-speech synthesis using intuitive prosodic features14 Sep 2020 0 repositories listed
-
What the Future Brings: Investigating the Impact of Lookahead for Incremental Neural TTS4 Sep 2020 0 repositories listed
-
Voice Conversion by Cascading Automatic Speech Recognition and Text-to-Speech Synthesis with Prosody Transfer3 Sep 2020 0 repositories listed
-
Textual Echo Cancellation13 Aug 2020 0 repositories listed
-
Bunched LPCNet : Vocoder for Low-cost Neural Text-To-Speech Systems11 Aug 2020 0 repositories listed
-
Unsupervised Learning For Sequence-to-sequence Text-to-speech For Low-resource Languages11 Aug 2020 0 repositories listed
-
LRSpeech: Extremely Low-Resource Speech Synthesis and Recognition9 Aug 2020 0 repositories listed
-
Incremental Text to Speech for Neural Sequence-to-Sequence Models using Reinforcement Learning7 Aug 2020 0 repositories listed
-
Multi-speaker Text-to-speech Synthesis Using Deep Gaussian Processes7 Aug 2020 0 repositories listed
-
Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability30 Jul 2020 0 repositories listed
-
A Transfer Learning End-to-End ArabicText-To-Speech (TTS) Deep Architecture22 Jul 2020 0 repositories listed
-
Normalizing Text using Language Modelling based on Phonetics and String Similarity25 Jun 2020 0 repositories listed
-
Generic Indic Text-to-speech Synthesisers with Rapid Adaptation in an End-to-end Framework12 Jun 2020 0 repositories listed
-
NAUTILUS: a Versatile Voice Cloning System22 May 2020 0 repositories listed
-
Cross-lingual Multispeaker Text-to-Speech under Limited-Data Scenario21 May 2020 0 repositories listed
-
Investigation of learning abilities on linguistic features in sequence-to-sequence text-to-speech synthesis20 May 2020 0 repositories listed
-
Improving Accent Conversion with Reference Encoder and End-To-End Text-To-Speech19 May 2020 0 repositories listed
-
Knowledge-and-Data-Driven Amplitude Spectrum Prediction for Hierarchical Neural Vocoders18 May 2020 0 repositories listed
-
Semi-supervised Learning for Multi-speaker Text-to-speech Synthesis Using Discrete Speech Representation16 May 2020 0 repositories listed
-
JDI-T: Jointly trained Duration Informed Transformer for Text-To-Speech without Explicit Alignment15 May 2020 0 repositories listed
-
You Do Not Need More Data: Improving End-To-End Speech Recognition by Text-To-Speech Data Augmentation14 May 2020 0 repositories listed
-
AdaDurIAN: Few-shot Adaptation for Neural Text-to-Speech with DurIAN12 May 2020 0 repositories listed
-
DiscreTalk: Text-to-Speech as a Machine Translation Problem12 May 2020 0 repositories listed
-
Burmese Speech Corpus, Finite-State Text Normalization and Pronunciation Grammars with an Application to Text-to-Speech1 May 2020 0 repositories listed
-
Corpus Generation for Voice Command in Smart Home and the Effect of Speech Synthesis on End-to-End SLU1 May 2020 0 repositories listed
-
Crowdsourcing Latin American Spanish for Low-Resource Text-to-Speech1 May 2020 0 repositories listed
-
Development and Evaluation of Speech Synthesis Corpora for Latvian1 May 2020 0 repositories listed
-
IndicSpeech: Text-to-Speech Corpus for Indian Languages1 May 2020 0 repositories listed
-
Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment: the Case of Gascon Occitan1 May 2020 0 repositories listed
-
1 May 2020 0 repositories listed
-
Open-source Multi-speaker Speech Corpora for Building Gujarati, Kannada, Malayalam, Marathi, Tamil and Telugu Speech Synthesis Systems1 May 2020 0 repositories listed
-
Style Variation as a Vantage Point for Code-Switching1 May 2020 0 repositories listed
-
CopyCat: Many-to-Many Fine-Grained Prosody Transfer for Neural Text-to-Speech30 Apr 2020 0 repositories listed
-
A Study of Non-autoregressive Model for Sequence Generation22 Apr 2020 0 repositories listed
-
Data Processing for Optimizing Naturalness of Vietnamese Text-to-speech System20 Apr 2020 0 repositories listed
-
Generating Multilingual Voices Using Speaker Space Translation Based on Bilingual Speaker Data10 Apr 2020 0 repositories listed
-
Scalable Multilingual Frontend for TTS10 Apr 2020 0 repositories listed
-
Improving Readability for Automatic Speech Recognition Transcription9 Apr 2020 0 repositories listed
-
Statistical Context-Dependent Units Boundary Correction for Corpus-based Unit-Selection Text-to-Speech5 Mar 2020 0 repositories listed
-
GraphTTS: graph-to-sequence modelling in neural text-to-speech4 Mar 2020 0 repositories listed
-
Fully-hierarchical fine-grained prosody modeling for interpretable speech synthesis6 Feb 2020 0 repositories listed
-
Generating diverse and natural text-to-speech samples using a quantized fine-grained VAE and auto-regressive prosody prior6 Feb 2020 0 repositories listed
-
BOFFIN TTS: Few-Shot Speaker Adaptation by Bayesian Optimization4 Feb 2020 0 repositories listed
-
2 Feb 2020 0 repositories listed
-
From Speech-to-Speech Translation to Automatic Dubbing19 Jan 2020 0 repositories listed
-
Parallel Neural Text-to-Speech1 Jan 2020 0 repositories listed
-
Smart Summarizer for Blind People1 Jan 2020 0 repositories listed
-
Singing Synthesis: with a little help from my attention12 Dec 2019 0 repositories listed
-
Towards Robust Neural Vocoding for Speech Generation: A Survey5 Dec 2019 0 repositories listed
-
Dynamic Prosody Generation for Speech Synthesis using Linguistics-Driven Acoustic Embedding Selection2 Dec 2019 0 repositories listed
-
Using VAEs and Normalizing Flows for One-shot Text-To-Speech Synthesis of Expressive Speech28 Nov 2019 0 repositories listed
-
Cross-lingual Multi-speaker Text-to-speech Synthesis for Voice Cloning without Using Parallel Corpus for Unseen Speakers26 Nov 2019 0 repositories listed
-
Prosody Transfer in Neural Text to Speech Using Global Pitch and Loudness Features21 Nov 2019 0 repositories listed
-
A unified sequence-to-sequence front-end model for Mandarin text-to-speech synthesis11 Nov 2019 0 repositories listed
-
Incremental Text-to-Speech Synthesis with Prefix-to-Prefix Framework7 Nov 2019 0 repositories listed
-
Teacher-Student Training for Robust Tacotron-based TTS7 Nov 2019 0 repositories listed
-
A System for Diacritizing Four Varieties of Arabic1 Nov 2019 0 repositories listed
-
Effect of choice of probability distribution, randomness, and search methods for alignment modeling in sequence-to-sequence text-to-speech synthesis using hard alignment28 Oct 2019 0 repositories listed
-
Unsupervised pre-training for sequence to sequence speech recognition28 Oct 2019 0 repositories listed
-
Multi-Reference Neural TTS Stylization with Adversarial Cycle Consistency25 Oct 2019 0 repositories listed
-
G2G: TTS-Driven Pronunciation Learning for Graphemic Hybrid ASR22 Oct 2019 0 repositories listed
-
The Theory behind Controllable Expressive Speech Synthesis: a Cross-disciplinary Approach14 Oct 2019 0 repositories listed
-
Semi-Supervised Generative Modeling for Controllable Speech Synthesis3 Oct 2019 0 repositories listed
-
Bootstrapping non-parallel voice conversion from speaker-adaptive text-to-speech14 Sep 2019 0 repositories listed
-
Modular Meta-Learning with Shrinkage12 Sep 2019 0 repositories listed
-
Evaluating Long-form Text-to-Speech: Comparing the Ratings of Sentences and Paragraphs9 Sep 2019 0 repositories listed
-
Neural Network-Based Modeling of Phonetic Durations6 Sep 2019 0 repositories listed
-
A Large-Scale User Study of an Alexa Prize Chatbot: Effect of TTS Dynamism on Perceived Quality of Social Dialog1 Sep 2019 0 repositories listed
-
Initial investigation of an encoder-decoder end-to-end TTS framework using marginalization of monotonic hard latent alignments30 Aug 2019 0 repositories listed
-
Neural Harmonic-plus-Noise Waveform Model with Trainable Maximum Voice Frequency for Text-to-Speech Synthesis27 Aug 2019 0 repositories listed
-
From Text to Sound: A Preliminary Study on Retrieving Sound Effects to Radio Stories20 Aug 2019 0 repositories listed
-
Hierarchical Sequence to Sequence Voice Conversion with Limited Data15 Jul 2019 0 repositories listed
-
M3D-GAN: Multi-Modal Multi-Domain Translation with Universal Attention9 Jul 2019 0 repositories listed
-
A Methodology for Controlling the Emotional Expressiveness in Synthetic Speech -- a Deep Learning approach5 Jul 2019 0 repositories listed
-
A Novel Approach to OCR using Image Recognition based Classification for Ancient Tamil Inscriptions in Temples4 Jul 2019 0 repositories listed
-
Fine-grained robust prosody transfer for single-speaker neural text-to-speech4 Jul 2019 0 repositories listed
-
Polyphone Disambiguation for Mandarin Chinese Using Conditional Neural Network with Multi-level Embedding Features3 Jul 2019 0 repositories listed