Browse State-of-the-Art › text-to-speech › Papers, page 12
text-to-speech
Papers archive 2025-07-28
archive papers tagged: 1,413 · with a code link: 395 · where Syntology ran a sample: 106 (95 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (106 of 1,413 tagged: 95 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 12 of 15: papers 1,101 to 1,200 of 1,413, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
3 Oct 2021 0 repositories listed
-
Incorporating speaker embedding and post-filter network for improving speaker similarity of personalized speech synthesis system1 Oct 2021 0 repositories listed
-
Conditioning Sequence-to-sequence Networks with Learned Activations29 Sep 2021 0 repositories listed
-
Guided-TTS:Text-to-Speech with Untranscribed Speech29 Sep 2021 0 repositories listed
-
FlowVocoder: A small Footprint Neural Vocoder based Normalizing flow for Speech Synthesis27 Sep 2021 0 repositories listed
-
A Proposal of Automatic Error Correction in Text24 Sep 2021 0 repositories listed
-
Low-Latency Incremental Text-to-Speech Synthesis with Distilled Context Prediction Network22 Sep 2021 0 repositories listed
-
On-device neural speech synthesis17 Sep 2021 0 repositories listed
-
Referee: Towards reference-free cross-speaker style transfer with low-quality data for expressive speech synthesis8 Sep 2021 0 repositories listed
-
A Unified Transformer-based Framework for Duplex Text Normalization23 Aug 2021 0 repositories listed
-
Fighting Game Commentator with Pitch and Loudness Adjustment Utilizing Highlight Cues18 Aug 2021 0 repositories listed
-
GC-TTS: Few-shot Speaker Adaptation with Geometric Constraints16 Aug 2021 0 repositories listed
-
Enhancing audio quality for expressive Neural Text-to-Speech13 Aug 2021 0 repositories listed
-
RW-Resnet: A Novel Speech Anti-Spoofing Model Using Raw Waveform12 Aug 2021 0 repositories listed
-
AnyoneNet: Synchronized Speech and Talking Head Generation for Arbitrary Person9 Aug 2021 0 repositories listed
-
A Speech-enabled Fixed-phrase Translator for Healthcare Accessibility1 Aug 2021 0 repositories listed
-
A Survey on Audio Synthesis and Audio-Visual Multimodal Processing1 Aug 2021 0 repositories listed
-
BTS: Back TranScription for Speech-to-Text Post-Processor using Text-to-Speech-to-Text1 Aug 2021 0 repositories listed
-
Cross-speaker Style Transfer with Prosody Bottleneck in Neural Speech Synthesis27 Jul 2021 0 repositories listed
-
Digital Einstein Experience: Fast Text-to-Speech for Conversational AI21 Jul 2021 0 repositories listed
-
On Prosody Modeling for ASR+TTS based Voice Conversion20 Jul 2021 0 repositories listed
-
Federated Learning with Dynamic Transformer for Text to Speech9 Jul 2021 0 repositories listed
-
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style6 Jul 2021 0 repositories listed
-
Location, Location: Enhancing the Evaluation of Text-to-Speech Synthesis Using the Rapid Prosody Transcription Paradigm6 Jul 2021 0 repositories listed
-
GANSpeech: Adversarial Training for High-Fidelity Multi-Speaker Speech Synthesis29 Jun 2021 0 repositories listed
-
Hierarchical Context-Aware Transformers for Non-Autoregressive Text to Speech29 Jun 2021 0 repositories listed
-
Multi-Scale Spectrogram Modelling for Neural Text-to-Speech29 Jun 2021 0 repositories listed
-
Non-Autoregressive TTS with Explicit Duration Modelling for Low-Resource Highly Expressive Speech24 Jun 2021 0 repositories listed
-
Non-native English lexicon creation for bilingual speech synthesis21 Jun 2021 0 repositories listed
-
Advances in Speech Vocoding for Text-to-Speech with Continuous Parameters19 Jun 2021 0 repositories listed
-
17 Jun 2021 0 repositories listed
-
Improving the expressiveness of neural vocoding with non-affine Normalizing Flows16 Jun 2021 0 repositories listed
-
ADEPT: A Dataset for Evaluating Prosody Transfer15 Jun 2021 0 repositories listed
-
Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis15 Jun 2021 0 repositories listed
-
A learned conditional prior for the VAE acoustic space of a TTS system14 Jun 2021 0 repositories listed
-
SynthASR: Unlocking Synthetic Data for Speech Recognition14 Jun 2021 0 repositories listed
-
Improving multi-speaker TTS prosody variance with a residual encoder and normalizing flows10 Jun 2021 0 repositories listed
-
Speech BERT Embedding For Improving Prosody in Neural TTS8 Jun 2021 0 repositories listed
-
Data Augmentation Methods for End-to-end Speech Recognition on Distant-Talk Scenarios7 Jun 2021 0 repositories listed
-
Reinforce-Aligner: Reinforcement Alignment Search for Robust End-to-End Text-to-Speech5 Jun 2021 0 repositories listed
-
An objective evaluation of the effects of recording conditions and speaker characteristics in multi-speaker deep neural speech synthesis3 Jun 2021 0 repositories listed
-
Speaker verification-derived loss and data augmentation for DNN-based multispeaker speech synthesis3 Jun 2021 0 repositories listed
-
Dual Script E2E framework for Multilingual and Code-Switching ASR2 Jun 2021 0 repositories listed
-
21 May 2021 0 repositories listed
-
Learning Robust Latent Representations for Controllable Speech Synthesis10 May 2021 0 repositories listed
-
Talrómur: A large Icelandic TTS corpus1 May 2021 0 repositories listed
-
On Addressing Practical Challenges for RNN-Transducer27 Apr 2021 0 repositories listed
-
Enhancing Word-Level Semantic Representation via Dependency Structure for Expressive Text-to-Speech Synthesis14 Apr 2021 0 repositories listed
-
Non-autoregressive sequence-to-sequence voice conversion14 Apr 2021 0 repositories listed
-
Comparing the Benefit of Synthetic Training Data for Various Automatic Speech Recognition Architectures12 Apr 2021 0 repositories listed
-
Exploring Machine Speech Chain for Domain Adaptation and Few-Shot Speaker Adaptation8 Apr 2021 0 repositories listed
-
Flavored Tacotron: Conditional Learning for Prosodic-linguistic Features8 Apr 2021 0 repositories listed
-
Grapheme-to-Phoneme Transformer Model for Transfer Learning Dialects8 Apr 2021 0 repositories listed
-
Hi-Fi Multi-Speaker English TTS Dataset3 Apr 2021 0 repositories listed
-
Reinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability3 Apr 2021 0 repositories listed
-
Expressive Text-to-Speech using Style Tag1 Apr 2021 0 repositories listed
-
Fast DCTTS: Efficient Deep Convolutional Text-to-Speech1 Apr 2021 0 repositories listed
-
Multi-rate attention architecture for fast streamable Text-to-speech spectrum modeling1 Apr 2021 0 repositories listed
-
Continual Speaker Adaptation for Text-to-Speech Synthesis26 Mar 2021 0 repositories listed
-
GAN Vocoder: Multi-Resolution Discriminator Is All You Need9 Mar 2021 0 repositories listed
-
A Neural Text-to-Speech Model Utilizing Broadcast Data Mixed with Background Music4 Mar 2021 0 repositories listed
-
Model architectures to extrapolate emotional expressions in DNN-based text-to-speech20 Feb 2021 0 repositories listed
-
Alternate Endings: Improving Prosody for Incremental Neural TTS with Predicted Future Text Input19 Feb 2021 0 repositories listed
-
AudioVisual Speech Synthesis: A brief literature review18 Feb 2021 0 repositories listed
-
VARA-TTS: Non-Autoregressive Text-to-Speech Synthesis based on Very Deep VAE with Residual Attention12 Feb 2021 0 repositories listed
-
Voice Cloning: a Multi-Speaker Text-to-Speech Synthesis Approach based on Transfer Learning10 Feb 2021 0 repositories listed
-
Towards Natural and Controllable Cross-Lingual Voice Conversion Based on Neural TTS Model and Phonetic Posteriorgram3 Feb 2021 0 repositories listed
-
Expressive Neural Voice Cloning30 Jan 2021 0 repositories listed
-
Triple M: A Practical Text-to-speech Synthesis System With Multi-guidance Attention And Multi-band Multi-time LPCNet30 Jan 2021 0 repositories listed
-
EmoCat: Language-agnostic Emotional Voice Conversion14 Jan 2021 0 repositories listed
-
Generating coherent spontaneous speech and gesture from text14 Jan 2021 0 repositories listed
-
Whispered and Lombard Neural Speech Synthesis13 Jan 2021 0 repositories listed
-
1 Jan 2021 0 repositories listed
-
Detection of Lexical Stress Errors in Non-Native (L2) English with Data Augmentation and Attention29 Dec 2020 0 repositories listed
-
Denoising Text to Speech with Frame-Level Noise Modeling17 Dec 2020 0 repositories listed
-
Parallel WaveNet conditioned on VAE latent vectors17 Dec 2020 0 repositories listed
-
Syntactic representation learning for neural network based TTS with syntactic parse tree traversal13 Dec 2020 0 repositories listed
-
Using previous acoustic context to improve Text-to-Speech synthesis7 Dec 2020 0 repositories listed
-
GraphPB: Graphical Representations of Prosody Boundary in Speech Synthesis3 Dec 2020 0 repositories listed
-
Text-to-speech for the hearing impaired3 Dec 2020 0 repositories listed
-
Development of Smartcall Vietnamese Text-to-Speech for VLSP 20201 Dec 2020 0 repositories listed
-
Improving prosodic phrasing of Vietnamese text-to-speech systems1 Dec 2020 0 repositories listed
-
Vietnamese Text-To-Speech Shared Task VLSP 2020: Remaining problems with state-of-the-art techniques1 Dec 2020 0 repositories listed
-
Bootstrap an end-to-end ASR system by multilingual training, transfer learning, text-to-text mapping and synthetic audio25 Nov 2020 0 repositories listed
-
FBWave: Efficient and Scalable Neural Vocoders for Streaming Text-To-Speech on the Edge25 Nov 2020 0 repositories listed
-
Synth2Aug: Cross-domain speaker recognition with TTS synthesized speech24 Nov 2020 0 repositories listed
-
Using Synthetic Audio to Improve The Recognition of Out-Of-Vocabulary Words in End-To-End ASR Systems23 Nov 2020 0 repositories listed
-
Deep Shallow Fusion for RNN-T Personalization16 Nov 2020 0 repositories listed
-
Using IPA-Based Tacotron for Data Efficient Cross-Lingual Speaker Adaptation and Pronunciation Enhancement12 Nov 2020 0 repositories listed
-
Low-resource expressive text-to-speech using data augmentation11 Nov 2020 0 repositories listed
-
Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS10 Nov 2020 0 repositories listed
-
Fine-grained Style Modeling, Transfer and Prediction in Text-to-Speech Synthesis via Phone-Level Content-Style Disentanglement8 Nov 2020 0 repositories listed
-
Improving Prosody Modelling with Cross-Utterance BERT Embeddings for End-to-end Speech Synthesis6 Nov 2020 0 repositories listed
-
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework4 Nov 2020 0 repositories listed
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
Prosodic Representation Learning and Contextual Sampling for Neural Text-to-Speech4 Nov 2020 0 repositories listed
-
Training Wake Word Detection with Synthesized Speech Data on Confusion Words3 Nov 2020 0 repositories listed
-
Learning to Maximize Speech Quality Directly Using MOS Prediction for Neural Text-to-Speech2 Nov 2020 0 repositories listed
-
Learning from Explanations and Demonstrations: A Pilot Study1 Nov 2020 0 repositories listed
-
DeviceTTS: A Small-Footprint, Fast, Stable Network for On-Device Text-to-Speech29 Oct 2020 0 repositories listed