Browse State-of-the-Art › Speech Synthesis › Papers, page 11
Speech Synthesis
Papers archive 2025-07-28
archive papers tagged: 1,249 · with a code link: 366 · where Syntology ran a sample: 101 (85 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (101 of 1,249 tagged: 85 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 11 of 13: papers 1,001 to 1,100 of 1,249, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
DNN-based Speaker Embedding Using Subjective Inter-speaker Similarity for Multi-speaker Modeling in Speech Synthesis19 Jul 2019 0 repositories listed
-
Multi-Speaker End-to-End Speech Synthesis9 Jul 2019 0 repositories listed
-
Évaluation objective de plongements pour la synthèse de parole guidée par réseaux de neurones (Objective evaluation of embeddings for speech synthesis guided by neural networks)1 Jul 2019 0 repositories listed
-
End-to-End Emotional Speech Synthesis Using Style Tokens and Semi-Supervised Training26 Jun 2019 0 repositories listed
-
26 Jun 2019 0 repositories listed
-
A Unified Speaker Adaptation Method for Speech Synthesis using Transcribed and Untranscribed Speech with Backpropagation18 Jun 2019 0 repositories listed
-
Towards Transfer Learning for End-to-End Speech Synthesis from Deep Pre-Trained Language Models17 Jun 2019 0 repositories listed
-
Listening while Speaking and Visualizing: Improving ASR through Multimodal Chain3 Jun 2019 0 repositories listed
-
Neural Models of Text Normalization for Speech Applications1 Jun 2019 0 repositories listed
-
Neural Text Normalization with Subword Units1 Jun 2019 0 repositories listed
-
Permanent Magnetic Articulograph (PMA) vs Electromagnetic Articulograph (EMA) in Articulation-to-Speech Synthesis for Silent Speech Interface1 Jun 2019 0 repositories listed
-
Speaker Anonymization Using X-vector and Neural Waveform Models30 May 2019 0 repositories listed
-
Video-to-Video Translation for Visual Speech Synthesis28 May 2019 0 repositories listed
-
ET-GAN: Cross-Language Emotion Transfer Based on Cycle-Consistent Generative Adversarial Networks27 May 2019 0 repositories listed
-
CHiVE: Varying Prosody in Speech Synthesis with a Linguistically Driven Dynamic Hierarchical Conditional Variational Network17 May 2019 0 repositories listed
-
Speaker-Independent Speech-Driven Visual Speech Synthesis using Domain-Adapted Acoustic Models15 May 2019 0 repositories listed
-
Neural source-filter waveform models for statistical parametric speech synthesis27 Apr 2019 0 repositories listed
-
Latent Variable Algorithms for Multimodal Learning and Sensor Fusion23 Apr 2019 0 repositories listed
-
Unsupervised acoustic unit discovery for speech synthesis using discrete latent-variable neural networks16 Apr 2019 0 repositories listed
-
A high quality and phonetic balanced speech corpus for Vietnamese11 Apr 2019 0 repositories listed
-
Deep Learning the EEG Manifold for Phonological Categorization from Active Thoughts8 Apr 2019 0 repositories listed
-
SPEAK YOUR MIND! Towards Imagined Speech Recognition With Hierarchical Deep Learning8 Apr 2019 0 repositories listed
-
WaveCycleGAN2: Time-domain Neural Post-filter for Speech Waveform Generation5 Apr 2019 0 repositories listed
-
Multi-reference Tacotron by Intercross Training for Style Disentangling,Transfer and Control in Speech Synthesis4 Apr 2019 0 repositories listed
-
Speech denoising by parametric resynthesis2 Apr 2019 0 repositories listed
-
Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet29 Mar 2019 0 repositories listed
-
Generative adversarial network-based glottal waveform model for statistical parametric speech synthesis14 Mar 2019 0 repositories listed
-
Deep Text-to-Speech System with Seq2Seq Model11 Mar 2019 0 repositories listed
-
The Virtual Doctor: An Interactive Artificial Intelligence based on Deep Learning for Non-Invasive Prediction of Diabetes9 Mar 2019 0 repositories listed
-
Securing Voice-driven Interfaces against Fake (Cloned) Audio Attacks18 Feb 2019 0 repositories listed
-
Bytes are All You Need: End-to-End Multilingual Speech Recognition and Synthesis with Bytes22 Nov 2018 0 repositories listed
-
Effect of data reduction on sequence-to-sequence neural TTS15 Nov 2018 0 repositories listed
-
AttS2S-VC: Sequence-to-Sequence Voice Conversion with Attention and Context Preservation Mechanisms9 Nov 2018 0 repositories listed
-
ExcitNet vocoder: A neural excitation model for parametric speech synthesis systems9 Nov 2018 0 repositories listed
-
Speaker-adaptive neural vocoders for parametric speech synthesis systems8 Nov 2018 0 repositories listed
-
Investigating context features hidden in End-to-End TTS4 Nov 2018 0 repositories listed
-
End-to-End Feedback Loss in Speech Chain Framework via Straight-Through Estimator31 Oct 2018 0 repositories listed
-
Disentangling Correlated Speaker and Noise for Speech Synthesis via Data Augmentation and Adversarial Factorization30 Oct 2018 0 repositories listed
-
How to make someone speak a language that they don't know.30 Oct 2018 0 repositories listed
-
Waveform generation for text-to-speech synthesis using pitch-synchronous multi-scale generative adversarial networks30 Oct 2018 0 repositories listed
-
Neural source-filter-based waveform model for statistical parametric speech synthesis29 Oct 2018 0 repositories listed
-
Speaking style adaptation in Text-To-Speech synthesis using Sequence-to-sequence models with attention29 Oct 2018 0 repositories listed
-
A Challenge Set and Methods for Noun-Verb Ambiguity1 Oct 2018 0 repositories listed
-
WaveCycleGAN: Synthetic-to-natural speech waveform conversion using cycle-consistent adversarial networks25 Sep 2018 0 repositories listed
-
Hindi-English Code-Switching Speech Corpus24 Sep 2018 0 repositories listed
-
Self-Attention Linguistic-Acoustic Decoder31 Aug 2018 0 repositories listed
-
Semi-Supervised Training for Improving Data Efficiency in End-to-End Speech Synthesis30 Aug 2018 0 repositories listed
-
Fast Spectrogram Inversion using Multi-head Convolutional Neural Networks20 Aug 2018 0 repositories listed
-
Multimodal speech synthesis architecture for unsupervised speaker adaptation20 Aug 2018 0 repositories listed
-
Predicting Expressive Speaking Style From Text In End-To-End Speech Synthesis4 Aug 2018 0 repositories listed
-
Investigating accuracy of pitch-accent annotations in neural network-based speech synthesis and denoising effects2 Aug 2018 0 repositories listed
-
Indigenous language technologies in Canada: Assessment, challenges, and successes1 Aug 2018 0 repositories listed
-
Scaling and bias codes for modeling speaker-adaptive DNN-based speech synthesis systems31 Jul 2018 0 repositories listed
-
Wasserstein GAN and Waveform Loss-based Acoustic Model Training for Multi-speaker Text-to-Speech Synthesis Systems Using a WaveNet Vocoder31 Jul 2018 0 repositories listed
-
Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis30 Jul 2018 0 repositories listed
-
Analysing Shortcomings of Statistical Parametric Speech Synthesis28 Jul 2018 0 repositories listed
-
Forward Attention in Sequence-to-sequence Acoustic Modelling for Speech Synthesis18 Jul 2018 0 repositories listed
-
Syllabification by Phone Categorization15 Jul 2018 0 repositories listed
-
An Analysis of the Effect of Emotional Speech Synthesis on Non-Task-Oriented Dialogue System1 Jul 2018 0 repositories listed
-
EMPHASIS: An Emotional Phoneme-based Acoustic Model for Speech Synthesis System26 Jun 2018 0 repositories listed
-
Multi-task WaveNet: A Multi-task Generative Model for Statistical Parametric Speech Synthesis without Fundamental Frequency Conditions22 Jun 2018 0 repositories listed
-
ASR-based Features for Emotion Recognition: A Transfer Learning Approach23 May 2018 0 repositories listed
-
A Regression Model of Recurrent Deep Neural Networks for Noise Robust Estimation of the Fundamental Frequency Contour of Speech8 May 2018 0 repositories listed
-
Building Open Javanese and Sundanese Corpora for Multilingual Text-to-Speech1 May 2018 0 repositories listed
-
Construction of English-French Multimodal Affective Conversational Corpus from TV Dramas1 May 2018 0 repositories listed
-
CPJD Corpus: Crowdsourced Parallel Speech Corpus of Japanese Dialects1 May 2018 0 repositories listed
-
Design and Development of Speech Corpora for Air Traffic Control Training1 May 2018 0 repositories listed
-
From `Solved Problems' to New Challenges: A Report on LDC Activities1 May 2018 0 repositories listed
-
Improving homograph disambiguation with supervised machine learning1 May 2018 0 repositories listed
-
Literality and cognitive effort: Japanese and Spanish1 May 2018 0 repositories listed
-
SynPaFlex-Corpus: An Expressive French Audiobooks Corpus dedicated to expressive speech synthesis.1 May 2018 0 repositories listed
-
Speaker-independent raw waveform model for glottal excitation25 Apr 2018 0 repositories listed
-
A comparison of recent waveform generation and acoustic modeling methods for neural-network-based speech synthesis7 Apr 2018 0 repositories listed
-
Expressive Speech Synthesis via Modeling Expressions with Variational Autoencoder6 Apr 2018 0 repositories listed
-
High-quality nonparallel voice conversion based on cycle-consistent adversarial network2 Apr 2018 0 repositories listed
-
Machine Speech Chain with One-shot Speaker Adaptation28 Mar 2018 0 repositories listed
-
Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama's voice using GAN, WaveNet and low-quality found data2 Mar 2018 0 repositories listed
-
Deep Feed-forward Sequential Memory Networks for Speech Synthesis26 Feb 2018 0 repositories listed
-
Fitting New Speakers Based on a Short Untranscribed Sample20 Feb 2018 0 repositories listed
-
HybridNet: A Hybrid Neural Architecture to Speed-up Autoregressive Models1 Jan 2018 0 repositories listed
-
POLICY DRIVEN GENERATIVE ADVERSARIAL NETWORKS FOR ACCENTED SPEECH GENERATION1 Jan 2018 0 repositories listed
-
pyiwn: A Python based API to access Indian Language WordNets1 Jan 2018 0 repositories listed
-
Synthesizing Audio for Hindi WordNet1 Jan 2018 0 repositories listed
-
23 Dec 2017 0 repositories listed
-
Creating New Language and Voice Components for the Updated MaryTTS Text-to-Speech Synthesis Platform13 Dec 2017 0 repositories listed
-
完全基於類神經網路之語音合成系統初步研究 (A Preliminary Study on Fully Neural Network-based Speech Synthesis System) [In Chinese]1 Nov 2017 0 repositories listed
-
SUT System Description for Anti-Spoofing 2017 Challenge1 Nov 2017 0 repositories listed
-
Uncovering Latent Style Factors for Expressive Speech Synthesis1 Nov 2017 0 repositories listed
-
Fast and Accurate Decision Trees for Natural Language Processing Tasks1 Sep 2017 0 repositories listed
-
Lexicon for Natural Language Generation in Spanish Adapted to Alternative and Augmentative Communication1 Sep 2017 0 repositories listed
-
Refer-iTTS: A System for Referring in Spoken Installments to Objects in Real-World Images1 Sep 2017 0 repositories listed
-
Listening while Speaking: Speech Chain by Deep Learning16 Jul 2017 0 repositories listed
-
Hidden-Markov-Model Based Speech Enhancement4 Jul 2017 0 repositories listed
-
PyDial: A Multi-domain Statistical Dialogue System Toolkit1 Jul 2017 0 repositories listed
-
A Variational EM Method for Pole-Zero Modeling of Speech with Mixed Block Sparse and Gaussian Excitation24 Jun 2017 0 repositories listed
-
I Probe, Therefore I Am: Designing a Virtual Journalist with Human Emotions18 May 2017 0 repositories listed
-
Aligning phonemes using finte-state methods1 May 2017 0 repositories listed
-
Building and using language resources and infrastructure to develop e-learning programs for a minority language1 May 2017 0 repositories listed
-
Sampling-based speech parameter generation using moment-matching networks12 Apr 2017 0 repositories listed
-
Voice Conversion Using Sequence-to-Sequence Learning of Context Posterior Probabilities10 Apr 2017 0 repositories listed