Browse State-of-the-Art › Speech Synthesis › Papers, page 9
Speech Synthesis
Papers archive 2025-07-28
archive papers tagged: 1,249 · with a code link: 366 · where Syntology ran a sample: 101 (85 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (101 of 1,249 tagged: 85 with a run with no instrument failure, 16 where every run was a failure of Syntology's instrument)
Page 9 of 13: papers 801 to 900 of 1,249, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Cross-lingual Low Resource Speaker Adaptation Using Phonological Features17 Nov 2021 0 repositories listed
-
High Quality Streaming Speech Synthesis with Low, Sentence-Length-Independent Latency17 Nov 2021 0 repositories listed
-
Modeling speech recognition and synthesis simultaneously: Encoding and decoding lexical and sublexical semantic information into speech with no access to speech data16 Nov 2021 0 repositories listed
-
Speech Synthesis for Low Resource Languages using Transliteration Enabled Transfer Learning16 Nov 2021 0 repositories listed
-
Improving Prosody for Unseen Texts in Speech Synthesis by Utilizing Linguistic Information and Noisy Data15 Nov 2021 0 repositories listed
-
Assessing Evaluation Metrics for Speech-to-Speech Translation26 Oct 2021 0 repositories listed
-
Synt++: Utilizing Imperfect Synthetic Data to Improve Speech Recognition21 Oct 2021 0 repositories listed
-
Direct Simultaneous Speech-to-Speech Translation with Variational Monotonic Multihead Attention15 Oct 2021 0 repositories listed
-
From Start to Finish: Latency Reduction Strategies for Incremental Speech Synthesis in Simultaneous Speech-to-Speech Translation15 Oct 2021 0 repositories listed
-
SingGAN: Generative Adversarial Network For High-Fidelity Singing Voice Generation14 Oct 2021 0 repositories listed
-
DeepA: A Deep Neural Analyzer For Speech And Singing Vocoding13 Oct 2021 0 repositories listed
-
LaughNet: synthesizing laughter utterances from waveform silhouettes and a single laughter example11 Oct 2021 0 repositories listed
-
Using multiple reference audios and style embedding constraints for speech synthesis9 Oct 2021 0 repositories listed
-
Environment Aware Text-to-Speech Synthesis8 Oct 2021 0 repositories listed
-
Cloning one's voice using very limited data in the wild7 Oct 2021 0 repositories listed
-
VisualTTS: TTS with Accurate Lip-Speech Synchronization for Automatic Voice Over7 Oct 2021 0 repositories listed
-
GANtron: Emotional Speech Synthesis with Generative Adversarial Networks6 Oct 2021 0 repositories listed
-
Prosody-TTS: An end-to-end speech synthesis system with prosody control6 Oct 2021 0 repositories listed
-
On the Interplay Between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis4 Oct 2021 0 repositories listed
-
3 Oct 2021 0 repositories listed
-
Incorporating speaker embedding and post-filter network for improving speaker similarity of personalized speech synthesis system1 Oct 2021 0 repositories listed
-
Conditioning Sequence-to-sequence Networks with Learned Activations29 Sep 2021 0 repositories listed
-
Guided-TTS:Text-to-Speech with Untranscribed Speech29 Sep 2021 0 repositories listed
-
Speech-MLP: a simple MLP architecture for speech processing29 Sep 2021 0 repositories listed
-
SynCLR: A Synthesis Framework for Contrastive Learning of out-of-domain Speech Representations29 Sep 2021 0 repositories listed
-
FlowVocoder: A small Footprint Neural Vocoder based Normalizing flow for Speech Synthesis27 Sep 2021 0 repositories listed
-
Low-Latency Incremental Text-to-Speech Synthesis with Distilled Context Prediction Network22 Sep 2021 0 repositories listed
-
"Hello, It's Me": Deep Learning-based Speech Synthesis Attacks in the Real World20 Sep 2021 0 repositories listed
-
On-device neural speech synthesis17 Sep 2021 0 repositories listed
-
Referee: Towards reference-free cross-speaker style transfer with low-quality data for expressive speech synthesis8 Sep 2021 0 repositories listed
-
Physiological-Physical Feature Fusion for Automatic Voice Spoofing Detection1 Sep 2021 0 repositories listed
-
Neural Sequence-to-Sequence Speech Synthesis Using a Hidden Semi-Markov Model Based Structured Attention Mechanism31 Aug 2021 0 repositories listed
-
Full Attention Bidirectional Deep Learning Structure for Single Channel Speech Enhancement27 Aug 2021 0 repositories listed
-
A Unified Transformer-based Framework for Duplex Text Normalization23 Aug 2021 0 repositories listed
-
Enhancing audio quality for expressive Neural Text-to-Speech13 Aug 2021 0 repositories listed
-
A Streamwise GAN Vocoder for Wideband Speech Coding at Very Low Bit Rate9 Aug 2021 0 repositories listed
-
Improved pronunciation prediction accuracy using morphology1 Aug 2021 0 repositories listed
-
Cross-speaker Style Transfer with Prosody Bottleneck in Neural Speech Synthesis27 Jul 2021 0 repositories listed
-
Exploring the Potential of Lexical Paraphrases for Mitigating Noise-Induced Comprehension Errors18 Jul 2021 0 repositories listed
-
Location, Location: Enhancing the Evaluation of Text-to-Speech Synthesis Using the Rapid Prosody Transcription Paradigm6 Jul 2021 0 repositories listed
-
An Objective Evaluation Framework for Pathological Speech Synthesis1 Jul 2021 0 repositories listed
-
GANSpeech: Adversarial Training for High-Fidelity Multi-Speaker Speech Synthesis29 Jun 2021 0 repositories listed
-
Preliminary study on using vector quantization latent spaces for TTS/VC systems with consistent performance25 Jun 2021 0 repositories listed
-
Controllable Context-aware Conversational Speech Synthesis21 Jun 2021 0 repositories listed
-
Glow-WaveGAN: Learning Speech Representations from GAN-based Variational Auto-Encoder For High Fidelity Flow-based Speech Synthesis21 Jun 2021 0 repositories listed
-
Non-native English lexicon creation for bilingual speech synthesis21 Jun 2021 0 repositories listed
-
UniTTS: Residual Learning of Unified Embedding Space for Speech Style Control21 Jun 2021 0 repositories listed
-
Advances in Speech Vocoding for Text-to-Speech with Continuous Parameters19 Jun 2021 0 repositories listed
-
17 Jun 2021 0 repositories listed
-
A Flow-Based Neural Network for Time Domain Speech Enhancement16 Jun 2021 0 repositories listed
-
Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis15 Jun 2021 0 repositories listed
-
Pathological voice adaptation with autoencoder-based voice conversion15 Jun 2021 0 repositories listed
-
Continuous Wavelet Vocoder-based Decomposition of Parametric Speech Waveform Synthesis12 Jun 2021 0 repositories listed
-
Sprachsynthese -- State-of-the-Art in englischer und deutscher Sprache11 Jun 2021 0 repositories listed
-
Learning to Efficiently Sample from Diffusion Probabilistic Models7 Jun 2021 0 repositories listed
-
Mathematical Vocoder Algorithm : Modified Spectral Inversion for Efficient Neural Speech Synthesis6 Jun 2021 0 repositories listed
-
An objective evaluation of the effects of recording conditions and speaker characteristics in multi-speaker deep neural speech synthesis3 Jun 2021 0 repositories listed
-
Speaker verification-derived loss and data augmentation for DNN-based multispeaker speech synthesis3 Jun 2021 0 repositories listed
-
Dual Script E2E framework for Multilingual and Code-Switching ASR2 Jun 2021 0 repositories listed
-
21 May 2021 0 repositories listed
-
Learning Robust Latent Representations for Controllable Speech Synthesis10 May 2021 0 repositories listed
-
Towards a practical lip-to-speech conversion system using deep neural networks and mobile application frontend29 Apr 2021 0 repositories listed
-
End-to-End Video-To-Speech Synthesis using Generative Adversarial Networks27 Apr 2021 0 repositories listed
-
An Adaptive Learning based Generative Adversarial Network for One-To-One Voice Conversion25 Apr 2021 0 repositories listed
-
Review of end-to-end speech synthesis technology based on deep learning20 Apr 2021 0 repositories listed
-
Enhancing Word-Level Semantic Representation via Dependency Structure for Expressive Text-to-Speech Synthesis14 Apr 2021 0 repositories listed
-
Flavored Tacotron: Conditional Learning for Prosodic-linguistic Features8 Apr 2021 0 repositories listed
-
Towards Multi-Scale Style Control for Expressive Speech Synthesis8 Apr 2021 0 repositories listed
-
Reinforcement Learning for Emotional Text-to-Speech Synthesis with Improved Emotion Discriminability3 Apr 2021 0 repositories listed
-
Continual Speaker Adaptation for Text-to-Speech Synthesis26 Mar 2021 0 repositories listed
-
21 Mar 2021 0 repositories listed
-
Alternate Endings: Improving Prosody for Incremental Neural TTS with Predicted Future Text Input19 Feb 2021 0 repositories listed
-
AudioVisual Speech Synthesis: A brief literature review18 Feb 2021 0 repositories listed
-
VARA-TTS: Non-Autoregressive Text-to-Speech Synthesis based on Very Deep VAE with Residual Attention12 Feb 2021 0 repositories listed
-
11 Feb 2021 0 repositories listed
-
Voice Cloning: a Multi-Speaker Text-to-Speech Synthesis Approach based on Transfer Learning10 Feb 2021 0 repositories listed
-
Generacion de voces artificiales infantiles en castellano con acento costarricense2 Feb 2021 0 repositories listed
-
SPEAK WITH YOUR HANDS Using Continuous Hand Gestures to control Articulatory Speech Synthesizer2 Feb 2021 0 repositories listed
-
Universal Neural Vocoding with Parallel WaveNet1 Feb 2021 0 repositories listed
-
Expressive Neural Voice Cloning30 Jan 2021 0 repositories listed
-
Triple M: A Practical Text-to-speech Synthesis System With Multi-guidance Attention And Multi-band Multi-time LPCNet30 Jan 2021 0 repositories listed
-
Generating coherent spontaneous speech and gesture from text14 Jan 2021 0 repositories listed
-
Whispered and Lombard Neural Speech Synthesis13 Jan 2021 0 repositories listed
-
Speech Synthesis as Augmentation for Low-Resource ASR23 Dec 2020 0 repositories listed
-
Parallel WaveNet conditioned on VAE latent vectors17 Dec 2020 0 repositories listed
-
Few Shot Adaptive Normalization Driven Multi-Speaker Speech Synthesis14 Dec 2020 0 repositories listed
-
Using previous acoustic context to improve Text-to-Speech synthesis7 Dec 2020 0 repositories listed
-
GraphPB: Graphical Representations of Prosody Boundary in Speech Synthesis3 Dec 2020 0 repositories listed
-
German-Arabic Speech-to-Speech Translation for Psychiatric Diagnosis1 Dec 2020 0 repositories listed
-
基於深度學習之中文文字轉台語語音合成系統初步探討 (A Preliminary Study on Deep Learning-based Chinese Text to Taiwanese Speech Synthesis System)1 Dec 2020 0 repositories listed
-
Sentiment Analysis for Emotional Speech Synthesis in a News Dialogue System1 Dec 2020 0 repositories listed
-
TaL: a synchronised multi-speaker corpus of ultrasound tongue imaging, audio, and lip videos19 Nov 2020 0 repositories listed
-
Pretraining Strategies, Waveform Model Choice, and Acoustic Configurations for Multi-Speaker End-to-End Speech Synthesis10 Nov 2020 0 repositories listed
-
Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS10 Nov 2020 0 repositories listed
-
Using GANs to Synthesise Minimum Training Data for Deepfake Generation10 Nov 2020 0 repositories listed
-
Fine-grained Style Modeling, Transfer and Prediction in Text-to-Speech Synthesis via Phone-Level Content-Style Disentanglement8 Nov 2020 0 repositories listed
-
Improving Prosody Modelling with Cross-Utterance BERT Embeddings for End-to-end Speech Synthesis6 Nov 2020 0 repositories listed
-
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework4 Nov 2020 0 repositories listed
-
Incremental Machine Speech Chain Towards Enabling Listening while Speaking in Real-time4 Nov 2020 0 repositories listed
-
Prosodic Representation Learning and Contextual Sampling for Neural Text-to-Speech4 Nov 2020 0 repositories listed