Browse State-of-the-Art › Voice Conversion › Papers, page 4
Voice Conversion
Papers archive 2025-07-28
archive papers tagged: 520 · with a code link: 175 · where Syntology ran a sample: 41 (32 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (41 of 520 tagged: 32 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)
Page 4 of 6: papers 301 to 400 of 520, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Face-Driven Zero-Shot Voice Conversion with Memory-based Face-Voice Alignment18 Sep 2023 0 repositories listed
-
PromptVC: Flexible Stylistic Voice Conversion in Latent Space Driven by Natural Language Prompts17 Sep 2023 0 repositories listed
-
Cross-lingual Knowledge Distillation via Flow-based Voice Conversion for Robust Polyglot Text-To-Speech15 Sep 2023 0 repositories listed
-
Improving Voice Conversion for Dissimilar Speakers Using Perceptual Losses15 Sep 2023 0 repositories listed
-
Parallel and Limited Data Voice Conversion Using Stochastic Variational Deep Kernel Learning8 Sep 2023 0 repositories listed
-
Stylebook: Content-Dependent Speaking Style Modeling for Any-to-Any Voice Conversion using Only Speech Data6 Sep 2023 0 repositories listed
-
MSM-VC: High-fidelity Source Style Transfer for Non-Parallel Voice Conversion by Multi-scale Style Modeling3 Sep 2023 0 repositories listed
-
Learning Speech Representation From Contrastive Token-Acoustic Pretraining1 Sep 2023 0 repositories listed
-
Generalizable Zero-Shot Speaker Adaptive Speech Synthesis with Disentangled Representations24 Aug 2023 0 repositories listed
-
24 Aug 2023 0 repositories listed
-
Effects of Convolutional Autoencoder Bottleneck Width on StarGAN-based Singing Technique Conversion19 Aug 2023 0 repositories listed
-
SLMGAN: Exploiting Speech Language Model Representations for Unsupervised Zero-Shot Voice Conversion in GANs18 Jul 2023 0 repositories listed
-
Deep Learning-based F0 Synthesis for Speaker Anonymization29 Jun 2023 0 repositories listed
-
Fake the Real: Backdoor Attack on Deep Speech Classification via Voice Conversion28 Jun 2023 0 repositories listed
-
Two-Stage Voice Anonymization for Enhanced Privacy28 Jun 2023 0 repositories listed
-
Automatic Speech Disentanglement for Voice Conversion using Rank Module and Speech Augmentation21 Jun 2023 0 repositories listed
-
LM-VC: Zero-shot Voice Conversion via Speech Generation based on Language Models18 Jun 2023 0 repositories listed
-
ALO-VC: Any-to-any Low-latency One-shot Voice Conversion1 Jun 2023 0 repositories listed
-
Make-A-Voice: Unified Voice Synthesis With Discrete Representation30 May 2023 0 repositories listed
-
Creating Personalized Synthetic Voices from Post-Glossectomy Speech with Guided Diffusion Models27 May 2023 0 repositories listed
-
Iteratively Improving Speech Recognition and Voice Conversion24 May 2023 0 repositories listed
-
DualVC: Dual-mode Voice Conversion using Intra-model Knowledge Distillation and Hybrid Predictive Coding21 May 2023 0 repositories listed
-
Data Augmentation for Diverse Voice Conversion in Noisy Environments18 May 2023 0 repositories listed
-
Adversarial Speaker Disentanglement Using Unannotated External Data for Self-supervised Representation Based Voice Conversion16 May 2023 0 repositories listed
-
Multi-level Temporal-channel Speaker Retrieval for Zero-shot Voice Conversion12 May 2023 0 repositories listed
-
AlignSTS: Speech-to-Singing Conversion via Cross-Modal Alignment8 May 2023 0 repositories listed
-
Evaluation of Speaker Anonymization on Emotional Speech15 Apr 2023 0 repositories listed
-
Self-Supervised Representations for Singing Voice Conversion21 Mar 2023 0 repositories listed
-
A Comparative Analysis Of Latent Regressor Losses For Singing Voice Conversion27 Feb 2023 0 repositories listed
-
Cross-modal Face- and Voice-style Transfer27 Feb 2023 0 repositories listed
-
Catch You and I Can: Revealing Source Voiceprint Against Voice Conversion24 Feb 2023 0 repositories listed
-
Nonparallel Emotional Voice Conversion For Unseen Speaker-Emotion Pairs Using Dual Domain Adversarial Network & Virtual Domain Pairing21 Feb 2023 0 repositories listed
-
ACE-VC: Adaptive and Controllable Voice Conversion using Explicitly Disentangled Self-supervised Speech Representations16 Feb 2023 0 repositories listed
-
Modelling low-resource accents without accent-specific TTS frontend11 Jan 2023 0 repositories listed
-
UnifySpeech: A Unified Framework for Zero-shot Text-to-Speech and Voice Conversion10 Jan 2023 0 repositories listed
-
VSVC: Backdoor attack against Keyword Spotting based on Voiceprint Selection and Voice Conversion20 Dec 2022 0 repositories listed
-
Disentangling Prosody Representations with Unsupervised Speech Reconstruction14 Dec 2022 0 repositories listed
-
Disentangled Feature Learning for Real-Time Neural Speech Coding22 Nov 2022 0 repositories listed
-
Audio Anti-spoofing Using a Simple Attention Module and Joint Optimization Based on Additive Angular Margin Loss and Meta-learning17 Nov 2022 0 repositories listed
-
Delivering Speaking Style in Low-resource Voice Conversion with Multi-factor Constraints16 Nov 2022 0 repositories listed
-
Improved disentangled speech representations using contrastive learning in factorized hierarchical variational autoencoder15 Nov 2022 0 repositories listed
-
Expressive-VC: Highly Expressive Voice Conversion with Attention Fusion of Bottleneck and Perturbation Features9 Nov 2022 0 repositories listed
-
Preserving background sound in noise-robust voice conversion via multi-task learning6 Nov 2022 0 repositories listed
-
Combining Automatic Speaker Verification and Prosody Analysis for Synthetic Speech Detection31 Oct 2022 0 repositories listed
-
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation31 Oct 2022 0 repositories listed
-
Streaming Voice Conversion Via Intermediate Bottleneck Features And Non-streaming Teacher Guidance27 Oct 2022 0 repositories listed
-
V-Cloak: Intelligibility-, Naturalness- & Timbre-Preserving Real-Time Voice Anonymization27 Oct 2022 0 repositories listed
-
Disentangled Speech Representation Learning for One-Shot Cross-lingual Voice Conversion Using β-VAE25 Oct 2022 0 repositories listed
-
MetaSpeech: Speech Effects Switch Along with Environment for Metaverse25 Oct 2022 0 repositories listed
-
Mixed-EVC: Mixed Emotion Synthesis and Control in Voice Conversion25 Oct 2022 0 repositories listed
-
DisC-VC: Disentangled and F0-Controllable Neural Voice Conversion20 Oct 2022 0 repositories listed
-
Robust One-Shot Singing Voice Conversion20 Oct 2022 0 repositories listed
-
Boosting Star-GANs for Voice Conversion with Contrastive Discriminator21 Sep 2022 0 repositories listed
-
Non-Parallel Voice Conversion for ASR Augmentation15 Sep 2022 0 repositories listed
-
Using Rater and System Metadata to Explain Variance in the VoiceMOS Challenge 2022 Dataset14 Sep 2022 0 repositories listed
-
Investigation into Target Speaking Rate Adaptation for Voice Conversion5 Sep 2022 0 repositories listed
-
Are disentangled representations all you need to build speaker anonymization systems?22 Aug 2022 0 repositories listed
-
Differentiable WORLD Synthesizer-based Neural Vocoder With Application To End-To-End Audio Style Transfer15 Aug 2022 0 repositories listed
-
TGAVC: Improving Autoencoder Voice Conversion with Text-Guided and Adversarial Training8 Aug 2022 0 repositories listed
-
Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation29 Jul 2022 0 repositories listed
-
Transplantation of Conversational Speaking Style with Interjections in Sequence-to-Sequence Speech Synthesis25 Jul 2022 0 repositories listed
-
GlowVC: Mel-spectrogram space disentangling model for language-independent text-free voice conversion4 Jul 2022 0 repositories listed
-
A Hierarchical Speaker Representation Framework for One-shot Singing Voice Conversion28 Jun 2022 0 repositories listed
-
Comparison of Speech Representations for the MOS Prediction System28 Jun 2022 0 repositories listed
-
Identifying Source Speakers for Voice Conversion based Spoofing Attacks on Speaker Verification Systems18 Jun 2022 0 repositories listed
-
End-to-End Voice Conversion with Information Perturbation15 Jun 2022 0 repositories listed
-
Face-Dubbing++: Lip-Synchronous, Voice Preserving Translation of Videos9 Jun 2022 0 repositories listed
-
Investigating Inter- and Intra-speaker Voice Conversion using Audiobooks1 Jun 2022 0 repositories listed
-
Attentive activation function for improving end-to-end spoofing countermeasure systems3 May 2022 0 repositories listed
-
Cross-Speaker Emotion Transfer for Low-Resource Text-to-Speech Using Non-Parallel Voice Conversion with Pitch-Shift Data Augmentation21 Apr 2022 0 repositories listed
-
Audio Deep Fake Detection System with Neural Stitching for ADD 202219 Apr 2022 0 repositories listed
-
Time Domain Adversarial Voice Conversion for ADD 202219 Apr 2022 0 repositories listed
-
The PartialSpoof Database and Countermeasures for the Detection of Short Fake Speech Segments Embedded in an Utterance11 Apr 2022 0 repositories listed
-
Representation Selective Self-distillation and wav2vec 2.0 Feature Exploration for Spoof-aware Speaker Verification6 Apr 2022 0 repositories listed
-
Disentangled Speech Representation Learning Based on Factorized Hierarchical Variational Autoencoder with Self-Supervised Objective5 Apr 2022 0 repositories listed
-
Anti-Spoofing Using Transfer Learning with Variational Information Bottleneck4 Apr 2022 0 repositories listed
-
Self-Supervised Speech Representations Preserve Speech Characteristics while Anonymizing Voices4 Apr 2022 0 repositories listed
-
WavThruVec: Latent speech representation as intermediate features for neural speech synthesis31 Mar 2022 0 repositories listed
-
Enhancing Zero-Shot Many to Many Voice Conversion with Self-Attention VAE30 Mar 2022 0 repositories listed
-
An Overview & Analysis of Sequence-to-Sequence Emotional Voice Conversion29 Mar 2022 0 repositories listed
-
Analysis of Voice Conversion and Code-Switching Synthesis Using VQ-VAE28 Mar 2022 0 repositories listed
-
Disentangleing Content and Fine-grained Prosody Information via Hybrid ASR Bottleneck Features for Voice Conversion24 Mar 2022 0 repositories listed
-
Separating Content from Speaker Identity in Speech for the Assessment of Cognitive Impairments21 Mar 2022 0 repositories listed
-
Improve few-shot voice cloning using multi-modal learning18 Mar 2022 0 repositories listed
-
Text-free non-parallel many-to-many voice conversion using normalising flows15 Mar 2022 0 repositories listed
-
VCVTS: Multi-speaker Video-to-Speech synthesis via cross-modal knowledge transfer from voice conversion18 Feb 2022 0 repositories listed
-
Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module16 Feb 2022 0 repositories listed
-
Partially Fake Audio Detection by Self-attention-based Fake Span Discovery14 Feb 2022 0 repositories listed
-
Cross-speaker style transfer for text-to-speech using data augmentation10 Feb 2022 0 repositories listed
-
Invertible Voice Conversion26 Jan 2022 0 repositories listed
-
The Effectiveness of Time Stretching for Enhancing Dysarthric Speech for Improved Dysarthric Speech Recognition13 Jan 2022 0 repositories listed
-
Emotion Intensity and its Control for Emotional Voice Conversion10 Jan 2022 0 repositories listed
-
Adversarial Transformation of Spoofing Attacks for Voice Biometrics4 Jan 2022 0 repositories listed
-
IQDUBBING: Prosody modeling based on discrete self-supervised speech representation for expressive voice conversion2 Jan 2022 0 repositories listed
-
The exploitation of Multiple Feature Extraction Techniques for Speaker Identification in Emotional States under Disguised Voices15 Dec 2021 0 repositories listed
-
Training Robust Zero-Shot Voice Conversion Models with Self-supervised Features8 Dec 2021 0 repositories listed
-
Conditional Deep Hierarchical Variational Autoencoder for Voice Conversion6 Dec 2021 0 repositories listed
-
VoiceMixer: Adversarial Voice Style Mixup1 Dec 2021 0 repositories listed
-
One-shot Voice Conversion For Style Transfer Based On Speaker Adaptation24 Nov 2021 0 repositories listed
-
AC-VC: Non-parallel Low Latency Phonetic Posteriorgrams Based Voice Conversion12 Nov 2021 0 repositories listed