{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/voice-conversion-using-convolutional-neural","title":"Voice Conversion using Convolutional Neural Networks","arxiv_id":"1610.08927","date":"2016-10-27","proceeding":null,"authors":["Shariq Mobin","Joan Bruna"],"abstract":"The human auditory system is able to distinguish the vocal source of\nthousands of speakers, yet not much is known about what features the auditory\nsystem uses to do this. Fourier Transforms are capable of capturing the pitch\nand harmonic structure of the speaker but this alone proves insufficient at\nidentifying speakers uniquely. The remaining structure, often referred to as\ntimbre, is critical to identifying speakers but we understood little about it.\nIn this paper we use recent advances in neural networks in order to manipulate\nthe voice of one speaker into another by transforming not only the pitch of the\nspeaker, but the timbre. We review generative models built with neural networks\nas well as architectures for creating neural networks that learn analogies. Our\npreliminary results converting voices from one speaker to another are\nencouraging.","url_abs":"http://arxiv.org/abs/1610.08927v1","url_pdf":"http://arxiv.org/pdf/1610.08927v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"voice-conversion-using-convolutional-neural","repo_url":"https://github.com/ShariqM/smcnn","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[{"task_slug":"voice-conversion","task_name":"Voice Conversion"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}