Papers › Wavetable Synthesis Using CVAE for Timbre Control Based on Semantic Label

Wavetable Synthesis Using CVAE for Timbre Control Based on Semantic Label

24 Oct 2024arXiv:2410.18628archive 2025-07-28

Tsugumasa Yutani, Yuya Yamamoto, Shuyo Nakatani, Hiroko Terasawa

Synthesizers are essential in modern music production. However, their complex timbre parameters, often filled with technical terms, require expertise. This research introduces a method of timbre control in wavetable synthesis that is intuitive and sensible and utilizes semantic labels. Using a conditional variational autoencoder (CVAE), users can select a wavetable and define the timbre with labels such as bright, warm, and rich. The CVAE model, featuring convolutional and upsampling layers, effectively captures the wavetable nuances, ensuring real-time performance owing to their processing in the time domain. Experiments demonstrate that this approach allows for real-time, effective control of the timbre of the wavetable using semantic inputs and aims for intuitive timbre control through data-based semantic control.

PaperPDFCode

Code

tsugumasa320/wavetablecvae officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Datasets

Introduced by this paper, per the archive.

Single cycle waveforms

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

cVAE

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections