Browse State-of-the-Art › Audio Generation › Papers, page 3
Audio Generation
Papers archive 2025-07-28
archive papers tagged: 270 · with a code link: 124 · where Syntology ran a sample: 57 (49 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (57 of 270 tagged: 49 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 3 of 3: papers 201 to 270 of 270, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Autoregressive Diffusion Transformer for Text-to-Speech Synthesis8 Jun 2024 0 repositories listed
-
Creative Text-to-Audio Generation via Synthesizer Programming1 Jun 2024 0 repositories listed
-
A Survey of Deep Learning Audio Generation Methods31 May 2024 0 repositories listed
-
C3LLM: Conditional Multimodal Content Generation Using Large Language Models25 May 2024 0 repositories listed
-
The Rarity of Musical Audio Signals Within the Space of Possible Audio Generation23 May 2024 0 repositories listed
-
Visual Echoes: A Simple Unified Transformer for Audio-Visual Generation23 May 2024 0 repositories listed
-
11 May 2024 0 repositories listed Syntology 16 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 2 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Leveraging AI to Generate Audio for User-generated Content in Video Games25 Apr 2024 0 repositories listed
-
Music Style Transfer With Diffusion Model23 Apr 2024 0 repositories listed
-
LD-Pruner: Efficient Pruning of Latent Diffusion Models using Task-Agnostic Insights18 Apr 2024 0 repositories listed
-
Synthetic training set generation using text-to-audio models for environmental sound classification26 Mar 2024 0 repositories listed
-
Text-to-Audio Generation Synchronized with Videos8 Mar 2024 0 repositories listed
-
(Un)paired signal-to-signal translation with 1D conditional GANs5 Mar 2024 0 repositories listed
-
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models2 Mar 2024 0 repositories listed
-
Seeing and Hearing: Open-domain Visual-Audio Generation with Diffusion Latent Aligners27 Feb 2024 0 repositories listed
-
Classification Diffusion Models: Revitalizing Density Ratio Estimation15 Feb 2024 0 repositories listed
-
Analyzing Neural Network-Based Generative Diffusion Models through Convex Optimization3 Feb 2024 0 repositories listed
-
Bass Accompaniment Generation via Latent Diffusion2 Feb 2024 0 repositories listed
-
ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering14 Jan 2024 0 repositories listed
-
Masked Audio Generation using a Single Non-Autoregressive Transformer9 Jan 2024 0 repositories listed
-
Efficient Parallel Audio Generation using Group Masked Language Modeling2 Jan 2024 0 repositories listed
-
Cyclic Learning for Binaural Audio Generation and Localization1 Jan 2024 0 repositories listed
-
25 Dec 2023 0 repositories listed
-
Diffusion-EXR: Controllable Review Generation for Explainable Recommendation via Diffusion Models24 Dec 2023 0 repositories listed
-
CMMD: Contrastive Multi-Modal Diffusion for Video-Audio Conditional Modeling8 Dec 2023 0 repositories listed
-
SEFGAN: Harvesting the Power of Normalizing Flows and GANs for Efficient High-Quality Speech Enhancement4 Dec 2023 0 repositories listed
-
tinyCLAP: Distilling Constrastive Language-Audio Pretrained Models24 Nov 2023 0 repositories listed
-
Cross-modal Generative Model for Visual-Guided Binaural Stereo Generation13 Nov 2023 0 repositories listed
-
In-Context Prompt Editing For Conditional Audio Generation1 Nov 2023 0 repositories listed
-
On The Open Prompt Challenge In Conditional Audio Generation1 Nov 2023 0 repositories listed
-
Audio Editing with Non-Rigid Text Prompts19 Oct 2023 0 repositories listed
-
FoleyGen: Visually-Guided Audio Generation19 Sep 2023 0 repositories listed
-
Enhance audio generation controllability through representation similarity regularization15 Sep 2023 0 repositories listed
-
14 Sep 2023 0 repositories listed
-
Advances in machine-learning-based sampling motivated by lattice quantum chromodynamics3 Sep 2023 0 repositories listed
-
Audio Generation with Multiple Conditional Diffusion Model23 Aug 2023 0 repositories listed
-
IteraTTA: An interface for exploring both text prompts and audio priors in generating music with text-to-audio models24 Jul 2023 0 repositories listed
-
A Demand-Driven Perspective on Generative Audio AI10 Jul 2023 0 repositories listed
-
LM-VC: Zero-shot Voice Conversion via Speech Generation based on Language Models18 Jun 2023 0 repositories listed
-
DiffAVA: Personalized Text-to-Audio Generation with Visual Alignment22 May 2023 0 repositories listed
-
Leveraging Pre-trained AudioLDM for Sound Generation: A Benchmark Study7 Mar 2023 0 repositories listed
-
SingSong: Generating musical accompaniments from singing30 Jan 2023 0 repositories listed
-
SUPERB @ SLT 2022: Challenge on Generalization and Efficiency of Self-Supervised Speech Representation Learning16 Oct 2022 0 repositories listed
-
Audio Deepfake Attribution: An Initial Dataset and Investigation21 Aug 2022 0 repositories listed
-
Adversarial Audio Synthesis with Complex-valued Polynomial Networks14 Jun 2022 0 repositories listed
-
FlexLip: A Controllable Text-to-Lip System7 Jun 2022 0 repositories listed
-
On Target Representation in Continuous-output Neural Machine Translation1 May 2022 0 repositories listed
-
Streamable Neural Audio Synthesis With Non-Causal Convolutions14 Apr 2022 0 repositories listed
-
ADD 2022: the First Audio Deep Synthesis Detection Challenge17 Feb 2022 0 repositories listed
-
Soundify: Matching Sound Effects to Video17 Dec 2021 0 repositories listed
-
Geometry-Aware Multi-Task Learning for Binaural Audio Generation from Video21 Nov 2021 0 repositories listed
-
An investigation of pre-upsampling generative modelling and Generative Adversarial Networks in audio super resolution30 Sep 2021 0 repositories listed
-
Depth Infused Binaural Audio Generation using Hierarchical Cross-Modal Attention10 Aug 2021 0 repositories listed
-
CRASH: Raw Audio Score-based Generative Modeling for Controllable High-resolution Drum Sound Synthesis14 Jun 2021 0 repositories listed
-
Exploiting Audio-Visual Consistency with Partial Supervision for Spatial Audio Generation3 May 2021 0 repositories listed
-
Visually Informed Binaural Audio Generation without Binaural Audios13 Apr 2021 0 repositories listed
-
A Comprehensive Survey on Deep Music Generation: Multi-level Representations, Algorithms, Evaluations, and Future Directions13 Nov 2020 0 repositories listed
-
NU-GAN: High resolution neural upsampling with GAN22 Oct 2020 0 repositories listed
-
Audio Dequantization for High Fidelity Audio Generation in Flow-based Neural Vocoder16 Aug 2020 0 repositories listed
-
Incremental Text to Speech for Neural Sequence-to-Sequence Models using Reinforcement Learning7 Aug 2020 0 repositories listed
-
Neural Granular Sound Synthesis4 Aug 2020 0 repositories listed
-
Sep-Stereo: Visually Guided Stereophonic Audio Generation by Associating Source Separation20 Jul 2020 0 repositories listed
-
High-Fidelity Audio Generation and Representation Learning with Guided Adversarial Autoencoder1 Jun 2020 0 repositories listed
-
Guided Generative Adversarial Neural Network for Representation Learning and High Fidelity Audio Generation using Fewer Labelled Audio Data5 Mar 2020 0 repositories listed
-
Cross-modal variational inference for bijective signal-symbol translation10 Feb 2020 0 repositories listed
-
FastWave: Accelerating Autoregressive Convolutional Neural Networks on FPGA9 Feb 2020 0 repositories listed
-
Transferring neural speech waveform synthesizers to musical instrument sounds generation27 Oct 2019 0 repositories listed
-
Progressive Upsampling Audio Synthesis via Effective Adversarial Training25 Sep 2019 0 repositories listed
-
Voice command generation using Progressive Wavegans13 Mar 2019 0 repositories listed
-
CMCGAN: A Uniform Framework for Cross-Modal Visual-Audio Mutual Generation22 Nov 2017 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.