Browse State-of-the-Art › Audio Generation › Papers, page 2
Audio Generation
Papers archive 2025-07-28
archive papers tagged: 270 · with a code link: 124 · where Syntology ran a sample: 57 (49 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (57 of 270 tagged: 49 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 270, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
29 Mar 2023 1 repository listed
-
4 Feb 2023 1 repository listed
-
30 Jan 2023 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
30 Sep 2022 1 repository listed
-
20 Jul 2022 1 repository listed
-
10 May 2022 1 repository listed Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
25 Oct 2021 1 repository listed
-
11 Jul 2021 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
11 Jun 2021 1 repository listed
-
11 Jun 2021 1 repository listed
-
23 Feb 2021 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
1 Jan 2021 1 repository listed
-
3 Dec 2020 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
23 Jun 2020 1 repository listed
-
11 Jun 2020 1 repository listed
-
18 May 2020 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
26 Dec 2019 1 repository listed
-
27 Nov 2019 1 repository listed
-
14 Nov 2019 1 repository listed Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
11 Feb 2019 1 repository listed
-
29 Oct 2018 1 repository listed
-
27 Sep 2018 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples)
-
27 Aug 2018 1 repository listed
-
FreeAudio: Training-Free Timing Planning for Controllable Long-Form Text-to-Audio Generation11 Jul 2025 0 repositories listed
-
Step-by-Step Video-to-Audio Synthesis via Negative Audio Guidance26 Jun 2025 0 repositories listed
-
Kling-Foley: Multimodal Diffusion Transformer for High-Quality Video-to-Audio Generation24 Jun 2025 0 repositories listed
-
LiLAC: A Lightweight Latent ControlNet for Musical Audio Generation13 Jun 2025 0 repositories listed
-
ViSAGe: Video-to-Spatial Audio Generation13 Jun 2025 0 repositories listed
-
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations6 Jun 2025 0 repositories listed
-
Sounding that Object: Interactive Object-Aware Image to Audio Generation4 Jun 2025 0 repositories listed
-
DGMO: Training-Free Audio Source Separation through Diffusion-Guided Mask Optimization3 Jun 2025 0 repositories listed
-
InfiniteAudio: Infinite-Length Audio Generation with Consistency3 Jun 2025 0 repositories listed
-
IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling31 May 2025 0 repositories listed
-
AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion28 May 2025 0 repositories listed
-
EnvSDD: Benchmarking Environmental Sound Deepfake Detection25 May 2025 0 repositories listed
-
MMMG: a Comprehensive and Reliable Evaluation Suite for Multitask Multimodal Generation23 May 2025 0 repositories listed
-
Unified Cross-modal Translation of Score Images, Symbolic Music, and Performance Audio19 May 2025 0 repositories listed
-
DPN-GAN: Inducing Periodic Activations in Generative Adversarial Networks for High-Fidelity Audio Synthesis14 May 2025 0 repositories listed
-
TACOS: Temporally-aligned Audio CaptiOnS for Language-Audio Pretraining12 May 2025 0 repositories listed
-
Discrete Optimal Transport and Voice Conversion7 May 2025 0 repositories listed
-
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients6 May 2025 0 repositories listed
-
On the Design of Diffusion-based Neural Speech Codecs11 Apr 2025 0 repositories listed
-
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models6 Apr 2025 0 repositories listed
-
DeepAudio-V1:Towards Multi-Modal Multi-Stage End-to-End Video to Speech and Audio Generation28 Mar 2025 0 repositories listed
-
DeepSound-V1: Start to Think Step-by-Step in the Audio Generation from Videos28 Mar 2025 0 repositories listed
-
Make Some Noise: Towards LLM audio reasoning and generation using sound tokens28 Mar 2025 0 repositories listed
-
DiffGAP: A Lightweight Diffusion Module in Contrastive Space for Bridging Cross-Model Gap15 Mar 2025 0 repositories listed
-
AudioX: Diffusion Transformer for Anything-to-Audio Generation13 Mar 2025 0 repositories listed
-
TA-V2A: Textually Assisted Video-to-Audio Generation12 Mar 2025 0 repositories listed
-
ReelWave: Multi-Agentic Movie Sound Generation through Multimodal LLM Conversation10 Mar 2025 0 repositories listed
-
Synchronized Video-to-Audio Generation via Mel Quantization-Continuum Decomposition10 Mar 2025 0 repositories listed
-
Speech Audio Generation from dynamic MRI via a Knowledge Enhanced Conditional Variational Autoencoder9 Mar 2025 0 repositories listed
-
DualSpec: Text-to-spatial-audio Generation via Dual-Spectrogram Guided Diffusion Model26 Feb 2025 0 repositories listed
-
Towards efficient quantum algorithms for diffusion probability models20 Feb 2025 0 repositories listed
-
AudioSpa: Spatializing Sound Events with Text16 Feb 2025 0 repositories listed
-
UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation6 Feb 2025 0 repositories listed
-
CosyAudio: Improving Audio Generation with Confidence Scores and Synthetic Captions28 Jan 2025 0 repositories listed
-
CCStereo: Audio-Visual Contextual and Contrastive Learning for Binaural Audio Generation6 Jan 2025 0 repositories listed
-
Nonparametric estimation of a factorizable density using diffusion models3 Jan 2025 0 repositories listed
-
Animate and Sound an Image1 Jan 2025 0 repositories listed
-
Foley-Flow: Coordinated Video-to-Audio Generation with Masked Audio-Visual Alignment and Dynamic Conditional Flows1 Jan 2025 0 repositories listed
-
Tri-Ergon: Fine-grained Video-to-Audio Generation with Multi-modal Conditions and LUFS Control29 Dec 2024 0 repositories listed
-
VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis26 Dec 2024 0 repositories listed
-
Smooth-Foley: Creating Continuous Sound for Video-to-Audio Generation Under Semantic Guidance24 Dec 2024 0 repositories listed
-
FolAI: Synchronized Foley Sound Generation with Semantic and Temporal Alignment19 Dec 2024 0 repositories listed
-
VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation14 Dec 2024 0 repositories listed
-
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls12 Dec 2024 0 repositories listed
-
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding5 Dec 2024 0 repositories listed
-
Continuous Autoregressive Models with Noise Augmentation Avoid Error Accumulation27 Nov 2024 0 repositories listed
-
Video-Guided Foley Sound Generation with Multimodal Controls26 Nov 2024 0 repositories listed
-
PerceiverS: A Multi-Scale Perceiver with Effective Segmentation for Long-Term Expressive Symbolic Music Generation13 Nov 2024 0 repositories listed
-
Audiobox TTA-RAG: Improving Zero-Shot and Few-Shot Text-To-Audio with Retrieval-Augmented Generation7 Nov 2024 0 repositories listed
-
The NPU-HWC System for the ISCSLP 2024 Inspirational and Convincing Audio Generation Challenge31 Oct 2024 0 repositories listed
-
Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation14 Oct 2024 0 repositories listed
-
Audio-Agent: Leveraging LLMs For Audio Generation, Editing and Composition4 Oct 2024 0 repositories listed
-
Did You Hear That? Introducing AADG: A Framework for Generating Benchmark Data in Audio Anomaly Detection4 Oct 2024 0 repositories listed
-
MDSGen: Fast and Efficient Masked Diffusion Temporal-Aware Transformers for Open-Domain Sound Generation3 Oct 2024 0 repositories listed
-
Analyzing and Mitigating Inconsistency in Discrete Audio Tokens for Neural Codec Language Models28 Sep 2024 0 repositories listed
-
Video-to-Audio Generation with Fine-grained Temporal Semantics23 Sep 2024 0 repositories listed
-
PTQ4ADM: Post-Training Quantization for Efficient Text Conditional Audio Diffusion Models20 Sep 2024 0 repositories listed
-
AudioComposer: Towards Fine-grained Audio Generation with Natural Language Descriptions19 Sep 2024 0 repositories listed
-
NDVQ: Robust Neural Audio Codec with Normal Distribution-Based Vector Quantization19 Sep 2024 0 repositories listed
-
EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer17 Sep 2024 0 repositories listed
-
Learning Source Disentanglement in Neural Audio Codec17 Sep 2024 0 repositories listed
-
Text Prompt is Not Enough: Sound Event Enhanced Prompt Adapter for Target Style Audio Generation14 Sep 2024 0 repositories listed
-
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization5 Sep 2024 0 repositories listed
-
Applications and Advances of Artificial Intelligence in Music Generation:A Review3 Sep 2024 0 repositories listed
-
Video-Foley: Two-Stage Video-To-Sound Generation via Temporal Event Condition For Foley Sound21 Aug 2024 0 repositories listed
-
Demystifying the Communication Characteristics for Distributed Transformer Models19 Aug 2024 0 repositories listed
-
Connective Viewpoints of Signal-to-Noise Diffusion Models8 Aug 2024 0 repositories listed
-
Braille-to-Speech Generator: Audio Generation Based on Joint Fine-Tuning of CLIP and Fastspeech219 Jul 2024 0 repositories listed
-
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control18 Jul 2024 0 repositories listed
-
Modeling and Driving Human Body Soundfields through Acoustic Primitives18 Jul 2024 0 repositories listed
-
Video-to-Audio Generation with Hidden Alignment10 Jul 2024 0 repositories listed
-
SOAF: Scene Occlusion-aware Neural Acoustic Field2 Jul 2024 0 repositories listed
-
Provable Statistical Rates for Consistency Diffusion Models23 Jun 2024 0 repositories listed
-
Action2Sound: Ambient-Aware Generation of Action Sounds from Egocentric Videos13 Jun 2024 0 repositories listed
-
Codecfake: An Initial Dataset for Detecting LLM-based Deepfake Audio12 Jun 2024 0 repositories listed
-
AV-DiT: Efficient Audio-Visual Diffusion Transformer for Joint Audio and Video Generation11 Jun 2024 0 repositories listed
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.