EMID: An Emotional Aligned Dataset in Audio-Visual Modality Paper • 2308.07622 • Published Aug 15, 2023 • 1
WritingBench: A Comprehensive Benchmark for Generative Writing Paper • 2503.05244 • Published Mar 7, 2025 • 22
MM-StoryAgent: Immersive Narrated Storybook Video Generation with a Multi-Agent Paradigm across Text, Image and Audio Paper • 2503.05242 • Published Mar 7, 2025 • 1
LARA-Gen: Enabling Continuous Emotion Control for Music Generation Models via Latent Affective Representation Alignment Paper • 2510.05875 • Published Oct 7, 2025
DashengTokenizer: One layer is enough for unified audio understanding and generation Paper • 2602.23765 • Published Feb 27
Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text Paper • 2605.27838 • Published May 27
UniFlow-Audio: Unified Flow Matching for Audio Generation from Omni-Modalities Paper • 2509.24391 • Published Sep 29, 2025
DashengAudioGen Collection Audio scene generators for Text-to-Speech + Text-to-Music + Text-to-Sound,, • 2 items • Updated Jun 1