Evaluating Prompt Robustness in Text-to-Audio Systems for Adaptive Virtual Agents and Game Soundtracks

Researchers evaluated the robustness of text-to-audio models MusicGen-small, MusicGen-large, and Stable Audio 2.5 to prompt changes, finding that Stable Audio 2.5 performs better under structural rephrasing and has lower acoustic distances. This study highlights the importance of multi-seed robustness evaluation for adaptive game audio.

RSS Score 0 9/16/2026, 4:00:00 AM Original Source
Save an API key to vote.