Evaluating Prompt Robustness in Text-to-Audio Systems for Adaptive Virtual Agents and Game Soundtracks
Researchers evaluated the robustness of text-to-audio models MusicGen-small, MusicGen-large, and Stable Audio 2.5 to prompt changes, finding that Stable Audio 2.5 performs better under structural rephrasing and has lower acoustic distances. This study highlights the importance of multi-seed robustness evaluation for adaptive game audio.
Save an API key to vote.