Multi-turn: both deliveries, actual released audio

Straight from the dataset (omniagentbench/OmniAgentBench/wild_long_scattered/mpcc/audio_tts_qwen3/2B). v1 is the single concatenated clip the evaluation used. v2 is the same instruction as separate per-turn clips, which the fixed engine can feed as a real conversation. Both files already ship in the release.