Side-by-side experimentation of native speech-to-speech models on the same scenario.
Pick an arm and a scenario, connect, then talk.