Artificial Intelligence #voice agents#post-interruption recovery
IHBench: Evaluating Post-Interruption Recovery in Voice Agents with Structured Workflows
IHBench, a new benchmark from researchers including Salimi et al., evaluates how voice agents recover after interruptions in structured enterprise workflows. The benchmark tests 27 audio-language models from OpenAI, Google, and the open-weight community, finding that closed-weight models are consistently more robust, degrading 3.3x more slowly in long conversations.
Jun 22, 2026 1 source