Talk2Agent: Benchmarking Voice Interfaces for Text Agents
Talk2Agent is a benchmark for evaluating how effectively voice interfaces convey human-spoken instructions to LLM-based computer-use agents. It builds human-spoken versions of tasks and evaluates various voice interfaces, including ASR models, audio-capable LLMs, and contextual biasing. Talk2Agent proposes an execution-free, task-conditioned evaluation framework to measure task-relevant information retention after voice interfaces.
Save an API key to vote.