Voice Agents

Momentum

8 papers in the last four weeks, up 100% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 34

All topics
CardsList
  1. VAmoS Part Deux: Harder, More Realistic Voice-Agent Simulation

    Sep 29, 2026Joshua Meyer, Sahar Shayegan, Ritiz Tambi +5Voice AgentsDialogue Benchmarks

  2. Voices as Handles: Reasoning about Speaker Identity with Frozen Text LLMs

    Sep 29, 2026Runqiu Xu, Zhisheng Zheng, David HarwathVoice AgentsSpeaker

  3. Qwen-Audio-3.1-Realtime: Towards Reliable Agentic Voice Interaction

    Sep 21, 2026Lujia Bao, Qian Chen, Luyao Cheng +15Voice AgentsAudio Understanding

  4. MTVA-Bench: Evaluating the Language Model Inside Cascaded Voice Agents

    Sep 17, 2026Pritish Mishra, Ishaan Kumar, Akshat Mandoli +1Voice AgentsDialogue Benchmarks

  5. MP-Bench: Evaluating Voice Agents as a Multiparty Conversation Participant

    Sep 14, 2026Yi-Jen Shih, Shih-Yun Shan Kuan, Guan-Ting Lin +7Voice AgentsDialogue Benchmarks

  6. Benchmarking LLM Judges for Voice-Agent Evaluation: Reliability, Calibration, and Human Oversight

    Aug 25, 2026Anupam Purwar, Shashank Singh, Kritika SrivastavaLlm-As-A-JudgeLarge Language Model Judges

  7. DuplexWorld: Can voice agents help you get through the day?

    Aug 11, 2026Aryan Vijay Bhosale, Harshit Rajgarhia, Akhil Pothanapalli +3Voice AgentsMultilingual Agents

  8. FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

    Aug 6, 2026Aman Dalmia, Sanskriti Midha, Jigar DoshiVoice AgentsHindi

  9. Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models

    Aug 3, 2026Yuwen Wang, Tian-Hao Zhang, Minghao Cai +7Audio UnderstandingVoice Agents

  10. VAmoS Bench: Voice Agent Simulation Bench

    Jul 29, 2026Joshua Meyer, Sahar Shayegan, Ritiz Tambi +5Voice AgentsDialogue Benchmarks

  11. Designed Vocalizations Dataset: Sound-Designed Human and Animal Voices for Non-human Voice Conversion

    Jul 23, 2026Seolhee Lee, Minsu Kang, Yangsun Lee +3Voice ConversionVocalizations

  12. MCP Server Architecture Patterns for LLM-Integrated Applications

    Jun 29, 2026Carson Rodrigues, Oysturn VasModel Context ProtocolVoice Agents

  13. Real-Time Voice AI Hears but Does Not Listen

    Jun 24, 2026Martijn Bartelds, Federico Bianchi, James ZouVoice AgentsVocalizations

  14. CORTIS: Text-Only Adaptation of Spoken Language Models for Task-Oriented Voice Agents

    Jun 19, 2026Youngwon Choi, Hyeonyu Kim, Taeyoun Kwon +2Voice AgentsText-Only Adaptation

  15. IHBench: Evaluating Post-Interruption Recovery in Voice Agents with Structured Workflows

    Jun 17, 2026Ahmad Salimi, Wentao Ma, Yuzhi Tang +3Dialogue BenchmarksVoice Agents

  16. Exploring How Agent Voice Accents Shape Human-AI Collaboration in K-12 Group Learning

    Jun 11, 2026Prerna Ravi, Carúmey Stevens, Ben Hurt +3Voice AgentsCollaboration

  17. Audio-Mind: An Auditable Agentic Framework for Audio Understanding

    May 27, 2026Yucheng Wang, Jing Peng, Hanqi Li +6Audio UnderstandingVoice Agents

  18. MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes

    May 7, 2026Maximillian Chen, Xuanming Zhang, Michael Peng +3Voice AgentsSmart Homes

  19. Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment

    Apr 21, 2026Huanchen Cai, Sten TernströmText-To-Speech SynthesisVoice Agents

  20. Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection

    Apr 16, 2026Meng Chen, Kun Wang, Li Lu +2Large Audio Language ModelsHijacking

  21. Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

    Nov 10, 2025Vidya Srinivas, Zachary Englhardt, Vikram Iyer +1Full-Duplex Voice AgentsVoice Agents

  22. Robust Wake-Up Word Detection by Two-stage Multi-resolution Ensembles

    Oct 17, 2023Fernando López, Jordi Luque, Carlos Segura +1Sound Event DetectionUnsupervised Detection

  23. EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents

    Date pendingTara Bogavelli, Gabrielle Gauthier Melançon, Katrina Stankiewicz +10Voice AgentsEvaluation Benchmarks