VoiceAgentRAG: Solving the RAG Latency Bottleneck in Real-Time Voice Agents Using Dual-Agent Architectures
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Jielin, Zhang, Jianguo, Chen, Zixiang, Yang, Liangwei, Zhu, Ming, Tan, Juntao, Chen, Haolin, Zhao, Wenting, Murthy, Rithesh, Ram, Roshan, Prabhakar, Akshara, Heinecke, Shelby, Xiong, Caiming, Savarese, Silvio, Wang, Huan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Building Enterprise Realtime Voice Agents from Scratch: A Technical Tutorial
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
AudioCapBench: Quick Evaluation on Audio Captioning across Sound, Music, and Speech
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
Whisper-AuT: Domain-Adapted Audio Encoder for Efficient Audio-LLM Training
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
Enterprise Sales Copilot: Enabling Real-Time AI Support with Automatic Information Retrieval in Live Sales Calls
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
von: Qiu, Jielin, et al.
Veröffentlicht: (2026)
RealUserSim: Bridging the Reality Gap in Agent Benchmarking via Grounded User Simulation
von: Zhu, Ming, et al.
Veröffentlicht: (2026)
von: Zhu, Ming, et al.
Veröffentlicht: (2026)
LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering
von: Qiu, Jielin, et al.
Veröffentlicht: (2025)
von: Qiu, Jielin, et al.
Veröffentlicht: (2025)
Prompt Optimization Via Diffusion Language Models
von: Wang, Shiyu, et al.
Veröffentlicht: (2026)
von: Wang, Shiyu, et al.
Veröffentlicht: (2026)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
von: Murthy, Rithesh, et al.
Veröffentlicht: (2025)
von: Murthy, Rithesh, et al.
Veröffentlicht: (2025)
Position: Vector Prompt Interfaces Should Be Exposed to Enable Customization of Large Language Models
von: Yang, Liangwei, et al.
Veröffentlicht: (2026)
von: Yang, Liangwei, et al.
Veröffentlicht: (2026)
Enterprise Deep Research: Steerable Multi-Agent Deep Research for Enterprise Analytics
von: Prabhakar, Akshara, et al.
Veröffentlicht: (2025)
von: Prabhakar, Akshara, et al.
Veröffentlicht: (2025)
Test-Time Adaptation for LLM Agents via Environment Interaction
von: Chen, Arthur, et al.
Veröffentlicht: (2025)
von: Chen, Arthur, et al.
Veröffentlicht: (2025)
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
UserBench: An Interactive Gym Environment for User-Centric Agents
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
UserRL: Training Interactive User-Centric Agent via Reinforcement Learning
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
PRACT: Optimizing Principled Reasoning and Acting of LLM Agent
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
xLAM: A Family of Large Action Models to Empower AI Agent Systems
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2024)
AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
LoCoBench: A Benchmark for Long-Context Large Language Models in Complex Software Engineering
von: Qiu, Jielin, et al.
Veröffentlicht: (2025)
von: Qiu, Jielin, et al.
Veröffentlicht: (2025)
Personalized Multi-task Training for Recommender System
von: Yang, Liangwei, et al.
Veröffentlicht: (2024)
von: Yang, Liangwei, et al.
Veröffentlicht: (2024)
APIGen-MT: Agentic Pipeline for Multi-Turn Data Generation via Simulated Agent-Human Interplay
von: Prabhakar, Akshara, et al.
Veröffentlicht: (2025)
von: Prabhakar, Akshara, et al.
Veröffentlicht: (2025)
Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization
von: Yao, Weiran, et al.
Veröffentlicht: (2023)
von: Yao, Weiran, et al.
Veröffentlicht: (2023)
xRouter: Training Cost-Aware LLMs Orchestration System via Reinforcement Learning
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Entropy-Based Block Pruning for Efficient Large Language Models
von: Yang, Liangwei, et al.
Veröffentlicht: (2025)
von: Yang, Liangwei, et al.
Veröffentlicht: (2025)
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
GeoGNN: Quantifying and Mitigating Semantic Drift in Text-Attributed Graphs
von: Yang, Liangwei, et al.
Veröffentlicht: (2025)
von: Yang, Liangwei, et al.
Veröffentlicht: (2025)
PersonaBench: Evaluating AI Models on Understanding Personal Information through Accessing (Synthetic) Private User Data
von: Tan, Juntao, et al.
Veröffentlicht: (2025)
von: Tan, Juntao, et al.
Veröffentlicht: (2025)
REX: Rapid Exploration and eXploitation for AI Agents
von: Murthy, Rithesh, et al.
Veröffentlicht: (2023)
von: Murthy, Rithesh, et al.
Veröffentlicht: (2023)
ToolLibGen: Scalable Automatic Tool Creation and Aggregation for LLM Reasoning
von: Yue, Murong, et al.
Veröffentlicht: (2025)
von: Yue, Murong, et al.
Veröffentlicht: (2025)
ToolScan: A Benchmark for Characterizing Errors in Tool-Use LLMs
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
von: Kokane, Shirley, et al.
Veröffentlicht: (2024)
Diversity Empowers Intelligence: Integrating Expertise of Software Engineering Agents
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
von: Zhang, Kexun, et al.
Veröffentlicht: (2024)
ActionStudio: A Lightweight Framework for Data and Training of Large Action Models
von: Zhang, Jianguo, et al.
Veröffentlicht: (2025)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2025)
CRMArena: Understanding the Capacity of LLM Agents to Perform Professional CRM Tasks in Realistic Environments
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2024)
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2024)
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
von: Liu, Zuxin, et al.
Veröffentlicht: (2024)
CRMArena-Pro: Holistic Assessment of LLM Agents Across Diverse Business Scenarios and Interactions
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2025)
von: Huang, Kung-Hsiang, et al.
Veröffentlicht: (2025)
MobileAIBench: Benchmarking LLMs and LMMs for On-Device Use Cases
von: Murthy, Rithesh, et al.
Veröffentlicht: (2024)
von: Murthy, Rithesh, et al.
Veröffentlicht: (2024)
i-LAVA: Insights on Low Latency Voice-2-Voice Architecture for Agents
von: Purwar, Anupam, et al.
Veröffentlicht: (2025)
von: Purwar, Anupam, et al.
Veröffentlicht: (2025)
CoDA: Coding LM via Diffusion Adaptation
von: Chen, Haolin, et al.
Veröffentlicht: (2025)
von: Chen, Haolin, et al.
Veröffentlicht: (2025)
Asynchronous Tool Usage for Real-Time Agents
von: Ginart, Antonio A., et al.
Veröffentlicht: (2024)
von: Ginart, Antonio A., et al.
Veröffentlicht: (2024)
CBM-RAG: Demonstrating Enhanced Interpretability in Radiology Report Generation with Multi-Agent RAG and Concept Bottleneck Models
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2025)
von: Alam, Hasan Md Tusfiqur, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Building Enterprise Realtime Voice Agents from Scratch: A Technical Tutorial
von: Qiu, Jielin, et al.
Veröffentlicht: (2026) -
AudioCapBench: Quick Evaluation on Audio Captioning across Sound, Music, and Speech
von: Qiu, Jielin, et al.
Veröffentlicht: (2026) -
Whisper-AuT: Domain-Adapted Audio Encoder for Efficient Audio-LLM Training
von: Qiu, Jielin, et al.
Veröffentlicht: (2026) -
Enterprise Sales Copilot: Enabling Real-Time AI Support with Automatic Information Retrieval in Live Sales Calls
von: Qiu, Jielin, et al.
Veröffentlicht: (2026) -
RealUserSim: Bridging the Reality Gap in Agent Benchmarking via Grounded User Simulation
von: Zhu, Ming, et al.
Veröffentlicht: (2026)