$τ$-Voice: Benchmarking Full-Duplex Voice Agents on Real-World Domains
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ray, Soham, Dhandhania, Keshav, Barres, Victor, Narasimhan, Karthik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
$τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
von: Barres, Victor, et al.
Veröffentlicht: (2025)
von: Barres, Victor, et al.
Veröffentlicht: (2025)
IntrinsicVoice: Empowering LLMs with Intrinsic Real-time Voice Interaction Abilities
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
i-LAVA: Insights on Low Latency Voice-2-Voice Architecture for Agents
von: Purwar, Anupam, et al.
Veröffentlicht: (2025)
von: Purwar, Anupam, et al.
Veröffentlicht: (2025)
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
von: Yao, Shunyu, et al.
Veröffentlicht: (2024)
von: Yao, Shunyu, et al.
Veröffentlicht: (2024)
Physics-Guided Deepfake Detection for Voice Authentication Systems
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)
Spectral Masking and Interpolation Attack (SMIA): A Black-box Adversarial Attack against Voice Authentication and Anti-Spoofing Systems
von: Kamel, Kamel, et al.
Veröffentlicht: (2025)
von: Kamel, Kamel, et al.
Veröffentlicht: (2025)
VoiceWukong: Benchmarking Deepfake Voice Detection
von: Yan, Ziwei, et al.
Veröffentlicht: (2024)
von: Yan, Ziwei, et al.
Veröffentlicht: (2024)
An Agent-Based Framework for Automated Higher-Voice Harmony Generation
von: Ganapathy, Nia D'Souza, et al.
Veröffentlicht: (2025)
von: Ganapathy, Nia D'Souza, et al.
Veröffentlicht: (2025)
From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench
von: Xu, Ke, et al.
Veröffentlicht: (2026)
von: Xu, Ke, et al.
Veröffentlicht: (2026)
Voila: Voice-Language Foundation Models for Real-Time Autonomous Interaction and Voice Role-Play
von: Shi, Yemin, et al.
Veröffentlicht: (2025)
von: Shi, Yemin, et al.
Veröffentlicht: (2025)
StyleStream: Real-Time Zero-Shot Voice Style Conversion
von: Liu, Yisi, et al.
Veröffentlicht: (2026)
von: Liu, Yisi, et al.
Veröffentlicht: (2026)
$τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge
von: Shi, Quan, et al.
Veröffentlicht: (2026)
von: Shi, Quan, et al.
Veröffentlicht: (2026)
Fairness-Aware Partial-label Domain Adaptation for Voice Classification of Parkinson's and ALS
von: Francesconi, Arianna, et al.
Veröffentlicht: (2026)
von: Francesconi, Arianna, et al.
Veröffentlicht: (2026)
VoiceBench: Benchmarking LLM-Based Voice Assistants
von: Chen, Yiming, et al.
Veröffentlicht: (2024)
von: Chen, Yiming, et al.
Veröffentlicht: (2024)
YingMusic-SVC: Real-World Robust Zero-Shot Singing Voice Conversion with Flow-GRPO and Singing-Specific Inductive Biases
von: Chen, Gongyu, et al.
Veröffentlicht: (2025)
von: Chen, Gongyu, et al.
Veröffentlicht: (2025)
Voice Privacy from an Attribute-based Perspective
von: Rahman, Mehtab Ur, et al.
Veröffentlicht: (2026)
von: Rahman, Mehtab Ur, et al.
Veröffentlicht: (2026)
Probabilistic Verification of Voice Anti-Spoofing Models
von: Kushnir, Evgeny, et al.
Veröffentlicht: (2026)
von: Kushnir, Evgeny, et al.
Veröffentlicht: (2026)
The First Voice Timbre Attribute Detection Challenge
von: Chen, Liping, et al.
Veröffentlicht: (2025)
von: Chen, Liping, et al.
Veröffentlicht: (2025)
MOSS-VoiceGenerator: Create Realistic Voices with Natural Language Descriptions
von: Huang, Kexin, et al.
Veröffentlicht: (2026)
von: Huang, Kexin, et al.
Veröffentlicht: (2026)
R2-SVC: Towards Real-World Robust and Expressive Zero-shot Singing Voice Conversion
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
An Intelligent AI glasses System with Multi-Agent Architecture for Real-Time Voice Processing and Task Execution
von: Chen, Sheng-Kai, et al.
Veröffentlicht: (2026)
von: Chen, Sheng-Kai, et al.
Veröffentlicht: (2026)
Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models
von: Riera, Pablo, et al.
Veröffentlicht: (2026)
von: Riera, Pablo, et al.
Veröffentlicht: (2026)
Self Voice Conversion as an Attack against Neural Audio Watermarking
von: Özer, Yigitcan, et al.
Veröffentlicht: (2026)
von: Özer, Yigitcan, et al.
Veröffentlicht: (2026)
Voice Biomarkers for Depression and Anxiety
von: Abramenko, Oleksii, et al.
Veröffentlicht: (2026)
von: Abramenko, Oleksii, et al.
Veröffentlicht: (2026)
EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions
von: Modi, Smit Nautambhai, et al.
Veröffentlicht: (2026)
von: Modi, Smit Nautambhai, et al.
Veröffentlicht: (2026)
LTS-VoiceAgent: A Listen-Think-Speak Framework for Efficient Streaming Voice Interaction via Semantic Triggering and Incremental Reasoning
von: Zou, Wenhao, et al.
Veröffentlicht: (2026)
von: Zou, Wenhao, et al.
Veröffentlicht: (2026)
A Real-Time Voice Activity Detection Based On Lightweight Neural
von: Jia, Jidong, et al.
Veröffentlicht: (2024)
von: Jia, Jidong, et al.
Veröffentlicht: (2024)
Neural Multi-Speaker Voice Cloning for Nepali in Low-Resource Settings
von: Shrestha, Aayush M., et al.
Veröffentlicht: (2026)
von: Shrestha, Aayush M., et al.
Veröffentlicht: (2026)
DAST: A Dual-Stream Voice Anonymization Attacker with Staged Training
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2026)
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2026)
UniSS: Unified Expressive Speech-to-Speech Translation with Your Voice
von: Cheng, Sitong, et al.
Veröffentlicht: (2025)
von: Cheng, Sitong, et al.
Veröffentlicht: (2025)
SegReConcat: A Data Augmentation Method for Voice Anonymization Attack
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2025)
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2025)
Voiced-Aware Style Extraction and Style Direction Adjustment for Expressive Text-to-Speech
von: Kim, Nam-Gyu
Veröffentlicht: (2025)
von: Kim, Nam-Gyu
Veröffentlicht: (2025)
VoiceGRPO: Modern MoE Transformers with Group Relative Policy Optimization GRPO for AI Voice Health Care Applications on Voice Pathology Detection
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
von: Togootogtokh, Enkhtogtokh, et al.
Veröffentlicht: (2025)
VoicePrompter: Robust Zero-Shot Voice Conversion with Voice Prompt and Conditional Flow Matching
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
von: Choi, Ha-Yeong, et al.
Veröffentlicht: (2025)
AI-Driven Acoustic Voice Biomarker-Based Hierarchical Classification of Benign Laryngeal Voice Disorders from Sustained Vowels
von: Annabestani, Mohsen, et al.
Veröffentlicht: (2025)
von: Annabestani, Mohsen, et al.
Veröffentlicht: (2025)
Voice Cloning: Comprehensive Survey
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
von: Azzuni, Hussam, et al.
Veröffentlicht: (2025)
Proactive Detection of Voice Cloning with Localized Watermarking
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
von: Roman, Robin San, et al.
Veröffentlicht: (2024)
Voices of the Mountains: Deep Learning-Based Vocal Error Detection System for Kurdish Maqams
von: Khairaldeen, Darvan Shvan, et al.
Veröffentlicht: (2026)
von: Khairaldeen, Darvan Shvan, et al.
Veröffentlicht: (2026)
VividVoice: A Unified Framework for Scene-Aware Visually-Driven Speech Synthesis
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyuan, et al.
Veröffentlicht: (2026)
A Lightweight Pipeline for Noisy Speech Voice Cloning and Accurate Lip Sync Synthesis
von: Amir, Javeria, et al.
Veröffentlicht: (2025)
von: Amir, Javeria, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
$τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
von: Barres, Victor, et al.
Veröffentlicht: (2025) -
IntrinsicVoice: Empowering LLMs with Intrinsic Real-time Voice Interaction Abilities
von: Zhang, Xin, et al.
Veröffentlicht: (2024) -
i-LAVA: Insights on Low Latency Voice-2-Voice Architecture for Agents
von: Purwar, Anupam, et al.
Veröffentlicht: (2025) -
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
von: Yao, Shunyu, et al.
Veröffentlicht: (2024) -
Physics-Guided Deepfake Detection for Voice Authentication Systems
von: Mohammadi, Alireza, et al.
Veröffentlicht: (2025)