Better Slow than Sorry: Introducing Positive Friction for Reliable Dialogue Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | İnan, Mert, Sicilia, Anthony, Dey, Suvodip, Dongre, Vardhan, Srinivasan, Tejas, Thomason, Jesse, Tür, Gökhan, Hakkani-Tür, Dilek, Alikhani, Malihe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue
por: Dongre, Vardhan, et al.
Publicado: (2026)
por: Dongre, Vardhan, et al.
Publicado: (2026)
Simulating User Agents for Embodied Conversational-AI
por: Philipov, Daniel, et al.
Publicado: (2024)
por: Philipov, Daniel, et al.
Publicado: (2024)
Plan Verification for LLM-Based Embodied Task Completion Agents
por: Hariharan, Ananth, et al.
Publicado: (2025)
por: Hariharan, Ananth, et al.
Publicado: (2025)
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents
por: Dongre, Vardhan, et al.
Publicado: (2024)
por: Dongre, Vardhan, et al.
Publicado: (2024)
Confidence Estimation for LLM-Based Dialogue State Tracking
por: Sun, Yi-Jyun, et al.
Publicado: (2024)
por: Sun, Yi-Jyun, et al.
Publicado: (2024)
Know Your Mistakes: Towards Preventing Overreliance on Task-Oriented Conversational AI Through Accountability Modeling
por: Dey, Suvodip, et al.
Publicado: (2025)
por: Dey, Suvodip, et al.
Publicado: (2025)
TD-EVAL: Revisiting Task-Oriented Dialogue Evaluation by Combining Turn-Level Precision with Dialogue-Level Comparisons
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
MIRAGE: A Benchmark for Multimodal Information-Seeking and Reasoning in Agricultural Expert-Guided Conversations
por: Dongre, Vardhan, et al.
Publicado: (2025)
por: Dongre, Vardhan, et al.
Publicado: (2025)
Accounting for Sycophancy in Language Model Uncertainty Estimation
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems
por: Kazi, Taaha, et al.
Publicado: (2024)
por: Kazi, Taaha, et al.
Publicado: (2024)
A Desideratum for Conversational Agents: Capabilities, Challenges, and Future Directions
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
Generating Signed Language Instructions in Large-Scale Dialogue Systems
por: İnan, Mert, et al.
Publicado: (2024)
por: İnan, Mert, et al.
Publicado: (2024)
From Context to Action: Analysis of the Impact of State Representation and Context on the Generalization of Multi-Turn Web Navigation Agents
por: Tiwary, Nalin, et al.
Publicado: (2024)
por: Tiwary, Nalin, et al.
Publicado: (2024)
Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models
por: Bozdag, Nimet Beyza, et al.
Publicado: (2025)
por: Bozdag, Nimet Beyza, et al.
Publicado: (2025)
SiLVERScore: Semantically-Aware Embeddings for Sign Language Generation Evaluation
por: Imai, Saki, et al.
Publicado: (2025)
por: Imai, Saki, et al.
Publicado: (2025)
Measuring How (Not Just Whether) VLMs Build Common Ground
por: Imai, Saki, et al.
Publicado: (2025)
por: Imai, Saki, et al.
Publicado: (2025)
ATOD: An Evaluation Framework and Benchmark for Agentic Task-Oriented Dialogue Systems
por: Zhang, Yifei, et al.
Publicado: (2026)
por: Zhang, Yifei, et al.
Publicado: (2026)
Language Specific Knowledge: Do Models Know Better in X than in English?
por: Agarwal, Ishika, et al.
Publicado: (2025)
por: Agarwal, Ishika, et al.
Publicado: (2025)
Do LLMs Encode Functional Importance of Reasoning Tokens?
por: Singh, Janvijay, et al.
Publicado: (2026)
por: Singh, Janvijay, et al.
Publicado: (2026)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
por: Agarwal, Ishika, et al.
Publicado: (2025)
por: Agarwal, Ishika, et al.
Publicado: (2025)
When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction
por: Dongre, Vardhan, et al.
Publicado: (2026)
por: Dongre, Vardhan, et al.
Publicado: (2026)
Self-Improving LLM Agents at Test-Time
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
por: Acikgoz, Emre Can, et al.
Publicado: (2025)
From Fact to Judgment: Investigating the Impact of Task Framing on LLM Conviction in Dialogue Systems
por: Rabbani, Parisa, et al.
Publicado: (2025)
por: Rabbani, Parisa, et al.
Publicado: (2025)
YourBench: Easy Custom Evaluation Sets for Everyone
por: Shashidhar, Sumuk, et al.
Publicado: (2025)
por: Shashidhar, Sumuk, et al.
Publicado: (2025)
Goal Alignment in LLM-Based User Simulators for Conversational AI
por: Mehri, Shuhaib, et al.
Publicado: (2025)
por: Mehri, Shuhaib, et al.
Publicado: (2025)
Drift No More? Context Equilibria in Multi-Turn LLM Interactions
por: Dongre, Vardhan, et al.
Publicado: (2025)
por: Dongre, Vardhan, et al.
Publicado: (2025)
Tool-R0: Self-Evolving LLM Agents for Tool-Learning from Zero Data
por: Acikgoz, Emre Can, et al.
Publicado: (2026)
por: Acikgoz, Emre Can, et al.
Publicado: (2026)
SIMU: Selective Influence Machine Unlearning
por: Agarwal, Anu, et al.
Publicado: (2025)
por: Agarwal, Anu, et al.
Publicado: (2025)
Question Generation for Assessing Early Literacy Reading Comprehension
por: Yang, Xiaocheng, et al.
Publicado: (2025)
por: Yang, Xiaocheng, et al.
Publicado: (2025)
Identifying & Interactively Refining Ambiguous User Goals for Data Visualization Code Generation
por: İnan, Mert, et al.
Publicado: (2025)
por: İnan, Mert, et al.
Publicado: (2025)
Evaluating Theory of (an uncertain) Mind: Predicting the Uncertain Beliefs of Others in Conversation Forecasting
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
Eliciting Uncertainty in Chain-of-Thought to Mitigate Bias against Forecasting Harmful User Behaviors
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance
por: Srinivasan, Tejas, et al.
Publicado: (2025)
por: Srinivasan, Tejas, et al.
Publicado: (2025)
Learning to Generate Context-Sensitive Backchannel Smiles for Embodied AI Agents with Applications in Mental Health Dialogues
por: Bilalpur, Maneesh, et al.
Publicado: (2024)
por: Bilalpur, Maneesh, et al.
Publicado: (2024)
Including Facial Expressions in Contextual Embeddings for Sign Language Generation
por: Viegas, Carla, et al.
Publicado: (2022)
por: Viegas, Carla, et al.
Publicado: (2022)
ReIn: Conversational Error Recovery with Reasoning Inception
por: Kim, Takyoung, et al.
Publicado: (2026)
por: Kim, Takyoung, et al.
Publicado: (2026)
Dialog Flow Induction for Constrainable LLM-Based Chatbots
por: Agrawal, Stuti, et al.
Publicado: (2024)
por: Agrawal, Stuti, et al.
Publicado: (2024)
ToolRL: Reward is All Tool Learning Needs
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
SMART: Self-Aware Agent for Tool Overuse Mitigation
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
Beyond Sample-Level Feedback: Using Reference-Level Feedback to Guide Data Synthesis
por: Mehri, Shuhaib, et al.
Publicado: (2025)
por: Mehri, Shuhaib, et al.
Publicado: (2025)
Ejemplares similares
-
Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue
por: Dongre, Vardhan, et al.
Publicado: (2026) -
Simulating User Agents for Embodied Conversational-AI
por: Philipov, Daniel, et al.
Publicado: (2024) -
Plan Verification for LLM-Based Embodied Task Completion Agents
por: Hariharan, Ananth, et al.
Publicado: (2025) -
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents
por: Dongre, Vardhan, et al.
Publicado: (2024) -
Confidence Estimation for LLM-Based Dialogue State Tracking
por: Sun, Yi-Jyun, et al.
Publicado: (2024)