Goal Alignment in LLM-Based User Simulators for Conversational AI
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mehri, Shuhaib, Yang, Xiaocheng, Kim, Takyoung, Tur, Gokhan, Mehri, Shikib, Hakkani-Tür, Dilek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models
von: Bozdag, Nimet Beyza, et al.
Veröffentlicht: (2025)
von: Bozdag, Nimet Beyza, et al.
Veröffentlicht: (2025)
Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2026)
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2026)
User Preference Modeling for Conversational LLM Agents: Weak Rewards from Retrieval-Augmented Interaction
von: Hao, Yuren, et al.
Veröffentlicht: (2026)
von: Hao, Yuren, et al.
Veröffentlicht: (2026)
AURA: A Diagnostic Framework for Tracking User Satisfaction of Interactive Planning Agents
von: Kim, Takyoung, et al.
Veröffentlicht: (2025)
von: Kim, Takyoung, et al.
Veröffentlicht: (2025)
Simulating User Agents for Embodied Conversational-AI
von: Philipov, Daniel, et al.
Veröffentlicht: (2024)
von: Philipov, Daniel, et al.
Veröffentlicht: (2024)
MultiSessionCollab: Learning User Preferences with Memory to Improve Long-Term Collaboration
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2026)
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2026)
Must Read: A Comprehensive Survey of Computational Persuasion
von: Bozdag, Nimet Beyza, et al.
Veröffentlicht: (2025)
von: Bozdag, Nimet Beyza, et al.
Veröffentlicht: (2025)
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents
von: Dongre, Vardhan, et al.
Veröffentlicht: (2024)
von: Dongre, Vardhan, et al.
Veröffentlicht: (2024)
ReIn: Conversational Error Recovery with Reasoning Inception
von: Kim, Takyoung, et al.
Veröffentlicht: (2026)
von: Kim, Takyoung, et al.
Veröffentlicht: (2026)
Confidence Estimation for LLM-Based Dialogue State Tracking
von: Sun, Yi-Jyun, et al.
Veröffentlicht: (2024)
von: Sun, Yi-Jyun, et al.
Veröffentlicht: (2024)
Beyond Sample-Level Feedback: Using Reference-Level Feedback to Guide Data Synthesis
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2025)
TD-EVAL: Revisiting Task-Oriented Dialogue Evaluation by Combining Turn-Level Precision with Dialogue-Level Comparisons
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
Question Generation for Assessing Early Literacy Reading Comprehension
von: Yang, Xiaocheng, et al.
Veröffentlicht: (2025)
von: Yang, Xiaocheng, et al.
Veröffentlicht: (2025)
Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems
von: Kazi, Taaha, et al.
Veröffentlicht: (2024)
von: Kazi, Taaha, et al.
Veröffentlicht: (2024)
PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators
von: Hoang, Nguyen Khoi, et al.
Veröffentlicht: (2026)
von: Hoang, Nguyen Khoi, et al.
Veröffentlicht: (2026)
Measuring and Mitigating the Distributional Gap Between Real and Simulated User Behaviors
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2026)
von: Mehri, Shuhaib, et al.
Veröffentlicht: (2026)
Self-Improving LLM Agents at Test-Time
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
von: Agarwal, Ishika, et al.
Veröffentlicht: (2025)
von: Agarwal, Ishika, et al.
Veröffentlicht: (2025)
Do LLMs Encode Functional Importance of Reasoning Tokens?
von: Singh, Janvijay, et al.
Veröffentlicht: (2026)
von: Singh, Janvijay, et al.
Veröffentlicht: (2026)
Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue
von: Dongre, Vardhan, et al.
Veröffentlicht: (2026)
von: Dongre, Vardhan, et al.
Veröffentlicht: (2026)
Plan Verification for LLM-Based Embodied Task Completion Agents
von: Hariharan, Ananth, et al.
Veröffentlicht: (2025)
von: Hariharan, Ananth, et al.
Veröffentlicht: (2025)
A Desideratum for Conversational Agents: Capabilities, Challenges, and Future Directions
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
Know Your Mistakes: Towards Preventing Overreliance on Task-Oriented Conversational AI Through Accountability Modeling
von: Dey, Suvodip, et al.
Veröffentlicht: (2025)
von: Dey, Suvodip, et al.
Veröffentlicht: (2025)
MAC: A Multi-Agent Framework for Interactive User Clarification in Multi-turn Conversations
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
MIRAGE: A Benchmark for Multimodal Information-Seeking and Reasoning in Agricultural Expert-Guided Conversations
von: Dongre, Vardhan, et al.
Veröffentlicht: (2025)
von: Dongre, Vardhan, et al.
Veröffentlicht: (2025)
Can a Single Model Master Both Multi-turn Conversations and Tool Use? CoALM: A Unified Conversational Agentic Language Model
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
YourBench: Easy Custom Evaluation Sets for Everyone
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2025)
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2025)
ATOD: An Evaluation Framework and Benchmark for Agentic Task-Oriented Dialogue Systems
von: Zhang, Yifei, et al.
Veröffentlicht: (2026)
von: Zhang, Yifei, et al.
Veröffentlicht: (2026)
Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic Interactions
von: Jang, Jihyoung, et al.
Veröffentlicht: (2025)
von: Jang, Jihyoung, et al.
Veröffentlicht: (2025)
ToolRL: Reward is All Tool Learning Needs
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
SMART: Self-Aware Agent for Tool Overuse Mitigation
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Unsupervised Human Preference Learning
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2024)
von: Shashidhar, Sumuk, et al.
Veröffentlicht: (2024)
Infogent: An Agent-Based Framework for Web Information Aggregation
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
von: Reddy, Revanth Gangi, et al.
Veröffentlicht: (2024)
SpeakRL: Synergizing Reasoning, Speaking, and Acting in Language Models with Reinforcement Learning
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
Premise-Augmented Reasoning Chains Improve Error Identification in Math reasoning with LLMs
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2025)
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2025)
On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization
von: Singh, Janvijay, et al.
Veröffentlicht: (2025)
von: Singh, Janvijay, et al.
Veröffentlicht: (2025)
Drift No More? Context Equilibria in Multi-Turn LLM Interactions
von: Dongre, Vardhan, et al.
Veröffentlicht: (2025)
von: Dongre, Vardhan, et al.
Veröffentlicht: (2025)
Aligning LLMs with Individual Preferences via Interaction
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
From Context to Action: Analysis of the Impact of State Representation and Context on the Generalization of Multi-Turn Web Navigation Agents
von: Tiwary, Nalin, et al.
Veröffentlicht: (2024)
von: Tiwary, Nalin, et al.
Veröffentlicht: (2024)
When Attention Closes: How LLMs Lose the Thread in Multi-Turn Interaction
von: Dongre, Vardhan, et al.
Veröffentlicht: (2026)
von: Dongre, Vardhan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models
von: Bozdag, Nimet Beyza, et al.
Veröffentlicht: (2025) -
Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration
von: Kargupta, Priyanka, et al.
Veröffentlicht: (2026) -
User Preference Modeling for Conversational LLM Agents: Weak Rewards from Retrieval-Augmented Interaction
von: Hao, Yuren, et al.
Veröffentlicht: (2026) -
AURA: A Diagnostic Framework for Tracking User Satisfaction of Interactive Planning Agents
von: Kim, Takyoung, et al.
Veröffentlicht: (2025) -
Simulating User Agents for Embodied Conversational-AI
von: Philipov, Daniel, et al.
Veröffentlicht: (2024)