Mind the Goal: Data-Efficient Goal-Oriented Evaluation of Conversational Agents and Chatbots using Teacher Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Piskala, Deepak Babu, Chen, Sharlene, Patel, Udita, Kalra, Parul, Castrillo, Rafael |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
par: Piskala, Deepak Babu
Publié: (2026)
par: Piskala, Deepak Babu
Publié: (2026)
THELMA: Task Based Holistic Evaluation of Large Language Model Applications-RAG Question Answering
par: Patel, Udita, et autres
Publié: (2025)
par: Patel, Udita, et autres
Publié: (2025)
PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
par: Piskala, Deepak Babu
Publié: (2025)
par: Piskala, Deepak Babu
Publié: (2025)
Beyond Ontology in Dialogue State Tracking for Goal-Oriented Chatbot
par: Lee, Sejin, et autres
Publié: (2024)
par: Lee, Sejin, et autres
Publié: (2024)
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
par: Piskala, Deepak Babu
Publié: (2026)
par: Piskala, Deepak Babu
Publié: (2026)
Towards Goal-Oriented Agents for Evolving Problems Observed via Conversation
par: Free, Michael, et autres
Publié: (2024)
par: Free, Michael, et autres
Publié: (2024)
Contextual ASR Error Handling with LLMs Augmentation for Goal-Oriented Conversational AI
par: Asano, Yuya, et autres
Publié: (2025)
par: Asano, Yuya, et autres
Publié: (2025)
Plan-Grounded Large Language Models for Dual Goal Conversational Settings
par: Glória-Silva, Diogo, et autres
Publié: (2024)
par: Glória-Silva, Diogo, et autres
Publié: (2024)
Uncertainty as a Planning Signal: Multi-Turn Decision Making for Goal-Oriented Conversation
par: Ling, Xinyi, et autres
Publié: (2026)
par: Ling, Xinyi, et autres
Publié: (2026)
Feedback-Aware Monte Carlo Tree Search for Efficient Information Seeking in Goal-Oriented Conversations
par: Chopra, Harshita, et autres
Publié: (2025)
par: Chopra, Harshita, et autres
Publié: (2025)
SelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
par: Yang, Ruihan, et autres
Publié: (2024)
par: Yang, Ruihan, et autres
Publié: (2024)
A Goal-Oriented Chatbot for Engaging the Elderly Through Family Photo Conversations
par: Chung, Raymond, et autres
Publié: (2026)
par: Chung, Raymond, et autres
Publié: (2026)
The AI Roles Continuum: Blurring the Boundary Between Research and Engineering
par: Piskala, Deepak Babu
Publié: (2025)
par: Piskala, Deepak Babu
Publié: (2025)
From Everything-is-a-File to Files-Are-All-You-Need: How Unix Philosophy Informs the Design of Agentic AI Systems
par: Piskala, Deepak Babu
Publié: (2026)
par: Piskala, Deepak Babu
Publié: (2026)
Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks
par: Rahman, Hasibur, et autres
Publié: (2025)
par: Rahman, Hasibur, et autres
Publié: (2025)
A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents
par: Arghal, Raghu, et autres
Publié: (2026)
par: Arghal, Raghu, et autres
Publié: (2026)
Evaluating the Goal-Directedness of Large Language Models
par: Everitt, Tom, et autres
Publié: (2025)
par: Everitt, Tom, et autres
Publié: (2025)
Goal Alignment in LLM-Based User Simulators for Conversational AI
par: Mehri, Shuhaib, et autres
Publié: (2025)
par: Mehri, Shuhaib, et autres
Publié: (2025)
HierTOD: A Task-Oriented Dialogue System Driven by Hierarchical Goals
par: Mo, Lingbo, et autres
Publié: (2024)
par: Mo, Lingbo, et autres
Publié: (2024)
Decoupling Strategy and Execution in Task-Focused Dialogue via Goal-Oriented Preference Optimization
par: Xu, Jingyi, et autres
Publié: (2026)
par: Xu, Jingyi, et autres
Publié: (2026)
A Survey of Personality, Persona, and Profile in Conversational Agents and Chatbots
par: Sutcliffe, Richard
Publié: (2023)
par: Sutcliffe, Richard
Publié: (2023)
Dynamic LLM Routing and Selection based on User Preferences: Balancing Performance, Cost, and Ethics
par: Piskala, Deepak Babu, et autres
Publié: (2025)
par: Piskala, Deepak Babu, et autres
Publié: (2025)
Asymmetric Goal Drift in Coding Agents Under Value Conflict
par: Saebo, Magnus, et autres
Publié: (2026)
par: Saebo, Magnus, et autres
Publié: (2026)
Evaluating the Paperclip Maximizer: Are RL-Based Language Models More Likely to Pursue Instrumental Goals?
par: He, Yufei, et autres
Publié: (2025)
par: He, Yufei, et autres
Publié: (2025)
Controllable and Reliable Knowledge-Intensive Task-Oriented Conversational Agents with Declarative Genie Worksheets
par: Joshi, Harshit, et autres
Publié: (2024)
par: Joshi, Harshit, et autres
Publié: (2024)
Goal-Directed Search Outperforms Goal-Agnostic Memory Compression in Long-Context Memory Tasks
par: Zheng, Yicong, et autres
Publié: (2025)
par: Zheng, Yicong, et autres
Publié: (2025)
Identifying User Goals from UI Trajectories
par: Berkovitch, Omri, et autres
Publié: (2024)
par: Berkovitch, Omri, et autres
Publié: (2024)
Flexible Agent Alignment with Goal Inference from Open-Ended Dialog
par: Ma, Rachel, et autres
Publié: (2025)
par: Ma, Rachel, et autres
Publié: (2025)
Personalized Learning Path Planning with Goal-Driven Learner State Modeling
par: Lim, Joy Jia Yin, et autres
Publié: (2025)
par: Lim, Joy Jia Yin, et autres
Publié: (2025)
Substance over Style: Evaluating Proactive Conversational Coaching Agents
par: Srinivas, Vidya, et autres
Publié: (2025)
par: Srinivas, Vidya, et autres
Publié: (2025)
Adaptive GoGI-Skip: Coupling Goal-Gradient Importance with Dynamic Uncertainty for Efficient Reasoning
par: Zhuang, Ren
Publié: (2025)
par: Zhuang, Ren
Publié: (2025)
Can Large Language Model Summarizers Adapt to Diverse Scientific Communication Goals?
par: Fonseca, Marcio, et autres
Publié: (2024)
par: Fonseca, Marcio, et autres
Publié: (2024)
Towards Goal-oriented Prompt Engineering for Large Language Models: A Survey
par: Li, Haochen, et autres
Publié: (2024)
par: Li, Haochen, et autres
Publié: (2024)
Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems
par: Panigrahy, Deepak, et autres
Publié: (2026)
par: Panigrahy, Deepak, et autres
Publié: (2026)
Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies and Their Limitations
par: Arnaudo, Anna, et autres
Publié: (2026)
par: Arnaudo, Anna, et autres
Publié: (2026)
GOAT: A Training Framework for Goal-Oriented Agent with Tools
par: Min, Hyunji, et autres
Publié: (2025)
par: Min, Hyunji, et autres
Publié: (2025)
GoalNet: Goal Areas Oriented Pedestrian Trajectory Prediction
par: Fadillah, Amar, et autres
Publié: (2024)
par: Fadillah, Amar, et autres
Publié: (2024)
Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems
par: Liang, Jiazhou, et autres
Publié: (2026)
par: Liang, Jiazhou, et autres
Publié: (2026)
Goal Hijacking Attack on Large Language Models via Pseudo-Conversation Injection
par: Chen, Zheng, et autres
Publié: (2024)
par: Chen, Zheng, et autres
Publié: (2024)
ReflAct: World-Grounded Decision Making in LLM Agents via Goal-State Reflection
par: Kim, Jeonghye, et autres
Publié: (2025)
par: Kim, Jeonghye, et autres
Publié: (2025)
Documents similaires
-
MAPLE: A Sub-Agent Architecture for Memory, Learning, and Personalization in Agentic AI Systems
par: Piskala, Deepak Babu
Publié: (2026) -
THELMA: Task Based Holistic Evaluation of Large Language Model Applications-RAG Question Answering
par: Patel, Udita, et autres
Publié: (2025) -
PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
par: Piskala, Deepak Babu
Publié: (2025) -
Beyond Ontology in Dialogue State Tracking for Goal-Oriented Chatbot
par: Lee, Sejin, et autres
Publié: (2024) -
Spec-Driven Development:From Code to Contract in the Age of AI Coding Assistants
par: Piskala, Deepak Babu
Publié: (2026)