Guardado en:
| Autores principales: | Pachtrachai, Krittin, Pornpichitsuwan, Petmongkon, Modecrua, Wachiravit, Kraisingkorn, Touchapon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.15859 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Cloning a Conversational Voice AI Agent from Call\,Recording Datasets for Telesales
por: Kaewtawee, Krittanon, et al.
Publicado: (2025)
por: Kaewtawee, Krittanon, et al.
Publicado: (2025)
Multi-Turn Reinforcement Learning for Tool-Calling Agents with Iterative Reward Calibration
por: Modecrua, Wachiravit, et al.
Publicado: (2026)
por: Modecrua, Wachiravit, et al.
Publicado: (2026)
ROAD: Reflective Optimization via Automated Debugging for Zero-Shot Agent Alignment
por: Temyingyong, Natchaya, et al.
Publicado: (2025)
por: Temyingyong, Natchaya, et al.
Publicado: (2025)
Evaluating the Performance of RAG Methods for Conversational AI in the Airport Domain
por: Li, Yuyang, et al.
Publicado: (2025)
por: Li, Yuyang, et al.
Publicado: (2025)
Detecting Ambiguities to Guide Query Rewrite for Robust Conversations in Enterprise AI Assistants
por: Tanjim, Md Mehrab, et al.
Publicado: (2025)
por: Tanjim, Md Mehrab, et al.
Publicado: (2025)
From AI Assistant to AI Scientist: Autonomous Discovery of LLM-RL Algorithms with LLM Agents
por: Xia, Sirui, et al.
Publicado: (2026)
por: Xia, Sirui, et al.
Publicado: (2026)
PersonaLens: A Benchmark for Personalization Evaluation in Conversational AI Assistants
por: Zhao, Zheng, et al.
Publicado: (2025)
por: Zhao, Zheng, et al.
Publicado: (2025)
AI Knowledge Assist: An Automated Approach for the Creation of Knowledge Bases for Conversational AI Agents
por: Laskar, Md Tahmid Rahman, et al.
Publicado: (2025)
por: Laskar, Md Tahmid Rahman, et al.
Publicado: (2025)
LLM Agent Meets Agentic AI: Can LLM Agents Simulate Customers to Evaluate Agentic-AI-based Shopping Assistants?
por: Sun, Lu, et al.
Publicado: (2025)
por: Sun, Lu, et al.
Publicado: (2025)
Mentalic Net: Development of RAG-based Conversational AI and Evaluation Framework for Mental Health Support
por: Dutta, Anandi, et al.
Publicado: (2025)
por: Dutta, Anandi, et al.
Publicado: (2025)
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
por: Shin, Hagyeong, et al.
Publicado: (2025)
por: Shin, Hagyeong, et al.
Publicado: (2025)
ORAssistant: A Custom RAG-based Conversational Assistant for OpenROAD
por: Kaintura, Aviral, et al.
Publicado: (2024)
por: Kaintura, Aviral, et al.
Publicado: (2024)
Information Extraction from Conversation Transcripts: Neuro-Symbolic vs. LLM
por: Kwak, Alice Saebom, et al.
Publicado: (2025)
por: Kwak, Alice Saebom, et al.
Publicado: (2025)
Evaluation and Incident Prevention in an Enterprise AI Assistant
por: Maharaj, Akash V., et al.
Publicado: (2025)
por: Maharaj, Akash V., et al.
Publicado: (2025)
TravelAgent: An AI Assistant for Personalized Travel Planning
por: Chen, Aili, et al.
Publicado: (2024)
por: Chen, Aili, et al.
Publicado: (2024)
Measuring Data Science Automation: A Survey of Evaluation Tools for AI Assistants and Agents
por: Testini, Irene, et al.
Publicado: (2025)
por: Testini, Irene, et al.
Publicado: (2025)
Affordable AI Assistants with Knowledge Graph of Thoughts
por: Besta, Maciej, et al.
Publicado: (2025)
por: Besta, Maciej, et al.
Publicado: (2025)
Agent-Testing Agent: A Meta-Agent for Automated Testing and Evaluation of Conversational AI Agents
por: Komoravolu, Sameer, et al.
Publicado: (2025)
por: Komoravolu, Sameer, et al.
Publicado: (2025)
NyayaAI: An AI-Powered Legal Assistant Using Multi-Agent Architecture and Retrieval-Augmented Generation
por: Deepanshu, et al.
Publicado: (2026)
por: Deepanshu, et al.
Publicado: (2026)
A RAG-Based Institutional Assistant
por: Kuratomi, Gustavo, et al.
Publicado: (2025)
por: Kuratomi, Gustavo, et al.
Publicado: (2025)
AI-Powered Assistant for Long-Term Access to RHIC Knowledge
por: Atif, Mohammad, et al.
Publicado: (2025)
por: Atif, Mohammad, et al.
Publicado: (2025)
IntellAgent: A Multi-Agent Framework for Evaluating Conversational AI Systems
por: Levi, Elad, et al.
Publicado: (2025)
por: Levi, Elad, et al.
Publicado: (2025)
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
por: Gritta, Milan, et al.
Publicado: (2024)
por: Gritta, Milan, et al.
Publicado: (2024)
Developing an AI Assistant for Knowledge Management and Workforce Training in State DOTs
por: Amaram, Divija, et al.
Publicado: (2026)
por: Amaram, Divija, et al.
Publicado: (2026)
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process
por: Kim, Jaewoong, et al.
Publicado: (2024)
por: Kim, Jaewoong, et al.
Publicado: (2024)
Turning Conversations into Workflows: A Framework to Extract and Evaluate Dialog Workflows for Service AI Agents
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
por: Choubey, Prafulla Kumar, et al.
Publicado: (2025)
SALAD: Smart AI Language Assistant Daily
por: Nihal, Ragib Amin, et al.
Publicado: (2024)
por: Nihal, Ragib Amin, et al.
Publicado: (2024)
Chatty-KG: A Multi-Agent AI System for On-Demand Conversational Question Answering over Knowledge Graphs
por: Omar, Reham, et al.
Publicado: (2025)
por: Omar, Reham, et al.
Publicado: (2025)
$τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge
por: Shi, Quan, et al.
Publicado: (2026)
por: Shi, Quan, et al.
Publicado: (2026)
TREC iKAT 2023: A Test Collection for Evaluating Conversational and Interactive Knowledge Assistants
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
por: Aliannejadi, Mohammad, et al.
Publicado: (2024)
MITRA: An AI Assistant for Knowledge Retrieval in Physics Collaborations
por: Mallampalli, Abhishikth, et al.
Publicado: (2026)
por: Mallampalli, Abhishikth, et al.
Publicado: (2026)
Designing LMS and Instructional Strategies for Integrating Generative-Conversational AI
por: Ra, Elias, et al.
Publicado: (2025)
por: Ra, Elias, et al.
Publicado: (2025)
SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?
por: Dou, Yao, et al.
Publicado: (2025)
por: Dou, Yao, et al.
Publicado: (2025)
Simulating User Agents for Embodied Conversational-AI
por: Philipov, Daniel, et al.
Publicado: (2024)
por: Philipov, Daniel, et al.
Publicado: (2024)
HybridRAG: Integrating Knowledge Graphs and Vector Retrieval Augmented Generation for Efficient Information Extraction
por: Sarmah, Bhaskarjit, et al.
Publicado: (2024)
por: Sarmah, Bhaskarjit, et al.
Publicado: (2024)
Towards More Standardized AI Evaluation: From Models to Agents
por: Filali, Ali El, et al.
Publicado: (2026)
por: Filali, Ali El, et al.
Publicado: (2026)
AI Agents for Conversational Patient Triage: Preliminary Simulation-Based Evaluation with Real-World EHR Data
por: Rashidian, Sina, et al.
Publicado: (2025)
por: Rashidian, Sina, et al.
Publicado: (2025)
User Modeling Challenges in Interactive AI Assistant Systems
por: Su, Megan, et al.
Publicado: (2024)
por: Su, Megan, et al.
Publicado: (2024)
Fusion-Eval: Integrating Assistant Evaluators with LLMs
por: Shu, Lei, et al.
Publicado: (2023)
por: Shu, Lei, et al.
Publicado: (2023)
Preserving Knowledge Invariance: Rethinking Robustness Evaluation of Open Information Extraction
por: Qi, Ji, et al.
Publicado: (2023)
por: Qi, Ji, et al.
Publicado: (2023)
Ejemplares similares
-
Cloning a Conversational Voice AI Agent from Call\,Recording Datasets for Telesales
por: Kaewtawee, Krittanon, et al.
Publicado: (2025) -
Multi-Turn Reinforcement Learning for Tool-Calling Agents with Iterative Reward Calibration
por: Modecrua, Wachiravit, et al.
Publicado: (2026) -
ROAD: Reflective Optimization via Automated Debugging for Zero-Shot Agent Alignment
por: Temyingyong, Natchaya, et al.
Publicado: (2025) -
Evaluating the Performance of RAG Methods for Conversational AI in the Airport Domain
por: Li, Yuyang, et al.
Publicado: (2025) -
Detecting Ambiguities to Guide Query Rewrite for Robust Conversations in Enterprise AI Assistants
por: Tanjim, Md Mehrab, et al.
Publicado: (2025)