$τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge
Fuente:
arXiv
Salvato in:
| Autori principali: | Shi, Quan, Zytek, Alexandra, Razavi, Pedram, Narasimhan, Karthik, Barres, Victor |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
$τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
di: Barres, Victor, et al.
Pubblicazione: (2025)
di: Barres, Victor, et al.
Pubblicazione: (2025)
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
di: Yao, Shunyu, et al.
Pubblicazione: (2024)
di: Yao, Shunyu, et al.
Pubblicazione: (2024)
Reasoning over User Preferences: Knowledge Graph-Augmented LLMs for Explainable Conversational Recommendations
di: Qiu, Zhangchi, et al.
Pubblicazione: (2024)
di: Qiu, Zhangchi, et al.
Pubblicazione: (2024)
TREC iKAT 2023: A Test Collection for Evaluating Conversational and Interactive Knowledge Assistants
di: Aliannejadi, Mohammad, et al.
Pubblicazione: (2024)
di: Aliannejadi, Mohammad, et al.
Pubblicazione: (2024)
Deterministic Legal Agents: A Canonical Primitive API for Auditable Reasoning over Temporal Knowledge Graphs
di: de Martim, Hudson
Pubblicazione: (2025)
di: de Martim, Hudson
Pubblicazione: (2025)
Evaluating the External and Parametric Knowledge Fusion of Large Language Models
di: Zhang, Hao, et al.
Pubblicazione: (2024)
di: Zhang, Hao, et al.
Pubblicazione: (2024)
Enhancing Supply Chain Visibility with Knowledge Graphs and Large Language Models
di: AlMahri, Sara, et al.
Pubblicazione: (2024)
di: AlMahri, Sara, et al.
Pubblicazione: (2024)
Zep: A Temporal Knowledge Graph Architecture for Agent Memory
di: Rasmussen, Preston, et al.
Pubblicazione: (2025)
di: Rasmussen, Preston, et al.
Pubblicazione: (2025)
Knowledge Graphs and Pre-trained Language Models enhanced Representation Learning for Conversational Recommender Systems
di: Qiu, Zhangchi, et al.
Pubblicazione: (2023)
di: Qiu, Zhangchi, et al.
Pubblicazione: (2023)
Bias-Aware Agent: Enhancing Fairness in AI-Driven Knowledge Retrieval
di: Singh, Karanbir, et al.
Pubblicazione: (2025)
di: Singh, Karanbir, et al.
Pubblicazione: (2025)
Joint Knowledge Editing for Information Enrichment and Probability Promotion
di: Shi, Wenhang, et al.
Pubblicazione: (2024)
di: Shi, Wenhang, et al.
Pubblicazione: (2024)
FlyAOC: Evaluating Agentic Ontology Curation of Drosophila Scientific Knowledge Bases
di: Zhang, Xingjian, et al.
Pubblicazione: (2026)
di: Zhang, Xingjian, et al.
Pubblicazione: (2026)
HybridRAG: A Practical LLM-based ChatBot Framework based on Pre-Generated Q&A over Raw Unstructured Documents
di: Kim, Sungmoon, et al.
Pubblicazione: (2025)
di: Kim, Sungmoon, et al.
Pubblicazione: (2025)
A Bi-Encoder LSTM Model For Learning Unstructured Dialogs
di: Brahman, Danny, et al.
Pubblicazione: (2021)
di: Brahman, Danny, et al.
Pubblicazione: (2021)
Tug-of-War Between Knowledge: Exploring and Resolving Knowledge Conflicts in Retrieval-Augmented Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
KG-LLM-Bench: A Scalable Benchmark for Evaluating LLM Reasoning on Textualized Knowledge Graphs
di: Markowitz, Elan, et al.
Pubblicazione: (2025)
di: Markowitz, Elan, et al.
Pubblicazione: (2025)
Comparative Analysis of Neural Retriever-Reranker Pipelines for Retrieval-Augmented Generation over Knowledge Graphs in E-commerce Applications
di: Rumble, Teri, et al.
Pubblicazione: (2025)
di: Rumble, Teri, et al.
Pubblicazione: (2025)
Reasoning on Efficient Knowledge Paths:Knowledge Graph Guides Large Language Model for Domain Question Answering
di: Wang, Yuqi, et al.
Pubblicazione: (2024)
di: Wang, Yuqi, et al.
Pubblicazione: (2024)
ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions
di: Ye, Jingheng, et al.
Pubblicazione: (2024)
di: Ye, Jingheng, et al.
Pubblicazione: (2024)
Evaluating improvements on using Large Language Models (LLMs) for property extraction in the Open Research Knowledge Graph (ORKG)
di: Schaftner, Sandra
Pubblicazione: (2025)
di: Schaftner, Sandra
Pubblicazione: (2025)
Enhancing LLM Medical Coding with Structured External Knowledge
di: Gan, Yidong, et al.
Pubblicazione: (2026)
di: Gan, Yidong, et al.
Pubblicazione: (2026)
TRAWL: External Knowledge-Enhanced Recommendation with LLM Assistance
di: Luo, Weiqing, et al.
Pubblicazione: (2024)
di: Luo, Weiqing, et al.
Pubblicazione: (2024)
Multilingual Information Retrieval with a Monolingual Knowledge Base
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
di: Zhuang, Yingying, et al.
Pubblicazione: (2025)
KG-RAG: Bridging the Gap Between Knowledge and Creativity
di: Sanmartin, Diego
Pubblicazione: (2024)
di: Sanmartin, Diego
Pubblicazione: (2024)
Analyzing the Influence of Knowledge Graph Information on Relation Extraction
di: Möller, Cedric, et al.
Pubblicazione: (2025)
di: Möller, Cedric, et al.
Pubblicazione: (2025)
KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking
di: Kim, Juyeon, et al.
Pubblicazione: (2025)
di: Kim, Juyeon, et al.
Pubblicazione: (2025)
Building FKG.in: a Knowledge Graph for Indian Food
di: Gupta, Saransh Kumar, et al.
Pubblicazione: (2024)
di: Gupta, Saransh Kumar, et al.
Pubblicazione: (2024)
Distilling Reasoning Without Knowledge: A Framework for Reliable LLMs
di: Kietkajornrit, Auksarapak, et al.
Pubblicazione: (2026)
di: Kietkajornrit, Auksarapak, et al.
Pubblicazione: (2026)
Enhancing LLM Generation with Knowledge Hypergraph for Evidence-Based Medicine
di: Dou, Chengfeng, et al.
Pubblicazione: (2025)
di: Dou, Chengfeng, et al.
Pubblicazione: (2025)
Knowledge Management for Automobile Failure Analysis Using Graph RAG
di: Ojima, Yuta, et al.
Pubblicazione: (2024)
di: Ojima, Yuta, et al.
Pubblicazione: (2024)
Comparing Knowledge Sources for Open-Domain Scientific Claim Verification
di: Vladika, Juraj, et al.
Pubblicazione: (2024)
di: Vladika, Juraj, et al.
Pubblicazione: (2024)
AI-Powered Assistant for Long-Term Access to RHIC Knowledge
di: Atif, Mohammad, et al.
Pubblicazione: (2025)
di: Atif, Mohammad, et al.
Pubblicazione: (2025)
Resisting Contextual Interference in RAG via Parametric-Knowledge Reinforcement
di: Lin, Chenyu, et al.
Pubblicazione: (2025)
di: Lin, Chenyu, et al.
Pubblicazione: (2025)
RAG-KG-IL: A Multi-Agent Hybrid Framework for Reducing Hallucinations and Enhancing LLM Reasoning through RAG and Incremental Knowledge Graph Learning Integration
di: Yu, Hong Qing, et al.
Pubblicazione: (2025)
di: Yu, Hong Qing, et al.
Pubblicazione: (2025)
Evaluating Large Language Models as Generative User Simulators for Conversational Recommendation
di: Yoon, Se-eun, et al.
Pubblicazione: (2024)
di: Yoon, Se-eun, et al.
Pubblicazione: (2024)
DynaRAG: Bridging Static and Dynamic Knowledge in Retrieval-Augmented Generation
di: Liang, Penghao, et al.
Pubblicazione: (2026)
di: Liang, Penghao, et al.
Pubblicazione: (2026)
Semi-Automated Knowledge Engineering and Process Mapping for Total Airport Management
di: Teo, Darryl, et al.
Pubblicazione: (2026)
di: Teo, Darryl, et al.
Pubblicazione: (2026)
Developing an AI Assistant for Knowledge Management and Workforce Training in State DOTs
di: Amaram, Divija, et al.
Pubblicazione: (2026)
di: Amaram, Divija, et al.
Pubblicazione: (2026)
Training the Knowledge Base through Evidence Distillation and Write-Back Enrichment
di: Lu, Yuxing, et al.
Pubblicazione: (2026)
di: Lu, Yuxing, et al.
Pubblicazione: (2026)
CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG
di: Tian, Yang, et al.
Pubblicazione: (2025)
di: Tian, Yang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
$τ^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
di: Barres, Victor, et al.
Pubblicazione: (2025) -
$τ$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
di: Yao, Shunyu, et al.
Pubblicazione: (2024) -
Reasoning over User Preferences: Knowledge Graph-Augmented LLMs for Explainable Conversational Recommendations
di: Qiu, Zhangchi, et al.
Pubblicazione: (2024) -
TREC iKAT 2023: A Test Collection for Evaluating Conversational and Interactive Knowledge Assistants
di: Aliannejadi, Mohammad, et al.
Pubblicazione: (2024) -
Deterministic Legal Agents: A Canonical Primitive API for Auditable Reasoning over Temporal Knowledge Graphs
di: de Martim, Hudson
Pubblicazione: (2025)