iAgentBench: Benchmarking Sensemaking Capabilities of Information-Seeking Agents on High-Traffic Topics
Fuente:
arXiv
Salvato in:
| Autori principali: | Dammu, Preetam Prabhu Srikar, Palkhiwala, Arnav, Roosta, Tanya, Shah, Chirag |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Static Evaluation: Rethinking the Assessment of Personalized Agent Adaptability in Information Retrieval
di: Kaur, Kirandeep, et al.
Pubblicazione: (2025)
di: Kaur, Kirandeep, et al.
Pubblicazione: (2025)
Dynamic Evaluation Framework for Personalized and Trustworthy Agents: A Multi-Session Approach to Preference Adaptability
di: Shah, Chirag, et al.
Pubblicazione: (2025)
di: Shah, Chirag, et al.
Pubblicazione: (2025)
Dynamic-KGQA: A Scalable Framework for Generating Adaptive Question Answering Datasets
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2025)
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2025)
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
di: Wu, Bin, et al.
Pubblicazione: (2026)
di: Wu, Bin, et al.
Pubblicazione: (2026)
Behind the Prompt: The Agent-User Problem in Information Retrieval
di: Zerhoudi, Saber, et al.
Pubblicazione: (2026)
di: Zerhoudi, Saber, et al.
Pubblicazione: (2026)
Agentic SPARQL: Evaluating SPARQL-MCP-powered Intelligent Agents on the Federated KGQA Benchmark
di: Dobriy, Daniel, et al.
Pubblicazione: (2026)
di: Dobriy, Daniel, et al.
Pubblicazione: (2026)
LLMGreenRec: LLM-Based Multi-Agent Recommender System for Sustainable E-Commerce
di: Nguyen, Hao N., et al.
Pubblicazione: (2026)
di: Nguyen, Hao N., et al.
Pubblicazione: (2026)
A Learnable Agent Collaboration Network Framework for Personalized Multimodal AI Search Engine
di: Shi, Yunxiao, et al.
Pubblicazione: (2024)
di: Shi, Yunxiao, et al.
Pubblicazione: (2024)
TWICE: An LLM Agent Framework for Simulating Personalized User Tweeting Behavior with Long-term Temporal Features
di: Jin, Bingrui, et al.
Pubblicazione: (2025)
di: Jin, Bingrui, et al.
Pubblicazione: (2025)
GAAMA: Graph Augmented Associative Memory for Agents
di: Paul, Swarna Kamal, et al.
Pubblicazione: (2026)
di: Paul, Swarna Kamal, et al.
Pubblicazione: (2026)
Knowledge Graph Enhanced Language Agents for Recommendation
di: Guo, Taicheng, et al.
Pubblicazione: (2024)
di: Guo, Taicheng, et al.
Pubblicazione: (2024)
SynthAgent: A Multi-Agent LLM Framework for Realistic Patient Simulation -- A Case Study in Obesity with Mental Health Comorbidities
di: Aghaee, Arman, et al.
Pubblicazione: (2026)
di: Aghaee, Arman, et al.
Pubblicazione: (2026)
ClinQueryAgent: A Conversational Agent for Population Health Management
di: Boyle, Joseph S., et al.
Pubblicazione: (2026)
di: Boyle, Joseph S., et al.
Pubblicazione: (2026)
Evaluating Supply Chain Resilience During Pandemic Using Agent-based Simulation
di: Lazebnik, Teddy
Pubblicazione: (2024)
di: Lazebnik, Teddy
Pubblicazione: (2024)
Nested Browser-Use Learning for Agentic Information Seeking
di: Li, Baixuan, et al.
Pubblicazione: (2025)
di: Li, Baixuan, et al.
Pubblicazione: (2025)
Multi-Agent Video Recommenders: Evolution, Patterns, and Open Challenges
di: Ranganathan, Srivaths, et al.
Pubblicazione: (2026)
di: Ranganathan, Srivaths, et al.
Pubblicazione: (2026)
Towards Efficient Hypergraph and Multi-LLM Agent Recommender Systems
di: Mukande, Tendai, et al.
Pubblicazione: (2025)
di: Mukande, Tendai, et al.
Pubblicazione: (2025)
AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents
di: Jin, Jiarui, et al.
Pubblicazione: (2026)
di: Jin, Jiarui, et al.
Pubblicazione: (2026)
From Fluent to Verifiable: Claim-Level Auditability for Deep Research Agents
di: Rasheed, Razeen A, et al.
Pubblicazione: (2026)
di: Rasheed, Razeen A, et al.
Pubblicazione: (2026)
A Multi-Agent Orchestration Framework for Venture Capital Due Diligence
di: Alexandrou, Grigorios, et al.
Pubblicazione: (2026)
di: Alexandrou, Grigorios, et al.
Pubblicazione: (2026)
A Hierarchical Multi-Agent System for Autonomous Discovery in Geoscientific Data Archives
di: Pantiukhin, Dmitrii, et al.
Pubblicazione: (2026)
di: Pantiukhin, Dmitrii, et al.
Pubblicazione: (2026)
Divide by Question, Conquer by Agent: SPLIT-RAG with Question-Driven Graph Partitioning
di: Yang, Ruiyi, et al.
Pubblicazione: (2025)
di: Yang, Ruiyi, et al.
Pubblicazione: (2025)
LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence
di: Vinogradov, Vlad, et al.
Pubblicazione: (2025)
di: Vinogradov, Vlad, et al.
Pubblicazione: (2025)
Parallelism Meets Adaptiveness: Scalable Documents Understanding in Multi-Agent LLM Systems
di: Xia, Chengxuan, et al.
Pubblicazione: (2025)
di: Xia, Chengxuan, et al.
Pubblicazione: (2025)
Agent4POI: Agentic Context-Conditioned Affordance Reasoning for Multimodal Point-of-Interest Recommendation
di: Wang, Jinze, et al.
Pubblicazione: (2026)
di: Wang, Jinze, et al.
Pubblicazione: (2026)
MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation
di: Wu, Chuanjie, et al.
Pubblicazione: (2026)
di: Wu, Chuanjie, et al.
Pubblicazione: (2026)
DrunkAgent: Stealthy Memory Corruption in LLM-Powered Recommender Agents
di: Yang, Shiyi, et al.
Pubblicazione: (2025)
di: Yang, Shiyi, et al.
Pubblicazione: (2025)
NutriOrion: A Hierarchical Multi-Agent Framework for Personalized Nutrition Intervention Grounded in Clinical Guidelines
di: Wu, Junwei, et al.
Pubblicazione: (2026)
di: Wu, Junwei, et al.
Pubblicazione: (2026)
PlotEdit: Natural Language-Driven Accessible Chart Editing in PDFs via Multimodal LLM Agents
di: Goswami, Kanika, et al.
Pubblicazione: (2025)
di: Goswami, Kanika, et al.
Pubblicazione: (2025)
Physics-Grounded Multi-Agent Architecture for Traceable, Risk-Aware Human-AI Decision Support in Manufacturing
di: Hoang, Danny, et al.
Pubblicazione: (2026)
di: Hoang, Danny, et al.
Pubblicazione: (2026)
ClaimDB: A Fact Verification Benchmark over Large Structured Data
di: Theologitis, Michael, et al.
Pubblicazione: (2026)
di: Theologitis, Michael, et al.
Pubblicazione: (2026)
Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation
di: Lazebnik, Teddy, et al.
Pubblicazione: (2025)
di: Lazebnik, Teddy, et al.
Pubblicazione: (2025)
FitText: Evolving Agent Tool Ecologies via Memetic Retrieval
di: Zheng, Kyle, et al.
Pubblicazione: (2026)
di: Zheng, Kyle, et al.
Pubblicazione: (2026)
paper.json: A Coordination Convention for LLM-Agent-Actionable Papers
di: Canedo, Arquimedes
Pubblicazione: (2026)
di: Canedo, Arquimedes
Pubblicazione: (2026)
Personalized Recommendation Systems using Multimodal, Autonomous, Multi Agent Systems
di: Thakkar, Param, et al.
Pubblicazione: (2024)
di: Thakkar, Param, et al.
Pubblicazione: (2024)
Osprey: Production-Ready Agentic AI for Safety-Critical Control Systems
di: Hellert, Thorsten, et al.
Pubblicazione: (2025)
di: Hellert, Thorsten, et al.
Pubblicazione: (2025)
CogPlanner: Unveiling the Potential of Agentic Multimodal Retrieval Augmented Generation with Planning
di: Yu, Xiaohan, et al.
Pubblicazione: (2025)
di: Yu, Xiaohan, et al.
Pubblicazione: (2025)
Ultra Low-Cost Two-Stage Multimodal System for Non-Normative Behavior Detection
di: Lu, Albert, et al.
Pubblicazione: (2024)
di: Lu, Albert, et al.
Pubblicazione: (2024)
Caesar: Deep Agentic Web Exploration for Creative Answer Synthesis
di: Liang, Jason, et al.
Pubblicazione: (2026)
di: Liang, Jason, et al.
Pubblicazione: (2026)
From Generation to Attribution: Music AI Agent Architectures for the Post-Streaming Era
di: Kim, Wonil, et al.
Pubblicazione: (2025)
di: Kim, Wonil, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Beyond Static Evaluation: Rethinking the Assessment of Personalized Agent Adaptability in Information Retrieval
di: Kaur, Kirandeep, et al.
Pubblicazione: (2025) -
Dynamic Evaluation Framework for Personalized and Trustworthy Agents: A Multi-Session Approach to Preference Adaptability
di: Shah, Chirag, et al.
Pubblicazione: (2025) -
Dynamic-KGQA: A Scalable Framework for Generating Adaptive Question Answering Datasets
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2025) -
AgentSearchBench: A Benchmark for AI Agent Search in the Wild
di: Wu, Bin, et al.
Pubblicazione: (2026) -
Behind the Prompt: The Agent-User Problem in Information Retrieval
di: Zerhoudi, Saber, et al.
Pubblicazione: (2026)