HumanMCP: A Human-Like Query Dataset for Evaluating MCP Tool Retrieval Performance
Fuente:
arXiv
Guardado en:
| Autores principales: | Laddha, Shubh, Changbencharoen, Lucas, Kuptivej, Win, Shringla, Surya, Vaidheeswaran, Archana, Bhaskar, Yash |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Federated k-Means over Networks
por: Yang, Xu, et al.
Publicado: (2025)
por: Yang, Xu, et al.
Publicado: (2025)
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
por: Berman, Shmuel, et al.
Publicado: (2024)
por: Berman, Shmuel, et al.
Publicado: (2024)
Emergent Collective Memory in Decentralized Multi-Agent AI Systems
por: Khushiyant
Publicado: (2025)
por: Khushiyant
Publicado: (2025)
Adaptive Multi-Stage Patent Claim Generation with Unified Quality Assessment
por: Liang, Chen-Wei, et al.
Publicado: (2026)
por: Liang, Chen-Wei, et al.
Publicado: (2026)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026)
por: Jia, Xiao
Publicado: (2026)
ARISE: Agentic Rubric-Guided Iterative Survey Engine for Automated Scholarly Paper Generation
por: Wang, Zi, et al.
Publicado: (2025)
por: Wang, Zi, et al.
Publicado: (2025)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
por: Kalušev, Vladimir, et al.
Publicado: (2026)
por: Kalušev, Vladimir, et al.
Publicado: (2026)
Adaptive Minds: Empowering Agents with LoRA-as-Tools
por: Shekar, Pavan C, et al.
Publicado: (2025)
por: Shekar, Pavan C, et al.
Publicado: (2025)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
por: Koubaa, Anis, et al.
Publicado: (2025)
por: Koubaa, Anis, et al.
Publicado: (2025)
EBIC: an open source software for high-dimensional and big data biclustering analyses
por: Orzechowski, Patryk, et al.
Publicado: (2018)
por: Orzechowski, Patryk, et al.
Publicado: (2018)
SPD-RAG: Sub-Agent Per Document Retrieval-Augmented Generation
por: Akay, Yagiz Can, et al.
Publicado: (2026)
por: Akay, Yagiz Can, et al.
Publicado: (2026)
CHORUS: An Agentic Framework for Generating Realistic Deliberation Data
por: Koursaris, A., et al.
Publicado: (2026)
por: Koursaris, A., et al.
Publicado: (2026)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
por: Alanazi, Ahmed, et al.
Publicado: (2025)
por: Alanazi, Ahmed, et al.
Publicado: (2025)
The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection
por: Repantis, Vyzantinos, et al.
Publicado: (2026)
por: Repantis, Vyzantinos, et al.
Publicado: (2026)
Toward AI VIS Co-Scientists: A General and End-to-End Agent Harness for Solving Complex Data Visualization Tasks
por: Miao, Haichao, et al.
Publicado: (2026)
por: Miao, Haichao, et al.
Publicado: (2026)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
por: Pepe, Alberto, et al.
Publicado: (2026)
por: Pepe, Alberto, et al.
Publicado: (2026)
Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
por: Singh, Diyansha
Publicado: (2026)
por: Singh, Diyansha
Publicado: (2026)
AI Agents: Evolution, Architecture, and Real-World Applications
por: Krishnan, Naveen
Publicado: (2025)
por: Krishnan, Naveen
Publicado: (2025)
A Multi-Agent Framework for Medical AI: Leveraging Fine-Tuned GPT, LLaMA, and DeepSeek R1 for Evidence-Based and Bias-Aware Clinical Query Processing
por: Nourmohammadi, Naeimeh, et al.
Publicado: (2026)
por: Nourmohammadi, Naeimeh, et al.
Publicado: (2026)
Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions
por: Pu, Qingwen, et al.
Publicado: (2026)
por: Pu, Qingwen, et al.
Publicado: (2026)
Supporting software engineering tasks with agentic AI: Demonstration on document retrieval and test scenario generation
por: Kica, Marian, et al.
Publicado: (2026)
por: Kica, Marian, et al.
Publicado: (2026)
MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional
por: Cruz, Roberto, et al.
Publicado: (2026)
por: Cruz, Roberto, et al.
Publicado: (2026)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
por: Huo, Dongjie, et al.
Publicado: (2026)
por: Huo, Dongjie, et al.
Publicado: (2026)
On the relativistic viability of multi-automaton systems: essential concepts, challenges and prospects
por: Băbeanu, Alexandru-Ionuţ
Publicado: (2024)
por: Băbeanu, Alexandru-Ionuţ
Publicado: (2024)
Client-Conditional Federated Learning via Local Training Data Statistics
por: Brännvall, Rickard
Publicado: (2026)
por: Brännvall, Rickard
Publicado: (2026)
Model Callers for Transforming Predictive and Generative AI Applications
por: Dalal, Mukesh
Publicado: (2024)
por: Dalal, Mukesh
Publicado: (2024)
Bridging Industrial Expertise and XR with LLM-Powered Conversational Agents
por: Tomkou, Despina, et al.
Publicado: (2025)
por: Tomkou, Despina, et al.
Publicado: (2025)
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
Decoding Fake Narratives in Spreading Hateful Stories: A Dual-Head RoBERTa Model with Multi-Task Learning
por: Bhaskar, Yash, et al.
Publicado: (2025)
por: Bhaskar, Yash, et al.
Publicado: (2025)
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
Scalable Heterogeneous Graph Foundation Models for Data-Driven Optimal Power Flow in Smart Grids
por: Pasini, Massimiliano Lupo, et al.
Publicado: (2026)
por: Pasini, Massimiliano Lupo, et al.
Publicado: (2026)
Your Data, My Model: Learning Who Really Helps in Federated Learning
por: Abdurakhmanova, Shamsiiat, et al.
Publicado: (2024)
por: Abdurakhmanova, Shamsiiat, et al.
Publicado: (2024)
Dynamic Dual-Granularity Skill Bank for Agentic RL
por: Tu, Songjun, et al.
Publicado: (2026)
por: Tu, Songjun, et al.
Publicado: (2026)
Geist in the Machine: Simulating Recognition and Inner Dialogue in AI-Mediated Teaching and Research
por: Magee, Liam
Publicado: (2026)
por: Magee, Liam
Publicado: (2026)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
por: Wang, Zhen, et al.
Publicado: (2025)
por: Wang, Zhen, et al.
Publicado: (2025)
Manipulating Transformer-Based Models: Controllability, Steerability, and Robust Interventions
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis
por: Xi, Wang, et al.
Publicado: (2025)
por: Xi, Wang, et al.
Publicado: (2025)
Go Big or Go Home: Simulating Mobbing Behavior with Braitenbergian Robots
por: Sanoubari, Elaheh
Publicado: (2026)
por: Sanoubari, Elaheh
Publicado: (2026)
LTL Verification of Memoryful Neural Agents
por: Hosseini, Mehran, et al.
Publicado: (2025)
por: Hosseini, Mehran, et al.
Publicado: (2025)
Did You Check the Right Pocket? Cost-Sensitive Store Routing for Memory-Augmented Agents
por: Gaikwad, Madhava
Publicado: (2026)
por: Gaikwad, Madhava
Publicado: (2026)
Ejemplares similares
-
Federated k-Means over Networks
por: Yang, Xu, et al.
Publicado: (2025) -
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
por: Berman, Shmuel, et al.
Publicado: (2024) -
Emergent Collective Memory in Decentralized Multi-Agent AI Systems
por: Khushiyant
Publicado: (2025) -
Adaptive Multi-Stage Patent Claim Generation with Unified Quality Assessment
por: Liang, Chen-Wei, et al.
Publicado: (2026) -
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026)