Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
Fuente:
arXiv
Salvato in:
| Autori principali: | Fan, Flint Xiaofeng, Tan, Cheston, Wattenhofer, Roger, Ong, Yew-Soon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2025)
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2025)
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2024)
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2024)
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
di: Mak, Hei Yi, et al.
Pubblicazione: (2024)
di: Mak, Hei Yi, et al.
Pubblicazione: (2024)
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
di: Wei, Zhao, et al.
Pubblicazione: (2025)
di: Wei, Zhao, et al.
Pubblicazione: (2025)
Automated Large-scale CVRP Solver Design via LLM-assisted Flexible MCTS
di: Guo, Tong, et al.
Pubblicazione: (2026)
di: Guo, Tong, et al.
Pubblicazione: (2026)
Decentralized Federated Policy Gradient with Byzantine Fault-Tolerance and Provably Fast Convergence
di: Jordan, Philip, et al.
Pubblicazione: (2024)
di: Jordan, Philip, et al.
Pubblicazione: (2024)
Diversifying Policy Behaviors with Extrinsic Behavioral Curiosity
di: Wan, Zhenglin, et al.
Pubblicazione: (2024)
di: Wan, Zhenglin, et al.
Pubblicazione: (2024)
Multi-Task Learning with Multi-Task Optimization
di: Bai, Lu, et al.
Pubblicazione: (2024)
di: Bai, Lu, et al.
Pubblicazione: (2024)
Flow-Direct: Feedback-Efficient and Reusable Guidance for Flow Models via Non-Parametric Guidance Field
di: Tan, Kim Yong, et al.
Pubblicazione: (2026)
di: Tan, Kim Yong, et al.
Pubblicazione: (2026)
FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation
di: Jiang, Wenzheng, et al.
Pubblicazione: (2025)
di: Jiang, Wenzheng, et al.
Pubblicazione: (2025)
Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
di: Maloyan, Narek, et al.
Pubblicazione: (2026)
di: Maloyan, Narek, et al.
Pubblicazione: (2026)
Reasoning Effort and Problem Complexity: A Scaling Analysis in LLMs
di: Estermann, Benjamin, et al.
Pubblicazione: (2025)
di: Estermann, Benjamin, et al.
Pubblicazione: (2025)
Fast Direct: Query-Efficient Online Black-box Guidance for Diffusion-model Target Generation
di: Tan, Kim Yong, et al.
Pubblicazione: (2025)
di: Tan, Kim Yong, et al.
Pubblicazione: (2025)
Distributional Multi-objective Black-box Optimization for Diffusion-model Inference-time Multi-Target Generation
di: Tan, Kim Yong, et al.
Pubblicazione: (2025)
di: Tan, Kim Yong, et al.
Pubblicazione: (2025)
TangramSR: Can Vision-Language Models Reason in Continuous Geometric Space?
di: Zong, Yikun, et al.
Pubblicazione: (2026)
di: Zong, Yikun, et al.
Pubblicazione: (2026)
Adaptive In-Context Learning with Large Language Models for Bundle Generation
di: Sun, Zhu, et al.
Pubblicazione: (2023)
di: Sun, Zhu, et al.
Pubblicazione: (2023)
GRASP: Gradient Realignment via Active Shared Perception for Multi-Agent Collaborative Optimization
di: Zhou, Sihan, et al.
Pubblicazione: (2026)
di: Zhou, Sihan, et al.
Pubblicazione: (2026)
MRP-LLM: Multitask Reflective Large Language Models for Privacy-Preserving Next POI Recommendation
di: Wu, Ziqing, et al.
Pubblicazione: (2024)
di: Wu, Ziqing, et al.
Pubblicazione: (2024)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
di: He, Ping, et al.
Pubblicazione: (2025)
di: He, Ping, et al.
Pubblicazione: (2025)
MCPToolBench++: A Large Scale AI Agent Model Context Protocol MCP Tool Use Benchmark
di: Fan, Shiqing, et al.
Pubblicazione: (2025)
di: Fan, Shiqing, et al.
Pubblicazione: (2025)
A Simple Yet Effective Approach for Diversified Session-Based Recommendation
di: Yin, Qing, et al.
Pubblicazione: (2024)
di: Yin, Qing, et al.
Pubblicazione: (2024)
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies
di: Berdoz, Frédéric, et al.
Pubblicazione: (2024)
di: Berdoz, Frédéric, et al.
Pubblicazione: (2024)
MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games
di: Xie, Yunfei, et al.
Pubblicazione: (2026)
di: Xie, Yunfei, et al.
Pubblicazione: (2026)
Agentic Spatio-Temporal Grounding via Collaborative Reasoning
di: Zhao, Heng, et al.
Pubblicazione: (2026)
di: Zhao, Heng, et al.
Pubblicazione: (2026)
Learning with Foresight: Enhancing Neural Routing Policy via Multi-Node Lookahead Prediction
di: Jiang, Xia, et al.
Pubblicazione: (2026)
di: Jiang, Xia, et al.
Pubblicazione: (2026)
Where to Move Next: Zero-shot Generalization of LLMs for Next POI Recommendation
di: Feng, Shanshan, et al.
Pubblicazione: (2024)
di: Feng, Shanshan, et al.
Pubblicazione: (2024)
The Digital Ecosystem of Beliefs: does evolution favour AI over humans?
di: Bossens, David M., et al.
Pubblicazione: (2024)
di: Bossens, David M., et al.
Pubblicazione: (2024)
SiamNAS: Siamese Surrogate Model for Dominance Relation Prediction in Multi-objective Neural Architecture Search
di: Zhou, Yuyang, et al.
Pubblicazione: (2025)
di: Zhou, Yuyang, et al.
Pubblicazione: (2025)
An Agentic Framework with LLMs for Solving Complex Vehicle Routing Problems
di: Zhang, Ni, et al.
Pubblicazione: (2025)
di: Zhang, Ni, et al.
Pubblicazione: (2025)
MazeMate: An LLM-Powered Chatbot to Support Computational Thinking in Gamified Programming Learning
di: Hou, Chenyu, et al.
Pubblicazione: (2025)
di: Hou, Chenyu, et al.
Pubblicazione: (2025)
Policy Dispersion in Non-Markovian Environment
di: Qu, Bohao, et al.
Pubblicazione: (2023)
di: Qu, Bohao, et al.
Pubblicazione: (2023)
Human-Generative AI Collaborative Problem Solving Who Leads and How Students Perceive the Interactions
di: Zhu, Gaoxia, et al.
Pubblicazione: (2024)
di: Zhu, Gaoxia, et al.
Pubblicazione: (2024)
FinMCP-Bench: Benchmarking LLM Agents for Real-World Financial Tool Use under the Model Context Protocol
di: Zhu, Jie, et al.
Pubblicazione: (2026)
di: Zhu, Jie, et al.
Pubblicazione: (2026)
Social Learning through Interactions with Other Agents: A Survey
di: Hillier, Dylan, et al.
Pubblicazione: (2024)
di: Hillier, Dylan, et al.
Pubblicazione: (2024)
Large Language Models can be Guided to Evade AI-Generated Text Detection
di: Lu, Ning, et al.
Pubblicazione: (2023)
di: Lu, Ning, et al.
Pubblicazione: (2023)
AEye: A Visualization Tool for Image Datasets
di: Grötschla, Florian, et al.
Pubblicazione: (2024)
di: Grötschla, Florian, et al.
Pubblicazione: (2024)
Possibilistic Predictive Uncertainty for Deep Learning
di: Ni, Yao, et al.
Pubblicazione: (2026)
di: Ni, Yao, et al.
Pubblicazione: (2026)
Towards Trustworthy Vital Sign Forecasting: Leveraging Uncertainty for Prediction Intervals
di: Wang, Li Rong, et al.
Pubblicazione: (2025)
di: Wang, Li Rong, et al.
Pubblicazione: (2025)
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
Hierarchically Robust Zero-shot Vision-language Models
di: Dong, Junhao, et al.
Pubblicazione: (2026)
di: Dong, Junhao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2025) -
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
di: Fan, Flint Xiaofeng, et al.
Pubblicazione: (2024) -
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
di: Mak, Hei Yi, et al.
Pubblicazione: (2024) -
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
di: Wei, Zhao, et al.
Pubblicazione: (2025) -
Automated Large-scale CVRP Solver Design via LLM-assisted Flexible MCTS
di: Guo, Tong, et al.
Pubblicazione: (2026)