Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hongru, Qian, Cheng, Li, Manling, Qiu, Jiahao, Xue, Boyang, Wang, Mengdi, Ji, Heng, Storkey, Amos, Wong, Kam-Fai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
Acting Less is Reasoning More! Teaching Model to Act Efficiently
by: Wang, Hongru, et al.
Published: (2025)
by: Wang, Hongru, et al.
Published: (2025)
Enhancing Large Language Models Against Inductive Instructions with Dual-critique Prompting
by: Wang, Rui, et al.
Published: (2023)
by: Wang, Rui, et al.
Published: (2023)
Self-DC: When to Reason and When to Act? Self Divide-and-Conquer for Compositional Unknown Questions
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
MlingConf: A Comprehensive Study of Multilingual Confidence Estimation on Large Language Models
by: Xue, Boyang, et al.
Published: (2024)
by: Xue, Boyang, et al.
Published: (2024)
UniRetriever: Multi-task Candidates Selection for Various Context-Adaptive Conversational Retrieval
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
VLEU: a Method for Automatic Evaluation for Generalizability of Text-to-Image Models
by: Cao, Jingtao, et al.
Published: (2024)
by: Cao, Jingtao, et al.
Published: (2024)
Role Prompting Guided Domain Adaptation with General Capability Preserve for Large Language Models
by: Wang, Rui, et al.
Published: (2024)
by: Wang, Rui, et al.
Published: (2024)
AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
MlingConf: A Comprehensive Study of Multilingual Confidence Estimation on Large Language Models
by: Xue, Boyang, et al.
Published: (2024)
by: Xue, Boyang, et al.
Published: (2024)
UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models
by: Xue, Boyang, et al.
Published: (2024)
by: Xue, Boyang, et al.
Published: (2024)
Rationality Measurement and Theory for Reinforcement Learning Agents
by: Qian, Kejiang, et al.
Published: (2026)
by: Qian, Kejiang, et al.
Published: (2026)
Adversarial robustness of VAEs through the lens of local geometry
by: Khan, Asif, et al.
Published: (2022)
by: Khan, Asif, et al.
Published: (2022)
Label Noise: Correcting the Forward-Correction
by: Toner, William, et al.
Published: (2023)
by: Toner, William, et al.
Published: (2023)
Noisy Early Stopping for Noisy Labels
by: Toner, William, et al.
Published: (2024)
by: Toner, William, et al.
Published: (2024)
OSPC: Detecting Harmful Memes with Large Language Model as a Catalyst
by: Cao, Jingtao, et al.
Published: (2024)
by: Cao, Jingtao, et al.
Published: (2024)
DAST: Difficulty-Aware Self-Training on Large Language Models
by: Xue, Boyang, et al.
Published: (2025)
by: Xue, Boyang, et al.
Published: (2025)
ReliableMath: Benchmark of Reliable Mathematical Reasoning on Large Language Models
by: Xue, Boyang, et al.
Published: (2025)
by: Xue, Boyang, et al.
Published: (2025)
Approximate Bayesian Class-Conditional Models under Continuous Representation Shift
by: Lee, Thomas L., et al.
Published: (2023)
by: Lee, Thomas L., et al.
Published: (2023)
Chunking: Continual Learning is not just about Distribution Shift
by: Lee, Thomas L., et al.
Published: (2023)
by: Lee, Thomas L., et al.
Published: (2023)
A Survey of the Evolution of Language Model-Based Dialogue Systems: Data, Task and Models
by: Wang, Hongru, et al.
Published: (2023)
by: Wang, Hongru, et al.
Published: (2023)
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
by: Wang, Hongru, et al.
Published: (2025)
by: Wang, Hongru, et al.
Published: (2025)
Re-Invoke: Tool Invocation Rewriting for Zero-Shot Tool Retrieval
by: Chen, Yanfei, et al.
Published: (2024)
by: Chen, Yanfei, et al.
Published: (2024)
Quantum-Secured Device-Independent Global Positioning System
by: Kam, Chon-Fai, et al.
Published: (2025)
by: Kam, Chon-Fai, et al.
Published: (2025)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
by: Gupta, Akash, et al.
Published: (2025)
by: Gupta, Akash, et al.
Published: (2025)
Self-Guard: Empower the LLM to Safeguard Itself
by: Wang, Zezhong, et al.
Published: (2023)
by: Wang, Zezhong, et al.
Published: (2023)
PerLTQA: A Personal Long-Term Memory Dataset for Memory Classification, Retrieval, and Synthesis in Question Answering
by: Du, Yiming, et al.
Published: (2024)
by: Du, Yiming, et al.
Published: (2024)
When to Invoke: Refining LLM Fairness with Toxicity Assessment
by: Ren, Jing, et al.
Published: (2026)
by: Ren, Jing, et al.
Published: (2026)
Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst
by: Wang, Hongru, et al.
Published: (2025)
by: Wang, Hongru, et al.
Published: (2025)
ToolRL: Reward is All Tool Learning Needs
by: Qian, Cheng, et al.
Published: (2025)
by: Qian, Cheng, et al.
Published: (2025)
Nonlinear optical analogues of quantum phase transitions in a squeezing-enhanced LMG model
by: Kam, Chon-Fai
Published: (2025)
by: Kam, Chon-Fai
Published: (2025)
Majorana Constellations: A Geometric Lens on Multipartite Entanglement and Geometric Phases
by: Kam, Chon-Fai
Published: (2026)
by: Kam, Chon-Fai
Published: (2026)
Three-Axis Spin Squeezed States Associated with Excited-State Quantum Phase Transitions
by: Kam, Chon-Fai
Published: (2025)
by: Kam, Chon-Fai
Published: (2025)
Nonlinear optical realization of non-integrable phases accompanying quantum phase transitions
by: Kam, Chon-Fai
Published: (2025)
by: Kam, Chon-Fai
Published: (2025)
Admissibility and Standing in Artificial Epistemic Agents
by: Maley, Amos
Published: (2026)
by: Maley, Amos
Published: (2026)
Counterspeech for Mitigating the Influence of Media Bias: Comparing Human and LLM-Generated Responses
by: Lin, Luyang, et al.
Published: (2025)
by: Lin, Luyang, et al.
Published: (2025)
Investigating Bias in LLM-Based Bias Detection: Disparities between LLMs and Human Perception
by: Lin, Luyang, et al.
Published: (2024)
by: Lin, Luyang, et al.
Published: (2024)
ToolFlow: Boosting LLM Tool-Calling Through Natural and Coherent Dialogue Synthesis
by: Wang, Zezhong, et al.
Published: (2024)
by: Wang, Zezhong, et al.
Published: (2024)
New Necessary Conditions for Existence of Strong External Difference Families
by: Bao, Jingjun, et al.
Published: (2025)
by: Bao, Jingjun, et al.
Published: (2025)
WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
Similar Items
-
Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models
by: Wang, Rui, et al.
Published: (2025) -
Acting Less is Reasoning More! Teaching Model to Act Efficiently
by: Wang, Hongru, et al.
Published: (2025) -
Enhancing Large Language Models Against Inductive Instructions with Dual-critique Prompting
by: Wang, Rui, et al.
Published: (2023) -
Self-DC: When to Reason and When to Act? Self Divide-and-Conquer for Compositional Unknown Questions
by: Wang, Hongru, et al.
Published: (2024) -
MlingConf: A Comprehensive Study of Multilingual Confidence Estimation on Large Language Models
by: Xue, Boyang, et al.
Published: (2024)