Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Xueyang, Tie, Guiyao, Zhang, Guowen, Wang, Weidong, Zuo, Zhigang, Wu, Di, Chu, Duanfeng, Zhou, Pan, Gong, Neil Zhenqiang, Sun, Lichao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
von: Shi, Jiawen, et al.
Veröffentlicht: (2025)
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
A Survey of AI Scientists
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
von: Tie, Guiyao, et al.
Veröffentlicht: (2026)
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
MMLU-Reason: Benchmarking Multi-Task Multi-modal Language Understanding and Reasoning
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
BadToken: Token-level Backdoor Attacks to Multi-modal Large Language Models
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
von: Yuan, Zenghui, et al.
Veröffentlicht: (2025)
Enabling Extensible Embodied Capabilities with Tools
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2026)
Self-Cognition in Large Language Models: An Exploratory Study
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
ObliInjection: Order-Oblivious Prompt Injection Attack to LLM Agents with Multi-source Data
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
ConvoyLLM: Dynamic Multi-Lane Convoy Control Using LLMs
von: Lu, Liping, et al.
Veröffentlicht: (2025)
von: Lu, Liping, et al.
Veröffentlicht: (2025)
CHARMS: A Cognitive Hierarchical Agent for Reasoning and Motion Stylization in Autonomous Driving
von: Wang, Jingyi, et al.
Veröffentlicht: (2025)
von: Wang, Jingyi, et al.
Veröffentlicht: (2025)
Evaluating LLM-based Personal Information Extraction and Countermeasures
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
WebInject: Prompt Injection Attack to Web Agents
von: Wang, Xilong, et al.
Veröffentlicht: (2025)
von: Wang, Xilong, et al.
Veröffentlicht: (2025)
MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
von: Huang, Yue, et al.
Veröffentlicht: (2023)
von: Huang, Yue, et al.
Veröffentlicht: (2023)
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling
von: Zhou, Jialong, et al.
Veröffentlicht: (2025)
von: Zhou, Jialong, et al.
Veröffentlicht: (2025)
StringLLM: Understanding the String Processing Capability of Large Language Models
von: Wang, Xilong, et al.
Veröffentlicht: (2024)
von: Wang, Xilong, et al.
Veröffentlicht: (2024)
The Necessity of a Unified Framework for LLM-Based Agent Evaluation
von: Zhu, Pengyu, et al.
Veröffentlicht: (2026)
von: Zhu, Pengyu, et al.
Veröffentlicht: (2026)
Multi-Agent Trajectory Prediction with Difficulty-Guided Feature Enhancement Network
von: Xin, Guipeng, et al.
Veröffentlicht: (2024)
von: Xin, Guipeng, et al.
Veröffentlicht: (2024)
Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents
von: Yang, Zhejian, et al.
Veröffentlicht: (2025)
von: Yang, Zhejian, et al.
Veröffentlicht: (2025)
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
von: Zhang, Ziqian, et al.
Veröffentlicht: (2026)
CorruptEncoder: Data Poisoning based Backdoor Attacks to Contrastive Learning
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2022)
von: Zhang, Jinghuai, et al.
Veröffentlicht: (2022)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
SafeText: Safe Text-to-image Models via Aligning the Text Encoder
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2025)
Model Poisoning Attacks to Federated Learning via Multi-Round Consistency
von: Xie, Yueqi, et al.
Veröffentlicht: (2024)
von: Xie, Yueqi, et al.
Veröffentlicht: (2024)
Competitive Advantage Attacks to Decentralized Federated Learning
von: Jia, Yuqi, et al.
Veröffentlicht: (2023)
von: Jia, Yuqi, et al.
Veröffentlicht: (2023)
Provably Robust Federated Reinforcement Learning
von: Fang, Minghong, et al.
Veröffentlicht: (2025)
von: Fang, Minghong, et al.
Veröffentlicht: (2025)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
von: Huang, Yanbin, et al.
Veröffentlicht: (2026)
Revisiting the Necessity of Lengthy Chain-of-Thought in Vision-centric Reasoning Generalization
von: Du, Yifan, et al.
Veröffentlicht: (2025)
von: Du, Yifan, et al.
Veröffentlicht: (2025)
CLIP-SENet: CLIP-based Semantic Enhancement Network for Vehicle Re-identification
von: Lu, Liping, et al.
Veröffentlicht: (2025)
von: Lu, Liping, et al.
Veröffentlicht: (2025)
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
Merger-as-a-Stealer: Stealing Targeted PII from Aligned LLMs with Model Merging
von: Lu, Lin, et al.
Veröffentlicht: (2025)
von: Lu, Lin, et al.
Veröffentlicht: (2025)
Virtual Context: Enhancing Jailbreak Attacks with Special Token Injection
von: Zhou, Yuqi, et al.
Veröffentlicht: (2024)
von: Zhou, Yuqi, et al.
Veröffentlicht: (2024)
Creation of Three‐Scroll Hidden Conservative Lorenz‐Like Chaotic Flows
von: Guiyao Ke
Veröffentlicht: (2024)
von: Guiyao Ke
Veröffentlicht: (2024)
Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?
von: Fan, Chenrui, et al.
Veröffentlicht: (2025)
von: Fan, Chenrui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
Prompt Injection Attack to Tool Selection in LLM Agents
von: Shi, Jiawen, et al.
Veröffentlicht: (2025) -
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
A Survey of AI Scientists
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)