The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Baixuan, Zheng, Tianshi, Wang, Zhaowei, Tsang, Hong Ting, Wang, Weiqi, Fang, Tianqing, Song, Yangqiu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction
di: Han, Kaiqiao, et al.
Pubblicazione: (2024)
di: Han, Kaiqiao, et al.
Pubblicazione: (2024)
KnowShiftQA: How Robust are RAG Systems when Textbook Knowledge Shifts in K-12 Education?
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
From Automation to Autonomy: A Survey on Large Language Models in Scientific Discovery
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
CritiCal: Can Critique Help LLM Uncertainty or Confidence Calibration?
di: Zong, Qing, et al.
Pubblicazione: (2025)
di: Zong, Qing, et al.
Pubblicazione: (2025)
ConKE: Conceptualization-Augmented Knowledge Editing in Large Language Models for Commonsense Reasoning
di: Zhang, Liyu, et al.
Pubblicazione: (2024)
di: Zhang, Liyu, et al.
Pubblicazione: (2024)
AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
CKBP v2: Better Annotation and Reasoning for Commonsense Knowledge Base Population
di: Fang, Tianqing, et al.
Pubblicazione: (2023)
di: Fang, Tianqing, et al.
Pubblicazione: (2023)
Towards Multi-Agent Reasoning Systems for Collaborative Expertise Delegation: An Exploratory Design Study
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
di: Xu, Baixuan, et al.
Pubblicazione: (2025)
Acquiring and Modelling Abstract Commonsense Knowledge via Conceptualization
di: He, Mutian, et al.
Pubblicazione: (2022)
di: He, Mutian, et al.
Pubblicazione: (2022)
ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
di: Zong, Qing, et al.
Pubblicazione: (2024)
di: Zong, Qing, et al.
Pubblicazione: (2024)
CANDLE: Iterative Conceptualization and Instantiation Distillation from Large Language Models for Commonsense Reasoning
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
INFERENCEDYNAMICS: Efficient Routing Across LLMs through Structured Capability and Knowledge Profiling
di: Shi, Haochen, et al.
Pubblicazione: (2025)
di: Shi, Haochen, et al.
Pubblicazione: (2025)
AbsPyramid: Benchmarking the Abstraction Ability of Language Models with a Unified Entailment Graph
di: Wang, Zhaowei, et al.
Pubblicazione: (2023)
di: Wang, Zhaowei, et al.
Pubblicazione: (2023)
ConstraintChecker: A Plugin for Large Language Models to Reason on Commonsense Knowledge Bases
di: Do, Quyet V., et al.
Pubblicazione: (2024)
di: Do, Quyet V., et al.
Pubblicazione: (2024)
Legal Rule Induction: Towards Generalizable Principle Discovery from Analogous Judicial Precedents
di: Fan, Wei, et al.
Pubblicazione: (2025)
di: Fan, Wei, et al.
Pubblicazione: (2025)
CLR-Fact: Evaluating the Complex Logical Reasoning Capability of Large Language Models over Factual Knowledge
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
di: Zheng, Tianshi, et al.
Pubblicazione: (2024)
Transformers for Complex Query Answering over Knowledge Hypergraphs
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
di: Tsang, Hong Ting, et al.
Pubblicazione: (2025)
NewtonBench: Benchmarking Generalizable Scientific Law Discovery in LLM Agents
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
MARS: Benchmarking the Metaphysical Reasoning Abilities of Language Models with a Multi-task Evaluation Dataset
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
Getting Sick After Seeing a Doctor? Diagnosing and Mitigating Knowledge Conflicts in Event Temporal Reasoning
di: Fang, Tianqing, et al.
Pubblicazione: (2023)
di: Fang, Tianqing, et al.
Pubblicazione: (2023)
Structuring the Unstructured: A Systematic Review of Text-to-Structure Generation for Agentic AI with a Universal Evaluation Framework
di: Deng, Zheye, et al.
Pubblicazione: (2025)
di: Deng, Zheye, et al.
Pubblicazione: (2025)
On the Role of Entity and Event Level Conceptualization in Generalizable Reasoning: A Survey of Tasks, Methods, Applications, and Future Directions
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
di: Wang, Weiqi, et al.
Pubblicazione: (2024)
EcomEdit: An Automated E-commerce Knowledge Editing Framework for Enhanced Product and Purchase Intention Understanding
di: Lau, Ching Ming Samuel, et al.
Pubblicazione: (2024)
di: Lau, Ching Ming Samuel, et al.
Pubblicazione: (2024)
ChatGPT Evaluation on Sentence Level Relations: A Focus on Temporal, Causal, and Discourse Relations
di: Chan, Chunkit, et al.
Pubblicazione: (2023)
di: Chan, Chunkit, et al.
Pubblicazione: (2023)
EntailE: Introducing Textual Entailment in Commonsense Knowledge Graph Completion
di: Su, Ying, et al.
Pubblicazione: (2024)
di: Su, Ying, et al.
Pubblicazione: (2024)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2025)
di: Erdogan, Lutfi Eren, et al.
Pubblicazione: (2025)
SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning
di: Zheng, Tianshi, et al.
Pubblicazione: (2026)
di: Zheng, Tianshi, et al.
Pubblicazione: (2026)
KNOWCOMP POKEMON Team at DialAM-2024: A Two-Stage Pipeline for Detecting Relations in Dialogical Argument Mining
di: Zheng, Zihao, et al.
Pubblicazione: (2024)
di: Zheng, Zihao, et al.
Pubblicazione: (2024)
Empowering LLMs with Parameterized Skills for Adversarial Long-Horizon Planning
di: Cui, Sijia, et al.
Pubblicazione: (2025)
di: Cui, Sijia, et al.
Pubblicazione: (2025)
AutoSchemaKG: Autonomous Knowledge Graph Construction through Dynamic Schema Induction from Web-Scale Corpora
di: Bai, Jiaxin, et al.
Pubblicazione: (2025)
di: Bai, Jiaxin, et al.
Pubblicazione: (2025)
Text-Tuple-Table: Towards Information Integration in Text-to-Table Generation via Global Tuple Extraction
di: Deng, Zheye, et al.
Pubblicazione: (2024)
di: Deng, Zheye, et al.
Pubblicazione: (2024)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
di: Wang, Zehong, et al.
Pubblicazione: (2026)
di: Wang, Zehong, et al.
Pubblicazione: (2026)
Complex Reasoning over Logical Queries on Commonsense Knowledge Graphs
di: Fang, Tianqing, et al.
Pubblicazione: (2024)
di: Fang, Tianqing, et al.
Pubblicazione: (2024)
Patterns Over Principles: The Fragility of Inductive Reasoning in LLMs under Noisy Observations
di: Li, Chunyang, et al.
Pubblicazione: (2025)
di: Li, Chunyang, et al.
Pubblicazione: (2025)
MIND: Multimodal Shopping Intention Distillation from Large Vision-language Models for E-commerce Purchase Understanding
di: Xu, Baixuan, et al.
Pubblicazione: (2024)
di: Xu, Baixuan, et al.
Pubblicazione: (2024)
Structured Preference Optimization for Vision-Language Long-Horizon Task Planning
di: Liang, Xiwen, et al.
Pubblicazione: (2025)
di: Liang, Xiwen, et al.
Pubblicazione: (2025)
NAACL: Noise-AwAre Verbal Confidence Calibration for Robust LLMs in RAG Systems
di: Liu, Jiayu, et al.
Pubblicazione: (2026)
di: Liu, Jiayu, et al.
Pubblicazione: (2026)
The Curse of CoT: On the Limitations of Chain-of-Thought in In-Context Learning
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
di: Zheng, Tianshi, et al.
Pubblicazione: (2025)
Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?
di: Liu, Jiayu, et al.
Pubblicazione: (2025)
di: Liu, Jiayu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction
di: Han, Kaiqiao, et al.
Pubblicazione: (2024) -
KnowShiftQA: How Robust are RAG Systems when Textbook Knowledge Shifts in K-12 Education?
di: Zheng, Tianshi, et al.
Pubblicazione: (2024) -
From Automation to Autonomy: A Survey on Large Language Models in Scientific Discovery
di: Zheng, Tianshi, et al.
Pubblicazione: (2025) -
CritiCal: Can Critique Help LLM Uncertainty or Confidence Calibration?
di: Zong, Qing, et al.
Pubblicazione: (2025) -
ConKE: Conceptualization-Augmented Knowledge Editing in Large Language Models for Commonsense Reasoning
di: Zhang, Liyu, et al.
Pubblicazione: (2024)