Adapting Like Humans: A Metacognitive Agent with Test-time Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yang, He, Zhiyuan, Huang, Yuxuan, Xiao, Zhuhanling, Yu, Chao, Fang, Meng, Shao, Kun, Wang, Jun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking
por: Lee, Ka Yiu, et al.
Publicado: (2026)
por: Lee, Ka Yiu, et al.
Publicado: (2026)
CoT2-Meta: Budgeted Metacognitive Control for Test-Time Reasoning
por: Ma, Siyuan, et al.
Publicado: (2026)
por: Ma, Siyuan, et al.
Publicado: (2026)
From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company
por: Yu, Zhengxu, et al.
Publicado: (2026)
por: Yu, Zhengxu, et al.
Publicado: (2026)
Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents
por: Ding, Renhua, et al.
Publicado: (2025)
por: Ding, Renhua, et al.
Publicado: (2025)
Traffic-R1: Reinforced LLMs Bring Human-Like Reasoning to Traffic Signal Control Systems
por: Zou, Xingchen, et al.
Publicado: (2025)
por: Zou, Xingchen, et al.
Publicado: (2025)
Deep Research Agents: A Systematic Examination And Roadmap
por: Huang, Yuxuan, et al.
Publicado: (2025)
por: Huang, Yuxuan, et al.
Publicado: (2025)
Web2BigTable: A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction
por: Huang, Yuxuan, et al.
Publicado: (2026)
por: Huang, Yuxuan, et al.
Publicado: (2026)
Beyond Meta-Reasoning: Metacognitive Consolidation for Self-Improving LLM Reasoning
por: Zhuang, Ziqing, et al.
Publicado: (2026)
por: Zhuang, Ziqing, et al.
Publicado: (2026)
Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets
por: Zhang, Simpson, et al.
Publicado: (2025)
por: Zhang, Simpson, et al.
Publicado: (2025)
Adaptive Collaboration with Humans: Metacognitive Policy Optimization for Multi-Agent LLMs with Continual Learning
por: Yang, Wei, et al.
Publicado: (2026)
por: Yang, Wei, et al.
Publicado: (2026)
GTA1: GUI Test-time Scaling Agent
por: Yang, Yan, et al.
Publicado: (2025)
por: Yang, Yan, et al.
Publicado: (2025)
PRISM-MCTS: Learning from Reasoning Trajectories with Metacognitive Reflection
por: Cheng, Siyuan, et al.
Publicado: (2026)
por: Cheng, Siyuan, et al.
Publicado: (2026)
ARISE: An Adaptive Resolution-Aware Metric for Test-Time Scaling Evaluation in Large Reasoning Models
por: Yin, Zhangyue, et al.
Publicado: (2025)
por: Yin, Zhangyue, et al.
Publicado: (2025)
Know What You Know: Metacognitive Entropy Calibration for Verifiable RL Reasoning
por: Zhao, Qiannian, et al.
Publicado: (2026)
por: Zhao, Qiannian, et al.
Publicado: (2026)
Social-R1: Towards Human-like Social Reasoning in LLMs
por: Wu, Jincenzi, et al.
Publicado: (2026)
por: Wu, Jincenzi, et al.
Publicado: (2026)
Toward Safety-First Human-Like Decision Making for Autonomous Vehicles in Time-Varying Traffic Flow
por: Wang, Xiao, et al.
Publicado: (2025)
por: Wang, Xiao, et al.
Publicado: (2025)
Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents
por: Xu, Tianshi, et al.
Publicado: (2026)
por: Xu, Tianshi, et al.
Publicado: (2026)
Visual Generation Unlocks Human-Like Reasoning through Multimodal World Models
por: Wu, Jialong, et al.
Publicado: (2026)
por: Wu, Jialong, et al.
Publicado: (2026)
Exploring the Potential of Metacognitive Support Agents for Human-AI Co-Creation
por: Gmeiner, Frederic, et al.
Publicado: (2025)
por: Gmeiner, Frederic, et al.
Publicado: (2025)
Fully Test-time Adaptation for Tabular Data
por: Zhou, Zhi, et al.
Publicado: (2024)
por: Zhou, Zhi, et al.
Publicado: (2024)
Towards Self-Evolving Benchmarks: Synthesizing Agent Trajectories via Test-Time Exploration under Validate-by-Reproduce Paradigm
por: Guo, Dadi, et al.
Publicado: (2025)
por: Guo, Dadi, et al.
Publicado: (2025)
Simulating Human-Like Learning Dynamics with LLM-Empowered Agents
por: Yuan, Yu, et al.
Publicado: (2025)
por: Yuan, Yu, et al.
Publicado: (2025)
Hi-Agent: Hierarchical Vision-Language Agents for Mobile Device Control
por: Wu, Zhe, et al.
Publicado: (2025)
por: Wu, Zhe, et al.
Publicado: (2025)
ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning
por: Liao, Huanxuan, et al.
Publicado: (2026)
por: Liao, Huanxuan, et al.
Publicado: (2026)
Meta-R1: Empowering Large Reasoning Models with Metacognition
por: Dong, Haonan, et al.
Publicado: (2025)
por: Dong, Haonan, et al.
Publicado: (2025)
Language Models Coupled with Metacognition Can Outperform Reasoning Models
por: Khandelwal, Vedant, et al.
Publicado: (2025)
por: Khandelwal, Vedant, et al.
Publicado: (2025)
Scaling Test-time Compute for LLM Agents
por: Zhu, King, et al.
Publicado: (2025)
por: Zhu, King, et al.
Publicado: (2025)
AgentSafe: Safeguarding Large Language Model-based Multi-agent Systems via Hierarchical Data Management
por: Mao, Junyuan, et al.
Publicado: (2025)
por: Mao, Junyuan, et al.
Publicado: (2025)
Dynamic Sampling that Adapts: Self-Aware Iterative Data Persistent Optimization for Mathematical Reasoning
por: Rao, Jun, et al.
Publicado: (2025)
por: Rao, Jun, et al.
Publicado: (2025)
Spiral of Silence in Large Language Model Agents
por: Zhong, Mingze, et al.
Publicado: (2025)
por: Zhong, Mingze, et al.
Publicado: (2025)
GraphAgent: Agentic Graph Language Assistant
por: Yang, Yuhao, et al.
Publicado: (2024)
por: Yang, Yuhao, et al.
Publicado: (2024)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
por: Trae Research Team, et al.
Publicado: (2025)
por: Trae Research Team, et al.
Publicado: (2025)
LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
CogGuide: Human-Like Guidance for Zero-Shot Omni-Modal Reasoning
por: Shou, Zhou-Peng, et al.
Publicado: (2025)
por: Shou, Zhou-Peng, et al.
Publicado: (2025)
Truly Self-Improving Agents Require Intrinsic Metacognitive Learning
por: Liu, Tennison, et al.
Publicado: (2025)
por: Liu, Tennison, et al.
Publicado: (2025)
MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
por: Chen, Ruihan, et al.
Publicado: (2025)
por: Chen, Ruihan, et al.
Publicado: (2025)
Beyond Chemical QA: Evaluating LLM's Chemical Reasoning with Modular Chemical Operations
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
por: Verma, Gaurav, et al.
Publicado: (2024)
por: Verma, Gaurav, et al.
Publicado: (2024)
Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification
por: Wan, Yuxuan, et al.
Publicado: (2026)
por: Wan, Yuxuan, et al.
Publicado: (2026)
Collaborative Multi-Agent Test-Time Reinforcement Learning for Reasoning
por: Hu, Zhiyuan, et al.
Publicado: (2026)
por: Hu, Zhiyuan, et al.
Publicado: (2026)
Ejemplares similares
-
InfoSeeker: A Scalable Hierarchical Parallel Agent Framework for Web Information Seeking
por: Lee, Ka Yiu, et al.
Publicado: (2026) -
CoT2-Meta: Budgeted Metacognitive Control for Test-Time Reasoning
por: Ma, Siyuan, et al.
Publicado: (2026) -
From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company
por: Yu, Zhengxu, et al.
Publicado: (2026) -
Invisible to Humans, Triggered by Agents: Stealthy Jailbreak Attacks on Mobile Vision-Language Agents
por: Ding, Renhua, et al.
Publicado: (2025) -
Traffic-R1: Reinforced LLMs Bring Human-Like Reasoning to Traffic Signal Control Systems
por: Zou, Xingchen, et al.
Publicado: (2025)