Profile-Aware Maneuvering: A Dynamic Multi-Agent System for Robust GAIA Problem Solving by AWorld
Fuente:
arXiv
Salvato in:
| Autori principali: | Xie, Zhitian, Wu, Qintong, Yu, Chengyue, Zhuang, Chenyi, Gu, Jinjie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AWorld: Orchestrating the Training Recipe for Agentic AI
di: Yu, Chengyue, et al.
Pubblicazione: (2025)
di: Yu, Chengyue, et al.
Pubblicazione: (2025)
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
di: Yu, Chengyue, et al.
Pubblicazione: (2024)
di: Yu, Chengyue, et al.
Pubblicazione: (2024)
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
di: Zhao, Yao, et al.
Pubblicazione: (2023)
di: Zhao, Yao, et al.
Pubblicazione: (2023)
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
di: Tan, Zhiwen, et al.
Pubblicazione: (2025)
di: Tan, Zhiwen, et al.
Pubblicazione: (2025)
Recon-Act: A Self-Evolving Multi-Agent Browser-Use System via Web Reconnaissance, Tool Generation, and Task Execution
di: He, Kaiwen, et al.
Pubblicazione: (2025)
di: He, Kaiwen, et al.
Pubblicazione: (2025)
Don't Just Fine-tune the Agent, Tune the Environment
di: Lu, Siyuan, et al.
Pubblicazione: (2025)
di: Lu, Siyuan, et al.
Pubblicazione: (2025)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
di: Xie, Zhitian, et al.
Pubblicazione: (2024)
di: Xie, Zhitian, et al.
Pubblicazione: (2024)
LiveAgentBench: Comprehensive Benchmarking of Agentic Systems Across 104 Real-World Challenges
di: Li, Hao, et al.
Pubblicazione: (2026)
di: Li, Hao, et al.
Pubblicazione: (2026)
MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants
di: Zhang, Zuhao, et al.
Pubblicazione: (2026)
di: Zhang, Zuhao, et al.
Pubblicazione: (2026)
Large Multimodal Model Compression via Efficient Pruning and Distillation at AntGroup
di: Wang, Maolin, et al.
Pubblicazione: (2023)
di: Wang, Maolin, et al.
Pubblicazione: (2023)
DNS-Rec: Data-aware Neural Architecture Search for Recommender Systems
di: Zhang, Sheng, et al.
Pubblicazione: (2024)
di: Zhang, Sheng, et al.
Pubblicazione: (2024)
Tree-based RAG-Agent Recommendation System: A Case Study in Medical Test Data
di: Yang, Yahe, et al.
Pubblicazione: (2025)
di: Yang, Yahe, et al.
Pubblicazione: (2025)
GAIA: A Foundation Model for Operational Atmospheric Dynamics
di: Asanjan, Ata Akbari, et al.
Pubblicazione: (2025)
di: Asanjan, Ata Akbari, et al.
Pubblicazione: (2025)
V2P: Visual Attention Calibration for GUI Grounding via Background Suppression and Center Peaking
di: Chen, Jikai, et al.
Pubblicazione: (2026)
di: Chen, Jikai, et al.
Pubblicazione: (2026)
V2P: Visual Attention Calibration for GUI Grounding via Background Suppression and Center Peaking
di: Chen, Jikai, et al.
Pubblicazione: (2025)
di: Chen, Jikai, et al.
Pubblicazione: (2025)
Unleashing the Denoising Capability of Diffusion Prior for Solving Inverse Problems
di: Zhang, Jiawei, et al.
Pubblicazione: (2024)
di: Zhang, Jiawei, et al.
Pubblicazione: (2024)
Multi-Agent Deep Research: Training Multi-Agent Systems with M-GRPO
di: Hong, Haoyang, et al.
Pubblicazione: (2025)
di: Hong, Haoyang, et al.
Pubblicazione: (2025)
MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System
di: Xu, Yuliang, et al.
Pubblicazione: (2026)
di: Xu, Yuliang, et al.
Pubblicazione: (2026)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
di: Parmar, Mihir, et al.
Pubblicazione: (2025)
Reasoning through Exploration: A Reinforcement Learning Framework for Robust Function Calling
di: Hao, Bingguang, et al.
Pubblicazione: (2025)
di: Hao, Bingguang, et al.
Pubblicazione: (2025)
Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems
di: Wu, Zejian Eric, et al.
Pubblicazione: (2026)
di: Wu, Zejian Eric, et al.
Pubblicazione: (2026)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
di: Alavi, Khashayar, et al.
Pubblicazione: (2025)
di: Alavi, Khashayar, et al.
Pubblicazione: (2025)
Solving Multi-Agent Multi-Goal Path Finding Problems in Polynomial Time
di: Edelkamp, Stefan
Pubblicazione: (2025)
di: Edelkamp, Stefan
Pubblicazione: (2025)
StressWeb: A Diagnostic Benchmark for Web Agent Robustness under Realistic Interaction Variability
di: Bai, Haoyue, et al.
Pubblicazione: (2026)
di: Bai, Haoyue, et al.
Pubblicazione: (2026)
EngiAgent: Fully Connected Coordination of LLM Agents for Solving Open-ended Engineering Problems with Feasible Solutions
di: Zhou, Xiyuan, et al.
Pubblicazione: (2026)
di: Zhou, Xiyuan, et al.
Pubblicazione: (2026)
GAIA-v2-LILT: Multilingual Adaptation of Agent Benchmark beyond Translation
di: Kim, Yunsu, et al.
Pubblicazione: (2026)
di: Kim, Yunsu, et al.
Pubblicazione: (2026)
MapCoder: Multi-Agent Code Generation for Competitive Problem Solving
di: Islam, Md. Ashraful, et al.
Pubblicazione: (2024)
di: Islam, Md. Ashraful, et al.
Pubblicazione: (2024)
SwiftSolve: A Self-Iterative, Complexity-Aware Multi-Agent Framework for Competitive Programming
di: Singh, Adhyayan Veer, et al.
Pubblicazione: (2025)
di: Singh, Adhyayan Veer, et al.
Pubblicazione: (2025)
GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models
di: Wang, Shaokang, et al.
Pubblicazione: (2026)
di: Wang, Shaokang, et al.
Pubblicazione: (2026)
Literature Review Of Multi-Agent Debate For Problem-Solving
di: Tillmann, Arne
Pubblicazione: (2025)
di: Tillmann, Arne
Pubblicazione: (2025)
ProSEA: Problem Solving via Exploration Agents
di: Nguyen, William, et al.
Pubblicazione: (2025)
di: Nguyen, William, et al.
Pubblicazione: (2025)
GAIA: Categorical Foundations of Generative AI
di: Mahadevan, Sridhar
Pubblicazione: (2024)
di: Mahadevan, Sridhar
Pubblicazione: (2024)
EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving
di: Zhou, Xiyuan, et al.
Pubblicazione: (2025)
di: Zhou, Xiyuan, et al.
Pubblicazione: (2025)
MACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical Problems
di: Lei, Bin, et al.
Pubblicazione: (2024)
di: Lei, Bin, et al.
Pubblicazione: (2024)
Applying Multi-Agent Negotiation to Solve the Production Routing Problem With Privacy Preserving
di: Biasoto, Luiza Pellin, et al.
Pubblicazione: (2024)
di: Biasoto, Luiza Pellin, et al.
Pubblicazione: (2024)
Enhancing LLM Problem Solving via Tutor-Student Multi-Agent Interaction
di: Özdemir, Nurullah Eymen, et al.
Pubblicazione: (2026)
di: Özdemir, Nurullah Eymen, et al.
Pubblicazione: (2026)
KABB: Knowledge-Aware Bayesian Bandits for Dynamic Expert Coordination in Multi-Agent Systems
di: Zhang, Jusheng, et al.
Pubblicazione: (2025)
di: Zhang, Jusheng, et al.
Pubblicazione: (2025)
Multi-LLM Collaborative Search for Complex Problem Solving
di: Yang, Sen, et al.
Pubblicazione: (2025)
di: Yang, Sen, et al.
Pubblicazione: (2025)
DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving
di: Tong, Yuxuan, et al.
Pubblicazione: (2024)
di: Tong, Yuxuan, et al.
Pubblicazione: (2024)
Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving
di: Gao, Songyang, et al.
Pubblicazione: (2025)
di: Gao, Songyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AWorld: Orchestrating the Training Recipe for Agentic AI
di: Yu, Chengyue, et al.
Pubblicazione: (2025) -
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
di: Yu, Chengyue, et al.
Pubblicazione: (2024) -
Lookahead: An Inference Acceleration Framework for Large Language Model with Lossless Generation Accuracy
di: Zhao, Yao, et al.
Pubblicazione: (2023) -
RAG-R1: Incentivizing the Search and Reasoning Capabilities of LLMs through Multi-query Parallelism
di: Tan, Zhiwen, et al.
Pubblicazione: (2025) -
Recon-Act: A Self-Evolving Multi-Agent Browser-Use System via Web Reconnaissance, Tool Generation, and Task Execution
di: He, Kaiwen, et al.
Pubblicazione: (2025)