Cradle: Empowering Foundation Agents Towards General Computer Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Weihao, Zhang, Wentao, Xu, Xinrun, Xia, Haochong, Ding, Ziluo, Li, Boyu, Zhou, Bohan, Yue, Junpeng, Jiang, Jiechuan, Li, Yewen, An, Ruyi, Qin, Molei, Zong, Chuqiao, Zheng, Longtao, Wu, Yujie, Chai, Xiaoqiang, Bi, Yifei, Xie, Tianbao, Gu, Pengjie, Li, Xiyun, Zhang, Ceyao, Tian, Long, Wang, Chaojie, Wang, Xinrun, Karlsson, Börje F., An, Bo, Yan, Shuicheng, Lu, Zongqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
von: Yue, Junpeng, et al.
Veröffentlicht: (2024)
von: Yue, Junpeng, et al.
Veröffentlicht: (2024)
FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures Trading
von: Qin, Molei, et al.
Veröffentlicht: (2025)
von: Qin, Molei, et al.
Veröffentlicht: (2025)
A Survey on Game Playing Agents and Large Models: Methods, Applications, and Challenges
von: Xu, Xinrun, et al.
Veröffentlicht: (2024)
von: Xu, Xinrun, et al.
Veröffentlicht: (2024)
MacroHFT: Memory Augmented Context-aware Reinforcement Learning On High Frequency Trading
von: Zong, Chuqiao, et al.
Veröffentlicht: (2024)
von: Zong, Chuqiao, et al.
Veröffentlicht: (2024)
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
von: Li, Boyu, et al.
Veröffentlicht: (2026)
von: Li, Boyu, et al.
Veröffentlicht: (2026)
SELU: Self-Learning Embodied MLLMs in Unknown Environments
von: Li, Boyu, et al.
Veröffentlicht: (2024)
von: Li, Boyu, et al.
Veröffentlicht: (2024)
Resultant: Incremental Effectiveness on Likelihood for Unsupervised Out-of-Distribution Detection
von: Li, Yewen, et al.
Veröffentlicht: (2024)
von: Li, Yewen, et al.
Veröffentlicht: (2024)
History Is Not Enough: An Adaptive Dataflow System for Financial Time-Series Synthesis
von: Xia, Haochong, et al.
Veröffentlicht: (2026)
von: Xia, Haochong, et al.
Veröffentlicht: (2026)
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
Market-GAN: Adding Control to Financial Market Data Generation with Semantic Context
von: Xia, Haochong, et al.
Veröffentlicht: (2023)
von: Xia, Haochong, et al.
Veröffentlicht: (2023)
AgentStudio: A Toolkit for Building General Virtual Agents
von: Zheng, Longtao, et al.
Veröffentlicht: (2024)
von: Zheng, Longtao, et al.
Veröffentlicht: (2024)
Towards Proprioception-Aware Embodied Planning for Dual-Arm Humanoid Robots
von: Li, Boyu, et al.
Veröffentlicht: (2025)
von: Li, Boyu, et al.
Veröffentlicht: (2025)
True Knowledge Comes from Practice: Aligning LLMs with Embodied Environments via Reinforcement Learning
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
von: Tan, Weihao, et al.
Veröffentlicht: (2024)
Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control
von: Zheng, Longtao, et al.
Veröffentlicht: (2023)
von: Zheng, Longtao, et al.
Veröffentlicht: (2023)
RL from Physical Feedback: Aligning Large Motion Models with Humanoid Control
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
Egocentric Vision Language Planning
von: Fang, Zhirui, et al.
Veröffentlicht: (2024)
von: Fang, Zhirui, et al.
Veröffentlicht: (2024)
From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and Generalist
von: Zhang, Wentao, et al.
Veröffentlicht: (2024)
von: Zhang, Wentao, et al.
Veröffentlicht: (2024)
Fully Decentralized Cooperative Multi-Agent Reinforcement Learning: A Survey
von: Jiang, Jiechuan, et al.
Veröffentlicht: (2024)
von: Jiang, Jiechuan, et al.
Veröffentlicht: (2024)
FinWorld: An All-in-One Open-Source Platform for End-to-End Financial AI Research and Deployment
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning
von: Liu, Jiazheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiazheng, et al.
Veröffentlicht: (2025)
Reinforcement Learning Friendly Vision-Language Model for Minecraft
von: Jiang, Haobin, et al.
Veröffentlicht: (2023)
von: Jiang, Haobin, et al.
Veröffentlicht: (2023)
Beyond Hungarian: Match-Free Supervision for End-to-End Object Detection
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2026)
JAEGER: Dual-Level Humanoid Whole-Body Controller
von: Ding, Ziluo, et al.
Veröffentlicht: (2025)
von: Ding, Ziluo, et al.
Veröffentlicht: (2025)
Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
DualTHOR: A Dual-Arm Humanoid Simulation Platform for Contingency-Aware Planning
von: Li, Boyu, et al.
Veröffentlicht: (2025)
von: Li, Boyu, et al.
Veröffentlicht: (2025)
EgoActor: Grounding Task Planning into Spatial-aware Egocentric Actions for Humanoid Robots via Visual-Language Models
von: Bai, Yu, et al.
Veröffentlicht: (2026)
von: Bai, Yu, et al.
Veröffentlicht: (2026)
GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
von: Deng, Yanchen, et al.
Veröffentlicht: (2025)
Why Regression? Binary Encoding Classification Brings Confidence to Stock Market Index Price Prediction
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
von: Jiang, Junzhe, et al.
Veröffentlicht: (2025)
Bayesian Robust Financial Trading with Adversarial Synthetic Market Data
von: Xia, Haochong, et al.
Veröffentlicht: (2026)
von: Xia, Haochong, et al.
Veröffentlicht: (2026)
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
von: Xu, Xinrun, et al.
Veröffentlicht: (2025)
MindRef: Mimicking Human Memory for Hierarchical Reference Retrieval with Fine-Grained Location Awareness
von: Wang, Ye, et al.
Veröffentlicht: (2024)
von: Wang, Ye, et al.
Veröffentlicht: (2024)
StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley
von: Tan, Weihao, et al.
Veröffentlicht: (2025)
von: Tan, Weihao, et al.
Veröffentlicht: (2025)
PC-BEV: An Efficient Polar-Cartesian BEV Fusion Framework for LiDAR Semantic Segmentation
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2024)
von: Qiu, Shoumeng, et al.
Veröffentlicht: (2024)
In the Cradle of Pharmacometric Methodology: Introducing Population PKPD Modeling, Simultaneous Analysis, and the Effect‐Compartment Model—Commentary on Sheiner et al.
von: Mats O. Karlsson
Veröffentlicht: (2025)
von: Mats O. Karlsson
Veröffentlicht: (2025)
ICPRL: Acquiring Physical Intuition from Interactive Control
von: Xu, Xinrun, et al.
Veröffentlicht: (2026)
von: Xu, Xinrun, et al.
Veröffentlicht: (2026)
keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM
von: Wang, Chaojie, et al.
Veröffentlicht: (2023)
von: Wang, Chaojie, et al.
Veröffentlicht: (2023)
SAM-OCTA: Prompting Segment-Anything for OCTA Image Segmentation
von: Chen, Xinrun, et al.
Veröffentlicht: (2023)
von: Chen, Xinrun, et al.
Veröffentlicht: (2023)
AdaGAT: Adaptive Guidance Adversarial Training for the Robustness of Deep Neural Networks
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
von: Yue, Junpeng, et al.
Veröffentlicht: (2024) -
FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures Trading
von: Qin, Molei, et al.
Veröffentlicht: (2025) -
A Survey on Game Playing Agents and Large Models: Methods, Applications, and Challenges
von: Xu, Xinrun, et al.
Veröffentlicht: (2024) -
MacroHFT: Memory Augmented Context-aware Reinforcement Learning On High Frequency Trading
von: Zong, Chuqiao, et al.
Veröffentlicht: (2024) -
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
von: Li, Boyu, et al.
Veröffentlicht: (2026)