Position Paper: Agent AI Towards a Holistic Intelligence
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Qiuyuan, Wake, Naoki, Sarkar, Bidipta, Durante, Zane, Gong, Ran, Taori, Rohan, Noda, Yusuke, Terzopoulos, Demetri, Kuno, Noboru, Famoti, Ade, Llorens, Ashley, Langford, John, Vo, Hoi, Fei-Fei, Li, Ikeuchi, Katsu, Gao, Jianfeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Interactive Agent Foundation Model
by: Durante, Zane, et al.
Published: (2024)
by: Durante, Zane, et al.
Published: (2024)
Agent AI: Surveying the Horizons of Multimodal Interaction
by: Durante, Zane, et al.
Published: (2024)
by: Durante, Zane, et al.
Published: (2024)
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering
by: Guo, Danfeng, et al.
Published: (2024)
by: Guo, Danfeng, et al.
Published: (2024)
Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
RL-Driven Data Generation for Robust Vision-Based Dexterous Grasping
by: Kanehira, Atsushi, et al.
Published: (2025)
by: Kanehira, Atsushi, et al.
Published: (2025)
Plan-and-Act using Large Language Models for Interactive Agreement
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
GPT-4V(ision) for Robotics: Multimodal Task Planning from Human Demonstration
by: Wake, Naoki, et al.
Published: (2023)
by: Wake, Naoki, et al.
Published: (2023)
Open-Vocabulary Action Localization with Iterative Visual Prompting
by: Wake, Naoki, et al.
Published: (2024)
by: Wake, Naoki, et al.
Published: (2024)
VLM-driven Behavior Tree for Context-aware Task Planning
by: Wake, Naoki, et al.
Published: (2025)
by: Wake, Naoki, et al.
Published: (2025)
A Taxonomy of Self-Handover
by: Wake, Naoki, et al.
Published: (2025)
by: Wake, Naoki, et al.
Published: (2025)
IK Seed Generator for Dual-Arm Human-like Physicality Robot with Mobile Base
by: Takamatsu, Jun, et al.
Published: (2025)
by: Takamatsu, Jun, et al.
Published: (2025)
Inverse Attention Agents for Multi-Agent Systems
by: Long, Qian, et al.
Published: (2024)
by: Long, Qian, et al.
Published: (2024)
Designing Library of Skill-Agents for Hardware-Level Reusability
by: Takamatsu, Jun, et al.
Published: (2024)
by: Takamatsu, Jun, et al.
Published: (2024)
The Essential Role of Causality in Foundation World Models for Embodied AI
by: Gupta, Tarun, et al.
Published: (2024)
by: Gupta, Tarun, et al.
Published: (2024)
Learning Neural Force Manifolds for Sim2Real Robotic Symmetrical Paper Folding
by: Choi, Andrew, et al.
Published: (2023)
by: Choi, Andrew, et al.
Published: (2023)
A Human-centric Framework for Debating the Ethics of AI Consciousness Under Uncertainty
by: Ziheng, Zhou, et al.
Published: (2025)
by: Ziheng, Zhou, et al.
Published: (2025)
Trajectories of second language student classroom engagement: Profiles and correlates
by: Hoi Vo
Published: (2026)
by: Hoi Vo
Published: (2026)
VideoWeave: A Data-Centric Approach for Efficient Video Understanding
by: Durante, Zane, et al.
Published: (2026)
by: Durante, Zane, et al.
Published: (2026)
Modality-Driven Design for Multi-Step Dexterous Manipulation: Insights from Neuroscience
by: Wake, Naoki, et al.
Published: (2024)
by: Wake, Naoki, et al.
Published: (2024)
APriCoT: Action Primitives based on Contact-state Transition for In-Hand Tool Manipulation
by: Saito, Daichi, et al.
Published: (2024)
by: Saito, Daichi, et al.
Published: (2024)
TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
by: Long, Qian, et al.
Published: (2024)
by: Long, Qian, et al.
Published: (2024)
mBEST: Realtime Deformable Linear Object Detection Through Minimal Bending Energy Skeleton Pixel Traversals
by: Choi, Andrew, et al.
Published: (2023)
by: Choi, Andrew, et al.
Published: (2023)
Unstructured Moving Least Squares Material Point Methods: A Stable Kernel Approach With Continuous Gradient Reconstruction on General Unstructured Tessellations
by: Cao, Yadi, et al.
Published: (2023)
by: Cao, Yadi, et al.
Published: (2023)
Cross-Slice Attention and Evidential Critical Loss for Uncertainty-Aware Prostate Cancer Detection
by: Hung, Alex Ling Yu, et al.
Published: (2024)
by: Hung, Alex Ling Yu, et al.
Published: (2024)
Towards Fine-Grained Video Question Answering
by: Dai, Wei, et al.
Published: (2025)
by: Dai, Wei, et al.
Published: (2025)
Few-Shot Classification of Interactive Activities of Daily Living (InteractADL)
by: Durante, Zane, et al.
Published: (2024)
by: Durante, Zane, et al.
Published: (2024)
On Approximating the Weighted Region Problem in Square Tessellations
by: Kakimura, Naonori, et al.
Published: (2024)
by: Kakimura, Naonori, et al.
Published: (2024)
Pneumatic bladder links with wide range of motion joints for articulated inflatable robots
by: Uchiyama, Katsu, et al.
Published: (2025)
by: Uchiyama, Katsu, et al.
Published: (2025)
Review of Deep Learning Applications to Structural Proteomics Enabled by Cryogenic Electron Microscopy and Tomography
by: Zhou, Brady K., et al.
Published: (2025)
by: Zhou, Brady K., et al.
Published: (2025)
AnimaMimic: Imitating 3D Animation from Video Priors
by: Xie, Tianyi, et al.
Published: (2025)
by: Xie, Tianyi, et al.
Published: (2025)
CFSum: A Transformer-Based Multi-Modal Video Summarization Framework With Coarse-Fine Fusion
by: Guo, Yaowei, et al.
Published: (2025)
by: Guo, Yaowei, et al.
Published: (2025)
A back-linked Fabry-Perot interferometer for space-borne gravitational wave observations
by: Izumi, Kiwamu, et al.
Published: (2020)
by: Izumi, Kiwamu, et al.
Published: (2020)
Impact Analysis of Holistic Vibration on Cellular Adaptation
by: Rohan Vashisht, et al.
Published: (2018)
by: Rohan Vashisht, et al.
Published: (2018)
Simple Role Assignment is Extraordinarily Effective for Safety Alignment
by: Ziheng, Zhou, et al.
Published: (2026)
by: Ziheng, Zhou, et al.
Published: (2026)
MPM Lite: Linear Kernels and Integration without Particles
by: Feng, Xiang, et al.
Published: (2026)
by: Feng, Xiang, et al.
Published: (2026)
HourVideo: 1-Hour Video-Language Understanding
by: Chandrasegaran, Keshigeyan, et al.
Published: (2024)
by: Chandrasegaran, Keshigeyan, et al.
Published: (2024)
Wonderland: Navigating 3D Scenes from a Single Image
by: Liang, Hanwen, et al.
Published: (2024)
by: Liang, Hanwen, et al.
Published: (2024)
Task Value, Teacher Enthusiasm, and Student Engagement in Online Second Language Learning: A Latent Moderated Model
by: Hoi Vo, et al.
Published: (2026)
by: Hoi Vo, et al.
Published: (2026)
Lysine-Cysteine-Serine-Tryptophan Inserted into the DNA-Binding Domain of Human Mineralocorticoid Receptor Increases Transcriptional Activation by Aldosterone
by: Katsu, Yoshinao, et al.
Published: (2024)
by: Katsu, Yoshinao, et al.
Published: (2024)
Physically Grounded Vision-Language Models for Robotic Manipulation
by: Gao, Jensen, et al.
Published: (2023)
by: Gao, Jensen, et al.
Published: (2023)
Similar Items
-
An Interactive Agent Foundation Model
by: Durante, Zane, et al.
Published: (2024) -
Agent AI: Surveying the Horizons of Multimodal Interaction
by: Durante, Zane, et al.
Published: (2024) -
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering
by: Guo, Danfeng, et al.
Published: (2024) -
Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
by: Sasabuchi, Kazuhiro, et al.
Published: (2025) -
RL-Driven Data Generation for Robust Vision-Based Dexterous Grasping
by: Kanehira, Atsushi, et al.
Published: (2025)