Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Yafei, Xie, Quanting, Jain, Vidhi, Francis, Jonathan, Patrikar, Jay, Keetha, Nikhil, Kim, Seungchan, Xie, Yaqi, Zhang, Tianyi, Fang, Hao-Shu, Zhao, Shibo, Omidshafiei, Shayegan, Kim, Dong-Ki, Agha-mohammadi, Ali-akbar, Sycara, Katia, Johnson-Roberson, Matthew, Batra, Dhruv, Wang, Xiaolong, Scherer, Sebastian, Wang, Chen, Kira, Zsolt, Xia, Fei, Bisk, Yonatan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GrndCtrl: Grounding World Models via Self-Supervised Reward Alignment
by: He, Haoyang, et al.
Published: (2025)
by: He, Haoyang, et al.
Published: (2025)
Delay-Aware Diffusion Policy: Bridging the Observation-Execution Gap in Dynamic Tasks
by: Liao, Aileen, et al.
Published: (2025)
by: Liao, Aileen, et al.
Published: (2025)
StageACT: Stage-Conditioned Imitation for Robust Humanoid Door Opening
by: Lee, Moonyoung, et al.
Published: (2025)
by: Lee, Moonyoung, et al.
Published: (2025)
Reasoning about the Unseen for Efficient Outdoor Object Navigation
by: Xie, Quanting, et al.
Published: (2023)
by: Xie, Quanting, et al.
Published: (2023)
Co-Me: Confidence-Guided Token Merging for Visual Geometric Transformers
by: Chen, Yutian, et al.
Published: (2025)
by: Chen, Yutian, et al.
Published: (2025)
SayComply: Grounding Field Robotic Tasks in Operational Compliance through Retrieval-Based Language Models
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
RAVEN: Resilient Aerial Navigation via Open-Set Semantic Memory and Behavior Adaptation
by: Kim, Seungchan, et al.
Published: (2025)
by: Kim, Seungchan, et al.
Published: (2025)
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
by: Korekata, Ryosuke, et al.
Published: (2025)
by: Korekata, Ryosuke, et al.
Published: (2025)
Staircase Localization for Autonomous Exploration in Urban Environments
by: Kim, Jinrae, et al.
Published: (2024)
by: Kim, Jinrae, et al.
Published: (2024)
FALCON: Learning Force-Adaptive Humanoid Loco-Manipulation
by: Zhang, Yuanhang, et al.
Published: (2025)
by: Zhang, Yuanhang, et al.
Published: (2025)
ANAVI: Audio Noise Awareness using Visuals of Indoor environments for NAVIgation
by: Jain, Vidhi, et al.
Published: (2024)
by: Jain, Vidhi, et al.
Published: (2024)
Don't Run with Scissors: Pruning Breaks VLA Models but They Can Be Recovered
by: Jabbour, Jason, et al.
Published: (2025)
by: Jabbour, Jason, et al.
Published: (2025)
MM-SeR: Multimodal Self-Refinement for Lightweight Image Captioning
by: Song, Junha, et al.
Published: (2025)
by: Song, Junha, et al.
Published: (2025)
World Model Failure Classification and Anomaly Detection for Autonomous Inspection
by: Ho, Michelle, et al.
Published: (2026)
by: Ho, Michelle, et al.
Published: (2026)
Capability-aware Task Allocation and Team Formation Analysis for Cooperative Exploration of Complex Environments
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
FRAME: A Modular Framework for Autonomous Map Merging: Advancements in the Field
by: Stathoulopoulos, Nikolaos, et al.
Published: (2024)
by: Stathoulopoulos, Nikolaos, et al.
Published: (2024)
Enter the Mind Palace: Reasoning and Planning for Long-term Active Embodied Question Answering
by: Ginting, Muhammad Fadhil, et al.
Published: (2025)
by: Ginting, Muhammad Fadhil, et al.
Published: (2025)
Semantic Belief Behavior Graph: Enabling Autonomous Robot Inspection in Unknown Environments
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
by: Ginting, Muhammad Fadhil, et al.
Published: (2024)
Embodied-RAG: General Non-parametric Embodied Memory for Retrieval and Generation
by: Xie, Quanting, et al.
Published: (2024)
by: Xie, Quanting, et al.
Published: (2024)
Collision Avoidance Verification of Multiagent Systems with Learned Policies
by: Dong, Zihao, et al.
Published: (2024)
by: Dong, Zihao, et al.
Published: (2024)
Unsupervised Discovery of Long-Term Spatiotemporal Periodic Workflows in Human Activities
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
Jailbreaking Frontier Foundation Models Through Intention Deception
by: Wang, Xinhe, et al.
Published: (2026)
by: Wang, Xinhe, et al.
Published: (2026)
High Torque Density PCB Axial Flux Permanent Magnet Motor for Micro Robots
by: Wang, Jianren, et al.
Published: (2025)
by: Wang, Jianren, et al.
Published: (2025)
VENTURA: Adapting Image Diffusion Models for Unified Task Conditioned Navigation
by: Zhang, Arthur, et al.
Published: (2025)
by: Zhang, Arthur, et al.
Published: (2025)
ELLIPSE: Evidential Learning for Robust Waypoints and Uncertainties
by: Dong, Zihao, et al.
Published: (2026)
by: Dong, Zihao, et al.
Published: (2026)
Low Frequency Sampling in Model Predictive Path Integral Control
by: Vlahov, Bogdan, et al.
Published: (2024)
by: Vlahov, Bogdan, et al.
Published: (2024)
RayFronts: Open-Set Semantic Ray Frontiers for Online Scene Understanding and Exploration
by: Alama, Omar, et al.
Published: (2025)
by: Alama, Omar, et al.
Published: (2025)
Learning Model Successors
by: Chang, Yingshan, et al.
Published: (2025)
by: Chang, Yingshan, et al.
Published: (2025)
Language Models Need Inductive Biases to Count Inductively
by: Chang, Yingshan, et al.
Published: (2024)
by: Chang, Yingshan, et al.
Published: (2024)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
GL-NeRF: Gauss-Laguerre Quadrature Enables Training-Free NeRF Acceleration
by: Yong, Silong, et al.
Published: (2024)
by: Yong, Silong, et al.
Published: (2024)
Seeing the Unseen: Visual Common Sense for Semantic Placement
by: Ramrakhya, Ram, et al.
Published: (2024)
by: Ramrakhya, Ram, et al.
Published: (2024)
Risk-aware Meta-level Decision Making for Exploration Under Uncertainty
by: Ott, Joshua, et al.
Published: (2022)
by: Ott, Joshua, et al.
Published: (2022)
ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
by: Elawady, Ahmad, et al.
Published: (2024)
by: Elawady, Ahmad, et al.
Published: (2024)
World2Rules: A Neuro-Symbolic Framework for Learning World-Governing Safety Rules for Aviation
by: Wang, Haichuan, et al.
Published: (2026)
by: Wang, Haichuan, et al.
Published: (2026)
AutoODD: Agentic Audits via Bayesian Red Teaming in Black-Box Models
by: Martin, Rebecca, et al.
Published: (2025)
by: Martin, Rebecca, et al.
Published: (2025)
MapEx: Indoor Structure Exploration with Probabilistic Information Gain from Global Map Predictions
by: Ho, Cherie, et al.
Published: (2024)
by: Ho, Cherie, et al.
Published: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
by: Yenamandra, Sriram, et al.
Published: (2023)
by: Yenamandra, Sriram, et al.
Published: (2023)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
by: Li, Samuel, et al.
Published: (2024)
by: Li, Samuel, et al.
Published: (2024)
Similar Items
-
GrndCtrl: Grounding World Models via Self-Supervised Reward Alignment
by: He, Haoyang, et al.
Published: (2025) -
Delay-Aware Diffusion Policy: Bridging the Observation-Execution Gap in Dynamic Tasks
by: Liao, Aileen, et al.
Published: (2025) -
StageACT: Stage-Conditioned Imitation for Robust Humanoid Door Opening
by: Lee, Moonyoung, et al.
Published: (2025) -
Reasoning about the Unseen for Efficient Outdoor Object Navigation
by: Xie, Quanting, et al.
Published: (2023) -
Co-Me: Confidence-Guided Token Merging for Visual Geometric Transformers
by: Chen, Yutian, et al.
Published: (2025)