Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?
Fuente:
arXiv
Salvato in:
| Autori principali: | Majumdar, Arjun, Yadav, Karmesh, Arnaud, Sergio, Ma, Yecheng Jason, Chen, Claire, Silwal, Sneha, Jain, Aryan, Berges, Vincent-Pierre, Abbeel, Pieter, Malik, Jitendra, Batra, Dhruv, Lin, Yixin, Maksymets, Oleksandr, Rajeswaran, Aravind, Meier, Franziska |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
di: Silwal, Sneha, et al.
Pubblicazione: (2023)
di: Silwal, Sneha, et al.
Pubblicazione: (2023)
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
di: Shentu, Yide, et al.
Pubblicazione: (2024)
di: Shentu, Yide, et al.
Pubblicazione: (2024)
Semi-Supervised One-Shot Imitation Learning
di: Wu, Philipp, et al.
Pubblicazione: (2024)
di: Wu, Philipp, et al.
Pubblicazione: (2024)
Interactive Task Planning with Language Models
di: Li, Boyi, et al.
Pubblicazione: (2023)
di: Li, Boyi, et al.
Pubblicazione: (2023)
From Thousands to Billions: 3D Visual Language Grounding via Render-Supervised Distillation from 2D VLMs
di: Cao, Ang, et al.
Pubblicazione: (2025)
di: Cao, Ang, et al.
Pubblicazione: (2025)
ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
di: Elawady, Ahmad, et al.
Pubblicazione: (2024)
di: Elawady, Ahmad, et al.
Pubblicazione: (2024)
Deep Sensorimotor Control by Imitating Predictive Models of Human Motion
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2025)
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2025)
ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation
di: Heng, Liang, et al.
Pubblicazione: (2025)
di: Heng, Liang, et al.
Pubblicazione: (2025)
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
di: Arnaud, Sergio, et al.
Pubblicazione: (2025)
di: Arnaud, Sergio, et al.
Pubblicazione: (2025)
How to Peel with a Knife: Aligning Fine-Grained Manipulation with Human Preference
di: Lin, Toru, et al.
Pubblicazione: (2026)
di: Lin, Toru, et al.
Pubblicazione: (2026)
Twisting Lids Off with Two Hands
di: Lin, Toru, et al.
Pubblicazione: (2024)
di: Lin, Toru, et al.
Pubblicazione: (2024)
End-to-end RL Improves Dexterous Grasping Policies
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
di: Singh, Ritvik, et al.
Pubblicazione: (2025)
Rodrigues Network for Learning Robot Actions
di: Zhang, Jialiang, et al.
Pubblicazione: (2025)
di: Zhang, Jialiang, et al.
Pubblicazione: (2025)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
di: Lancaster, Patrick, et al.
Pubblicazione: (2023)
di: Lancaster, Patrick, et al.
Pubblicazione: (2023)
Hand-Object Interaction Pretraining from Videos
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
di: Singh, Himanshu Gaurav, et al.
Pubblicazione: (2024)
Animal models for abdominal aortic aneurysms: Where we are and where we need to go
di: Kangli Tian, et al.
Pubblicazione: (2025)
di: Kangli Tian, et al.
Pubblicazione: (2025)
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control
di: Gupta, Gunshi, et al.
Pubblicazione: (2024)
di: Gupta, Gunshi, et al.
Pubblicazione: (2024)
OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction
di: Huang, Huang, et al.
Pubblicazione: (2025)
di: Huang, Huang, et al.
Pubblicazione: (2025)
DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization
di: Tang, Yikai, et al.
Pubblicazione: (2025)
di: Tang, Yikai, et al.
Pubblicazione: (2025)
Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning
di: Cheng, Ziheng, et al.
Pubblicazione: (2026)
di: Cheng, Ziheng, et al.
Pubblicazione: (2026)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
di: Seo, Younggyo, et al.
Pubblicazione: (2024)
di: Seo, Younggyo, et al.
Pubblicazione: (2024)
Embodied AI Agents: Modeling the World
di: Fung, Pascale, et al.
Pubblicazione: (2025)
di: Fung, Pascale, et al.
Pubblicazione: (2025)
SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending
di: Kuang, Yuxuan, et al.
Pubblicazione: (2025)
di: Kuang, Yuxuan, et al.
Pubblicazione: (2025)
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
di: Gupta, Gunshi, et al.
Pubblicazione: (2025)
di: Gupta, Gunshi, et al.
Pubblicazione: (2025)
FindingDory: A Benchmark to Evaluate Memory in Embodied Agents
di: Yadav, Karmesh, et al.
Pubblicazione: (2025)
di: Yadav, Karmesh, et al.
Pubblicazione: (2025)
Neuroinflammation in Alzheimer's Disease: Mechanisms, Impact and Emerging Therapies
di: Aryan Malik
Pubblicazione: (2025)
di: Aryan Malik
Pubblicazione: (2025)
Lightning Grasp: High Performance Procedural Grasp Synthesis with Contact Fields
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2025)
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2025)
Offline Imitation Learning Through Graph Search and Retrieval
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2024)
di: Yin, Zhao-Heng, et al.
Pubblicazione: (2024)
SURFACE TENSIONS : Roads, Potholes and the Embodied Politics of Driving in Urban India
di: Sneha Annavarapu
Pubblicazione: (2025)
di: Sneha Annavarapu
Pubblicazione: (2025)
Where we are, where we are going
di: Teresa Bermejo-Vicedo
Pubblicazione: (2020)
di: Teresa Bermejo-Vicedo
Pubblicazione: (2020)
DexGarmentLab: Dexterous Garment Manipulation Environment with Generalizable Policy
di: Wang, Yuran, et al.
Pubblicazione: (2025)
di: Wang, Yuran, et al.
Pubblicazione: (2025)
A Stable Whitening Optimizer for Efficient Neural Network Training
di: Frans, Kevin, et al.
Pubblicazione: (2025)
di: Frans, Kevin, et al.
Pubblicazione: (2025)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
di: Lee, Vint, et al.
Pubblicazione: (2023)
di: Lee, Vint, et al.
Pubblicazione: (2023)
Reward-Conditioned Reinforcement Learning
di: Nauman, Michal, et al.
Pubblicazione: (2026)
di: Nauman, Michal, et al.
Pubblicazione: (2026)
What Really Matters in Matrix-Whitening Optimizers?
di: Frans, Kevin, et al.
Pubblicazione: (2025)
di: Frans, Kevin, et al.
Pubblicazione: (2025)
SEMDICE: Off-policy State Entropy Maximization via Stationary Distribution Correction Estimation
di: Lee, Jongmin, et al.
Pubblicazione: (2025)
di: Lee, Jongmin, et al.
Pubblicazione: (2025)
Where have all the platelets gone? A simple solution to the suboptimal performance of PRP tubes that contain a thixotropic gel separator
di: Jaya Krishna Rose Batra, et al.
Pubblicazione: (2024)
di: Jaya Krishna Rose Batra, et al.
Pubblicazione: (2024)
The Sound of Simulation: Learning Multimodal Sim-to-Real Robot Policies with Generative Audio
di: Wang, Renhao, et al.
Pubblicazione: (2025)
di: Wang, Renhao, et al.
Pubblicazione: (2025)
Introduction - Language Planning: Where have we been? Where might we be going?
di: Richard B. Baldauf Jr.
Pubblicazione: (2012)
di: Richard B. Baldauf Jr.
Pubblicazione: (2012)
Hematological cytomorphology: Where we are
di: G. Zini
Pubblicazione: (2024)
di: G. Zini
Pubblicazione: (2024)
Documenti analoghi
-
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
di: Silwal, Sneha, et al.
Pubblicazione: (2023) -
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
di: Shentu, Yide, et al.
Pubblicazione: (2024) -
Semi-Supervised One-Shot Imitation Learning
di: Wu, Philipp, et al.
Pubblicazione: (2024) -
Interactive Task Planning with Language Models
di: Li, Boyi, et al.
Pubblicazione: (2023) -
From Thousands to Billions: 3D Visual Language Grounding via Render-Supervised Distillation from 2D VLMs
di: Cao, Ang, et al.
Pubblicazione: (2025)