Virtual avatar generation models as world navigators
Fuente:
arXiv
Guardado en:
| Autor principal: | Mandava, Sai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition
por: Luo, Shengcheng, et al.
Publicado: (2024)
por: Luo, Shengcheng, et al.
Publicado: (2024)
An Approach to Systematic Data Acquisition and Data-Driven Simulation for the Safety Testing of Automated Driving Functions
por: Eisemann, Leon, et al.
Publicado: (2024)
por: Eisemann, Leon, et al.
Publicado: (2024)
ShelfHelp: Empowering Humans to Perform Vision-Independent Manipulation Tasks with a Socially Assistive Robotic Cane
por: Agrawal, Shivendra, et al.
Publicado: (2024)
por: Agrawal, Shivendra, et al.
Publicado: (2024)
Intelligent Control of Robotic X-ray Devices using a Language-promptable Digital Twin
por: Killeen, Benjamin D., et al.
Publicado: (2024)
por: Killeen, Benjamin D., et al.
Publicado: (2024)
RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
por: Garg, Sourav, et al.
Publicado: (2024)
por: Garg, Sourav, et al.
Publicado: (2024)
Explainable AI for Safe and Trustworthy Autonomous Driving: A Systematic Review
por: Kuznietsov, Anton, et al.
Publicado: (2024)
por: Kuznietsov, Anton, et al.
Publicado: (2024)
Monocular 3D Object Position Estimation with VLMs for Human-Robot Interaction
por: Wahl, Ari, et al.
Publicado: (2026)
por: Wahl, Ari, et al.
Publicado: (2026)
HAPI: A Model for Learning Robot Facial Expressions from Human Preferences
por: Yang, Dongsheng, et al.
Publicado: (2025)
por: Yang, Dongsheng, et al.
Publicado: (2025)
ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement
por: Su, Chenyu, et al.
Publicado: (2025)
por: Su, Chenyu, et al.
Publicado: (2025)
HOSt3R: Keypoint-free Hand-Object 3D Reconstruction from RGB images
por: Swamy, Anilkumar, et al.
Publicado: (2025)
por: Swamy, Anilkumar, et al.
Publicado: (2025)
Aria Gen 2 Pilot Dataset
por: Kong, Chen, et al.
Publicado: (2025)
por: Kong, Chen, et al.
Publicado: (2025)
AI Guide Dog: Egocentric Path Prediction on Smartphone
por: Jadhav, Aishwarya, et al.
Publicado: (2025)
por: Jadhav, Aishwarya, et al.
Publicado: (2025)
Magma: A Foundation Model for Multimodal AI Agents
por: Yang, Jianwei, et al.
Publicado: (2025)
por: Yang, Jianwei, et al.
Publicado: (2025)
Mind2Drive: Predicting Driver Intentions from EEG in Real-world On-Road Driving
por: Alosaimi, Ghadah, et al.
Publicado: (2026)
por: Alosaimi, Ghadah, et al.
Publicado: (2026)
Predicting User Grasp Intentions in Virtual Reality
por: Zeng, Linghao
Publicado: (2025)
por: Zeng, Linghao
Publicado: (2025)
Logic-Free Building Automation: Learning the Control of Room Facilities with Wall Switches and Ceiling Camera
por: Ochiai, Hideya, et al.
Publicado: (2024)
por: Ochiai, Hideya, et al.
Publicado: (2024)
SurgBox: Agent-Driven Operating Room Sandbox with Surgery Copilot
por: Wu, Jinlin, et al.
Publicado: (2024)
por: Wu, Jinlin, et al.
Publicado: (2024)
GarmentLab: A Unified Simulation and Benchmark for Garment Manipulation
por: Lu, Haoran, et al.
Publicado: (2024)
por: Lu, Haoran, et al.
Publicado: (2024)
A Backbone for Long-Horizon Robot Task Understanding
por: Chen, Xiaoshuai, et al.
Publicado: (2024)
por: Chen, Xiaoshuai, et al.
Publicado: (2024)
Creativity and Visual Communication from Machine to Musician: Sharing a Score through a Robotic Camera
por: Greer, Ross, et al.
Publicado: (2024)
por: Greer, Ross, et al.
Publicado: (2024)
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
por: Han, Kyungtae, et al.
Publicado: (2025)
por: Han, Kyungtae, et al.
Publicado: (2025)
CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence
por: Zeng, Tianle, et al.
Publicado: (2026)
por: Zeng, Tianle, et al.
Publicado: (2026)
iLearnRobot: An Interactive Learning-Based Multi-Modal Robot with Continuous Improvement
por: Wang, Kohou, et al.
Publicado: (2025)
por: Wang, Kohou, et al.
Publicado: (2025)
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
por: Jha, Saurav, et al.
Publicado: (2025)
por: Jha, Saurav, et al.
Publicado: (2025)
Gaze Detection and Analysis for Initiating Joint Activity in Industrial Human-Robot Collaboration
por: Prajod, Pooja, et al.
Publicado: (2023)
por: Prajod, Pooja, et al.
Publicado: (2023)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
por: Yakolli, Nivedan, et al.
Publicado: (2025)
por: Yakolli, Nivedan, et al.
Publicado: (2025)
VLM-driven Behavior Tree for Context-aware Task Planning
por: Wake, Naoki, et al.
Publicado: (2025)
por: Wake, Naoki, et al.
Publicado: (2025)
Multi-face emotion detection for effective Human-Robot Interaction
por: Yahyaoui, Mohamed Ala, et al.
Publicado: (2025)
por: Yahyaoui, Mohamed Ala, et al.
Publicado: (2025)
Skeleton-Based Transformer for Classification of Errors and Better Feedback in Low Back Pain Physical Rehabilitation Exercises
por: Marusic, Aleksa, et al.
Publicado: (2025)
por: Marusic, Aleksa, et al.
Publicado: (2025)
ICPR 2024 Competition on Rider Intention Prediction
por: Gangisetty, Shankar, et al.
Publicado: (2025)
por: Gangisetty, Shankar, et al.
Publicado: (2025)
DEXOP: A Device for Robotic Transfer of Dexterous Human Manipulation
por: Fang, Hao-Shu, et al.
Publicado: (2025)
por: Fang, Hao-Shu, et al.
Publicado: (2025)
IndEgo: A Dataset of Industrial Scenarios and Collaborative Work for Egocentric Assistants
por: Chavan, Vivek, et al.
Publicado: (2025)
por: Chavan, Vivek, et al.
Publicado: (2025)
Chaining text-to-image and large language model: A novel approach for generating personalized e-commerce banners
por: Vashishtha, Shanu, et al.
Publicado: (2024)
por: Vashishtha, Shanu, et al.
Publicado: (2024)
Human-in-the-Loop Segmentation of Multi-species Coral Imagery
por: Raine, Scarlett, et al.
Publicado: (2024)
por: Raine, Scarlett, et al.
Publicado: (2024)
Benchmarking Adaptive Intelligence and Computer Vision on Human-Robot Collaboration
por: Saraj, Salaar, et al.
Publicado: (2024)
por: Saraj, Salaar, et al.
Publicado: (2024)
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
por: Sezer, Berk, et al.
Publicado: (2026)
por: Sezer, Berk, et al.
Publicado: (2026)
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
por: Jalayer, Reza, et al.
Publicado: (2025)
por: Jalayer, Reza, et al.
Publicado: (2025)
Visuo-Acoustic Hand Pose and Contact Estimation
por: Mao, Yuemin, et al.
Publicado: (2025)
por: Mao, Yuemin, et al.
Publicado: (2025)
Give me scissors: Collision-Free Dual-Arm Surgical Assistive Robot for Instrument Delivery
por: Luo, Xuejin, et al.
Publicado: (2026)
por: Luo, Xuejin, et al.
Publicado: (2026)
Toward a Surgeon-in-the-Loop Ophthalmic Robotic Apprentice using Reinforcement and Imitation Learning
por: Gomaa, Amr, et al.
Publicado: (2023)
por: Gomaa, Amr, et al.
Publicado: (2023)
Ejemplares similares
-
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition
por: Luo, Shengcheng, et al.
Publicado: (2024) -
An Approach to Systematic Data Acquisition and Data-Driven Simulation for the Safety Testing of Automated Driving Functions
por: Eisemann, Leon, et al.
Publicado: (2024) -
ShelfHelp: Empowering Humans to Perform Vision-Independent Manipulation Tasks with a Socially Assistive Robotic Cane
por: Agrawal, Shivendra, et al.
Publicado: (2024) -
Intelligent Control of Robotic X-ray Devices using a Language-promptable Digital Twin
por: Killeen, Benjamin D., et al.
Publicado: (2024) -
RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
por: Garg, Sourav, et al.
Publicado: (2024)