VLH: Vision-Language-Haptics Foundation Model
Fuente:
arXiv
Saved in:
| Main Authors: | Fuentes, Luis Francisco Moreno, Khan, Muhammad Haris, Cabrera, Miguel Altamirano, Serpiva, Valerii, Iarchuk, Dmitri, Mahmoud, Yara, Tokmurziyev, Issatay, Tsetserukou, Dzmitry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HapticVLM: VLM-Driven Texture Recognition Aimed at Intelligent Haptic Interaction
by: Khan, Muhammad Haris, et al.
Published: (2025)
by: Khan, Muhammad Haris, et al.
Published: (2025)
LLM-Glasses: GenAI-driven Glasses with Haptic Feedback for Navigation of Visually Impaired People
by: Tokmurziyev, Issatay, et al.
Published: (2025)
by: Tokmurziyev, Issatay, et al.
Published: (2025)
FlyHaptics: Flying Multi-contact Haptic Interface
by: Moreno, Luis, et al.
Published: (2025)
by: Moreno, Luis, et al.
Published: (2025)
GazeRace: Revolutionizing Remote Piloting with Eye-Gaze Control
by: Tokmurziyev, Issatay, et al.
Published: (2024)
by: Tokmurziyev, Issatay, et al.
Published: (2024)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
by: Khan, Muhamamd Haris, et al.
Published: (2025)
by: Khan, Muhamamd Haris, et al.
Published: (2025)
GazeGrasp: DNN-Driven Robotic Grasping with Wearable Eye-Gaze Interface
by: Tokmurziyev, Issatay, et al.
Published: (2025)
by: Tokmurziyev, Issatay, et al.
Published: (2025)
HumanDiffusion: A Vision-Based Diffusion Trajectory Planner with Human-Conditioned Goals for Search and Rescue UAV
by: Batool, Faryal, et al.
Published: (2026)
by: Batool, Faryal, et al.
Published: (2026)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
by: Mahmoud, Yara, et al.
Published: (2025)
by: Mahmoud, Yara, et al.
Published: (2025)
Musinger: Communication of Music over a Distance with Wearable Haptic Display and Touch Sensitive Surface
by: Cabrera, Miguel Altamirano, et al.
Published: (2024)
by: Cabrera, Miguel Altamirano, et al.
Published: (2024)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
by: Mahmoud, Yara, et al.
Published: (2026)
by: Mahmoud, Yara, et al.
Published: (2026)
GuideTouch: An Obstacle Avoidance Device with Tactile Feedback for Visually Impaired
by: Kozlov, Timofei, et al.
Published: (2026)
by: Kozlov, Timofei, et al.
Published: (2026)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
by: Lykov, Artem, et al.
Published: (2025)
by: Lykov, Artem, et al.
Published: (2025)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
by: Serpiva, Valerii, et al.
Published: (2025)
by: Serpiva, Valerii, et al.
Published: (2025)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
by: Yaqoot, Yasheerah, et al.
Published: (2025)
by: Yaqoot, Yasheerah, et al.
Published: (2025)
HoverAI: An Embodied Aerial Agent for Natural Human-Drone Interaction
by: Jin, Yuhua, et al.
Published: (2026)
by: Jin, Yuhua, et al.
Published: (2026)
AgilePilot: DRL-Based Drone Agent for Real-Time Motion Planning in Dynamic Environments by Leveraging Object Detection
by: Khan, Roohan Ahmed, et al.
Published: (2025)
by: Khan, Roohan Ahmed, et al.
Published: (2025)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
by: Serpiva, Valerii, et al.
Published: (2025)
by: Serpiva, Valerii, et al.
Published: (2025)
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
by: Lykov, Artem, et al.
Published: (2025)
by: Lykov, Artem, et al.
Published: (2025)
Evolution 6.0: Robot Evolution through Generative Design
by: Khan, Muhammad Haris, et al.
Published: (2025)
by: Khan, Muhammad Haris, et al.
Published: (2025)
DiffusionCinema: Text-to-Aerial Cinematography
by: Serpiva, Valerii, et al.
Published: (2026)
by: Serpiva, Valerii, et al.
Published: (2026)
Action Agent: Agentic Video Generation Meets Flow-Constrained Diffusion
by: Sam, Jeffrin, et al.
Published: (2026)
by: Sam, Jeffrin, et al.
Published: (2026)
Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
by: Serpiva, Valerii, et al.
Published: (2024)
by: Serpiva, Valerii, et al.
Published: (2024)
GenerativeMPC: VLM-RAG-guided Whole-Body MPC with Virtual Impedance for Bimanual Mobile Manipulation
by: Fernando, Marcelino Julio, et al.
Published: (2026)
by: Fernando, Marcelino Julio, et al.
Published: (2026)
DiffusionAnything: End-to-End In-context Diffusion Learning for Unified Navigation and Pre-Grasp Motion
by: Zhura, Iana, et al.
Published: (2026)
by: Zhura, Iana, et al.
Published: (2026)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
by: Serpiva, Valerii, et al.
Published: (2026)
by: Serpiva, Valerii, et al.
Published: (2026)
FiDTouch: A 3D Wearable Haptic Display for the Finger Pad
by: Trinitatova, Daria, et al.
Published: (2025)
by: Trinitatova, Daria, et al.
Published: (2025)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
by: Gbagbe, Koffivi Fidèle, et al.
Published: (2024)
by: Gbagbe, Koffivi Fidèle, et al.
Published: (2024)
DeepXPalm: Tilt and Position Rendering using Palm-worn Haptic Display and CNN-based Tactile Pattern Recognition
by: Miguel, Altamirano Cabrera, et al.
Published: (2022)
by: Miguel, Altamirano Cabrera, et al.
Published: (2022)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
Hybrid F' and ROS2 Architecture for Vision-Based Autonomous Flight: Design and Experimental Validation
by: Metwally, Abdelrahman, et al.
Published: (2026)
by: Metwally, Abdelrahman, et al.
Published: (2026)
SafeSwarm: Decentralized Safe RL for the Swarm of Drones Landing in Dense Crowds
by: Tadevosyan, Grik, et al.
Published: (2025)
by: Tadevosyan, Grik, et al.
Published: (2025)
AttentionSwarm: Reinforcement Learning with Attention Control Barier Function for Crazyflie Drones in Dynamic Environments
by: Tadevosyan, Grik, et al.
Published: (2025)
by: Tadevosyan, Grik, et al.
Published: (2025)
GraspSense: Physically Grounded Grasp and Grip Planning for a Dexterous Robotic Hand via Language-Guided Perception and Force Maps
by: Semenyakina, Elizaveta, et al.
Published: (2026)
by: Semenyakina, Elizaveta, et al.
Published: (2026)
Prosthetic Hand Manipulation System Based on EMG and Eye Tracking Powered by the Neuromorphic Processor AltAi
by: Akinshin, Roman, et al.
Published: (2026)
by: Akinshin, Roman, et al.
Published: (2026)
DroneVLA: VLA based Aerial Manipulation
by: Mehboob, Fawad, et al.
Published: (2026)
by: Mehboob, Fawad, et al.
Published: (2026)
ImpedanceGPT: VLM-driven Impedance Control of Swarm of Mini-drones for Intelligent Navigation in Dynamic Environment
by: Batool, Faryal, et al.
Published: (2025)
by: Batool, Faryal, et al.
Published: (2025)
Glove2UAV: A Wearable IMU-Based Glove for Intuitive Control of UAV
by: Habel, Amir, et al.
Published: (2026)
by: Habel, Amir, et al.
Published: (2026)
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
by: Gubernatorov, Konstantin, et al.
Published: (2026)
by: Gubernatorov, Konstantin, et al.
Published: (2026)
Similar Items
-
HapticVLM: VLM-Driven Texture Recognition Aimed at Intelligent Haptic Interaction
by: Khan, Muhammad Haris, et al.
Published: (2025) -
LLM-Glasses: GenAI-driven Glasses with Haptic Feedback for Navigation of Visually Impaired People
by: Tokmurziyev, Issatay, et al.
Published: (2025) -
FlyHaptics: Flying Multi-contact Haptic Interface
by: Moreno, Luis, et al.
Published: (2025) -
GazeRace: Revolutionizing Remote Piloting with Eye-Gaze Control
by: Tokmurziyev, Issatay, et al.
Published: (2024) -
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
by: Khan, Muhamamd Haris, et al.
Published: (2025)