IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Lian, Shijie, Yu, Bin, Lin, Xiaopeng, Shen, Zhaolong, Yang, Laurence Tianruo, Jin, Yurun, Liu, Haishan, Wu, Changti, Yuan, Hang, Huang, Cong, Chen, Kai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Euclid's Gift: Enhancing Spatial Perception and Reasoning in Vision-Language Models via Geometric Surrogate Tasks
by: Lian, Shijie, et al.
Published: (2025)
by: Lian, Shijie, et al.
Published: (2025)
DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry
by: Wu, Changti, et al.
Published: (2025)
by: Wu, Changti, et al.
Published: (2025)
FrameSkip: Learning from Fewer but More Informative Frames in VLA Training
by: Yu, Bin, et al.
Published: (2026)
by: Yu, Bin, et al.
Published: (2026)
TwinBrainVLA: Unleashing the Potential of Generalist VLMs for Embodied Tasks via Asymmetric Mixture-of-Transformers
by: Yu, Bin, et al.
Published: (2026)
by: Yu, Bin, et al.
Published: (2026)
PhysBrain 1.0 Technical Report
by: Lian, Shijie, et al.
Published: (2026)
by: Lian, Shijie, et al.
Published: (2026)
3D-Mix for VLA: A Plug-and-Play Module for Integrating VGGT-based 3D Information into Vision-Language-Action Models
by: Yu, Bin, et al.
Published: (2026)
by: Yu, Bin, et al.
Published: (2026)
ScalSelect: Scalable Training-Free Multimodal Data Selection for Efficient Visual Instruction Tuning
by: Wu, Changti, et al.
Published: (2026)
by: Wu, Changti, et al.
Published: (2026)
Diving into Underwater: Segment Anything Model Guided Underwater Salient Instance Segmentation and A Large-scale Dataset
by: Lian, Shijie, et al.
Published: (2024)
by: Lian, Shijie, et al.
Published: (2024)
PhysBrain: Human Egocentric Data as a Bridge from Vision Language Models to Physical Intelligence
by: Lin, Xiaopeng, et al.
Published: (2025)
by: Lin, Xiaopeng, et al.
Published: (2025)
IntentCoding: Amplifying User Intent in Code Generation
by: Fang, Zheng, et al.
Published: (2026)
by: Fang, Zheng, et al.
Published: (2026)
Detecting Conversational Mental Manipulation with Intent-Aware Prompting
by: Ma, Jiayuan, et al.
Published: (2024)
by: Ma, Jiayuan, et al.
Published: (2024)
Reading with Intent -- Neutralizing Intent
by: Reichman, Benjamin, et al.
Published: (2025)
by: Reichman, Benjamin, et al.
Published: (2025)
Few-Shot Query Intent Detection via Relation-Aware Prompt Learning
by: Zhang, Liang, et al.
Published: (2025)
by: Zhang, Liang, et al.
Published: (2025)
TUGS: Physics-based Compact Representation of Underwater Scenes by Tensorized Gaussian
by: Lian, Shijie, et al.
Published: (2025)
by: Lian, Shijie, et al.
Published: (2025)
From Intents to Conversations: Generating Intent-Driven Dialogues with Contrastive Learning for Multi-Turn Classification
by: Liu, Junhua, et al.
Published: (2024)
by: Liu, Junhua, et al.
Published: (2024)
Conformal Intent Classification and Clarification for Fast and Accurate Intent Recognition
by: Hengst, Floris den, et al.
Published: (2024)
by: Hengst, Floris den, et al.
Published: (2024)
TrajSelector: Harnessing Latent Representations for Efficient and Effective Best-of-N in Large Reasoning Model
by: Yu, Bin, et al.
Published: (2025)
by: Yu, Bin, et al.
Published: (2025)
IntentGPT: Few-shot Intent Discovery with Large Language Models
by: Rodriguez, Juan A., et al.
Published: (2024)
by: Rodriguez, Juan A., et al.
Published: (2024)
Tracking with Human-Intent Reasoning
by: Zhu, Jiawen, et al.
Published: (2023)
by: Zhu, Jiawen, et al.
Published: (2023)
AnyUser: Translating Sketched User Intent into Domestic Robots
by: Yang, Songyuan, et al.
Published: (2026)
by: Yang, Songyuan, et al.
Published: (2026)
IntentFlow: Investigating Fluid Dynamics of Intent Communication in Generative AI
by: Kim, Yoonsu, et al.
Published: (2025)
by: Kim, Yoonsu, et al.
Published: (2025)
Exploring the Vulnerability of the Content Moderation Guardrail in Large Language Models via Intent Manipulation
by: Zhuang, Jun, et al.
Published: (2025)
by: Zhuang, Jun, et al.
Published: (2025)
Probabilistic Human Intent Prediction for Mobile Manipulation: An Evaluation with Human-Inspired Constraints
by: Contreras, Cesar Alan, et al.
Published: (2025)
by: Contreras, Cesar Alan, et al.
Published: (2025)
Intent Matters: Enhancing AI Tutoring with Fine-Grained Pedagogical Intent Annotation
by: Petukhova, Kseniia, et al.
Published: (2025)
by: Petukhova, Kseniia, et al.
Published: (2025)
Mirror Skin: In Situ Visualization of Robot Touch Intent on Robotic Skin
by: Wagmann, David, et al.
Published: (2025)
by: Wagmann, David, et al.
Published: (2025)
Intent-Driven UAM Rescheduling
by: Kim, Jeongseok, et al.
Published: (2025)
by: Kim, Jeongseok, et al.
Published: (2025)
Known Intents, New Combinations: Clause-Factorized Decoding for Compositional Multi-Intent Detection
by: Nandy, Abhilash
Published: (2026)
by: Nandy, Abhilash
Published: (2026)
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
by: Chen, Yi, et al.
Published: (2026)
by: Chen, Yi, et al.
Published: (2026)
Distill: Uncovering the True Intent behind Human-Robot Communication
by: Li, Ting, et al.
Published: (2026)
by: Li, Ting, et al.
Published: (2026)
Confidence-based Intent Prediction for Teleoperation in Bimanual Robotic Suturing
by: Hu, Zhaoyang Jacopo, et al.
Published: (2025)
by: Hu, Zhaoyang Jacopo, et al.
Published: (2025)
IntentGrasp: A Comprehensive Benchmark for Intent Understanding
by: Yin, Yuwei, et al.
Published: (2026)
by: Yin, Yuwei, et al.
Published: (2026)
Intent Detection in the Age of LLMs
by: Arora, Gaurav, et al.
Published: (2024)
by: Arora, Gaurav, et al.
Published: (2024)
IntentContinuum: Using LLMs to Support Intent-Based Computing Across the Compute Continuum
by: Akbari, Negin, et al.
Published: (2025)
by: Akbari, Negin, et al.
Published: (2025)
Brickify: Enabling Expressive Design Intent Specification through Direct Manipulation on Design Tokens
by: Shi, Xinyu, et al.
Published: (2025)
by: Shi, Xinyu, et al.
Published: (2025)
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
by: Zhang, Shiduo, et al.
Published: (2024)
by: Zhang, Shiduo, et al.
Published: (2024)
An Interdisciplinary Review of Commonsense Reasoning and Intent Detection
by: Sakib, Md Nazmus
Published: (2025)
by: Sakib, Md Nazmus
Published: (2025)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
U-Fold: Dynamic Intent-Aware Context Folding for User-Centric Agents
by: Su, Jin, et al.
Published: (2026)
by: Su, Jin, et al.
Published: (2026)
BloomIntent: Automating Search Evaluation with LLM-Generated Fine-Grained User Intents
by: Choi, Yoonseo, et al.
Published: (2025)
by: Choi, Yoonseo, et al.
Published: (2025)
An Analysis of Intent-Based Markets
by: Chitra, Tarun, et al.
Published: (2024)
by: Chitra, Tarun, et al.
Published: (2024)
Similar Items
-
Euclid's Gift: Enhancing Spatial Perception and Reasoning in Vision-Language Models via Geometric Surrogate Tasks
by: Lian, Shijie, et al.
Published: (2025) -
DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry
by: Wu, Changti, et al.
Published: (2025) -
FrameSkip: Learning from Fewer but More Informative Frames in VLA Training
by: Yu, Bin, et al.
Published: (2026) -
TwinBrainVLA: Unleashing the Potential of Generalist VLMs for Embodied Tasks via Asymmetric Mixture-of-Transformers
by: Yu, Bin, et al.
Published: (2026) -
PhysBrain 1.0 Technical Report
by: Lian, Shijie, et al.
Published: (2026)