Exploring 3D Reasoning-Driven Planning: From Implicit Human Intentions to Route-Aware Activity Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Xueying, Li, Wenhao, Zhang, Xiaoqin, Shao, Ling, Lu, Shijian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
Multimodal 3D Reasoning Segmentation with Complex Scenes
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction
von: Qian, Lin, et al.
Veröffentlicht: (2026)
von: Qian, Lin, et al.
Veröffentlicht: (2026)
L3DR: 3D-aware LiDAR Diffusion and Rectification
von: Liu, Quan, et al.
Veröffentlicht: (2026)
von: Liu, Quan, et al.
Veröffentlicht: (2026)
Weakly Supervised Monocular 3D Detection with a Single-View Image
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
SketchPlan: Diffusion Based Drone Planning From Human Sketches
von: Norelius, Sixten, et al.
Veröffentlicht: (2025)
von: Norelius, Sixten, et al.
Veröffentlicht: (2025)
A Survey of Label-Efficient Deep Learning for 3D Point Clouds
von: Xiao, Aoran, et al.
Veröffentlicht: (2023)
von: Xiao, Aoran, et al.
Veröffentlicht: (2023)
From Human Intention to Action Prediction: Intention-Driven End-to-End Autonomous Driving
von: Zheng, Huan, et al.
Veröffentlicht: (2025)
von: Zheng, Huan, et al.
Veröffentlicht: (2025)
PCR-GS: COLMAP-Free 3D Gaussian Splatting via Pose Co-Regularizations
von: Wei, Yu, et al.
Veröffentlicht: (2025)
von: Wei, Yu, et al.
Veröffentlicht: (2025)
MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
Context-Nav: Context-Driven Exploration and Viewpoint-Aware 3D Spatial Reasoning for Instance Navigation
von: Jang, Won Shik, et al.
Veröffentlicht: (2026)
von: Jang, Won Shik, et al.
Veröffentlicht: (2026)
LLM-Grounded Dynamic Task Planning with Hierarchical Temporal Logic for Human-Aware Multi-Robot Collaboration
von: Hu, Shuyuan, et al.
Veröffentlicht: (2026)
von: Hu, Shuyuan, et al.
Veröffentlicht: (2026)
RAP: 3D Rasterization Augmented End-to-End Planning
von: Feng, Lan, et al.
Veröffentlicht: (2025)
von: Feng, Lan, et al.
Veröffentlicht: (2025)
A Modern Take on Visual Relationship Reasoning for Grasp Planning
von: Rabino, Paolo, et al.
Veröffentlicht: (2024)
von: Rabino, Paolo, et al.
Veröffentlicht: (2024)
Generalized Trajectory Scoring for End-to-end Multimodal Planning
von: Li, Zhenxin, et al.
Veröffentlicht: (2025)
von: Li, Zhenxin, et al.
Veröffentlicht: (2025)
D3D-VLP: Dynamic 3D Vision-Language-Planning Model for Embodied Grounding and Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
PlanTRansformer: Unified Prediction and Planning with Goal-conditioned Transformer
von: Selzer, Constantin, et al.
Veröffentlicht: (2026)
von: Selzer, Constantin, et al.
Veröffentlicht: (2026)
LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
von: Huang, Zhe, et al.
Veröffentlicht: (2024)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement Learning
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
EgoPlan-Bench: Benchmarking Multimodal Large Language Models for Human-Level Planning
von: Chen, Yi, et al.
Veröffentlicht: (2023)
von: Chen, Yi, et al.
Veröffentlicht: (2023)
Solving Motion Planning Tasks with a Scalable Generative Model
von: Hu, Yihan, et al.
Veröffentlicht: (2024)
von: Hu, Yihan, et al.
Veröffentlicht: (2024)
ToDRE: Effective Visual Token Pruning via Token Diversity and Task Relevance
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs
von: Jiang, Xueying, et al.
Veröffentlicht: (2026)
von: Jiang, Xueying, et al.
Veröffentlicht: (2026)
6-DoF Grasp Planning using Fast 3D Reconstruction and Grasp Quality CNN
von: Avigal, Yahav, et al.
Veröffentlicht: (2020)
von: Avigal, Yahav, et al.
Veröffentlicht: (2020)
Exploiting Priors from 3D Diffusion Models for RGB-Based One-Shot View Planning
von: Pan, Sicong, et al.
Veröffentlicht: (2024)
von: Pan, Sicong, et al.
Veröffentlicht: (2024)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
Echo Planning for Autonomous Driving: From Current Observations to Future Trajectories and Back
von: Sun, Jintao, et al.
Veröffentlicht: (2025)
von: Sun, Jintao, et al.
Veröffentlicht: (2025)
Driving Intents Amplify Planning-Oriented Reinforcement Learning
von: Lu, Hengtong, et al.
Veröffentlicht: (2026)
von: Lu, Hengtong, et al.
Veröffentlicht: (2026)
Multi-view Pose Fusion for Occlusion-Aware 3D Human Pose Estimation
von: Bragagnolo, Laura, et al.
Veröffentlicht: (2024)
von: Bragagnolo, Laura, et al.
Veröffentlicht: (2024)
HOI4ABOT: Human-Object Interaction Anticipation for Human Intention Reading Collaborative roBOTs
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2023)
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2023)
Cognitive-Hierarchy Guided End-to-End Planning for Autonomous Driving
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
von: Wang, Zhennan, et al.
Veröffentlicht: (2025)
Semantics-Aware Next-best-view Planning for Efficient Search and Detection of Task-relevant Plant Parts
von: Burusa, Akshay K., et al.
Veröffentlicht: (2023)
von: Burusa, Akshay K., et al.
Veröffentlicht: (2023)
MomaGraph: State-Aware Unified Scene Graphs with Vision-Language Model for Embodied Task Planning
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
von: Ju, Yuanchen, et al.
Veröffentlicht: (2025)
Exploring 3D Human Pose Estimation and Forecasting from the Robot's Perspective: The HARPER Dataset
von: Avogaro, Andrea, et al.
Veröffentlicht: (2024)
von: Avogaro, Andrea, et al.
Veröffentlicht: (2024)
EEG-Driven Intention Decoding: Offline Deep Learning Benchmarking on a Robotic Rover
von: Alosaimi, Ghadah, et al.
Veröffentlicht: (2026)
von: Alosaimi, Ghadah, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders
von: Jiang, Xueying, et al.
Veröffentlicht: (2024) -
Multimodal 3D Reasoning Segmentation with Complex Scenes
von: Jiang, Xueying, et al.
Veröffentlicht: (2024) -
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
von: Li, Wenhao, et al.
Veröffentlicht: (2026) -
IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction
von: Qian, Lin, et al.
Veröffentlicht: (2026) -
L3DR: 3D-aware LiDAR Diffusion and Rectification
von: Liu, Quan, et al.
Veröffentlicht: (2026)