Towards Task-Oriented Flying: Framework, Infrastructure, and Principles
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Kangyao, Wang, Hao, Chen, Jingyu, Chen, Jintao, Luo, Yu, Guo, Di, Zhang, Xiangkui, Ji, Xiangyang, Liu, Huaping |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PAS-SLAM: A Visual SLAM System for Planar Ambiguous Scenes
by: Hu, Xinggang, et al.
Published: (2024)
by: Hu, Xinggang, et al.
Published: (2024)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026)
by: Hu, Xinggang, et al.
Published: (2026)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
by: Liu, An-Lun, et al.
Published: (2025)
by: Liu, An-Lun, et al.
Published: (2025)
Task-Oriented 6-DoF Grasp Pose Detection in Clutters
by: Wang, An-Lan, et al.
Published: (2025)
by: Wang, An-Lan, et al.
Published: (2025)
TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation
by: Zhao, Hongxiang, et al.
Published: (2025)
by: Zhao, Hongxiang, et al.
Published: (2025)
Stimulate the Potential of Robots via Competition
by: Huang, Kangyao, et al.
Published: (2024)
by: Huang, Kangyao, et al.
Published: (2024)
TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation
by: Zhao, Sizhe, et al.
Published: (2026)
by: Zhao, Sizhe, et al.
Published: (2026)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
by: Wang, Guokang, et al.
Published: (2024)
by: Wang, Guokang, et al.
Published: (2024)
UAV-Based Infrastructure Inspections: A Literature Review and Proposed Framework for AEC+FM
by: Nikkhah, Amir Farzin, et al.
Published: (2026)
by: Nikkhah, Amir Farzin, et al.
Published: (2026)
MesaTask: Towards Task-Driven Tabletop Scene Generation via 3D Spatial Reasoning
by: Hao, Jinkun, et al.
Published: (2025)
by: Hao, Jinkun, et al.
Published: (2025)
FlyPose: Towards Robust Human Pose Estimation From Aerial Views
by: Farooq, Hassaan, et al.
Published: (2026)
by: Farooq, Hassaan, et al.
Published: (2026)
Adaptive Articulated Object Manipulation On The Fly with Foundation Model Reasoning and Part Grounding
by: Zhang, Xiaojie, et al.
Published: (2025)
by: Zhang, Xiaojie, et al.
Published: (2025)
Towards Immersive Human-X Interaction: A Real-Time Framework for Physically Plausible Motion Synthesis
by: Ji, Kaiyang, et al.
Published: (2025)
by: Ji, Kaiyang, et al.
Published: (2025)
TP-MDDN: Task-Preferenced Multi-Demand-Driven Navigation with Autonomous Decision-Making
by: Li, Shanshan, et al.
Published: (2025)
by: Li, Shanshan, et al.
Published: (2025)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
UW-SDF: Exploiting Hybrid Geometric Priors for Neural SDF Reconstruction from Underwater Multi-view Monocular Images
by: Chen, Zeyu, et al.
Published: (2024)
by: Chen, Zeyu, et al.
Published: (2024)
FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation
by: Guo, Jun, et al.
Published: (2025)
by: Guo, Jun, et al.
Published: (2025)
LongFly: Long-Horizon UAV Vision-and-Language Navigation with Spatiotemporal Context Integration
by: Jiang, Wen, et al.
Published: (2025)
by: Jiang, Wen, et al.
Published: (2025)
On-the-Fly SfM: What you capture is What you get
by: Zhan, Zongqian, et al.
Published: (2023)
by: Zhan, Zongqian, et al.
Published: (2023)
GFreeDet: Exploiting Gaussian Splatting and Foundation Models for Model-free Unseen Object Detection in the BOP Challenge 2024
by: Liu, Xingyu, et al.
Published: (2024)
by: Liu, Xingyu, et al.
Published: (2024)
Multi-Agent 3D Map Reconstruction and Change Detection in Microgravity with Free-Flying Robots
by: Dinkel, Holly, et al.
Published: (2023)
by: Dinkel, Holly, et al.
Published: (2023)
iFlyBot-VLA Technical Report
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
Towards Next-Generation SLAM: A Survey on 3DGS-SLAM Focusing on Performance, Robustness, and Future Directions
by: Wang, Li, et al.
Published: (2026)
by: Wang, Li, et al.
Published: (2026)
SuperFusion: Multilevel LiDAR-Camera Fusion for Long-Range HD Map Generation
by: Dong, Hao, et al.
Published: (2022)
by: Dong, Hao, et al.
Published: (2022)
ManipArena: Comprehensive Real-world Evaluation of Reasoning-Oriented Generalist Robot Manipulation
by: Sun, Yu, et al.
Published: (2026)
by: Sun, Yu, et al.
Published: (2026)
Solving Motion Planning Tasks with a Scalable Generative Model
by: Hu, Yihan, et al.
Published: (2024)
by: Hu, Yihan, et al.
Published: (2024)
You Only Estimate Once: Unified, One-stage, Real-Time Category-level Articulated Object 6D Pose Estimation for Robotic Grasping
by: Huang, Jingshun, et al.
Published: (2025)
by: Huang, Jingshun, et al.
Published: (2025)
ODYSSEY: Open-World Quadrupeds Exploration and Manipulation for Long-Horizon Tasks
by: Wang, Kaijun, et al.
Published: (2025)
by: Wang, Kaijun, et al.
Published: (2025)
Arcadia: Toward a Full-Lifecycle Framework for Embodied Lifelong Learning
by: Gao, Minghe, et al.
Published: (2025)
by: Gao, Minghe, et al.
Published: (2025)
Digital and Robotic Twinning for Validation of Proximity Operations and Formation Flying
by: Ahmed, Z., et al.
Published: (2025)
by: Ahmed, Z., et al.
Published: (2025)
EventFly: Event Camera Perception from Ground to the Sky
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Target-Oriented Object Grasping via Multimodal Human Guidance
by: Xie, Pengwei, et al.
Published: (2024)
by: Xie, Pengwei, et al.
Published: (2024)
Vision-Only Gaussian Splatting for Collaborative Semantic Occupancy Prediction
by: Chen, Cheng, et al.
Published: (2025)
by: Chen, Cheng, et al.
Published: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
by: Bai, Yongjie, et al.
Published: (2025)
by: Bai, Yongjie, et al.
Published: (2025)
Optimizing Drug Delivery in Smart Pharmacies: A Novel Framework of Multi-Stage Grasping Network Combined with Adaptive Robotics Mechanism
by: Tang, Rui, et al.
Published: (2024)
by: Tang, Rui, et al.
Published: (2024)
Learning on the Fly: Replay-Based Continual Object Perception for Indoor Drones
by: Nae, Sebastian-Ion, et al.
Published: (2026)
by: Nae, Sebastian-Ion, et al.
Published: (2026)
Event-Aided Sharp Radiance Field Reconstruction for Fast-Flying Drones
by: Zou, Rong, et al.
Published: (2026)
by: Zou, Rong, et al.
Published: (2026)
GraspLDP: Towards Generalizable Grasping Policy via Latent Diffusion
by: Xiang, Enda, et al.
Published: (2026)
by: Xiang, Enda, et al.
Published: (2026)
A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning
by: Zhang, Zixin, et al.
Published: (2025)
by: Zhang, Zixin, et al.
Published: (2025)
Similar Items
-
PAS-SLAM: A Visual SLAM System for Planar Ambiguous Scenes
by: Hu, Xinggang, et al.
Published: (2024) -
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026) -
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
by: Liu, An-Lun, et al.
Published: (2025) -
Task-Oriented 6-DoF Grasp Pose Detection in Clutters
by: Wang, An-Lan, et al.
Published: (2025) -
TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation
by: Zhao, Hongxiang, et al.
Published: (2025)