BFA: Best-Feature-Aware Fusion for Multi-View Fine-grained Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lan, Zihan, Mao, Weixin, Li, Haosheng, Wang, Le, Wang, Tiancai, Fan, Haoqiang, Yoshie, Osamu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BFA++: Hierarchical Best-Feature-Aware Token Prune for Multi-View Vision Language Action Model
von: Li, Haosheng, et al.
Veröffentlicht: (2026)
von: Li, Haosheng, et al.
Veröffentlicht: (2026)
Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation
von: Li, Haosheng, et al.
Veröffentlicht: (2024)
von: Li, Haosheng, et al.
Veröffentlicht: (2024)
RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
von: Mao, Yiming, et al.
Veröffentlicht: (2026)
von: Mao, Yiming, et al.
Veröffentlicht: (2026)
RoboGSim: A Real2Sim2Real Robotic Gaussian Splatting Simulator
von: Li, Xinhai, et al.
Veröffentlicht: (2024)
von: Li, Xinhai, et al.
Veröffentlicht: (2024)
Overlap-Aware Feature Learning for Robust Unsupervised Domain Adaptation for 3D Semantic Segmentation
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
SegGrasp: Zero-Shot Task-Oriented Grasping via Semantic and Geometric Guided Segmentation
von: Li, Haosheng, et al.
Veröffentlicht: (2024)
von: Li, Haosheng, et al.
Veröffentlicht: (2024)
SparseDFF: Sparse-View Feature Distillation for One-Shot Dexterous Manipulation
von: Wang, Qianxu, et al.
Veröffentlicht: (2023)
von: Wang, Qianxu, et al.
Veröffentlicht: (2023)
Hestia: Voxel-Face-Aware Hierarchical Next-Best-View Acquisition for Efficient 3D Reconstruction
von: Lu, Cheng-You, et al.
Veröffentlicht: (2025)
von: Lu, Cheng-You, et al.
Veröffentlicht: (2025)
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
von: Huang, Binyuan, et al.
Veröffentlicht: (2024)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
PEAfowl: Perception-Enhanced Multi-View Vision-Language-Action for Bimanual Manipulation
von: Fan, Qingyu, et al.
Veröffentlicht: (2026)
von: Fan, Qingyu, et al.
Veröffentlicht: (2026)
M4Diffuser: Multi-View Diffusion Policy with Manipulability-Aware Control for Robust Mobile Manipulation
von: Dong, Ju, et al.
Veröffentlicht: (2025)
von: Dong, Ju, et al.
Veröffentlicht: (2025)
PADriver: Towards Personalized Autonomous Driving
von: Kou, Genghua, et al.
Veröffentlicht: (2025)
von: Kou, Genghua, et al.
Veröffentlicht: (2025)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects
von: Chen, Jiahong, et al.
Veröffentlicht: (2025)
von: Chen, Jiahong, et al.
Veröffentlicht: (2025)
Learning Actionable Manipulation Recovery via Counterfactual Failure Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026)
von: Li, Dayou, et al.
Veröffentlicht: (2026)
FlowRAM: Grounding Flow Matching Policy with Region-Aware Mamba Framework for Robotic Manipulation
von: Wang, Sen, et al.
Veröffentlicht: (2025)
von: Wang, Sen, et al.
Veröffentlicht: (2025)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation
von: Shao, Dian, et al.
Veröffentlicht: (2026)
von: Shao, Dian, et al.
Veröffentlicht: (2026)
Low Resolution Next Best View for Robot Packing
von: Preziosa, Giuseppe Fabio, et al.
Veröffentlicht: (2025)
von: Preziosa, Giuseppe Fabio, et al.
Veröffentlicht: (2025)
Symmetry-Aware Fusion of Vision and Tactile Sensing via Bilateral Force Priors for Robotic Manipulation
von: Lee, Wonju, et al.
Veröffentlicht: (2026)
von: Lee, Wonju, et al.
Veröffentlicht: (2026)
OccFusion: Multi-Sensor Fusion Framework for 3D Semantic Occupancy Prediction
von: Ming, Zhenxing, et al.
Veröffentlicht: (2024)
von: Ming, Zhenxing, et al.
Veröffentlicht: (2024)
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
von: Feng, Ruoxuan, et al.
Veröffentlicht: (2024)
von: Feng, Ruoxuan, et al.
Veröffentlicht: (2024)
Active Next-Best-View Optimization for Risk-Averse Path Planning
von: Khass, Amirhossein Mollaei, et al.
Veröffentlicht: (2025)
von: Khass, Amirhossein Mollaei, et al.
Veröffentlicht: (2025)
From Instruction to Event: Sound-Triggered Mobile Manipulation
von: Ju, Hao, et al.
Veröffentlicht: (2026)
von: Ju, Hao, et al.
Veröffentlicht: (2026)
Enhancing UAV Search under Occlusion using Next Best View Planning
von: Strand, Sigrid Helene, et al.
Veröffentlicht: (2025)
von: Strand, Sigrid Helene, et al.
Veröffentlicht: (2025)
Boundary Exploration of Next Best View Policy in 3D Robotic Scanning
von: Li, Leihui, et al.
Veröffentlicht: (2024)
von: Li, Leihui, et al.
Veröffentlicht: (2024)
GenNBV: Generalizable Next-Best-View Policy for Active 3D Reconstruction
von: Chen, Xiao, et al.
Veröffentlicht: (2024)
von: Chen, Xiao, et al.
Veröffentlicht: (2024)
Active Implicit Object Reconstruction using Uncertainty-guided Next-Best-View Optimization
von: Yan, Dongyu, et al.
Veröffentlicht: (2023)
von: Yan, Dongyu, et al.
Veröffentlicht: (2023)
POp-GS: Next Best View in 3D-Gaussian Splatting with P-Optimality
von: Wilson, Joey, et al.
Veröffentlicht: (2025)
von: Wilson, Joey, et al.
Veröffentlicht: (2025)
LLaDA-VLA: Vision Language Diffusion Action Models
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
von: Wen, Yuqing, et al.
Veröffentlicht: (2025)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
von: Wang, Jiasen, et al.
Veröffentlicht: (2024)
von: Wang, Jiasen, et al.
Veröffentlicht: (2024)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation
von: Liu, Youzhi, et al.
Veröffentlicht: (2024)
von: Liu, Youzhi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BFA++: Hierarchical Best-Feature-Aware Token Prune for Multi-View Vision Language Action Model
von: Li, Haosheng, et al.
Veröffentlicht: (2026) -
Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation
von: Li, Haosheng, et al.
Veröffentlicht: (2024) -
RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
von: Mao, Weixin, et al.
Veröffentlicht: (2024) -
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025) -
Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness
von: Wang, Haochen, et al.
Veröffentlicht: (2025)