VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Haowen, Zhang, Shaolong, Li, Mingyang, Ma, Chengzhong, Chen, Xinzhe, Cui, Qiongjie, Chen, Xingyu, Liu, Zeyang, Lan, Xuguang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation
by: Li, Mingyang, et al.
Published: (2026)
by: Li, Mingyang, et al.
Published: (2026)
PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations
by: Sun, Haowen, et al.
Published: (2025)
by: Sun, Haowen, et al.
Published: (2025)
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
by: Chen, Xinzhe, et al.
Published: (2026)
by: Chen, Xinzhe, et al.
Published: (2026)
AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter
by: Tang, Yingbo, et al.
Published: (2025)
by: Tang, Yingbo, et al.
Published: (2025)
Bridging Simulation and Reality: Cross-Domain Transfer with Semantic 2D Gaussian Splatting
by: Tang, Jian, et al.
Published: (2025)
by: Tang, Jian, et al.
Published: (2025)
REGNet V2: End-to-End REgion-based Grasp Detection Network for Grippers of Different Sizes in Point Clouds
by: Zhao, Binglei, et al.
Published: (2024)
by: Zhao, Binglei, et al.
Published: (2024)
OpenVox: Real-time Instance-level Open-vocabulary Probabilistic Voxel Representation
by: Deng, Yinan, et al.
Published: (2025)
by: Deng, Yinan, et al.
Published: (2025)
AffordDP: Generalizable Diffusion Policy with Transferable Affordance
by: Wu, Shijie, et al.
Published: (2024)
by: Wu, Shijie, et al.
Published: (2024)
DexDiff: Towards Extrinsic Dexterity Manipulation of Ungraspable Objects in Unrestricted Environments
by: Ma, Chengzhong, et al.
Published: (2024)
by: Ma, Chengzhong, et al.
Published: (2024)
AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance
by: Wei, Yi-Lin, et al.
Published: (2025)
by: Wei, Yi-Lin, et al.
Published: (2025)
RoboAfford++: A Generative AI-Enhanced Dataset for Multimodal Affordance Learning in Robotic Manipulation and Navigation
by: Hao, Xiaoshuai, et al.
Published: (2025)
by: Hao, Xiaoshuai, et al.
Published: (2025)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
by: Zhu, Xiaomeng, et al.
Published: (2025)
by: Zhu, Xiaomeng, et al.
Published: (2025)
Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance
by: Wang, Runze, et al.
Published: (2026)
by: Wang, Runze, et al.
Published: (2026)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
by: Chu, Hengshuo, et al.
Published: (2025)
by: Chu, Hengshuo, et al.
Published: (2025)
AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis
by: Wu, Xiaofei, et al.
Published: (2026)
by: Wu, Xiaofei, et al.
Published: (2026)
EqvAfford: SE(3) Equivariance for Point-Level Affordance Learning
by: Chen, Yue, et al.
Published: (2024)
by: Chen, Yue, et al.
Published: (2024)
OVAL-Prompt: Open-Vocabulary Affordance Localization for Robot Manipulation through LLM Affordance-Grounding
by: Tong, Edmond, et al.
Published: (2024)
by: Tong, Edmond, et al.
Published: (2024)
OVAL-Grasp: Open-Vocabulary Affordance Localization for Task Oriented Grasping
by: Tong, Edmond, et al.
Published: (2025)
by: Tong, Edmond, et al.
Published: (2025)
PreAfford: Universal Affordance-Based Pre-Grasping for Diverse Objects and Environments
by: Ding, Kairui, et al.
Published: (2024)
by: Ding, Kairui, et al.
Published: (2024)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
Toward Reliable Sim-to-Real Predictability for MoE-based Robust Quadrupedal Locomotion
by: Wu, Tianyang, et al.
Published: (2026)
by: Wu, Tianyang, et al.
Published: (2026)
Affordance-Guided Coarse-to-Fine Exploration for Base Placement in Open-Vocabulary Mobile Manipulation
by: Lin, Tzu-Jung, et al.
Published: (2025)
by: Lin, Tzu-Jung, et al.
Published: (2025)
AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment
by: Kong, Weijie, et al.
Published: (2026)
by: Kong, Weijie, et al.
Published: (2026)
FUS3DMaps: Scalable and Accurate Open-Vocabulary Semantic Mapping by 3D Fusion of Voxel- and Instance-Level Layers
by: Homberger, Timon, et al.
Published: (2026)
by: Homberger, Timon, et al.
Published: (2026)
AffordTissue: Dense Affordance Prediction for Tool-Action Specific Tissue Interaction
by: Maksutova, Aiza, et al.
Published: (2026)
by: Maksutova, Aiza, et al.
Published: (2026)
ECHO: Continuous Hierarchical Memory for Vision-Language-Action Models
by: Hu, Yanbin, et al.
Published: (2026)
by: Hu, Yanbin, et al.
Published: (2026)
VoxAct-B: Voxel-Based Acting and Stabilizing Policy for Bimanual Manipulation
by: Liu, I-Chun Arthur, et al.
Published: (2024)
by: Liu, I-Chun Arthur, et al.
Published: (2024)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
by: Tian, Tongxuan, et al.
Published: (2025)
by: Tian, Tongxuan, et al.
Published: (2025)
VoxNeRF: Bridging Voxel Representation and Neural Radiance Fields for Enhanced Indoor View Synthesis
by: Wang, Sen, et al.
Published: (2023)
by: Wang, Sen, et al.
Published: (2023)
AffordGen: Generating Diverse Demonstrations for Generalizable Object Manipulation with Afford Correspondence
by: Zhang, Jiawei, et al.
Published: (2026)
by: Zhang, Jiawei, et al.
Published: (2026)
OVGrasp: Open-Vocabulary Grasping Assistance via Multimodal Intent Detection
by: Hu, Chen, et al.
Published: (2025)
by: Hu, Chen, et al.
Published: (2025)
LIO-GVM: an Accurate, Tightly-Coupled Lidar-Inertial Odometry with Gaussian Voxel Map
by: Ji, Xingyu, et al.
Published: (2023)
by: Ji, Xingyu, et al.
Published: (2023)
OVExp: Open Vocabulary Exploration for Object-Oriented Navigation
by: Wei, Meng, et al.
Published: (2024)
by: Wei, Meng, et al.
Published: (2024)
Real-Time Roadway Obstacle Detection for Electric Scooters Using Deep Learning and Multi-Sensor Fusion
by: Zheng, Zeyang, et al.
Published: (2025)
by: Zheng, Zeyang, et al.
Published: (2025)
CT-VoxelMap: Efficient Continuous-Time LiDAR-Inertial Odometry with Probabilistic Adaptive Voxel Mapping
by: Zhao, Lei, et al.
Published: (2026)
by: Zhao, Lei, et al.
Published: (2026)
Open-Vocabulary Part-Based Grasping
by: van Oort, Tjeard, et al.
Published: (2024)
by: van Oort, Tjeard, et al.
Published: (2024)
SAGA: Open-World Mobile Manipulation via Structured Affordance Grounding
by: Fang, Kuan, et al.
Published: (2025)
by: Fang, Kuan, et al.
Published: (2025)
One-Step Flow Policy: Self-Distillation for Fast Visuomotor Policies
by: Li, Shaolong, et al.
Published: (2026)
by: Li, Shaolong, et al.
Published: (2026)
DRIVE-Nav: Directional Reasoning, Inspection, and Verification for Efficient Open-Vocabulary Navigation
by: Gao, Maoguo, et al.
Published: (2026)
by: Gao, Maoguo, et al.
Published: (2026)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
Similar Items
-
AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation
by: Li, Mingyang, et al.
Published: (2026) -
PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations
by: Sun, Haowen, et al.
Published: (2025) -
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
by: Chen, Xinzhe, et al.
Published: (2026) -
AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter
by: Tang, Yingbo, et al.
Published: (2025) -
Bridging Simulation and Reality: Cross-Domain Transfer with Semantic 2D Gaussian Splatting
by: Tang, Jian, et al.
Published: (2025)