SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Posheng, Cheng, Powen, Faure, Gueter Josmy, Su, Hung-Ting, Hsu, Winston H. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning
by: Chunhachatrachai, Pawat, et al.
Published: (2026)
by: Chunhachatrachai, Pawat, et al.
Published: (2026)
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
by: Faure, Gueter Josmy, et al.
Published: (2026)
by: Faure, Gueter Josmy, et al.
Published: (2026)
HERMES: temporal-coHERent long-forM understanding with Episodes and Semantics
by: Faure, Gueter Josmy, et al.
Published: (2024)
by: Faure, Gueter Josmy, et al.
Published: (2024)
MovieCORE: COgnitive REasoning in Movies
by: Faure, Gueter Josmy, et al.
Published: (2025)
by: Faure, Gueter Josmy, et al.
Published: (2025)
Improving Generalization Ability for 3D Object Detection by Learning Sparsity-invariant Features
by: Lu, Hsin-Cheng, et al.
Published: (2025)
by: Lu, Hsin-Cheng, et al.
Published: (2025)
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
by: Su, Hung-Ting, et al.
Published: (2026)
by: Su, Hung-Ting, et al.
Published: (2026)
FunGraph: Functionality Aware 3D Scene Graphs for Language-Prompted Scene Interaction
by: Rotondi, Dennis, et al.
Published: (2025)
by: Rotondi, Dennis, et al.
Published: (2025)
WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection
by: Tsou, Tsung-Lin, et al.
Published: (2023)
by: Tsou, Tsung-Lin, et al.
Published: (2023)
MesaTask: Towards Task-Driven Tabletop Scene Generation via 3D Spatial Reasoning
by: Hao, Jinkun, et al.
Published: (2025)
by: Hao, Jinkun, et al.
Published: (2025)
Context-Aware Replanning with Pre-explored Semantic Map for Object Navigation
by: Ko, Po-Chen, et al.
Published: (2024)
by: Ko, Po-Chen, et al.
Published: (2024)
FunHOI: Annotation-Free 3D Hand-Object Interaction Generation via Functional Text Guidanc
by: Tian, Yongqi, et al.
Published: (2025)
by: Tian, Yongqi, et al.
Published: (2025)
FunGrasp: Functional Grasping for Diverse Dexterous Hands
by: Huang, Linyi, et al.
Published: (2024)
by: Huang, Linyi, et al.
Published: (2024)
ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints
by: Chen, Pei-An, et al.
Published: (2026)
by: Chen, Pei-An, et al.
Published: (2026)
Nav-R1: Reasoning and Navigation in Embodied Scenes
by: Liu, Qingxiang, et al.
Published: (2025)
by: Liu, Qingxiang, et al.
Published: (2025)
TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
by: Feng, ZhiYuan, et al.
Published: (2026)
by: Feng, ZhiYuan, et al.
Published: (2026)
Evaluating Robustness of Visual Representations for Object Assembly Task Requiring Spatio-Geometrical Reasoning
by: Ku, Chahyon, et al.
Published: (2023)
by: Ku, Chahyon, et al.
Published: (2023)
Probabilistic 3D Multi-Object Cooperative Tracking for Autonomous Driving via Differentiable Multi-Sensor Kalman Filter
by: Chiu, Hsu-kuang, et al.
Published: (2023)
by: Chiu, Hsu-kuang, et al.
Published: (2023)
Belief Scene Graphs: Expanding Partial Scenes with Objects through Computation of Expectation
by: Saucedo, Mario A. V., et al.
Published: (2024)
by: Saucedo, Mario A. V., et al.
Published: (2024)
DeformGS: Scene Flow in Highly Deformable Scenes for Deformable Object Manipulation
by: Duisterhof, Bardienus P., et al.
Published: (2023)
by: Duisterhof, Bardienus P., et al.
Published: (2023)
Affordance-Guided Coarse-to-Fine Exploration for Base Placement in Open-Vocabulary Mobile Manipulation
by: Lin, Tzu-Jung, et al.
Published: (2025)
by: Lin, Tzu-Jung, et al.
Published: (2025)
CompassAD: Intent-Driven 3D Affordance Grounding in Functionally Competing Objects
by: Li, Jingliang, et al.
Published: (2026)
by: Li, Jingliang, et al.
Published: (2026)
EmbodiedVSR: Dynamic Scene Graph-Guided Chain-of-Thought Reasoning for Visual Spatial Tasks
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
AED: Adaptable Error Detection for Few-shot Imitation Policy
by: Yeh, Jia-Fong, et al.
Published: (2024)
by: Yeh, Jia-Fong, et al.
Published: (2024)
PIGEON: VLM-Driven Object Navigation via Points of Interest Selection
by: Peng, Cheng, et al.
Published: (2025)
by: Peng, Cheng, et al.
Published: (2025)
VICtoR: Learning Hierarchical Vision-Instruction Correlation Rewards for Long-horizon Manipulation
by: Hung, Kuo-Han, et al.
Published: (2024)
by: Hung, Kuo-Han, et al.
Published: (2024)
SG-Tailor: Inter-Object Commonsense Relationship Reasoning for Scene Graph Manipulation
by: Shang, Haoliang, et al.
Published: (2025)
by: Shang, Haoliang, et al.
Published: (2025)
Improved Scene Landmark Detection for Camera Localization
by: Do, Tien, et al.
Published: (2024)
by: Do, Tien, et al.
Published: (2024)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
Zero-shot Reconstruction of In-Scene Object Manipulation from Video
by: Lin, Dixuan, et al.
Published: (2025)
by: Lin, Dixuan, et al.
Published: (2025)
Efficient Multi-Task Scene Analysis with RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
ASHiTA: Automatic Scene-grounded HIerarchical Task Analysis
by: Chang, Yun, et al.
Published: (2025)
by: Chang, Yun, et al.
Published: (2025)
VISO-Grasp: Vision-Language Informed Spatial Object-centric 6-DoF Active View Planning and Grasping in Clutter and Invisibility
by: Shi, Yitian, et al.
Published: (2025)
by: Shi, Yitian, et al.
Published: (2025)
Online 3D Scene Reconstruction Using Neural Object Priors
by: Chabal, Thomas, et al.
Published: (2025)
by: Chabal, Thomas, et al.
Published: (2025)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
Open Scene Graphs for Open World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2024)
by: Loo, Joel, et al.
Published: (2024)
Open Scene Graphs for Open-World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2025)
by: Loo, Joel, et al.
Published: (2025)
Tether: Autonomous Functional Play with Correspondence-Driven Trajectory Warping
by: Liang, William, et al.
Published: (2026)
by: Liang, William, et al.
Published: (2026)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
by: Rim, Patrick, et al.
Published: (2026)
by: Rim, Patrick, et al.
Published: (2026)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
by: Liu, An-Lun, et al.
Published: (2025)
by: Liu, An-Lun, et al.
Published: (2025)
Queryable 3D Scene Representation: A Multi-Modal Framework for Semantic Reasoning and Robotic Task Planning
by: Li, Xun, et al.
Published: (2025)
by: Li, Xun, et al.
Published: (2025)
Similar Items
-
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning
by: Chunhachatrachai, Pawat, et al.
Published: (2026) -
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
by: Faure, Gueter Josmy, et al.
Published: (2026) -
HERMES: temporal-coHERent long-forM understanding with Episodes and Semantics
by: Faure, Gueter Josmy, et al.
Published: (2024) -
MovieCORE: COgnitive REasoning in Movies
by: Faure, Gueter Josmy, et al.
Published: (2025) -
Improving Generalization Ability for 3D Object Detection by Learning Sparsity-invariant Features
by: Lu, Hsin-Cheng, et al.
Published: (2025)