SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
Fuente:
arXiv
Saved in:
| Main Authors: | Maillard, Léopold, Engelmann, Francis, Durand, Tom, Pan, Boxiao, You, Yang, Litany, Or, Guibas, Leonidas, Ovsjanikov, Maks |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LACONIC: A 3D Layout Adapter for Controllable Image Creation
by: Maillard, Léopold, et al.
Published: (2025)
by: Maillard, Léopold, et al.
Published: (2025)
DeBaRA: Denoising-Based 3D Room Arrangement Generation
by: Maillard, Léopold, et al.
Published: (2024)
by: Maillard, Léopold, et al.
Published: (2024)
Beyond Prompts: Unconditional 3D Inversion for Out-of-Distribution Shapes
by: Chen, Victoria Yue, et al.
Published: (2026)
by: Chen, Victoria Yue, et al.
Published: (2026)
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances
by: Wang, Qirui, et al.
Published: (2026)
by: Wang, Qirui, et al.
Published: (2026)
Dynamic Reflections: Probing Video Representations with Text Alignment
by: Zhu, Tyler, et al.
Published: (2025)
by: Zhu, Tyler, et al.
Published: (2025)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
FILTR: Extracting Topological Features from Pretrained 3D Models
by: Martinez, Louis, et al.
Published: (2026)
by: Martinez, Louis, et al.
Published: (2026)
Memory-Scalable and Simplified Functional Map Learning
by: Magnet, Robin, et al.
Published: (2024)
by: Magnet, Robin, et al.
Published: (2024)
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
by: Rozenberszki, David, et al.
Published: (2023)
by: Rozenberszki, David, et al.
Published: (2023)
FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
by: Delitzas, Alexandros, et al.
Published: (2026)
by: Delitzas, Alexandros, et al.
Published: (2026)
Shape Non-rigid Kinematics (SNK): A Zero-Shot Method for Non-Rigid Shape Matching via Unsupervised Functional Map Regularized Reconstruction
by: Attaiki, Souhaib, et al.
Published: (2024)
by: Attaiki, Souhaib, et al.
Published: (2024)
Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features
by: Wimmer, Thomas, et al.
Published: (2023)
by: Wimmer, Thomas, et al.
Published: (2023)
Grounding 3D Scene Affordance From Egocentric Interactions
by: Liu, Cuiyu, et al.
Published: (2024)
by: Liu, Cuiyu, et al.
Published: (2024)
Img2CAD: Reverse Engineering 3D CAD Models from Images through VLM-Assisted Conditional Factorization
by: You, Yang, et al.
Published: (2024)
by: You, Yang, et al.
Published: (2024)
Animal Pose Labeling Using General-Purpose Point Trackers
by: Pan, Zhuoyang, et al.
Published: (2025)
by: Pan, Zhuoyang, et al.
Published: (2025)
Smoothed Graph Contrastive Learning via Seamless Proximity Integration
by: Behmanesh, Maysam, et al.
Published: (2024)
by: Behmanesh, Maysam, et al.
Published: (2024)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
by: Bieri, Valentin, et al.
Published: (2025)
by: Bieri, Valentin, et al.
Published: (2025)
LookOut: Real-World Humanoid Egocentric Navigation
by: Pan, Boxiao, et al.
Published: (2025)
by: Pan, Boxiao, et al.
Published: (2025)
OCH3R: Object-Centric Holistic 3D Reconstruction
by: Du, Yi, et al.
Published: (2026)
by: Du, Yi, et al.
Published: (2026)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
by: Lei, Jiahui, et al.
Published: (2025)
by: Lei, Jiahui, et al.
Published: (2025)
Generative Drifting is Secretly Score Matching: a Spectral and Variational Perspective
by: Turan, Erkan, et al.
Published: (2026)
by: Turan, Erkan, et al.
Published: (2026)
Self-Supervised Dual Contouring
by: Sundararaman, Ramana, et al.
Published: (2024)
by: Sundararaman, Ramana, et al.
Published: (2024)
FourieRF: Few-Shot NeRFs via Progressive Fourier Frequency Control
by: Gomez, Diego, et al.
Published: (2025)
by: Gomez, Diego, et al.
Published: (2025)
Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication
by: Behmanesh, Maysam, et al.
Published: (2025)
by: Behmanesh, Maysam, et al.
Published: (2025)
To Supervise or Not to Supervise: Understanding and Addressing the Key Challenges of Point Cloud Transfer Learning
by: Hadgi, Souhail, et al.
Published: (2024)
by: Hadgi, Souhail, et al.
Published: (2024)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
by: Jia, Wanjun, et al.
Published: (2025)
by: Jia, Wanjun, et al.
Published: (2025)
Global Motion Corresponder for 3D Point-Based Scene Interpolation under Large Motion
by: Lin, Junru, et al.
Published: (2025)
by: Lin, Junru, et al.
Published: (2025)
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025)
by: Lascheit, Kai, et al.
Published: (2025)
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
by: Huang, Ian, et al.
Published: (2024)
by: Huang, Ian, et al.
Published: (2024)
OpenSUN3D: 1st Workshop Challenge on Open-Vocabulary 3D Scene Understanding
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
Agentic Scene Policies: Unifying Space, Semantics, and Affordances for Robot Action
by: Morin, Sacha, et al.
Published: (2025)
by: Morin, Sacha, et al.
Published: (2025)
A3R: Agentic Affordance Reasoning via Cross-Dimensional Evidence in 3D Gaussian Scenes
by: Li, Di, et al.
Published: (2026)
by: Li, Di, et al.
Published: (2026)
MultiPhys: Multi-Person Physics-aware 3D Motion Estimation
by: Ugrinovic, Nicolas, et al.
Published: (2024)
by: Ugrinovic, Nicolas, et al.
Published: (2024)
Neural Attention Field: Emerging Point Relevance in 3D Scenes for One-Shot Dexterous Grasping
by: Wang, Qianxu, et al.
Published: (2024)
by: Wang, Qianxu, et al.
Published: (2024)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026)
by: Hu, Xinggang, et al.
Published: (2026)
Open-Vocabulary Functional 3D Scene Graphs for Real-World Indoor Spaces
by: Zhang, Chenyangguang, et al.
Published: (2025)
by: Zhang, Chenyangguang, et al.
Published: (2025)
Multiview Equivariance Improves 3D Correspondence Understanding with Minimal Feature Finetuning
by: You, Yang, et al.
Published: (2024)
by: You, Yang, et al.
Published: (2024)
Similar Items
-
LACONIC: A 3D Layout Adapter for Controllable Image Creation
by: Maillard, Léopold, et al.
Published: (2025) -
DeBaRA: Denoising-Based 3D Room Arrangement Generation
by: Maillard, Léopold, et al.
Published: (2024) -
Beyond Prompts: Unconditional 3D Inversion for Out-of-Distribution Shapes
by: Chen, Victoria Yue, et al.
Published: (2026) -
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025) -
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
by: Fedele, Elisabetta, et al.
Published: (2025)