FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Delitzas, Alexandros, Zhang, Chenyangguang, Gavryushin, Alexey, Di Mario, Tommaso, Sun, Boyang, Dabral, Rishabh, Guibas, Leonidas, Theobalt, Christian, Pollefeys, Marc, Engelmann, Francis, Barath, Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026)
by: Hu, Xinggang, et al.
Published: (2026)
Open-Vocabulary Functional 3D Scene Graphs for Real-World Indoor Spaces
by: Zhang, Chenyangguang, et al.
Published: (2025)
by: Zhang, Chenyangguang, et al.
Published: (2025)
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025)
by: Lascheit, Kai, et al.
Published: (2025)
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints
by: Zhang, Chenyangguang, et al.
Published: (2026)
by: Zhang, Chenyangguang, et al.
Published: (2026)
REACT3D: Recovering Articulations for Interactive Physical 3D Scenes
by: Huang, Zhao, et al.
Published: (2025)
by: Huang, Zhao, et al.
Published: (2025)
SIGHT: Synthesizing Image-Text Conditioned and Geometry-Guided 3D Hand-Object Trajectories
by: Gavryushin, Alexey, et al.
Published: (2025)
by: Gavryushin, Alexey, et al.
Published: (2025)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
by: Gopal, Suhas, et al.
Published: (2025)
by: Gopal, Suhas, et al.
Published: (2025)
FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video
by: Camiletto, Andrea Boscolo, et al.
Published: (2025)
by: Camiletto, Andrea Boscolo, et al.
Published: (2025)
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
by: Fedele, Elisabetta, et al.
Published: (2025)
by: Fedele, Elisabetta, et al.
Published: (2025)
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
by: Aytekin, Ayce Idil, et al.
Published: (2025)
by: Aytekin, Ayce Idil, et al.
Published: (2025)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
by: Gavryushin, Alexey, et al.
Published: (2025)
by: Gavryushin, Alexey, et al.
Published: (2025)
MIBURI: Towards Expressive Interactive Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2026)
by: Mughal, M. Hamza, et al.
Published: (2026)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
UnLoc: Leveraging Depth Uncertainties for Floorplan Localization
by: Wüest, Matthias, et al.
Published: (2025)
by: Wüest, Matthias, et al.
Published: (2025)
Real-time Free-view Human Rendering from Sparse-view RGB Videos using Double Unprojected Textures
by: Sun, Guoxing, et al.
Published: (2024)
by: Sun, Guoxing, et al.
Published: (2024)
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
by: Ghosh, Anindita, et al.
Published: (2023)
by: Ghosh, Anindita, et al.
Published: (2023)
CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration
by: Meric, Adil, et al.
Published: (2026)
by: Meric, Adil, et al.
Published: (2026)
Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
MetaCap: Meta-learning Priors from Multi-View Imagery for Sparse-view Human Performance Capture and Rendering
by: Sun, Guoxing, et al.
Published: (2024)
by: Sun, Guoxing, et al.
Published: (2024)
Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions
by: Aytekin, Ayce Idil, et al.
Published: (2026)
by: Aytekin, Ayce Idil, et al.
Published: (2026)
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
by: Ghosh, Anindita, et al.
Published: (2026)
by: Ghosh, Anindita, et al.
Published: (2026)
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
by: Maillard, Léopold, et al.
Published: (2026)
by: Maillard, Léopold, et al.
Published: (2026)
"Where am I?" Scene Retrieval with Language
by: Chen, Jiaqi, et al.
Published: (2024)
by: Chen, Jiaqi, et al.
Published: (2024)
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
by: Zhang, Wanyue, et al.
Published: (2023)
by: Zhang, Wanyue, et al.
Published: (2023)
ARKit LabelMaker: A New Scale for Indoor 3D Scene Understanding
by: Ji, Guangda, et al.
Published: (2024)
by: Ji, Guangda, et al.
Published: (2024)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
Gravity-aligned Rotation Averaging with Circular Regression
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Learning to Make Keypoints Sub-Pixel Accurate
by: Kim, Shinjeong, et al.
Published: (2024)
by: Kim, Shinjeong, et al.
Published: (2024)
BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects
by: Zhang, Wanyue, et al.
Published: (2024)
by: Zhang, Wanyue, et al.
Published: (2024)
EmbodMocap: In-the-Wild 4D Human-Scene Reconstruction for Embodied Agents
by: Wang, Wenjia, et al.
Published: (2026)
by: Wang, Wenjia, et al.
Published: (2026)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
CrossOver: 3D Scene Cross-Modal Alignment
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
by: Takmaz, Ayca, et al.
Published: (2024)
by: Takmaz, Ayca, et al.
Published: (2024)
PractiLight: Practical Light Control Using Foundational Diffusion Models
by: Erel, Yotam, et al.
Published: (2025)
by: Erel, Yotam, et al.
Published: (2025)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
by: Bieri, Valentin, et al.
Published: (2025)
by: Bieri, Valentin, et al.
Published: (2025)
No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos
by: Balice, Matteo, et al.
Published: (2026)
by: Balice, Matteo, et al.
Published: (2026)
LookOut: Real-World Humanoid Egocentric Navigation
by: Pan, Boxiao, et al.
Published: (2025)
by: Pan, Boxiao, et al.
Published: (2025)
Similar Items
-
SuperDec: 3D Scene Decomposition with Superquadric Primitives
by: Fedele, Elisabetta, et al.
Published: (2025) -
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026) -
Open-Vocabulary Functional 3D Scene Graphs for Real-World Indoor Spaces
by: Zhang, Chenyangguang, et al.
Published: (2025) -
Robust Human Registration with Body Part Segmentation on Noisy Point Clouds
by: Lascheit, Kai, et al.
Published: (2025) -
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints
by: Zhang, Chenyangguang, et al.
Published: (2026)