Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
Fuente:
arXiv
Saved in:
| Main Authors: | Udugama, U. V. B. L., Vosselman, George, Nex, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
M2H-MX: Multi-Task Semantic and Geometric Perception for Real-Time Monocular 3D Scene Graph Construction
by: Udugama, U. V. B. L., et al.
Published: (2026)
by: Udugama, U. V. B. L., et al.
Published: (2026)
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
by: Udugama, U. V. B. L, et al.
Published: (2025)
by: Udugama, U. V. B. L, et al.
Published: (2025)
Mono-hydra: Real-time 3D scene graph construction from monocular camera input with IMU
by: Udugama, U. V. B. L., et al.
Published: (2023)
by: Udugama, U. V. B. L., et al.
Published: (2023)
A Comparison of Multi-View Stereo Methods for Photogrammetric 3D Reconstruction: From Traditional to Learning-Based Approaches
by: Li, Yawen, et al.
Published: (2026)
by: Li, Yawen, et al.
Published: (2026)
ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging
by: Iz, Selim Ahmet, et al.
Published: (2026)
by: Iz, Selim Ahmet, et al.
Published: (2026)
Clio: Real-time Task-Driven Open-Set 3D Scene Graphs
by: Maggio, Dominic, et al.
Published: (2024)
by: Maggio, Dominic, et al.
Published: (2024)
Feasibility of Indoor Frame-Wise Lidar Semantic Segmentation via Distillation from Visual Foundation Model
by: Wu, Haiyang, et al.
Published: (2026)
by: Wu, Haiyang, et al.
Published: (2026)
Terra: Hierarchical Terrain-Aware 3D Scene Graph for Task-Agnostic Outdoor Mapping
by: Samuelson, Chad R., et al.
Published: (2025)
by: Samuelson, Chad R., et al.
Published: (2025)
LEXI-SG: Monocular 3D Scene Graph Mapping with Room-Guided Feed-Forward Reconstruction
by: Kassab, Christina, et al.
Published: (2026)
by: Kassab, Christina, et al.
Published: (2026)
VINGS-Mono: Visual-Inertial Gaussian Splatting Monocular SLAM in Large Scenes
by: Wu, Ke, et al.
Published: (2025)
by: Wu, Ke, et al.
Published: (2025)
REACT: Real-time Efficient Attribute Clustering and Transfer for Updatable 3D Scene Graph
by: Nguyen, Phuoc, et al.
Published: (2025)
by: Nguyen, Phuoc, et al.
Published: (2025)
MonoGlass3D: Monocular 3D Glass Detection with Plane Regression and Adaptive Feature Fusion
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
MonoLSS: Learnable Sample Selection For Monocular 3D Detection
by: Li, Zhenjia, et al.
Published: (2023)
by: Li, Zhenjia, et al.
Published: (2023)
MA3DSG: Multi-Agent 3D Scene Graph Generation for Large-Scale Indoor Environments
by: Kim, Yirum, et al.
Published: (2026)
by: Kim, Yirum, et al.
Published: (2026)
MoD-SLAM: Monocular Dense Mapping for Unbounded 3D Scene Reconstruction
by: Zhou, Heng, et al.
Published: (2024)
by: Zhou, Heng, et al.
Published: (2024)
Real-time Monocular 2D and 3D Perception of Endoluminal Scenes for Controlling Flexible Robotic Endoscopic Instruments
by: Wei, Ruofeng, et al.
Published: (2026)
by: Wei, Ruofeng, et al.
Published: (2026)
OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding
by: Zhu, Siting, et al.
Published: (2026)
by: Zhu, Siting, et al.
Published: (2026)
MonoSpheres: Large-Scale Monocular SLAM-Based UAV Exploration through Perception-Coupled Mapping and Planning
by: Musil, Tomáš, et al.
Published: (2025)
by: Musil, Tomáš, et al.
Published: (2025)
MonoSLAM: Robust Monocular SLAM with Global Structure Optimization
by: Jiang, Bingzheng, et al.
Published: (2025)
by: Jiang, Bingzheng, et al.
Published: (2025)
GLEAM: Learning Generalizable Exploration Policy for Active Mapping in Complex 3D Indoor Scenes
by: Chen, Xiao, et al.
Published: (2025)
by: Chen, Xiao, et al.
Published: (2025)
Real-Time Monocular Scene Analysis for UAV in Outdoor Environments
by: AlaaEldin, Yara
Published: (2026)
by: AlaaEldin, Yara
Published: (2026)
Relationship-Aware Hierarchical 3D Scene Graph for Task Reasoning
by: Puigjaner, Albert Gassol, et al.
Published: (2026)
by: Puigjaner, Albert Gassol, et al.
Published: (2026)
MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
Incremental Semantics-Aided Meshing from LiDAR-Inertial Odometry and RGB Direct Label Transfer
by: Affan, Muhammad, et al.
Published: (2026)
by: Affan, Muhammad, et al.
Published: (2026)
Indoor and Outdoor 3D Scene Graph Generation via Language-Enabled Spatial Ontologies
by: Strader, Jared, et al.
Published: (2023)
by: Strader, Jared, et al.
Published: (2023)
Flash-Mono: Feed-Forward Accelerated Gaussian Splatting Monocular SLAM
by: Zhang, Zicheng, et al.
Published: (2026)
by: Zhang, Zicheng, et al.
Published: (2026)
MonoEM-GS: Monocular Expectation-Maximization Gaussian Splatting SLAM
by: Kruzhkov, Evgenii, et al.
Published: (2026)
by: Kruzhkov, Evgenii, et al.
Published: (2026)
A Survey on Monocular Re-Localization: From the Perspective of Scene Map Representation
by: Miao, Jinyu, et al.
Published: (2023)
by: Miao, Jinyu, et al.
Published: (2023)
Leveraging Computation of Expectation Models for Commonsense Affordance Estimation on 3D Scene Graphs
by: Saucedo, Mario A. V., et al.
Published: (2024)
by: Saucedo, Mario A. V., et al.
Published: (2024)
Hierarchical and Holistic Open-Vocabulary Functional 3D Scene Graphs for Indoor Spaces
by: Hu, Xinggang, et al.
Published: (2026)
by: Hu, Xinggang, et al.
Published: (2026)
MonoRace: Winning Champion-Level Drone Racing with Robust Monocular AI
by: Bahnam, Stavrow A., et al.
Published: (2026)
by: Bahnam, Stavrow A., et al.
Published: (2026)
MR-COGraphs: Communication-efficient Multi-Robot Open-vocabulary Mapping System via 3D Scene Graphs
by: Gu, Qiuyi, et al.
Published: (2024)
by: Gu, Qiuyi, et al.
Published: (2024)
SPADE: Towards Scalable Path Planning Architecture on Actionable Multi-Domain 3D Scene Graphs
by: Viswanathan, Vignesh Kottayam, et al.
Published: (2025)
by: Viswanathan, Vignesh Kottayam, et al.
Published: (2025)
Real-Time 3D Vision-Language Embedding Mapping
by: Rauch, Christian, et al.
Published: (2025)
by: Rauch, Christian, et al.
Published: (2025)
MonoMPC: Monocular Vision Based Navigation with Learned Collision Model and Risk-Aware Model Predictive Control
by: Sharma, Basant, et al.
Published: (2025)
by: Sharma, Basant, et al.
Published: (2025)
Traversability-aware Consistent Situational Graphs for Indoor Localization and Mapping
by: Kim, Jeewon, et al.
Published: (2025)
by: Kim, Jeewon, et al.
Published: (2025)
RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots
by: Modi, Giorgia, et al.
Published: (2026)
by: Modi, Giorgia, et al.
Published: (2026)
Semantic Region Aware Autonomous Exploration for Multi-Type Map Construction in Unknown Indoor Environments
by: Mao, Jianfang
Published: (2024)
by: Mao, Jianfang
Published: (2024)
Towards Terrain-Aware Task-Driven 3D Scene Graph Generation in Outdoor Environments
by: Samuelson, Chad R, et al.
Published: (2025)
by: Samuelson, Chad R, et al.
Published: (2025)
FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand
by: Maggio, Dominic, et al.
Published: (2026)
by: Maggio, Dominic, et al.
Published: (2026)
Similar Items
-
M2H-MX: Multi-Task Semantic and Geometric Perception for Real-Time Monocular 3D Scene Graph Construction
by: Udugama, U. V. B. L., et al.
Published: (2026) -
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
by: Udugama, U. V. B. L, et al.
Published: (2025) -
Mono-hydra: Real-time 3D scene graph construction from monocular camera input with IMU
by: Udugama, U. V. B. L., et al.
Published: (2023) -
A Comparison of Multi-View Stereo Methods for Photogrammetric 3D Reconstruction: From Traditional to Learning-Based Approaches
by: Li, Yawen, et al.
Published: (2026) -
ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imaging
by: Iz, Selim Ahmet, et al.
Published: (2026)