Safe-Construct: Redefining Construction Safety Violation Recognition as 3D Multi-View Engagement Task
Fuente:
arXiv
Saved in:
| Main Authors: | Chharia, Aviral, Ren, Tianyu, Furuhata, Tomotake, Shimada, Kenji |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MV-SSM: Multi-View State Space Modeling for 3D Human Pose Estimation
by: Chharia, Aviral, et al.
Published: (2025)
by: Chharia, Aviral, et al.
Published: (2025)
Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation
by: Chharia, Aviral, et al.
Published: (2026)
by: Chharia, Aviral, et al.
Published: (2026)
Hamba: Single-view 3D Hand Reconstruction with Graph-guided Bi-Scanning Mamba
by: Dong, Haoye, et al.
Published: (2024)
by: Dong, Haoye, et al.
Published: (2024)
Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
by: Udugama, U. V. B. L., et al.
Published: (2026)
by: Udugama, U. V. B. L., et al.
Published: (2026)
Human-VDM: Learning Single-Image 3D Human Gaussian Splatting from Video Diffusion Models
by: Liu, Zhibin, et al.
Published: (2024)
by: Liu, Zhibin, et al.
Published: (2024)
A Computer Vision Approach for Autonomous Cars to Drive Safe at Construction Zone
by: Ahammed, Abu Shad, et al.
Published: (2024)
by: Ahammed, Abu Shad, et al.
Published: (2024)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
by: Wang, Jiasen, et al.
Published: (2024)
by: Wang, Jiasen, et al.
Published: (2024)
DSM: Constructing a Diverse Semantic Map for 3D Visual Grounding
by: Xie, Qinghongbing, et al.
Published: (2025)
by: Xie, Qinghongbing, et al.
Published: (2025)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
by: Mahdavian, Mohammad, et al.
Published: (2026)
by: Mahdavian, Mohammad, et al.
Published: (2026)
SparseGrasp: Robotic Grasping via 3D Semantic Gaussian Splatting from Sparse Multi-View RGB Images
by: Yu, Junqiu, et al.
Published: (2024)
by: Yu, Junqiu, et al.
Published: (2024)
LS-HAR: Language Supervised Human Action Recognition with Salient Fusion, Construction Sites as a Use-Case
by: Mahdavian, Mohammad, et al.
Published: (2024)
by: Mahdavian, Mohammad, et al.
Published: (2024)
MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning
by: Xu, Tianyu, et al.
Published: (2025)
by: Xu, Tianyu, et al.
Published: (2025)
MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition
by: Xu, Zhengyi, et al.
Published: (2026)
by: Xu, Zhengyi, et al.
Published: (2026)
SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection
by: Papais, Sandro, et al.
Published: (2026)
by: Papais, Sandro, et al.
Published: (2026)
Segmentation Dataset for Reinforced Concrete Construction
by: Schmidt, Patrick, et al.
Published: (2024)
by: Schmidt, Patrick, et al.
Published: (2024)
Robotic Arm Platform for Multi-View Image Acquisition and 3D Reconstruction in Minimally Invasive Surgery
by: Saikia, Alexander, et al.
Published: (2024)
by: Saikia, Alexander, et al.
Published: (2024)
Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition
by: Shang, Tianyi, et al.
Published: (2025)
by: Shang, Tianyi, et al.
Published: (2025)
LiDAR-EVS: Enhance Extrapolated View Synthesis for 3D Gaussian Splatting with Pseudo-LiDAR Supervision
by: Huang, Yiming, et al.
Published: (2026)
by: Huang, Yiming, et al.
Published: (2026)
FastOcc: Accelerating 3D Occupancy Prediction by Fusing the 2D Bird's-Eye View and Perspective View
by: Hou, Jiawei, et al.
Published: (2024)
by: Hou, Jiawei, et al.
Published: (2024)
Systematic Evaluation of Novel View Synthesis for Video Place Recognition
by: Mahmud, Muhammad Zawad, et al.
Published: (2026)
by: Mahmud, Muhammad Zawad, et al.
Published: (2026)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
SLAM for Indoor Mapping of Wide Area Construction Environments
by: Ress, Vincent, et al.
Published: (2024)
by: Ress, Vincent, et al.
Published: (2024)
LM-MCVT: A Lightweight Multi-modal Multi-view Convolutional-Vision Transformer Approach for 3D Object Recognition
by: Xiong, Songsong, et al.
Published: (2025)
by: Xiong, Songsong, et al.
Published: (2025)
Efficient Multi-Task Scene Analysis with RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
MrGS: Multi-modal Radiance Fields with 3D Gaussian Splatting for RGB-Thermal Novel View Synthesis
by: Kweon, Minseong, et al.
Published: (2025)
by: Kweon, Minseong, et al.
Published: (2025)
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025)
by: Kim, Young Hun, et al.
Published: (2025)
Monocular Visual Place Recognition in LiDAR Maps via Cross-Modal State Space Model and Multi-View Matching
by: Yao, Gongxin, et al.
Published: (2024)
by: Yao, Gongxin, et al.
Published: (2024)
InspecSafe-V1: A Multimodal Benchmark for Safety Assessment in Industrial Inspection Scenarios
by: Liu, Zeyi, et al.
Published: (2026)
by: Liu, Zeyi, et al.
Published: (2026)
LiDAR-BEVMTN: Real-Time LiDAR Bird's-Eye View Multi-Task Perception Network for Autonomous Driving
by: Mohapatra, Sambit, et al.
Published: (2023)
by: Mohapatra, Sambit, et al.
Published: (2023)
When LLMs step into the 3D World: A Survey and Meta-Analysis of 3D Tasks via Multi-modal Large Language Models
by: Ma, Xianzheng, et al.
Published: (2024)
by: Ma, Xianzheng, et al.
Published: (2024)
DuoSpaceNet: Leveraging Both Bird's-Eye-View and Perspective View Representations for 3D Object Detection
by: Huang, Zhe, et al.
Published: (2024)
by: Huang, Zhe, et al.
Published: (2024)
Scalable 3D Registration via Truncated Entry-wise Absolute Residuals
by: Huang, Tianyu, et al.
Published: (2024)
by: Huang, Tianyu, et al.
Published: (2024)
Are Open-Vocabulary Models Ready for Detection of MEP Elements on Construction Sites
by: Abdalwhab, Abdalwhab, et al.
Published: (2025)
by: Abdalwhab, Abdalwhab, et al.
Published: (2025)
GrowSplat: Constructing Temporal Digital Twins of Plants with Gaussian Splats
by: Adebola, Simeon, et al.
Published: (2025)
by: Adebola, Simeon, et al.
Published: (2025)
BIM-Constrained Optimization for Accurate Localization and Deviation Correction in Construction Monitoring
by: Bikandi-Noya, Asier, et al.
Published: (2025)
by: Bikandi-Noya, Asier, et al.
Published: (2025)
Direct Robot Configuration Space Construction using Convolutional Encoder-Decoders
by: Benka, Christopher, et al.
Published: (2023)
by: Benka, Christopher, et al.
Published: (2023)
Impact of Localization Errors on Label Quality for Online HD Map Construction
by: Blumberg, Alexander, et al.
Published: (2026)
by: Blumberg, Alexander, et al.
Published: (2026)
Active 6D Pose Estimation for Textureless Objects using Multi-View RGB Frames
by: Yang, Jun, et al.
Published: (2025)
by: Yang, Jun, et al.
Published: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
by: Bai, Yongjie, et al.
Published: (2025)
by: Bai, Yongjie, et al.
Published: (2025)
Boundary Exploration of Next Best View Policy in 3D Robotic Scanning
by: Li, Leihui, et al.
Published: (2024)
by: Li, Leihui, et al.
Published: (2024)
Similar Items
-
MV-SSM: Multi-View State Space Modeling for 3D Human Pose Estimation
by: Chharia, Aviral, et al.
Published: (2025) -
Multi-view Consistent 3D Gaussian Head Avatars 'without' Multi-view Generation
by: Chharia, Aviral, et al.
Published: (2026) -
Hamba: Single-view 3D Hand Reconstruction with Graph-guided Bi-Scanning Mamba
by: Dong, Haoye, et al.
Published: (2024) -
Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
by: Udugama, U. V. B. L., et al.
Published: (2026) -
Human-VDM: Learning Single-Image 3D Human Gaussian Splatting from Video Diffusion Models
by: Liu, Zhibin, et al.
Published: (2024)