Patient4D: Temporally Consistent Patient Body Mesh Recovery from Monocular Operating Room Video
Fuente:
arXiv
Saved in:
| Main Authors: | Tu, Mingxiao, Jung, Hoijoon, Moghadam, Alireza, Kyme, Andre, Kim, Jinman |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-modal 3D Pose and Shape Estimation with Computed Tomography
by: Tu, Mingxiao, et al.
Published: (2025)
by: Tu, Mingxiao, et al.
Published: (2025)
SMPL-GPTexture: Dual-View 3D Human Texture Estimation using Text-to-Image Generation Models
by: Tu, Mingxiao, et al.
Published: (2025)
by: Tu, Mingxiao, et al.
Published: (2025)
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026)
by: Cen, Jiaxin, et al.
Published: (2026)
Monocular Mesh Recovery and Body Measurement of Female Saanen Goats
by: Jin, Bo, et al.
Published: (2026)
by: Jin, Bo, et al.
Published: (2026)
DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
SAM-Body4D: Training-Free 4D Human Body Mesh Recovery from Videos
by: Gao, Mingqi, et al.
Published: (2025)
by: Gao, Mingqi, et al.
Published: (2025)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
M3DHMR: Monocular 3D Hand Mesh Recovery
by: Lin, Yihong, et al.
Published: (2025)
by: Lin, Yihong, et al.
Published: (2025)
3D Hand Mesh Recovery from Monocular RGB in Camera Space
by: Li, Haonan, et al.
Published: (2024)
by: Li, Haonan, et al.
Published: (2024)
SAM 3D Body: Robust Full-Body Human Mesh Recovery
by: Yang, Xitong, et al.
Published: (2026)
by: Yang, Xitong, et al.
Published: (2026)
STAF: 3D Human Mesh Recovery from Video with Spatio-Temporal Alignment Fusion
by: Yao, Wei, et al.
Published: (2024)
by: Yao, Wei, et al.
Published: (2024)
Reinforcing Consistency in Video MLLMs with Structured Rewards
by: Quan, Yihao, et al.
Published: (2026)
by: Quan, Yihao, et al.
Published: (2026)
Monocular Models are Strong Learners for Multi-View Human Mesh Recovery
by: Xie, Haoyu, et al.
Published: (2026)
by: Xie, Haoyu, et al.
Published: (2026)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
STATIC : Surface Temporal Affine for TIme Consistency in Video Monocular Depth Estimation
by: Yang, Sunghun, et al.
Published: (2024)
by: Yang, Sunghun, et al.
Published: (2024)
MetricHMSR:Metric Human Mesh and Scene Recovery from Monocular Images
by: Song, Chentao, et al.
Published: (2025)
by: Song, Chentao, et al.
Published: (2025)
PTC-Depth: Pose-Refined Monocular Depth Estimation with Temporal Consistency
by: Han, Leezy, et al.
Published: (2026)
by: Han, Leezy, et al.
Published: (2026)
Fast SAM 3D Body: Accelerating SAM 3D Body for Real-Time Full-Body Human Mesh Recovery
by: Yang, Timing, et al.
Published: (2026)
by: Yang, Timing, et al.
Published: (2026)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
HyDAR-Pano3D: A Hybrid Disentangled Anatomical Recovery Framework for Panoramic-to-3D Reconstruction
by: Yue, Yaoyao, et al.
Published: (2026)
by: Yue, Yaoyao, et al.
Published: (2026)
Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos
by: Wei, Dingkun, et al.
Published: (2026)
by: Wei, Dingkun, et al.
Published: (2026)
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
by: Yao, Wei, et al.
Published: (2023)
by: Yao, Wei, et al.
Published: (2023)
MonoSAOD: Monocular 3D Object Detection with Sparsely Annotated Label
by: Jung, Junyoung, et al.
Published: (2026)
by: Jung, Junyoung, et al.
Published: (2026)
Tessellation GS: Neural Mesh Gaussians for Robust Monocular Reconstruction of Dynamic Objects
by: Tao, Shuohan, et al.
Published: (2025)
by: Tao, Shuohan, et al.
Published: (2025)
Egocentric Whole-Body Human Mesh Recovery with Prior-Guided Learning
by: Na, Soyeon, et al.
Published: (2026)
by: Na, Soyeon, et al.
Published: (2026)
On the Consistency of Video Large Language Models in Temporal Comprehension
by: Jung, Minjoon, et al.
Published: (2024)
by: Jung, Minjoon, et al.
Published: (2024)
Transferring Relative Monocular Depth to Surgical Vision with Temporal Consistency
by: Budd, Charlie, et al.
Published: (2024)
by: Budd, Charlie, et al.
Published: (2024)
GVT2RPM: An Empirical Study for General Video Transformer Adaptation to Remote Physiological Measurement
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Multi-RoI Human Mesh Recovery with Camera Consistency and Contrastive Losses
by: Nie, Yongwei, et al.
Published: (2024)
by: Nie, Yongwei, et al.
Published: (2024)
GP-4DGS: Probabilistic 4D Gaussian Splatting from Monocular Video via Variational Gaussian Processes
by: Kim, Mijeong, et al.
Published: (2026)
by: Kim, Mijeong, et al.
Published: (2026)
Ctrl-Room: Controllable Text-to-3D Room Meshes Generation with Layout Constraints
by: Fang, Chuan, et al.
Published: (2023)
by: Fang, Chuan, et al.
Published: (2023)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
CT4D: Consistent Text-to-4D Generation with Animatable Meshes
by: Chen, Ce, et al.
Published: (2024)
by: Chen, Ce, et al.
Published: (2024)
Geometry-guided Online 3D Video Synthesis with Multi-View Temporal Consistency
by: Ha, Hyunho, et al.
Published: (2025)
by: Ha, Hyunho, et al.
Published: (2025)
DreamMesh4D: Video-to-4D Generation with Sparse-Controlled Gaussian-Mesh Hybrid Representation
by: Li, Zhiqi, et al.
Published: (2024)
by: Li, Zhiqi, et al.
Published: (2024)
AR4D: Autoregressive 4D Generation from Monocular Videos
by: Zhu, Hanxin, et al.
Published: (2025)
by: Zhu, Hanxin, et al.
Published: (2025)
Divide and Fuse: Body Part Mesh Recovery from Partially Visible Human Images
by: Luan, Tianyu, et al.
Published: (2024)
by: Luan, Tianyu, et al.
Published: (2024)
PanORama: Multiview Consistent Panoptic Segmentation in Operating Rooms
by: Gürbüz, Tuna, et al.
Published: (2026)
by: Gürbüz, Tuna, et al.
Published: (2026)
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
by: Li, Mingxiao, et al.
Published: (2025)
by: Li, Mingxiao, et al.
Published: (2025)
Drive Any Mesh: 4D Latent Diffusion for Mesh Deformation from Video
by: Shi, Yahao, et al.
Published: (2025)
by: Shi, Yahao, et al.
Published: (2025)
Similar Items
-
Multi-modal 3D Pose and Shape Estimation with Computed Tomography
by: Tu, Mingxiao, et al.
Published: (2025) -
SMPL-GPTexture: Dual-View 3D Human Texture Estimation using Text-to-Image Generation Models
by: Tu, Mingxiao, et al.
Published: (2025) -
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026) -
Monocular Mesh Recovery and Body Measurement of Female Saanen Goats
by: Jin, Bo, et al.
Published: (2026) -
DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos
by: Shen, Wenhao, et al.
Published: (2026)