Coherent Human-Scene Reconstruction from Multi-Person Multi-View Video in a Single Pass
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Sangmin, Hwang, Minhyuk, Cha, Geonho, Wee, Dongyoon, Park, Jaesik |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Humans as a Calibration Pattern: Dynamic 3D Scene Reconstruction from Unsynchronized and Uncalibrated Videos
by: Choi, Changwoon, et al.
Published: (2024)
by: Choi, Changwoon, et al.
Published: (2024)
Regularizing Dynamic Radiance Fields with Kinematic Fields
by: Im, Woobin, et al.
Published: (2024)
by: Im, Woobin, et al.
Published: (2024)
HOIGS: Human-Object Interaction Gaussian Splatting
by: Kim, Taewoo, et al.
Published: (2026)
by: Kim, Taewoo, et al.
Published: (2026)
CoCoGaussian: Leveraging Circle of Confusion for Gaussian Splatting from Defocused Images
by: Lee, Jungho, et al.
Published: (2024)
by: Lee, Jungho, et al.
Published: (2024)
SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models
by: Chung, Jiwoo, et al.
Published: (2026)
by: Chung, Jiwoo, et al.
Published: (2026)
3D Reconstruction of Interacting Multi-Person in Clothing from a Single Image
by: Cha, Junuk, et al.
Published: (2024)
by: Cha, Junuk, et al.
Published: (2024)
EditSplat: Multi-View Fusion and Attention-Guided Optimization for View-Consistent 3D Scene Editing with 3D Gaussian Splatting
by: Lee, Dong In, et al.
Published: (2024)
by: Lee, Dong In, et al.
Published: (2024)
MultiEgo: A Multi-View Egocentric Video Dataset for 4D Scene Reconstruction
by: Li, Bate, et al.
Published: (2025)
by: Li, Bate, et al.
Published: (2025)
Metropolis-Hastings Sampling for 3D Gaussian Reconstruction
by: Kim, Hyunjin, et al.
Published: (2025)
by: Kim, Hyunjin, et al.
Published: (2025)
MV2Cyl: Reconstructing 3D Extrusion Cylinders from Multi-View Images
by: Hong, Eunji, et al.
Published: (2024)
by: Hong, Eunji, et al.
Published: (2024)
Motion-Oriented Compositional Neural Radiance Fields for Monocular Dynamic Human Modeling
by: Kim, Jaehyeok, et al.
Published: (2024)
by: Kim, Jaehyeok, et al.
Published: (2024)
Video-Oasis: Rethinking Evaluation of Video Understanding
by: Lim, Geuntaek, et al.
Published: (2026)
by: Lim, Geuntaek, et al.
Published: (2026)
Dense Dynamic Scene Reconstruction and Camera Pose Estimation from Multi-View Videos
by: Sun, Shuo, et al.
Published: (2026)
by: Sun, Shuo, et al.
Published: (2026)
Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs
by: Hyun, Jeongseok, et al.
Published: (2025)
by: Hyun, Jeongseok, et al.
Published: (2025)
SEAL-pose: Enhancing 3D Human Pose Estimation via a Learned Loss for Structural Consistency
by: Kim, Yeonsung, et al.
Published: (2026)
by: Kim, Yeonsung, et al.
Published: (2026)
MV-TAP: Tracking Any Point in Multi-View Videos
by: Koo, Jahyeok, et al.
Published: (2025)
by: Koo, Jahyeok, et al.
Published: (2025)
ControlFace: Harnessing Facial Parametric Control for Face Rigging
by: Jang, Wooseok, et al.
Published: (2024)
by: Jang, Wooseok, et al.
Published: (2024)
FreeTimeGS++: Secrets of Dynamic Gaussian Splatting and Their Principles
by: Lee, Lucas Yunkyu, et al.
Published: (2026)
by: Lee, Lucas Yunkyu, et al.
Published: (2026)
Deep Cost Ray Fusion for Sparse Depth Video Completion
by: Kim, Jungeon, et al.
Published: (2024)
by: Kim, Jungeon, et al.
Published: (2024)
Direct Reward Fine-Tuning on Poses for Single Image to 3D Human in the Wild
by: Do, Seunguk, et al.
Published: (2026)
by: Do, Seunguk, et al.
Published: (2026)
Learning Multi-View Spatial Reasoning from Cross-View Relations
by: Jeong, Suchae, et al.
Published: (2026)
by: Jeong, Suchae, et al.
Published: (2026)
Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing
by: Shin, Joonghyuk, et al.
Published: (2025)
by: Shin, Joonghyuk, et al.
Published: (2025)
UniSH: Unifying Scene and Human Reconstruction in a Feed-Forward Pass
by: Li, Mengfei, et al.
Published: (2026)
by: Li, Mengfei, et al.
Published: (2026)
Seeing Beyond Views: Multi-View Driving Scene Video Generation with Holistic Attention
by: Lu, Hannan, et al.
Published: (2024)
by: Lu, Hannan, et al.
Published: (2024)
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
by: Huang, Buzhen, et al.
Published: (2024)
by: Huang, Buzhen, et al.
Published: (2024)
A Simple Baseline with Single-encoder for Referring Image Segmentation
by: Yu, Seonghoon, et al.
Published: (2024)
by: Yu, Seonghoon, et al.
Published: (2024)
Recovering Dynamic 3D Sketches from Videos
by: Lee, Jaeah, et al.
Published: (2025)
by: Lee, Jaeah, et al.
Published: (2025)
Fast Underwater Scene Reconstruction using Multi-View Stereo and Physical Imaging
by: Hu, Shuyi, et al.
Published: (2025)
by: Hu, Shuyi, et al.
Published: (2025)
Classification Matters: Improving Video Action Detection with Class-Specific Attention
by: Lee, Jinsung, et al.
Published: (2024)
by: Lee, Jinsung, et al.
Published: (2024)
Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation
by: Han, Su Ho, et al.
Published: (2025)
by: Han, Su Ho, et al.
Published: (2025)
ActMVS: Active Scene Reconstruction with Monocular Multi-View Stereo
by: Pu, Guo, et al.
Published: (2026)
by: Pu, Guo, et al.
Published: (2026)
Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval
by: Ko, Dohwan, et al.
Published: (2025)
by: Ko, Dohwan, et al.
Published: (2025)
GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction
by: Schmid, Katharina, et al.
Published: (2026)
by: Schmid, Katharina, et al.
Published: (2026)
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
by: Wu, Guile, et al.
Published: (2025)
by: Wu, Guile, et al.
Published: (2025)
Single-View Scene Point Cloud Human Grasp Generation
by: Wang, Yan-Kang, et al.
Published: (2024)
by: Wang, Yan-Kang, et al.
Published: (2024)
360 in the Wild: Dataset for Depth Prediction and View Synthesis
by: Park, Kibaek, et al.
Published: (2024)
by: Park, Kibaek, et al.
Published: (2024)
Learning from Synchronization: Self-Supervised Uncalibrated Multi-View Person Association in Challenging Scenes
by: Chen, Keqi, et al.
Published: (2025)
by: Chen, Keqi, et al.
Published: (2025)
SceneCrafter: Controllable Multi-View Driving Scene Editing
by: Zhu, Zehao, et al.
Published: (2025)
by: Zhu, Zehao, et al.
Published: (2025)
Leveraging Temporal Contextualization for Video Action Recognition
by: Kim, Minji, et al.
Published: (2024)
by: Kim, Minji, et al.
Published: (2024)
VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling
by: Go, Hyojun, et al.
Published: (2025)
by: Go, Hyojun, et al.
Published: (2025)
Similar Items
-
Humans as a Calibration Pattern: Dynamic 3D Scene Reconstruction from Unsynchronized and Uncalibrated Videos
by: Choi, Changwoon, et al.
Published: (2024) -
Regularizing Dynamic Radiance Fields with Kinematic Fields
by: Im, Woobin, et al.
Published: (2024) -
HOIGS: Human-Object Interaction Gaussian Splatting
by: Kim, Taewoo, et al.
Published: (2026) -
CoCoGaussian: Leveraging Circle of Confusion for Gaussian Splatting from Defocused Images
by: Lee, Jungho, et al.
Published: (2024) -
SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models
by: Chung, Jiwoo, et al.
Published: (2026)