Reconstructing People, Places, and Cameras
Fuente:
arXiv
Saved in:
| Main Authors: | Müller, Lea, Choi, Hongsuk, Zhang, Anthony, Yi, Brent, Malik, Jitendra, Kanazawa, Angjoo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Human-level 3D shape perception emerges from multi-view learning
by: Bonnen, Tyler, et al.
Published: (2026)
by: Bonnen, Tyler, et al.
Published: (2026)
Synergy and Synchrony in Couple Dances
by: Maluleke, Vongani, et al.
Published: (2024)
by: Maluleke, Vongani, et al.
Published: (2024)
Estimating Body and Hand Motion in an Ego-sensed World
by: Yi, Brent, et al.
Published: (2024)
by: Yi, Brent, et al.
Published: (2024)
Visual Imitation Enables Contextual Humanoid Control
by: Allshire, Arthur, et al.
Published: (2025)
by: Allshire, Arthur, et al.
Published: (2025)
Cameras as Relative Positional Encoding
by: Li, Ruilong, et al.
Published: (2025)
by: Li, Ruilong, et al.
Published: (2025)
Diffusion Forcing for Multi-Agent Interaction Sequence Modeling
by: Maluleke, Vongani H., et al.
Published: (2025)
by: Maluleke, Vongani H., et al.
Published: (2025)
Viser: Imperative, Web-based 3D Visualization in Python
by: Yi, Brent, et al.
Published: (2025)
by: Yi, Brent, et al.
Published: (2025)
Predict-Optimize-Distill: A Self-Improving Cycle for 4D Object Understanding
by: Wu, Mingxuan, et al.
Published: (2025)
by: Wu, Mingxuan, et al.
Published: (2025)
Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction
by: Kerr, Justin, et al.
Published: (2024)
by: Kerr, Justin, et al.
Published: (2024)
Splatfacto-W: A Nerfstudio Implementation of Gaussian Splatting for Unconstrained Photo Collections
by: Xu, Congrong, et al.
Published: (2024)
by: Xu, Congrong, et al.
Published: (2024)
SOAR: Self-Occluded Avatar Recovery from a Single Video In the Wild
by: Pan, Zhuoyang, et al.
Published: (2024)
by: Pan, Zhuoyang, et al.
Published: (2024)
HandNeRF: Learning to Reconstruct Hand-Object Interaction Scene from a Single RGB Image
by: Choi, Hongsuk, et al.
Published: (2023)
by: Choi, Hongsuk, et al.
Published: (2023)
Shape of Motion: 4D Reconstruction from a Single Video
by: Wang, Qianqian, et al.
Published: (2024)
by: Wang, Qianqian, et al.
Published: (2024)
Generating Continual Human Motion in Diverse 3D Scenes
by: Mir, Aymen, et al.
Published: (2023)
by: Mir, Aymen, et al.
Published: (2023)
St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
by: Feng, Haiwen, et al.
Published: (2025)
by: Feng, Haiwen, et al.
Published: (2025)
Continuous 3D Perception Model with Persistent State
by: Wang, Qianqian, et al.
Published: (2025)
by: Wang, Qianqian, et al.
Published: (2025)
Eye, Robot: Learning to Look to Act with a BC-RL Perception-Action Loop
by: Kerr, Justin, et al.
Published: (2025)
by: Kerr, Justin, et al.
Published: (2025)
The More You See in 2D, the More You Perceive in 3D
by: Han, Xinyang, et al.
Published: (2024)
by: Han, Xinyang, et al.
Published: (2024)
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
by: Yang, Gengshan, et al.
Published: (2024)
by: Yang, Gengshan, et al.
Published: (2024)
Reconstructing Hand-Held Objects in 3D from Images and Videos
by: Wu, Jane, et al.
Published: (2024)
by: Wu, Jane, et al.
Published: (2024)
TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos
by: Liu, Jinpeng, et al.
Published: (2026)
by: Liu, Jinpeng, et al.
Published: (2026)
Fillerbuster: Unified Generative Scene Completion Model for Casual Captures
by: Weber, Ethan, et al.
Published: (2025)
by: Weber, Ethan, et al.
Published: (2025)
NeRF-XL: Scaling NeRFs with Multiple GPUs
by: Li, Ruilong, et al.
Published: (2024)
by: Li, Ruilong, et al.
Published: (2024)
Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind
by: Plizzari, Chiara, et al.
Published: (2024)
by: Plizzari, Chiara, et al.
Published: (2024)
gsplat: An Open-Source Library for Gaussian Splatting
by: Ye, Vickie, et al.
Published: (2024)
by: Ye, Vickie, et al.
Published: (2024)
Segment Any Motion in Videos
by: Huang, Nan, et al.
Published: (2025)
by: Huang, Nan, et al.
Published: (2025)
Toon3D: Seeing Cartoons from New Perspectives
by: Weber, Ethan, et al.
Published: (2024)
by: Weber, Ethan, et al.
Published: (2024)
From Audio to Photoreal Embodiment: Synthesizing Humans in Conversations
by: Ng, Evonne, et al.
Published: (2024)
by: Ng, Evonne, et al.
Published: (2024)
Synthesizing Moving People with 3D Control
by: Li, Boyi, et al.
Published: (2024)
by: Li, Boyi, et al.
Published: (2024)
Decentralized Diffusion Models
by: McAllister, David, et al.
Published: (2025)
by: McAllister, David, et al.
Published: (2025)
GARField: Group Anything with Radiance Fields
by: Kim, Chung Min, et al.
Published: (2024)
by: Kim, Chung Min, et al.
Published: (2024)
Self-Improving 4D Perception via Self-Distillation
by: Huang, Nan, et al.
Published: (2026)
by: Huang, Nan, et al.
Published: (2026)
CameraHMR: Aligning People with Perspective
by: Patel, Priyanka, et al.
Published: (2024)
by: Patel, Priyanka, et al.
Published: (2024)
Adaptive Human Trajectory Prediction via Latent Corridors
by: Thakkar, Neerja, et al.
Published: (2023)
by: Thakkar, Neerja, et al.
Published: (2023)
MegaSaM: Accurate, Fast, and Robust Structure and Motion from Casual Dynamic Videos
by: Li, Zhengqi, et al.
Published: (2024)
by: Li, Zhengqi, et al.
Published: (2024)
Learning Visuotactile Skills with Two Multifingered Hands
by: Lin, Toru, et al.
Published: (2024)
by: Lin, Toru, et al.
Published: (2024)
Tracking by Predicting 3-D Gaussians Over Time
by: Baranwal, Tanish, et al.
Published: (2025)
by: Baranwal, Tanish, et al.
Published: (2025)
Spatiotemporal Multi-Camera Calibration using Freely Moving People
by: Lee, Sang-Eun, et al.
Published: (2025)
by: Lee, Sang-Eun, et al.
Published: (2025)
Ensemble-Based Event Camera Place Recognition Under Varying Illumination
by: Joseph, Therese, et al.
Published: (2025)
by: Joseph, Therese, et al.
Published: (2025)
Poly-Autoregressive Prediction for Modeling Interactions
by: Thakkar, Neerja, et al.
Published: (2025)
by: Thakkar, Neerja, et al.
Published: (2025)
Similar Items
-
Human-level 3D shape perception emerges from multi-view learning
by: Bonnen, Tyler, et al.
Published: (2026) -
Synergy and Synchrony in Couple Dances
by: Maluleke, Vongani, et al.
Published: (2024) -
Estimating Body and Hand Motion in an Ego-sensed World
by: Yi, Brent, et al.
Published: (2024) -
Visual Imitation Enables Contextual Humanoid Control
by: Allshire, Arthur, et al.
Published: (2025) -
Cameras as Relative Positional Encoding
by: Li, Ruilong, et al.
Published: (2025)