GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Gwanghyun, Li, Xueting, Yuan, Ye, Nagano, Koki, Li, Tianye, Kautz, Jan, Chun, Se Young, Iqbal, Umar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdaHuman: Animatable Detailed 3D Human Generation with Compositional Multiview Diffusion
by: Huang, Yangyi, et al.
Published: (2025)
by: Huang, Yangyi, et al.
Published: (2025)
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
by: Jeong, Wongi, et al.
Published: (2026)
by: Jeong, Wongi, et al.
Published: (2026)
GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning
by: Yuan, Ye, et al.
Published: (2023)
by: Yuan, Ye, et al.
Published: (2023)
Coherent3D: Coherent 3D Portrait Video Reconstruction via Triplane Fusion
by: Wang, Shengze, et al.
Published: (2024)
by: Wang, Shengze, et al.
Published: (2024)
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
by: Kim, Heechang, et al.
Published: (2025)
by: Kim, Heechang, et al.
Published: (2025)
Dream, Lift, Animate: From Single Images to Animatable Gaussian Avatars
by: Bühler, Marcel C., et al.
Published: (2025)
by: Bühler, Marcel C., et al.
Published: (2025)
DC-VSR: Spatially and Temporally Consistent Video Super-Resolution with Video Diffusion Prior
by: Han, Janghyeok, et al.
Published: (2025)
by: Han, Janghyeok, et al.
Published: (2025)
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
by: Jeong, Wongi, et al.
Published: (2025)
by: Jeong, Wongi, et al.
Published: (2025)
MOST: MR reconstruction Optimization for multiple downStream Tasks via continual learning
by: Jeong, Hwihun, et al.
Published: (2024)
by: Jeong, Hwihun, et al.
Published: (2024)
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024)
by: Kim, Gwanghyun, et al.
Published: (2024)
Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition
by: Kim, Sungnyun, et al.
Published: (2024)
by: Kim, Sungnyun, et al.
Published: (2024)
Improving Temporal Consistency and Fidelity at Inference-time in Perceptual Video Restoration by Zero-shot Image-based Diffusion Models
by: Rahimi, Nasrin, et al.
Published: (2025)
by: Rahimi, Nasrin, et al.
Published: (2025)
COIN: Control-Inpainting Diffusion Prior for Human and Camera Motion Estimation
by: Li, Jiefeng, et al.
Published: (2024)
by: Li, Jiefeng, et al.
Published: (2024)
Generative Human Video Compression with Multi-granularity Temporal Trajectory Factorization
by: Yin, Shanzhi, et al.
Published: (2024)
by: Yin, Shanzhi, et al.
Published: (2024)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024)
by: Kim, Gwanghyun, et al.
Published: (2024)
Adaptive Selection of Sampling-Reconstruction in Fourier Compressed Sensing
by: Hong, Seongmin, et al.
Published: (2024)
by: Hong, Seongmin, et al.
Published: (2024)
Efficient and robust 3D blind harmonization for large domain gaps
by: Jeong, Hwihun, et al.
Published: (2025)
by: Jeong, Hwihun, et al.
Published: (2025)
MoCha-LD: Temporal-Chunked Latent Diffusion with Motion-Aware Consistency Regularization for Long Video Generation
by: Jin, Haopeng
Published: (2026)
by: Jin, Haopeng
Published: (2026)
Semantic and Temporal Integration in Latent Diffusion Space for High-Fidelity Video Super-Resolution
by: Wang, Yiwen, et al.
Published: (2025)
by: Wang, Yiwen, et al.
Published: (2025)
HyTIP: Hybrid Temporal Information Propagation for Masked Conditional Residual Video Coding
by: Chen, Yi-Hsin, et al.
Published: (2025)
by: Chen, Yi-Hsin, et al.
Published: (2025)
A Diffusion-Driven Temporal Super-Resolution and Spatial Consistency Enhancement Framework for 4D MRI imaging
by: Zhou, Xuanru, et al.
Published: (2025)
by: Zhou, Xuanru, et al.
Published: (2025)
MediViSTA: Medical Video Segmentation via Temporal Fusion SAM Adaptation for Echocardiography
by: Kim, Sekeun, et al.
Published: (2023)
by: Kim, Sekeun, et al.
Published: (2023)
Convergent Complex Quasi-Newton Proximal Methods for Gradient-Driven Denoisers in Compressed Sensing MRI Reconstruction
by: Hong, Tao, et al.
Published: (2025)
by: Hong, Tao, et al.
Published: (2025)
Fast and accurate sparse-view CBCT reconstruction using meta-learned neural attenuation field and hash-encoding regularization
by: Shin, Heejun, et al.
Published: (2023)
by: Shin, Heejun, et al.
Published: (2023)
DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction
by: Li, Siyu, et al.
Published: (2024)
by: Li, Siyu, et al.
Published: (2024)
A Target-Free Harmonization Method for MRI
by: Kim, Minjun, et al.
Published: (2026)
by: Kim, Minjun, et al.
Published: (2026)
Spatial Decomposition and Temporal Fusion based Inter Prediction for Learned Video Compression
by: Sheng, Xihua, et al.
Published: (2024)
by: Sheng, Xihua, et al.
Published: (2024)
Deep Internal Learning: Deep Learning from a Single Input
by: Tirer, Tom, et al.
Published: (2023)
by: Tirer, Tom, et al.
Published: (2023)
Diffusion-based Perceptual Neural Video Compression with Temporal Diffusion Information Reuse
by: Ma, Wenzhuo, et al.
Published: (2025)
by: Ma, Wenzhuo, et al.
Published: (2025)
TVRN: Invertible Neural Networks for Compression-Aware Temporal Video Rescaling
by: Feng, Xinmin, et al.
Published: (2026)
by: Feng, Xinmin, et al.
Published: (2026)
Identity-Consistent Diffusion Network for Grading Knee Osteoarthritis Progression in Radiographic Imaging
by: Wu, Wenhua, et al.
Published: (2024)
by: Wu, Wenhua, et al.
Published: (2024)
Spatio-Temporal Representation Decoupling and Enhancement for Federated Instrument Segmentation in Surgical Videos
by: Fang, Zheng, et al.
Published: (2025)
by: Fang, Zheng, et al.
Published: (2025)
NeuralLVC: Neural Lossless Video Compression via Masked Diffusion with Temporal Conditioning
by: Uricchio, Tiberio, et al.
Published: (2026)
by: Uricchio, Tiberio, et al.
Published: (2026)
FreqPrior: Improving Video Diffusion Models with Frequency Filtering Gaussian Noise
by: Yuan, Yunlong, et al.
Published: (2025)
by: Yuan, Yunlong, et al.
Published: (2025)
Sampling-Pattern-Agnostic MRI Reconstruction through Adaptive Consistency Enforcement with Diffusion Model
by: Malyala, Anurag, et al.
Published: (2024)
by: Malyala, Anurag, et al.
Published: (2024)
Image Motion Blur Removal in the Temporal Dimension with Video Diffusion Models
by: Pang, Wang, et al.
Published: (2025)
by: Pang, Wang, et al.
Published: (2025)
A Multi-Scale Spatial-Temporal Network for Wireless Video Transmission
by: Zhou, Xinyi, et al.
Published: (2024)
by: Zhou, Xinyi, et al.
Published: (2024)
Spatial-Temporal Multi-level Association for Video Object Segmentation
by: Miao, Deshui, et al.
Published: (2024)
by: Miao, Deshui, et al.
Published: (2024)
Wireless Semantic Communications for Video Conferencing
by: Jiang, Peiwen, et al.
Published: (2022)
by: Jiang, Peiwen, et al.
Published: (2022)
OD-VAE: An Omni-dimensional Video Compressor for Improving Latent Video Diffusion Model
by: Chen, Liuhan, et al.
Published: (2024)
by: Chen, Liuhan, et al.
Published: (2024)
Similar Items
-
AdaHuman: Animatable Detailed 3D Human Generation with Compositional Multiview Diffusion
by: Huang, Yangyi, et al.
Published: (2025) -
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
by: Jeong, Wongi, et al.
Published: (2026) -
GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning
by: Yuan, Ye, et al.
Published: (2023) -
Coherent3D: Coherent 3D Portrait Video Reconstruction via Triplane Fusion
by: Wang, Shengze, et al.
Published: (2024) -
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
by: Kim, Heechang, et al.
Published: (2025)