COIN: Control-Inpainting Diffusion Prior for Human and Camera Motion Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jiefeng, Yuan, Ye, Rempe, Davis, Zhang, Haotian, Molchanov, Pavlo, Lu, Cewu, Kautz, Jan, Iqbal, Umar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GENMO: A GENeralist Model for Human MOtion
by: Li, Jiefeng, et al.
Published: (2025)
by: Li, Jiefeng, et al.
Published: (2025)
AdaHuman: Animatable Detailed 3D Human Generation with Compositional Multiview Diffusion
by: Huang, Yangyi, et al.
Published: (2025)
by: Huang, Yangyi, et al.
Published: (2025)
Kimodo: Scaling Controllable Human Motion Generation
by: Rempe, Davis, et al.
Published: (2026)
by: Rempe, Davis, et al.
Published: (2026)
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
by: Petrovich, Mathis, et al.
Published: (2024)
by: Petrovich, Mathis, et al.
Published: (2024)
AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One
by: Ranzinger, Mike, et al.
Published: (2023)
by: Ranzinger, Mike, et al.
Published: (2023)
ShapeBoost: Boosting Human Shape Estimation with Part-Based Parameterization and Clothing-Preserving Augmentation
by: Bian, Siyuan, et al.
Published: (2024)
by: Bian, Siyuan, et al.
Published: (2024)
GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion
by: Kim, Gwanghyun, et al.
Published: (2025)
by: Kim, Gwanghyun, et al.
Published: (2025)
Generating Human Interaction Motions in Scenes with Text Control
by: Yi, Hongwei, et al.
Published: (2024)
by: Yi, Hongwei, et al.
Published: (2024)
FeatSharp: Your Vision Model Features, Sharper
by: Ranzinger, Mike, et al.
Published: (2025)
by: Ranzinger, Mike, et al.
Published: (2025)
Exploiting Motion Prior for Accurate Pose Estimation of Dashboard Cameras
by: Lu, Yipeng, et al.
Published: (2024)
by: Lu, Yipeng, et al.
Published: (2024)
C-RADIOv4 (Tech Report)
by: Ranzinger, Mike, et al.
Published: (2026)
by: Ranzinger, Mike, et al.
Published: (2026)
HumanOLAT: A Large-Scale Dataset for Full-Body Human Relighting and Novel-View Synthesis
by: Teufel, Timo, et al.
Published: (2025)
by: Teufel, Timo, et al.
Published: (2025)
LITA: Language Instructed Temporal-Localization Assistant
by: Huang, De-An, et al.
Published: (2024)
by: Huang, De-An, et al.
Published: (2024)
RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models
by: Heinrich, Greg, et al.
Published: (2024)
by: Heinrich, Greg, et al.
Published: (2024)
GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning
by: Yuan, Ye, et al.
Published: (2023)
by: Yuan, Ye, et al.
Published: (2023)
VILA: On Pre-training for Visual Language Models
by: Lin, Ji, et al.
Published: (2023)
by: Lin, Ji, et al.
Published: (2023)
FasterViT: Fast Vision Transformers with Hierarchical Attention
by: Hatamizadeh, Ali, et al.
Published: (2023)
by: Hatamizadeh, Ali, et al.
Published: (2023)
RI3D: Few-Shot Gaussian Splatting With Repair and Inpainting Diffusion Priors
by: Paliwal, Avinash, et al.
Published: (2025)
by: Paliwal, Avinash, et al.
Published: (2025)
VILA$^2$: VILA Augmented VILA
by: Fang, Yunhao, et al.
Published: (2024)
by: Fang, Yunhao, et al.
Published: (2024)
TwinTURBO: Semi-Supervised Fine-Tuning of Foundation Models via Mutual Information Decompositions for Downstream Task and Latent Spaces
by: Quétant, Guillaume, et al.
Published: (2025)
by: Quétant, Guillaume, et al.
Published: (2025)
Stereo-Inertial Poser: Towards Metric-Accurate Shape-Aware Motion Capture Using Sparse IMUs and a Single Stereo Camera
by: Tang, Tutian, et al.
Published: (2026)
by: Tang, Tutian, et al.
Published: (2026)
NeRF Inpainting with Geometric Diffusion Prior and Balanced Score Distillation
by: Zhang, Menglin, et al.
Published: (2024)
by: Zhang, Menglin, et al.
Published: (2024)
InpaintHuman: Reconstructing Occluded Humans with Multi-Scale UV Mapping and Identity-Preserving Diffusion Inpainting
by: Fan, Jinlong, et al.
Published: (2026)
by: Fan, Jinlong, et al.
Published: (2026)
SimAvatar: Simulation-Ready Avatars with Layered Hair and Clothing
by: Li, Xueting, et al.
Published: (2024)
by: Li, Xueting, et al.
Published: (2024)
Scaling Vision Pre-Training to 4K Resolution
by: Shi, Baifeng, et al.
Published: (2025)
by: Shi, Baifeng, et al.
Published: (2025)
Bridging the Gap between Human Motion and Action Semantics via Kinematic Phrases
by: Liu, Xinpeng, et al.
Published: (2023)
by: Liu, Xinpeng, et al.
Published: (2023)
Boosting Camera Motion Control for Video Diffusion Transformers
by: Cheong, Soon Yau, et al.
Published: (2024)
by: Cheong, Soon Yau, et al.
Published: (2024)
X-VILA: Cross-Modality Alignment for Large Language Model
by: Ye, Hanrong, et al.
Published: (2024)
by: Ye, Hanrong, et al.
Published: (2024)
PALUM: Part-based Attention Learning for Unified Motion Retargeting
by: Liu, Siqi, et al.
Published: (2026)
by: Liu, Siqi, et al.
Published: (2026)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
by: Wang, Xinkai, et al.
Published: (2026)
by: Wang, Xinkai, et al.
Published: (2026)
SOMA: Unifying Parametric Human Body Models
by: Saito, Jun, et al.
Published: (2026)
by: Saito, Jun, et al.
Published: (2026)
CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
by: Xu, Dejia, et al.
Published: (2024)
by: Xu, Dejia, et al.
Published: (2024)
Estimating 2D Camera Motion with Hybrid Motion Basis
by: Li, Haipeng, et al.
Published: (2025)
by: Li, Haipeng, et al.
Published: (2025)
ControlEvents: Controllable Synthesis of Event Camera Datawith Foundational Prior from Image Diffusion Models
by: Hu, Yixuan, et al.
Published: (2025)
by: Hu, Yixuan, et al.
Published: (2025)
Controllable Localized Face Anonymization Via Diffusion Inpainting
by: Salar, Ali, et al.
Published: (2025)
by: Salar, Ali, et al.
Published: (2025)
ADM: Accelerated Diffusion Model via Estimated Priors for Robust Motion Prediction under Uncertainties
by: Li, Jiahui, et al.
Published: (2024)
by: Li, Jiahui, et al.
Published: (2024)
Consistent and Optimal Solution to Camera Motion Estimation
by: Zeng, Guangyang, et al.
Published: (2024)
by: Zeng, Guangyang, et al.
Published: (2024)
MultiCOIN: Multi-Modal COntrollable Video INbetweening
by: Tanveer, Maham, et al.
Published: (2025)
by: Tanveer, Maham, et al.
Published: (2025)
DiffCap: Diffusion-based Real-time Human Motion Capture using Sparse IMUs and a Monocular Camera
by: Pan, Shaohua, et al.
Published: (2025)
by: Pan, Shaohua, et al.
Published: (2025)
StableMotion: Repurposing Diffusion-Based Image Priors for Motion Estimation
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
Similar Items
-
GENMO: A GENeralist Model for Human MOtion
by: Li, Jiefeng, et al.
Published: (2025) -
AdaHuman: Animatable Detailed 3D Human Generation with Compositional Multiview Diffusion
by: Huang, Yangyi, et al.
Published: (2025) -
Kimodo: Scaling Controllable Human Motion Generation
by: Rempe, Davis, et al.
Published: (2026) -
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
by: Petrovich, Mathis, et al.
Published: (2024) -
AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One
by: Ranzinger, Mike, et al.
Published: (2023)