Synergistic Global-space Camera and Human Reconstruction from Videos
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhao, Yizhou, Wang, Tuanfeng Y., Raj, Bhiksha, Xu, Min, Yang, Jimei, Huang, Chun-Hao Paul |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Boosting Camera Motion Control for Video Diffusion Transformers
par: Cheong, Soon Yau, et autres
Publié: (2024)
par: Cheong, Soon Yau, et autres
Publié: (2024)
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
par: Huang, Zhening, et autres
Publié: (2025)
par: Huang, Zhening, et autres
Publié: (2025)
MASIV: Toward Material-Agnostic System Identification from Videos
par: Zhao, Yizhou, et autres
Publié: (2025)
par: Zhao, Yizhou, et autres
Publié: (2025)
ActAnywhere: Subject-Aware Video Background Generation
par: Pan, Boxiao, et autres
Publié: (2024)
par: Pan, Boxiao, et autres
Publié: (2024)
Pattern Guided UV Recovery for Realistic Video Garment Texturing
par: Zhan, Youyi, et autres
Publié: (2024)
par: Zhan, Youyi, et autres
Publié: (2024)
Comprehensive Relighting: Generalizable and Consistent Monocular Human Relighting and Harmonization
par: Wang, Junying, et autres
Publié: (2025)
par: Wang, Junying, et autres
Publié: (2025)
OmniCam: Unified Multimodal Video Generation via Camera Control
par: Yang, Xiaoda, et autres
Publié: (2025)
par: Yang, Xiaoda, et autres
Publié: (2025)
V-RGBX: Video Editing with Accurate Controls over Intrinsic Properties
par: Fang, Ye, et autres
Publié: (2025)
par: Fang, Ye, et autres
Publié: (2025)
ControlVAR: Exploring Controllable Visual Autoregressive Modeling
par: Li, Xiang, et autres
Publié: (2024)
par: Li, Xiang, et autres
Publié: (2024)
TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos
par: Liu, Jinpeng, et autres
Publié: (2026)
par: Liu, Jinpeng, et autres
Publié: (2026)
D-CoDe: Scaling Image-Pretrained VLMs to Video via Dynamic Compression and Question Decomposition
par: Huang, Yiyang, et autres
Publié: (2025)
par: Huang, Yiyang, et autres
Publié: (2025)
Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting
par: Zhao, Yizhou, et autres
Publié: (2025)
par: Zhao, Yizhou, et autres
Publié: (2025)
I2VControl-Camera: Precise Video Camera Control with Adjustable Motion Strength
par: Feng, Wanquan, et autres
Publié: (2024)
par: Feng, Wanquan, et autres
Publié: (2024)
A Survey of 3D Reconstruction with Event Cameras
par: Xu, Chuanzhi, et autres
Publié: (2025)
par: Xu, Chuanzhi, et autres
Publié: (2025)
JOG3R: Towards 3D-Consistent Video Generators
par: Huang, Chun-Hao Paul, et autres
Publié: (2025)
par: Huang, Chun-Hao Paul, et autres
Publié: (2025)
Robust Latent Matters: Boosting Image Generation with Sampling Error Synthesis
par: Qiu, Kai, et autres
Publié: (2025)
par: Qiu, Kai, et autres
Publié: (2025)
Scalable Benchmarking and Robust Learning for Noise-Free Ego-Motion and 3D Reconstruction from Noisy Video
par: Xu, Xiaohao, et autres
Publié: (2025)
par: Xu, Xiaohao, et autres
Publié: (2025)
Virtually Being: Customizing Camera-Controllable Video Diffusion Models with Multi-View Performance Captures
par: Xu, Yuancheng, et autres
Publié: (2025)
par: Xu, Yuancheng, et autres
Publié: (2025)
Echo4DIR: 4D Implicit Heart Reconstruction from 2D Echocardiography Videos
par: Liu, Yanan, et autres
Publié: (2026)
par: Liu, Yanan, et autres
Publié: (2026)
Neurons: Emulating the Human Visual Cortex Improves Fidelity and Interpretability in fMRI-to-Video Reconstruction
par: Wang, Haonan, et autres
Publié: (2025)
par: Wang, Haonan, et autres
Publié: (2025)
SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework
par: Wu, Tianshu, et autres
Publié: (2026)
par: Wu, Tianshu, et autres
Publié: (2026)
FreeOrbit4D: Training-Free Arbitrary Camera Redirection for Monocular Videos via Foreground-Complete 4D Reconstruction
par: Cao, Wei, et autres
Publié: (2026)
par: Cao, Wei, et autres
Publié: (2026)
GloTSFormer: Global Video Text Spotting Transformer
par: Wang, Han, et autres
Publié: (2024)
par: Wang, Han, et autres
Publié: (2024)
VideoJudge: Bootstrapping Enables Scalable Supervision of MLLM-as-a-Judge for Video Understanding
par: Waheed, Abdul, et autres
Publié: (2025)
par: Waheed, Abdul, et autres
Publié: (2025)
CHRIS: Clothed Human Reconstruction with Side View Consistency
par: Liu, Dong, et autres
Publié: (2025)
par: Liu, Dong, et autres
Publié: (2025)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
par: He, Hao, et autres
Publié: (2024)
par: He, Hao, et autres
Publié: (2024)
VerLM: Explaining Face Verification Using Natural Language
par: Hannan, Syed Abdul, et autres
Publié: (2026)
par: Hannan, Syed Abdul, et autres
Publié: (2026)
GenFusion: Closing the Loop between Reconstruction and Generation via Videos
par: Wu, Sibo, et autres
Publié: (2025)
par: Wu, Sibo, et autres
Publié: (2025)
Template-Free Single-View 3D Human Digitalization with Diffusion-Guided LRM
par: Weng, Zhenzhen, et autres
Publié: (2024)
par: Weng, Zhenzhen, et autres
Publié: (2024)
WinT3R: Window-Based Streaming Reconstruction with Camera Token Pool
par: Li, Zizun, et autres
Publié: (2025)
par: Li, Zizun, et autres
Publié: (2025)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
par: Jeong, Hyeonho, et autres
Publié: (2024)
par: Jeong, Hyeonho, et autres
Publié: (2024)
Video Forgery Detection for Surveillance Cameras: A Review
par: Tayfor, Noor B., et autres
Publié: (2025)
par: Tayfor, Noor B., et autres
Publié: (2025)
Geometry-Guided Camera Motion Understanding in VideoLLMs
par: Feng, Haoan, et autres
Publié: (2026)
par: Feng, Haoan, et autres
Publié: (2026)
PRISM: A Unified Framework for Photorealistic Reconstruction and Intrinsic Scene Modeling
par: Dirik, Alara, et autres
Publié: (2025)
par: Dirik, Alara, et autres
Publié: (2025)
Animating the Past: Reconstruct Trilobite via Video Generation
par: Wu, Xiaoran, et autres
Publié: (2024)
par: Wu, Xiaoran, et autres
Publié: (2024)
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
par: Song, Chenxi, et autres
Publié: (2025)
par: Song, Chenxi, et autres
Publié: (2025)
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
par: Chen, Hao, et autres
Publié: (2022)
par: Chen, Hao, et autres
Publié: (2022)
Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks
par: Chen, Hao, et autres
Publié: (2023)
par: Chen, Hao, et autres
Publié: (2023)
ReGenNet: Towards Human Action-Reaction Synthesis
par: Xu, Liang, et autres
Publié: (2024)
par: Xu, Liang, et autres
Publié: (2024)
Distorted or Fabricated? A Survey on Hallucination in Video LLMs
par: Huang, Yiyang, et autres
Publié: (2026)
par: Huang, Yiyang, et autres
Publié: (2026)
Documents similaires
-
Boosting Camera Motion Control for Video Diffusion Transformers
par: Cheong, Soon Yau, et autres
Publié: (2024) -
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
par: Huang, Zhening, et autres
Publié: (2025) -
MASIV: Toward Material-Agnostic System Identification from Videos
par: Zhao, Yizhou, et autres
Publié: (2025) -
ActAnywhere: Subject-Aware Video Background Generation
par: Pan, Boxiao, et autres
Publié: (2024) -
Pattern Guided UV Recovery for Realistic Video Garment Texturing
par: Zhan, Youyi, et autres
Publié: (2024)