PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular Videos
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Yufei, Kephart, Jeffrey O., Cui, Zijun, Ji, Qiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Diffusion-based 3D Hand Motion Recovery with Intuitive Physics
por: Zhang, Yufei, et al.
Publicado: (2025)
por: Zhang, Yufei, et al.
Publicado: (2025)
Weakly-Supervised 3D Hand Reconstruction with Knowledge Prior and Uncertainty Guidance
por: Zhang, Yufei, et al.
Publicado: (2024)
por: Zhang, Yufei, et al.
Publicado: (2024)
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
por: Ismayilzada, Elkhan, et al.
Publicado: (2026)
por: Ismayilzada, Elkhan, et al.
Publicado: (2026)
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
por: Xing, Chaoyue, et al.
Publicado: (2026)
por: Xing, Chaoyue, et al.
Publicado: (2026)
Beyond Motion Pattern: An Empirical Study of Physical Forces for Human Motion Understanding
por: Dao, Anh, et al.
Publicado: (2025)
por: Dao, Anh, et al.
Publicado: (2025)
PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments
por: Peng, Kebin, et al.
Publicado: (2024)
por: Peng, Kebin, et al.
Publicado: (2024)
MultiPhys: Multi-Person Physics-aware 3D Motion Estimation
por: Ugrinovic, Nicolas, et al.
Publicado: (2024)
por: Ugrinovic, Nicolas, et al.
Publicado: (2024)
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos
por: Lu, Jiahao, et al.
Publicado: (2024)
por: Lu, Jiahao, et al.
Publicado: (2024)
IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular Videos
por: Li, Yuan, et al.
Publicado: (2025)
por: Li, Yuan, et al.
Publicado: (2025)
PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos
por: Cao, Meng, et al.
Publicado: (2024)
por: Cao, Meng, et al.
Publicado: (2024)
D$^3$-Human: Dynamic Disentangled Digital Human from Monocular Video
por: Chen, Honghu, et al.
Publicado: (2025)
por: Chen, Honghu, et al.
Publicado: (2025)
Physics Informed Human Posture Estimation Based on 3D Landmarks from Monocular RGB-Videos
por: Leuthold, Tobias, et al.
Publicado: (2025)
por: Leuthold, Tobias, et al.
Publicado: (2025)
MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos
por: Rho, Daniel, et al.
Publicado: (2026)
por: Rho, Daniel, et al.
Publicado: (2026)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
por: Ji, Sihui, et al.
Publicado: (2025)
por: Ji, Sihui, et al.
Publicado: (2025)
Confidence-aware Monocular Depth Estimation for Minimally Invasive Surgery
por: Asad, Muhammad, et al.
Publicado: (2026)
por: Asad, Muhammad, et al.
Publicado: (2026)
Transtreaming: Adaptive Delay-aware Transformer for Real-time Streaming Perception
por: Zhang, Xiang, et al.
Publicado: (2024)
por: Zhang, Xiang, et al.
Publicado: (2024)
The Devil is in the Edges: Monocular Depth Estimation with Edge-aware Consistency Fusion
por: Li, Pengzhi, et al.
Publicado: (2024)
por: Li, Pengzhi, et al.
Publicado: (2024)
Predicting 4D Hand Trajectory from Monocular Videos
por: Ye, Yufei, et al.
Publicado: (2025)
por: Ye, Yufei, et al.
Publicado: (2025)
UniCon3R: Unified Contact-aware 4D Human-Scene Reconstruction from Monocular Video
por: Sur, Tanuj, et al.
Publicado: (2026)
por: Sur, Tanuj, et al.
Publicado: (2026)
IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray Tracing
por: Wang, Shaofei, et al.
Publicado: (2023)
por: Wang, Shaofei, et al.
Publicado: (2023)
Decoupling Dynamic Monocular Videos for Dynamic View Synthesis
por: You, Meng, et al.
Publicado: (2023)
por: You, Meng, et al.
Publicado: (2023)
Spectral Compression Transformer with Line Pose Graph for Monocular 3D Human Pose Estimation
por: Zheng, Zenghao, et al.
Publicado: (2025)
por: Zheng, Zenghao, et al.
Publicado: (2025)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
por: Wang, Chen, et al.
Publicado: (2025)
por: Wang, Chen, et al.
Publicado: (2025)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
por: Zhang, Haoze, et al.
Publicado: (2025)
por: Zhang, Haoze, et al.
Publicado: (2025)
Dynamic Gaussian Splatting from Defocused and Motion-blurred Monocular Videos
por: Zhang, Xuankai, et al.
Publicado: (2025)
por: Zhang, Xuankai, et al.
Publicado: (2025)
PhysRVG: Physics-Aware Unified Reinforcement Learning for Video Generative Models
por: Zhang, Qiyuan, et al.
Publicado: (2026)
por: Zhang, Qiyuan, et al.
Publicado: (2026)
CTNeRF: Cross-Time Transformer for Dynamic Neural Radiance Field from Monocular Video
por: Miao, Xingyu, et al.
Publicado: (2024)
por: Miao, Xingyu, et al.
Publicado: (2024)
Map-Mono-Ego: Map-Grounded Global Human Pose Estimation from Monocular Egocentric Video
por: Deguchi, Hiroyuki, et al.
Publicado: (2026)
por: Deguchi, Hiroyuki, et al.
Publicado: (2026)
Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos
por: Wei, Dingkun, et al.
Publicado: (2026)
por: Wei, Dingkun, et al.
Publicado: (2026)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
por: Bao, Yiming, et al.
Publicado: (2024)
por: Bao, Yiming, et al.
Publicado: (2024)
I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions
por: Zhao, Chengfeng, et al.
Publicado: (2023)
por: Zhao, Chengfeng, et al.
Publicado: (2023)
BadDepth: Backdoor Attacks Against Monocular Depth Estimation in the Physical World
por: Guo, Ji, et al.
Publicado: (2025)
por: Guo, Ji, et al.
Publicado: (2025)
DBMovi-GS: Dynamic View Synthesis from Blurry Monocular Video via Sparse-Controlled Gaussian Splatting
por: Song, Yeon-Ji, et al.
Publicado: (2025)
por: Song, Yeon-Ji, et al.
Publicado: (2025)
StableDPT: Temporal Stable Monocular Video Depth Estimation
por: Sobko, Ivan, et al.
Publicado: (2026)
por: Sobko, Ivan, et al.
Publicado: (2026)
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
por: Yasarla, Rajeev, et al.
Publicado: (2023)
por: Yasarla, Rajeev, et al.
Publicado: (2023)
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
por: Wang, Haiping, et al.
Publicado: (2023)
por: Wang, Haiping, et al.
Publicado: (2023)
Expressive Gaussian Human Avatars from Monocular RGB Video
por: Hu, Hezhen, et al.
Publicado: (2024)
por: Hu, Hezhen, et al.
Publicado: (2024)
PhysFlow: Unleashing the Potential of Multi-modal Foundation Models and Video Diffusion for 4D Dynamic Physical Scene Simulation
por: Liu, Zhuoman, et al.
Publicado: (2024)
por: Liu, Zhuoman, et al.
Publicado: (2024)
ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video
por: Wang, Boyuan, et al.
Publicado: (2026)
por: Wang, Boyuan, et al.
Publicado: (2026)
Interpretable Vision Transformers in Monocular Depth Estimation via SVDA
por: Arampatzakis, Vasileios, et al.
Publicado: (2026)
por: Arampatzakis, Vasileios, et al.
Publicado: (2026)
Ejemplares similares
-
Diffusion-based 3D Hand Motion Recovery with Intuitive Physics
por: Zhang, Yufei, et al.
Publicado: (2025) -
Weakly-Supervised 3D Hand Reconstruction with Knowledge Prior and Uncertainty Guidance
por: Zhang, Yufei, et al.
Publicado: (2024) -
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
por: Ismayilzada, Elkhan, et al.
Publicado: (2026) -
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
por: Xing, Chaoyue, et al.
Publicado: (2026) -
Beyond Motion Pattern: An Empirical Study of Physical Forces for Human Motion Understanding
por: Dao, Anh, et al.
Publicado: (2025)