WATCH: World-aware Allied Trajectory and pose reconstruction for Camera and Human
Fuente:
arXiv
Saved in:
| Main Authors: | Ying, Qijun, Hu, Zhongyuan, Zhang, Rui, Li, Ronghui, Lu, Yu, Zeng, Zijiao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation
by: Wang, Yuelei, et al.
Published: (2024)
by: Wang, Yuelei, et al.
Published: (2024)
From Camera to World: A Plug-and-Play Module for Human Mesh Transformation
by: Ma, Changhai, et al.
Published: (2025)
by: Ma, Changhai, et al.
Published: (2025)
UCM: Unifying Camera Control and Memory with Time-aware Positional Encoding Warping for World Models
by: Xu, Tianxing, et al.
Published: (2026)
by: Xu, Tianxing, et al.
Published: (2026)
Dive Deeper into Rectifying Homography for Stereo Camera Online Self-Calibration
by: Zhao, Hongbo, et al.
Published: (2023)
by: Zhao, Hongbo, et al.
Published: (2023)
SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model
by: Du, Jiayuan, et al.
Published: (2025)
by: Du, Jiayuan, et al.
Published: (2025)
Dynamic Patch-aware Enrichment Transformer for Occluded Person Re-Identification
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
NeuroCine: Decoding Vivid Video Sequences from Human Brain Activties
by: Sun, Jingyuan, et al.
Published: (2024)
by: Sun, Jingyuan, et al.
Published: (2024)
Harmonious Group Choreography with Trajectory-Controllable Diffusion
by: Dai, Yuqin, et al.
Published: (2024)
by: Dai, Yuqin, et al.
Published: (2024)
AToM: Aligning Text-to-Motion Model at Event-Level with GPT-4Vision Reward
by: Han, Haonan, et al.
Published: (2024)
by: Han, Haonan, et al.
Published: (2024)
GeoWATCH for Detecting Heavy Construction in Heterogeneous Time Series of Satellite Images
by: Crall, Jon, et al.
Published: (2024)
by: Crall, Jon, et al.
Published: (2024)
DePatch: Towards Robust Adversarial Patch for Evading Person Detectors in the Real World
by: Cheng, Jikang, et al.
Published: (2024)
by: Cheng, Jikang, et al.
Published: (2024)
Equivariant symmetry-aware head pose estimation for fetal MRI
by: Muthukrishnan, Ramya, et al.
Published: (2025)
by: Muthukrishnan, Ramya, et al.
Published: (2025)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
by: YU, Mark, et al.
Published: (2025)
by: YU, Mark, et al.
Published: (2025)
CameraMaster: Unified Camera Semantic-Parameter Control for Photography Retouching
by: Yang, Qirui, et al.
Published: (2025)
by: Yang, Qirui, et al.
Published: (2025)
Online,Target-Free LiDAR-Camera Extrinsic Calibration via Cross-Modal Mask Matching
by: Huang, Zhiwei, et al.
Published: (2024)
by: Huang, Zhiwei, et al.
Published: (2024)
WATCH: Wide-Area Archaeological Site Tracking for Change Detection
by: Tadesse, Girmaw Abebe, et al.
Published: (2026)
by: Tadesse, Girmaw Abebe, et al.
Published: (2026)
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
by: Wang, Weizhuo, et al.
Published: (2024)
by: Wang, Weizhuo, et al.
Published: (2024)
EmoFace: Audio-driven Emotional 3D Face Animation
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
Humans as Checkerboards: Calibrating Camera Motion Scale for World-Coordinate Human Mesh Recovery
by: Yang, Fengyuan, et al.
Published: (2024)
by: Yang, Fengyuan, et al.
Published: (2024)
MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation
by: Lei, Guojun, et al.
Published: (2025)
by: Lei, Guojun, et al.
Published: (2025)
GeoGS3D: Single-view 3D Reconstruction via Geometric-aware Diffusion Model and Gaussian Splatting
by: Feng, Qijun, et al.
Published: (2024)
by: Feng, Qijun, et al.
Published: (2024)
EF-Calib: Spatiotemporal Calibration of Event- and Frame-Based Cameras Using Continuous-Time Trajectories
by: Wang, Shaoan, et al.
Published: (2024)
by: Wang, Shaoan, et al.
Published: (2024)
Hybrid bundle-adjusting 3D Gaussians for view consistent rendering with pose optimization
by: Guo, Yanan, et al.
Published: (2024)
by: Guo, Yanan, et al.
Published: (2024)
ALIVE: Animate Your World with Lifelike Audio-Video Generation
by: Guo, Ying, et al.
Published: (2026)
by: Guo, Ying, et al.
Published: (2026)
DST-Calib: A Dual-Path, Self-Supervised, Target-Free LiDAR-Camera Extrinsic Calibration Network
by: Huang, Zhiwei, et al.
Published: (2026)
by: Huang, Zhiwei, et al.
Published: (2026)
MGStream: Motion-aware 3D Gaussian for Streamable Dynamic Scene Reconstruction
by: Bao, Zhenyu, et al.
Published: (2025)
by: Bao, Zhenyu, et al.
Published: (2025)
InfiniteDance: Scalable 3D Dance Generation Towards in-the-wild Generalization
by: Li, Ronghui, et al.
Published: (2026)
by: Li, Ronghui, et al.
Published: (2026)
3D Trajectory Reconstruction of Moving Points Based on Asynchronous Cameras
by: Huang, Huayu, et al.
Published: (2025)
by: Huang, Huayu, et al.
Published: (2025)
Trajectory-aware Shifted State Space Models for Online Video Super-Resolution
by: Zhu, Qiang, et al.
Published: (2025)
by: Zhu, Qiang, et al.
Published: (2025)
VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation
by: Li, Mengtian, et al.
Published: (2026)
by: Li, Mengtian, et al.
Published: (2026)
Camera-aware Label Refinement for Unsupervised Person Re-identification
by: Li, Pengna, et al.
Published: (2024)
by: Li, Pengna, et al.
Published: (2024)
Cross-view and Cross-pose Completion for 3D Human Understanding
by: Armando, Matthieu, et al.
Published: (2023)
by: Armando, Matthieu, et al.
Published: (2023)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
3D Trajectory Reconstruction of Moving Points Based on a Monocular Camera
by: Huang, Huayu, et al.
Published: (2025)
by: Huang, Huayu, et al.
Published: (2025)
Cross-Modality Gait Recognition: Bridging LiDAR and Camera Modalities for Human Identification
by: Wang, Rui, et al.
Published: (2024)
by: Wang, Rui, et al.
Published: (2024)
Multi-Camera Asynchronous Ball Localization and Trajectory Prediction with Factor Graphs and Human Poses
by: Xiao, Qingyu, et al.
Published: (2024)
by: Xiao, Qingyu, et al.
Published: (2024)
$\text{PKS}^4$:Parallel Kinematic Selective State Space Scanners for Efficient Video Understanding
by: Zeng, Lingjie, et al.
Published: (2026)
by: Zeng, Lingjie, et al.
Published: (2026)
WorldCraft: From Camera Navigation to Object Manipulation in Interactive Video World Models
by: Gu, Bohai, et al.
Published: (2026)
by: Gu, Bohai, et al.
Published: (2026)
WoVoGen: World Volume-aware Diffusion for Controllable Multi-camera Driving Scene Generation
by: Lu, Jiachen, et al.
Published: (2023)
by: Lu, Jiachen, et al.
Published: (2023)
CAMOT: Camera Angle-aware Multi-Object Tracking
by: Limanta, Felix, et al.
Published: (2024)
by: Limanta, Felix, et al.
Published: (2024)
Similar Items
-
CPA: Camera-pose-awareness Diffusion Transformer for Video Generation
by: Wang, Yuelei, et al.
Published: (2024) -
From Camera to World: A Plug-and-Play Module for Human Mesh Transformation
by: Ma, Changhai, et al.
Published: (2025) -
UCM: Unifying Camera Control and Memory with Time-aware Positional Encoding Warping for World Models
by: Xu, Tianxing, et al.
Published: (2026) -
Dive Deeper into Rectifying Homography for Stereo Camera Online Self-Calibration
by: Zhao, Hongbo, et al.
Published: (2023) -
SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model
by: Du, Jiayuan, et al.
Published: (2025)