Probing into Camera Control of Video Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hou, Chen, Rupprecht, Christian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos
by: Wimbauer, Felix, et al.
Published: (2025)
by: Wimbauer, Felix, et al.
Published: (2025)
Training-free Camera Control for Video Generation
by: Hou, Chen, et al.
Published: (2024)
by: Hou, Chen, et al.
Published: (2024)
Epipolar Geometry Improves Video Generation Models
by: Kupyn, Orest, et al.
Published: (2025)
by: Kupyn, Orest, et al.
Published: (2025)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026)
by: Ge, Wenhang, et al.
Published: (2026)
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
by: Zhou, Jensen, et al.
Published: (2025)
by: Zhou, Jensen, et al.
Published: (2025)
Learning Camera Movement Control from Real-World Drone Videos
by: Hou, Yunzhong, et al.
Published: (2024)
by: Hou, Yunzhong, et al.
Published: (2024)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
by: He, Hao, et al.
Published: (2024)
by: He, Hao, et al.
Published: (2024)
CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation
by: Zhao, Haoyu, et al.
Published: (2026)
by: Zhao, Haoyu, et al.
Published: (2026)
Unbiasing through Textual Descriptions: Mitigating Representation Bias in Video Benchmarks
by: Shvetsova, Nina, et al.
Published: (2025)
by: Shvetsova, Nina, et al.
Published: (2025)
OmniCamera: A Unified Framework for Multi-task Video Generation with Arbitrary Camera Control
by: Wang, Yukun, et al.
Published: (2026)
by: Wang, Yukun, et al.
Published: (2026)
FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views
by: Zhang, Shangzhan, et al.
Published: (2025)
by: Zhang, Shangzhan, et al.
Published: (2025)
DualCamCtrl: Dual-Branch Diffusion Model for Geometry-Aware Camera-Controlled Video Generation
by: Zhang, Hongfei, et al.
Published: (2025)
by: Zhang, Hongfei, et al.
Published: (2025)
Diffusion Models for Open-Vocabulary Segmentation
by: Karazija, Laurynas, et al.
Published: (2023)
by: Karazija, Laurynas, et al.
Published: (2023)
Unified Camera Positional Encoding for Controlled Video Generation
by: Zhang, Cheng, et al.
Published: (2025)
by: Zhang, Cheng, et al.
Published: (2025)
CamI2V: Camera-Controlled Image-to-Video Diffusion Model
by: Zheng, Guangcong, et al.
Published: (2024)
by: Zheng, Guangcong, et al.
Published: (2024)
Latent-Reframe: Enabling Camera Control for Video Diffusion Model without Training
by: Zhou, Zhenghong, et al.
Published: (2024)
by: Zhou, Zhenghong, et al.
Published: (2024)
HowToCaption: Prompting LLMs to Transform Video Annotations at Scale
by: Shvetsova, Nina, et al.
Published: (2023)
by: Shvetsova, Nina, et al.
Published: (2023)
On Geometric Understanding and Learned Priors in Feed-forward 3D Reconstruction Models
by: Bratulić, Jelena, et al.
Published: (2025)
by: Bratulić, Jelena, et al.
Published: (2025)
FloVD: Optical Flow Meets Video Diffusion Model for Enhanced Camera-Controlled Video Synthesis
by: Jin, Wonjoon, et al.
Published: (2025)
by: Jin, Wonjoon, et al.
Published: (2025)
Direct-a-Video: Customized Video Generation with User-Directed Camera Movement and Object Motion
by: Yang, Shiyuan, et al.
Published: (2024)
by: Yang, Shiyuan, et al.
Published: (2024)
SHIC: Shape-Image Correspondences with no Keypoint Supervision
by: Shtedritski, Aleksandar, et al.
Published: (2024)
by: Shtedritski, Aleksandar, et al.
Published: (2024)
PoseDiffusion: Solving Pose Estimation via Diffusion-aided Bundle Adjustment
by: Wang, Jianyuan, et al.
Published: (2023)
by: Wang, Jianyuan, et al.
Published: (2023)
Splatter Image: Ultra-Fast Single-View 3D Reconstruction
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
by: Kupyn, Orest, et al.
Published: (2025)
by: Kupyn, Orest, et al.
Published: (2025)
CameraNoise: Enabling Faithful Camera Control in Video Diffusion through Geometry-Flow-Guided Noise Warping
by: Zhao, Haoyu, et al.
Published: (2026)
by: Zhao, Haoyu, et al.
Published: (2026)
CoTracker3: Simpler and Better Point Tracking by Pseudo-Labelling Real Videos
by: Karaev, Nikita, et al.
Published: (2024)
by: Karaev, Nikita, et al.
Published: (2024)
Dataset Enhancement with Instance-Level Augmentations
by: Kupyn, Orest, et al.
Published: (2024)
by: Kupyn, Orest, et al.
Published: (2024)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
by: Xu, Youcan, et al.
Published: (2026)
by: Xu, Youcan, et al.
Published: (2026)
GimbalDiffusion: Gravity-Aware Camera Control for Video Generation
by: Fortier-Chouinard, Frédéric, et al.
Published: (2025)
by: Fortier-Chouinard, Frédéric, et al.
Published: (2025)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
by: Kim, Min-Jung, et al.
Published: (2025)
by: Kim, Min-Jung, et al.
Published: (2025)
Taming Camera-Controlled Video Generation with Verifiable Geometry Reward
by: Wang, Zhaoqing, et al.
Published: (2025)
by: Wang, Zhaoqing, et al.
Published: (2025)
Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics
by: Li, Ruining, et al.
Published: (2024)
by: Li, Ruining, et al.
Published: (2024)
Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
Light-X: Generative 4D Video Rendering with Camera and Illumination Control
by: Liu, Tianqi, et al.
Published: (2025)
by: Liu, Tianqi, et al.
Published: (2025)
SymphoMotion: Joint Control of Camera Motion and Object Dynamics for Coherent Video Generation
by: Zhang, Guiyu, et al.
Published: (2026)
by: Zhang, Guiyu, et al.
Published: (2026)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
by: Jiang, Hong, et al.
Published: (2026)
by: Jiang, Hong, et al.
Published: (2026)
MyGo: Consistent and Controllable Multi-View Driving Video Generation with Camera Control
by: Yao, Yining, et al.
Published: (2024)
by: Yao, Yining, et al.
Published: (2024)
PostCam: Camera-Controllable Novel-View Video Generation with Query-Shared Cross-Attention
by: Chen, Yipeng, et al.
Published: (2025)
by: Chen, Yipeng, et al.
Published: (2025)
CamLit: Unified Video Diffusion with Explicit Camera and Lighting Control
by: Kuang, Zhiyi, et al.
Published: (2026)
by: Kuang, Zhiyi, et al.
Published: (2026)
Zero-Shot Personalized Camera Motion Control for Image-to-Video Synthesis
by: Guhan, Pooja, et al.
Published: (2025)
by: Guhan, Pooja, et al.
Published: (2025)
Similar Items
-
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos
by: Wimbauer, Felix, et al.
Published: (2025) -
Training-free Camera Control for Video Generation
by: Hou, Chen, et al.
Published: (2024) -
Epipolar Geometry Improves Video Generation Models
by: Kupyn, Orest, et al.
Published: (2025) -
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026) -
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
by: Zhou, Jensen, et al.
Published: (2025)