Saved in:
| Main Authors: | Lichy, Daniel, Su, Hang, Badki, Abhishek, Kautz, Jan, Gallo, Orazio |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.12074 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FoVA-Depth: Field-of-View Agnostic Depth Estimation for Cross-Dataset Generalization
by: Lichy, Daniel, et al.
Published: (2024)
by: Lichy, Daniel, et al.
Published: (2024)
L4P: Towards Unified Low-Level 4D Vision Perception
by: Badki, Abhishek, et al.
Published: (2025)
by: Badki, Abhishek, et al.
Published: (2025)
Zero-Shot Monocular Scene Flow Estimation in the Wild
by: Liang, Yiqing, et al.
Published: (2025)
by: Liang, Yiqing, et al.
Published: (2025)
CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
by: Xu, Dejia, et al.
Published: (2024)
by: Xu, Dejia, et al.
Published: (2024)
FoundationStereo: Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
DeltaCam: Differential Intrinsic Camera Modeling for Video Generation
by: Mandal, Debabrata, et al.
Published: (2026)
by: Mandal, Debabrata, et al.
Published: (2026)
MambaVision: A Hybrid Mamba-Transformer Vision Backbone
by: Hatamizadeh, Ali, et al.
Published: (2024)
by: Hatamizadeh, Ali, et al.
Published: (2024)
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning
by: Wu, Hang, et al.
Published: (2026)
by: Wu, Hang, et al.
Published: (2026)
WorldCam: Interactive Autoregressive 3D Gaming Worlds with Camera Pose as a Unifying Geometric Representation
by: Nam, Jisu, et al.
Published: (2026)
by: Nam, Jisu, et al.
Published: (2026)
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos
by: Wimbauer, Felix, et al.
Published: (2025)
by: Wimbauer, Felix, et al.
Published: (2025)
DiffPhysCam: Differentiable Physics-Based Camera Simulation for Inverse Rendering and Embodied AI
by: Chen, Bo-Hsun, et al.
Published: (2025)
by: Chen, Bo-Hsun, et al.
Published: (2025)
ChatCam: Empowering Camera Control through Conversational AI
by: Liu, Xinhang, et al.
Published: (2024)
by: Liu, Xinhang, et al.
Published: (2024)
DriveCamSim: Generalizable Camera Simulation via Explicit Camera Modeling for Autonomous Driving
by: Sun, Wenchao, et al.
Published: (2025)
by: Sun, Wenchao, et al.
Published: (2025)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026)
by: Ge, Wenhang, et al.
Published: (2026)
MultiCam: On-the-fly Multi-Camera Pose Estimation Using Spatiotemporal Overlaps of Known Objects
by: Li, Shiyu, et al.
Published: (2026)
by: Li, Shiyu, et al.
Published: (2026)
CamPro: Camera-based Anti-Facial Recognition
by: Zhu, Wenjun, et al.
Published: (2023)
by: Zhu, Wenjun, et al.
Published: (2023)
Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM
by: Pan, Zirui, et al.
Published: (2025)
by: Pan, Zirui, et al.
Published: (2025)
CamLit: Unified Video Diffusion with Explicit Camera and Lighting Control
by: Kuang, Zhiyi, et al.
Published: (2026)
by: Kuang, Zhiyi, et al.
Published: (2026)
Motion-Aware Optical Camera Communication with Event Cameras
by: Su, Hang, et al.
Published: (2024)
by: Su, Hang, et al.
Published: (2024)
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
by: Yao, Manyi, et al.
Published: (2025)
by: Yao, Manyi, et al.
Published: (2025)
CamI2V: Camera-Controlled Image-to-Video Diffusion Model
by: Zheng, Guangcong, et al.
Published: (2024)
by: Zheng, Guangcong, et al.
Published: (2024)
FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning
by: Lyu, Weijie, et al.
Published: (2026)
by: Lyu, Weijie, et al.
Published: (2026)
CamCloneMaster: Enabling Reference-based Camera Control for Video Generation
by: Luo, Yawen, et al.
Published: (2025)
by: Luo, Yawen, et al.
Published: (2025)
CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion
by: Ji, Chenhao, et al.
Published: (2025)
by: Ji, Chenhao, et al.
Published: (2025)
CamFreeDiff: Camera-free Image to Panorama Generation with Diffusion Model
by: Yuan, Xiaoding, et al.
Published: (2024)
by: Yuan, Xiaoding, et al.
Published: (2024)
VividCam: Learning Unconventional Camera Motions from Virtual Synthetic Videos
by: Wu, Qiucheng, et al.
Published: (2025)
by: Wu, Qiucheng, et al.
Published: (2025)
An N-Point Linear Solver for Line and Motion Estimation with Event Cameras
by: Gao, Ling, et al.
Published: (2024)
by: Gao, Ling, et al.
Published: (2024)
Learning Camera-Agnostic White-Balance Preferences
by: Zhao, Luxi, et al.
Published: (2025)
by: Zhao, Luxi, et al.
Published: (2025)
GeoDANO: Geometric VLM with Domain Agnostic Vision Encoder
by: Cho, Seunghyuk, et al.
Published: (2025)
by: Cho, Seunghyuk, et al.
Published: (2025)
HyperCam: Low-Power Onboard Computer Vision for IoT Cameras
by: Lee, Chae Young, et al.
Published: (2025)
by: Lee, Chae Young, et al.
Published: (2025)
Avatar Fingerprinting for Authorized Use of Synthetic Talking-Head Videos
by: Prashnani, Ekta, et al.
Published: (2023)
by: Prashnani, Ekta, et al.
Published: (2023)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
by: Xu, Youcan, et al.
Published: (2026)
by: Xu, Youcan, et al.
Published: (2026)
RadarCam-Depth: Radar-Camera Fusion for Depth Estimation with Learned Metric Scale
by: Li, Han, et al.
Published: (2024)
by: Li, Han, et al.
Published: (2024)
SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints
by: Bai, Jianhong, et al.
Published: (2024)
by: Bai, Jianhong, et al.
Published: (2024)
ReCamMaster: Camera-Controlled Generative Rendering from A Single Video
by: Bai, Jianhong, et al.
Published: (2025)
by: Bai, Jianhong, et al.
Published: (2025)
AM-RADIO: Agglomerative Vision Foundation Model -- Reduce All Domains Into One
by: Ranzinger, Mike, et al.
Published: (2023)
by: Ranzinger, Mike, et al.
Published: (2023)
OmniCam: Unified Multimodal Video Generation via Camera Control
by: Yang, Xiaoda, et al.
Published: (2025)
by: Yang, Xiaoda, et al.
Published: (2025)
GS-ProCams: Gaussian Splatting-based Projector-Camera Systems
by: Deng, Qingyue, et al.
Published: (2024)
by: Deng, Qingyue, et al.
Published: (2024)
MemCam: Memory-Augmented Camera Control for Consistent Video Generation
by: Gao, Xinhang, et al.
Published: (2026)
by: Gao, Xinhang, et al.
Published: (2026)
HitoMi-Cam: A Shape-Agnostic Person Detection Method Using the Spectral Characteristics of Clothing
by: Ono, Shuji
Published: (2025)
by: Ono, Shuji
Published: (2025)
Similar Items
-
FoVA-Depth: Field-of-View Agnostic Depth Estimation for Cross-Dataset Generalization
by: Lichy, Daniel, et al.
Published: (2024) -
L4P: Towards Unified Low-Level 4D Vision Perception
by: Badki, Abhishek, et al.
Published: (2025) -
Zero-Shot Monocular Scene Flow Estimation in the Wild
by: Liang, Yiqing, et al.
Published: (2025) -
CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
by: Xu, Dejia, et al.
Published: (2024) -
FoundationStereo: Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)