ObjCtrl-2.5D: Training-free Object Control with Camera Poses
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zhouxia, Lan, Yushi, Zhou, Shangchen, Loy, Chen Change |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement
by: Luo, Yihang, et al.
Published: (2024)
by: Luo, Yihang, et al.
Published: (2024)
4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere
by: Luo, Yihang, et al.
Published: (2026)
by: Luo, Yihang, et al.
Published: (2026)
GaussianAnything: Interactive Point Cloud Flow Matching For 3D Object Generation
by: Lan, Yushi, et al.
Published: (2024)
by: Lan, Yushi, et al.
Published: (2024)
Denoising as Adaptation: Noise-Space Domain Adaptation for Image Restoration
by: Liao, Kang, et al.
Published: (2024)
by: Liao, Kang, et al.
Published: (2024)
Precise Object and Effect Removal with Adaptive Target-Aware Attention
by: Zhao, Jixin, et al.
Published: (2025)
by: Zhao, Jixin, et al.
Published: (2025)
Trans-Adapter: A Plug-and-Play Framework for Transparent Image Inpainting
by: Dai, Yuekun, et al.
Published: (2025)
by: Dai, Yuekun, et al.
Published: (2025)
Control Color: Multimodal Diffusion-based Interactive Image Colorization
by: Liang, Zhexin, et al.
Published: (2024)
by: Liang, Zhexin, et al.
Published: (2024)
SAR3D: Autoregressive 3D Object Generation and Understanding via Multi-scale 3D VQVAE
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
STream3R: Scalable Sequential 3D Reconstruction with Causal Transformer
by: Lan, Yushi, et al.
Published: (2025)
by: Lan, Yushi, et al.
Published: (2025)
LN3DIFF++: Scalable Latent Neural Fields Diffusion for Speedy 3D Generation
by: Lan, Yushi, et al.
Published: (2024)
by: Lan, Yushi, et al.
Published: (2024)
Learning Inclusion Matching for Animation Paint Bucket Colorization
by: Dai, Yuekun, et al.
Published: (2024)
by: Dai, Yuekun, et al.
Published: (2024)
MatAnyone: Stable Video Matting with Consistent Memory Propagation
by: Yang, Peiqing, et al.
Published: (2025)
by: Yang, Peiqing, et al.
Published: (2025)
Exploiting Diffusion Prior for Real-World Image Super-Resolution
by: Wang, Jianyi, et al.
Published: (2023)
by: Wang, Jianyi, et al.
Published: (2023)
Paint Bucket Colorization Using Anime Character Color Design Sheets
by: Dai, Yuekun, et al.
Published: (2024)
by: Dai, Yuekun, et al.
Published: (2024)
LightCtrl: Training-free Controllable Video Relighting
by: Peng, Yizuo, et al.
Published: (2026)
by: Peng, Yizuo, et al.
Published: (2026)
Next Visual Granularity Generation
by: Wang, Yikai, et al.
Published: (2025)
by: Wang, Yikai, et al.
Published: (2025)
MVIP-NeRF: Multi-view 3D Inpainting on NeRF Scenes via Diffusion Prior
by: Chen, Honghua, et al.
Published: (2024)
by: Chen, Honghua, et al.
Published: (2024)
CameraCtrl: Enabling Camera Control for Text-to-Video Generation
by: He, Hao, et al.
Published: (2024)
by: He, Hao, et al.
Published: (2024)
DST-Det: Simple Dynamic Self-Training for Open-Vocabulary Object Detection
by: Xu, Shilin, et al.
Published: (2023)
by: Xu, Shilin, et al.
Published: (2023)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
by: Wang, Zhouxia, et al.
Published: (2023)
by: Wang, Zhouxia, et al.
Published: (2023)
Learning 3D Garment Animation from Trajectories of A Piece of Cloth
by: Shao, Yidi, et al.
Published: (2025)
by: Shao, Yidi, et al.
Published: (2025)
GausSim: Foreseeing Reality by Gaussian Simulator for Elastic Objects
by: Shao, Yidi, et al.
Published: (2024)
by: Shao, Yidi, et al.
Published: (2024)
PnP-U3D: Plug-and-Play 3D Framework Bridging Autoregression and Diffusion for Unified Understanding and Generation
by: Chen, Yongwei, et al.
Published: (2026)
by: Chen, Yongwei, et al.
Published: (2026)
SeedVR2: One-Step Video Restoration via Diffusion Adversarial Post-Training
by: Wang, Jianyi, et al.
Published: (2025)
by: Wang, Jianyi, et al.
Published: (2025)
Thinking with Camera: A Unified Multimodal Model for Camera-Centric Understanding and Generation
by: Liao, Kang, et al.
Published: (2025)
by: Liao, Kang, et al.
Published: (2025)
Contextual Object Detection with Multimodal Large Language Models
by: Zang, Yuhang, et al.
Published: (2023)
by: Zang, Yuhang, et al.
Published: (2023)
DifFace: Blind Face Restoration with Diffused Error Contraction
by: Yue, Zongsheng, et al.
Published: (2022)
by: Yue, Zongsheng, et al.
Published: (2022)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
Kalman-Inspired Feature Propagation for Video Face Super-Resolution
by: Feng, Ruicheng, et al.
Published: (2024)
by: Feng, Ruicheng, et al.
Published: (2024)
Generalizable Implicit Motion Modeling for Video Frame Interpolation
by: Guo, Zujin, et al.
Published: (2024)
by: Guo, Zujin, et al.
Published: (2024)
AITTI: Learning Adaptive Inclusive Token for Text-to-Image Generation
by: Hou, Xinyu, et al.
Published: (2024)
by: Hou, Xinyu, et al.
Published: (2024)
ObjFiller3D: Scaling 3D Object Inpainting to Dense Multi-View Consistency
by: Feng, Haitang, et al.
Published: (2025)
by: Feng, Haitang, et al.
Published: (2025)
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
by: Yang, Shuai, et al.
Published: (2024)
by: Yang, Shuai, et al.
Published: (2024)
EdgeSAM: Prompt-In-the-Loop Distillation for SAM
by: Zhou, Chong, et al.
Published: (2023)
by: Zhou, Chong, et al.
Published: (2023)
Training-free Camera Control for Video Generation
by: Hou, Chen, et al.
Published: (2024)
by: Hou, Chen, et al.
Published: (2024)
DualCamCtrl: Dual-Branch Diffusion Model for Geometry-Aware Camera-Controlled Video Generation
by: Zhang, Hongfei, et al.
Published: (2025)
by: Zhang, Hongfei, et al.
Published: (2025)
CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control
by: Popov, Stefan, et al.
Published: (2025)
by: Popov, Stefan, et al.
Published: (2025)
ObjEmbed: Towards Universal Multimodal Object Embeddings
by: Fu, Shenghao, et al.
Published: (2026)
by: Fu, Shenghao, et al.
Published: (2026)
Controllable Human-centric Keyframe Interpolation with Generative Prior
by: Guo, Zujin, et al.
Published: (2025)
by: Guo, Zujin, et al.
Published: (2025)
Efficient Diffusion Model for Image Restoration by Residual Shifting
by: Yue, Zongsheng, et al.
Published: (2024)
by: Yue, Zongsheng, et al.
Published: (2024)
Similar Items
-
3DEnhancer: Consistent Multi-View Diffusion for 3D Enhancement
by: Luo, Yihang, et al.
Published: (2024) -
4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere
by: Luo, Yihang, et al.
Published: (2026) -
GaussianAnything: Interactive Point Cloud Flow Matching For 3D Object Generation
by: Lan, Yushi, et al.
Published: (2024) -
Denoising as Adaptation: Noise-Space Domain Adaptation for Image Restoration
by: Liao, Kang, et al.
Published: (2024) -
Precise Object and Effect Removal with Adaptive Target-Aware Attention
by: Zhao, Jixin, et al.
Published: (2025)