I2V3D: Controllable image-to-video generation with 3D guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhiyuan, Chen, Dongdong, Liao, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
Chat2Layout: Interactive 3D Furniture Layout with a Multimodal LLM
by: Wang, Can, et al.
Published: (2024)
by: Wang, Can, et al.
Published: (2024)
C3DAG: Controlled 3D Animal Generation using 3D pose guidance
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
Anti-I2V: Safeguarding your photos from malicious image-to-video generation
by: Vu, Duc, et al.
Published: (2026)
by: Vu, Duc, et al.
Published: (2026)
Multi-robot autonomous 3D reconstruction using Gaussian splatting with Semantic guidance
by: Zeng, Jing, et al.
Published: (2024)
by: Zeng, Jing, et al.
Published: (2024)
Distractor-free Generalizable 3D Gaussian Splatting
by: Bao, Yanqi, et al.
Published: (2024)
by: Bao, Yanqi, et al.
Published: (2024)
Omni-3DEdit: Generalized Versatile 3D Editing in One-Pass
by: Liyi, Chen, et al.
Published: (2026)
by: Liyi, Chen, et al.
Published: (2026)
OrientDream: Streamlining Text-to-3D Generation with Explicit Orientation Control
by: Huang, Yuzhong, et al.
Published: (2024)
by: Huang, Yuzhong, et al.
Published: (2024)
SGEdit: Bridging LLM with Text2Image Generative Model for Scene Graph-based Image Editing
by: Zhang, Zhiyuan, et al.
Published: (2024)
by: Zhang, Zhiyuan, et al.
Published: (2024)
Live image-based neurosurgical guidance and roadmap generation using unsupervised embedding
by: Sarwin, Gary, et al.
Published: (2023)
by: Sarwin, Gary, et al.
Published: (2023)
Geometry aware 3D generation from in-the-wild images in ImageNet
by: Shen, Qijia, et al.
Published: (2024)
by: Shen, Qijia, et al.
Published: (2024)
3D microstructural generation from 2D images of cement paste using generative adversarial networks
by: Zhao, Xin, et al.
Published: (2022)
by: Zhao, Xin, et al.
Published: (2022)
3DProxyImg: Controllable 3D-Aware Animation Synthesis from Single Image via 2D-3D Aligned Proxy Embedding
by: Zhu, Yupeng, et al.
Published: (2025)
by: Zhu, Yupeng, et al.
Published: (2025)
3D representation in 512-Byte:Variational tokenizer is the key for autoregressive 3D generation
by: Zhang, Jinzhi, et al.
Published: (2024)
by: Zhang, Jinzhi, et al.
Published: (2024)
Brain3D: Generating 3D Objects from fMRI
by: Yang, Yuankun, et al.
Published: (2024)
by: Yang, Yuankun, et al.
Published: (2024)
V3D: Video Diffusion Models are Effective 3D Generators
by: Chen, Zilong, et al.
Published: (2024)
by: Chen, Zilong, et al.
Published: (2024)
EZIGen: Enhancing zero-shot personalized image generation with precise subject encoding and decoupled guidance
by: Duan, Zicheng, et al.
Published: (2024)
by: Duan, Zicheng, et al.
Published: (2024)
C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing
by: Tao, Zeng, et al.
Published: (2025)
by: Tao, Zeng, et al.
Published: (2025)
V2V3D: View-to-View Denoised 3D Reconstruction for Light-Field Microscopy
by: Zhao, Jiayin, et al.
Published: (2025)
by: Zhao, Jiayin, et al.
Published: (2025)
3D Dynamics-Aware Manipulation: Endowing Manipulation Policies with 3D Foresight
by: He, Yuxin, et al.
Published: (2025)
by: He, Yuxin, et al.
Published: (2025)
SegDINO3D: 3D Instance Segmentation Empowered by Both Image-Level and Object-Level 2D Features
by: Qu, Jinyuan, et al.
Published: (2025)
by: Qu, Jinyuan, et al.
Published: (2025)
Dream3DAvatar: Text-Controlled 3D Avatar Reconstruction from a Single Image
by: Liu, Gaofeng, et al.
Published: (2025)
by: Liu, Gaofeng, et al.
Published: (2025)
Layout-your-3D: Controllable and Precise 3D Generation with 2D Blueprint
by: Zhou, Junwei, et al.
Published: (2024)
by: Zhou, Junwei, et al.
Published: (2024)
3DSceneEditor: Controllable 3D Scene Editing with Gaussian Splatting
by: Yan, Ziyang, et al.
Published: (2024)
by: Yan, Ziyang, et al.
Published: (2024)
Online 3D reconstruction and dense tracking in endoscopic videos
by: Hayoz, Michel, et al.
Published: (2024)
by: Hayoz, Michel, et al.
Published: (2024)
PanopticNeRF-360: Panoramic 3D-to-2D Label Transfer in Urban Scenes
by: Fu, Xiao, et al.
Published: (2023)
by: Fu, Xiao, et al.
Published: (2023)
Text-to-3D Generation by 2D Editing
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
One2Scene: Geometric Consistent Explorable 3D Scene Generation from a Single Image
by: Wang, Pengfei, et al.
Published: (2026)
by: Wang, Pengfei, et al.
Published: (2026)
3DEgo: 3D Editing on the Go!
by: Khalid, Umar, et al.
Published: (2024)
by: Khalid, Umar, et al.
Published: (2024)
Towards 3D heart mesh generation using contactless radar imaging and physics-informed neural network
by: Li, Jinye, et al.
Published: (2026)
by: Li, Jinye, et al.
Published: (2026)
ScenDi: 3D-to-2D Scene Diffusion Cascades for Urban Generation
by: Guo, Hanlei, et al.
Published: (2026)
by: Guo, Hanlei, et al.
Published: (2026)
V2Edit: Versatile Video Diffusion Editor for Videos and 3D Scenes
by: Zhang, Yanming, et al.
Published: (2025)
by: Zhang, Yanming, et al.
Published: (2025)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
DoF-Gaussian: Controllable Depth-of-Field for 3D Gaussian Splatting
by: Shen, Liao, et al.
Published: (2025)
by: Shen, Liao, et al.
Published: (2025)
Text2NeRF: Text-Driven 3D Scene Generation with Neural Radiance Fields
by: Zhang, Jingbo, et al.
Published: (2023)
by: Zhang, Jingbo, et al.
Published: (2023)
PhyScene3D: Physically Consistent Interactive 3D Tabletop Scene Generation
by: Chen, Weixing, et al.
Published: (2026)
by: Chen, Weixing, et al.
Published: (2026)
Dream4D: Lifting Camera-Controlled I2V towards Spatiotemporally Consistent 4D Generation
by: Liu, Xiaoyan, et al.
Published: (2025)
by: Liu, Xiaoyan, et al.
Published: (2025)
RAPS-3D: Efficient interactive segmentation for 3D radiological imaging
by: Danielou, Théo, et al.
Published: (2025)
by: Danielou, Théo, et al.
Published: (2025)
T3DNet: Compressing Point Cloud Models for Lightweight 3D Recognition
by: Yang, Zhiyuan, et al.
Published: (2024)
by: Yang, Zhiyuan, et al.
Published: (2024)
Gaussian Control with Hierarchical Semantic Graphs in 3D Human Recovery
by: Wang, Hongsheng, et al.
Published: (2024)
by: Wang, Hongsheng, et al.
Published: (2024)
Similar Items
-
FlexTraj: Image-to-Video Generation with Flexible Point Trajectory Control
by: Zhang, Zhiyuan, et al.
Published: (2025) -
Chat2Layout: Interactive 3D Furniture Layout with a Multimodal LLM
by: Wang, Can, et al.
Published: (2024) -
C3DAG: Controlled 3D Animal Generation using 3D pose guidance
by: Mishra, Sandeep, et al.
Published: (2024) -
Anti-I2V: Safeguarding your photos from malicious image-to-video generation
by: Vu, Duc, et al.
Published: (2026) -
Multi-robot autonomous 3D reconstruction using Gaussian splatting with Semantic guidance
by: Zeng, Jing, et al.
Published: (2024)