Guardado en:
| Autores principales: | Liu, Zhiheng, Ouyang, Hao, Wang, Qiuyu, Cheng, Ka Leong, Xiao, Jie, Zhu, Kai, Xue, Nan, Liu, Yu, Shen, Yujun, Cao, Yang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.11613 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DepthLab: From Partial to Complete
por: Liu, Zhiheng, et al.
Publicado: (2024)
por: Liu, Zhiheng, et al.
Publicado: (2024)
MagicQuill: An Intelligent Interactive Image Editing System
por: Liu, Zichen, et al.
Publicado: (2024)
por: Liu, Zichen, et al.
Publicado: (2024)
AniDoc: Animation Creation Made Easier
por: Meng, Yihao, et al.
Publicado: (2024)
por: Meng, Yihao, et al.
Publicado: (2024)
MangaNinja: Line Art Colorization with Precise Reference Following
por: Liu, Zhiheng, et al.
Publicado: (2025)
por: Liu, Zhiheng, et al.
Publicado: (2025)
Edicho: Consistent Image Editing in the Wild
por: Bai, Qingyan, et al.
Publicado: (2024)
por: Bai, Qingyan, et al.
Publicado: (2024)
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
por: Wang, Hanlin, et al.
Publicado: (2024)
por: Wang, Hanlin, et al.
Publicado: (2024)
Learning Naturally Aggregated Appearance for Efficient 3D Editing
por: Cheng, Ka Leong, et al.
Publicado: (2023)
por: Cheng, Ka Leong, et al.
Publicado: (2023)
VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization
por: Fang, Zixun, et al.
Publicado: (2025)
por: Fang, Zixun, et al.
Publicado: (2025)
Calligrapher: Freestyle Text Image Customization
por: Ma, Yue, et al.
Publicado: (2025)
por: Ma, Yue, et al.
Publicado: (2025)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
por: Lu, Yunhong, et al.
Publicado: (2025)
por: Lu, Yunhong, et al.
Publicado: (2025)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
por: Meng, Yihao, et al.
Publicado: (2025)
por: Meng, Yihao, et al.
Publicado: (2025)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
por: Meng, Yihao, et al.
Publicado: (2026)
por: Meng, Yihao, et al.
Publicado: (2026)
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
por: Bai, Qingyan, et al.
Publicado: (2025)
por: Bai, Qingyan, et al.
Publicado: (2025)
MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
por: Wang, Hanlin, et al.
Publicado: (2025)
por: Wang, Hanlin, et al.
Publicado: (2025)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
por: Shao, Jiahao, et al.
Publicado: (2024)
por: Shao, Jiahao, et al.
Publicado: (2024)
NeRF Inpainting with Geometric Diffusion Prior and Balanced Score Distillation
por: Zhang, Menglin, et al.
Publicado: (2024)
por: Zhang, Menglin, et al.
Publicado: (2024)
P‐5.6: Synergistic Low‐Light Image Enhancement: A Fusion of Dark Channel Dehazing and K‐means Clustering
por: Nan Xue, et al.
Publicado: (2024)
por: Nan Xue, et al.
Publicado: (2024)
Rectified Diffusion Guidance for Conditional Generation
por: Xia, Mengfei, et al.
Publicado: (2024)
por: Xia, Mengfei, et al.
Publicado: (2024)
Geometric Context Transformer for Streaming 3D Reconstruction
por: Chen, Lin-Zhuo, et al.
Publicado: (2026)
por: Chen, Lin-Zhuo, et al.
Publicado: (2026)
RI3D: Few-Shot Gaussian Splatting With Repair and Inpainting Diffusion Priors
por: Paliwal, Avinash, et al.
Publicado: (2025)
por: Paliwal, Avinash, et al.
Publicado: (2025)
UCD: Unconditional Discriminator Promotes Nash Equilibrium in GANs
por: Xia, Mengfei, et al.
Publicado: (2025)
por: Xia, Mengfei, et al.
Publicado: (2025)
Framer: Interactive Frame Interpolation
por: Wang, Wen, et al.
Publicado: (2024)
por: Wang, Wen, et al.
Publicado: (2024)
EraserDiT: Fast Video Inpainting with Diffusion Transformer Model
por: Liu, Jie, et al.
Publicado: (2025)
por: Liu, Jie, et al.
Publicado: (2025)
CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
por: Ouyang, Hao, et al.
Publicado: (2023)
por: Ouyang, Hao, et al.
Publicado: (2023)
Gaussian Belief Propagation Network for Depth Completion
por: Tang, Jie, et al.
Publicado: (2026)
por: Tang, Jie, et al.
Publicado: (2026)
Advancing Open-source World Models
por: Robbyant Team, et al.
Publicado: (2026)
por: Robbyant Team, et al.
Publicado: (2026)
A Recovery Theory for Diffusion Priors: Deterministic Analysis of the Implicit Prior Algorithm
por: Leong, Oscar, et al.
Publicado: (2025)
por: Leong, Oscar, et al.
Publicado: (2025)
InpDiffusion: Image Inpainting Localization via Conditional Diffusion Models
por: Wang, Kai, et al.
Publicado: (2025)
por: Wang, Kai, et al.
Publicado: (2025)
Diffusion-Based Depth Inpainting for Transparent and Reflective Objects
por: Sun, Tianyu, et al.
Publicado: (2024)
por: Sun, Tianyu, et al.
Publicado: (2024)
Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models
por: Wang, Wen, et al.
Publicado: (2025)
por: Wang, Wen, et al.
Publicado: (2025)
Structured Diffusion Models with Mixture of Gaussians as Prior Distribution
por: Jia, Nanshan, et al.
Publicado: (2024)
por: Jia, Nanshan, et al.
Publicado: (2024)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
por: Shriram, Jaidev, et al.
Publicado: (2024)
por: Shriram, Jaidev, et al.
Publicado: (2024)
Stealing Stable Diffusion Prior for Robust Monocular Depth Estimation
por: Mao, Yifan, et al.
Publicado: (2024)
por: Mao, Yifan, et al.
Publicado: (2024)
Seeing through Satellite Images at Street Views
por: Qian, Ming, et al.
Publicado: (2025)
por: Qian, Ming, et al.
Publicado: (2025)
3D Gaussian Inpainting with Depth-Guided Cross-View Consistency
por: Huang, Sheng-Yu, et al.
Publicado: (2025)
por: Huang, Sheng-Yu, et al.
Publicado: (2025)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
por: Fang, Zixun, et al.
Publicado: (2025)
por: Fang, Zixun, et al.
Publicado: (2025)
DISCO: Language-Guided Manipulation with Diffusion Policies and Constrained Inpainting
por: Hao, Ce, et al.
Publicado: (2024)
por: Hao, Ce, et al.
Publicado: (2024)
Distilling Textual Priors from LLM to Efficient Image Fusion
por: Zhang, Ran, et al.
Publicado: (2025)
por: Zhang, Ran, et al.
Publicado: (2025)
Masked Depth Modeling for Spatial Perception
por: Tan, Bin, et al.
Publicado: (2026)
por: Tan, Bin, et al.
Publicado: (2026)
Ejemplares similares
-
DepthLab: From Partial to Complete
por: Liu, Zhiheng, et al.
Publicado: (2024) -
MagicQuill: An Intelligent Interactive Image Editing System
por: Liu, Zichen, et al.
Publicado: (2024) -
AniDoc: Animation Creation Made Easier
por: Meng, Yihao, et al.
Publicado: (2024) -
MangaNinja: Line Art Colorization with Precise Reference Following
por: Liu, Zhiheng, et al.
Publicado: (2025) -
Edicho: Consistent Image Editing in the Wild
por: Bai, Qingyan, et al.
Publicado: (2024)