The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hanlin, Ouyang, Hao, Wang, Qiuyu, Yu, Yue, Meng, Yihao, Wang, Wen, Cheng, Ka Leong, Ma, Shuailei, Bai, Qingyan, Li, Yixuan, Chen, Cheng, Zeng, Yanhong, Zhu, Xing, Shen, Yujun, Chen, Qifeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
by: Wang, Hanlin, et al.
Published: (2024)
by: Wang, Hanlin, et al.
Published: (2024)
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
by: Bai, Qingyan, et al.
Published: (2025)
by: Bai, Qingyan, et al.
Published: (2025)
MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Edicho: Consistent Image Editing in the Wild
by: Bai, Qingyan, et al.
Published: (2024)
by: Bai, Qingyan, et al.
Published: (2024)
Calligrapher: Freestyle Text Image Customization
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
by: Meng, Yihao, et al.
Published: (2025)
by: Meng, Yihao, et al.
Published: (2025)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
by: Meng, Yihao, et al.
Published: (2026)
by: Meng, Yihao, et al.
Published: (2026)
AniDoc: Animation Creation Made Easier
by: Meng, Yihao, et al.
Published: (2024)
by: Meng, Yihao, et al.
Published: (2024)
Learning Naturally Aggregated Appearance for Efficient 3D Editing
by: Cheng, Ka Leong, et al.
Published: (2023)
by: Cheng, Ka Leong, et al.
Published: (2023)
Advancing Open-source World Models
by: Robbyant Team, et al.
Published: (2026)
by: Robbyant Team, et al.
Published: (2026)
MagicQuill: An Intelligent Interactive Image Editing System
by: Liu, Zichen, et al.
Published: (2024)
by: Liu, Zichen, et al.
Published: (2024)
DepthLab: From Partial to Complete
by: Liu, Zhiheng, et al.
Published: (2024)
by: Liu, Zhiheng, et al.
Published: (2024)
CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
by: Ouyang, Hao, et al.
Published: (2023)
by: Ouyang, Hao, et al.
Published: (2023)
Real-time 3D-aware Portrait Editing from a Single Image
by: Bai, Qingyan, et al.
Published: (2024)
by: Bai, Qingyan, et al.
Published: (2024)
MangaNinja: Line Art Colorization with Precise Reference Following
by: Liu, Zhiheng, et al.
Published: (2025)
by: Liu, Zhiheng, et al.
Published: (2025)
Towards Degradation-Robust Reconstruction in Generalizable NeRF
by: Park, Chan Ho, et al.
Published: (2024)
by: Park, Chan Ho, et al.
Published: (2024)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
by: Lu, Yunhong, et al.
Published: (2025)
by: Lu, Yunhong, et al.
Published: (2025)
Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation
by: Chen, Qihua, et al.
Published: (2024)
by: Chen, Qihua, et al.
Published: (2024)
InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
by: Liu, Zhiheng, et al.
Published: (2024)
by: Liu, Zhiheng, et al.
Published: (2024)
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
by: Yuan, Runtian, et al.
Published: (2025)
by: Yuan, Runtian, et al.
Published: (2025)
STeCa: Step-level Trajectory Calibration for LLM Agent Learning
by: Wang, Hanlin, et al.
Published: (2025)
by: Wang, Hanlin, et al.
Published: (2025)
DiffBrush:Just Painting the Art by Your Hands
by: Chu, Jiaming, et al.
Published: (2025)
by: Chu, Jiaming, et al.
Published: (2025)
Spectroscopic Investigations of a Vandalized Contemporary Acrylic Painting on Canvas Using Model Paintings and Chemometrics
by: Jana Striova, et al.
Published: (2025)
by: Jana Striova, et al.
Published: (2025)
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
by: Zhang, Yiming, et al.
Published: (2023)
by: Zhang, Yiming, et al.
Published: (2023)
Follow-Your-Color: Multi-Instance Sketch Colorization
by: Zhang, Yinhan, et al.
Published: (2025)
by: Zhang, Yinhan, et al.
Published: (2025)
Social-Transmotion: Promptable Human Trajectory Prediction
by: Saadatnejad, Saeed, et al.
Published: (2023)
by: Saadatnejad, Saeed, et al.
Published: (2023)
LeftRefill: Filling Right Canvas based on Left Reference through Generalized Text-to-Image Diffusion Model
by: Cao, Chenjie, et al.
Published: (2023)
by: Cao, Chenjie, et al.
Published: (2023)
Skill-Evolving Grounded Reasoning for Free-Text Promptable 3D Medical Image Segmentation
by: Zhang, Tongrui, et al.
Published: (2026)
by: Zhang, Tongrui, et al.
Published: (2026)
Degradation of Synthetic Restoration Materials by Xerotolerant/Xerophilic Fungi Contaminating Canvas Paintings.
by: Kujović, Amela, et al.
Published: (2025)
by: Kujović, Amela, et al.
Published: (2025)
Make-It-Vivid: Dressing Your Animatable Biped Cartoon Characters from Text
by: Tang, Junshu, et al.
Published: (2024)
by: Tang, Junshu, et al.
Published: (2024)
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
by: Ma, Yue, et al.
Published: (2023)
by: Ma, Yue, et al.
Published: (2023)
HeadArtist: Text-conditioned 3D Head Generation with Self Score Distillation
by: Liu, Hongyu, et al.
Published: (2023)
by: Liu, Hongyu, et al.
Published: (2023)
Framer: Interactive Frame Interpolation
by: Wang, Wen, et al.
Published: (2024)
by: Wang, Wen, et al.
Published: (2024)
Comment on “The Development Trajectories of Media Exposure and Cognitive Function in Older Adults: An Analysis Based on Latent Growth Models”
by: Qiuyu Wang, et al.
Published: (2026)
by: Qiuyu Wang, et al.
Published: (2026)
MV-SAM: Multi-view Promptable Segmentation using Pointmap Guidance
by: Jeong, Yoonwoo, et al.
Published: (2026)
by: Jeong, Yoonwoo, et al.
Published: (2026)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
by: Huang, Yiming, et al.
Published: (2025)
by: Huang, Yiming, et al.
Published: (2025)
No Two Devils Alike: Unveiling Distinct Mechanisms of Fine-tuning Attacks
by: Leong, Chak Tou, et al.
Published: (2024)
by: Leong, Chak Tou, et al.
Published: (2024)
Inverse Painting: Reconstructing The Painting Process
by: Chen, Bowei, et al.
Published: (2024)
by: Chen, Bowei, et al.
Published: (2024)
MaPa: Text-driven Photorealistic Material Painting for 3D Shapes
by: Zhang, Shangzan, et al.
Published: (2024)
by: Zhang, Shangzan, et al.
Published: (2024)
PLUTO: Pushing the Limit of Imitation Learning-based Planning for Autonomous Driving
by: Cheng, Jie, et al.
Published: (2024)
by: Cheng, Jie, et al.
Published: (2024)
Similar Items
-
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
by: Wang, Hanlin, et al.
Published: (2024) -
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
by: Bai, Qingyan, et al.
Published: (2025) -
MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
by: Liu, Zichen, et al.
Published: (2025) -
Edicho: Consistent Image Editing in the Wild
by: Bai, Qingyan, et al.
Published: (2024) -
Calligrapher: Freestyle Text Image Customization
by: Ma, Yue, et al.
Published: (2025)