ProcessPainter: Learn Painting Process from Sequence Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Yiren, Huang, Shijie, Yao, Chen, Ye, Xiaojun, Ci, Hai, Liu, Jiaming, Zhang, Yuxuan, Shou, Mike Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
von: Huang, Shijie, et al.
Veröffentlicht: (2025)
Steganalysis on Digital Watermarking: Is Your Defense Truly Impervious?
von: Yang, Pei, et al.
Veröffentlicht: (2024)
von: Yang, Pei, et al.
Veröffentlicht: (2024)
X-Humanoid: Robotize Human Videos to Generate Humanoid Videos at Scale
von: Yang, Pei, et al.
Veröffentlicht: (2025)
von: Yang, Pei, et al.
Veröffentlicht: (2025)
IDProtector: An Adversarial Noise Encoder to Protect Against ID-Preserving Image Generation
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification
von: Ci, Hai, et al.
Veröffentlicht: (2024)
von: Ci, Hai, et al.
Veröffentlicht: (2024)
Loomis Painter: Reconstructing the Painting Process
von: Pobitzer, Markus, et al.
Veröffentlicht: (2025)
von: Pobitzer, Markus, et al.
Veröffentlicht: (2025)
UENR-600K: A Large-Scale Physically Grounded Dataset for Nighttime Video Deraining
von: Yang, Pei, et al.
Veröffentlicht: (2026)
von: Yang, Pei, et al.
Veröffentlicht: (2026)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
von: Ci, Hai, et al.
Veröffentlicht: (2025)
von: Ci, Hai, et al.
Veröffentlicht: (2025)
Anti-Reference: Universal and Immediate Defense Against Reference-Based Generation
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
MakeAnything: Harnessing Diffusion Transformers for Multi-Domain Procedural Sequence Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
WMAdapter: Adding WaterMark Control to Latent Diffusion Models
von: Ci, Hai, et al.
Veröffentlicht: (2024)
von: Ci, Hai, et al.
Veröffentlicht: (2024)
LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
Impossible Videos
von: Bai, Zechen, et al.
Veröffentlicht: (2025)
von: Bai, Zechen, et al.
Veröffentlicht: (2025)
AnimatePainter: A Self-Supervised Rendering Framework for Reconstructing Painting Process
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
Edit2Perceive: Image Editing Diffusion Models Are Strong Dense Perceivers
von: Shi, Yiqing, et al.
Veröffentlicht: (2025)
von: Shi, Yiqing, et al.
Veröffentlicht: (2025)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
von: Song, Yiren, et al.
Veröffentlicht: (2024)
von: Song, Yiren, et al.
Veröffentlicht: (2024)
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
RobotSeg: A Model and Dataset for Segmenting Robots in Image and Video
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
Image Watermarks are Removable Using Controllable Regeneration from Clean Noise
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
von: Liu, Yepeng, et al.
Veröffentlicht: (2024)
StreamingEffect: Real-Time Human-Centric Video Effect Generation
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
Mitty: Diffusion-based Human-to-Robot Video Generation
von: Song, Yiren, et al.
Veröffentlicht: (2025)
von: Song, Yiren, et al.
Veröffentlicht: (2025)
OmniPSD: Layered PSD Generation with Diffusion Transformer
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
Inverse Painting: Reconstructing The Painting Process
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
DiffSeg30k: A Multi-Turn Diffusion Editing Benchmark for Localized AIGC Detection
von: Ci, Hai, et al.
Veröffentlicht: (2025)
von: Ci, Hai, et al.
Veröffentlicht: (2025)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
AttentionPainter: An Efficient and Adaptive Stroke Predictor for Scene Painting
von: Tang, Yizhe, et al.
Veröffentlicht: (2024)
von: Tang, Yizhe, et al.
Veröffentlicht: (2024)
Loom: Diffusion-Transformer for Interleaved Generation
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
von: Ye, Mingcheng, et al.
Veröffentlicht: (2025)
Personalized Vision via Visual In-Context Learning
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
Stable-Makeup: When Real-World Makeup Transfer Meets Diffusion Model
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2024)
SWEET: Sparse World Modeling with Image Editing for Embodied Task Execution
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptation
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
Unlocking the Latent Canvas: Eliciting and Benchmarking Symbolic Visual Expression in LLMs
von: Zheng, Yiren, et al.
Veröffentlicht: (2026)
von: Zheng, Yiren, et al.
Veröffentlicht: (2026)
Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration
von: Song, Yiren, et al.
Veröffentlicht: (2026)
von: Song, Yiren, et al.
Veröffentlicht: (2026)
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
EasyControl: Adding Efficient and Flexible Control for Diffusion Transformer
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2025)
PaintFlow: A Unified Framework for Interactive Oil Paintings Editing and Generation
von: Hu, Zhangli, et al.
Veröffentlicht: (2025)
von: Hu, Zhangli, et al.
Veröffentlicht: (2025)
PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion
von: Gao, Heyuan, et al.
Veröffentlicht: (2026)
von: Gao, Heyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PhotoDoodle: Learning Artistic Image Editing from Few-Shot Pairwise Data
von: Huang, Shijie, et al.
Veröffentlicht: (2025) -
Steganalysis on Digital Watermarking: Is Your Defense Truly Impervious?
von: Yang, Pei, et al.
Veröffentlicht: (2024) -
X-Humanoid: Robotize Human Videos to Generate Humanoid Videos at Scale
von: Yang, Pei, et al.
Veröffentlicht: (2025) -
IDProtector: An Adversarial Noise Encoder to Protect Against ID-Preserving Image Generation
von: Song, Yiren, et al.
Veröffentlicht: (2024) -
RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification
von: Ci, Hai, et al.
Veröffentlicht: (2024)