DreamO: A Unified Framework for Image Customization
Fuente:
arXiv
Saved in:
| Main Authors: | Mou, Chong, Wu, Yanze, Wu, Wenxu, Guo, Zinan, Zhang, Pengze, Cheng, Yufeng, Luo, Yiming, Ding, Fei, Zhang, Shiwen, Li, Xinghui, Li, Mengtian, Liu, Mingcong, Zhang, Yi, Wu, Shaojin, Zhao, Songtao, Zhang, Jian, He, Qian, Wu, Xinglong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MUSAR: Exploring Multi-Subject Customization from Single-Subject Dataset via Attention Routing
by: Guo, Zinan, et al.
Published: (2025)
by: Guo, Zinan, et al.
Published: (2025)
InstructX: Towards Unified Visual Editing with MLLM Guidance
by: Mou, Chong, et al.
Published: (2025)
by: Mou, Chong, et al.
Published: (2025)
UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward
by: Cheng, Yufeng, et al.
Published: (2025)
by: Cheng, Yufeng, et al.
Published: (2025)
DreamID: High-Fidelity and Fast diffusion-based Face Swapping via Triplet ID Group Learning
by: Ye, Fulong, et al.
Published: (2025)
by: Ye, Fulong, et al.
Published: (2025)
OmniTransfer: All-in-one Framework for Spatio-temporal Video Transfer
by: Zhang, Pengze, et al.
Published: (2026)
by: Zhang, Pengze, et al.
Published: (2026)
USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
by: Wu, Shaojin, et al.
Published: (2025)
by: Wu, Shaojin, et al.
Published: (2025)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
by: Wu, Shaojin, et al.
Published: (2025)
by: Wu, Shaojin, et al.
Published: (2025)
CDST: Color Disentangled Style Transfer for Universal Style Reference Customization
by: Zhang, Shiwen, et al.
Published: (2025)
by: Zhang, Shiwen, et al.
Published: (2025)
PuLID: Pure and Lightning ID Customization via Contrastive Alignment
by: Guo, Zinan, et al.
Published: (2024)
by: Guo, Zinan, et al.
Published: (2024)
DreamStyle: A Unified Framework for Video Stylization
by: Li, Mengtian, et al.
Published: (2026)
by: Li, Mengtian, et al.
Published: (2026)
DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer
by: Guo, Xu, et al.
Published: (2026)
by: Guo, Xu, et al.
Published: (2026)
DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation
by: Guo, Xu, et al.
Published: (2026)
by: Guo, Xu, et al.
Published: (2026)
DreamOmni: Unified Image Generation and Editing
by: Xia, Bin, et al.
Published: (2024)
by: Xia, Bin, et al.
Published: (2024)
RealCustom++: Representing Images as Real Textual Word for Real-Time Customization
by: Mao, Zhendong, et al.
Published: (2024)
by: Mao, Zhendong, et al.
Published: (2024)
AnyDressing: Customizable Multi-Garment Virtual Dressing via Latent Diffusion Models
by: Li, Xinghui, et al.
Published: (2024)
by: Li, Xinghui, et al.
Published: (2024)
DreamPoster: A Unified Framework for Image-Conditioned Generative Poster Design
by: Hu, Xiwei, et al.
Published: (2025)
by: Hu, Xiwei, et al.
Published: (2025)
Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset
by: Chen, Zhuowei, et al.
Published: (2025)
by: Chen, Zhuowei, et al.
Published: (2025)
VMix: Improving Text-to-Image Diffusion Model with Cross-Attention Mixing Control
by: Wu, Shaojin, et al.
Published: (2024)
by: Wu, Shaojin, et al.
Published: (2024)
DreamVE: Unified Instruction-based Image and Video Editing
by: Xia, Bin, et al.
Published: (2025)
by: Xia, Bin, et al.
Published: (2025)
Four New Bisabolane‐Type Sesquiterpenoids From Sanguisorba officinalis
by: Longlong Wu, et al.
Published: (2025)
by: Longlong Wu, et al.
Published: (2025)
Multilayer Routing and Resource Assignment in Spatial Channel Networks (SCNs): Oriented Toward the Massive SDM Era
by: Yang, Mingcong, et al.
Published: (2020)
by: Yang, Mingcong, et al.
Published: (2020)
Recent Advances in Electrodeposited Iridium and Ruthenium‐Based Electrocatalysts for Acidic Water Electrolysis
by: Wenxu Qi, et al.
Published: (2026)
by: Wenxu Qi, et al.
Published: (2026)
OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models
by: Chen, Jinshu, et al.
Published: (2025)
by: Chen, Jinshu, et al.
Published: (2025)
RealCustom: Narrowing Real Text Word for Real-Time Open-Domain Text-to-Image Customization
by: Huang, Mengqi, et al.
Published: (2024)
by: Huang, Mengqi, et al.
Published: (2024)
DreamRelation: Bridging Customization and Relation Generation
by: Shi, Qingyu, et al.
Published: (2024)
by: Shi, Qingyu, et al.
Published: (2024)
Charon: A Unified and Fine-Grained Simulator for Large-Scale LLM Training and Inference
by: Yang, Mengtian, et al.
Published: (2026)
by: Yang, Mengtian, et al.
Published: (2026)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
by: Dharmarajan, Karthik, et al.
Published: (2025)
by: Dharmarajan, Karthik, et al.
Published: (2025)
Lance: Unified Multimodal Modeling by Multi-Task Synergy
by: Fu, Fengyi, et al.
Published: (2026)
by: Fu, Fengyi, et al.
Published: (2026)
DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing
by: Wang, Weitao, et al.
Published: (2025)
by: Wang, Weitao, et al.
Published: (2025)
Bounding the number of reticulation events for displaying multiple trees in a phylogenetic network
by: Wu, Yufeng, et al.
Published: (2024)
by: Wu, Yufeng, et al.
Published: (2024)
ReEXplore: Improving MLLMs for Embodied Exploration with Contextualized Retrospective Experience Replay
by: Zhang, Gengyuan, et al.
Published: (2025)
by: Zhang, Gengyuan, et al.
Published: (2025)
Joint identification of spatially variable genes via a network-assisted Bayesian regularization approach
by: Wu, Mingcong, et al.
Published: (2024)
by: Wu, Mingcong, et al.
Published: (2024)
DreamOmni2: Multimodal Instruction-based Editing and Generation
by: Xia, Bin, et al.
Published: (2025)
by: Xia, Bin, et al.
Published: (2025)
Learning A Multi-Task Transformer Via Unified And Customized Instruction Tuning For Chest Radiograph Interpretation
by: Xu, Lijian, et al.
Published: (2023)
by: Xu, Lijian, et al.
Published: (2023)
DreamRelation: Relation-Centric Video Customization
by: Wei, Yujie, et al.
Published: (2025)
by: Wei, Yujie, et al.
Published: (2025)
Large, Complex, and Realistic Safety Clothing and Helmet Detection: Dataset and Method
by: Yu, Fusheng, et al.
Published: (2023)
by: Yu, Fusheng, et al.
Published: (2023)
HyperLoRA: Parameter-Efficient Adaptive Generation for Portrait Synthesis
by: Li, Mengtian, et al.
Published: (2025)
by: Li, Mengtian, et al.
Published: (2025)
thuml/CompilerDream: CompilerDream Release v0.1
by: Jialong Wu
Published: (2025)
by: Jialong Wu
Published: (2025)
Semantic Correspondence: Unified Benchmarking and a Strong Baseline
by: Zhang, Kaiyan, et al.
Published: (2025)
by: Zhang, Kaiyan, et al.
Published: (2025)
MegaSR: Mining Customized Semantics and Expressive Guidance for Real-World Image Super-Resolution
by: Li, Xinrui, et al.
Published: (2025)
by: Li, Xinrui, et al.
Published: (2025)
Similar Items
-
MUSAR: Exploring Multi-Subject Customization from Single-Subject Dataset via Attention Routing
by: Guo, Zinan, et al.
Published: (2025) -
InstructX: Towards Unified Visual Editing with MLLM Guidance
by: Mou, Chong, et al.
Published: (2025) -
UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward
by: Cheng, Yufeng, et al.
Published: (2025) -
DreamID: High-Fidelity and Fast diffusion-based Face Swapping via Triplet ID Group Learning
by: Ye, Fulong, et al.
Published: (2025) -
OmniTransfer: All-in-one Framework for Spatio-temporal Video Transfer
by: Zhang, Pengze, et al.
Published: (2026)