UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Xi, Zhang, Zhifei, Zhang, He, Zhou, Yuqian, Kim, Soo Ye, Liu, Qing, Li, Yijun, Zhang, Jianming, Zhao, Nanxuan, Wang, Yilin, Ding, Hui, Lin, Zhe, Zhao, Hengshuang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
di: Ju, Xuan, et al.
Pubblicazione: (2025)
di: Ju, Xuan, et al.
Pubblicazione: (2025)
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
di: Dalva, Yusuf, et al.
Pubblicazione: (2024)
di: Dalva, Yusuf, et al.
Pubblicazione: (2024)
UniHuman: A Unified Model for Editing Human Images in the Wild
di: Li, Nannan, et al.
Pubblicazione: (2023)
di: Li, Nannan, et al.
Pubblicazione: (2023)
SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing
di: Gu, Jing, et al.
Pubblicazione: (2024)
di: Gu, Jing, et al.
Pubblicazione: (2024)
Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and Editing
di: Zhang, Shilong, et al.
Pubblicazione: (2025)
di: Zhang, Shilong, et al.
Pubblicazione: (2025)
Beyond Simple Edits: X-Planner for Complex Instruction-Based Image Editing
di: Yeh, Chun-Hsiao, et al.
Pubblicazione: (2025)
di: Yeh, Chun-Hsiao, et al.
Pubblicazione: (2025)
Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction
di: Cai, Yuanhao, et al.
Pubblicazione: (2024)
di: Cai, Yuanhao, et al.
Pubblicazione: (2024)
Tri-Prompting: Video Diffusion with Unified Control over Scene, Subject, and Motion
di: Zhou, Zhenghong, et al.
Pubblicazione: (2026)
di: Zhou, Zhenghong, et al.
Pubblicazione: (2026)
FINECAPTION: Compositional Image Captioning Focusing on Wherever You Want at Any Granularity
di: Hua, Hang, et al.
Pubblicazione: (2024)
di: Hua, Hang, et al.
Pubblicazione: (2024)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
di: Zhang, Peiying, et al.
Pubblicazione: (2025)
di: Zhang, Peiying, et al.
Pubblicazione: (2025)
UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
di: Liu, Zhe, et al.
Pubblicazione: (2025)
di: Liu, Zhe, et al.
Pubblicazione: (2025)
Generative Image Layer Decomposition with Visual Effects
di: Yang, Jinrui, et al.
Pubblicazione: (2024)
di: Yang, Jinrui, et al.
Pubblicazione: (2024)
Zero-shot Image Editing with Reference Imitation
di: Chen, Xi, et al.
Pubblicazione: (2024)
di: Chen, Xi, et al.
Pubblicazione: (2024)
UniMatch V2: Pushing the Limit of Semi-Supervised Semantic Segmentation
di: Yang, Lihe, et al.
Pubblicazione: (2024)
di: Yang, Lihe, et al.
Pubblicazione: (2024)
IRGPT: Understanding Real-world Infrared Image with Bi-cross-modal Curriculum on Large-scale Benchmark
di: Cao, Zhe, et al.
Pubblicazione: (2025)
di: Cao, Zhe, et al.
Pubblicazione: (2025)
FASTER: Rethinking Real-Time Flow VLAs
di: Lu, Yuxiang, et al.
Pubblicazione: (2026)
di: Lu, Yuxiang, et al.
Pubblicazione: (2026)
ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services
di: Ji, Fengxian, et al.
Pubblicazione: (2026)
di: Ji, Fengxian, et al.
Pubblicazione: (2026)
ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
di: Yu, Yongsheng, et al.
Pubblicazione: (2025)
Linear Image Generation by Synthesizing Exposure Brackets
di: Dai, Yuekun, et al.
Pubblicazione: (2026)
di: Dai, Yuekun, et al.
Pubblicazione: (2026)
Animate-X++: Universal Character Image Animation with Dynamic Backgrounds
di: Tan, Shuai, et al.
Pubblicazione: (2025)
di: Tan, Shuai, et al.
Pubblicazione: (2025)
Learning an Image Editing Model without Image Editing Pairs
di: Kumari, Nupur, et al.
Pubblicazione: (2025)
di: Kumari, Nupur, et al.
Pubblicazione: (2025)
FocalClick-XL: Towards Unified and High-quality Interactive Segmentation
di: Chen, Xi, et al.
Pubblicazione: (2025)
di: Chen, Xi, et al.
Pubblicazione: (2025)
IMPRINT: Generative Object Compositing by Learning Identity-Preserving Representation
di: Song, Yizhi, et al.
Pubblicazione: (2024)
di: Song, Yizhi, et al.
Pubblicazione: (2024)
Thinking Outside the BBox: Unconstrained Generative Object Compositing
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2024)
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2024)
Text-to-Vector Generation with Neural Path Representation
di: Zhang, Peiying, et al.
Pubblicazione: (2024)
di: Zhang, Peiying, et al.
Pubblicazione: (2024)
From Fragment to One Piece: A Survey on AI-Driven Graphic Design
di: Zou, Xingxing, et al.
Pubblicazione: (2025)
di: Zou, Xingxing, et al.
Pubblicazione: (2025)
Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation
di: Xie, Zhifei, et al.
Pubblicazione: (2026)
di: Xie, Zhifei, et al.
Pubblicazione: (2026)
Sim-to-Real Dynamic Object Manipulation on Conveyor Systems via Optimization Path Shaping
di: Li, Zhuoling, et al.
Pubblicazione: (2025)
di: Li, Zhuoling, et al.
Pubblicazione: (2025)
Uni-Fusion: Universal Continuous Mapping
di: Yuan, Yijun, et al.
Pubblicazione: (2023)
di: Yuan, Yijun, et al.
Pubblicazione: (2023)
Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment
di: Song, Yizhi, et al.
Pubblicazione: (2024)
di: Song, Yizhi, et al.
Pubblicazione: (2024)
Restoring Real-World Images with an Internal Detail Enhancement Diffusion Model
di: Xiao, Peng, et al.
Pubblicazione: (2025)
di: Xiao, Peng, et al.
Pubblicazione: (2025)
Kernel Adversarial Learning for Real-world Image Super-resolution
di: Wang, Hu, et al.
Pubblicazione: (2021)
di: Wang, Hu, et al.
Pubblicazione: (2021)
Notational Animating: An Interactive Approach to Creating and Editing Animation Keyframes
di: Shi, Xinyu, et al.
Pubblicazione: (2026)
di: Shi, Xinyu, et al.
Pubblicazione: (2026)
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions
di: Cai, Yuanhao, et al.
Pubblicazione: (2025)
di: Cai, Yuanhao, et al.
Pubblicazione: (2025)
UniMERNet: A Universal Network for Real-World Mathematical Expression Recognition
di: Wang, Bin, et al.
Pubblicazione: (2024)
di: Wang, Bin, et al.
Pubblicazione: (2024)
UniPAD: A Universal Pre-training Paradigm for Autonomous Driving
di: Yang, Honghui, et al.
Pubblicazione: (2023)
di: Yang, Honghui, et al.
Pubblicazione: (2023)
CompSplat: Compression-aware 3D Gaussian Splatting for Real-world Video
di: Song, Hojun, et al.
Pubblicazione: (2026)
di: Song, Hojun, et al.
Pubblicazione: (2026)
Multitwine: Multi-Object Compositing with Text and Layout Control
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2025)
di: Tarrés, Gemma Canet, et al.
Pubblicazione: (2025)
GroupDiff: Diffusion-based Group Portrait Editing
di: Jiang, Yuming, et al.
Pubblicazione: (2024)
di: Jiang, Yuming, et al.
Pubblicazione: (2024)
Bezier Splatting for Fast and Differentiable Vector Graphics Rendering
di: Liu, Xi, et al.
Pubblicazione: (2025)
di: Liu, Xi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning
di: Ju, Xuan, et al.
Pubblicazione: (2025) -
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
di: Dalva, Yusuf, et al.
Pubblicazione: (2024) -
UniHuman: A Unified Model for Editing Human Images in the Wild
di: Li, Nannan, et al.
Pubblicazione: (2023) -
SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing
di: Gu, Jing, et al.
Pubblicazione: (2024) -
Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and Editing
di: Zhang, Shilong, et al.
Pubblicazione: (2025)