OMG: Occlusion-friendly Personalized Multi-concept Generation in Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Zhe, Zhang, Yong, Yang, Tianyu, Wang, Tao, Zhang, Kaihao, Wu, Bizhu, Chen, Guanying, Liu, Wei, Luo, Wenhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
MaterialMVP: Illumination-Invariant Material Generation via Multi-view PBR Diffusion
by: He, Zebin, et al.
Published: (2025)
by: He, Zebin, et al.
Published: (2025)
Towards Real-World Blind Face Restoration with Generative Diffusion Prior
by: Chen, Xiaoxu, et al.
Published: (2023)
by: Chen, Xiaoxu, et al.
Published: (2023)
Dual Teacher Knowledge Distillation with Domain Alignment for Face Anti-spoofing
by: Kong, Zhe, et al.
Published: (2024)
by: Kong, Zhe, et al.
Published: (2024)
MB-TaylorFormer V2: Improved Multi-branch Linear Transformer Expanded by Taylor Formula for Image Restoration
by: Jin, Zhi, et al.
Published: (2025)
by: Jin, Zhi, et al.
Published: (2025)
MRD: Multi-resolution Retrieval-Detection Fusion for High-Resolution Image Understanding
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal
by: Wang, Tao, et al.
Published: (2024)
by: Wang, Tao, et al.
Published: (2024)
OMG-Seg: Is One Model Good Enough For All Segmentation?
by: Li, Xiangtai, et al.
Published: (2024)
by: Li, Xiangtai, et al.
Published: (2024)
OMG: Opacity Matters in Material Modeling with Gaussian Splatting
by: Yong, Silong, et al.
Published: (2025)
by: Yong, Silong, et al.
Published: (2025)
MotionMERGE: A Multi-granular Framework for Human Motion Editing, Reasoning, Generation, and Explanation
by: Wu, Bizhu, et al.
Published: (2026)
by: Wu, Bizhu, et al.
Published: (2026)
LLDif: Diffusion Models for Low-light Emotion Recognition
by: Wang, Zhifeng, et al.
Published: (2024)
by: Wang, Zhifeng, et al.
Published: (2024)
OMG: Towards Open-vocabulary Motion Generation via Mixture of Controllers
by: Liang, Han, et al.
Published: (2023)
by: Liang, Han, et al.
Published: (2023)
LRDif: Diffusion Models for Under-Display Camera Emotion Recognition
by: Wang, Zhifeng, et al.
Published: (2024)
by: Wang, Zhifeng, et al.
Published: (2024)
SPDiffusion: Semantic Protection Diffusion Models for Multi-concept Text-to-image Generation
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
OMG-LLaVA: Bridging Image-level, Object-level, Pixel-level Reasoning and Understanding
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
AnyTalker: Scaling Multi-Person Talking Video Generation with Interactivity Refinement
by: Zhong, Zhizhou, et al.
Published: (2025)
by: Zhong, Zhizhou, et al.
Published: (2025)
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
by: Liu, Gongye, et al.
Published: (2026)
by: Liu, Gongye, et al.
Published: (2026)
MG-MotionLLM: A Unified Framework for Motion Comprehension and Generation across Multiple Granularities
by: Wu, Bizhu, et al.
Published: (2025)
by: Wu, Bizhu, et al.
Published: (2025)
OMG-Avatar: One-shot Multi-LOD Gaussian Head Avatar
by: Ren, Jianqiang, et al.
Published: (2026)
by: Ren, Jianqiang, et al.
Published: (2026)
FineMotion: A Dataset and Benchmark with both Spatial and Temporal Annotation for Fine-grained Motion Generation and Editing
by: Wu, Bizhu, et al.
Published: (2025)
by: Wu, Bizhu, et al.
Published: (2025)
GridFormer: Residual Dense Transformer with Grid Structure for Image Restoration in Adverse Weather Conditions
by: Wang, Tao, et al.
Published: (2023)
by: Wang, Tao, et al.
Published: (2023)
InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing
by: Yang, Shaoshu, et al.
Published: (2025)
by: Yang, Shaoshu, et al.
Published: (2025)
Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
MC$^2$: Multi-concept Guidance for Customized Multi-concept Generation
by: Jiang, Jiaxiu, et al.
Published: (2024)
by: Jiang, Jiaxiu, et al.
Published: (2024)
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
GO-MLVTON: Garment Occlusion-Aware Multi-Layer Virtual Try-On with Diffusion Models
by: Yu, Yang, et al.
Published: (2026)
by: Yu, Yang, et al.
Published: (2026)
A Survey on Personalized Content Synthesis with Diffusion Models
by: Zhang, Xulu, et al.
Published: (2024)
by: Zhang, Xulu, et al.
Published: (2024)
Stable Diffusion-Based Approach for Human De-Occlusion
by: Noh, Seung Young, et al.
Published: (2025)
by: Noh, Seung Young, et al.
Published: (2025)
LlamaSeg: Image Segmentation via Autoregressive Mask Generation
by: Deng, Jiru, et al.
Published: (2025)
by: Deng, Jiru, et al.
Published: (2025)
Foundation Cures Personalization: Improving Personalized Models' Prompt Consistency via Hidden Foundation Knowledge
by: Cai, Yiyang, et al.
Published: (2024)
by: Cai, Yiyang, et al.
Published: (2024)
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
by: Wang, Zihao, et al.
Published: (2026)
by: Wang, Zihao, et al.
Published: (2026)
Fourier-RWKV: A Multi-State Perception Network for Efficient Image Dehazing
by: Zheng, Lirong, et al.
Published: (2025)
by: Zheng, Lirong, et al.
Published: (2025)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023)
by: Luo, Run, et al.
Published: (2023)
Homography Guided Temporal Fusion for Road Line and Marking Segmentation
by: Wang, Shan, et al.
Published: (2024)
by: Wang, Shan, et al.
Published: (2024)
Generative Active Learning for Image Synthesis Personalization
by: Zhang, Xulu, et al.
Published: (2024)
by: Zhang, Xulu, et al.
Published: (2024)
All-in-one Weather-degraded Image Restoration via Adaptive Degradation-aware Self-prompting Model
by: Wen, Yuanbo, et al.
Published: (2024)
by: Wen, Yuanbo, et al.
Published: (2024)
Multi-Granularity Hand Action Detection
by: Zhe, Ting, et al.
Published: (2023)
by: Zhe, Ting, et al.
Published: (2023)
FineXtrol: Controllable Motion Generation via Fine-Grained Text
by: Shen, Keming, et al.
Published: (2025)
by: Shen, Keming, et al.
Published: (2025)
Similar Items
-
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
by: Kong, Zhe, et al.
Published: (2025) -
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
by: Kong, Zhe, et al.
Published: (2025) -
MaterialMVP: Illumination-Invariant Material Generation via Multi-view PBR Diffusion
by: He, Zebin, et al.
Published: (2025) -
Towards Real-World Blind Face Restoration with Generative Diffusion Prior
by: Chen, Xiaoxu, et al.
Published: (2023) -
Dual Teacher Knowledge Distillation with Domain Alignment for Face Anti-spoofing
by: Kong, Zhe, et al.
Published: (2024)