High-fidelity Person-centric Subject-to-Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yibin, Zhang, Weizhong, Zheng, Jianwei, Jin, Cheng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
by: Wang, Yibin, et al.
Published: (2024)
by: Wang, Yibin, et al.
Published: (2024)
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
by: Wang, Yibin, et al.
Published: (2024)
by: Wang, Yibin, et al.
Published: (2024)
DreamText: High Fidelity Scene Text Synthesis
by: Wang, Yibin, et al.
Published: (2024)
by: Wang, Yibin, et al.
Published: (2024)
Enhancing Object Coherence in Layout-to-Image Synthesis
by: Wang, Yibin, et al.
Published: (2023)
by: Wang, Yibin, et al.
Published: (2023)
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
by: Li, Xuanchen, et al.
Published: (2025)
by: Li, Xuanchen, et al.
Published: (2025)
A Framework For Image Synthesis Using Supervised Contrastive Learning
by: Liu, Yibin, et al.
Published: (2024)
by: Liu, Yibin, et al.
Published: (2024)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
by: Jin, Qiaoqiao, et al.
Published: (2025)
by: Jin, Qiaoqiao, et al.
Published: (2025)
DynASyn: Multi-Subject Personalization Enabling Dynamic Action Synthesis
by: Choi, Yongjin, et al.
Published: (2025)
by: Choi, Yongjin, et al.
Published: (2025)
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
by: Jang, Sangwon, et al.
Published: (2024)
by: Jang, Sangwon, et al.
Published: (2024)
CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images
by: Liu, Jian, et al.
Published: (2024)
by: Liu, Jian, et al.
Published: (2024)
xGen-VideoSyn-1: High-fidelity Text-to-Video Synthesis with Compressed Representations
by: Qin, Can, et al.
Published: (2024)
by: Qin, Can, et al.
Published: (2024)
Wonder3D++: Cross-domain Diffusion for High-fidelity 3D Generation from a Single Image
by: Yang, Yuxiao, et al.
Published: (2025)
by: Yang, Yuxiao, et al.
Published: (2025)
CustomTex: High-fidelity Indoor Scene Texturing via Multi-Reference Customization
by: Chen, Weilin, et al.
Published: (2026)
by: Chen, Weilin, et al.
Published: (2026)
Object-centric Binding in Contrastive Language-Image Pretraining
by: Assouel, Rim, et al.
Published: (2025)
by: Assouel, Rim, et al.
Published: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
by: Wang, Kuan-Chieh, et al.
Published: (2024)
by: Wang, Kuan-Chieh, et al.
Published: (2024)
Zero-shot High-fidelity and Pose-controllable Character Animation
by: Zhu, Bingwen, et al.
Published: (2024)
by: Zhu, Bingwen, et al.
Published: (2024)
Landmark-guided Diffusion Model for High-fidelity and Temporally Coherent Talking Head Generation
by: Tan, Jintao, et al.
Published: (2024)
by: Tan, Jintao, et al.
Published: (2024)
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
by: Zhong, Weizhi, et al.
Published: (2024)
by: Zhong, Weizhi, et al.
Published: (2024)
Few-Shot Medical Image Segmentation with High-Fidelity Prototypes
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
High-Quality 3D Creation from A Single Image Using Subject-Specific Knowledge Prior
by: Huang, Nan, et al.
Published: (2023)
by: Huang, Nan, et al.
Published: (2023)
ImageRAG: Enhancing Ultra High Resolution Remote Sensing Imagery Analysis with ImageRAG
by: Zhang, Zilun, et al.
Published: (2024)
by: Zhang, Zilun, et al.
Published: (2024)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
by: Lee, Donghwna, et al.
Published: (2024)
by: Lee, Donghwna, et al.
Published: (2024)
Towards Accurate and Interpretable Neuroblastoma Diagnosis via Contrastive Multi-scale Pathological Image Analysis
by: Zhu, Zhu, et al.
Published: (2025)
by: Zhu, Zhu, et al.
Published: (2025)
Concept Conductor: Orchestrating Multiple Personalized Concepts in Text-to-Image Synthesis
by: Yao, Zebin, et al.
Published: (2024)
by: Yao, Zebin, et al.
Published: (2024)
Evidential Graph Contrastive Alignment for Source-Free Blending-Target Domain Adaptation
by: Zheng, Juepeng, et al.
Published: (2024)
by: Zheng, Juepeng, et al.
Published: (2024)
Real-Time Person Image Synthesis Using a Flow Matching Model
by: Jeong, Jiwoo, et al.
Published: (2025)
by: Jeong, Jiwoo, et al.
Published: (2025)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
by: Xu, Yijia, et al.
Published: (2026)
by: Xu, Yijia, et al.
Published: (2026)
MM-Diff: High-Fidelity Image Personalization via Multi-Modal Condition Integration
by: Wei, Zhichao, et al.
Published: (2024)
by: Wei, Zhichao, et al.
Published: (2024)
C3S3: Complementary Competition and Contrastive Selection for Semi-Supervised Medical Image Segmentation
by: He, Jiaying, et al.
Published: (2025)
by: He, Jiaying, et al.
Published: (2025)
Flux Already Knows -- Activating Subject-Driven Image Generation without Training
by: Kang, Hao, et al.
Published: (2025)
by: Kang, Hao, et al.
Published: (2025)
When Identities Collapse: A Stress-Test Benchmark for Multi-Subject Personalization
by: Chen, Zhihan, et al.
Published: (2026)
by: Chen, Zhihan, et al.
Published: (2026)
Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data
by: Zhu, Xun, et al.
Published: (2025)
by: Zhu, Xun, et al.
Published: (2025)
GALA: A GlobAl-LocAl Approach for Multi-Source Active Domain Adaptation
by: Zheng, Juepeng, et al.
Published: (2025)
by: Zheng, Juepeng, et al.
Published: (2025)
Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization
by: Dunlop, Connor, et al.
Published: (2025)
by: Dunlop, Connor, et al.
Published: (2025)
Audio-centric Video Understanding Benchmark without Text Shortcut
by: Yang, Yudong, et al.
Published: (2025)
by: Yang, Yudong, et al.
Published: (2025)
Oasis: One Image is All You Need for Multimodal Instruction Data Synthesis
by: Zhang, Letian, et al.
Published: (2025)
by: Zhang, Letian, et al.
Published: (2025)
DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation
by: Hu, Zhenyu, et al.
Published: (2026)
by: Hu, Zhenyu, et al.
Published: (2026)
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing
by: Wang, Dianyi, et al.
Published: (2026)
by: Wang, Dianyi, et al.
Published: (2026)
Vehicle-centric Perception via Multimodal Structured Pre-training
by: Wu, Wentao, et al.
Published: (2025)
by: Wu, Wentao, et al.
Published: (2025)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
by: Zhang, Shizhou, et al.
Published: (2021)
by: Zhang, Shizhou, et al.
Published: (2021)
Similar Items
-
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
by: Wang, Yibin, et al.
Published: (2024) -
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
by: Wang, Yibin, et al.
Published: (2024) -
DreamText: High Fidelity Scene Text Synthesis
by: Wang, Yibin, et al.
Published: (2024) -
Enhancing Object Coherence in Layout-to-Image Synthesis
by: Wang, Yibin, et al.
Published: (2023) -
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
by: Li, Xuanchen, et al.
Published: (2025)