High-fidelity Person-centric Subject-to-Image Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yibin, Zhang, Weizhong, Zheng, Jianwei, Jin, Cheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
DreamText: High Fidelity Scene Text Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
Enhancing Object Coherence in Layout-to-Image Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2023)
von: Wang, Yibin, et al.
Veröffentlicht: (2023)
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
von: Li, Xuanchen, et al.
Veröffentlicht: (2025)
von: Li, Xuanchen, et al.
Veröffentlicht: (2025)
A Framework For Image Synthesis Using Supervised Contrastive Learning
von: Liu, Yibin, et al.
Veröffentlicht: (2024)
von: Liu, Yibin, et al.
Veröffentlicht: (2024)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
von: Jin, Qiaoqiao, et al.
Veröffentlicht: (2025)
DynASyn: Multi-Subject Personalization Enabling Dynamic Action Synthesis
von: Choi, Yongjin, et al.
Veröffentlicht: (2025)
von: Choi, Yongjin, et al.
Veröffentlicht: (2025)
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images
von: Liu, Jian, et al.
Veröffentlicht: (2024)
von: Liu, Jian, et al.
Veröffentlicht: (2024)
xGen-VideoSyn-1: High-fidelity Text-to-Video Synthesis with Compressed Representations
von: Qin, Can, et al.
Veröffentlicht: (2024)
von: Qin, Can, et al.
Veröffentlicht: (2024)
Wonder3D++: Cross-domain Diffusion for High-fidelity 3D Generation from a Single Image
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
CustomTex: High-fidelity Indoor Scene Texturing via Multi-Reference Customization
von: Chen, Weilin, et al.
Veröffentlicht: (2026)
von: Chen, Weilin, et al.
Veröffentlicht: (2026)
Object-centric Binding in Contrastive Language-Image Pretraining
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
Zero-shot High-fidelity and Pose-controllable Character Animation
von: Zhu, Bingwen, et al.
Veröffentlicht: (2024)
von: Zhu, Bingwen, et al.
Veröffentlicht: (2024)
Landmark-guided Diffusion Model for High-fidelity and Temporally Coherent Talking Head Generation
von: Tan, Jintao, et al.
Veröffentlicht: (2024)
von: Tan, Jintao, et al.
Veröffentlicht: (2024)
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
von: Zhong, Weizhi, et al.
Veröffentlicht: (2024)
Few-Shot Medical Image Segmentation with High-Fidelity Prototypes
von: Tang, Song, et al.
Veröffentlicht: (2024)
von: Tang, Song, et al.
Veröffentlicht: (2024)
High-Quality 3D Creation from A Single Image Using Subject-Specific Knowledge Prior
von: Huang, Nan, et al.
Veröffentlicht: (2023)
von: Huang, Nan, et al.
Veröffentlicht: (2023)
ImageRAG: Enhancing Ultra High Resolution Remote Sensing Imagery Analysis with ImageRAG
von: Zhang, Zilun, et al.
Veröffentlicht: (2024)
von: Zhang, Zilun, et al.
Veröffentlicht: (2024)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
von: Lee, Donghwna, et al.
Veröffentlicht: (2024)
von: Lee, Donghwna, et al.
Veröffentlicht: (2024)
Towards Accurate and Interpretable Neuroblastoma Diagnosis via Contrastive Multi-scale Pathological Image Analysis
von: Zhu, Zhu, et al.
Veröffentlicht: (2025)
von: Zhu, Zhu, et al.
Veröffentlicht: (2025)
Concept Conductor: Orchestrating Multiple Personalized Concepts in Text-to-Image Synthesis
von: Yao, Zebin, et al.
Veröffentlicht: (2024)
von: Yao, Zebin, et al.
Veröffentlicht: (2024)
Evidential Graph Contrastive Alignment for Source-Free Blending-Target Domain Adaptation
von: Zheng, Juepeng, et al.
Veröffentlicht: (2024)
von: Zheng, Juepeng, et al.
Veröffentlicht: (2024)
Real-Time Person Image Synthesis Using a Flow Matching Model
von: Jeong, Jiwoo, et al.
Veröffentlicht: (2025)
von: Jeong, Jiwoo, et al.
Veröffentlicht: (2025)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
MM-Diff: High-Fidelity Image Personalization via Multi-Modal Condition Integration
von: Wei, Zhichao, et al.
Veröffentlicht: (2024)
von: Wei, Zhichao, et al.
Veröffentlicht: (2024)
C3S3: Complementary Competition and Contrastive Selection for Semi-Supervised Medical Image Segmentation
von: He, Jiaying, et al.
Veröffentlicht: (2025)
von: He, Jiaying, et al.
Veröffentlicht: (2025)
Flux Already Knows -- Activating Subject-Driven Image Generation without Training
von: Kang, Hao, et al.
Veröffentlicht: (2025)
von: Kang, Hao, et al.
Veröffentlicht: (2025)
When Identities Collapse: A Stress-Test Benchmark for Multi-Subject Personalization
von: Chen, Zhihan, et al.
Veröffentlicht: (2026)
von: Chen, Zhihan, et al.
Veröffentlicht: (2026)
Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data
von: Zhu, Xun, et al.
Veröffentlicht: (2025)
von: Zhu, Xun, et al.
Veröffentlicht: (2025)
GALA: A GlobAl-LocAl Approach for Multi-Source Active Domain Adaptation
von: Zheng, Juepeng, et al.
Veröffentlicht: (2025)
von: Zheng, Juepeng, et al.
Veröffentlicht: (2025)
Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization
von: Dunlop, Connor, et al.
Veröffentlicht: (2025)
von: Dunlop, Connor, et al.
Veröffentlicht: (2025)
Audio-centric Video Understanding Benchmark without Text Shortcut
von: Yang, Yudong, et al.
Veröffentlicht: (2025)
von: Yang, Yudong, et al.
Veröffentlicht: (2025)
Oasis: One Image is All You Need for Multimodal Instruction Data Synthesis
von: Zhang, Letian, et al.
Veröffentlicht: (2025)
von: Zhang, Letian, et al.
Veröffentlicht: (2025)
DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation
von: Hu, Zhenyu, et al.
Veröffentlicht: (2026)
von: Hu, Zhenyu, et al.
Veröffentlicht: (2026)
DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing
von: Wang, Dianyi, et al.
Veröffentlicht: (2026)
von: Wang, Dianyi, et al.
Veröffentlicht: (2026)
Vehicle-centric Perception via Multimodal Structured Pre-training
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
von: Wu, Wentao, et al.
Veröffentlicht: (2025)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
von: Zhang, Shizhou, et al.
Veröffentlicht: (2021)
von: Zhang, Shizhou, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
von: Wang, Yibin, et al.
Veröffentlicht: (2024) -
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024) -
DreamText: High Fidelity Scene Text Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2024) -
Enhancing Object Coherence in Layout-to-Image Synthesis
von: Wang, Yibin, et al.
Veröffentlicht: (2023) -
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
von: Li, Xuanchen, et al.
Veröffentlicht: (2025)