DiffFAE: Advancing High-fidelity One-shot Facial Appearance Editing with Space-sensitive Customization and Semantic Preservation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qilin, Zhang, Jiangning, Xu, Chengming, Cao, Weijian, Tai, Ying, Han, Yue, Ge, Yanhao, Gu, Hong, Wang, Chengjie, Fu, Yanwei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
by: Wang, Qilin, et al.
Published: (2024)
by: Wang, Qilin, et al.
Published: (2024)
ArtWeaver: Advanced Dynamic Style Integration via Diffusion Model
by: Xu, Chengming, et al.
Published: (2024)
by: Xu, Chengming, et al.
Published: (2024)
FitDiT: Advancing the Authentic Garment Details for High-fidelity Virtual Try-on
by: Jiang, Boyuan, et al.
Published: (2024)
by: Jiang, Boyuan, et al.
Published: (2024)
A Generalist FaceX via Learning Unified Facial Representation
by: Han, Yue, et al.
Published: (2023)
by: Han, Yue, et al.
Published: (2023)
What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing
by: Lin, Hangyu, et al.
Published: (2026)
by: Lin, Hangyu, et al.
Published: (2026)
FFP-300K: Scaling First-Frame Propagation for Generalizable Video Editing
by: Huang, Xijie, et al.
Published: (2026)
by: Huang, Xijie, et al.
Published: (2026)
StrandDesigner: Towards Practical Strand Generation with Sketch Guidance
by: Zhang, Na, et al.
Published: (2025)
by: Zhang, Na, et al.
Published: (2025)
CustAny: Customizing Anything from A Single Example
by: Kong, Lingjie, et al.
Published: (2024)
by: Kong, Lingjie, et al.
Published: (2024)
IP-FaceDiff: Identity-Preserving Facial Video Editing with Diffusion
by: Anand, Tharun, et al.
Published: (2025)
by: Anand, Tharun, et al.
Published: (2025)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
by: Fu, Chencan, et al.
Published: (2024)
by: Fu, Chencan, et al.
Published: (2024)
RealTalk: Real-time and Realistic Audio-driven Face Generation with 3D Facial Prior-guided Identity Alignment Network
by: Ji, Xiaozhong, et al.
Published: (2024)
by: Ji, Xiaozhong, et al.
Published: (2024)
Semantic Frame Interpolation
by: Hong, Yijia, et al.
Published: (2025)
by: Hong, Yijia, et al.
Published: (2025)
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
by: Wang, Yuji, et al.
Published: (2025)
by: Wang, Yuji, et al.
Published: (2025)
Describe, Don't Dictate: Semantic Image Editing with Natural Language Intent
by: Ci, En, et al.
Published: (2025)
by: Ci, En, et al.
Published: (2025)
SwiftVideo: A Unified Framework for Few-Step Video Generation through Trajectory-Distribution Alignment
by: Sun, Yanxiao, et al.
Published: (2025)
by: Sun, Yanxiao, et al.
Published: (2025)
RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations
by: Ge, Yanhao, et al.
Published: (2026)
by: Ge, Yanhao, et al.
Published: (2026)
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
by: Luo, Weijian
Published: (2024)
by: Luo, Weijian
Published: (2024)
Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
by: Wang, Yuji, et al.
Published: (2025)
by: Wang, Yuji, et al.
Published: (2025)
DiffMagicFace: Identity Consistent Facial Editing of Real Videos
by: Yin, Huanghao, et al.
Published: (2026)
by: Yin, Huanghao, et al.
Published: (2026)
EmojiDiff: Advanced Facial Expression Control with High Identity Preservation in Portrait Generation
by: Jiang, Liangwei, et al.
Published: (2024)
by: Jiang, Liangwei, et al.
Published: (2024)
DisControlFace: Adding Disentangled Control to Diffusion Autoencoder for One-shot Explicit Facial Image Editing
by: Jia, Haozhe, et al.
Published: (2023)
by: Jia, Haozhe, et al.
Published: (2023)
Zero-shot High-fidelity and Pose-controllable Character Animation
by: Zhu, Bingwen, et al.
Published: (2024)
by: Zhu, Bingwen, et al.
Published: (2024)
ZeroDiff++: Substantial Unseen Visual-semantic Correlation in Zero-shot Learning
by: Ye, Zihan, et al.
Published: (2026)
by: Ye, Zihan, et al.
Published: (2026)
CtlGAN: Few-shot Artistic Portraits Generation with Contrastive Transfer Learning
by: Wang, Yue, et al.
Published: (2022)
by: Wang, Yue, et al.
Published: (2022)
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection
by: Chen, Xuhai, et al.
Published: (2023)
by: Chen, Xuhai, et al.
Published: (2023)
Face Adapter for Pre-Trained Diffusion Models with Fine-Grained ID and Attribute Control
by: Han, Yue, et al.
Published: (2024)
by: Han, Yue, et al.
Published: (2024)
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
by: Huang, Jiale, et al.
Published: (2024)
by: Huang, Jiale, et al.
Published: (2024)
Advancing Facial Stylization through Semantic Preservation Constraint and Pseudo-Paired Supervision
by: Lu, Zhanyi, et al.
Published: (2025)
by: Lu, Zhanyi, et al.
Published: (2025)
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection
by: Zhang, Jiangning, et al.
Published: (2023)
by: Zhang, Jiangning, et al.
Published: (2023)
One-shot Embroidery Customization via Contrastive LoRA Modulation
by: Ma, Jun, et al.
Published: (2025)
by: Ma, Jun, et al.
Published: (2025)
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
by: Yin, Zijin, et al.
Published: (2026)
by: Yin, Zijin, et al.
Published: (2026)
HiFiVFS: High Fidelity Video Face Swapping
by: Chen, Xu, et al.
Published: (2024)
by: Chen, Xu, et al.
Published: (2024)
Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation
by: Zhang, Jiangning, et al.
Published: (2025)
by: Zhang, Jiangning, et al.
Published: (2025)
Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning
by: He, Qingdong, et al.
Published: (2025)
by: He, Qingdong, et al.
Published: (2025)
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
by: Wu, Junyi, et al.
Published: (2026)
by: Wu, Junyi, et al.
Published: (2026)
AttDiff-GAN: A Hybrid Diffusion-GAN Framework for Facial Attribute Editing
by: Huang, Wenmin, et al.
Published: (2026)
by: Huang, Wenmin, et al.
Published: (2026)
Monocular Facial Appearance Capture in the Wild
by: Xu, Yingyan, et al.
Published: (2024)
by: Xu, Yingyan, et al.
Published: (2024)
ReactDiff: Latent Diffusion for Facial Reaction Generation
by: Li, Jiaming, et al.
Published: (2025)
by: Li, Jiaming, et al.
Published: (2025)
VTBench: Comprehensive Benchmark Suite Towards Real-World Virtual Try-on Models
by: Xiaobin, Hu, et al.
Published: (2025)
by: Xiaobin, Hu, et al.
Published: (2025)
Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing
by: Xu, Pengcheng, et al.
Published: (2024)
by: Xu, Pengcheng, et al.
Published: (2024)
Similar Items
-
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
by: Wang, Qilin, et al.
Published: (2024) -
ArtWeaver: Advanced Dynamic Style Integration via Diffusion Model
by: Xu, Chengming, et al.
Published: (2024) -
FitDiT: Advancing the Authentic Garment Details for High-fidelity Virtual Try-on
by: Jiang, Boyuan, et al.
Published: (2024) -
A Generalist FaceX via Learning Unified Facial Representation
by: Han, Yue, et al.
Published: (2023) -
What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing
by: Lin, Hangyu, et al.
Published: (2026)