ZePo: Zero-Shot Portrait Stylization with Faster Sampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jin, Huang, Huaibo, Cao, Jie, He, Ran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vision Transformer with Super Token Sampling
von: Huang, Huaibo, et al.
Veröffentlicht: (2022)
von: Huang, Huaibo, et al.
Veröffentlicht: (2022)
Marmot: Object-Level Self-Correction via Multi-Agent Reasoning
von: Sun, Jiayang, et al.
Veröffentlicht: (2025)
von: Sun, Jiayang, et al.
Veröffentlicht: (2025)
InstaStyle: Inversion Noise of a Stylized Image is Secretly a Style Adviser
von: Cui, Xing, et al.
Veröffentlicht: (2023)
von: Cui, Xing, et al.
Veröffentlicht: (2023)
Breaking the Low-Rank Dilemma of Linear Attention
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
ZePT: Zero-Shot Pan-Tumor Segmentation via Query-Disentangling and Self-Prompting
von: Jiang, Yankai, et al.
Veröffentlicht: (2023)
von: Jiang, Yankai, et al.
Veröffentlicht: (2023)
ZeST: Zero-Shot Material Transfer from a Single Image
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2024)
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2024)
MagicStyle: Portrait Stylization Based on Reference Image
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
von: Deng, Zhaoli, et al.
Veröffentlicht: (2024)
DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis
von: Gu, Yuming, et al.
Veröffentlicht: (2023)
von: Gu, Yuming, et al.
Veröffentlicht: (2023)
Realtime Data-Efficient Portrait Stylization Based On Geometric Alignment
von: Wang, Xinrui, et al.
Veröffentlicht: (2022)
von: Wang, Xinrui, et al.
Veröffentlicht: (2022)
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
Rectifying Magnitude Neglect in Linear Attention
von: Fan, Qihang, et al.
Veröffentlicht: (2025)
von: Fan, Qihang, et al.
Veröffentlicht: (2025)
Random Wins All: Rethinking Grouping Strategies for Vision Tokens
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
NOFT: Test-Time Noise Finetune via Information Bottleneck for Highly Correlated Asset Creation
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Lightweight Vision Transformer with Bidirectional Interaction
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2025)
ZDySS -- Zero-Shot Dynamic Scene Stylization using Gaussian Splatting
von: Saroha, Abhishek, et al.
Veröffentlicht: (2025)
von: Saroha, Abhishek, et al.
Veröffentlicht: (2025)
RMT: Retentive Networks Meet Vision Transformers
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
Think 360°: Evaluating the Width-centric Reasoning Capability of MLLMs Beyond Depth
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
Advancing Vision Transformer with Enhanced Spatial Priors
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
ZeST: an LLM-based Zero-Shot Traversability Navigation for Unknown Environments
von: Gummadi, Shreya, et al.
Veröffentlicht: (2025)
von: Gummadi, Shreya, et al.
Veröffentlicht: (2025)
Hybrid Fusion: One-Minute Efficient Training for Zero-Shot Cross-Domain Image Fusion
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
von: Zhang, Ran, et al.
Veröffentlicht: (2026)
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
InfoBFR: Real-World Blind Face Restoration via Information Bottleneck
von: Gao, Nan, et al.
Veröffentlicht: (2025)
von: Gao, Nan, et al.
Veröffentlicht: (2025)
Parallel Augmentation and Dual Enhancement for Occluded Person Re-identification
von: Wang, Zi, et al.
Veröffentlicht: (2022)
von: Wang, Zi, et al.
Veröffentlicht: (2022)
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
StylizedGS: Controllable Stylization for 3D Gaussian Splatting
von: Zhang, Dingxi, et al.
Veröffentlicht: (2024)
von: Zhang, Dingxi, et al.
Veröffentlicht: (2024)
ZeBROD: Zero-Retraining Based Recognition and Object Detection Framework
von: Hidayatullah, Priyanto, et al.
Veröffentlicht: (2025)
von: Hidayatullah, Priyanto, et al.
Veröffentlicht: (2025)
Unlocking the Potential of Difficulty Prior in RL-based Multimodal Reasoning
von: Chen, Mingrui, et al.
Veröffentlicht: (2025)
von: Chen, Mingrui, et al.
Veröffentlicht: (2025)
MVPBench: A Multi-Video Perception Evaluation Benchmark for Multi-Modal Video Understanding
von: Bai, Purui, et al.
Veröffentlicht: (2026)
von: Bai, Purui, et al.
Veröffentlicht: (2026)
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
von: Qi, Jinwei, et al.
Veröffentlicht: (2025)
CSCNET: Class-Specified Cascaded Network for Compositional Zero-Shot Learning
von: Zhang, Yanyi, et al.
Veröffentlicht: (2024)
von: Zhang, Yanyi, et al.
Veröffentlicht: (2024)
A Framework for Portrait Stylization with Skin-Tone Awareness and Nudity Identification
von: Kim, Seungkwon, et al.
Veröffentlicht: (2024)
von: Kim, Seungkwon, et al.
Veröffentlicht: (2024)
EchoShot: Multi-Shot Portrait Video Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
One-Shot Structure-Aware Stylized Image Synthesis
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
Visual Anchors Are Strong Information Aggregators For Multimodal Large Language Model
von: Liu, Haogeng, et al.
Veröffentlicht: (2024)
von: Liu, Haogeng, et al.
Veröffentlicht: (2024)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
von: He, Xuanhua, et al.
Veröffentlicht: (2024)
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Vision Transformer with Super Token Sampling
von: Huang, Huaibo, et al.
Veröffentlicht: (2022) -
Marmot: Object-Level Self-Correction via Multi-Agent Reasoning
von: Sun, Jiayang, et al.
Veröffentlicht: (2025) -
InstaStyle: Inversion Noise of a Stylized Image is Secretly a Style Adviser
von: Cui, Xing, et al.
Veröffentlicht: (2023) -
Breaking the Low-Rank Dilemma of Linear Attention
von: Fan, Qihang, et al.
Veröffentlicht: (2024) -
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2024)