A training-free framework for high-fidelity appearance transfer via diffusion transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Shengrong, Wang, Ye, Wu, Song, Ma, Rui, Wang, Qian, Wang, Lanjun, Yi, Zili |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OSTAF: A One-Shot Tuning Method for Improved Attribute-Focused T2I Personalization
by: Wang, Ye, et al.
Published: (2024)
by: Wang, Ye, et al.
Published: (2024)
Anywhere: A Multi-Agent Framework for User-Guided, Reliable, and Diverse Foreground-Conditioned Image Generation
by: Xie, Tianyidan, et al.
Published: (2024)
by: Xie, Tianyidan, et al.
Published: (2024)
DiTraj: training-free trajectory control for video diffusion transformer
by: Lei, Cheng, et al.
Published: (2025)
by: Lei, Cheng, et al.
Published: (2025)
Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance
by: Wu, Song, et al.
Published: (2026)
by: Wu, Song, et al.
Published: (2026)
Revealing Vulnerabilities in Stable Diffusion via Targeted Attacks
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
OmniStyle: Filtering High Quality Style Transfer Data at Scale
by: Wang, Ye, et al.
Published: (2025)
by: Wang, Ye, et al.
Published: (2025)
FreeControl: Efficient, Training-Free Structural Control via One-Step Attention Extraction
by: Lin, Jiang, et al.
Published: (2025)
by: Lin, Jiang, et al.
Published: (2025)
One-Shot Learning for Pose-Guided Person Image Synthesis in the Wild
by: Fan, Dongqi, et al.
Published: (2024)
by: Fan, Dongqi, et al.
Published: (2024)
GHOST 2.0: generative high-fidelity one shot transfer of heads
by: Groshev, Alexander, et al.
Published: (2025)
by: Groshev, Alexander, et al.
Published: (2025)
MVOC: a training-free multiple video object composition method with diffusion models
by: Wang, Wei, et al.
Published: (2024)
by: Wang, Wei, et al.
Published: (2024)
MarkPlugger: Generalizable Watermark Framework for Latent Diffusion Models without Retraining
by: Zhang, Guokai, et al.
Published: (2024)
by: Zhang, Guokai, et al.
Published: (2024)
OmniStyle2: Learning to Stylize by Learning to Destylize
by: Wang, Ye, et al.
Published: (2025)
by: Wang, Ye, et al.
Published: (2025)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
by: Chen, Xinyu, et al.
Published: (2026)
by: Chen, Xinyu, et al.
Published: (2026)
Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On
by: Shao, Dingbao, et al.
Published: (2026)
by: Shao, Dingbao, et al.
Published: (2026)
Domain Adaptation from Generated Multi-Weather Images for Unsupervised Maritime Object Classification
by: Song, Dan, et al.
Published: (2025)
by: Song, Dan, et al.
Published: (2025)
Training-free image inversion for one-step diffusion models
by: Wu, Tao, et al.
Published: (2026)
by: Wu, Tao, et al.
Published: (2026)
Self-supervised transformer-based pre-training method with General Plant Infection dataset
by: Wang, Zhengle, et al.
Published: (2024)
by: Wang, Zhengle, et al.
Published: (2024)
X-Part: high fidelity and structure coherent shape decomposition
by: Yan, Xinhao, et al.
Published: (2025)
by: Yan, Xinhao, et al.
Published: (2025)
HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images
by: Choi, Sungik, et al.
Published: (2024)
by: Choi, Sungik, et al.
Published: (2024)
Inversion-Free Video Style Transfer with Trajectory Reset Attention Control and Content-Style Bridging
by: Lin, Jiang, et al.
Published: (2025)
by: Lin, Jiang, et al.
Published: (2025)
Metaphor-based Jailbreak Attacks on Text-to-Image Models
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
ETC: training-free diffusion models acceleration with Error-aware Trend Consistency
by: Xie, Jiajian, et al.
Published: (2025)
by: Xie, Jiajian, et al.
Published: (2025)
SemanticHuman-HD: High-Resolution Semantic Disentangled 3D Human Generation
by: Zheng, Peng, et al.
Published: (2024)
by: Zheng, Peng, et al.
Published: (2024)
VQAttack: Transferable Adversarial Attacks on Visual Question Answering via Pre-trained Models
by: Yin, Ziyi, et al.
Published: (2024)
by: Yin, Ziyi, et al.
Published: (2024)
From Zero to Detail: Deconstructing Ultra-High-Definition Image Restoration from Progressive Spectral Perspective
by: Zhao, Chen, et al.
Published: (2025)
by: Zhao, Chen, et al.
Published: (2025)
FreeLoRA: Enabling Training-Free LoRA Fusion for Autoregressive Multi-Subject Personalization
by: Zheng, Peng, et al.
Published: (2025)
by: Zheng, Peng, et al.
Published: (2025)
Towards Deconfounded Image-Text Matching with Causal Inference
by: Li, Wenhui, et al.
Published: (2024)
by: Li, Wenhui, et al.
Published: (2024)
DM-FNet: Unified multimodal medical image fusion via diffusion process-trained encoder-decoder
by: He, Dan, et al.
Published: (2025)
by: He, Dan, et al.
Published: (2025)
VFM-SDM: A vision foundation model-based framework for training-free, marker-free, and calibration-free structural displacement measurement
by: Xian, Qingyu, et al.
Published: (2026)
by: Xian, Qingyu, et al.
Published: (2026)
R3D-AD: Reconstruction via Diffusion for 3D Anomaly Detection
by: Zhou, Zheyuan, et al.
Published: (2024)
by: Zhou, Zheyuan, et al.
Published: (2024)
Scalable and Generalizable Correspondence Pruning via Geometry-Consistent Pre-training
by: Liao, Tangfei, et al.
Published: (2024)
by: Liao, Tangfei, et al.
Published: (2024)
DenseTrack: Drone-based Crowd Tracking via Density-aware Motion-appearance Synergy
by: Lei, Yi, et al.
Published: (2024)
by: Lei, Yi, et al.
Published: (2024)
PairHuman: A High-Fidelity Photographic Dataset for Customized Dual-Person Generation
by: Pan, Ting, et al.
Published: (2025)
by: Pan, Ting, et al.
Published: (2025)
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
by: Wang, Haiguang, et al.
Published: (2025)
by: Wang, Haiguang, et al.
Published: (2025)
RaSim: A Range-aware High-fidelity RGB-D Data Simulation Pipeline for Real-world Applications
by: Liu, Xingyu, et al.
Published: (2024)
by: Liu, Xingyu, et al.
Published: (2024)
OSInsert: Towards High-authenticity and High-fidelity Image Composition
by: Wang, Jingyuan, et al.
Published: (2026)
by: Wang, Jingyuan, et al.
Published: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
FaceCom: Towards High-fidelity 3D Facial Shape Completion via Optimization and Inpainting Guidance
by: Li, Yinglong, et al.
Published: (2024)
by: Li, Yinglong, et al.
Published: (2024)
Sharpening Your Density Fields: Spiking Neuron Aided Fast Geometry Learning
by: Gu, Yi, et al.
Published: (2024)
by: Gu, Yi, et al.
Published: (2024)
Similar Items
-
OSTAF: A One-Shot Tuning Method for Improved Attribute-Focused T2I Personalization
by: Wang, Ye, et al.
Published: (2024) -
Anywhere: A Multi-Agent Framework for User-Guided, Reliable, and Diverse Foreground-Conditioned Image Generation
by: Xie, Tianyidan, et al.
Published: (2024) -
DiTraj: training-free trajectory control for video diffusion transformer
by: Lei, Cheng, et al.
Published: (2025) -
Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance
by: Wu, Song, et al.
Published: (2026) -
Revealing Vulnerabilities in Stable Diffusion via Targeted Attacks
by: Zhang, Chenyu, et al.
Published: (2024)