InstructVTON: Optimal Auto-Masking and Natural-Language-Guided Interactive Style Control for Inpainting-Based Virtual Try-On
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Julien, Qiu, Shuwen, Li, Qi, Xu, Xingzi, Seyfioglu, Mehmet Saygin, Asadi, Kavosh, Bouyarmane, Karim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Style-Instructed Mask-Free Virtual Try On
by: Zhang, Mengqi, et al.
Published: (2026)
by: Zhang, Mengqi, et al.
Published: (2026)
DiT-VTON: Diffusion Transformer Framework for Unified Multi-Category Virtual Try-On and Virtual Try-All with Integrated Image Editing
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
DEFT-VTON: Efficient Virtual Try-On with Consistent Generalised H-Transform
by: Xu, Xingzi, et al.
Published: (2025)
by: Xu, Xingzi, et al.
Published: (2025)
Efficient Encoder-Free Pose Conditioning and Pose Control for Virtual Try-On
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
Diffuse to Choose: Enriching Image Conditioned Inpainting in Latent Diffusion Models for Virtual Try-All
by: Seyfioglu, Mehmet Saygin, et al.
Published: (2024)
by: Seyfioglu, Mehmet Saygin, et al.
Published: (2024)
When Rubrics Fail: Error Enumeration as Reward in Reference-Free RL Post-Training for Virtual Try-On
by: Ikezogwo, Wisdom, et al.
Published: (2026)
by: Ikezogwo, Wisdom, et al.
Published: (2026)
Adjoint sharding for very long context training of state space models
by: Xu, Xingzi, et al.
Published: (2025)
by: Xu, Xingzi, et al.
Published: (2025)
C2-DPO: Constrained Controlled Direct Preference Optimization
by: Asadi, Kavosh, et al.
Published: (2025)
by: Asadi, Kavosh, et al.
Published: (2025)
Learning to Reason Efficiently with Discounted Reinforcement Learning
by: Ayoub, Alex, et al.
Published: (2025)
by: Ayoub, Alex, et al.
Published: (2025)
VTON-IT: Virtual Try-On using Image Translation
by: Adhikari, Santosh, et al.
Published: (2023)
by: Adhikari, Santosh, et al.
Published: (2023)
FLDM-VTON: Faithful Latent Diffusion Model for Virtual Try-on
by: Wang, Chenhui, et al.
Published: (2024)
by: Wang, Chenhui, et al.
Published: (2024)
MFP-VTON: Enhancing Mask-Free Person-to-Person Virtual Try-On via Diffusion Transformer
by: Shen, Le, et al.
Published: (2025)
by: Shen, Le, et al.
Published: (2025)
Mobile-VTON: High-Fidelity On-Device Virtual Try-On
by: Wan, Zhenchen, et al.
Published: (2026)
by: Wan, Zhenchen, et al.
Published: (2026)
OmniVTON: Training-Free Universal Virtual Try-On
by: Yang, Zhaotong, et al.
Published: (2025)
by: Yang, Zhaotong, et al.
Published: (2025)
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
by: Zhang, Xuanpu, et al.
Published: (2024)
by: Zhang, Xuanpu, et al.
Published: (2024)
FW-VTON: Flattening-and-Warping for Person-to-Person Virtual Try-on
by: Wang, Zheng, et al.
Published: (2025)
by: Wang, Zheng, et al.
Published: (2025)
AvatarVTON: 4D Virtual Try-On for Animatable Avatars
by: Jiang, Zicheng, et al.
Published: (2025)
by: Jiang, Zicheng, et al.
Published: (2025)
MV-VTON: Multi-View Virtual Try-On with Diffusion Models
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
ACDG-VTON: Accurate and Contained Diffusion Generation for Virtual Try-On
by: Zhang, Jeffrey, et al.
Published: (2024)
by: Zhang, Jeffrey, et al.
Published: (2024)
MC-VTON: Minimal Control Virtual Try-On Diffusion Transformer
by: Luan, Junsheng, et al.
Published: (2025)
by: Luan, Junsheng, et al.
Published: (2025)
VTON-HandFit: Virtual Try-on for Arbitrary Hand Pose Guided by Hand Priors Embedding
by: Liang, Yujie, et al.
Published: (2024)
by: Liang, Yujie, et al.
Published: (2024)
HF-VTON: High-Fidelity Virtual Try-On via Consistent Geometric and Semantic Alignment
by: Meng, Ming, et al.
Published: (2025)
by: Meng, Ming, et al.
Published: (2025)
GS-VTON: Controllable 3D Virtual Try-on with Gaussian Splatting
by: Cao, Yukang, et al.
Published: (2024)
by: Cao, Yukang, et al.
Published: (2024)
SIFT-VTON: Geometric Correspondence Supervision on Cross-Attention for Virtual Try-On
by: Takemoto, Kosuke, et al.
Published: (2026)
by: Takemoto, Kosuke, et al.
Published: (2026)
Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos
by: Seyfioglu, Mehmet Saygin, et al.
Published: (2023)
by: Seyfioglu, Mehmet Saygin, et al.
Published: (2023)
CrossVTON: Mimicking the Logic Reasoning on Cross-category Virtual Try-on guided by Tri-zone Priors
by: Luo, Donghao, et al.
Published: (2025)
by: Luo, Donghao, et al.
Published: (2025)
OmniVTON++: Training-Free Universal Virtual Try-On with Principal Pose Guidance
by: Yang, Zhaotong, et al.
Published: (2026)
by: Yang, Zhaotong, et al.
Published: (2026)
DreamVTON: Customizing 3D Virtual Try-on with Personalized Diffusion Models
by: Xie, Zhenyu, et al.
Published: (2024)
by: Xie, Zhenyu, et al.
Published: (2024)
VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction
by: He, Zijian, et al.
Published: (2025)
by: He, Zijian, et al.
Published: (2025)
DS-VTON: An Enhanced Dual-Scale Coarse-to-Fine Framework for Virtual Try-On
by: Sun, Xianbing, et al.
Published: (2025)
by: Sun, Xianbing, et al.
Published: (2025)
CatVTON: Concatenation Is All You Need for Virtual Try-On with Diffusion Models
by: Chong, Zheng, et al.
Published: (2024)
by: Chong, Zheng, et al.
Published: (2024)
D$^4$-VTON: Dynamic Semantics Disentangling for Differential Diffusion based Virtual Try-On
by: Yang, Zhaotong, et al.
Published: (2024)
by: Yang, Zhaotong, et al.
Published: (2024)
DH-VTON: Deep Text-Driven Virtual Try-On via Hybrid Attention Learning
by: Wei, Jiabao, et al.
Published: (2024)
by: Wei, Jiabao, et al.
Published: (2024)
LPH-VTON: Resolving the Structure-Texture Dilemma of Virtual Try-On via Latent Process Handover
by: Liu, Yixin, et al.
Published: (2026)
by: Liu, Yixin, et al.
Published: (2026)
MuGa-VTON: Multi-Garment Virtual Try-On via Diffusion Transformers with Prompt Customization
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
MedicalNarratives: Connecting Medical Vision and Language with Localized Narratives
by: Ikezogwo, Wisdom O., et al.
Published: (2025)
by: Ikezogwo, Wisdom O., et al.
Published: (2025)
OpenVTON-Bench: A Large-Scale High-Resolution Benchmark for Controllable Virtual Try-On Evaluation
by: Li, Jin, et al.
Published: (2026)
by: Li, Jin, et al.
Published: (2026)
OmniTry: Virtual Try-On Anything without Masks
by: Feng, Yutong, et al.
Published: (2025)
by: Feng, Yutong, et al.
Published: (2025)
Displacement-Resistant Extensions of DPO with Nonconvex $f$-Divergences
by: Pipano, Idan, et al.
Published: (2026)
by: Pipano, Idan, et al.
Published: (2026)
GaussianVTON: 3D Human Virtual Try-ON via Multi-Stage Gaussian Splatting Editing with Image Prompting
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
Similar Items
-
Style-Instructed Mask-Free Virtual Try On
by: Zhang, Mengqi, et al.
Published: (2026) -
DiT-VTON: Diffusion Transformer Framework for Unified Multi-Category Virtual Try-On and Virtual Try-All with Integrated Image Editing
by: Li, Qi, et al.
Published: (2025) -
DEFT-VTON: Efficient Virtual Try-On with Consistent Generalised H-Transform
by: Xu, Xingzi, et al.
Published: (2025) -
Efficient Encoder-Free Pose Conditioning and Pose Control for Virtual Try-On
by: Li, Qi, et al.
Published: (2025) -
Diffuse to Choose: Enriching Image Conditioned Inpainting in Latent Diffusion Models for Virtual Try-All
by: Seyfioglu, Mehmet Saygin, et al.
Published: (2024)