Diffusion Model-Based Size Variable Virtual Try-On Technology and Evaluation Method
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Shufang, Qian, Hang, Ni, Minxue, Li, Yaxuan, Ding, Wenxin, Liu, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Virtual Try-On with Garment-focused Diffusion Models
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
von: Wan, Siqi, et al.
Veröffentlicht: (2024)
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
Smart Fitting Room: A One-stop Framework for Matching-aware Virtual Try-on
von: Yu, Mingzhe, et al.
Veröffentlicht: (2024)
von: Yu, Mingzhe, et al.
Veröffentlicht: (2024)
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
von: Li, Dong, et al.
Veröffentlicht: (2025)
von: Li, Dong, et al.
Veröffentlicht: (2025)
3MDiT: Unified Tri-Modal Diffusion Transformer for Text-Driven Synchronized Audio-Video Generation
von: Li, Yaoru, et al.
Veröffentlicht: (2025)
von: Li, Yaoru, et al.
Veröffentlicht: (2025)
Period-conscious Time-series Reconstruction under Local Differential Privacy
von: Wang, Yaxuan, et al.
Veröffentlicht: (2026)
von: Wang, Yaxuan, et al.
Veröffentlicht: (2026)
Exploring the Robustness of Decision-Level Through Adversarial Attacks on LLM-Based Embodied Models
von: Liu, Shuyuan, et al.
Veröffentlicht: (2024)
von: Liu, Shuyuan, et al.
Veröffentlicht: (2024)
A Video Steganography for H.265/HEVC Based on Multiple CU Size and Block Structure Distortion
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Research on Piano Timbre Transformation System Based on Diffusion Model
von: Hsu, Chun-Chieh, et al.
Veröffentlicht: (2026)
von: Hsu, Chun-Chieh, et al.
Veröffentlicht: (2026)
A Subjective Quality Evaluation of 3D Mesh with Dynamic Level of Detail in Virtual Reality
von: Nguyen, Duc, et al.
Veröffentlicht: (2024)
von: Nguyen, Duc, et al.
Veröffentlicht: (2024)
Building and Evaluating a Realistic Virtual World for Large Scale Urban Exploration from 360° Videos
von: Takenawa, Mizuki, et al.
Veröffentlicht: (2025)
von: Takenawa, Mizuki, et al.
Veröffentlicht: (2025)
ConCLVD: Controllable Chinese Landscape Video Generation via Diffusion Model
von: Liu, Dingming, et al.
Veröffentlicht: (2024)
von: Liu, Dingming, et al.
Veröffentlicht: (2024)
EV-NVC: Efficient Variable bitrate Neural Video Compression
von: Hu, Yongcun, et al.
Veröffentlicht: (2025)
von: Hu, Yongcun, et al.
Veröffentlicht: (2025)
ConvBench: A Multi-Turn Conversation Evaluation Benchmark with Hierarchical Capability for Large Vision-Language Models
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
Tri-Subspaces Disentanglement for Multimodal Sentiment Analysis
von: Meng, Chunlei, et al.
Veröffentlicht: (2026)
von: Meng, Chunlei, et al.
Veröffentlicht: (2026)
PRISM: Exposing and Resolving Spurious Isolation in Federated Multimodal Continual Learning
von: Wu, Beining, et al.
Veröffentlicht: (2026)
von: Wu, Beining, et al.
Veröffentlicht: (2026)
Evaluating the Usability of Microgestures for Text Editing Tasks in Virtual Reality
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
von: Ma, Hongjian, et al.
Veröffentlicht: (2026)
Towards Alleviating Text-to-Image Retrieval Hallucination for CLIP in Zero-shot Learning
von: Wang, Hanyao, et al.
Veröffentlicht: (2024)
von: Wang, Hanyao, et al.
Veröffentlicht: (2024)
A Tri-Dynamic Preprocessing Framework for UGC Video Compression
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Rethinking Bjøntegaard Delta for Compression Efficiency Evaluation: Are We Calculating It Precisely and Reliably?
von: Hang, Xinyu, et al.
Veröffentlicht: (2024)
von: Hang, Xinyu, et al.
Veröffentlicht: (2024)
GestureHYDRA: Semantic Co-speech Gesture Synthesis via Hybrid Modality Diffusion Transformer and Cascaded-Synchronized Retrieval-Augmented Generation
von: Yang, Quanwei, et al.
Veröffentlicht: (2025)
von: Yang, Quanwei, et al.
Veröffentlicht: (2025)
Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning
von: Wang, Youze, et al.
Veröffentlicht: (2023)
von: Wang, Youze, et al.
Veröffentlicht: (2023)
DiffCL: A Diffusion-Based Contrastive Learning Framework with Semantic Alignment for Multimodal Recommendations
von: Song, Qiya, et al.
Veröffentlicht: (2025)
von: Song, Qiya, et al.
Veröffentlicht: (2025)
RoSMM: A Robust and Secure Multi-Modal Watermarking Framework for Diffusion Models
von: Fang, ZhongLi, et al.
Veröffentlicht: (2025)
von: Fang, ZhongLi, et al.
Veröffentlicht: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
von: Wang, Sen, et al.
Veröffentlicht: (2024)
von: Wang, Sen, et al.
Veröffentlicht: (2024)
REAL: Realism Evaluation of Text-to-Image Generation Models for Effective Data Augmentation
von: Li, Ran, et al.
Veröffentlicht: (2025)
von: Li, Ran, et al.
Veröffentlicht: (2025)
AIM: Let Any Multi-modal Large Language Models Embrace Efficient In-Context Learning
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
StyleSpeaker: Audio-Enhanced Fine-Grained Style Modeling for Speech-Driven 3D Facial Animation
von: Yang, An, et al.
Veröffentlicht: (2025)
von: Yang, An, et al.
Veröffentlicht: (2025)
Personalized Playback Technology: How Short Video Services Create Excellent User Experience
von: Deng, Weihui, et al.
Veröffentlicht: (2024)
von: Deng, Weihui, et al.
Veröffentlicht: (2024)
MetaDragonBoat: Exploring Paddling Techniques of Virtual Dragon Boating in a Metaverse Campus
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
Realistic Virtual Flood Experience System Using 360° Videos and 3D City Models Constructed from Building Footprints
von: Banno, Tatsuro, et al.
Veröffentlicht: (2026)
von: Banno, Tatsuro, et al.
Veröffentlicht: (2026)
MM-InstructEval: Zero-Shot Evaluation of (Multimodal) Large Language Models on Multimodal Reasoning Tasks
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
von: Yang, Xiaocui, et al.
Veröffentlicht: (2024)
PC-JND: Subjective Study and Dataset on Just Noticeable Difference for Point Clouds in 6DoF Virtual Reality
von: Fan, Chunling, et al.
Veröffentlicht: (2025)
von: Fan, Chunling, et al.
Veröffentlicht: (2025)
DIP: Diffusion Learning of Inconsistency Pattern for General DeepFake Detection
von: Nie, Fan, et al.
Veröffentlicht: (2024)
von: Nie, Fan, et al.
Veröffentlicht: (2024)
Subjective Quality Assessment of Dynamic 3D Meshes in Virtual Reality Environment
von: Nguyen, Duc V., et al.
Veröffentlicht: (2026)
von: Nguyen, Duc V., et al.
Veröffentlicht: (2026)
Investigating Conceptual Blending of a Diffusion Model for Improving Nonword-to-Image Generation
von: Matsuhira, Chihaya, et al.
Veröffentlicht: (2024)
von: Matsuhira, Chihaya, et al.
Veröffentlicht: (2024)
Language-oriented Semantic Communication for Image Transmission with Fine-Tuned Diffusion Model
von: Wei, Xinfeng, et al.
Veröffentlicht: (2024)
von: Wei, Xinfeng, et al.
Veröffentlicht: (2024)
Harmonizing Pixels and Melodies: Maestro-Guided Film Score Generation and Composition Style Transfer
von: Qi, F., et al.
Veröffentlicht: (2024)
von: Qi, F., et al.
Veröffentlicht: (2024)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
von: Li, Haobo, et al.
Veröffentlicht: (2024)
von: Li, Haobo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Improving Virtual Try-On with Garment-focused Diffusion Models
von: Wan, Siqi, et al.
Veröffentlicht: (2024) -
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025) -
Smart Fitting Room: A One-stop Framework for Matching-aware Virtual Try-on
von: Yu, Mingzhe, et al.
Veröffentlicht: (2024) -
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
von: Li, Dong, et al.
Veröffentlicht: (2025) -
3MDiT: Unified Tri-Modal Diffusion Transformer for Text-Driven Synchronized Audio-Video Generation
von: Li, Yaoru, et al.
Veröffentlicht: (2025)