ShoeModel: Learning to Wear on the User-specified Shoes via Diffusion Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Binghui, Li, Wenyu, Geng, Yifeng, Xie, Xuansong, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
Strictly-ID-Preserved and Controllable Accessory Advertising Image Generation
von: Xue, Youze, et al.
Veröffentlicht: (2024)
von: Xue, Youze, et al.
Veröffentlicht: (2024)
AnyText: Multilingual Visual Text Generation And Editing
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
von: Zhang, Yabo, et al.
Veröffentlicht: (2024)
von: Zhang, Yabo, et al.
Veröffentlicht: (2024)
Shoe Style-Invariant and Ground-Aware Learning for Dense Foot Contact Estimation
von: Jung, Daniel Sungho, et al.
Veröffentlicht: (2025)
von: Jung, Daniel Sungho, et al.
Veröffentlicht: (2025)
Put Myself in Your Shoes: Lifting the Egocentric Perspective from Exocentric Videos
von: Luo, Mi, et al.
Veröffentlicht: (2024)
von: Luo, Mi, et al.
Veröffentlicht: (2024)
3D Reconstruction of Shoes for Augmented Reality
von: Shrestha, Pratik, et al.
Veröffentlicht: (2025)
von: Shrestha, Pratik, et al.
Veröffentlicht: (2025)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
WordArt Designer API: User-Driven Artistic Typography Synthesis with Large Language Models on ModelScope
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
Overcoming Topology Agnosticism: Enhancing Skeleton-Based Action Recognition through Redefined Skeletal Topology Awareness
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2023)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2023)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
von: Qiao, Weidong, et al.
Veröffentlicht: (2026)
von: Qiao, Weidong, et al.
Veröffentlicht: (2026)
MV-VTON: Multi-View Virtual Try-On with Diffusion Models
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
U-VAP: User-specified Visual Appearance Personalization via Decoupled Self Augmentation
von: Wu, You, et al.
Veröffentlicht: (2024)
von: Wu, You, et al.
Veröffentlicht: (2024)
Tracking with Human-Intent Reasoning
von: Zhu, Jiawen, et al.
Veröffentlicht: (2023)
von: Zhu, Jiawen, et al.
Veröffentlicht: (2023)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
von: Yang, Feng, et al.
Veröffentlicht: (2025)
von: Yang, Feng, et al.
Veröffentlicht: (2025)
DiffusionGAN3D: Boosting Text-guided 3D Generation and Domain Adaptation by Combining 3D GANs and Diffusion Priors
von: Lei, Biwen, et al.
Veröffentlicht: (2023)
von: Lei, Biwen, et al.
Veröffentlicht: (2023)
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
von: He, Junjie, et al.
Veröffentlicht: (2025)
von: He, Junjie, et al.
Veröffentlicht: (2025)
Pixel-Aware Stable Diffusion for Realistic Image Super-resolution and Personalized Stylization
von: Yang, Tao, et al.
Veröffentlicht: (2023)
von: Yang, Tao, et al.
Veröffentlicht: (2023)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
von: Li, Xiaoming, et al.
Veröffentlicht: (2025)
von: Li, Xiaoming, et al.
Veröffentlicht: (2025)
SplatWeaver: Learning to Allocate Gaussian Primitives for Generalizable Novel View Synthesis
von: Wan, Yecong, et al.
Veröffentlicht: (2026)
von: Wan, Yecong, et al.
Veröffentlicht: (2026)
AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning
von: Yuan, Shihao, et al.
Veröffentlicht: (2025)
von: Yuan, Shihao, et al.
Veröffentlicht: (2025)
GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection
von: Yao, Hang, et al.
Veröffentlicht: (2024)
von: Yao, Hang, et al.
Veröffentlicht: (2024)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2024)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
ScrollScape: Unlocking 32K Image Generation With Video Diffusion Priors
von: Yu, Haodong, et al.
Veröffentlicht: (2026)
von: Yu, Haodong, et al.
Veröffentlicht: (2026)
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
von: Liu, Chunpu, et al.
Veröffentlicht: (2024)
von: Liu, Chunpu, et al.
Veröffentlicht: (2024)
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
ACE: Anti-Editing Concept Erasure in Text-to-Image Models
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
von: Wang, Zihao, et al.
Veröffentlicht: (2025)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
von: Ni, Minheng, et al.
Veröffentlicht: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
von: Zhang, Zhilu, et al.
Veröffentlicht: (2024)
von: Zhang, Zhilu, et al.
Veröffentlicht: (2024)
Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models
von: Huang, Zitong, et al.
Veröffentlicht: (2026)
von: Huang, Zitong, et al.
Veröffentlicht: (2026)
En3D: An Enhanced Generative Model for Sculpting 3D Humans from 2D Synthetic Data
von: Men, Yifang, et al.
Veröffentlicht: (2024)
von: Men, Yifang, et al.
Veröffentlicht: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
von: Xu, Wan, et al.
Veröffentlicht: (2023)
von: Xu, Wan, et al.
Veröffentlicht: (2023)
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
von: Huang, Tianyu, et al.
Veröffentlicht: (2024)
von: Huang, Tianyu, et al.
Veröffentlicht: (2024)
IMWA: Iterative Model Weight Averaging Benefits Class-Imbalanced Learning Tasks
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
von: Huang, Zitong, et al.
Veröffentlicht: (2024)
RobustMVS: Single Domain Generalized Deep Multi-view Stereo
von: Xu, Hongbin, et al.
Veröffentlicht: (2024)
von: Xu, Hongbin, et al.
Veröffentlicht: (2024)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
von: Chu, Jisheng, et al.
Veröffentlicht: (2026)
von: Chu, Jisheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
von: Chen, Binghui, et al.
Veröffentlicht: (2024) -
Strictly-ID-Preserved and Controllable Accessory Advertising Image Generation
von: Xue, Youze, et al.
Veröffentlicht: (2024) -
AnyText: Multilingual Visual Text Generation And Editing
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023) -
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
von: Zhang, Yabo, et al.
Veröffentlicht: (2024) -
Shoe Style-Invariant and Ground-Aware Learning for Dense Foot Contact Estimation
von: Jung, Daniel Sungho, et al.
Veröffentlicht: (2025)