HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
Fuente:
arXiv
Salvato in:
| Autori principali: | Narasimhaswamy, Supreeth, Bhattacharya, Uttaran, Chen, Xiang, Dasgupta, Ishita, Mitra, Saayan, Hoai, Minh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HOIST-Former: Hand-held Objects Identification, Segmentation, and Tracking in the Wild
di: Narasimhaswamy, Supreeth, et al.
Pubblicazione: (2024)
di: Narasimhaswamy, Supreeth, et al.
Pubblicazione: (2024)
SKALD: Learning-Based Shot Assembly for Coherent Multi-Shot Video Creation
di: Lu, Chen Yi, et al.
Pubblicazione: (2025)
di: Lu, Chen Yi, et al.
Pubblicazione: (2025)
Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions
di: Liao, Ting-Hsuan, et al.
Pubblicazione: (2025)
di: Liao, Ting-Hsuan, et al.
Pubblicazione: (2025)
Hand1000: Generating Realistic Hands from Text with Only 1,000 Images
di: Zhang, Haozhuo, et al.
Pubblicazione: (2024)
di: Zhang, Haozhuo, et al.
Pubblicazione: (2024)
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
di: Park, Junho, et al.
Pubblicazione: (2024)
di: Park, Junho, et al.
Pubblicazione: (2024)
Detecting Precise Hand Touch Moments in Egocentric Video
di: Nguyen, Huy Anh, et al.
Pubblicazione: (2026)
di: Nguyen, Huy Anh, et al.
Pubblicazione: (2026)
Text-guided Controllable Diffusion for Realistic Camouflage Images Generation
di: Qian, Yuhang, et al.
Pubblicazione: (2025)
di: Qian, Yuhang, et al.
Pubblicazione: (2025)
HanDrawer: Leveraging Spatial Information to Render Realistic Hands Using a Conditional Diffusion Model in Single Stage
di: Fu, Qifan, et al.
Pubblicazione: (2025)
di: Fu, Qifan, et al.
Pubblicazione: (2025)
A Penalty Goes a Long Way: Measuring Lexical Diversity in Synthetic Texts Under Prompt-Influenced Length Variations
di: Deshpande, Vijeta, et al.
Pubblicazione: (2025)
di: Deshpande, Vijeta, et al.
Pubblicazione: (2025)
Efficient and Robust Registration on the 3D Special Euclidean Group
di: Bhattacharya, Uttaran, et al.
Pubblicazione: (2019)
di: Bhattacharya, Uttaran, et al.
Pubblicazione: (2019)
Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models
di: Raji, Fadlullah, et al.
Pubblicazione: (2026)
di: Raji, Fadlullah, et al.
Pubblicazione: (2026)
Upcycling Text-to-Image Diffusion Models for Multi-Task Capabilities
di: Chavhan, Ruchika, et al.
Pubblicazione: (2025)
di: Chavhan, Ruchika, et al.
Pubblicazione: (2025)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
di: Liu, Huijie, et al.
Pubblicazione: (2025)
di: Liu, Huijie, et al.
Pubblicazione: (2025)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
di: Deng, Fan, et al.
Pubblicazione: (2024)
di: Deng, Fan, et al.
Pubblicazione: (2024)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
di: Xu, Yijia, et al.
Pubblicazione: (2026)
di: Xu, Yijia, et al.
Pubblicazione: (2026)
Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion
di: Kim, Jiwon, et al.
Pubblicazione: (2025)
di: Kim, Jiwon, et al.
Pubblicazione: (2025)
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
di: Bhattacharya, Uttaran, et al.
Pubblicazione: (2024)
di: Bhattacharya, Uttaran, et al.
Pubblicazione: (2024)
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
di: Kwon, Patrick, et al.
Pubblicazione: (2024)
di: Kwon, Patrick, et al.
Pubblicazione: (2024)
PathoGen: Diffusion-Based Synthesis of Realistic Lesions in Histopathology Images
di: Koohi-Moghadam, Mohamad, et al.
Pubblicazione: (2026)
di: Koohi-Moghadam, Mohamad, et al.
Pubblicazione: (2026)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
di: Tong, Lei, et al.
Pubblicazione: (2025)
di: Tong, Lei, et al.
Pubblicazione: (2025)
VidSketch: Hand-drawn Sketch-Driven Video Generation with Diffusion Control
di: Jiang, Lifan, et al.
Pubblicazione: (2025)
di: Jiang, Lifan, et al.
Pubblicazione: (2025)
Grounding Text-to-Image Diffusion Models for Controlled High-Quality Image Generation
di: Süleyman, Ahmad, et al.
Pubblicazione: (2025)
di: Süleyman, Ahmad, et al.
Pubblicazione: (2025)
Eye-for-an-eye: Appearance Transfer with Semantic Correspondence in Diffusion Models
di: Go, Sooyeon, et al.
Pubblicazione: (2024)
di: Go, Sooyeon, et al.
Pubblicazione: (2024)
Closer to Ground Truth: Realistic Shape and Appearance Labeled Data Generation for Unsupervised Underwater Image Segmentation
di: Jelea, Andrei, et al.
Pubblicazione: (2025)
di: Jelea, Andrei, et al.
Pubblicazione: (2025)
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
di: Gao, Xiang, et al.
Pubblicazione: (2024)
di: Gao, Xiang, et al.
Pubblicazione: (2024)
PhysHanDI: Physics-Based Reconstruction of Hand-Deformable Object Interactions
di: Lee, Jihyun, et al.
Pubblicazione: (2026)
di: Lee, Jihyun, et al.
Pubblicazione: (2026)
LSSGen: Leveraging Latent Space Scaling in Flow and Diffusion for Efficient Text to Image Generation
di: Tang, Jyun-Ze, et al.
Pubblicazione: (2025)
di: Tang, Jyun-Ze, et al.
Pubblicazione: (2025)
SimMotionEdit: Text-Based Human Motion Editing with Motion Similarity Prediction
di: Li, Zhengyuan, et al.
Pubblicazione: (2025)
di: Li, Zhengyuan, et al.
Pubblicazione: (2025)
Uncovering the Text Embedding in Text-to-Image Diffusion Models
di: Yu, Hu, et al.
Pubblicazione: (2024)
di: Yu, Hu, et al.
Pubblicazione: (2024)
Towards Realistic Scene Generation with LiDAR Diffusion Models
di: Ran, Haoxi, et al.
Pubblicazione: (2024)
di: Ran, Haoxi, et al.
Pubblicazione: (2024)
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
di: Wang, Yuxuan, et al.
Pubblicazione: (2025)
On the Scalability of Diffusion-based Text-to-Image Generation
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
di: Huang, Yifeng, et al.
Pubblicazione: (2023)
di: Huang, Yifeng, et al.
Pubblicazione: (2023)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
di: Wu, Fan, et al.
Pubblicazione: (2025)
di: Wu, Fan, et al.
Pubblicazione: (2025)
Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting
di: Zeng, Weili, et al.
Pubblicazione: (2024)
di: Zeng, Weili, et al.
Pubblicazione: (2024)
MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models
di: Zhou, Donghao, et al.
Pubblicazione: (2024)
di: Zhou, Donghao, et al.
Pubblicazione: (2024)
Development and Enhancement of Text-to-Image Diffusion Models
di: Sahu, Rajdeep Roshan
Pubblicazione: (2025)
di: Sahu, Rajdeep Roshan
Pubblicazione: (2025)
AID: Attention Interpolation of Text-to-Image Diffusion
di: He, Qiyuan, et al.
Pubblicazione: (2024)
di: He, Qiyuan, et al.
Pubblicazione: (2024)
DualMat: PBR Material Estimation via Coherent Dual-Path Diffusion
di: Huang, Yifeng, et al.
Pubblicazione: (2025)
di: Huang, Yifeng, et al.
Pubblicazione: (2025)
Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning
di: Zhang, Yu, et al.
Pubblicazione: (2025)
di: Zhang, Yu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
HOIST-Former: Hand-held Objects Identification, Segmentation, and Tracking in the Wild
di: Narasimhaswamy, Supreeth, et al.
Pubblicazione: (2024) -
SKALD: Learning-Based Shot Assembly for Coherent Multi-Shot Video Creation
di: Lu, Chen Yi, et al.
Pubblicazione: (2025) -
Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions
di: Liao, Ting-Hsuan, et al.
Pubblicazione: (2025) -
Hand1000: Generating Realistic Hands from Text with Only 1,000 Images
di: Zhang, Haozhuo, et al.
Pubblicazione: (2024) -
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
di: Park, Junho, et al.
Pubblicazione: (2024)