THOM: Generating Physically Plausible Hand-Object Meshes From Text
Fuente:
arXiv
Salvato in:
| Autori principali: | Jeong, Uyoung, Tiruneh, Yihalem Yimolal, Chang, Hyung Jin, Baek, Seungryul, Kim, Kwang In |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
di: Kim, Junsu, et al.
Pubblicazione: (2024)
di: Kim, Junsu, et al.
Pubblicazione: (2024)
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
di: Jeong, Uyoung, et al.
Pubblicazione: (2023)
di: Jeong, Uyoung, et al.
Pubblicazione: (2023)
QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects
di: Ismayilzada, Elkhan, et al.
Pubblicazione: (2025)
di: Ismayilzada, Elkhan, et al.
Pubblicazione: (2025)
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
di: Jeong, Uyoung, et al.
Pubblicazione: (2025)
di: Jeong, Uyoung, et al.
Pubblicazione: (2025)
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
di: Sayem, MD Khalequzzaman Chowdhury, et al.
Pubblicazione: (2026)
di: Sayem, MD Khalequzzaman Chowdhury, et al.
Pubblicazione: (2026)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
di: Cha, Junuk, et al.
Pubblicazione: (2024)
di: Cha, Junuk, et al.
Pubblicazione: (2024)
VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model
di: Kim, Junsu, et al.
Pubblicazione: (2024)
di: Kim, Junsu, et al.
Pubblicazione: (2024)
Exploiting Style Latent Flows for Generalizing Deepfake Video Detection
di: Choi, Jongwook, et al.
Pubblicazione: (2024)
di: Choi, Jongwook, et al.
Pubblicazione: (2024)
B-RIGHT: Benchmark Re-evaluation for Integrity in Generalized Human-Object Interaction Testing
di: Jang, Yoojin, et al.
Pubblicazione: (2025)
di: Jang, Yoojin, et al.
Pubblicazione: (2025)
Can Synthetic Images Conquer Forgetting? Beyond Unexplored Doubts in Few-Shot Class-Incremental Learning
di: Kim, Junsu, et al.
Pubblicazione: (2025)
di: Kim, Junsu, et al.
Pubblicazione: (2025)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
di: Fan, Zicong, et al.
Pubblicazione: (2024)
di: Fan, Zicong, et al.
Pubblicazione: (2024)
As-Plausible-As-Possible: Plausibility-Aware Mesh Deformation Using 2D Diffusion Priors
di: Yoo, Seungwoo, et al.
Pubblicazione: (2023)
di: Yoo, Seungwoo, et al.
Pubblicazione: (2023)
Beyond Synthetic Replays: Turning Diffusion Features into Few-Shot Class-Incremental Learning Knowledge
di: Kim, Junsu, et al.
Pubblicazione: (2025)
di: Kim, Junsu, et al.
Pubblicazione: (2025)
NCRF: Neural Contact Radiance Fields for Free-Viewpoint Rendering of Hand-Object Interaction
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
NL2Contact: Natural Language Guided 3D Hand-Object Contact Modeling with Diffusion Model
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
di: Zhang, Zhongqun, et al.
Pubblicazione: (2024)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
di: Xu, Hao, et al.
Pubblicazione: (2024)
di: Xu, Hao, et al.
Pubblicazione: (2024)
Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility
di: Hao, Yutong, et al.
Pubblicazione: (2025)
di: Hao, Yutong, et al.
Pubblicazione: (2025)
PersonaBooth: Personalized Text-to-Motion Generation
di: Kim, Boeun, et al.
Pubblicazione: (2025)
di: Kim, Boeun, et al.
Pubblicazione: (2025)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
di: Kim, Manjin, et al.
Pubblicazione: (2026)
di: Kim, Manjin, et al.
Pubblicazione: (2026)
3D Hand Mesh-Guided AI-Generated Malformed Hand Refinement with Hand Pose Transformation via Diffusion Model
di: Feng, Chen-Bin, et al.
Pubblicazione: (2025)
di: Feng, Chen-Bin, et al.
Pubblicazione: (2025)
TextGaze: Gaze-Controllable Face Generation with Natural Language
di: Wang, Hengfei, et al.
Pubblicazione: (2024)
di: Wang, Hengfei, et al.
Pubblicazione: (2024)
Beyond Spatial Frequency: Pixel-wise Temporal Frequency-based Deepfake Video Detection
di: Kim, Taehoon, et al.
Pubblicazione: (2025)
di: Kim, Taehoon, et al.
Pubblicazione: (2025)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
di: Cha, Junuk, et al.
Pubblicazione: (2025)
di: Cha, Junuk, et al.
Pubblicazione: (2025)
Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models
di: Jin, Hyundong, et al.
Pubblicazione: (2025)
di: Jin, Hyundong, et al.
Pubblicazione: (2025)
From Generated Human Videos to Physically Plausible Robot Trajectories
di: Ni, James, et al.
Pubblicazione: (2025)
di: Ni, James, et al.
Pubblicazione: (2025)
Revisiting Reliability in the Reasoning-based Pose Estimation Benchmark
di: Kim, Junsu, et al.
Pubblicazione: (2025)
di: Kim, Junsu, et al.
Pubblicazione: (2025)
Text2Relight: Creative Portrait Relighting with Text Guidance
di: Cha, Junuk, et al.
Pubblicazione: (2024)
di: Cha, Junuk, et al.
Pubblicazione: (2024)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
di: Han, Guangyi, et al.
Pubblicazione: (2025)
di: Han, Guangyi, et al.
Pubblicazione: (2025)
From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models
di: Gong, Weile, et al.
Pubblicazione: (2026)
di: Gong, Weile, et al.
Pubblicazione: (2026)
PhysPart: Physically Plausible Part Completion for Interactable Objects
di: Luo, Rundong, et al.
Pubblicazione: (2024)
di: Luo, Rundong, et al.
Pubblicazione: (2024)
OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance
di: Wang, Cong, et al.
Pubblicazione: (2026)
di: Wang, Cong, et al.
Pubblicazione: (2026)
TIGeR: Text-Instructed Generation and Refinement for Template-Free Hand-Object Interaction
di: Huang, Yiyao, et al.
Pubblicazione: (2025)
di: Huang, Yiyao, et al.
Pubblicazione: (2025)
3D Reconstruction of Interacting Multi-Person in Clothing from a Single Image
di: Cha, Junuk, et al.
Pubblicazione: (2024)
di: Cha, Junuk, et al.
Pubblicazione: (2024)
MMPhysVideo: Scaling Physical Plausibility in Video Generation via Joint Multimodal Modeling
di: Lin, Shubo, et al.
Pubblicazione: (2026)
di: Lin, Shubo, et al.
Pubblicazione: (2026)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
di: Tse, Tze Ho Elden, et al.
Pubblicazione: (2025)
di: Tse, Tze Ho Elden, et al.
Pubblicazione: (2025)
STMR: Spiral Transformer for Hand Mesh Reconstruction
di: Xie, Huilong, et al.
Pubblicazione: (2024)
di: Xie, Huilong, et al.
Pubblicazione: (2024)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
di: Zhou, Bohan, et al.
Pubblicazione: (2025)
di: Zhou, Bohan, et al.
Pubblicazione: (2025)
BIGS: Bimanual Category-agnostic Interaction Reconstruction from Monocular Videos via 3D Gaussian Splatting
di: On, Jeongwan, et al.
Pubblicazione: (2025)
di: On, Jeongwan, et al.
Pubblicazione: (2025)
A Self-Supervised Approach on Motion Calibration for Enhancing Physical Plausibility in Text-to-Motion
di: Shim, Gahyeon, et al.
Pubblicazione: (2026)
di: Shim, Gahyeon, et al.
Pubblicazione: (2026)
TextMesh4D: Text-to-4D Mesh Generation via Jacobian Deformation Field
di: Dai, Sisi, et al.
Pubblicazione: (2025)
di: Dai, Sisi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
di: Kim, Junsu, et al.
Pubblicazione: (2024) -
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
di: Jeong, Uyoung, et al.
Pubblicazione: (2023) -
QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects
di: Ismayilzada, Elkhan, et al.
Pubblicazione: (2025) -
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
di: Jeong, Uyoung, et al.
Pubblicazione: (2025) -
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
di: Sayem, MD Khalequzzaman Chowdhury, et al.
Pubblicazione: (2026)