Gespeichert in:
| Hauptverfasser: | Fan, Siyuan, Du, Bo, Cai, Xiantao, Peng, Bo, Sun, Longling |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2408.03302 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
3D Human Interaction Generation: A Survey
von: Fan, Siyuan, et al.
Veröffentlicht: (2025)
von: Fan, Siyuan, et al.
Veröffentlicht: (2025)
Controllable Text-to-Motion Generation via Modular Body-Part Phase Control
von: Dai, Minyue, et al.
Veröffentlicht: (2026)
von: Dai, Minyue, et al.
Veröffentlicht: (2026)
ParCo: Part-Coordinating Text-to-Motion Synthesis
von: Zou, Qiran, et al.
Veröffentlicht: (2024)
von: Zou, Qiran, et al.
Veröffentlicht: (2024)
SFA: Scan, Focus, and Amplify toward Guidance-aware Answering for Video TextVQA
von: He, Haibin, et al.
Veröffentlicht: (2025)
von: He, Haibin, et al.
Veröffentlicht: (2025)
ParTY: Part-Guidance for Expressive Text-to-Motion Synthesis
von: Heo, KunHo, et al.
Veröffentlicht: (2026)
von: Heo, KunHo, et al.
Veröffentlicht: (2026)
Autonomous Character-Scene Interaction Synthesis from Text Instruction
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2023)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2023)
Generating Human Interaction Motions in Scenes with Text Control
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility
von: Zhang, Xiantao
Veröffentlicht: (2025)
von: Zhang, Xiantao
Veröffentlicht: (2025)
Reasoning-OCR: Can Large Multimodal Models Solve Complex Logical Reasoning Problems from OCR Cues?
von: He, Haibin, et al.
Veröffentlicht: (2025)
von: He, Haibin, et al.
Veröffentlicht: (2025)
T3M: Text Guided 3D Human Motion Synthesis from Speech
von: Peng, Wenshuo, et al.
Veröffentlicht: (2024)
von: Peng, Wenshuo, et al.
Veröffentlicht: (2024)
InTeX: Interactive Text-to-texture Synthesis via Unified Depth-aware Inpainting
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
AnyText2: Visual Text Generation and Editing With Customizable Attributes
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
MaTe3D: Mask-guided Text-based 3D-aware Portrait Editing
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
von: Zhou, Kangneng, et al.
Veröffentlicht: (2023)
TextSSR: Diffusion-based Data Synthesis for Scene Text Recognition
von: Ye, Xingsong, et al.
Veröffentlicht: (2024)
von: Ye, Xingsong, et al.
Veröffentlicht: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
Text Data-Centric Image Captioning with Interactive Prompts
von: Wang, Yiyu, et al.
Veröffentlicht: (2024)
von: Wang, Yiyu, et al.
Veröffentlicht: (2024)
Rethink Sparse Signals for Pose-guided Text-to-image Generation
von: Xuan, Wenjie, et al.
Veröffentlicht: (2025)
von: Xuan, Wenjie, et al.
Veröffentlicht: (2025)
Articulate That Object Part (ATOP): 3D Part Articulation via Text and Motion Personalization
von: Vora, Aditya, et al.
Veröffentlicht: (2025)
von: Vora, Aditya, et al.
Veröffentlicht: (2025)
Hear the Scene: Audio-Enhanced Text Spotting
von: Li, Jing, et al.
Veröffentlicht: (2024)
von: Li, Jing, et al.
Veröffentlicht: (2024)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
Text to Blind Motion
von: Kim, Hee Jae, et al.
Veröffentlicht: (2024)
von: Kim, Hee Jae, et al.
Veröffentlicht: (2024)
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
von: He, Haibin, et al.
Veröffentlicht: (2026)
von: He, Haibin, et al.
Veröffentlicht: (2026)
ET-SAM: Efficient Point Prompt Prediction in SAM for Unified Scene Text Detection and Layout Analysis
von: Zhang, Xike, et al.
Veröffentlicht: (2026)
von: Zhang, Xike, et al.
Veröffentlicht: (2026)
Motion-aware Dynamic Graph Neural Network for Video Compressive Sensing
von: Lu, Ruiying, et al.
Veröffentlicht: (2022)
von: Lu, Ruiying, et al.
Veröffentlicht: (2022)
Autonomous LLM-Enhanced Adversarial Attack for Text-to-Motion
von: Miao, Honglei, et al.
Veröffentlicht: (2024)
von: Miao, Honglei, et al.
Veröffentlicht: (2024)
GUESS:GradUally Enriching SyntheSis for Text-Driven Human Motion Generation
von: Gao, Xuehao, et al.
Veröffentlicht: (2024)
von: Gao, Xuehao, et al.
Veröffentlicht: (2024)
Leveraging Text-to-Image Diffusion Models for Unsupervised Visual Object Tracking
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
von: Zhang, Zhengbo, et al.
Veröffentlicht: (2026)
SegVol: Universal and Interactive Volumetric Medical Image Segmentation
von: Du, Yuxin, et al.
Veröffentlicht: (2023)
von: Du, Yuxin, et al.
Veröffentlicht: (2023)
Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis
von: Gong, Jingyu, et al.
Veröffentlicht: (2024)
von: Gong, Jingyu, et al.
Veröffentlicht: (2024)
Towards Open Domain Text-Driven Synthesis of Multi-Person Motions
von: Shan, Mengyi, et al.
Veröffentlicht: (2024)
von: Shan, Mengyi, et al.
Veröffentlicht: (2024)
ConditionVideo: Training-Free Condition-Guided Text-to-Video Generation
von: Peng, Bo, et al.
Veröffentlicht: (2023)
von: Peng, Bo, et al.
Veröffentlicht: (2023)
PALUM: Part-based Attention Learning for Unified Motion Retargeting
von: Liu, Siqi, et al.
Veröffentlicht: (2026)
von: Liu, Siqi, et al.
Veröffentlicht: (2026)
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
von: Jin, Peng, et al.
Veröffentlicht: (2024)
von: Jin, Peng, et al.
Veröffentlicht: (2024)
Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image Modeling
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
von: Wang, Zixiao, et al.
Veröffentlicht: (2024)
Topology-Agnostic Animal Motion Generation from Text Prompt
von: Chen, Keyi, et al.
Veröffentlicht: (2025)
von: Chen, Keyi, et al.
Veröffentlicht: (2025)
Text2Place: Affordance-aware Text Guided Human Placement
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
Text-Video Multi-Grained Integration for Video Moment Montage
von: Yin, Zhihui, et al.
Veröffentlicht: (2024)
von: Yin, Zhihui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
3D Human Interaction Generation: A Survey
von: Fan, Siyuan, et al.
Veröffentlicht: (2025) -
Controllable Text-to-Motion Generation via Modular Body-Part Phase Control
von: Dai, Minyue, et al.
Veröffentlicht: (2026) -
ParCo: Part-Coordinating Text-to-Motion Synthesis
von: Zou, Qiran, et al.
Veröffentlicht: (2024) -
SFA: Scan, Focus, and Amplify toward Guidance-aware Answering for Video TextVQA
von: He, Haibin, et al.
Veröffentlicht: (2025) -
ParTY: Part-Guidance for Expressive Text-to-Motion Synthesis
von: Heo, KunHo, et al.
Veröffentlicht: (2026)