Fine-grained text-driven dual-human motion generation via dynamic hierarchical interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Mu, Wang, Yin, Leng, Zhiying, Liu, Jiapeng, Li, Frederick W. B., Liang, Xiaohui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fg-T2M++: LLMs-Augmented Fine-Grained Text Driven Human Motion Generation
by: Wang, Yin, et al.
Published: (2025)
by: Wang, Yin, et al.
Published: (2025)
MOST: Motion Diffusion Model for Rare Text via Temporal Clip Banzhaf Interaction
by: Wang, Yin, et al.
Published: (2025)
by: Wang, Yin, et al.
Published: (2025)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
Multimodal Priors-Augmented Text-Driven 3D Human-Object Interaction Generation
by: Wang, Yin, et al.
Published: (2026)
by: Wang, Yin, et al.
Published: (2026)
Cross-Temporal 3D Gaussian Splatting for Sparse-View Guided Scene Update
by: An, Zeyuan, et al.
Published: (2025)
by: An, Zeyuan, et al.
Published: (2025)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
RO-Bench: Large-scale robustness evaluation of MLLMs with text-driven counterfactual videos
by: Yang, Zixi, et al.
Published: (2025)
by: Yang, Zixi, et al.
Published: (2025)
MotionHiFlow: Text-to-motion via hierarchical flow matching
by: Li, Heng, et al.
Published: (2026)
by: Li, Heng, et al.
Published: (2026)
Uncertainty-aware Probabilistic 3D Human Motion Forecasting via Invertible Networks
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
Language-driven Fine-grained Retrieval
by: Wang, Shijie, et al.
Published: (2025)
by: Wang, Shijie, et al.
Published: (2025)
InstructAttribute: Fine-grained Object Attributes editing with Instruction
by: Yin, Xingxi, et al.
Published: (2025)
by: Yin, Xingxi, et al.
Published: (2025)
PHI: Bridging Domain Shift in Long-Term Action Quality Assessment via Progressive Hierarchical Instruction
by: Zhou, Kanglei, et al.
Published: (2025)
by: Zhou, Kanglei, et al.
Published: (2025)
MoDiPO: text-to-motion alignment via AI-feedback-driven Direct Preference Optimization
by: Pappa, Massimiliano, et al.
Published: (2024)
by: Pappa, Massimiliano, et al.
Published: (2024)
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
by: Xie, Chunyu, et al.
Published: (2025)
by: Xie, Chunyu, et al.
Published: (2025)
Towards more realistic human motion prediction with attention to motion coordination
by: Ding, Pengxiang, et al.
Published: (2024)
by: Ding, Pengxiang, et al.
Published: (2024)
Consistent text-to-image generation via scene de-contextualization
by: Tang, Song, et al.
Published: (2025)
by: Tang, Song, et al.
Published: (2025)
Multi-Resolution Haar Network: Enhancing human motion prediction via Haar transform
by: Lin, Li
Published: (2025)
by: Lin, Li
Published: (2025)
Continual Action Quality Assessment via Adaptive Manifold-Aligned Graph Regularization
by: Zhou, Kanglei, et al.
Published: (2025)
by: Zhou, Kanglei, et al.
Published: (2025)
Pulp Motion: Framing-aware multimodal camera and human motion generation
by: Courant, Robin, et al.
Published: (2025)
by: Courant, Robin, et al.
Published: (2025)
StreetTree: A Large-Scale Global Benchmark for Fine-Grained Tree Species Classification
by: Li, Jiapeng, et al.
Published: (2026)
by: Li, Jiapeng, et al.
Published: (2026)
FREAK: A Fine-grained Hallucination Evaluation Benchmark for Advanced MLLMs
by: Yin, Zhihan, et al.
Published: (2026)
by: Yin, Zhihan, et al.
Published: (2026)
Fine-grained Image Retrieval via Dual-Vision Adaptation
by: Jiang, Xin, et al.
Published: (2025)
by: Jiang, Xin, et al.
Published: (2025)
FireEdit: Fine-grained Instruction-based Image Editing via Region-aware Vision Language Model
by: Zhou, Jun, et al.
Published: (2025)
by: Zhou, Jun, et al.
Published: (2025)
Mesh-RFT: Enhancing Mesh Generation via Fine-grained Reinforcement Fine-Tuning
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
MAGR: Manifold-Aligned Graph Regularization for Continual Action Quality Assessment
by: Zhou, Kanglei, et al.
Published: (2024)
by: Zhou, Kanglei, et al.
Published: (2024)
KETA: Kinematic-Phrases-Enhanced Text-to-Motion Generation via Fine-grained Alignment
by: Jiang, Yu, et al.
Published: (2025)
by: Jiang, Yu, et al.
Published: (2025)
GeoMamba: A Geometry-driven MambaVision Framework and Dataset for Fine-grained Optical-SAR Object Retrieval
by: Fang, Tiantong, et al.
Published: (2026)
by: Fang, Tiantong, et al.
Published: (2026)
ODMixer: Fine-grained Spatial-temporal MLP for Metro Origin-Destination Prediction
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
FG-CLIP: Fine-Grained Visual and Textual Alignment
by: Xie, Chunyu, et al.
Published: (2025)
by: Xie, Chunyu, et al.
Published: (2025)
NewMove: Customizing text-to-video models with novel motions
by: Materzynska, Joanna, et al.
Published: (2023)
by: Materzynska, Joanna, et al.
Published: (2023)
Data-free Knowledge Distillation for Fine-grained Visual Categorization
by: Shao, Renrong, et al.
Published: (2024)
by: Shao, Renrong, et al.
Published: (2024)
Adversarial Reconstruction Feedback for Robust Fine-grained Generalization
by: Wang, Shijie, et al.
Published: (2025)
by: Wang, Shijie, et al.
Published: (2025)
Combo: Co-speech holistic 3D human motion generation and efficient customizable adaptation in harmony
by: Xu, Chao, et al.
Published: (2024)
by: Xu, Chao, et al.
Published: (2024)
SImpHAR: Advancing impedance-based human activity recognition using 3D simulation and text-to-motion models
by: Ray, Lala Shakti Swarup, et al.
Published: (2025)
by: Ray, Lala Shakti Swarup, et al.
Published: (2025)
MedFILIP: Medical Fine-grained Language-Image Pre-training
by: Liang, Xinjie, et al.
Published: (2025)
by: Liang, Xinjie, et al.
Published: (2025)
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
by: Zhang, Shi-Xue, et al.
Published: (2025)
by: Zhang, Shi-Xue, et al.
Published: (2025)
Little Strokes Fell Great Oaks: Boosting the Hierarchical Features for Multi-exposure Image Fusion
by: Mu, Pan, et al.
Published: (2024)
by: Mu, Pan, et al.
Published: (2024)
Distilling Monocular Foundation Model for Fine-grained Depth Completion
by: Liang, Yingping, et al.
Published: (2025)
by: Liang, Yingping, et al.
Published: (2025)
PestVL-Net: Enabling Multimodal Pest Learning via Fine-grained Vision-Language Interaction
by: Li, Xueheng, et al.
Published: (2026)
by: Li, Xueheng, et al.
Published: (2026)
A multimodal gesture recognition dataset for desktop human-computer interaction
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
Similar Items
-
Fg-T2M++: LLMs-Augmented Fine-Grained Text Driven Human Motion Generation
by: Wang, Yin, et al.
Published: (2025) -
MOST: Motion Diffusion Model for Rare Text via Temporal Clip Banzhaf Interaction
by: Wang, Yin, et al.
Published: (2025) -
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
by: Wang, Yin, et al.
Published: (2026) -
Multimodal Priors-Augmented Text-Driven 3D Human-Object Interaction Generation
by: Wang, Yin, et al.
Published: (2026) -
Cross-Temporal 3D Gaussian Splatting for Sparse-View Guided Scene Update
by: An, Zeyuan, et al.
Published: (2025)