Bilingual Text-to-Motion Generation: A New Benchmark and Baselines
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Weng, Wanjiang, Tan, Xiaofeng, Shu, Xiangbo, Xie, Guo-Sen, Zhou, Pan, Wang, Hongsong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReAlign: Bilingual Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025)
EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
SoPo: Text-to-Motion Generation Using Semi-Online Preference Optimization
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2024)
USDRL: Unified Skeleton-Based Dense Representation Learning with Multi-Grained Feature Decorrelation
von: Weng, Wanjiang, et al.
Veröffentlicht: (2024)
von: Weng, Wanjiang, et al.
Veröffentlicht: (2024)
Foundation Model for Skeleton-Based Human Action Understanding
von: Wang, Hongsong, et al.
Veröffentlicht: (2025)
von: Wang, Hongsong, et al.
Veröffentlicht: (2025)
Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition
von: Kuang, Jidong, et al.
Veröffentlicht: (2026)
von: Kuang, Jidong, et al.
Veröffentlicht: (2026)
Temporal Consistency-Aware Text-to-Motion Generation
von: Wang, Hongsong, et al.
Veröffentlicht: (2026)
von: Wang, Hongsong, et al.
Veröffentlicht: (2026)
PAMD: Plausibility-Aware Motion Diffusion Model for Long Dance Generation
von: Wang, Hongsong, et al.
Veröffentlicht: (2025)
von: Wang, Hongsong, et al.
Veröffentlicht: (2025)
FTMoMamba: Motion Generation with Frequency and Text State Space Models
von: Li, Chengjian, et al.
Veröffentlicht: (2024)
von: Li, Chengjian, et al.
Veröffentlicht: (2024)
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
DreamCS: Geometry-Aware Text-to-3D Generation with Unpaired 3D Reward Supervision
von: Zou, Xiandong, et al.
Veröffentlicht: (2025)
von: Zou, Xiandong, et al.
Veröffentlicht: (2025)
Kernel-Aware Graph Prompt Learning for Few-Shot Anomaly Detection
von: Tao, Fenfang, et al.
Veröffentlicht: (2024)
von: Tao, Fenfang, et al.
Veröffentlicht: (2024)
Vision-centric Token Compression in Large Language Model
von: Xing, Ling, et al.
Veröffentlicht: (2025)
von: Xing, Ling, et al.
Veröffentlicht: (2025)
TextInVision: Text and Prompt Complexity Driven Visual Text Generation Benchmark
von: Fallah, Forouzan, et al.
Veröffentlicht: (2025)
von: Fallah, Forouzan, et al.
Veröffentlicht: (2025)
LegalEval-Q: A New Benchmark for The Quality Evaluation of LLM-Generated Legal Text
von: yunhan, Li, et al.
Veröffentlicht: (2025)
von: yunhan, Li, et al.
Veröffentlicht: (2025)
Frequency-Guided Diffusion Model with Perturbation Training for Skeleton-Based Video Anomaly Detection
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2024)
Interleaved Scene Graphs for Interleaved Text-and-Image Generation Assessment
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
von: Chen, Dongping, et al.
Veröffentlicht: (2024)
Attack-Augmentation Mixing-Contrastive Skeletal Representation Learning
von: Xu, Binqian, et al.
Veröffentlicht: (2023)
von: Xu, Binqian, et al.
Veröffentlicht: (2023)
FedMLLM: Federated Fine-tuning MLLM on Multimodal Heterogeneity Data
von: Xu, Binqian, et al.
Veröffentlicht: (2024)
von: Xu, Binqian, et al.
Veröffentlicht: (2024)
Coordinate-Based Dual-Constrained Autoregressive Motion Generation
von: Ding, Kang, et al.
Veröffentlicht: (2026)
von: Ding, Kang, et al.
Veröffentlicht: (2026)
Text2Vis: A Challenging and Diverse Benchmark for Generating Multimodal Visualizations from Text
von: Rahman, Mizanur, et al.
Veröffentlicht: (2025)
von: Rahman, Mizanur, et al.
Veröffentlicht: (2025)
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries
von: Wu, Yin, et al.
Veröffentlicht: (2025)
von: Wu, Yin, et al.
Veröffentlicht: (2025)
Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs
von: Gao, Xin, et al.
Veröffentlicht: (2026)
von: Gao, Xin, et al.
Veröffentlicht: (2026)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
MMMG: A Massive, Multidisciplinary, Multi-Tier Generation Benchmark for Text-to-Image Reasoning
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
von: Luo, Yuxuan, et al.
Veröffentlicht: (2025)
AdaFPP: Adapt-Focused Bi-Propagating Prototype Learning for Panoramic Activity Recognition
von: Cao, Meiqi, et al.
Veröffentlicht: (2024)
von: Cao, Meiqi, et al.
Veröffentlicht: (2024)
MVP-Shot: Multi-Velocity Progressive-Alignment Framework for Few-Shot Action Recognition
von: Qu, Hongyu, et al.
Veröffentlicht: (2024)
von: Qu, Hongyu, et al.
Veröffentlicht: (2024)
Mojito: Motion Trajectory and Intensity Control for Video Generation
von: He, Xuehai, et al.
Veröffentlicht: (2024)
von: He, Xuehai, et al.
Veröffentlicht: (2024)
ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
MoTe: Learning Motion-Text Diffusion Model for Multiple Generation Tasks
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
Beyond Cropped Regions: New Benchmark and Corresponding Baseline for Chinese Scene Text Retrieval in Diverse Layouts
von: Li, Gengluo, et al.
Veröffentlicht: (2025)
von: Li, Gengluo, et al.
Veröffentlicht: (2025)
BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
T2MBench: A Benchmark for Out-of-Distribution Text-to-Motion Generation
von: Yang, Bin, et al.
Veröffentlicht: (2026)
von: Yang, Bin, et al.
Veröffentlicht: (2026)
LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models
von: Wang, Chenglin, et al.
Veröffentlicht: (2026)
von: Wang, Chenglin, et al.
Veröffentlicht: (2026)
Draw ALL Your Imagine: A Holistic Benchmark and Agent Framework for Complex Instruction-based Image Generation
von: Zhou, Yucheng, et al.
Veröffentlicht: (2025)
von: Zhou, Yucheng, et al.
Veröffentlicht: (2025)
VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effects
von: Gao, Xiangbo, et al.
Veröffentlicht: (2026)
von: Gao, Xiangbo, et al.
Veröffentlicht: (2026)
TextTIGER: Text-based Intelligent Generation with Entity Prompt Refinement for Text-to-Image Generation
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
von: Ozaki, Shintaro, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ReAlign: Bilingual Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025) -
ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
von: Weng, Wanjiang, et al.
Veröffentlicht: (2025) -
EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026) -
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026) -
SoPo: Text-to-Motion Generation Using Semi-Online Preference Optimization
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2024)