Light-T2M: A Lightweight and Fast Model for Text-to-motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Ling-An, Huang, Guohong, Wu, Gaojie, Zheng, Wei-Shi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Explicit Joint-level Interaction Modeling with Mamba for Text-guided HOI Generation
by: Huang, Guohong, et al.
Published: (2025)
by: Huang, Guohong, et al.
Published: (2025)
Progressive Human Motion Generation Based on Text and Few Motion Frames
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
FLUID: A Fine-Grained Lightweight Urban Signalized-Intersection Dataset of Dense Conflict Trajectories
by: Chen, Yiyang, et al.
Published: (2025)
by: Chen, Yiyang, et al.
Published: (2025)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
by: Li, Yuan-Ming, et al.
Published: (2025)
by: Li, Yuan-Ming, et al.
Published: (2025)
MotionHiFlow: Text-to-motion via hierarchical flow matching
by: Li, Heng, et al.
Published: (2026)
by: Li, Heng, et al.
Published: (2026)
MMD-Thinker: Adaptive Multi-Dimensional Thinking for Multimodal Misinformation Detection
by: Wu, Junjie, et al.
Published: (2025)
by: Wu, Junjie, et al.
Published: (2025)
Light4GS: Lightweight Compact 4D Gaussian Splatting Generation via Context Model
by: Liu, Mufan, et al.
Published: (2025)
by: Liu, Mufan, et al.
Published: (2025)
DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation
by: Yan, Junkai, et al.
Published: (2024)
by: Yan, Junkai, et al.
Published: (2024)
Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models
by: Wu, Sihao, et al.
Published: (2025)
by: Wu, Sihao, et al.
Published: (2025)
Instilling Multi-round Thinking to Text-guided Image Generation
by: Zeng, Lidong, et al.
Published: (2024)
by: Zeng, Lidong, et al.
Published: (2024)
FastLightGen: Fast and Light Video Generation with Fewer Steps and Parameters
by: Shao, Shitong, et al.
Published: (2026)
by: Shao, Shitong, et al.
Published: (2026)
Continual Action Assessment via Task-Consistent Score-Discriminative Feature Distribution Modeling
by: Li, Yuan-Ming, et al.
Published: (2023)
by: Li, Yuan-Ming, et al.
Published: (2023)
TPA3D: Triplane Attention for Fast Text-to-3D Generation
by: Wu, Bin-Shih, et al.
Published: (2023)
by: Wu, Bin-Shih, et al.
Published: (2023)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
LightAVSeg: Lightweight Audio-Visual Segmentation
by: Zhong, Qing, et al.
Published: (2026)
by: Zhong, Qing, et al.
Published: (2026)
A Lightweight Real-Time Low-Light Enhancement Network for Embedded Automotive Vision Systems
by: Chen, Yuhan, et al.
Published: (2025)
by: Chen, Yuhan, et al.
Published: (2025)
Fast Prompt Alignment for Text-to-Image Generation
by: Mrini, Khalil, et al.
Published: (2024)
by: Mrini, Khalil, et al.
Published: (2024)
TextBlockV2: Towards Precise-Detection-Free Scene Text Spotting with Pre-trained Language Model
by: Lyu, Jiahao, et al.
Published: (2024)
by: Lyu, Jiahao, et al.
Published: (2024)
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
by: Wu, Xun, et al.
Published: (2024)
by: Wu, Xun, et al.
Published: (2024)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
by: Shi, Tiandong, et al.
Published: (2026)
by: Shi, Tiandong, et al.
Published: (2026)
OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
by: Lin, Gaojie, et al.
Published: (2025)
by: Lin, Gaojie, et al.
Published: (2025)
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
by: Hsiao, Teng-Fang, et al.
Published: (2025)
by: Hsiao, Teng-Fang, et al.
Published: (2025)
Chest-Diffusion: A Light-Weight Text-to-Image Model for Report-to-CXR Generation
by: Huang, Peng, et al.
Published: (2024)
by: Huang, Peng, et al.
Published: (2024)
Global Modeling Matters: A Fast, Lightweight and Effective Baseline for Efficient Image Restoration
by: Jiang, Xingyu, et al.
Published: (2025)
by: Jiang, Xingyu, et al.
Published: (2025)
Structure Observation Driven Image-Text Contrastive Learning for Computed Tomography Report Generation
by: Liu, Hong, et al.
Published: (2026)
by: Liu, Hong, et al.
Published: (2026)
Optimising Event-Driven Spiking Neural Network with Regularisation and Cutoff
by: Wu, Dengyu, et al.
Published: (2023)
by: Wu, Dengyu, et al.
Published: (2023)
PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation
by: Lei, Nan, et al.
Published: (2026)
by: Lei, Nan, et al.
Published: (2026)
Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation
by: Liu, Yifei, et al.
Published: (2026)
by: Liu, Yifei, et al.
Published: (2026)
Learning Visual Generative Priors without Text
by: Ma, Shuailei, et al.
Published: (2024)
by: Ma, Shuailei, et al.
Published: (2024)
M2D2M: Multi-Motion Generation from Text with Discrete Diffusion Models
by: Chi, Seunggeun, et al.
Published: (2024)
by: Chi, Seunggeun, et al.
Published: (2024)
TurboFill: Adapting Few-step Text-to-image Model for Fast Image Inpainting
by: Xie, Liangbin, et al.
Published: (2025)
by: Xie, Liangbin, et al.
Published: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
by: Shi, Junyu, et al.
Published: (2025)
by: Shi, Junyu, et al.
Published: (2025)
DreamLite: A Lightweight On-Device Unified Model for Image Generation and Editing
by: Feng, Kailai, et al.
Published: (2026)
by: Feng, Kailai, et al.
Published: (2026)
MVLight: Relightable Text-to-3D Generation via Light-conditioned Multi-View Diffusion
by: Shim, Dongseok, et al.
Published: (2024)
by: Shim, Dongseok, et al.
Published: (2024)
Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
by: Wu, Sihao, et al.
Published: (2025)
by: Wu, Sihao, et al.
Published: (2025)
Insight-A: Attribution-aware for Multimodal Misinformation Detection
by: Wu, Junjie, et al.
Published: (2025)
by: Wu, Junjie, et al.
Published: (2025)
Generative Texture Filtering
by: Zheng, Rongjia, et al.
Published: (2026)
by: Zheng, Rongjia, et al.
Published: (2026)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
by: Zhao, Lin, et al.
Published: (2024)
by: Zhao, Lin, et al.
Published: (2024)
Similar Items
-
Efficient Explicit Joint-level Interaction Modeling with Mamba for Text-guided HOI Generation
by: Huang, Guohong, et al.
Published: (2025) -
Progressive Human Motion Generation Based on Text and Few Motion Frames
by: Zeng, Ling-An, et al.
Published: (2025) -
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
by: Zeng, Ling-An, et al.
Published: (2025) -
FLUID: A Fine-Grained Lightweight Urban Signalized-Intersection Dataset of Dense Conflict Trajectories
by: Chen, Yiyang, et al.
Published: (2025) -
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
by: Li, Yuan-Ming, et al.
Published: (2025)