Light-T2M: A Lightweight and Fast Model for Text-to-motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Ling-An, Huang, Guohong, Wu, Gaojie, Zheng, Wei-Shi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Explicit Joint-level Interaction Modeling with Mamba for Text-guided HOI Generation
von: Huang, Guohong, et al.
Veröffentlicht: (2025)
von: Huang, Guohong, et al.
Veröffentlicht: (2025)
Progressive Human Motion Generation Based on Text and Few Motion Frames
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025)
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025)
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025)
FLUID: A Fine-Grained Lightweight Urban Signalized-Intersection Dataset of Dense Conflict Trajectories
von: Chen, Yiyang, et al.
Veröffentlicht: (2025)
von: Chen, Yiyang, et al.
Veröffentlicht: (2025)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2025)
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2025)
MotionHiFlow: Text-to-motion via hierarchical flow matching
von: Li, Heng, et al.
Veröffentlicht: (2026)
von: Li, Heng, et al.
Veröffentlicht: (2026)
MMD-Thinker: Adaptive Multi-Dimensional Thinking for Multimodal Misinformation Detection
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
Light4GS: Lightweight Compact 4D Gaussian Splatting Generation via Context Model
von: Liu, Mufan, et al.
Veröffentlicht: (2025)
von: Liu, Mufan, et al.
Veröffentlicht: (2025)
DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation
von: Yan, Junkai, et al.
Veröffentlicht: (2024)
von: Yan, Junkai, et al.
Veröffentlicht: (2024)
Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
Instilling Multi-round Thinking to Text-guided Image Generation
von: Zeng, Lidong, et al.
Veröffentlicht: (2024)
von: Zeng, Lidong, et al.
Veröffentlicht: (2024)
FastLightGen: Fast and Light Video Generation with Fewer Steps and Parameters
von: Shao, Shitong, et al.
Veröffentlicht: (2026)
von: Shao, Shitong, et al.
Veröffentlicht: (2026)
Continual Action Assessment via Task-Consistent Score-Discriminative Feature Distribution Modeling
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2023)
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2023)
TPA3D: Triplane Attention for Fast Text-to-3D Generation
von: Wu, Bin-Shih, et al.
Veröffentlicht: (2023)
von: Wu, Bin-Shih, et al.
Veröffentlicht: (2023)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
LightAVSeg: Lightweight Audio-Visual Segmentation
von: Zhong, Qing, et al.
Veröffentlicht: (2026)
von: Zhong, Qing, et al.
Veröffentlicht: (2026)
A Lightweight Real-Time Low-Light Enhancement Network for Embedded Automotive Vision Systems
von: Chen, Yuhan, et al.
Veröffentlicht: (2025)
von: Chen, Yuhan, et al.
Veröffentlicht: (2025)
Fast Prompt Alignment for Text-to-Image Generation
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
von: Mrini, Khalil, et al.
Veröffentlicht: (2024)
TextBlockV2: Towards Precise-Detection-Free Scene Text Spotting with Pre-trained Language Model
von: Lyu, Jiahao, et al.
Veröffentlicht: (2024)
von: Lyu, Jiahao, et al.
Veröffentlicht: (2024)
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
von: Wu, Xun, et al.
Veröffentlicht: (2024)
von: Wu, Xun, et al.
Veröffentlicht: (2024)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
von: Jiang, Lutao, et al.
Veröffentlicht: (2024)
von: Jiang, Lutao, et al.
Veröffentlicht: (2024)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
von: Lin, Gaojie, et al.
Veröffentlicht: (2025)
von: Lin, Gaojie, et al.
Veröffentlicht: (2025)
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2025)
von: Hsiao, Teng-Fang, et al.
Veröffentlicht: (2025)
Chest-Diffusion: A Light-Weight Text-to-Image Model for Report-to-CXR Generation
von: Huang, Peng, et al.
Veröffentlicht: (2024)
von: Huang, Peng, et al.
Veröffentlicht: (2024)
Global Modeling Matters: A Fast, Lightweight and Effective Baseline for Efficient Image Restoration
von: Jiang, Xingyu, et al.
Veröffentlicht: (2025)
von: Jiang, Xingyu, et al.
Veröffentlicht: (2025)
Structure Observation Driven Image-Text Contrastive Learning for Computed Tomography Report Generation
von: Liu, Hong, et al.
Veröffentlicht: (2026)
von: Liu, Hong, et al.
Veröffentlicht: (2026)
Optimising Event-Driven Spiking Neural Network with Regularisation and Cutoff
von: Wu, Dengyu, et al.
Veröffentlicht: (2023)
von: Wu, Dengyu, et al.
Veröffentlicht: (2023)
PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation
von: Lei, Nan, et al.
Veröffentlicht: (2026)
von: Lei, Nan, et al.
Veröffentlicht: (2026)
Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation
von: Liu, Yifei, et al.
Veröffentlicht: (2026)
von: Liu, Yifei, et al.
Veröffentlicht: (2026)
Learning Visual Generative Priors without Text
von: Ma, Shuailei, et al.
Veröffentlicht: (2024)
von: Ma, Shuailei, et al.
Veröffentlicht: (2024)
M2D2M: Multi-Motion Generation from Text with Discrete Diffusion Models
von: Chi, Seunggeun, et al.
Veröffentlicht: (2024)
von: Chi, Seunggeun, et al.
Veröffentlicht: (2024)
TurboFill: Adapting Few-step Text-to-image Model for Fast Image Inpainting
von: Xie, Liangbin, et al.
Veröffentlicht: (2025)
von: Xie, Liangbin, et al.
Veröffentlicht: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
DreamLite: A Lightweight On-Device Unified Model for Image Generation and Editing
von: Feng, Kailai, et al.
Veröffentlicht: (2026)
von: Feng, Kailai, et al.
Veröffentlicht: (2026)
MVLight: Relightable Text-to-3D Generation via Light-conditioned Multi-View Diffusion
von: Shim, Dongseok, et al.
Veröffentlicht: (2024)
von: Shim, Dongseok, et al.
Veröffentlicht: (2024)
Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
von: Wu, Sihao, et al.
Veröffentlicht: (2025)
Insight-A: Attribution-aware for Multimodal Misinformation Detection
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
Generative Texture Filtering
von: Zheng, Rongjia, et al.
Veröffentlicht: (2026)
von: Zheng, Rongjia, et al.
Veröffentlicht: (2026)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
von: Zhao, Lin, et al.
Veröffentlicht: (2024)
von: Zhao, Lin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Efficient Explicit Joint-level Interaction Modeling with Mamba for Text-guided HOI Generation
von: Huang, Guohong, et al.
Veröffentlicht: (2025) -
Progressive Human Motion Generation Based on Text and Few Motion Frames
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025) -
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
von: Zeng, Ling-An, et al.
Veröffentlicht: (2025) -
FLUID: A Fine-Grained Lightweight Urban Signalized-Intersection Dataset of Dense Conflict Trajectories
von: Chen, Yiyang, et al.
Veröffentlicht: (2025) -
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2025)