LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zhe, Yuan, Weihao, He, Yisheng, Qiu, Lingteng, Zhu, Shenhao, Gu, Xiaodong, Shen, Weichao, Dong, Yuan, Dong, Zilong, Yang, Laurence T. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
von: Li, Zhe, et al.
Veröffentlicht: (2024)
von: Li, Zhe, et al.
Veröffentlicht: (2024)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
von: Yuan, Weihao, et al.
Veröffentlicht: (2024)
von: Yuan, Weihao, et al.
Veröffentlicht: (2024)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
von: Qiu, Lingteng, et al.
Veröffentlicht: (2024)
von: Qiu, Lingteng, et al.
Veröffentlicht: (2024)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
von: Zhu, Shenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Shenhao, et al.
Veröffentlicht: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
von: Zuo, Qi, et al.
Veröffentlicht: (2024)
LaMP-Cap: Personalized Figure Caption Generation With Multimodal Figure Profiles
von: Ng, Ho Yin 'Sam', et al.
Veröffentlicht: (2025)
von: Ng, Ho Yin 'Sam', et al.
Veröffentlicht: (2025)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
von: Wang, Xinkai, et al.
Veröffentlicht: (2026)
von: Wang, Xinkai, et al.
Veröffentlicht: (2026)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
von: He, Yisheng, et al.
Veröffentlicht: (2025)
von: He, Yisheng, et al.
Veröffentlicht: (2025)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
von: He, Yisheng, et al.
Veröffentlicht: (2024)
von: He, Yisheng, et al.
Veröffentlicht: (2024)
Large Depth Completion Model from Sparse Observations
von: Yu, Zhu, et al.
Veröffentlicht: (2026)
von: Yu, Zhu, et al.
Veröffentlicht: (2026)
LaMP: When Large Language Models Meet Personalization
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
von: Han, Xiaoguang, et al.
Veröffentlicht: (2024)
von: Han, Xiaoguang, et al.
Veröffentlicht: (2024)
HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction
von: Gu, Xiaodong, et al.
Veröffentlicht: (2024)
von: Gu, Xiaodong, et al.
Veröffentlicht: (2024)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
von: Chen, Minglin, et al.
Veröffentlicht: (2024)
von: Chen, Minglin, et al.
Veröffentlicht: (2024)
LaMP-Val: Large Language Models Empower Personalized Valuation in Auction
von: Sun, Jie, et al.
Veröffentlicht: (2024)
von: Sun, Jie, et al.
Veröffentlicht: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
von: Li, Peng, et al.
Veröffentlicht: (2025)
von: Li, Peng, et al.
Veröffentlicht: (2025)
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
von: Wu, Yushuang, et al.
Veröffentlicht: (2024)
von: Wu, Yushuang, et al.
Veröffentlicht: (2024)
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
von: Hu, Yingdong, et al.
Veröffentlicht: (2025)
von: Hu, Yingdong, et al.
Veröffentlicht: (2025)
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
GIC: Gaussian-Informed Continuum for Physical Property Identification and Simulation
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes
von: Zhao, Zhengyi, et al.
Veröffentlicht: (2024)
von: Zhao, Zhengyi, et al.
Veröffentlicht: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
von: Ye, Chongjie, et al.
Veröffentlicht: (2024)
von: Ye, Chongjie, et al.
Veröffentlicht: (2024)
Towards Fine-Grained Human Motion Video Captioning
von: Song, Guorui, et al.
Veröffentlicht: (2025)
von: Song, Guorui, et al.
Veröffentlicht: (2025)
Dense Motion Captioning
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
MoCHA: Denoising Caption Supervision for Motion-Text Retrieval
von: Warner, Nikolai, et al.
Veröffentlicht: (2026)
von: Warner, Nikolai, et al.
Veröffentlicht: (2026)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
FingerCap: Fine-grained Finger-level Hand Motion Captioning
von: Shen, Xin, et al.
Veröffentlicht: (2025)
von: Shen, Xin, et al.
Veröffentlicht: (2025)
HyPlaneHead: Rethinking Tri-plane-like Representations in Full-Head Image Synthesis
von: Li, Heyuan, et al.
Veröffentlicht: (2025)
von: Li, Heyuan, et al.
Veröffentlicht: (2025)
MotionRAG: Motion Retrieval-Augmented Image-to-Video Generation
von: Zhu, Chenhui, et al.
Veröffentlicht: (2025)
von: Zhu, Chenhui, et al.
Veröffentlicht: (2025)
Exploring Motion-Language Alignment for Text-driven Motion Generation
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
Neural MP: A Generalist Neural Motion Planner
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
PPT: Pretraining with Pseudo-Labeled Trajectories for Motion Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
von: He, Chao, et al.
Veröffentlicht: (2025)
von: He, Chao, et al.
Veröffentlicht: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
von: Fang, Haopeng, et al.
Veröffentlicht: (2024)
von: Fang, Haopeng, et al.
Veröffentlicht: (2024)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2025)
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
von: Li, Zhe, et al.
Veröffentlicht: (2024) -
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
von: Yuan, Weihao, et al.
Veröffentlicht: (2024) -
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
von: Li, Zhe, et al.
Veröffentlicht: (2025) -
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
von: Qiu, Lingteng, et al.
Veröffentlicht: (2024) -
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
von: Zhu, Shenhao, et al.
Veröffentlicht: (2024)