MoST: Multi-modality Scene Tokenization for Motion Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mu, Norman, Ji, Jingwei, Yang, Zhenpei, Harada, Nate, Tang, Haotian, Chen, Kan, Qi, Charles R., Ge, Runzhou, Goel, Kratarth, Yang, Zoey, Ettinger, Scott, Al-Rfou, Rami, Anguelov, Dragomir, Zhou, Yin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Scaling Motion Forecasting Models with Ensemble Distillation
von: Ettinger, Scott, et al.
Veröffentlicht: (2024)
von: Ettinger, Scott, et al.
Veröffentlicht: (2024)
WOMD-LiDAR: Raw Sensor Dataset Benchmark for Motion Forecasting
von: Chen, Kan, et al.
Veröffentlicht: (2023)
von: Chen, Kan, et al.
Veröffentlicht: (2023)
Scaling Laws of Motion Forecasting and Planning -- Technical Report
von: Baniodeh, Mustafa, et al.
Veröffentlicht: (2025)
von: Baniodeh, Mustafa, et al.
Veröffentlicht: (2025)
MoST: Motion Style Transformer between Diverse Action Contents
von: Kim, Boeun, et al.
Veröffentlicht: (2024)
von: Kim, Boeun, et al.
Veröffentlicht: (2024)
MoST: Mixing Speech and Text with Modality-Aware Mixture of Experts
von: Lou, Yuxuan, et al.
Veröffentlicht: (2026)
von: Lou, Yuxuan, et al.
Veröffentlicht: (2026)
Direct Post-Training Preference Alignment for Multi-Agent Motion Generation Models Using Implicit Feedback from Pre-training Demonstrations
von: Tian, Ran, et al.
Veröffentlicht: (2025)
von: Tian, Ran, et al.
Veröffentlicht: (2025)
MoST: Efficient Monarch Sparse Tuning for 3D Representation Learning
von: Han, Xu, et al.
Veröffentlicht: (2025)
von: Han, Xu, et al.
Veröffentlicht: (2025)
SceneCrafter: Controllable Multi-View Driving Scene Editing
von: Zhu, Zehao, et al.
Veröffentlicht: (2025)
von: Zhu, Zehao, et al.
Veröffentlicht: (2025)
MoST: Measuring of Sustainable Throughput method for Large Language Models as a Service
von: Anonymous
Veröffentlicht: (2026)
von: Anonymous
Veröffentlicht: (2026)
Enhanced Motion Forecasting with Plug-and-Play Multimodal Large Language Models
von: Luo, Katie, et al.
Veröffentlicht: (2025)
von: Luo, Katie, et al.
Veröffentlicht: (2025)
Collective Electrostatics and Band Alignment in Janus MoSTe nanotubes
von: Sadanandan, Adithya, et al.
Veröffentlicht: (2026)
von: Sadanandan, Adithya, et al.
Veröffentlicht: (2026)
Scene Reconstruction as Mapping Priors for 3D Detection
von: Fu, Yang, et al.
Veröffentlicht: (2026)
von: Fu, Yang, et al.
Veröffentlicht: (2026)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
von: Leng, Zhaoqi, et al.
Veröffentlicht: (2024)
von: Leng, Zhaoqi, et al.
Veröffentlicht: (2024)
State‐owned Enterprises and the Politics of Financializing Infrastructure Development in Indonesia: De‐risking at the Limit?
von: Dimitar Anguelov
Veröffentlicht: (2024)
von: Dimitar Anguelov
Veröffentlicht: (2024)
SceneDiffuser++: City-Scale Traffic Simulation via a Generative World Model
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
Drive&Gen: Co-Evaluating End-to-End Driving and Video Generation Models
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
LET-3D-AP: Longitudinal Error Tolerant 3D Average Precision for Camera-Only 3D Detection
von: Hung, Wei-Chih, et al.
Veröffentlicht: (2022)
von: Hung, Wei-Chih, et al.
Veröffentlicht: (2022)
Hypergraph-based Motion Generation with Multi-modal Interaction Relational Reasoning
von: Wu, Keshu, et al.
Veröffentlicht: (2024)
von: Wu, Keshu, et al.
Veröffentlicht: (2024)
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
von: Ghosh, Anindita, et al.
Veröffentlicht: (2026)
von: Ghosh, Anindita, et al.
Veröffentlicht: (2026)
Generative Motion Infilling From Imprecisely Timed Keyframes
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
von: Perozzi, Bryan, et al.
Veröffentlicht: (2024)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
von: Lei, Jiahui, et al.
Veröffentlicht: (2025)
von: Lei, Jiahui, et al.
Veröffentlicht: (2025)
Stationary states of aggregation-diffusion equations with compactly supported attraction kernels: radial symmetry and mass-independent boundedness
von: Anguelov, Roumen, et al.
Veröffentlicht: (2024)
von: Anguelov, Roumen, et al.
Veröffentlicht: (2024)
PaMoSplat: Part-Aware Motion-Guided Gaussian Splatting for Dynamic Scene Reconstruction
von: Deng, Yinan, et al.
Veröffentlicht: (2026)
von: Deng, Yinan, et al.
Veröffentlicht: (2026)
Tokenizing Motion: A Generative Approach for Scene Dynamics Compression
von: Yin, Shanzhi, et al.
Veröffentlicht: (2024)
von: Yin, Shanzhi, et al.
Veröffentlicht: (2024)
Enterprise Architecture as a Dynamic Capability for Scalable and Sustainable Generative AI adoption: Bridging Innovation and Governance in Large Organisations
von: Ettinger, Alexander
Veröffentlicht: (2025)
von: Ettinger, Alexander
Veröffentlicht: (2025)
Diseño desde el ser humano. Richard Neutra y su proyecto para América Latina
von: Catherine Ettinger
Veröffentlicht: (2018)
von: Catherine Ettinger
Veröffentlicht: (2018)
Explaining Canada's Unsurprising Response to Russia's Invasion of Ukraine, 2022-2023
von: Aaron Ettinger
Veröffentlicht: (2023)
von: Aaron Ettinger
Veröffentlicht: (2023)
Classical localization problem: a survey
von: Zhou, Zoey
Veröffentlicht: (2025)
von: Zhou, Zoey
Veröffentlicht: (2025)
When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models
von: Li, Yanhong, et al.
Veröffentlicht: (2024)
von: Li, Yanhong, et al.
Veröffentlicht: (2024)
Any 3D Scene is Worth 1K Tokens: 3D-Grounded Representation for Scene Generation at Scale
von: Wei, Dongxu, et al.
Veröffentlicht: (2026)
von: Wei, Dongxu, et al.
Veröffentlicht: (2026)
Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
SceneDiffuser: Efficient and Controllable Driving Simulation Initialization and Rollout
von: Jiang, Chiyu Max, et al.
Veröffentlicht: (2024)
von: Jiang, Chiyu Max, et al.
Veröffentlicht: (2024)
MoAngelo: Motion-Aware Neural Surface Reconstruction for Dynamic Scenes
von: Ebbed, Mohamed, et al.
Veröffentlicht: (2025)
von: Ebbed, Mohamed, et al.
Veröffentlicht: (2025)
SceneProp: Combining Neural Network and Markov Random Field for Scene-Graph Grounding
von: Otani, Keita, et al.
Veröffentlicht: (2025)
von: Otani, Keita, et al.
Veröffentlicht: (2025)
DiffMoE: Dynamic Token Selection for Scalable Diffusion Transformers
von: Shi, Minglei, et al.
Veröffentlicht: (2025)
von: Shi, Minglei, et al.
Veröffentlicht: (2025)
S4-Driver: Scalable Self-Supervised Driving Multimodal Large Language Modelwith Spatio-Temporal Visual Representation
von: Xie, Yichen, et al.
Veröffentlicht: (2025)
von: Xie, Yichen, et al.
Veröffentlicht: (2025)
RoleMotion: A Large-Scale Dataset towards Robust Scene-Specific Role-Playing Motion Synthesis with Fine-grained Descriptions
von: Peng, Junran, et al.
Veröffentlicht: (2025)
von: Peng, Junran, et al.
Veröffentlicht: (2025)
Skeleton‐Parted 3D Human Motion Prediction Using ST ‐ GC and Attention Mechanisms
von: Shubin Yang, et al.
Veröffentlicht: (2026)
von: Shubin Yang, et al.
Veröffentlicht: (2026)
Images of Order. Descriptions of Domestic Architecture in Mission Era California
von: Catherine R. Ettinger
Veröffentlicht: (2007)
von: Catherine R. Ettinger
Veröffentlicht: (2007)
Ähnliche Einträge
-
Scaling Motion Forecasting Models with Ensemble Distillation
von: Ettinger, Scott, et al.
Veröffentlicht: (2024) -
WOMD-LiDAR: Raw Sensor Dataset Benchmark for Motion Forecasting
von: Chen, Kan, et al.
Veröffentlicht: (2023) -
Scaling Laws of Motion Forecasting and Planning -- Technical Report
von: Baniodeh, Mustafa, et al.
Veröffentlicht: (2025) -
MoST: Motion Style Transformer between Diverse Action Contents
von: Kim, Boeun, et al.
Veröffentlicht: (2024) -
MoST: Mixing Speech and Text with Modality-Aware Mixture of Experts
von: Lou, Yuxuan, et al.
Veröffentlicht: (2026)