UMO: Unified In-Context Learning Unlocks Motion Foundation Model Priors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cong, Xiaoyan, Li, Zekun, Dou, Zhiyang, Li, Hongyu, Taheri, Omid, Guo, Chuan, Mittal, Abhay, An, Sizhe, Komura, Taku, Matusik, Wojciech, Black, Michael J., Sridhar, Srinath |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens
von: Li, Zekun, et al.
Veröffentlicht: (2026)
von: Li, Zekun, et al.
Veröffentlicht: (2026)
Soft Anisotropic Diagrams for Differentiable Image Representation
von: Iinbor, Laki, et al.
Veröffentlicht: (2026)
von: Iinbor, Laki, et al.
Veröffentlicht: (2026)
CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects
von: Pi, Huaijin, et al.
Veröffentlicht: (2025)
von: Pi, Huaijin, et al.
Veröffentlicht: (2025)
TLControl: Trajectory and Language Control for Human Motion Synthesis
von: Wan, Weilin, et al.
Veröffentlicht: (2023)
von: Wan, Weilin, et al.
Veröffentlicht: (2023)
Art3D: Training-Free 3D Generation from Flat-Colored Illustration
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
GenHSI: Controllable Generation of Human-Scene Interaction Videos
von: Li, Zekun, et al.
Veröffentlicht: (2025)
von: Li, Zekun, et al.
Veröffentlicht: (2025)
FUSION: Full-Body Unified Motion Prior for Body and Hands via Diffusion
von: Duran, Enes, et al.
Veröffentlicht: (2026)
von: Duran, Enes, et al.
Veröffentlicht: (2026)
MotionGlot: A Multi-Embodied Motion Generation Model
von: Harithas, Sudarshan, et al.
Veröffentlicht: (2024)
von: Harithas, Sudarshan, et al.
Veröffentlicht: (2024)
Motion2Motion: Cross-topology Motion Transfer with Sparse Correspondence
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2025)
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2025)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
PackUV: Packed Gaussian UV Maps for 4D Volumetric Video
von: Rai, Aashish, et al.
Veröffentlicht: (2026)
von: Rai, Aashish, et al.
Veröffentlicht: (2026)
Pay Attention and Move Better: Harnessing Attention for Interactive Motion Generation and Training-free Editing
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2024)
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2024)
EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation
von: Zhou, Wenyang, et al.
Veröffentlicht: (2023)
von: Zhou, Wenyang, et al.
Veröffentlicht: (2023)
RigidFormer: Learning Rigid Dynamics using Transformers
von: Dou, Zhiyang, et al.
Veröffentlicht: (2026)
von: Dou, Zhiyang, et al.
Veröffentlicht: (2026)
GeoPT: Scaling Physics Simulation via Lifted Geometric Pre-Training
von: Wu, Haixu, et al.
Veröffentlicht: (2026)
von: Wu, Haixu, et al.
Veröffentlicht: (2026)
DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
IAM: Identity-Aware Human Motion and Shape Joint Generation
von: Jia, Wenqi, et al.
Veröffentlicht: (2026)
von: Jia, Wenqi, et al.
Veröffentlicht: (2026)
Dynamic Realms: 4D Content Analysis, Recovery and Generation with Geometric, Topological and Physical Priors
von: Dou, Zhiyang
Veröffentlicht: (2024)
von: Dou, Zhiyang
Veröffentlicht: (2024)
MOSPA: Human Motion Generation Driven by Spatial Audio
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
WANDR: Intention-guided Human Motion Generation
von: Diomataris, Markos, et al.
Veröffentlicht: (2024)
von: Diomataris, Markos, et al.
Veröffentlicht: (2024)
Surf-D: Generating High-Quality Surfaces of Arbitrary Topologies Using Diffusion Models
von: Yu, Zhengming, et al.
Veröffentlicht: (2023)
von: Yu, Zhengming, et al.
Veröffentlicht: (2023)
HUMOS: Human Motion Model Conditioned on Body Shape
von: Tripathi, Shashank, et al.
Veröffentlicht: (2024)
von: Tripathi, Shashank, et al.
Veröffentlicht: (2024)
ViTa-Zero: Zero-shot Visuotactile Object 6D Pose Estimation
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
Moving by Looking: Towards Vision-Driven Avatar Motion Generation
von: Diomataris, Markos, et al.
Veröffentlicht: (2025)
von: Diomataris, Markos, et al.
Veröffentlicht: (2025)
GenHeld: Generating and Editing Handheld Objects
von: Min, Chaerin, et al.
Veröffentlicht: (2024)
von: Min, Chaerin, et al.
Veröffentlicht: (2024)
EgoSonics: Generating Synchronized Audio for Silent Egocentric Videos
von: Rai, Aashish, et al.
Veröffentlicht: (2024)
von: Rai, Aashish, et al.
Veröffentlicht: (2024)
RMD: A Simple Baseline for More General Human Motion Generation via Training-free Retrieval-Augmented Motion Diffuse
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
UNIC: Neural Garment Deformation Field for Real-time Clothed Character Animation
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization
von: Pan, Liang, et al.
Veröffentlicht: (2025)
von: Pan, Liang, et al.
Veröffentlicht: (2025)
CBIL: Collective Behavior Imitation Learning for Fish from Real Videos
von: Wu, Yifan, et al.
Veröffentlicht: (2025)
von: Wu, Yifan, et al.
Veröffentlicht: (2025)
EgoReAct: Egocentric Video-Driven 3D Human Reaction Generation
von: Zhang, Libo, et al.
Veröffentlicht: (2025)
von: Zhang, Libo, et al.
Veröffentlicht: (2025)
UniTac: Whole-Robot Touch Sensing Without Tactile Sensors
von: Fu, Wanjia, et al.
Veröffentlicht: (2025)
von: Fu, Wanjia, et al.
Veröffentlicht: (2025)
Foundation Molecular Grammar: Multi-Modal Foundation Models Induce Interpretable Molecular Graph Languages
von: Sun, Michael, et al.
Veröffentlicht: (2025)
von: Sun, Michael, et al.
Veröffentlicht: (2025)
Tracking-Guided 4D Generation: Foundation-Tracker Motion Priors for 3D Model Animation
von: Sun, Su, et al.
Veröffentlicht: (2025)
von: Sun, Su, et al.
Veröffentlicht: (2025)
V-HOP: Visuo-Haptic 6D Object Pose Tracking
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
SENC: Handling Self-collision in Neural Cloth Simulation
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
von: Liao, Zhouyingcheng, et al.
Veröffentlicht: (2024)
Half-Physics: Enabling Kinematic 3D Human Model with Physical Interactions
von: Siyao, Li, et al.
Veröffentlicht: (2025)
von: Siyao, Li, et al.
Veröffentlicht: (2025)
NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models
von: Albaba, Mert, et al.
Veröffentlicht: (2025)
von: Albaba, Mert, et al.
Veröffentlicht: (2025)
UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward
von: Cheng, Yufeng, et al.
Veröffentlicht: (2025)
von: Cheng, Yufeng, et al.
Veröffentlicht: (2025)
InteractVLM: 3D Interaction Reasoning from 2D Foundational Models
von: Dwivedi, Sai Kumar, et al.
Veröffentlicht: (2025)
von: Dwivedi, Sai Kumar, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens
von: Li, Zekun, et al.
Veröffentlicht: (2026) -
Soft Anisotropic Diagrams for Differentiable Image Representation
von: Iinbor, Laki, et al.
Veröffentlicht: (2026) -
CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects
von: Pi, Huaijin, et al.
Veröffentlicht: (2025) -
TLControl: Trajectory and Language Control for Human Motion Synthesis
von: Wan, Weilin, et al.
Veröffentlicht: (2023) -
Art3D: Training-Free 3D Generation from Flat-Colored Illustration
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)