GENMO: A GENeralist Model for Human MOtion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jiefeng, Cao, Jinkun, Zhang, Haotian, Rempe, Davis, Kautz, Jan, Iqbal, Umar, Yuan, Ye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
COIN: Control-Inpainting Diffusion Prior for Human and Camera Motion Estimation
von: Li, Jiefeng, et al.
Veröffentlicht: (2024)
von: Li, Jiefeng, et al.
Veröffentlicht: (2024)
Kimodo: Scaling Controllable Human Motion Generation
von: Rempe, Davis, et al.
Veröffentlicht: (2026)
von: Rempe, Davis, et al.
Veröffentlicht: (2026)
SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control
von: Luo, Zhengyi, et al.
Veröffentlicht: (2025)
von: Luo, Zhengyi, et al.
Veröffentlicht: (2025)
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
von: Petrovich, Mathis, et al.
Veröffentlicht: (2024)
von: Petrovich, Mathis, et al.
Veröffentlicht: (2024)
WHAC: World-grounded Humans and Cameras
von: Yin, Wanqi, et al.
Veröffentlicht: (2024)
von: Yin, Wanqi, et al.
Veröffentlicht: (2024)
AddBiomechanics Dataset: Capturing the Physics of Human Motion at Scale
von: Werling, Keenon, et al.
Veröffentlicht: (2024)
von: Werling, Keenon, et al.
Veröffentlicht: (2024)
ImDy: Human Inverse Dynamics from Imitated Observations
von: Liu, Xinpeng, et al.
Veröffentlicht: (2024)
von: Liu, Xinpeng, et al.
Veröffentlicht: (2024)
GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning
von: Yuan, Ye, et al.
Veröffentlicht: (2023)
von: Yuan, Ye, et al.
Veröffentlicht: (2023)
SMPLOlympics: Sports Environments for Physically Simulated Humanoids
von: Luo, Zhengyi, et al.
Veröffentlicht: (2024)
von: Luo, Zhengyi, et al.
Veröffentlicht: (2024)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
von: Yang, Jianing, et al.
Veröffentlicht: (2025)
von: Yang, Jianing, et al.
Veröffentlicht: (2025)
SimAvatar: Simulation-Ready Avatars with Layered Hair and Clothing
von: Li, Xueting, et al.
Veröffentlicht: (2024)
von: Li, Xueting, et al.
Veröffentlicht: (2024)
Structure from Collision
von: Kaneko, Takuhiro
Veröffentlicht: (2025)
von: Kaneko, Takuhiro
Veröffentlicht: (2025)
Learning 3D-Gaussian Simulators from RGB Videos
von: Zhobro, Mikel, et al.
Veröffentlicht: (2025)
von: Zhobro, Mikel, et al.
Veröffentlicht: (2025)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
von: Zheng, Shuhong, et al.
Veröffentlicht: (2025)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2025)
IFG: Internet-Scale Guidance for Functional Grasping Generation
von: Liu, Ray Muxin, et al.
Veröffentlicht: (2025)
von: Liu, Ray Muxin, et al.
Veröffentlicht: (2025)
RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
von: Wei, Shenxing, et al.
Veröffentlicht: (2025)
Improving Physics-Augmented Continuum Neural Radiance Field-Based Geometry-Agnostic System Identification with Lagrangian Particle Optimization
von: Kaneko, Takuhiro
Veröffentlicht: (2024)
von: Kaneko, Takuhiro
Veröffentlicht: (2024)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
von: Zheng, Shuhong, et al.
Veröffentlicht: (2026)
RigidFormer: Learning Rigid Dynamics using Transformers
von: Dou, Zhiyang, et al.
Veröffentlicht: (2026)
von: Dou, Zhiyang, et al.
Veröffentlicht: (2026)
DyST: Towards Dynamic Neural Scene Representations on Real-World Videos
von: Seitzer, Maximilian, et al.
Veröffentlicht: (2023)
von: Seitzer, Maximilian, et al.
Veröffentlicht: (2023)
MoRight: Motion Control Done Right
von: Liu, Shaowei, et al.
Veröffentlicht: (2026)
von: Liu, Shaowei, et al.
Veröffentlicht: (2026)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
von: Ye, Weicai, et al.
Veröffentlicht: (2024)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
von: Ray, Arijit, et al.
Veröffentlicht: (2024)
Leveraging Foundation Models To learn the shape of semi-fluid deformable objects
von: Assal, Omar El, et al.
Veröffentlicht: (2024)
von: Assal, Omar El, et al.
Veröffentlicht: (2024)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
von: Shah, Ansh, et al.
Veröffentlicht: (2024)
von: Shah, Ansh, et al.
Veröffentlicht: (2024)
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
von: Trevithick, Alex, et al.
Veröffentlicht: (2024)
von: Trevithick, Alex, et al.
Veröffentlicht: (2024)
Generalizable Vision-Language Few-Shot Adaptation with Predictive Prompts and Negative Learning
von: Mandalika, Sriram
Veröffentlicht: (2025)
von: Mandalika, Sriram
Veröffentlicht: (2025)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
von: Pham, Phu, et al.
Veröffentlicht: (2024)
von: Pham, Phu, et al.
Veröffentlicht: (2024)
Physics-Based Motion Imitation with Adversarial Differential Discriminators
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
Free-Moving Object Reconstruction and Pose Estimation with Virtual Camera
von: Shi, Haixin, et al.
Veröffentlicht: (2024)
von: Shi, Haixin, et al.
Veröffentlicht: (2024)
Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
von: Dong, Zhao, et al.
Veröffentlicht: (2025)
von: Dong, Zhao, et al.
Veröffentlicht: (2025)
SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2026)
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2026)
Neural Implicit Representation for Building Digital Twins of Unknown Articulated Objects
von: Weng, Yijia, et al.
Veröffentlicht: (2024)
von: Weng, Yijia, et al.
Veröffentlicht: (2024)
VibraVerse: A Large-Scale Geometry-Acoustics Alignment Dataset for Physically-Consistent Multimodal Learning
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control
von: Mu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Mu, Yuxuan, et al.
Veröffentlicht: (2025)
Universal Humanoid Motion Representations for Physics-Based Control
von: Luo, Zhengyi, et al.
Veröffentlicht: (2023)
von: Luo, Zhengyi, et al.
Veröffentlicht: (2023)
Real-Time Simulated Avatar from Head-Mounted Sensors
von: Luo, Zhengyi, et al.
Veröffentlicht: (2024)
von: Luo, Zhengyi, et al.
Veröffentlicht: (2024)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
von: Xie, Kevin, et al.
Veröffentlicht: (2025)
von: Xie, Kevin, et al.
Veröffentlicht: (2025)
MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives
von: Wang, Tingwu, et al.
Veröffentlicht: (2026)
von: Wang, Tingwu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
COIN: Control-Inpainting Diffusion Prior for Human and Camera Motion Estimation
von: Li, Jiefeng, et al.
Veröffentlicht: (2024) -
Kimodo: Scaling Controllable Human Motion Generation
von: Rempe, Davis, et al.
Veröffentlicht: (2026) -
SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control
von: Luo, Zhengyi, et al.
Veröffentlicht: (2025) -
Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation
von: Petrovich, Mathis, et al.
Veröffentlicht: (2024) -
WHAC: World-grounded Humans and Cameras
von: Yin, Wanqi, et al.
Veröffentlicht: (2024)