MoFM: A Large-Scale Human Motion Foundation Model
Fuente:
arXiv
Saved in:
| Main Authors: | Baharani, Mohammadreza, Noghre, Ghazal Alinezhad, Pazho, Armin Danesh, Maldonado, Gabriel, Tabkhi, Hamed |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
by: Maldonado, Gabriel, et al.
Published: (2025)
by: Maldonado, Gabriel, et al.
Published: (2025)
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
by: Noghre, Ghazal Alinezhad, et al.
Published: (2025)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2025)
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024)
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
by: Yao, Shanle, et al.
Published: (2025)
by: Yao, Shanle, et al.
Published: (2025)
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
by: Pazho, Armin Danesh, et al.
Published: (2023)
by: Pazho, Armin Danesh, et al.
Published: (2023)
Evaluating the Effectiveness of Video Anomaly Detection in the Wild: Online Learning and Inference for Real-world Deployment
by: Yao, Shanle, et al.
Published: (2024)
by: Yao, Shanle, et al.
Published: (2024)
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
by: Maldonado, Gabriel, et al.
Published: (2025)
by: Maldonado, Gabriel, et al.
Published: (2025)
Shopformer: Transformer-Based Framework for Detecting Shoplifting via Human Pose
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
by: Rashvand, Narges, et al.
Published: (2025)
by: Rashvand, Narges, et al.
Published: (2025)
Towards Adaptive Human-centric Video Anomaly Detection: A Comprehensive Framework and A New Benchmark
by: Pazho, Armin Danesh, et al.
Published: (2024)
by: Pazho, Armin Danesh, et al.
Published: (2024)
From Lab to Field: Real-World Evaluation of an AI-Driven Smart Video Solution to Enhance Community Safety
by: Yao, Shanle, et al.
Published: (2023)
by: Yao, Shanle, et al.
Published: (2023)
Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild
by: Yao, Shanle, et al.
Published: (2026)
by: Yao, Shanle, et al.
Published: (2026)
From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection
by: Rashvand, Narges, et al.
Published: (2026)
by: Rashvand, Narges, et al.
Published: (2026)
Scaling Large Motion Models with Million-Level Human Motions
by: Wang, Ye, et al.
Published: (2024)
by: Wang, Ye, et al.
Published: (2024)
Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography
by: Haghighi, Tania, et al.
Published: (2026)
by: Haghighi, Tania, et al.
Published: (2026)
ScaMo: Exploring the Scaling Law in Autoregressive Motion Generation Model
by: Lu, Shunlin, et al.
Published: (2024)
by: Lu, Shunlin, et al.
Published: (2024)
TrajFM: A Vehicle Trajectory Foundation Model for Region and Task Transferability
by: Lin, Yan, et al.
Published: (2024)
by: Lin, Yan, et al.
Published: (2024)
From Offline to Periodic Adaptation for Pose-Based Shoplifting Detection in Real-world Retail Security
by: Yao, Shanle, et al.
Published: (2026)
by: Yao, Shanle, et al.
Published: (2026)
OCT-SelfNet: A Self-Supervised Framework with Multi-Modal Datasets for Generalized and Robust Retinal Disease Detection
by: Jannat, Fatema-E, et al.
Published: (2024)
by: Jannat, Fatema-E, et al.
Published: (2024)
SimpliHuMoN: Simplifying Human Motion Prediction
by: Agrawal, Aadya, et al.
Published: (2026)
by: Agrawal, Aadya, et al.
Published: (2026)
SHRUG-FM: Reliability-Aware Foundation Models for Earth Observation
by: Gonzalez-Calabuig, Maria, et al.
Published: (2025)
by: Gonzalez-Calabuig, Maria, et al.
Published: (2025)
AgriFM: A Multi-source Temporal Remote Sensing Foundation Model for Agriculture Mapping
by: Li, Wenyuan, et al.
Published: (2025)
by: Li, Wenyuan, et al.
Published: (2025)
Towards Large-Scale Training of Pathology Foundation Models
by: ai, kaiko., et al.
Published: (2024)
by: ai, kaiko., et al.
Published: (2024)
FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models
by: Zhang, Zhikai, et al.
Published: (2024)
by: Zhang, Zhikai, et al.
Published: (2024)
ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation
by: Hwang, Inwoo, et al.
Published: (2026)
by: Hwang, Inwoo, et al.
Published: (2026)
InkFM: A Foundational Model for Full-Page Online Handwritten Note Understanding
by: Fadeeva, Anastasiia, et al.
Published: (2025)
by: Fadeeva, Anastasiia, et al.
Published: (2025)
Temporal Visual Semantics-Induced Human Motion Understanding with Large Language Models
by: Xing, Zheng, et al.
Published: (2025)
by: Xing, Zheng, et al.
Published: (2025)
RoMo: A Large-Scale, Richly Organized Dataset and Semantic Taxonomy for Human Motion Generation
by: Zhang, Jiahao, et al.
Published: (2026)
by: Zhang, Jiahao, et al.
Published: (2026)
EchoFM: Foundation Model for Generalizable Echocardiogram Analysis
by: Kim, Sekeun, et al.
Published: (2024)
by: Kim, Sekeun, et al.
Published: (2024)
UniMoGen: Universal Motion Generation
by: Khani, Aliasghar, et al.
Published: (2025)
by: Khani, Aliasghar, et al.
Published: (2025)
EdgeVTP: Exploration of Latency-efficient Trajectory Prediction for Edge-based Embedded Vision Applications
by: Kim, Seungjin, et al.
Published: (2026)
by: Kim, Seungjin, et al.
Published: (2026)
Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders
by: Jiang, Yitong, et al.
Published: (2026)
by: Jiang, Yitong, et al.
Published: (2026)
Hard Cases Detection in Motion Prediction by Vision-Language Foundation Models
by: Yang, Yi, et al.
Published: (2024)
by: Yang, Yi, et al.
Published: (2024)
SeaMo: A Season-Aware Multimodal Foundation Model for Remote Sensing
by: Li, Xuyang, et al.
Published: (2024)
by: Li, Xuyang, et al.
Published: (2024)
EVLF-FM: Explainable Vision Language Foundation Model for Medicine
by: Bai, Yang, et al.
Published: (2025)
by: Bai, Yang, et al.
Published: (2025)
MedFM-Robust: Benchmarking Robustness of Medical Foundation Models
by: Cui, Xiangxiang, et al.
Published: (2026)
by: Cui, Xiangxiang, et al.
Published: (2026)
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
by: Yadav, Tanush, et al.
Published: (2026)
by: Yadav, Tanush, et al.
Published: (2026)
Training Video Foundation Models with NVIDIA NeMo
by: Patel, Zeeshan, et al.
Published: (2025)
by: Patel, Zeeshan, et al.
Published: (2025)
CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models
by: Dai, Wei, et al.
Published: (2025)
by: Dai, Wei, et al.
Published: (2025)
Similar Items
-
MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
by: Maldonado, Gabriel, et al.
Published: (2025) -
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024) -
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
by: Noghre, Ghazal Alinezhad, et al.
Published: (2025) -
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
by: Noghre, Ghazal Alinezhad, et al.
Published: (2024) -
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
by: Yao, Shanle, et al.
Published: (2025)