FMPose3D: monocular 3D pose estimation via flow matching
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Ti, Yu, Xiaohang, Mathis, Mackenzie Weygandt |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PRIMA: Boosting Animal Mesh Recovery with Biological Priors and Test-Time Adaptation
by: Yu, Xiaohang, et al.
Published: (2026)
by: Yu, Xiaohang, et al.
Published: (2026)
SuperAnimal pretrained pose estimation models for behavioral analysis
by: Ye, Shaokai, et al.
Published: (2022)
by: Ye, Shaokai, et al.
Published: (2022)
Spiking monocular event based 6D pose estimation for space application
by: Courtois, Jonathan, et al.
Published: (2025)
by: Courtois, Jonathan, et al.
Published: (2025)
Extended monocular 3D imaging
by: Shen, Zicheng, et al.
Published: (2025)
by: Shen, Zicheng, et al.
Published: (2025)
D4D: An RGBD diffusion model to boost monocular depth estimation
by: Papa, L., et al.
Published: (2024)
by: Papa, L., et al.
Published: (2024)
Distilling 3D distinctive local descriptors for 6D pose estimation
by: Hamza, Amir, et al.
Published: (2025)
by: Hamza, Amir, et al.
Published: (2025)
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2023)
by: Galanakis, Stathis, et al.
Published: (2023)
ProPLIKS: Probablistic 3D human body pose estimation
by: Shetty, Karthik, et al.
Published: (2024)
by: Shetty, Karthik, et al.
Published: (2024)
Flexible graph convolutional network for 3D human pose estimation
by: Shahjahan, Abu Taib Mohammed, et al.
Published: (2024)
by: Shahjahan, Abu Taib Mohammed, et al.
Published: (2024)
Adaptive graph Kolmogorov-Arnold network for 3D human pose estimation
by: Shahjahan, Abu Taib Mohammed, et al.
Published: (2025)
by: Shahjahan, Abu Taib Mohammed, et al.
Published: (2025)
Multi-hop graph transformer network for 3D human pose estimation
by: Islam, Zaedul, et al.
Published: (2024)
by: Islam, Zaedul, et al.
Published: (2024)
COSMU: Complete 3D human shape from monocular unconstrained images
by: Pesavento, Marco, et al.
Published: (2024)
by: Pesavento, Marco, et al.
Published: (2024)
Enhancing annotations for 5D apple pose estimation through 3D Gaussian Splatting (3DGS)
by: van de Ven, Robert, et al.
Published: (2025)
by: van de Ven, Robert, et al.
Published: (2025)
HUP-3D: A 3D multi-view synthetic dataset for assisted-egocentric hand-ultrasound pose estimation
by: Birlo, Manuel, et al.
Published: (2024)
by: Birlo, Manuel, et al.
Published: (2024)
Open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2023)
by: Corsetti, Jaime, et al.
Published: (2023)
Multi-person 3D pose estimation from unlabelled data
by: Rodriguez-Criado, Daniel, et al.
Published: (2022)
by: Rodriguez-Criado, Daniel, et al.
Published: (2022)
Deep learning for 3D human pose estimation and mesh recovery: A survey
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring
by: Shukla, Vandita, et al.
Published: (2026)
by: Shukla, Vandita, et al.
Published: (2026)
C3DAG: Controlled 3D Animal Generation using 3D pose guidance
by: Mishra, Sandeep, et al.
Published: (2024)
by: Mishra, Sandeep, et al.
Published: (2024)
MonoVisual3DFilter: 3D tomatoes' localisation with monocular cameras using histogram filters
by: Magalhães, Sandro Costa, et al.
Published: (2023)
by: Magalhães, Sandro Costa, et al.
Published: (2023)
High-resolution open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
Cross-view and Cross-pose Completion for 3D Human Understanding
by: Armando, Matthieu, et al.
Published: (2023)
by: Armando, Matthieu, et al.
Published: (2023)
Extending 3D body pose estimation for robotic-assistive therapies of autistic children
by: Santos, Laura, et al.
Published: (2024)
by: Santos, Laura, et al.
Published: (2024)
LLaVAction: evaluating and training multi-modal large language models for action understanding
by: Qi, Haozhe, et al.
Published: (2025)
by: Qi, Haozhe, et al.
Published: (2025)
METER: a mobile vision transformer architecture for monocular depth estimation
by: Papa, L., et al.
Published: (2024)
by: Papa, L., et al.
Published: (2024)
KBody: Towards general, robust, and aligned monocular whole-body estimation
by: Zioulis, Nikolaos, et al.
Published: (2023)
by: Zioulis, Nikolaos, et al.
Published: (2023)
Comparison of marker-less 2D image-based methods for infant pose estimation
by: Jahn, Lennart, et al.
Published: (2024)
by: Jahn, Lennart, et al.
Published: (2024)
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
by: Caraffa, Andrea, et al.
Published: (2025)
by: Caraffa, Andrea, et al.
Published: (2025)
EC-Depth: Exploring the consistency of self-supervised monocular depth estimation in challenging scenes
by: Song, Ziyang, et al.
Published: (2023)
by: Song, Ziyang, et al.
Published: (2023)
Self-localization on a 3D map by fusing global and local features from a monocular camera
by: Kikuchi, Satoshi, et al.
Published: (2025)
by: Kikuchi, Satoshi, et al.
Published: (2025)
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
by: Skorupski, Patryk, et al.
Published: (2025)
by: Skorupski, Patryk, et al.
Published: (2025)
CHIP: A multi-sensor dataset for 6D pose estimation of chairs in industrial settings
by: Nardon, Mattia, et al.
Published: (2025)
by: Nardon, Mattia, et al.
Published: (2025)
Portrait4D: Learning One-Shot 4D Head Avatar Synthesis using Synthetic Data
by: Deng, Yu, et al.
Published: (2023)
by: Deng, Yu, et al.
Published: (2023)
Hybrid bundle-adjusting 3D Gaussians for view consistent rendering with pose optimization
by: Guo, Yanan, et al.
Published: (2024)
by: Guo, Yanan, et al.
Published: (2024)
Domain adaptive pose estimation via multi-level alignment
by: Chen, Yugan, et al.
Published: (2024)
by: Chen, Yugan, et al.
Published: (2024)
Dual-Branch Graph Transformer Network for 3D Human Mesh Reconstruction from Video
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
A Cross-Dataset Study for Text-based 3D Human Motion Retrieval
by: Bensabath, Léore, et al.
Published: (2024)
by: Bensabath, Léore, et al.
Published: (2024)
Adversarially Robust Out-of-Distribution Detection Using Lyapunov-Stabilized Embeddings
by: Mirzaei, Hossein, et al.
Published: (2024)
by: Mirzaei, Hossein, et al.
Published: (2024)
Rooms from Motion: Un-posed Indoor 3D Object Detection as Localization and Mapping
by: Lazarow, Justin, et al.
Published: (2025)
by: Lazarow, Justin, et al.
Published: (2025)
MonoSOWA: Scalable monocular 3D Object detector Without human Annotations
by: Skvrna, Jan, et al.
Published: (2025)
by: Skvrna, Jan, et al.
Published: (2025)
Similar Items
-
PRIMA: Boosting Animal Mesh Recovery with Biological Priors and Test-Time Adaptation
by: Yu, Xiaohang, et al.
Published: (2026) -
SuperAnimal pretrained pose estimation models for behavioral analysis
by: Ye, Shaokai, et al.
Published: (2022) -
Spiking monocular event based 6D pose estimation for space application
by: Courtois, Jonathan, et al.
Published: (2025) -
Extended monocular 3D imaging
by: Shen, Zicheng, et al.
Published: (2025) -
D4D: An RGBD diffusion model to boost monocular depth estimation
by: Papa, L., et al.
Published: (2024)