KBody: Towards general, robust, and aligned monocular whole-body estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Zioulis, Nikolaos, O'Brien, James F. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Practical Single-shot Motion Synthesis
by: Roditakis, Konstantinos, et al.
Published: (2024)
by: Roditakis, Konstantinos, et al.
Published: (2024)
On the Skinning of Gaussian Avatars
by: Zioulis, Nikolaos, et al.
Published: (2025)
by: Zioulis, Nikolaos, et al.
Published: (2025)
METER: a mobile vision transformer architecture for monocular depth estimation
by: Papa, L., et al.
Published: (2024)
by: Papa, L., et al.
Published: (2024)
D4D: An RGBD diffusion model to boost monocular depth estimation
by: Papa, L., et al.
Published: (2024)
by: Papa, L., et al.
Published: (2024)
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
by: Skorupski, Patryk, et al.
Published: (2025)
by: Skorupski, Patryk, et al.
Published: (2025)
EC-Depth: Exploring the consistency of self-supervised monocular depth estimation in challenging scenes
by: Song, Ziyang, et al.
Published: (2023)
by: Song, Ziyang, et al.
Published: (2023)
FMPose3D: monocular 3D pose estimation via flow matching
by: Wang, Ti, et al.
Published: (2026)
by: Wang, Ti, et al.
Published: (2026)
Towards aligned body representations in vision models
by: Gizdov, Andrey, et al.
Published: (2025)
by: Gizdov, Andrey, et al.
Published: (2025)
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
by: Galanakis, Stathis, et al.
Published: (2023)
by: Galanakis, Stathis, et al.
Published: (2023)
Spiking monocular event based 6D pose estimation for space application
by: Courtois, Jonathan, et al.
Published: (2025)
by: Courtois, Jonathan, et al.
Published: (2025)
Extended monocular 3D imaging
by: Shen, Zicheng, et al.
Published: (2025)
by: Shen, Zicheng, et al.
Published: (2025)
Generalizing monocular colonoscopy image depth estimation by uncertainty-based global and local fusion network
by: Du, Sijia, et al.
Published: (2024)
by: Du, Sijia, et al.
Published: (2024)
PyMAF-X: Towards Well-aligned Full-body Model Regression from Monocular Images
by: Zhang, Hongwen, et al.
Published: (2022)
by: Zhang, Hongwen, et al.
Published: (2022)
COSMU: Complete 3D human shape from monocular unconstrained images
by: Pesavento, Marco, et al.
Published: (2024)
by: Pesavento, Marco, et al.
Published: (2024)
DRSM: efficient neural 4d decomposition for dynamic reconstruction in stationary monocular cameras
by: Xie, Weixing, et al.
Published: (2024)
by: Xie, Weixing, et al.
Published: (2024)
On the dynamic evolution of CLIP texture-shape bias and its relationship to human alignment and model robustness
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring
by: Shukla, Vandita, et al.
Published: (2026)
by: Shukla, Vandita, et al.
Published: (2026)
Domain adaptive pose estimation via multi-level alignment
by: Chen, Yugan, et al.
Published: (2024)
by: Chen, Yugan, et al.
Published: (2024)
Forest canopy height estimation from satellite RGB imagery using large-scale airborne LiDAR-derived training data and monocular depth estimation
by: Lai, Yongkang, et al.
Published: (2026)
by: Lai, Yongkang, et al.
Published: (2026)
Towards Scalable Human-aligned Benchmark for Text-guided Image Editing
by: Ryu, Suho, et al.
Published: (2025)
by: Ryu, Suho, et al.
Published: (2025)
MonoVisual3DFilter: 3D tomatoes' localisation with monocular cameras using histogram filters
by: Magalhães, Sandro Costa, et al.
Published: (2023)
by: Magalhães, Sandro Costa, et al.
Published: (2023)
T-LEAP: Occlusion-robust pose estimation of walking cows using temporal information
by: Russello, Helena, et al.
Published: (2021)
by: Russello, Helena, et al.
Published: (2021)
Self-localization on a 3D map by fusing global and local features from a monocular camera
by: Kikuchi, Satoshi, et al.
Published: (2025)
by: Kikuchi, Satoshi, et al.
Published: (2025)
A method for tissue-mask supported whole-body image registration in the UK Biobank
by: Utkueri, Yasemin, et al.
Published: (2025)
by: Utkueri, Yasemin, et al.
Published: (2025)
From Bias to Balance: Detecting Facial Expression Recognition Biases in Large Multimodal Foundation Models
by: Chhua, Kaylee, et al.
Published: (2024)
by: Chhua, Kaylee, et al.
Published: (2024)
ProPLIKS: Probablistic 3D human body pose estimation
by: Shetty, Karthik, et al.
Published: (2024)
by: Shetty, Karthik, et al.
Published: (2024)
An extremely coarse feedback signal is sufficient for learning human-aligned visual representations
by: Mehta, Yash, et al.
Published: (2026)
by: Mehta, Yash, et al.
Published: (2026)
FedPartWhole: Federated domain generalization via consistent part-whole hierarchies
by: Radwan, Ahmed, et al.
Published: (2024)
by: Radwan, Ahmed, et al.
Published: (2024)
Deconstructing Bias: A Multifaceted Framework for Diagnosing Cultural and Compositional Inequities in Text-to-Image Generative Models
by: Said, Muna Numan, et al.
Published: (2025)
by: Said, Muna Numan, et al.
Published: (2025)
Advancing Monocular Video-Based Gait Analysis Using Motion Imitation with Physics-Based Simulation
by: Smyrnakis, Nikolaos, et al.
Published: (2024)
by: Smyrnakis, Nikolaos, et al.
Published: (2024)
BeatFormer: Efficient motion-robust remote heart rate estimation through unsupervised spectral zoomed attention filters
by: Comas, Joaquim, et al.
Published: (2025)
by: Comas, Joaquim, et al.
Published: (2025)
Self-supervised video pretraining yields robust and more human-aligned visual representations
by: Parthasarathy, Nikhil, et al.
Published: (2022)
by: Parthasarathy, Nikhil, et al.
Published: (2022)
Prompt-aligned Gradient for Prompt Tuning
by: Zhu, Beier, et al.
Published: (2022)
by: Zhu, Beier, et al.
Published: (2022)
SpeechAct: Towards Generating Whole-body Motion from Speech
by: Zhang, Jinsong, et al.
Published: (2023)
by: Zhang, Jinsong, et al.
Published: (2023)
MonoSOWA: Scalable monocular 3D Object detector Without human Annotations
by: Skvrna, Jan, et al.
Published: (2025)
by: Skvrna, Jan, et al.
Published: (2025)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
by: Csizmadia, Daniel, et al.
Published: (2025)
by: Csizmadia, Daniel, et al.
Published: (2025)
Non-aligned supervision for Real Image Dehazing
by: Fan, Junkai, et al.
Published: (2023)
by: Fan, Junkai, et al.
Published: (2023)
Gravity-aligned Rotation Averaging with Circular Regression
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Learning to mask: Towards generalized face forgery detection
by: Fei, Jianwei, et al.
Published: (2022)
by: Fei, Jianwei, et al.
Published: (2022)
Digging into contrastive learning for robust depth estimation with diffusion models
by: Wang, Jiyuan, et al.
Published: (2024)
by: Wang, Jiyuan, et al.
Published: (2024)
Similar Items
-
Towards Practical Single-shot Motion Synthesis
by: Roditakis, Konstantinos, et al.
Published: (2024) -
On the Skinning of Gaussian Avatars
by: Zioulis, Nikolaos, et al.
Published: (2025) -
METER: a mobile vision transformer architecture for monocular depth estimation
by: Papa, L., et al.
Published: (2024) -
D4D: An RGBD diffusion model to boost monocular depth estimation
by: Papa, L., et al.
Published: (2024) -
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
by: Skorupski, Patryk, et al.
Published: (2025)