Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
Fuente:
arXiv
Saved in:
| Main Authors: | Karnoor, Sahil Bhandary, Choudhury, Romit Roy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
by: Sezer, Berk, et al.
Published: (2026)
by: Sezer, Berk, et al.
Published: (2026)
DeltaDorsal: Enhancing Hand Pose Estimation with Dorsal Features in Egocentric Views
by: Huang, William, et al.
Published: (2026)
by: Huang, William, et al.
Published: (2026)
Visuo-Acoustic Hand Pose and Contact Estimation
by: Mao, Yuemin, et al.
Published: (2025)
by: Mao, Yuemin, et al.
Published: (2025)
emg2pose: A Large and Diverse Benchmark for Surface Electromyographic Hand Pose Estimation
by: Salter, Sasha, et al.
Published: (2024)
by: Salter, Sasha, et al.
Published: (2024)
Human Motion Synthesis_ A Diffusion Approach for Motion Stitching and In-Betweening
by: Adewole, Michael, et al.
Published: (2024)
by: Adewole, Michael, et al.
Published: (2024)
Person Identification from Egocentric Human-Object Interactions using 3D Hand Pose
by: Hamza, Muhammad, et al.
Published: (2025)
by: Hamza, Muhammad, et al.
Published: (2025)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions
by: Xu, Vasco, et al.
Published: (2026)
by: Xu, Vasco, et al.
Published: (2026)
Dependency-Aware Discrete Diffusion for Scene Graph Generation
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
Continuous Human Action Recognition for Human-Machine Interaction: A Review
by: Gammulle, Harshala, et al.
Published: (2022)
by: Gammulle, Harshala, et al.
Published: (2022)
Zero-shot Emotion Annotation in Facial Images Using Large Multimodal Models: Benchmarking and Prospects for Multi-Class, Multi-Frame Approaches
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
ThermoHands: A Benchmark for 3D Hand Pose Estimation from Egocentric Thermal Images
by: Ding, Fangqiang, et al.
Published: (2024)
by: Ding, Fangqiang, et al.
Published: (2024)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
Human-like visual computing advances explainability and few-shot learning in deep neural networks for complex physiological data
by: Alahmadi, Alaa, et al.
Published: (2025)
by: Alahmadi, Alaa, et al.
Published: (2025)
PoseAugment: Generative Human Pose Data Augmentation with Physical Plausibility for IMU-based Motion Capture
by: Li, Zhuojun, et al.
Published: (2024)
by: Li, Zhuojun, et al.
Published: (2024)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
by: Vu, Evgeniia, et al.
Published: (2025)
by: Vu, Evgeniia, et al.
Published: (2025)
A Review of Driver Gaze Estimation and Application in Gaze Behavior Understanding
by: Sharma, Pavan Kumar, et al.
Published: (2023)
by: Sharma, Pavan Kumar, et al.
Published: (2023)
Evaluating the Evaluators: Towards Human-aligned Metrics for Missing Markers Reconstruction
by: Kucherenko, Taras, et al.
Published: (2024)
by: Kucherenko, Taras, et al.
Published: (2024)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic Propagation
by: Liu, Haofeng, et al.
Published: (2024)
by: Liu, Haofeng, et al.
Published: (2024)
Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition
by: Zhou, Zishu, et al.
Published: (2026)
by: Zhou, Zishu, et al.
Published: (2026)
Where Does My Model Underperform? A Human Evaluation of Slice Discovery Algorithms
by: Johnson, Nari, et al.
Published: (2023)
by: Johnson, Nari, et al.
Published: (2023)
Intervening in Black Box: Concept Bottleneck Model for Enhancing Human Neural Network Mutual Understanding
by: Xiong, Nuoye, et al.
Published: (2025)
by: Xiong, Nuoye, et al.
Published: (2025)
From Model Uncertainty to Human Attention: Localization-Aware Visual Cues for Scalable Annotation Review
by: Sbeyti, Moussa Kassem, et al.
Published: (2026)
by: Sbeyti, Moussa Kassem, et al.
Published: (2026)
Supporting Experts with a Multimodal Machine-Learning-Based Tool for Human Behavior Analysis of Conversational Videos
by: Arakawa, Riku, et al.
Published: (2024)
by: Arakawa, Riku, et al.
Published: (2024)
Timeline-based Process Discovery
by: Kaur, Harleen, et al.
Published: (2023)
by: Kaur, Harleen, et al.
Published: (2023)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
by: Hiranaka, Ayano, et al.
Published: (2024)
by: Hiranaka, Ayano, et al.
Published: (2024)
WheelPose: Data Synthesis Techniques to Improve Pose Estimation Performance on Wheelchair Users
by: Huang, William, et al.
Published: (2024)
by: Huang, William, et al.
Published: (2024)
Unsupervised visualization of image datasets using contrastive learning
by: Böhm, Jan Niklas, et al.
Published: (2022)
by: Böhm, Jan Niklas, et al.
Published: (2022)
VRMN-bD: A Multi-modal Natural Behavior Dataset of Immersive Human Fear Responses in VR Stand-up Interactive Games
by: Zhang, He, et al.
Published: (2024)
by: Zhang, He, et al.
Published: (2024)
Smartphone-based Eye Tracking System using Edge Intelligence and Model Optimisation
by: Gunawardena, Nishan, et al.
Published: (2024)
by: Gunawardena, Nishan, et al.
Published: (2024)
Assessing Graphical Perception of Image Embedding Models using Channel Effectiveness
by: Lee, Soohyun, et al.
Published: (2024)
by: Lee, Soohyun, et al.
Published: (2024)
FreeDrag: Feature Dragging for Reliable Point-based Image Editing
by: Ling, Pengyang, et al.
Published: (2023)
by: Ling, Pengyang, et al.
Published: (2023)
Improve accessibility for Low Vision and Blind people using Machine Learning and Computer Vision
by: Shukurov, Jasur
Published: (2024)
by: Shukurov, Jasur
Published: (2024)
Helios: An extremely low power event-based gesture recognition for always-on smart eyewear
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
A Deep Learning-based Multimodal Depth-Aware Dynamic Hand Gesture Recognition System
by: Mahmud, Hasan, et al.
Published: (2021)
by: Mahmud, Hasan, et al.
Published: (2021)
Strike the Balance: On-the-Fly Uncertainty based User Interactions for Long-Term Video Object Segmentation
by: Vujasinović, Stéphane, et al.
Published: (2024)
by: Vujasinović, Stéphane, et al.
Published: (2024)
ConvMixFormer- A Resource-efficient Convolution Mixer for Transformer-based Dynamic Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2024)
by: Garg, Mallika, et al.
Published: (2024)
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
by: Jalayer, Reza, et al.
Published: (2025)
by: Jalayer, Reza, et al.
Published: (2025)
Helios 2.0: A Robust, Ultra-Low Power Gesture Recognition System Optimised for Event-Sensor based Wearables
by: Bhattacharyya, Prarthana, et al.
Published: (2025)
by: Bhattacharyya, Prarthana, et al.
Published: (2025)
Similar Items
-
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
by: Sezer, Berk, et al.
Published: (2026) -
DeltaDorsal: Enhancing Hand Pose Estimation with Dorsal Features in Egocentric Views
by: Huang, William, et al.
Published: (2026) -
Visuo-Acoustic Hand Pose and Contact Estimation
by: Mao, Yuemin, et al.
Published: (2025) -
emg2pose: A Large and Diverse Benchmark for Surface Electromyographic Hand Pose Estimation
by: Salter, Sasha, et al.
Published: (2024) -
Human Motion Synthesis_ A Diffusion Approach for Motion Stitching and In-Betweening
by: Adewole, Michael, et al.
Published: (2024)