Learning Priors of Human Motion With Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Falqueto, Placido, Sanfeliu, Alberto, Palopoli, Luigi, Fontanelli, Daniele |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Surface Defect Identification using Bayesian Filtering on a 3D Mesh
von: Vedove, Matteo Dalle, et al.
Veröffentlicht: (2025)
von: Vedove, Matteo Dalle, et al.
Veröffentlicht: (2025)
Socially-Aware Opinion-Based Navigation with Oval Limit Cycles
von: d'Addato, Giulia, et al.
Veröffentlicht: (2024)
von: d'Addato, Giulia, et al.
Veröffentlicht: (2024)
Human motion trajectory prediction using the Social Force Model for real-time and low computational cost applications
von: Gil, Oscar, et al.
Veröffentlicht: (2023)
von: Gil, Oscar, et al.
Veröffentlicht: (2023)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
von: Wang, Xinkai, et al.
Veröffentlicht: (2026)
von: Wang, Xinkai, et al.
Veröffentlicht: (2026)
Synthetic-to-Real Self-supervised Robust Depth Estimation via Learning with Motion and Structure Priors
von: Yan, Weilong, et al.
Veröffentlicht: (2025)
von: Yan, Weilong, et al.
Veröffentlicht: (2025)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
von: Feng, Yuming, et al.
Veröffentlicht: (2024)
MP-SfM: Monocular Surface Priors for Robust Structure-from-Motion
von: Pataki, Zador, et al.
Veröffentlicht: (2025)
von: Pataki, Zador, et al.
Veröffentlicht: (2025)
AMPLIFY: Actionless Motion Priors for Robot Learning from Videos
von: Collins, Jeremy A., et al.
Veröffentlicht: (2025)
von: Collins, Jeremy A., et al.
Veröffentlicht: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
von: Wang, Kuanning, et al.
Veröffentlicht: (2026)
von: Wang, Kuanning, et al.
Veröffentlicht: (2026)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
von: Jeong, David C., et al.
Veröffentlicht: (2025)
von: Jeong, David C., et al.
Veröffentlicht: (2025)
ADM: Accelerated Diffusion Model via Estimated Priors for Robust Motion Prediction under Uncertainties
von: Li, Jiahui, et al.
Veröffentlicht: (2024)
von: Li, Jiahui, et al.
Veröffentlicht: (2024)
MaskAdapt: Learning Flexible Motion Adaptation via Mask-Invariant Prior for Physics-Based Characters
von: Park, Soomin, et al.
Veröffentlicht: (2026)
von: Park, Soomin, et al.
Veröffentlicht: (2026)
Video Generation with Learned Action Prior
von: Sarkar, Meenakshi, et al.
Veröffentlicht: (2024)
von: Sarkar, Meenakshi, et al.
Veröffentlicht: (2024)
MGTR: Multi-Granular Transformer for Motion Prediction with LiDAR
von: Gan, Yiqian, et al.
Veröffentlicht: (2023)
von: Gan, Yiqian, et al.
Veröffentlicht: (2023)
Multi-Transmotion: Pre-trained Model for Human Motion Prediction
von: Gao, Yang, et al.
Veröffentlicht: (2024)
von: Gao, Yang, et al.
Veröffentlicht: (2024)
Symmetry-Aware Fusion of Vision and Tactile Sensing via Bilateral Force Priors for Robotic Manipulation
von: Lee, Wonju, et al.
Veröffentlicht: (2026)
von: Lee, Wonju, et al.
Veröffentlicht: (2026)
Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning
von: Zhang, Mengtan, et al.
Veröffentlicht: (2025)
von: Zhang, Mengtan, et al.
Veröffentlicht: (2025)
HaltNav: Reactive Visual Halting over Lightweight Topological Priors for Robust Vision-Language Navigation
von: Yu, Zihui, et al.
Veröffentlicht: (2026)
von: Yu, Zihui, et al.
Veröffentlicht: (2026)
Look, Focus, Act: Efficient and Robust Robot Learning via Human Gaze and Foveated Vision Transformers
von: Chuang, Ian, et al.
Veröffentlicht: (2025)
von: Chuang, Ian, et al.
Veröffentlicht: (2025)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
von: Zhang, Chubin, et al.
Veröffentlicht: (2026)
von: Zhang, Chubin, et al.
Veröffentlicht: (2026)
Kimodo: Scaling Controllable Human Motion Generation
von: Rempe, Davis, et al.
Veröffentlicht: (2026)
von: Rempe, Davis, et al.
Veröffentlicht: (2026)
Unified Human Localization and Trajectory Prediction with Monocular Vision
von: Luan, Po-Chien, et al.
Veröffentlicht: (2025)
von: Luan, Po-Chien, et al.
Veröffentlicht: (2025)
MUT3R: Motion-aware Updating Transformer for Dynamic 3D Reconstruction
von: Shen, Guole, et al.
Veröffentlicht: (2025)
von: Shen, Guole, et al.
Veröffentlicht: (2025)
Learning Appearance and Motion Cues for Panoptic Tracking
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2025)
von: Hurtado, Juana Valeria, et al.
Veröffentlicht: (2025)
Uncertainty-aware Probabilistic 3D Human Motion Forecasting via Invertible Networks
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction
von: Shi, Zhonghao, et al.
Veröffentlicht: (2025)
von: Shi, Zhonghao, et al.
Veröffentlicht: (2025)
pySLAM: An Open-Source, Modular, and Extensible Framework for SLAM
von: Freda, Luigi
Veröffentlicht: (2025)
von: Freda, Luigi
Veröffentlicht: (2025)
PLVS: A SLAM System with Points, Lines, Volumetric Mapping, and 3D Incremental Segmentation
von: Freda, Luigi
Veröffentlicht: (2023)
von: Freda, Luigi
Veröffentlicht: (2023)
MPVO: Motion-Prior based Visual Odometry for PointGoal Navigation
von: Paul, Sayan, et al.
Veröffentlicht: (2024)
von: Paul, Sayan, et al.
Veröffentlicht: (2024)
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
von: Caramia, Donato, et al.
Veröffentlicht: (2025)
von: Caramia, Donato, et al.
Veröffentlicht: (2025)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
von: Li, Wenhao, et al.
Veröffentlicht: (2025)
Vision-Based Safe Human-Robot Collaboration with Uncertainty Guarantees
von: Thumm, Jakob, et al.
Veröffentlicht: (2026)
von: Thumm, Jakob, et al.
Veröffentlicht: (2026)
MOGRAS: Human Motion with Grasping in 3D Scenes
von: Bhosikar, Kunal, et al.
Veröffentlicht: (2025)
von: Bhosikar, Kunal, et al.
Veröffentlicht: (2025)
MOSPA: Human Motion Generation Driven by Spatial Audio
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
von: Xu, Shuyang, et al.
Veröffentlicht: (2025)
MAPLE: Encoding Dexterous Robotic Manipulation Priors Learned From Egocentric Videos
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
History-Enhanced Two-Stage Transformer for Aerial Vision-and-Language Navigation
von: Ding, Xichen, et al.
Veröffentlicht: (2025)
von: Ding, Xichen, et al.
Veröffentlicht: (2025)
Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
von: Hou, Zhi, et al.
Veröffentlicht: (2025)
von: Hou, Zhi, et al.
Veröffentlicht: (2025)
InterPrior: Scaling Generative Control for Physics-Based Human-Object Interactions
von: Xu, Sirui, et al.
Veröffentlicht: (2026)
von: Xu, Sirui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Surface Defect Identification using Bayesian Filtering on a 3D Mesh
von: Vedove, Matteo Dalle, et al.
Veröffentlicht: (2025) -
Socially-Aware Opinion-Based Navigation with Oval Limit Cycles
von: d'Addato, Giulia, et al.
Veröffentlicht: (2024) -
Human motion trajectory prediction using the Social Force Model for real-time and low computational cost applications
von: Gil, Oscar, et al.
Veröffentlicht: (2023) -
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
von: Wang, Xinkai, et al.
Veröffentlicht: (2026) -
Synthetic-to-Real Self-supervised Robust Depth Estimation via Learning with Motion and Structure Priors
von: Yan, Weilong, et al.
Veröffentlicht: (2025)