Action Motifs: Self-Supervised Hierarchical Representation of Human Body Movements
Fuente:
arXiv
Saved in:
| Main Authors: | Kinoshita, Genki, Nakamura, Shu, Kawahara, Ryo, Nobuhara, Shohei, Kawanishi, Yasutomo, Nishino, Ko |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REACH: Hand Pose Estimation from Room Corners
by: Nakamura, Shu, et al.
Published: (2026)
by: Nakamura, Shu, et al.
Published: (2026)
Spatiotemporal Multi-Camera Calibration using Freely Moving People
by: Lee, Sang-Eun, et al.
Published: (2025)
by: Lee, Sang-Eun, et al.
Published: (2025)
SteerPose: Simultaneous Extrinsic Camera Calibration and Matching from Articulation
by: Lee, Sang-Eun, et al.
Published: (2025)
by: Lee, Sang-Eun, et al.
Published: (2025)
Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation
by: Kinoshita, Genki, et al.
Published: (2023)
by: Kinoshita, Genki, et al.
Published: (2023)
SPIDeRS: Structured Polarization for Invisible Depth and Reflectance Sensing
by: Ichikawa, Tomoki, et al.
Published: (2023)
by: Ichikawa, Tomoki, et al.
Published: (2023)
RatBodyFormer: Rat Body Surface from Keypoints
by: Higami, Ayaka, et al.
Published: (2024)
by: Higami, Ayaka, et al.
Published: (2024)
Fooling Polarization-based Vision using Locally Controllable Polarizing Projection
by: Li, Zhuoxiao, et al.
Published: (2023)
by: Li, Zhuoxiao, et al.
Published: (2023)
Spatial Polarization Multiplexing: Single-Shot Invisible Shape and Reflectance Recovery
by: Ichikawa, Tomoki, et al.
Published: (2025)
by: Ichikawa, Tomoki, et al.
Published: (2025)
Multi-View Video-Based Learning: Leveraging Weak Labels for Frame-Level Perception
by: John, Vijay, et al.
Published: (2024)
by: John, Vijay, et al.
Published: (2024)
Action Selection Learning for Multi-label Multi-view Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2024)
by: Nguyen, Trung Thanh, et al.
Published: (2024)
M-PhyGs: Multi-Material Object Dynamics from Video
by: Wada, Norika, et al.
Published: (2025)
by: Wada, Norika, et al.
Published: (2025)
How good was my shot? Quantifying Player Skill Level in Table Tennis
by: Kubota, Akihiro, et al.
Published: (2026)
by: Kubota, Akihiro, et al.
Published: (2026)
One-Stage Open-Vocabulary Temporal Action Detection Leveraging Temporal Multi-scale and Action Label Features
by: Nguyen, Trung Thanh, et al.
Published: (2024)
by: Nguyen, Trung Thanh, et al.
Published: (2024)
KFD-NeRF: Rethinking Dynamic NeRF with Kalman Filter
by: Zhan, Yifan, et al.
Published: (2024)
by: Zhan, Yifan, et al.
Published: (2024)
MultiTSF: Transformer-based Sensor Fusion for Human-Centric Multi-view and Multi-modal Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
View-aware Cross-modal Distillation for Multi-view Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
Under One Sun: Multi-Object Generative Perception of Materials and Illumination
by: Yoshii, Nobuo, et al.
Published: (2026)
by: Yoshii, Nobuo, et al.
Published: (2026)
MultiSensor-Home: A Wide-area Multi-modal Multi-view Dataset for Action Recognition and Transformer-based Sensor Fusion
by: Nguyen, Trung Thanh, et al.
Published: (2025)
by: Nguyen, Trung Thanh, et al.
Published: (2025)
Tracking Small Birds by Detection Candidate Region Filtering and Detection History-aware Association
by: Liu, Tingwei, et al.
Published: (2024)
by: Liu, Tingwei, et al.
Published: (2024)
Detailed Geometry and Appearance from Opportunistic Motion
by: Hirai, Ryosuke, et al.
Published: (2026)
by: Hirai, Ryosuke, et al.
Published: (2026)
HeatFormer: A Neural Optimizer for Multiview Human Mesh Recovery
by: Matsubara, Yuto, et al.
Published: (2024)
by: Matsubara, Yuto, et al.
Published: (2024)
FROSS: Faster-than-Real-Time Online 3D Semantic Scene Graph Generation from RGB-D Images
by: Hou, Hao-Yu, et al.
Published: (2025)
by: Hou, Hao-Yu, et al.
Published: (2025)
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions
by: Inadumi, Shun, et al.
Published: (2024)
by: Inadumi, Shun, et al.
Published: (2024)
Physical Plausibility-aware Trajectory Prediction via Locomotion Embodiment
by: Taketsugu, Hiromu, et al.
Published: (2025)
by: Taketsugu, Hiromu, et al.
Published: (2025)
Self Supervised Networks for Learning Latent Space Representations of Human Body Scans and Motions
by: Hartman, Emmanuel, et al.
Published: (2024)
by: Hartman, Emmanuel, et al.
Published: (2024)
Class-agnostic 3D Segmentation by Granularity-Consistent Automatic 2D Mask Tracking
by: Wang, Juan, et al.
Published: (2025)
by: Wang, Juan, et al.
Published: (2025)
PBDyG: Position Based Dynamic Gaussians for Motion-Aware Clothed Human Avatars
by: Sasaki, Shota, et al.
Published: (2024)
by: Sasaki, Shota, et al.
Published: (2024)
Diffusion Reflectance Map: Single-Image Stochastic Inverse Rendering of Illumination and Reflectance
by: Enyo, Yuto, et al.
Published: (2023)
by: Enyo, Yuto, et al.
Published: (2023)
Representation Synthesis by Probabilistic Many-Valued Logic Operation in Self-Supervised Learning
by: Nakamura, Hiroki, et al.
Published: (2023)
by: Nakamura, Hiroki, et al.
Published: (2023)
SCAR: Self-Supervised Continuous Action Representation Learning
by: Liu, Hongjia, et al.
Published: (2026)
by: Liu, Hongjia, et al.
Published: (2026)
Hierarchical Action Learning for Weakly-Supervised Action Segmentation
by: Huang, Junxian, et al.
Published: (2026)
by: Huang, Junxian, et al.
Published: (2026)
Small Object Detection for Birds with Swin Transformer
by: Huo, Da, et al.
Published: (2025)
by: Huo, Da, et al.
Published: (2025)
Self-Supervised Skeleton-Based Action Representation Learning: A Benchmark and Beyond
by: Zhang, Jiahang, et al.
Published: (2024)
by: Zhang, Jiahang, et al.
Published: (2024)
Hierarchical Text-to-Vision Self Supervised Alignment for Improved Histopathology Representation Learning
by: Watawana, Hasindri, et al.
Published: (2024)
by: Watawana, Hasindri, et al.
Published: (2024)
ForestMamba: Sparse Mamba with Geometry-guided Queries for 3D Forest Point Cloud Segmentation
by: Nguyen, Trung Thanh, et al.
Published: (2026)
by: Nguyen, Trung Thanh, et al.
Published: (2026)
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
by: Liu, Mengyuan, et al.
Published: (2023)
by: Liu, Mengyuan, et al.
Published: (2023)
Correspondences of the Third Kind: Camera Pose Estimation from Object Reflection
by: Yamashita, Kohei, et al.
Published: (2023)
by: Yamashita, Kohei, et al.
Published: (2023)
Robust Human Trajectory Prediction via Self-Supervised Skeleton Representation Learning
by: Arashima, Taishu, et al.
Published: (2026)
by: Arashima, Taishu, et al.
Published: (2026)
Generative Perception of Shape and Material from Differential Motion
by: Han, Xinran Nicole, et al.
Published: (2025)
by: Han, Xinran Nicole, et al.
Published: (2025)
Multistable Shape from Shading Emerges from Patch Diffusion
by: Han, Xinran Nicole, et al.
Published: (2024)
by: Han, Xinran Nicole, et al.
Published: (2024)
Similar Items
-
REACH: Hand Pose Estimation from Room Corners
by: Nakamura, Shu, et al.
Published: (2026) -
Spatiotemporal Multi-Camera Calibration using Freely Moving People
by: Lee, Sang-Eun, et al.
Published: (2025) -
SteerPose: Simultaneous Extrinsic Camera Calibration and Matching from Articulation
by: Lee, Sang-Eun, et al.
Published: (2025) -
Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation
by: Kinoshita, Genki, et al.
Published: (2023) -
SPIDeRS: Structured Polarization for Invisible Depth and Reflectance Sensing
by: Ichikawa, Tomoki, et al.
Published: (2023)