Including Semantic Information via Word Embeddings for Skeleton-based Action Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Aganian, Dustin, Franze, Erik, Eisenbach, Markus, Gross, Horst-Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SkelVIT: Consensus of Vision Transformers for a Lightweight Skeleton-Based Action Recognition System
by: Karadag, Ozge Oztimur
Published: (2023)
by: Karadag, Ozge Oztimur
Published: (2023)
From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
by: Toupas, Petros, et al.
Published: (2023)
by: Toupas, Petros, et al.
Published: (2023)
RALACs: Action Recognition in Autonomous Vehicles using Interaction Encoding and Optical Flow
by: Zhou, Eddy, et al.
Published: (2022)
by: Zhou, Eddy, et al.
Published: (2022)
Robustness Evaluation of Machine Learning Models for Robot Arm Action Recognition in Noisy Environments
by: Motamedi, Elaheh, et al.
Published: (2024)
by: Motamedi, Elaheh, et al.
Published: (2024)
Efficient Prediction of Dense Visual Embeddings via Distillation and RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2026)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2026)
VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
by: Xu, Siyu, et al.
Published: (2025)
by: Xu, Siyu, et al.
Published: (2025)
Hybrid Training for Vision-Language-Action Models
by: Mazzaglia, Pietro, et al.
Published: (2025)
by: Mazzaglia, Pietro, et al.
Published: (2025)
Adaptive Hyper-Graph Convolution Network for Skeleton-based Human Action Recognition with Virtual Connections
by: Zhou, Youwei, et al.
Published: (2024)
by: Zhou, Youwei, et al.
Published: (2024)
CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition
by: Wen, Yuhang, et al.
Published: (2024)
by: Wen, Yuhang, et al.
Published: (2024)
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
by: Bazhenov, Artem, et al.
Published: (2024)
by: Bazhenov, Artem, et al.
Published: (2024)
SPAQ-DL-SLAM: Towards Optimizing Deep Learning-based SLAM for Resource-Constrained Embedded Platforms
by: Pudasaini, Niraj, et al.
Published: (2024)
by: Pudasaini, Niraj, et al.
Published: (2024)
Benchmarking Sensitivity of Continual Graph Learning for Skeleton-Based Action Recognition
by: Wei, Wei, et al.
Published: (2024)
by: Wei, Wei, et al.
Published: (2024)
Evolving Skeletons: Motion Dynamics in Action Recognition
by: Qiu, Jushang, et al.
Published: (2025)
by: Qiu, Jushang, et al.
Published: (2025)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
by: Liang, Zhixuan, et al.
Published: (2025)
by: Liang, Zhixuan, et al.
Published: (2025)
Crossway Diffusion: Improving Diffusion-based Visuomotor Policy via Self-supervised Learning
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
FEEL (Force-Enhanced Egocentric Learning): A Dataset for Physical Action Understanding
by: Dessalene, Eadom, et al.
Published: (2026)
by: Dessalene, Eadom, et al.
Published: (2026)
Exploring Single Domain Generalization of LiDAR-based Semantic Segmentation under Imperfect Labels
by: Kong, Weitong, et al.
Published: (2025)
by: Kong, Weitong, et al.
Published: (2025)
FALCON: Future-Aware Learning with Contextual Object-Centric Pretraining for UAV Action Recognition
by: Xian, Ruiqi, et al.
Published: (2024)
by: Xian, Ruiqi, et al.
Published: (2024)
MVSA-Net: Multi-View State-Action Recognition for Robust and Deployable Trajectory Generation
by: Asali, Ehsan, et al.
Published: (2023)
by: Asali, Ehsan, et al.
Published: (2023)
Grounding Foundational Vision Models with 3D Human Poses for Robust Action Recognition
by: Babey, Nicholas, et al.
Published: (2025)
by: Babey, Nicholas, et al.
Published: (2025)
Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition
by: Wen, Yuhang, et al.
Published: (2023)
by: Wen, Yuhang, et al.
Published: (2023)
Dense Policy: Bidirectional Autoregressive Learning of Actions
by: Su, Yue, et al.
Published: (2025)
by: Su, Yue, et al.
Published: (2025)
World Action Models are Zero-shot Policies
by: Ye, Seonghyeon, et al.
Published: (2026)
by: Ye, Seonghyeon, et al.
Published: (2026)
HFGCN:Hypergraph Fusion Graph Convolutional Networks for Skeleton-Based Action Recognition
by: Dong, Pengcheng, et al.
Published: (2025)
by: Dong, Pengcheng, et al.
Published: (2025)
Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation
by: Zhang, Wenbo, et al.
Published: (2025)
by: Zhang, Wenbo, et al.
Published: (2025)
Tactile Modality Fusion for Vision-Language-Action Models
by: Morissette, Charlotte, et al.
Published: (2026)
by: Morissette, Charlotte, et al.
Published: (2026)
Motus: A Unified Latent Action World Model
by: Bi, Hongzhe, et al.
Published: (2025)
by: Bi, Hongzhe, et al.
Published: (2025)
Few-shot Semantic Learning for Robust Multi-Biome 3D Semantic Mapping in Off-Road Environments
by: Atha, Deegan, et al.
Published: (2024)
by: Atha, Deegan, et al.
Published: (2024)
Skeleton-Based Human Action Recognition with Noisy Labels
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
SparseGrasp: Robotic Grasping via 3D Semantic Gaussian Splatting from Sparse Multi-View RGB Images
by: Yu, Junqiu, et al.
Published: (2024)
by: Yu, Junqiu, et al.
Published: (2024)
Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
by: Lin, Haitao, et al.
Published: (2026)
by: Lin, Haitao, et al.
Published: (2026)
Intelligent Sensing-to-Action for Robust Autonomy at the Edge: Opportunities and Challenges
by: Trivedi, Amit Ranjan, et al.
Published: (2025)
by: Trivedi, Amit Ranjan, et al.
Published: (2025)
Benchmarking Vision, Language, & Action Models on Robotic Learning Tasks
by: Guruprasad, Pranav, et al.
Published: (2024)
by: Guruprasad, Pranav, et al.
Published: (2024)
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
by: Niu, Dantong, et al.
Published: (2024)
by: Niu, Dantong, et al.
Published: (2024)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
by: Zhang, Zezhou, et al.
Published: (2026)
by: Zhang, Zezhou, et al.
Published: (2026)
Improving Vision-Language-Action Model with Online Reinforcement Learning
by: Guo, Yanjiang, et al.
Published: (2025)
by: Guo, Yanjiang, et al.
Published: (2025)
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
by: Tan, Hengkai, et al.
Published: (2025)
by: Tan, Hengkai, et al.
Published: (2025)
Focusing on What Matters: Object-Agent-centric Tokenization for Vision Language Action models
by: Bendikas, Rokas, et al.
Published: (2025)
by: Bendikas, Rokas, et al.
Published: (2025)
Diffusion-Based Environment-Aware Trajectory Prediction
by: Westny, Theodor, et al.
Published: (2024)
by: Westny, Theodor, et al.
Published: (2024)
Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression
by: Wang, Lirui, et al.
Published: (2025)
by: Wang, Lirui, et al.
Published: (2025)
Similar Items
-
SkelVIT: Consensus of Vision Transformers for a Lightweight Skeleton-Based Action Recognition System
by: Karadag, Ozge Oztimur
Published: (2023) -
From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
by: Toupas, Petros, et al.
Published: (2023) -
RALACs: Action Recognition in Autonomous Vehicles using Interaction Encoding and Optical Flow
by: Zhou, Eddy, et al.
Published: (2022) -
Robustness Evaluation of Machine Learning Models for Robot Arm Action Recognition in Noisy Environments
by: Motamedi, Elaheh, et al.
Published: (2024) -
Efficient Prediction of Dense Visual Embeddings via Distillation and RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2026)