Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Xing, Hao, Boey, Kai Zhe, Wu, Yuankai, Burschka, Darius, Cheng, Gordon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Open-World Human Action Segmentation Using Graph Convolutional Networks
by: Xing, Hao, et al.
Published: (2025)
by: Xing, Hao, et al.
Published: (2025)
Understanding Human Activity with Uncertainty Measure for Novelty in Graph Convolutional Networks
by: Xing, Hao, et al.
Published: (2024)
by: Xing, Hao, et al.
Published: (2024)
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
by: Xing, Hao, et al.
Published: (2024)
by: Xing, Hao, et al.
Published: (2024)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
by: Guo, Jianing, et al.
Published: (2025)
by: Guo, Jianing, et al.
Published: (2025)
Using The Concept Hierarchy for Household Action Recognition
by: Costinescu, Andrei, et al.
Published: (2024)
by: Costinescu, Andrei, et al.
Published: (2024)
Adaptation of Task Goal States from Prior Knowledge
by: Costinescu, Andrei, et al.
Published: (2025)
by: Costinescu, Andrei, et al.
Published: (2025)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
by: Zahid, Azizul, et al.
Published: (2025)
by: Zahid, Azizul, et al.
Published: (2025)
Friends Across Time: Multi-Scale Action Segmentation Transformer for Surgical Phase Recognition
by: Zhang, Bokai, et al.
Published: (2024)
by: Zhang, Bokai, et al.
Published: (2024)
MS-TCRNet: Multi-Stage Temporal Convolutional Recurrent Networks for Action Segmentation Using Sensor-Augmented Kinematics
by: Goldbraikh, Adam, et al.
Published: (2023)
by: Goldbraikh, Adam, et al.
Published: (2023)
Speeding Up Optimization-based Motion Planning through Deep Learning
by: Tenhumberg, Johannes, et al.
Published: (2023)
by: Tenhumberg, Johannes, et al.
Published: (2023)
Graph-Based Multi-Modal Sensor Fusion for Autonomous Driving
by: Sani, Depanshu, et al.
Published: (2024)
by: Sani, Depanshu, et al.
Published: (2024)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
by: Bhattacharya, Uttaran, et al.
Published: (2019)
by: Bhattacharya, Uttaran, et al.
Published: (2019)
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures
by: Belal, Mohammad, et al.
Published: (2024)
by: Belal, Mohammad, et al.
Published: (2024)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Accelerated Rotation-Invariant Convolution for UAV Image Segmentation
by: Manduhu, Manduhu, et al.
Published: (2025)
by: Manduhu, Manduhu, et al.
Published: (2025)
A Distributed Multi-Modal Sensing Approach for Human Activity Recognition in Real-Time Human-Robot Collaboration
by: Belcamino, Valerio, et al.
Published: (2026)
by: Belcamino, Valerio, et al.
Published: (2026)
Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence
by: Li, Wenxin, et al.
Published: (2025)
by: Li, Wenxin, et al.
Published: (2025)
Person Segmentation and Action Classification for Multi-Channel Hemisphere Field of View LiDAR Sensors
by: Seliunina, Svetlana, et al.
Published: (2024)
by: Seliunina, Svetlana, et al.
Published: (2024)
Physics-Encoded Graph Neural Networks for Deformation Prediction under Contact
by: Saleh, Mahdi, et al.
Published: (2024)
by: Saleh, Mahdi, et al.
Published: (2024)
LuSeg: Efficient Negative and Positive Obstacles Segmentation via Contrast-Driven Multi-Modal Feature Fusion on the Lunar
by: Jiao, Shuaifeng, et al.
Published: (2025)
by: Jiao, Shuaifeng, et al.
Published: (2025)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
Multimodal Graph Representation Learning for Robust Surgical Workflow Recognition with Adversarial Feature Disentanglement
by: Bai, Long, et al.
Published: (2025)
by: Bai, Long, et al.
Published: (2025)
Modality-Composable Diffusion Policy via Inference-Time Distribution-level Composition
by: Cao, Jiahang, et al.
Published: (2025)
by: Cao, Jiahang, et al.
Published: (2025)
Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy
by: Dai, Jiaheng, et al.
Published: (2026)
by: Dai, Jiaheng, et al.
Published: (2026)
Self-Contained Calibration of an Elastic Humanoid Upper Body Using Only a Head-Mounted RGB Camera
by: Tenhumberg, Johannes, et al.
Published: (2023)
by: Tenhumberg, Johannes, et al.
Published: (2023)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
by: Yin, Zhenhan, et al.
Published: (2025)
by: Yin, Zhenhan, et al.
Published: (2025)
PROBE: Probabilistic Occupancy BEV Encoding with Analytical Translation Robustness for 3D Place Recognition
by: Lee, Jinseop, et al.
Published: (2026)
by: Lee, Jinseop, et al.
Published: (2026)
UniBEV: Multi-modal 3D Object Detection with Uniform BEV Encoders for Robustness against Missing Sensor Modalities
by: Wang, Shiming, et al.
Published: (2023)
by: Wang, Shiming, et al.
Published: (2023)
Deep Learning-Based Multi-Modal Fusion for Robust Robot Perception and Navigation
by: Lai, Delun, et al.
Published: (2025)
by: Lai, Delun, et al.
Published: (2025)
Multi-Modal UAV Detection, Classification and Tracking Algorithm -- Technical Report for CVPR 2024 UG2 Challenge
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
by: Zhang, Chubin, et al.
Published: (2026)
by: Zhang, Chubin, et al.
Published: (2026)
Multi-Camera Asynchronous Ball Localization and Trajectory Prediction with Factor Graphs and Human Poses
by: Xiao, Qingyu, et al.
Published: (2024)
by: Xiao, Qingyu, et al.
Published: (2024)
Rodrigues Network for Learning Robot Actions
by: Zhang, Jialiang, et al.
Published: (2025)
by: Zhang, Jialiang, et al.
Published: (2025)
RALACs: Action Recognition in Autonomous Vehicles using Interaction Encoding and Optical Flow
by: Zhou, Eddy, et al.
Published: (2022)
by: Zhou, Eddy, et al.
Published: (2022)
Tactile Modality Fusion for Vision-Language-Action Models
by: Morissette, Charlotte, et al.
Published: (2026)
by: Morissette, Charlotte, et al.
Published: (2026)
Forging Spatial Intelligence: A Roadmap of Multi-Modal Data Pre-Training for Autonomous Systems
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
RLBind: Adversarial-Invariant Cross-Modal Alignment for Unified Robust Embeddings
by: Lu, Yuhong
Published: (2025)
by: Lu, Yuhong
Published: (2025)
PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction
by: Chen, Shizhe, et al.
Published: (2026)
by: Chen, Shizhe, et al.
Published: (2026)
ViTaPEs: Visuotactile Position Encodings for Cross-Modal Alignment in Multimodal Transformers
by: Lygerakis, Fotios, et al.
Published: (2025)
by: Lygerakis, Fotios, et al.
Published: (2025)
Improving the Successful Robotic Grasp Detection Using Convolutional Neural Networks
by: Hosseini, Hamed, et al.
Published: (2024)
by: Hosseini, Hamed, et al.
Published: (2024)
Similar Items
-
Towards Open-World Human Action Segmentation Using Graph Convolutional Networks
by: Xing, Hao, et al.
Published: (2025) -
Understanding Human Activity with Uncertainty Measure for Novelty in Graph Convolutional Networks
by: Xing, Hao, et al.
Published: (2024) -
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
by: Xing, Hao, et al.
Published: (2024) -
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
by: Guo, Jianing, et al.
Published: (2025) -
Using The Concept Hierarchy for Household Action Recognition
by: Costinescu, Andrei, et al.
Published: (2024)