Towards Open-World Human Action Segmentation Using Graph Convolutional Networks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Xing, Hao, Boey, Kai Zhe, Cheng, Gordon |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation
par: Xing, Hao, et autres
Publié: (2025)
par: Xing, Hao, et autres
Publié: (2025)
Understanding Human Activity with Uncertainty Measure for Novelty in Graph Convolutional Networks
par: Xing, Hao, et autres
Publié: (2024)
par: Xing, Hao, et autres
Publié: (2024)
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
par: Xing, Hao, et autres
Publié: (2024)
par: Xing, Hao, et autres
Publié: (2024)
Open-World Panoptic Segmentation
par: Sodano, Matteo, et autres
Publié: (2024)
par: Sodano, Matteo, et autres
Publié: (2024)
OpenSGA: Efficient 3D Scene Graph Alignment in the Open World
par: Chen, Gang, et autres
Publié: (2026)
par: Chen, Gang, et autres
Publié: (2026)
Towards Open-World Grasping with Large Vision-Language Models
par: Tziafas, Georgios, et autres
Publié: (2024)
par: Tziafas, Georgios, et autres
Publié: (2024)
Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models
par: Ruan, Bo-Kai, et autres
Publié: (2026)
par: Ruan, Bo-Kai, et autres
Publié: (2026)
MWM: Mobile World Models for Action-Conditioned Consistent Prediction
par: Yan, Han, et autres
Publié: (2026)
par: Yan, Han, et autres
Publié: (2026)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
par: Yin, Zhenhan, et autres
Publié: (2025)
par: Yin, Zhenhan, et autres
Publié: (2025)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
par: Bhattacharya, Uttaran, et autres
Publié: (2019)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
par: Wang, Linbo, et autres
Publié: (2026)
par: Wang, Linbo, et autres
Publié: (2026)
Improving the Successful Robotic Grasp Detection Using Convolutional Neural Networks
par: Hosseini, Hamed, et autres
Publié: (2024)
par: Hosseini, Hamed, et autres
Publié: (2024)
Vision-based Manipulation from Single Human Video with Open-World Object Graphs
par: Zhu, Yifeng, et autres
Publié: (2024)
par: Zhu, Yifeng, et autres
Publié: (2024)
Map-World: Masked Action planning and Path-Integral World Model for Autonomous Driving
par: Hu, Bin, et autres
Publié: (2025)
par: Hu, Bin, et autres
Publié: (2025)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
par: jia, Feiyang, et autres
Publié: (2026)
par: jia, Feiyang, et autres
Publié: (2026)
Accelerated Rotation-Invariant Convolution for UAV Image Segmentation
par: Manduhu, Manduhu, et autres
Publié: (2025)
par: Manduhu, Manduhu, et autres
Publié: (2025)
Virtual Community: An Open World for Humans, Robots, and Society
par: Zhou, Qinhong, et autres
Publié: (2025)
par: Zhou, Qinhong, et autres
Publié: (2025)
From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes
par: Zhang, Qifan, et autres
Publié: (2026)
par: Zhang, Qifan, et autres
Publié: (2026)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
par: Li, Yi, et autres
Publié: (2025)
par: Li, Yi, et autres
Publié: (2025)
World Guidance: World Modeling in Condition Space for Action Generation
par: Su, Yue, et autres
Publié: (2026)
par: Su, Yue, et autres
Publié: (2026)
Friends Across Time: Multi-Scale Action Segmentation Transformer for Surgical Phase Recognition
par: Zhang, Bokai, et autres
Publié: (2024)
par: Zhang, Bokai, et autres
Publié: (2024)
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures
par: Belal, Mohammad, et autres
Publié: (2024)
par: Belal, Mohammad, et autres
Publié: (2024)
MS-TCRNet: Multi-Stage Temporal Convolutional Recurrent Networks for Action Segmentation Using Sensor-Augmented Kinematics
par: Goldbraikh, Adam, et autres
Publié: (2023)
par: Goldbraikh, Adam, et autres
Publié: (2023)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
par: Yenamandra, Sriram, et autres
Publié: (2024)
par: Yenamandra, Sriram, et autres
Publié: (2024)
Latent Action Pretraining Through World Modeling
par: Tharwat, Bahey, et autres
Publié: (2025)
par: Tharwat, Bahey, et autres
Publié: (2025)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
par: Chai, Ying, et autres
Publié: (2025)
par: Chai, Ying, et autres
Publié: (2025)
3D-CDRGP: Towards Cross-Device Robotic Grasping Policy in 3D Open World
par: Zhao, Weiguang, et autres
Publié: (2024)
par: Zhao, Weiguang, et autres
Publié: (2024)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
par: Zhao, Baining, et autres
Publié: (2026)
par: Zhao, Baining, et autres
Publié: (2026)
ODYSSEY: Open-World Quadrupeds Exploration and Manipulation for Long-Horizon Tasks
par: Wang, Kaijun, et autres
Publié: (2025)
par: Wang, Kaijun, et autres
Publié: (2025)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
par: Wu, Yanmin, et autres
Publié: (2024)
par: Wu, Yanmin, et autres
Publié: (2024)
Galaxea Open-World Dataset and G0 Dual-System VLA Model
par: Jiang, Tao, et autres
Publié: (2025)
par: Jiang, Tao, et autres
Publié: (2025)
ACT-Bench: Towards Action Controllable World Models for Autonomous Driving
par: Arai, Hidehisa, et autres
Publié: (2024)
par: Arai, Hidehisa, et autres
Publié: (2024)
CyberDemo: Augmenting Simulated Human Demonstration for Real-World Dexterous Manipulation
par: Wang, Jun, et autres
Publié: (2024)
par: Wang, Jun, et autres
Publié: (2024)
A Step Toward World Models: A Survey on Robotic Manipulation
par: Zhang, Peng-Fei, et autres
Publié: (2025)
par: Zhang, Peng-Fei, et autres
Publié: (2025)
Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy
par: Dai, Jiaheng, et autres
Publié: (2026)
par: Dai, Jiaheng, et autres
Publié: (2026)
WoW: Towards a World omniscient World model Through Embodied Interaction
par: Chi, Xiaowei, et autres
Publié: (2025)
par: Chi, Xiaowei, et autres
Publié: (2025)
FASTer: Toward Efficient Autoregressive Vision Language Action Modeling via Neural Action Tokenization
par: Liu, Yicheng, et autres
Publié: (2025)
par: Liu, Yicheng, et autres
Publié: (2025)
DynoSLAM: Dynamic SLAM with Generative Graph Neural Networks for Real-World Social Navigation
par: Tokhchukov, Danil, et autres
Publié: (2026)
par: Tokhchukov, Danil, et autres
Publié: (2026)
Taxonomy-Aware Continual Semantic Segmentation in Hyperbolic Spaces for Open-World Perception
par: Hindel, Julia, et autres
Publié: (2024)
par: Hindel, Julia, et autres
Publié: (2024)
SNOW: Spatio-Temporal Scene Understanding with World Knowledge for Open-World Embodied Reasoning
par: Sohn, Tin Stribor, et autres
Publié: (2025)
par: Sohn, Tin Stribor, et autres
Publié: (2025)
Documents similaires
-
Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation
par: Xing, Hao, et autres
Publié: (2025) -
Understanding Human Activity with Uncertainty Measure for Novelty in Graph Convolutional Networks
par: Xing, Hao, et autres
Publié: (2024) -
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
par: Xing, Hao, et autres
Publié: (2024) -
Open-World Panoptic Segmentation
par: Sodano, Matteo, et autres
Publié: (2024) -
OpenSGA: Efficient 3D Scene Graph Alignment in the Open World
par: Chen, Gang, et autres
Publié: (2026)