Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yiping, Peng, Kunyu, Roitberg, Alina, Zhang, Jiaming, Zheng, Junwei, Liu, Ruiping, Chen, Yufan, Yang, Kailun, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
Skeleton-Based Human Action Recognition with Noisy Labels
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
Referring Atomic Video Action Recognition
by: Peng, Kunyu, et al.
Published: (2024)
by: Peng, Kunyu, et al.
Published: (2024)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
by: Liu, Ruiping, et al.
Published: (2022)
by: Liu, Ruiping, et al.
Published: (2022)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
by: Peng, Kunyu, et al.
Published: (2023)
by: Peng, Kunyu, et al.
Published: (2023)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
by: Fan, Linjuan, et al.
Published: (2025)
by: Fan, Linjuan, et al.
Published: (2025)
Towards Activated Muscle Group Estimation in the Wild
by: Peng, Kunyu, et al.
Published: (2023)
by: Peng, Kunyu, et al.
Published: (2023)
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
by: Teng, Fei, et al.
Published: (2023)
by: Teng, Fei, et al.
Published: (2023)
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes
by: Zhai, Jiajun, et al.
Published: (2026)
by: Zhai, Jiajun, et al.
Published: (2026)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
by: Lin, Kaixin, et al.
Published: (2026)
by: Lin, Kaixin, et al.
Published: (2026)
CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
Occlusion-Aware Seamless Segmentation
by: Cao, Yihong, et al.
Published: (2024)
by: Cao, Yihong, et al.
Published: (2024)
Mitigating Label Noise using Prompt-Based Hyperbolic Meta-Learning in Open-Set Domain Generalization
by: Peng, Kunyu, et al.
Published: (2024)
by: Peng, Kunyu, et al.
Published: (2024)
Exploring Event-based Human Pose Estimation with 3D Event Representations
by: Yin, Xiaoting, et al.
Published: (2023)
by: Yin, Xiaoting, et al.
Published: (2023)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024)
by: Zeng, Kang, et al.
Published: (2024)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
by: Shi, Hao, et al.
Published: (2023)
by: Shi, Hao, et al.
Published: (2023)
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation
by: Zhang, Jiaming, et al.
Published: (2022)
by: Zhang, Jiaming, et al.
Published: (2022)
Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence
by: Li, Wenxin, et al.
Published: (2025)
by: Li, Wenxin, et al.
Published: (2025)
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
by: Lin, Jiacheng, et al.
Published: (2024)
by: Lin, Jiacheng, et al.
Published: (2024)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
InterEdit: Navigating Text-Guided Multi-Human 3D Motion Editing
by: Yang, Yebin, et al.
Published: (2026)
by: Yang, Yebin, et al.
Published: (2026)
Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation
by: Zheng, Yuanfan, et al.
Published: (2026)
by: Zheng, Yuanfan, et al.
Published: (2026)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
by: Zheng, Junwei, et al.
Published: (2023)
by: Zheng, Junwei, et al.
Published: (2023)
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
by: Chen, Xuyang, et al.
Published: (2025)
by: Chen, Xuyang, et al.
Published: (2025)
Adaptive Resolution and Chroma Subsampling for Energy-Efficient Video Coding
by: Premkumar, Amritha, et al.
Published: (2026)
by: Premkumar, Amritha, et al.
Published: (2026)
Efficient Sub-pixel Motion Compensation in Learned Video Codecs
by: Ladune, Théo, et al.
Published: (2025)
by: Ladune, Théo, et al.
Published: (2025)
AV1 Motion Vector Fidelity and Application for Efficient Optical Flow
by: Zouein, Julien, et al.
Published: (2025)
by: Zouein, Julien, et al.
Published: (2025)
Unlocking Constraints: Source-Free Occlusion-Aware Seamless Segmentation
by: Cao, Yihong, et al.
Published: (2025)
by: Cao, Yihong, et al.
Published: (2025)
LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras
by: Teng, Fei, et al.
Published: (2024)
by: Teng, Fei, et al.
Published: (2024)
PAL: Prompting Analytic Learning with Missing Modality for Multi-Modal Class-Incremental Learning
by: Yue, Xianghu, et al.
Published: (2025)
by: Yue, Xianghu, et al.
Published: (2025)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Out-of-Distribution Semantic Occupancy Prediction
by: Zhang, Yuheng, et al.
Published: (2025)
by: Zhang, Yuheng, et al.
Published: (2025)
Region-Adaptive Learned Hierarchical Encoding for 3D Gaussian Splatting Data
by: Sridhara, Shashank N., et al.
Published: (2025)
by: Sridhara, Shashank N., et al.
Published: (2025)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise
by: Li, Wenxin, et al.
Published: (2026)
by: Li, Wenxin, et al.
Published: (2026)
Transform and Entropy Coding in AV2
by: Nalci, Alican, et al.
Published: (2026)
by: Nalci, Alican, et al.
Published: (2026)
Similar Items
-
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023) -
Skeleton-Based Human Action Recognition with Noisy Labels
by: Xu, Yi, et al.
Published: (2024) -
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
by: Peng, Kunyu, et al.
Published: (2025) -
Referring Atomic Video Action Recognition
by: Peng, Kunyu, et al.
Published: (2024) -
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)