SiT-MLP: A Simple MLP with Point-wise Topology Feature Learning for Skeleton-based Action Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shaojie, Yin, Jianqin, Dang, Yonghao, Fu, Jiajun |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Generically Contrastive Spatiotemporal Representation Enhancement for 3D Skeleton Action Recognition
by: Zhang, Shaojie, et al.
Published: (2023)
by: Zhang, Shaojie, et al.
Published: (2023)
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation
by: Wei, Wei, et al.
Published: (2025)
by: Wei, Wei, et al.
Published: (2025)
Kinematics Modeling Network for Video-based Human Pose Estimation
by: Dang, Yonghao, et al.
Published: (2022)
by: Dang, Yonghao, et al.
Published: (2022)
ActivityCLIP: Enhancing Group Activity Recognition by Mining Complementary Information from Text to Supplement Image Modality
by: Xu, Guoliang, et al.
Published: (2024)
by: Xu, Guoliang, et al.
Published: (2024)
A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition
by: Yin, Ruoqi, et al.
Published: (2023)
by: Yin, Ruoqi, et al.
Published: (2023)
MamKPD: A Simple Mamba Baseline for Real-Time 2D Keypoint Detection
by: Dang, Yonghao, et al.
Published: (2024)
by: Dang, Yonghao, et al.
Published: (2024)
SpiralMLP: A Lightweight Vision MLP Architecture
by: Mu, Haojie, et al.
Published: (2024)
by: Mu, Haojie, et al.
Published: (2024)
evMLP: An Efficient Event-Driven MLP Architecture for Vision
by: Zheng, Zhentan
Published: (2025)
by: Zheng, Zhentan
Published: (2025)
Leveraging the Video-level Semantic Consistency of Event for Audio-visual Event Localization
by: Jiang, Yuanyuan, et al.
Published: (2022)
by: Jiang, Yuanyuan, et al.
Published: (2022)
Physics-constrained Attack against Convolution-based Human Motion Prediction
by: Duan, Chengxu, et al.
Published: (2023)
by: Duan, Chengxu, et al.
Published: (2023)
Lateralization MLP: A Simple Brain-inspired Architecture for Diffusion
by: Hu, Zizhao, et al.
Published: (2024)
by: Hu, Zizhao, et al.
Published: (2024)
KAN or MLP? Point Cloud Shows the Way Forward
by: Shi, Yan, et al.
Published: (2025)
by: Shi, Yan, et al.
Published: (2025)
PointMT: Efficient Point Cloud Analysis with Hybrid MLP-Transformer Architecture
by: Zheng, Qiang, et al.
Published: (2024)
by: Zheng, Qiang, et al.
Published: (2024)
Variational Contrastive Learning for Skeleton-based Action Recognition
by: Nguyen, Dang Dinh, et al.
Published: (2026)
by: Nguyen, Dang Dinh, et al.
Published: (2026)
An Efficient MLP-based Point-guided Segmentation Network for Ore Images with Ambiguous Boundary
by: Sun, Guodong, et al.
Published: (2024)
by: Sun, Guodong, et al.
Published: (2024)
FW-GAN: Frequency-Driven Handwriting Synthesis with Wave-Modulated MLP Generator
by: Khoa, Huynh Tong Dang, et al.
Published: (2025)
by: Khoa, Huynh Tong Dang, et al.
Published: (2025)
D2-MLP: Dynamic Decomposed MLP Mixer for Medical Image Segmentation
by: Yang, Jin, et al.
Published: (2024)
by: Yang, Jin, et al.
Published: (2024)
A Sobel-Gradient MLP Baseline for Handwritten Character Recognition
by: Nouri, Azam
Published: (2025)
by: Nouri, Azam
Published: (2025)
Expressive Keypoints for Skeleton-based Action Recognition via Skeleton Transformation
by: Yang, Yijie, et al.
Published: (2024)
by: Yang, Yijie, et al.
Published: (2024)
ChebMixer: Efficient Graph Representation Learning with MLP Mixer
by: Kui, Xiaoyan, et al.
Published: (2024)
by: Kui, Xiaoyan, et al.
Published: (2024)
MLP Can Be A Good Transformer Learner
by: Lin, Sihao, et al.
Published: (2024)
by: Lin, Sihao, et al.
Published: (2024)
Adaptive MLP Pruning for Large Vision Transformers
by: Shen, Chengchao
Published: (2026)
by: Shen, Chengchao
Published: (2026)
An MLP Baseline for Handwriting Recognition Using Planar Curvature and Gradient Orientation
by: Nouri, Azam
Published: (2025)
by: Nouri, Azam
Published: (2025)
ESG-Net: Event-Aware Semantic Guided Network for Dense Audio-Visual Event Localization
by: Li, Huilai, et al.
Published: (2025)
by: Li, Huilai, et al.
Published: (2025)
ResDiT: Evoking the Intrinsic Resolution Scalability in Diffusion Transformers
by: Ma, Yiyang, et al.
Published: (2025)
by: Ma, Yiyang, et al.
Published: (2025)
PosMLP-Video: Spatial and Temporal Relative Position Encoding for Efficient Video Recognition
by: Hao, Yanbin, et al.
Published: (2024)
by: Hao, Yanbin, et al.
Published: (2024)
3DGAA: Realistic and Robust 3D Gaussian-based Adversarial Attack for Autonomous Driving
by: Zhang, Yixun, et al.
Published: (2025)
by: Zhang, Yixun, et al.
Published: (2025)
SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
by: Ma, Nanye, et al.
Published: (2024)
by: Ma, Nanye, et al.
Published: (2024)
L2HCount:Generalizing Crowd Counting from Low to High Crowd Density via Density Simulation
by: Xu, Guoliang, et al.
Published: (2025)
by: Xu, Guoliang, et al.
Published: (2025)
SkeletonContext: Skeleton-side Context Prompt Learning for Zero-Shot Skeleton-based Action Recognition
by: Wang, Ning, et al.
Published: (2026)
by: Wang, Ning, et al.
Published: (2026)
HDBN: A Novel Hybrid Dual-branch Network for Robust Skeleton-based Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Are GNNs Worth the Effort for IoT Botnet Detection? A Comparative Study of VAE-GNN vs. ViT-MLP and VAE-MLP Approaches
by: Wasswa, Hassan, et al.
Published: (2025)
by: Wasswa, Hassan, et al.
Published: (2025)
SkeletonAgent: An Agentic Interaction Framework for Skeleton-based Action Recognition
by: Liu, Hongda, et al.
Published: (2025)
by: Liu, Hongda, et al.
Published: (2025)
GraphMLP: A Graph MLP-Like Architecture for 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2022)
by: Li, Wenhao, et al.
Published: (2022)
3D Skeleton-Based Action Recognition: A Review
by: Liu, Mengyuan, et al.
Published: (2025)
by: Liu, Mengyuan, et al.
Published: (2025)
PointNet with KAN versus PointNet with MLP for 3D Classification and Segmentation of Point Sets
by: Kashefi, Ali
Published: (2024)
by: Kashefi, Ali
Published: (2024)
DHRNet: A Dual-Path Hierarchical Relation Network for Multi-Person Pose Estimation
by: Dang, Yonghao, et al.
Published: (2024)
by: Dang, Yonghao, et al.
Published: (2024)
CMIP-CIL: A Cross-Modal Benchmark for Image-Point Class Incremental Learning
by: Qi, Chao, et al.
Published: (2025)
by: Qi, Chao, et al.
Published: (2025)
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation
by: Zhang, Zongye, et al.
Published: (2025)
by: Zhang, Zongye, et al.
Published: (2025)
Cross-DINO: Cross the Deep MLP and Transformer for Small Object Detection
by: Cao, Guiping, et al.
Published: (2025)
by: Cao, Guiping, et al.
Published: (2025)
Similar Items
-
A Generically Contrastive Spatiotemporal Representation Enhancement for 3D Skeleton Action Recognition
by: Zhang, Shaojie, et al.
Published: (2023) -
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation
by: Wei, Wei, et al.
Published: (2025) -
Kinematics Modeling Network for Video-based Human Pose Estimation
by: Dang, Yonghao, et al.
Published: (2022) -
ActivityCLIP: Enhancing Group Activity Recognition by Mining Complementary Information from Text to Supplement Image Modality
by: Xu, Guoliang, et al.
Published: (2024) -
A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition
by: Yin, Ruoqi, et al.
Published: (2023)