Human Action Recognition in Still Images Using ConViT
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hosseyni, Seyed Rohollah, Seyedin, Sanaz, Taheri, Hasan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024)
Robust Ship Detection and Tracking Using Modified ViBe and Backwash Cancellation Algorithm
von: Saghafi, Mohammad Hassan, et al.
Veröffentlicht: (2026)
von: Saghafi, Mohammad Hassan, et al.
Veröffentlicht: (2026)
Leveraging Vision-Language Pre-training for Human Activity Recognition in Still Images
von: Mahanta, Cristina, et al.
Veröffentlicht: (2025)
von: Mahanta, Cristina, et al.
Veröffentlicht: (2025)
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer
von: Deressa, Deressa Wodajo, et al.
Veröffentlicht: (2023)
von: Deressa, Deressa Wodajo, et al.
Veröffentlicht: (2023)
MMeViT: Multi-Modal ensemble ViT for Post-Stroke Rehabilitation Action Recognition
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
Predicting Soccer Penalty Kick Direction Using Human Action Recognition
von: Freire-Obregón, David, et al.
Veröffentlicht: (2025)
von: Freire-Obregón, David, et al.
Veröffentlicht: (2025)
Continuous Sign Language Recognition Using Intra-inter Gloss Attention
von: Ranjbar, Hossein, et al.
Veröffentlicht: (2024)
von: Ranjbar, Hossein, et al.
Veröffentlicht: (2024)
ConViS-Bench: Estimating Video Similarity Through Semantic Concepts
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2025)
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2025)
Human Action Recognition without Human
von: Kataoka, Hirokatsu, et al.
Veröffentlicht: (2016)
von: Kataoka, Hirokatsu, et al.
Veröffentlicht: (2016)
ViConEx-Med: Visual Concept Explainability via Multi-Concept Token Transformer for Medical Image Analysis
von: Patrício, Cristiano, et al.
Veröffentlicht: (2025)
von: Patrício, Cristiano, et al.
Veröffentlicht: (2025)
ENet-21: An Optimized light CNN Structure for Lane Detection
von: Hosseini, Seyed Rasoul, et al.
Veröffentlicht: (2024)
von: Hosseini, Seyed Rasoul, et al.
Veröffentlicht: (2024)
Explore Human Parsing Modality for Action Recognition
von: Liu, Jinfu, et al.
Veröffentlicht: (2024)
von: Liu, Jinfu, et al.
Veröffentlicht: (2024)
ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
Human-Centric Transformer for Domain Adaptive Action Recognition
von: Lin, Kun-Yu, et al.
Veröffentlicht: (2024)
von: Lin, Kun-Yu, et al.
Veröffentlicht: (2024)
Active Generation Network of Human Skeleton for Action Recognition
von: Liu, Long, et al.
Veröffentlicht: (2024)
von: Liu, Long, et al.
Veröffentlicht: (2024)
Domain Generalization for Improved Human Activity Recognition in Office Space Videos Using Adaptive Pre-processing
von: Ghosh, Partho, et al.
Veröffentlicht: (2025)
von: Ghosh, Partho, et al.
Veröffentlicht: (2025)
Improvement of Human-Object Interaction Action Recognition Using Scene Information and Multi-Task Learning Approach
von: Shehata, Hesham M., et al.
Veröffentlicht: (2025)
von: Shehata, Hesham M., et al.
Veröffentlicht: (2025)
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
Human Action Recognition from Point Clouds over Time
von: Dickens, James
Veröffentlicht: (2025)
von: Dickens, James
Veröffentlicht: (2025)
SITAR: Semi-supervised Image Transformer for Action Recognition
von: Iqbal, Owais, et al.
Veröffentlicht: (2024)
von: Iqbal, Owais, et al.
Veröffentlicht: (2024)
ViTALS: Vision Transformer for Action Localization in Surgical Nephrectomy
von: Chandra, Soumyadeep, et al.
Veröffentlicht: (2024)
von: Chandra, Soumyadeep, et al.
Veröffentlicht: (2024)
Human Action Recognition (HAR) Using Skeleton-based Spatial Temporal Relative Transformer Network: ST-RTR
von: Mehmood, Faisal, et al.
Veröffentlicht: (2024)
von: Mehmood, Faisal, et al.
Veröffentlicht: (2024)
A Real-Time Human Action Recognition Model for Assisted Living
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
von: Wang, Yixuan, et al.
Veröffentlicht: (2025)
HabitAction: A Video Dataset for Human Habitual Behavior Recognition
von: Li, Hongwu, et al.
Veröffentlicht: (2024)
von: Li, Hongwu, et al.
Veröffentlicht: (2024)
YOWOv3: An Efficient and Generalized Framework for Human Action Detection and Recognition
von: Dang, Duc Manh Nguyen, et al.
Veröffentlicht: (2024)
von: Dang, Duc Manh Nguyen, et al.
Veröffentlicht: (2024)
SHARDeg: A Benchmark for Skeletal Human Action Recognition in Degraded Scenarios
von: Malzard, Simon, et al.
Veröffentlicht: (2025)
von: Malzard, Simon, et al.
Veröffentlicht: (2025)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
von: Kiray, Mert, et al.
Veröffentlicht: (2025)
von: Kiray, Mert, et al.
Veröffentlicht: (2025)
Towards Adaptive Fusion of Multimodal Deep Networks for Human Action Recognition
von: Yudistira, Novanto
Veröffentlicht: (2025)
von: Yudistira, Novanto
Veröffentlicht: (2025)
Video Domain Incremental Learning for Human Action Recognition in Home Environments
von: Hu, Yuanda, et al.
Veröffentlicht: (2024)
von: Hu, Yuanda, et al.
Veröffentlicht: (2024)
The Impact of the Single-Label Assumption in Image Recognition Benchmarking
von: Anzaku, Esla Timothy, et al.
Veröffentlicht: (2024)
von: Anzaku, Esla Timothy, et al.
Veröffentlicht: (2024)
Action Recognition Using Temporal Shift Module and Ensemble Learning
von: Duong, Anh-Kiet, et al.
Veröffentlicht: (2025)
von: Duong, Anh-Kiet, et al.
Veröffentlicht: (2025)
bViT: Investigating Single-Block Recurrence in Vision Transformers for Image Recognition
von: Byra, Michal, et al.
Veröffentlicht: (2026)
von: Byra, Michal, et al.
Veröffentlicht: (2026)
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
von: Liu, Mengyuan, et al.
Veröffentlicht: (2023)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2023)
SBF: An Effective Representation to Augment Skeleton for Video-based Human Action Recognition
von: Peng, Zhuoxuan, et al.
Veröffentlicht: (2026)
von: Peng, Zhuoxuan, et al.
Veröffentlicht: (2026)
Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?
von: Xie, Jianyang, et al.
Veröffentlicht: (2025)
von: Xie, Jianyang, et al.
Veröffentlicht: (2025)
Patch as Node: Human-Centric Graph Representation Learning for Multimodal Action Recognition
von: Liang, Zeyu, et al.
Veröffentlicht: (2025)
von: Liang, Zeyu, et al.
Veröffentlicht: (2025)
Advancing Human Action Recognition with Foundation Models trained on Unlabeled Public Videos
von: Qian, Yang, et al.
Veröffentlicht: (2024)
von: Qian, Yang, et al.
Veröffentlicht: (2024)
ViSTec: Video Modeling for Sports Technique Recognition and Tactical Analysis
von: He, Yuchen, et al.
Veröffentlicht: (2024)
von: He, Yuchen, et al.
Veröffentlicht: (2024)
LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment: Methods and Results
von: Li, Xin, et al.
Veröffentlicht: (2026)
von: Li, Xin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation
von: Hosseyni, S. Rohollah, et al.
Veröffentlicht: (2024) -
Robust Ship Detection and Tracking Using Modified ViBe and Backwash Cancellation Algorithm
von: Saghafi, Mohammad Hassan, et al.
Veröffentlicht: (2026) -
Leveraging Vision-Language Pre-training for Human Activity Recognition in Still Images
von: Mahanta, Cristina, et al.
Veröffentlicht: (2025) -
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer
von: Deressa, Deressa Wodajo, et al.
Veröffentlicht: (2023) -
MMeViT: Multi-Modal ensemble ViT for Post-Stroke Rehabilitation Action Recognition
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)