SMART-Vision: Survey of Modern Action Recognition Techniques in Vision
Fuente:
arXiv
Saved in:
| Main Authors: | AlShami, Ali K., Rabinowitz, Ryan, Lam, Khang, Shleibik, Yousra, Mersha, Melkamu, Boult, Terrance, Kalita, Jugal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pose2Trajectory: Using Transformers on Body Pose to Predict Tennis Player's Trajectory
by: AlShami, Ali K., et al.
Published: (2024)
by: AlShami, Ali K., et al.
Published: (2024)
Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
by: Mersha, Melkamu, et al.
Published: (2024)
by: Mersha, Melkamu, et al.
Published: (2024)
COOOL: Challenge Of Out-Of-Label A Novel Benchmark for Autonomous Driving
by: AlShami, Ali K., et al.
Published: (2024)
by: AlShami, Ali K., et al.
Published: (2024)
SGNetPose+: Stepwise Goal-Driven Networks with Pose Information for Trajectory Prediction in Autonomous Driving
by: Ghiya, Akshat, et al.
Published: (2025)
by: Ghiya, Akshat, et al.
Published: (2025)
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
by: Mersha, Melkamu Abay, et al.
Published: (2026)
by: Mersha, Melkamu Abay, et al.
Published: (2026)
A Unified Framework with Novel Metrics for Evaluating the Effectiveness of XAI Techniques in LLMs
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
ZoDIAC: Zoneout Dropout Injection Attention Calculation
by: Zohourianshahzadi, Zanyar, et al.
Published: (2022)
by: Zohourianshahzadi, Zanyar, et al.
Published: (2022)
COSTARR: Consolidated Open Set Technique with Attenuation for Robust Recognition
by: Rabinowitz, Ryan, et al.
Published: (2025)
by: Rabinowitz, Ryan, et al.
Published: (2025)
2COOOL: 2nd Workshop on the Challenge Of Out-Of-Label Hazards in Autonomous Driving
by: AlShami, Ali K., et al.
Published: (2025)
by: AlShami, Ali K., et al.
Published: (2025)
Adapting Feature Attenuation to NLP
by: Yang, Tianshuo, et al.
Published: (2026)
by: Yang, Tianshuo, et al.
Published: (2026)
Semantic-Driven Topic Modeling for Analyzing Creativity in Virtual Brainstorming
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
GHOST: Gaussian Hypothesis Open-Set Technique
by: Rabinowitz, Ryan, et al.
Published: (2025)
by: Rabinowitz, Ryan, et al.
Published: (2025)
Drug Repurposing Using Deep Embedded Clustering and Graph Neural Networks
by: Delzer, Luke, et al.
Published: (2025)
by: Delzer, Luke, et al.
Published: (2025)
Semantic-Driven Topic Modeling Using Transformer-Based Embeddings and Clustering Algorithms
by: Mersha, Melkamu Abay, et al.
Published: (2024)
by: Mersha, Melkamu Abay, et al.
Published: (2024)
Explainability in Neural Networks for Natural Language Processing Tasks
by: Mersha, Melkamu, et al.
Published: (2024)
by: Mersha, Melkamu, et al.
Published: (2024)
Toxicity in Online Platforms and AI Systems: A Survey of Needs, Challenges, Mitigations, and Future Directions
by: Khapre, Smita, et al.
Published: (2025)
by: Khapre, Smita, et al.
Published: (2025)
Deep Multi-Task Learning for Malware Image Classification
by: Bensaoud, Ahmed, et al.
Published: (2024)
by: Bensaoud, Ahmed, et al.
Published: (2024)
Ethio-Fake: Cutting-Edge Approaches to Combat Fake News in Under-Resourced Languages Using Explainable AI
by: Yigezu, Mesay Gemeda, et al.
Published: (2024)
by: Yigezu, Mesay Gemeda, et al.
Published: (2024)
Towards Human-Robot Teaming through Augmented Reality and Gaze-Based Attention Control
by: Shleibik, Yousra, et al.
Published: (2024)
by: Shleibik, Yousra, et al.
Published: (2024)
Saliency-Based Attention Shifting: A Framework for Improving Driver Situational Awareness of Out-of-Label Hazards
by: Shleibik, Yousra, et al.
Published: (2025)
by: Shleibik, Yousra, et al.
Published: (2025)
Explainable AI: XAI-Guided Context-Aware Data Augmentation
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
Action-Item-Driven Summarization of Long Meeting Transcripts
by: Golia, Logan, et al.
Published: (2023)
by: Golia, Logan, et al.
Published: (2023)
Survey on Vision-Language-Action Models
by: Adilkhanov, Adilzhan, et al.
Published: (2025)
by: Adilkhanov, Adilzhan, et al.
Published: (2025)
A Survey of Medical Vision-and-Language Applications and Their Techniques
by: Chen, Qi, et al.
Published: (2024)
by: Chen, Qi, et al.
Published: (2024)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
A Survey on Semantic Communication for Vision: Categories, Frameworks, Enabling Techniques, and Applications
by: Cheng, Runze, et al.
Published: (2026)
by: Cheng, Runze, et al.
Published: (2026)
Conformal Predictions for Human Action Recognition with Vision-Language Models
by: Tim, Bary, et al.
Published: (2025)
by: Tim, Bary, et al.
Published: (2025)
Image Recognition with Online Lightweight Vision Transformer: A Survey
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
Survey of Quantization Techniques for On-Device Vision-based Crack Detection
by: Zhang, Yuxuan, et al.
Published: (2025)
by: Zhang, Yuxuan, et al.
Published: (2025)
A Survey of Malware Detection Using Deep Learning
by: Bensaoud, Ahmed, et al.
Published: (2024)
by: Bensaoud, Ahmed, et al.
Published: (2024)
Analysis of Modern Computer Vision Models for Blood Cell Classification
by: Kim, Alexander, et al.
Published: (2024)
by: Kim, Alexander, et al.
Published: (2024)
Computer Vision Model Compression Techniques for Embedded Systems: A Survey
by: Lopes, Alexandre, et al.
Published: (2024)
by: Lopes, Alexandre, et al.
Published: (2024)
A Survey on Efficient Vision-Language-Action Models
by: Yu, Zhaoshu, et al.
Published: (2025)
by: Yu, Zhaoshu, et al.
Published: (2025)
Mamba in Vision: A Comprehensive Survey of Techniques and Applications
by: Rahman, Md Maklachur, et al.
Published: (2024)
by: Rahman, Md Maklachur, et al.
Published: (2024)
A Survey on Vision-Language-Action Models for Autonomous Driving
by: Jiang, Sicong, et al.
Published: (2025)
by: Jiang, Sicong, et al.
Published: (2025)
PVG: Progressive Vision Graph for Vision Recognition
by: Wu, Jiafu, et al.
Published: (2023)
by: Wu, Jiafu, et al.
Published: (2023)
Vision Mamba in Remote Sensing: A Comprehensive Survey of Techniques, Applications and Outlook
by: Bao, Muyi, et al.
Published: (2025)
by: Bao, Muyi, et al.
Published: (2025)
A Vision-Language Foundation Model for Leaf Disease Identification
by: Quoc, Khang Nguyen, et al.
Published: (2025)
by: Quoc, Khang Nguyen, et al.
Published: (2025)
TFLOP: Table Structure Recognition Framework with Layout Pointer Mechanism
by: Khang, Minsoo, et al.
Published: (2025)
by: Khang, Minsoo, et al.
Published: (2025)
Similar Items
-
Pose2Trajectory: Using Transformers on Body Pose to Predict Tennis Player's Trajectory
by: AlShami, Ali K., et al.
Published: (2024) -
Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
by: Mersha, Melkamu, et al.
Published: (2024) -
COOOL: Challenge Of Out-Of-Label A Novel Benchmark for Autonomous Driving
by: AlShami, Ali K., et al.
Published: (2024) -
SGNetPose+: Stepwise Goal-Driven Networks with Pose Information for Trajectory Prediction in Autonomous Driving
by: Ghiya, Akshat, et al.
Published: (2025) -
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
by: Mersha, Melkamu Abay, et al.
Published: (2026)