Saved in:
| Main Authors: | Shen, Xi, Gamboa, Julian, Hamidfar, Tabassom, Mitu, Shamima A., Shahriar, Selim M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.09939 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Shift, Scale and Rotation Invariant Multiple Object Detection using Balanced Joint Transform Correlator
by: Shen, Xi, et al.
Published: (2025)
by: Shen, Xi, et al.
Published: (2025)
Ultra-fast Real-time Target Recognition Using a Shift, Scale, and Rotation Invariant Hybrid Opto-electronic Joint Transform Correlator
by: Shen, Xi, et al.
Published: (2025)
by: Shen, Xi, et al.
Published: (2025)
Automatic Event Recognition Employing Optically Excited Electro-Nuclear Spin Coherence
by: Mitu, Shamima, et al.
Published: (2025)
by: Mitu, Shamima, et al.
Published: (2025)
Debiased Opto-electronic Joint Transform Correlator for Enhanced Real-Time Pattern Recognition
by: Gamboa, Julian, et al.
Published: (2025)
by: Gamboa, Julian, et al.
Published: (2025)
Opto-Atomic Spatio-Temporal Holographic Correlators for High-Speed 3D CNNs
by: Shen, Xi, et al.
Published: (2026)
by: Shen, Xi, et al.
Published: (2026)
HTR-VT: Handwritten Text Recognition with Vision Transformer
by: Li, Yuting, et al.
Published: (2024)
by: Li, Yuting, et al.
Published: (2024)
Action Recognition Using Temporal Shift Module and Ensemble Learning
by: Duong, Anh-Kiet, et al.
Published: (2025)
by: Duong, Anh-Kiet, et al.
Published: (2025)
Multi-Focus Temporal Shifting for Precise Event Spotting in Sports Videos
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
GATS: Gaussian Aware Temporal Scaling Transformer for Invariant 4D Spatio-Temporal Point Cloud Representation
by: Tian, Jiayi, et al.
Published: (2026)
by: Tian, Jiayi, et al.
Published: (2026)
Beyond Generation: Multi-Hop Reasoning for Factual Accuracy in Vision-Language Models
by: Hossain, Shamima
Published: (2025)
by: Hossain, Shamima
Published: (2025)
Learning Causal Domain-Invariant Temporal Dynamics for Few-Shot Action Recognition
by: Li, Yuke, et al.
Published: (2024)
by: Li, Yuke, et al.
Published: (2024)
Context-Aware Temporal Embedding of Objects in Video Data
by: Farhan, Ahnaf, et al.
Published: (2024)
by: Farhan, Ahnaf, et al.
Published: (2024)
Towards Large-Scale Pose-Invariant Face Recognition Using Face Defrontalization
by: Mesec, Patrik, et al.
Published: (2025)
by: Mesec, Patrik, et al.
Published: (2025)
MSSTNet: A Multi-Scale Spatio-Temporal CNN-Transformer Network for Dynamic Facial Expression Recognition
by: Wang, Linhuang, et al.
Published: (2024)
by: Wang, Linhuang, et al.
Published: (2024)
EventGait: Towards Robust Gait Recognition with Event Streams
by: Xu, Senyan, et al.
Published: (2026)
by: Xu, Senyan, et al.
Published: (2026)
Boosting Gesture Recognition with an Automatic Gesture Annotation Framework
by: Shen, Junxiao, et al.
Published: (2024)
by: Shen, Junxiao, et al.
Published: (2024)
WildIng: A Wildlife Image Invariant Representation Model for Geographical Domain Shift
by: Santamaria, Julian D., et al.
Published: (2026)
by: Santamaria, Julian D., et al.
Published: (2026)
A Critical Analysis on Machine Learning Techniques for Video-based Human Activity Recognition of Surveillance Systems: A Review
by: Jahan, Shahriar, et al.
Published: (2024)
by: Jahan, Shahriar, et al.
Published: (2024)
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024)
by: Do, Jeonghyeok, et al.
Published: (2024)
A Signer-Invariant Conformer and Multi-Scale Fusion Transformer for Continuous Sign Language Recognition
by: Haque, Md Rezwanul, et al.
Published: (2025)
by: Haque, Md Rezwanul, et al.
Published: (2025)
Boosting Continuous Emotion Recognition with Self-Pretraining using Masked Autoencoders, Temporal Convolutional Networks, and Transformers
by: Zhou, Weiwei, et al.
Published: (2024)
by: Zhou, Weiwei, et al.
Published: (2024)
TRACE: Temporal Grounding Video LLM via Causal Event Modeling
by: Guo, Yongxin, et al.
Published: (2024)
by: Guo, Yongxin, et al.
Published: (2024)
Event Transformer
by: Jiang, Bin, et al.
Published: (2022)
by: Jiang, Bin, et al.
Published: (2022)
Three-Stream Temporal-Shift Attention Network Based on Self-Knowledge Distillation for Micro-Expression Recognition
by: Zhu, Guanghao, et al.
Published: (2024)
by: Zhu, Guanghao, et al.
Published: (2024)
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition
by: Yang, Shu, et al.
Published: (2024)
by: Yang, Shu, et al.
Published: (2024)
Flexible and Efficient Spatio-Temporal Transformer for Sequential Visual Place Recognition
by: Kiu, Yu, et al.
Published: (2025)
by: Kiu, Yu, et al.
Published: (2025)
Automated Attendee Recognition System for Large-Scale Social Events or Conference Gathering
by: Motwani, Dhruv, et al.
Published: (2025)
by: Motwani, Dhruv, et al.
Published: (2025)
STMT: A Spatial-Temporal Mesh Transformer for MoCap-Based Action Recognition
by: Zhu, Xiaoyu, et al.
Published: (2023)
by: Zhu, Xiaoyu, et al.
Published: (2023)
A Paradigm Shift in Mouza Map Vectorization: A Human-Machine Collaboration Approach
by: Dhrubo, Mahir Shahriar, et al.
Published: (2024)
by: Dhrubo, Mahir Shahriar, et al.
Published: (2024)
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
by: Yamane, Taiga, et al.
Published: (2025)
by: Yamane, Taiga, et al.
Published: (2025)
EventSleep: Sleep Activity Recognition with Event Cameras
by: Plou, Carlos, et al.
Published: (2024)
by: Plou, Carlos, et al.
Published: (2024)
Heterogeneous Face Recognition Using Domain Invariant Units
by: George, Anjith, et al.
Published: (2024)
by: George, Anjith, et al.
Published: (2024)
Multispectral Texture Synthesis using RGB Convolutional Neural Networks
by: Ollivier, Sélim, et al.
Published: (2024)
by: Ollivier, Sélim, et al.
Published: (2024)
Local Precise Refinement: A Dual-Gated Mixture-of-Experts for Enhancing Foundation Model Generalization against Spectral Shifts
by: Chen, Xi, et al.
Published: (2026)
by: Chen, Xi, et al.
Published: (2026)
MSGL-Transformer: A Multi-Scale Global-Local Transformer for Rodent Social Behavior Recognition
by: Sharif, Muhammad Imran, et al.
Published: (2026)
by: Sharif, Muhammad Imran, et al.
Published: (2026)
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
NYC-Event-VPR: A Large-Scale High-Resolution Event-Based Visual Place Recognition Dataset in Dense Urban Environments
by: Pan, Taiyi, et al.
Published: (2024)
by: Pan, Taiyi, et al.
Published: (2024)
Filter or Compensate: Towards Invariant Representation from Distribution Shift for Anomaly Detection
by: Chen, Zining, et al.
Published: (2024)
by: Chen, Zining, et al.
Published: (2024)
Accurate Shift Invariant Convolutional Neural Networks Using Gaussian-Hermite Moments
by: Singh, Jaspreet, et al.
Published: (2026)
by: Singh, Jaspreet, et al.
Published: (2026)
Human Action Recognition (HAR) Using Skeleton-based Spatial Temporal Relative Transformer Network: ST-RTR
by: Mehmood, Faisal, et al.
Published: (2024)
by: Mehmood, Faisal, et al.
Published: (2024)
Similar Items
-
Shift, Scale and Rotation Invariant Multiple Object Detection using Balanced Joint Transform Correlator
by: Shen, Xi, et al.
Published: (2025) -
Ultra-fast Real-time Target Recognition Using a Shift, Scale, and Rotation Invariant Hybrid Opto-electronic Joint Transform Correlator
by: Shen, Xi, et al.
Published: (2025) -
Automatic Event Recognition Employing Optically Excited Electro-Nuclear Spin Coherence
by: Mitu, Shamima, et al.
Published: (2025) -
Debiased Opto-electronic Joint Transform Correlator for Enhanced Real-Time Pattern Recognition
by: Gamboa, Julian, et al.
Published: (2025) -
Opto-Atomic Spatio-Temporal Holographic Correlators for High-Speed 3D CNNs
by: Shen, Xi, et al.
Published: (2026)