Hands-on Evaluation of Visual Transformers for Object Recognition and Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Vlachogiannis, Dimitrios N., Koutsomitropoulos, Dimitrios A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Review and Implementation of Object Detection Models and Optimizations for Real-time Medical Mask Detection during the COVID-19 Pandemic
by: Gogou, Ioanna, et al.
Published: (2024)
by: Gogou, Ioanna, et al.
Published: (2024)
Survey on Hand Gesture Recognition from Visual Input
by: Linardakis, Manousos, et al.
Published: (2025)
by: Linardakis, Manousos, et al.
Published: (2025)
Finding the Subjective Truth: Collecting 2 Million Votes for Comprehensive Gen-AI Model Evaluation
by: Christodoulou, Dimitrios, et al.
Published: (2024)
by: Christodoulou, Dimitrios, et al.
Published: (2024)
Dynamic Distinction Learning: Adaptive Pseudo Anomalies for Video Anomaly Detection
by: Lappas, Demetris, et al.
Published: (2024)
by: Lappas, Demetris, et al.
Published: (2024)
Object Detection Approaches to Identifying Hand Images with High Forensic Values
by: Nguyen, Thanh Thi, et al.
Published: (2024)
by: Nguyen, Thanh Thi, et al.
Published: (2024)
Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition
by: Alhadidi, Taqwa, et al.
Published: (2024)
by: Alhadidi, Taqwa, et al.
Published: (2024)
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)
by: Doumanoglou, Alexandros, et al.
Published: (2023)
Evaluating LLM -- Generated Multimodal Diagnosis from Medical Images and Symptom Analysis
by: Panagoulias, Dimitrios P., et al.
Published: (2024)
by: Panagoulias, Dimitrios P., et al.
Published: (2024)
Source-Free Object Detection with Detection Transformer
by: Yao, Huizai, et al.
Published: (2025)
by: Yao, Huizai, et al.
Published: (2025)
Knowledge Amalgamation for Object Detection with Transformers
by: Zhang, Haofei, et al.
Published: (2022)
by: Zhang, Haofei, et al.
Published: (2022)
Dynamic Object Queries for Transformer-based Incremental Object Detection
by: Zhang, Jichuan, et al.
Published: (2024)
by: Zhang, Jichuan, et al.
Published: (2024)
TransCAD: A Hierarchical Transformer for CAD Sequence Inference from Point Clouds
by: Dupont, Elona, et al.
Published: (2024)
by: Dupont, Elona, et al.
Published: (2024)
Multimodal and Multiview Deep Fusion for Autonomous Marine Navigation
by: Dagdilelis, Dimitrios, et al.
Published: (2025)
by: Dagdilelis, Dimitrios, et al.
Published: (2025)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
by: Tse, Tze Ho Elden, et al.
Published: (2025)
by: Tse, Tze Ho Elden, et al.
Published: (2025)
Object Detection for Vehicle Dashcams using Transformers
by: Mustafa, Osama, et al.
Published: (2024)
by: Mustafa, Osama, et al.
Published: (2024)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
by: Kumar, Yogesh, et al.
Published: (2025)
by: Kumar, Yogesh, et al.
Published: (2025)
DynaHOI: Benchmarking Hand-Object Interaction for Dynamic Target
by: Hu, BoCheng, et al.
Published: (2026)
by: Hu, BoCheng, et al.
Published: (2026)
Train, Test, Re-evaluate: Schedule-Sensitive Evaluation of Generative Data for Hand Detection
by: Bhardwaj, Atmika, et al.
Published: (2026)
by: Bhardwaj, Atmika, et al.
Published: (2026)
HaSPeR: An Image Repository for Hand Shadow Puppet Recognition
by: Raiyan, Syed Rifat, et al.
Published: (2024)
by: Raiyan, Syed Rifat, et al.
Published: (2024)
Deep Learning in Cardiology
by: Bizopoulos, Paschalis, et al.
Published: (2019)
by: Bizopoulos, Paschalis, et al.
Published: (2019)
Hybrid Spiking Vision Transformer for Object Detection with Event Cameras
by: Xu, Qi, et al.
Published: (2025)
by: Xu, Qi, et al.
Published: (2025)
Teacher-Student Model for Detecting and Classifying Mitosis in the MIDOG 2025 Challenge
by: Choe, Seungho, et al.
Published: (2025)
by: Choe, Seungho, et al.
Published: (2025)
Hand-Object Interaction Pretraining from Videos
by: Singh, Himanshu Gaurav, et al.
Published: (2024)
by: Singh, Himanshu Gaurav, et al.
Published: (2024)
Beyond Recognition: Evaluating Visual Perspective Taking in Vision Language Models
by: Góral, Gracjan, et al.
Published: (2025)
by: Góral, Gracjan, et al.
Published: (2025)
Alpha Divergence Losses for Biometric Verification
by: Koutsianos, Dimitrios, et al.
Published: (2025)
by: Koutsianos, Dimitrios, et al.
Published: (2025)
PGcGAN: Pathological Gait-Conditioned GAN for Human Gait Synthesis
by: Chandrasekaran, Mritula, et al.
Published: (2026)
by: Chandrasekaran, Mritula, et al.
Published: (2026)
ARS-DETR: Aspect Ratio-Sensitive Detection Transformer for Aerial Oriented Object Detection
by: Zeng, Ying, et al.
Published: (2023)
by: Zeng, Ying, et al.
Published: (2023)
Research on Detection of Floating Objects in River and Lake Based on AI Intelligent Image Recognition
by: Zhang, Jingyu, et al.
Published: (2024)
by: Zhang, Jingyu, et al.
Published: (2024)
Online Hand Gesture Recognition Using 3D Convolutional Neural Networks
by: Qin, Yinghao, et al.
Published: (2026)
by: Qin, Yinghao, et al.
Published: (2026)
TGBFormer: Transformer-GraphFormer Blender Network for Video Object Detection
by: Qi, Qiang, et al.
Published: (2025)
by: Qi, Qiang, et al.
Published: (2025)
BOOTPLACE: Bootstrapped Object Placement with Detection Transformers
by: Zhou, Hang, et al.
Published: (2025)
by: Zhou, Hang, et al.
Published: (2025)
Learning Disentangled Representation in Object-Centric Models for Visual Dynamics Prediction via Transformers
by: Gandhi, Sanket, et al.
Published: (2024)
by: Gandhi, Sanket, et al.
Published: (2024)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
by: Nguyen, Hung Huy, et al.
Published: (2025)
by: Nguyen, Hung Huy, et al.
Published: (2025)
Visual Accommodation: Rethinking Image Scale as a Learnable Variable for Object Detection
by: Seo, Daeun, et al.
Published: (2024)
by: Seo, Daeun, et al.
Published: (2024)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
by: Li, Wenxi, et al.
Published: (2025)
by: Li, Wenxi, et al.
Published: (2025)
Real-Time Indoor Object Detection based on hybrid CNN-Transformer Approach
by: Laidoudi, Salah Eddine, et al.
Published: (2024)
by: Laidoudi, Salah Eddine, et al.
Published: (2024)
Uncertainty Quantification in Detection Transformers: Object-Level Calibration and Image-Level Reliability
by: Park, Young-Jin, et al.
Published: (2024)
by: Park, Young-Jin, et al.
Published: (2024)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
by: Messina, Nicola, et al.
Published: (2025)
by: Messina, Nicola, et al.
Published: (2025)
Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities
by: Mobbs, Rebecca, et al.
Published: (2025)
by: Mobbs, Rebecca, et al.
Published: (2025)
FMM-X3D: FPGA-based modeling and mapping of X3D for Human Action Recognition
by: Toupas, Petros, et al.
Published: (2023)
by: Toupas, Petros, et al.
Published: (2023)
Similar Items
-
A Review and Implementation of Object Detection Models and Optimizations for Real-time Medical Mask Detection during the COVID-19 Pandemic
by: Gogou, Ioanna, et al.
Published: (2024) -
Survey on Hand Gesture Recognition from Visual Input
by: Linardakis, Manousos, et al.
Published: (2025) -
Finding the Subjective Truth: Collecting 2 Million Votes for Comprehensive Gen-AI Model Evaluation
by: Christodoulou, Dimitrios, et al.
Published: (2024) -
Dynamic Distinction Learning: Adaptive Pseudo Anomalies for Video Anomaly Detection
by: Lappas, Demetris, et al.
Published: (2024) -
Object Detection Approaches to Identifying Hand Images with High Forensic Values
by: Nguyen, Thanh Thi, et al.
Published: (2024)