Guardado en:
| Autores principales: | Kerkouri, Mohamed Amine, Tliba, Marouane, Chetouani, Aladine, Bruno, Alessandro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.22049 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2026)
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2026)
Quantization Effects on Neural Networks Perception: How would quantization change the perceptual field of vision models?
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2024)
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2024)
Modeling Beyond MOS: Quality Assessment Models Must Integrate Context, Reasoning, and Multimodality
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2025)
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2025)
Shifting Focus: From Global Semantics to Local Prominent Features in Swin-Transformer for Knee Osteoarthritis Severity Assessment
por: Sekhri, Aymen, et al.
Publicado: (2024)
por: Sekhri, Aymen, et al.
Publicado: (2024)
Morphology-Aware KOA Classification: Integrating Graph Priors with Vision Models
por: Tliba, Marouane, et al.
Publicado: (2025)
por: Tliba, Marouane, et al.
Publicado: (2025)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
por: Wong, David C, et al.
Publicado: (2025)
por: Wong, David C, et al.
Publicado: (2025)
UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment
por: Wang, Zheng, et al.
Publicado: (2026)
por: Wang, Zheng, et al.
Publicado: (2026)
Multi-face emotion detection for effective Human-Robot Interaction
por: Yahyaoui, Mohamed Ala, et al.
Publicado: (2025)
por: Yahyaoui, Mohamed Ala, et al.
Publicado: (2025)
Several questions of visual generation in 2024
por: Gu, Shuyang
Publicado: (2024)
por: Gu, Shuyang
Publicado: (2024)
Real-Time Hand Gesture Recognition: Integrating Skeleton-Based Data Fusion and Multi-Stream CNN
por: Yusuf, Oluwaleke, et al.
Publicado: (2024)
por: Yusuf, Oluwaleke, et al.
Publicado: (2024)
CG-MER: A Card Game-based Multimodal dataset for Emotion Recognition
por: Farhat, Nessrine, et al.
Publicado: (2025)
por: Farhat, Nessrine, et al.
Publicado: (2025)
Category-aware EEG image generation based on wavelet transform and contrast semantic loss
por: Zhang, Enshang, et al.
Publicado: (2025)
por: Zhang, Enshang, et al.
Publicado: (2025)
CT-DegradBench: A Physics-Informed Benchmark for CT Degradation Detection and Severity Estimation
por: Taifour, Yousra Nabila, et al.
Publicado: (2026)
por: Taifour, Yousra Nabila, et al.
Publicado: (2026)
Cross-user activity recognition using deep domain adaptation with temporal relation information
por: Ye, Xiaozhou, et al.
Publicado: (2024)
por: Ye, Xiaozhou, et al.
Publicado: (2024)
GroundUp: Rapid Sketch-Based 3D City Massing
por: Unlu, Gizem Esra, et al.
Publicado: (2024)
por: Unlu, Gizem Esra, et al.
Publicado: (2024)
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
por: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Publicado: (2024)
por: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Publicado: (2024)
Accurate online action and gesture recognition system using detectors and Deep SPD Siamese Networks
por: Akremi, Mohamed Sanim, et al.
Publicado: (2025)
por: Akremi, Mohamed Sanim, et al.
Publicado: (2025)
Computer Vision for Objects used in Group Work: Challenges and Opportunities
por: Jung, Changsoo, et al.
Publicado: (2025)
por: Jung, Changsoo, et al.
Publicado: (2025)
Weak-Annotation of HAR Datasets using Vision Foundation Models
por: Bock, Marius, et al.
Publicado: (2024)
por: Bock, Marius, et al.
Publicado: (2024)
Generalized Pose Space Embeddings for Training In-the-Wild using Anaylis-by-Synthesis
por: Borer, Dominik, et al.
Publicado: (2024)
por: Borer, Dominik, et al.
Publicado: (2024)
CADDI: An in-Class Activity Detection Dataset using IMU data from low-cost sensors
por: Marquez-Carpintero, Luis, et al.
Publicado: (2025)
por: Marquez-Carpintero, Luis, et al.
Publicado: (2025)
Unsupervised learning of Data-driven Facial Expression Coding System (DFECS) using keypoint tracking
por: Tripathi, Shivansh Chandra, et al.
Publicado: (2024)
por: Tripathi, Shivansh Chandra, et al.
Publicado: (2024)
GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
por: Konrad, Robert, et al.
Publicado: (2024)
por: Konrad, Robert, et al.
Publicado: (2024)
Is Medieval Distant Viewing Possible? : Extending and Enriching Annotation of Legacy Image Collections using Visual Analytics
por: Meinecke, Christofer, et al.
Publicado: (2022)
por: Meinecke, Christofer, et al.
Publicado: (2022)
Implicit Search Intent Recognition using EEG and Eye Tracking: Novel Dataset and Cross-User Prediction
por: Sharma, Mansi, et al.
Publicado: (2025)
por: Sharma, Mansi, et al.
Publicado: (2025)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
por: Rocha, Tomás Silva Santos, et al.
Publicado: (2025)
por: Rocha, Tomás Silva Santos, et al.
Publicado: (2025)
Resource-Efficient Gesture Recognition using Low-Resolution Thermal Camera via Spiking Neural Networks and Sparse Segmentation
por: Safa, Ali, et al.
Publicado: (2024)
por: Safa, Ali, et al.
Publicado: (2024)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
por: Huang, Jinbin, et al.
Publicado: (2024)
por: Huang, Jinbin, et al.
Publicado: (2024)
How good are humans at detecting AI-generated images? Learnings from an experiment
por: Roca, Thomas, et al.
Publicado: (2025)
por: Roca, Thomas, et al.
Publicado: (2025)
Combining Transformers and CNNs for Efficient Object Detection in High-Resolution Satellite Imagery
por: Drapier, Nicolas, et al.
Publicado: (2025)
por: Drapier, Nicolas, et al.
Publicado: (2025)
The Visual Experience Dataset: Over 200 Recorded Hours of Integrated Eye Movement, Odometry, and Egocentric Video
por: Greene, Michelle R., et al.
Publicado: (2024)
por: Greene, Michelle R., et al.
Publicado: (2024)
CoCoG-2: Controllable generation of visual stimuli for understanding human concept representation
por: Wei, Chen, et al.
Publicado: (2024)
por: Wei, Chen, et al.
Publicado: (2024)
Accurate Eye Tracking from Dense 3D Surface Reconstructions using Single-Shot Deflectometry
por: Wang, Jiazhang, et al.
Publicado: (2023)
por: Wang, Jiazhang, et al.
Publicado: (2023)
Exploring Emotion Expression Recognition in Older Adults Interacting with a Virtual Coach
por: Palmero, Cristina, et al.
Publicado: (2023)
por: Palmero, Cristina, et al.
Publicado: (2023)
SCHEMA for Gemini 3 Pro Image: A Structured Methodology for Controlled AI Image Generation on Google's Native Multimodal Model
por: Cazzaniga, Luca
Publicado: (2026)
por: Cazzaniga, Luca
Publicado: (2026)
Viewpoint Recommendation for Point Cloud Labeling through Interaction Cost Modeling
por: Zhang, Yu, et al.
Publicado: (2026)
por: Zhang, Yu, et al.
Publicado: (2026)
SymbolSight: Minimizing Inter-Symbol Interference for Reading with Prosthetic Vision
por: Lesner, Jasmine, et al.
Publicado: (2026)
por: Lesner, Jasmine, et al.
Publicado: (2026)
Towards an End-to-End System for 3D Tracking of Physical Objects in Virtual Immersive Environments
por: Knapiński, Stanisław, et al.
Publicado: (2026)
por: Knapiński, Stanisław, et al.
Publicado: (2026)
MicroBi-ConvLSTM: An Ultra-Lightweight Efficient Model for Human Activity Recognition on Resource Constrained Devices
por: Mandal, Mridankan
Publicado: (2026)
por: Mandal, Mridankan
Publicado: (2026)
Real-Time Cellist Postural Evaluation With On-Device Computer Vision
por: Wang, Paolo, et al.
Publicado: (2026)
por: Wang, Paolo, et al.
Publicado: (2026)
Ejemplares similares
-
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2026) -
Quantization Effects on Neural Networks Perception: How would quantization change the perceptual field of vision models?
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2024) -
Modeling Beyond MOS: Quality Assessment Models Must Integrate Context, Reasoning, and Multimodality
por: Kerkouri, Mohamed Amine, et al.
Publicado: (2025) -
Shifting Focus: From Global Semantics to Local Prominent Features in Swin-Transformer for Knee Osteoarthritis Severity Assessment
por: Sekhri, Aymen, et al.
Publicado: (2024) -
Morphology-Aware KOA Classification: Integrating Graph Priors with Vision Models
por: Tliba, Marouane, et al.
Publicado: (2025)