AttZoom: Attention Zoom for Better Visual Features
Fuente:
arXiv
Saved in:
| Main Authors: | DeAlcala, Daniel, Morales, Aythami, Fierrez, Julian, Tolosana, Ruben |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Membership Inference Test: Auditing Training Data in Object Classification Models
by: Mancera, Gonzalo, et al.
Published: (2026)
by: Mancera, Gonzalo, et al.
Published: (2026)
MINT-Demo: Membership Inference Test Demonstrator
by: DeAlcala, Daniel, et al.
Published: (2025)
by: DeAlcala, Daniel, et al.
Published: (2025)
Active Membership Inference Test (aMINT): Enhancing Model Auditability with Multi-Task Learning
by: DeAlcala, Daniel, et al.
Published: (2025)
by: DeAlcala, Daniel, et al.
Published: (2025)
Is My Data in Your AI? Membership Inference Test (MINT) applied to Face Biometrics
by: DeAlcala, Daniel, et al.
Published: (2024)
by: DeAlcala, Daniel, et al.
Published: (2024)
Is My Text in Your AI Model? Gradient-based Membership Inference Test applied to LLMs
by: Mancera, Gonzalo, et al.
Published: (2025)
by: Mancera, Gonzalo, et al.
Published: (2025)
FakeIDet: Exploring Patches for Privacy-Preserving Fake ID Detection
by: Muñoz-Haro, Javier, et al.
Published: (2025)
by: Muñoz-Haro, Javier, et al.
Published: (2025)
mEBAL: A Multimodal Database for Eye Blink Detection and Attention Level Estimation
by: Daza, Roberto, et al.
Published: (2020)
by: Daza, Roberto, et al.
Published: (2020)
Privacy-Aware Detection of Fake Identity Documents: Methodology, Benchmark, and Improved Algorithms (FakeIDet2)
by: Muñoz-Haro, Javier, et al.
Published: (2025)
by: Muñoz-Haro, Javier, et al.
Published: (2025)
Balancing Tails when Comparing Distributions: Comprehensive Equity Index (CEI) with Application to Bias Evaluation in Operational Face Biometrics
by: Solano, Imanol, et al.
Published: (2025)
by: Solano, Imanol, et al.
Published: (2025)
Comprehensive Equity Index (CEI): Definition and Application to Bias Evaluation in Biometrics
by: Solano, Imanol, et al.
Published: (2024)
by: Solano, Imanol, et al.
Published: (2024)
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability
by: DeAndres-Tame, Ivan, et al.
Published: (2024)
by: DeAndres-Tame, Ivan, et al.
Published: (2024)
Exploring Deep Learning and Ultra-Widefield Imaging for Diabetic Retinopathy and Macular Edema
by: Jimenez-Lizcano, Pablo, et al.
Published: (2026)
by: Jimenez-Lizcano, Pablo, et al.
Published: (2026)
Just Zoom In: Cross-View Geo-Localization via Autoregressive Zooming
by: Erzurumlu, Yunus Talha, et al.
Published: (2026)
by: Erzurumlu, Yunus Talha, et al.
Published: (2026)
Is It Really You? Exploring Biometric Verification Scenarios in Photorealistic Talking-Head Avatar Videos
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2025)
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2025)
AdaZoom-GUI: Adaptive Zoom-based GUI Grounding with Instruction Refinement
by: Pei, Siqi, et al.
Published: (2026)
by: Pei, Siqi, et al.
Published: (2026)
Zoom in, Click out: Unlocking and Evaluating the Potential of Zooming for GUI Grounding
by: Jiang, Zhiyuan, et al.
Published: (2025)
by: Jiang, Zhiyuan, et al.
Published: (2025)
mEBAL2 Database and Benchmark: Image-based Multispectral Eyeblink Detection
by: Daza, Roberto, et al.
Published: (2023)
by: Daza, Roberto, et al.
Published: (2023)
DeepFace-Attention: Multimodal Face Biometrics for Attention Estimation with Application to e-Learning
by: Daza, Roberto, et al.
Published: (2024)
by: Daza, Roberto, et al.
Published: (2024)
Zoom and Shift are All You Need
by: Qin, Jiahao
Published: (2024)
by: Qin, Jiahao
Published: (2024)
SaFL: Sybil-aware Federated Learning with Application to Face Recognition
by: Ghafourian, Mahdi, et al.
Published: (2023)
by: Ghafourian, Mahdi, et al.
Published: (2023)
Seeing the Unseen: Zooming in the Dark with Event Cameras
by: Kai, Dachun, et al.
Published: (2026)
by: Kai, Dachun, et al.
Published: (2026)
Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception
by: Wei, Lai, et al.
Published: (2026)
by: Wei, Lai, et al.
Published: (2026)
Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines
by: Kim, Keon, et al.
Published: (2026)
by: Kim, Keon, et al.
Published: (2026)
Iterative Zoom-In: Temporal Interval Exploration for Long Video Understanding
by: Li, Chenglin, et al.
Published: (2025)
by: Li, Chenglin, et al.
Published: (2025)
Comparison of Visual Trackers for Biomechanical Analysis of Running
by: Gomez, Luis F., et al.
Published: (2025)
by: Gomez, Luis F., et al.
Published: (2025)
Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models
by: Thapa, Rahul, et al.
Published: (2024)
by: Thapa, Rahul, et al.
Published: (2024)
Are Vision-Language Models Ready for Dietary Assessment? Exploring the Next Frontier in AI-Powered Food Image Recognition
by: Romero-Tapiador, Sergio, et al.
Published: (2025)
by: Romero-Tapiador, Sergio, et al.
Published: (2025)
WonderZoom: Multi-Scale 3D World Generation
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding
by: Tang, Fei, et al.
Published: (2026)
by: Tang, Fei, et al.
Published: (2026)
Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming
by: Zhou, Yue, et al.
Published: (2026)
by: Zhou, Yue, et al.
Published: (2026)
Leveraging Avatar Fingerprinting: A Multi-Generator Photorealistic Talking-Head Public Database and Benchmark
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2026)
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2026)
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling
by: Liu, Xianjie, et al.
Published: (2025)
by: Liu, Xianjie, et al.
Published: (2025)
From Panel to Pixel: Zoom-In Vision-Language Pretraining from Biomedical Scientific Literature
by: Yuan, Kun, et al.
Published: (2025)
by: Yuan, Kun, et al.
Published: (2025)
Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models
by: Shi, Yuheng, et al.
Published: (2026)
by: Shi, Yuheng, et al.
Published: (2026)
VideoRun2D: Cost-Effective Markerless Motion Capture for Sprint Biomechanics
by: Garrido-Lopez, Gonzalo, et al.
Published: (2024)
by: Garrido-Lopez, Gonzalo, et al.
Published: (2024)
Benchmarking Graph Neural Networks for Document Layout Analysis in Public Affairs
by: Lopez-Duran, Miguel, et al.
Published: (2025)
by: Lopez-Duran, Miguel, et al.
Published: (2025)
Exploiting Multiple Representations: 3D Face Biometrics Fusion with Application to Surveillance
by: La Cava, Simone Maurizio, et al.
Published: (2025)
by: La Cava, Simone Maurizio, et al.
Published: (2025)
Exploring 3D Face Reconstruction and Fusion Methods for Face Verification: A Case-Study in Video Surveillance
by: La Cava, Simone Maurizio, et al.
Published: (2024)
by: La Cava, Simone Maurizio, et al.
Published: (2024)
Is Visual Realism Enough? Evaluating Gait Biometric Fidelity in Generative AI Human Animation
by: DeAndres-Tame, Ivan, et al.
Published: (2025)
by: DeAndres-Tame, Ivan, et al.
Published: (2025)
Similar Items
-
Membership Inference Test: Auditing Training Data in Object Classification Models
by: Mancera, Gonzalo, et al.
Published: (2026) -
MINT-Demo: Membership Inference Test Demonstrator
by: DeAlcala, Daniel, et al.
Published: (2025) -
Active Membership Inference Test (aMINT): Enhancing Model Auditability with Multi-Task Learning
by: DeAlcala, Daniel, et al.
Published: (2025) -
Is My Data in Your AI? Membership Inference Test (MINT) applied to Face Biometrics
by: DeAlcala, Daniel, et al.
Published: (2024) -
Is My Text in Your AI Model? Gradient-based Membership Inference Test applied to LLMs
by: Mancera, Gonzalo, et al.
Published: (2025)