Vision Transformers for Zero-Shot Clustering of Animal Images: A Comparative Benchmarking Study
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Markoff, Hugo, Bengtson, Stefan Hein, Ørsted, Michael |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Zero-Shot Wildlife Sorting Using Vision Transformers: Evaluating Clustering and Continuous Similarity Ordering
par: Markoff, Hugo, et autres
Publié: (2025)
par: Markoff, Hugo, et autres
Publié: (2025)
Hierarchical Re-Classification: Combining Animal Classification Models with Vision Transformers
par: Markoff, Hugo, et autres
Publié: (2025)
par: Markoff, Hugo, et autres
Publié: (2025)
Benchmarking Zero-Shot Recognition with Vision-Language Models: Challenges on Granularity and Specificity
par: Xu, Zhenlin, et autres
Publié: (2023)
par: Xu, Zhenlin, et autres
Publié: (2023)
Binary Verification for Zero-Shot Vision
par: Hu, Rongbin, et autres
Publié: (2025)
par: Hu, Rongbin, et autres
Publié: (2025)
Sea-ing Through Scattered Rays: Revisiting the Image Formation Model for Realistic Underwater Image Generation
par: Ismiroglou, Vasiliki, et autres
Publié: (2025)
par: Ismiroglou, Vasiliki, et autres
Publié: (2025)
VLAgeBench: Benchmarking Large Vision-Language Models for Zero-Shot Human Age Estimation
par: Sajib, Rakib Hossain, et autres
Publié: (2026)
par: Sajib, Rakib Hossain, et autres
Publié: (2026)
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
par: Sony, Redwan, et autres
Publié: (2025)
par: Sony, Redwan, et autres
Publié: (2025)
Rethinking Plant Disease Diagnosis: Bridging the Academic-Practical Gap with Vision Transformers and Zero-Shot Learning
par: Benabbas, Wassim, et autres
Publié: (2025)
par: Benabbas, Wassim, et autres
Publié: (2025)
LesionLocator: Zero-Shot Universal Tumor Segmentation and Tracking in 3D Whole-Body Imaging
par: Rokuss, Maximilian, et autres
Publié: (2025)
par: Rokuss, Maximilian, et autres
Publié: (2025)
HeatPrompt: Zero-Shot Vision-Language Modeling of Urban Heat Demand from Satellite Images
par: Thota, Kundan, et autres
Publié: (2026)
par: Thota, Kundan, et autres
Publié: (2026)
Benchmarking Unlearning for Vision Transformers
par: Zhao, Kairan, et autres
Publié: (2026)
par: Zhao, Kairan, et autres
Publié: (2026)
Efficient Zero-Shot AI-Generated Image Detection
par: Sonoda, Ryosuke, et autres
Publié: (2026)
par: Sonoda, Ryosuke, et autres
Publié: (2026)
Zero-Shot Vision-and-Language Navigation with Collision Mitigation in Continuous Environment
par: Jeong, Seongjun, et autres
Publié: (2024)
par: Jeong, Seongjun, et autres
Publié: (2024)
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis
par: Nagar, Aishik, et autres
Publié: (2024)
par: Nagar, Aishik, et autres
Publié: (2024)
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion
par: Allgeuer, Philipp, et autres
Publié: (2024)
par: Allgeuer, Philipp, et autres
Publié: (2024)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
par: Luo, Kun, et autres
Publié: (2026)
par: Luo, Kun, et autres
Publié: (2026)
SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild
par: Hu, Xuyi, et autres
Publié: (2026)
par: Hu, Xuyi, et autres
Publié: (2026)
ZSPAPrune: Zero-Shot Prompt-Aware Token Pruning for Vision-Language Models
par: Zhang, Pu, et autres
Publié: (2025)
par: Zhang, Pu, et autres
Publié: (2025)
TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation
par: Li, Dingbang, et autres
Publié: (2024)
par: Li, Dingbang, et autres
Publié: (2024)
LLM meets Vision-Language Models for Zero-Shot One-Class Classification
par: Bendou, Yassir, et autres
Publié: (2024)
par: Bendou, Yassir, et autres
Publié: (2024)
Intriguing Differences Between Zero-Shot and Systematic Evaluations of Vision-Language Transformer Models
par: Salman, Shaeke, et autres
Publié: (2024)
par: Salman, Shaeke, et autres
Publié: (2024)
A Framework for Evaluating Zero-Shot Image Generation in Concept-based Explainability
par: Astolfi, Giacomo, et autres
Publié: (2026)
par: Astolfi, Giacomo, et autres
Publié: (2026)
Are Video Models Ready as Zero-Shot Reasoners? An Empirical Study with the MME-CoF Benchmark
par: Guo, Ziyu, et autres
Publié: (2025)
par: Guo, Ziyu, et autres
Publié: (2025)
Transductive Zero-Shot and Few-Shot CLIP
par: Martin, Ségolène, et autres
Publié: (2024)
par: Martin, Ségolène, et autres
Publié: (2024)
PostEdit: Posterior Sampling for Efficient Zero-Shot Image Editing
par: Tian, Feng, et autres
Publié: (2024)
par: Tian, Feng, et autres
Publié: (2024)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
par: Park, SoYoung, et autres
Publié: (2025)
par: Park, SoYoung, et autres
Publié: (2025)
Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging
par: Shams, Montasir, et autres
Publié: (2025)
par: Shams, Montasir, et autres
Publié: (2025)
TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions
par: Chen, Ce, et autres
Publié: (2026)
par: Chen, Ce, et autres
Publié: (2026)
Navigating Data Scarcity using Foundation Models: A Benchmark of Few-Shot and Zero-Shot Learning Approaches in Medical Imaging
par: Woerner, Stefano, et autres
Publié: (2024)
par: Woerner, Stefano, et autres
Publié: (2024)
Benchmarking Zero-Shot Robustness of Multimodal Foundation Models: A Pilot Study
par: Wang, Chenguang, et autres
Publié: (2024)
par: Wang, Chenguang, et autres
Publié: (2024)
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
par: Yu, Lu, et autres
Publié: (2024)
par: Yu, Lu, et autres
Publié: (2024)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
par: Choi, Lucas, et autres
Publié: (2024)
par: Choi, Lucas, et autres
Publié: (2024)
Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval
par: Liu, Mingyu, et autres
Publié: (2026)
par: Liu, Mingyu, et autres
Publié: (2026)
Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors
par: Kuang, Zhengfei, et autres
Publié: (2024)
par: Kuang, Zhengfei, et autres
Publié: (2024)
Toward Ethical Facial Age Estimation: A Generalized Zero-Shot Benchmark Without Training on Children's Data
par: Petrucci, Caio, et autres
Publié: (2026)
par: Petrucci, Caio, et autres
Publié: (2026)
Efficient Masked Attention Transformer for Few-Shot Classification and Segmentation
par: Carrión-Ojeda, Dustin, et autres
Publié: (2025)
par: Carrión-Ojeda, Dustin, et autres
Publié: (2025)
Navigating the Trade-off: A Synthesis of Defensive Strategies for Zero-Shot Adversarial Robustness in Vision-Language Models
par: Xu, Zane, et autres
Publié: (2025)
par: Xu, Zane, et autres
Publié: (2025)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
par: Li, Wenxi, et autres
Publié: (2025)
par: Li, Wenxi, et autres
Publié: (2025)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
par: Kumar, Yogesh, et autres
Publié: (2025)
par: Kumar, Yogesh, et autres
Publié: (2025)
CompareBench: A Benchmark for Visual Comparison Reasoning in Vision-Language Models
par: Cai, Jie, et autres
Publié: (2025)
par: Cai, Jie, et autres
Publié: (2025)
Documents similaires
-
Zero-Shot Wildlife Sorting Using Vision Transformers: Evaluating Clustering and Continuous Similarity Ordering
par: Markoff, Hugo, et autres
Publié: (2025) -
Hierarchical Re-Classification: Combining Animal Classification Models with Vision Transformers
par: Markoff, Hugo, et autres
Publié: (2025) -
Benchmarking Zero-Shot Recognition with Vision-Language Models: Challenges on Granularity and Specificity
par: Xu, Zhenlin, et autres
Publié: (2023) -
Binary Verification for Zero-Shot Vision
par: Hu, Rongbin, et autres
Publié: (2025) -
Sea-ing Through Scattered Rays: Revisiting the Image Formation Model for Realistic Underwater Image Generation
par: Ismiroglou, Vasiliki, et autres
Publié: (2025)