Vision Backbone Efficient Selection for Image Classification in Low-Data Regimes
Fuente:
arXiv
Guardado en:
| Autores principales: | Guerin, Joris, Bansal, Shray, Shaban, Amirreza, Mann, Paulo, Gazula, Harshvardhan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hierarchical Classification for Automated Image Annotation of Coral Reef Benthic Structures
por: Blondin, Célia, et al.
Publicado: (2024)
por: Blondin, Célia, et al.
Publicado: (2024)
ViR: Towards Efficient Vision Retention Backbones
por: Hatamizadeh, Ali, et al.
Publicado: (2023)
por: Hatamizadeh, Ali, et al.
Publicado: (2023)
CPUBone: Efficient Vision Backbone Design for Devices with Low Parallelization Capabilities
por: Nottebaum, Moritz, et al.
Publicado: (2026)
por: Nottebaum, Moritz, et al.
Publicado: (2026)
Vision-LSTM: xLSTM as Generic Vision Backbone
por: Alkin, Benedikt, et al.
Publicado: (2024)
por: Alkin, Benedikt, et al.
Publicado: (2024)
CoMViT: An Efficient Vision Backbone for Supervised Classification in Medical Imaging
por: Safdar, Aon, et al.
Publicado: (2025)
por: Safdar, Aon, et al.
Publicado: (2025)
Can we Defend Against the Unknown? An Empirical Study About Threshold Selection for Neural Network Monitoring
por: Dang, Khoi Tran, et al.
Publicado: (2024)
por: Dang, Khoi Tran, et al.
Publicado: (2024)
Unlocking Pre-trained Image Backbones for Semantic Image Synthesis
por: Berrada, Tariq, et al.
Publicado: (2023)
por: Berrada, Tariq, et al.
Publicado: (2023)
EfficientTrain++: Generalized Curriculum Learning for Efficient Visual Backbone Training
por: Wang, Yulin, et al.
Publicado: (2024)
por: Wang, Yulin, et al.
Publicado: (2024)
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
por: Khazem, Salim
Publicado: (2026)
por: Khazem, Salim
Publicado: (2026)
Beyond MACs: Hardware Efficient Architecture Design for Vision Backbones
por: Nottebaum, Moritz, et al.
Publicado: (2026)
por: Nottebaum, Moritz, et al.
Publicado: (2026)
Improved Sub-Visible Particle Classification in Flow Imaging Microscopy via Generative AI-Based Image Synthesis
por: Ozbulak, Utku, et al.
Publicado: (2025)
por: Ozbulak, Utku, et al.
Publicado: (2025)
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models
por: Saini, Harshvardhan, et al.
Publicado: (2026)
por: Saini, Harshvardhan, et al.
Publicado: (2026)
Enhancing Fine-Grained Visual Recognition in the Low-Data Regime Through Feature Magnitude Regularization
por: Chapman, Avraham, et al.
Publicado: (2024)
por: Chapman, Avraham, et al.
Publicado: (2024)
Enhancing Few-Shot Image Classification through Learnable Multi-Scale Embedding and Attention Mechanisms
por: Askari, Fatemeh, et al.
Publicado: (2024)
por: Askari, Fatemeh, et al.
Publicado: (2024)
Revisiting the Integration of Convolution and Attention for Vision Backbone
por: Zhu, Lei, et al.
Publicado: (2024)
por: Zhu, Lei, et al.
Publicado: (2024)
Safety Monitoring of Machine Learning Perception Functions: a Survey
por: Ferreira, Raul Sena, et al.
Publicado: (2024)
por: Ferreira, Raul Sena, et al.
Publicado: (2024)
Pretraining Objective Matters in Extreme Low-Data FGVC: A Backbone-Controlled Study
por: Hackett, Alexander, et al.
Publicado: (2026)
por: Hackett, Alexander, et al.
Publicado: (2026)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
por: Sung, Yi-Lin, et al.
Publicado: (2023)
por: Sung, Yi-Lin, et al.
Publicado: (2023)
Adaptive Knowledge Distillation for Classification of Hand Images using Explainable Vision Transformers
por: Nguyen, Thanh Thi, et al.
Publicado: (2024)
por: Nguyen, Thanh Thi, et al.
Publicado: (2024)
Dual-Model Weight Selection and Self-Knowledge Distillation for Medical Image Classification
por: Tsutsumi, Ayaka, et al.
Publicado: (2025)
por: Tsutsumi, Ayaka, et al.
Publicado: (2025)
FILS: Self-Supervised Video Feature Prediction In Semantic Language Space
por: Ahmadian, Mona, et al.
Publicado: (2024)
por: Ahmadian, Mona, et al.
Publicado: (2024)
CountCLIP -- [Re] Teaching CLIP to Count to Ten
por: Mestha, Harshvardhan, et al.
Publicado: (2024)
por: Mestha, Harshvardhan, et al.
Publicado: (2024)
Beyond CNNs: Efficient Fine-Tuning of Multi-Modal LLMs for Object Detection on Low-Data Regimes
por: Elamon, Nirmal, et al.
Publicado: (2025)
por: Elamon, Nirmal, et al.
Publicado: (2025)
Capsule Vision Challenge 2024: Multi-Class Abnormality Classification for Video Capsule Endoscopy
por: Bansal, Aakarsh, et al.
Publicado: (2024)
por: Bansal, Aakarsh, et al.
Publicado: (2024)
Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling
por: Rodriguez-Opazo, Cristian, et al.
Publicado: (2024)
por: Rodriguez-Opazo, Cristian, et al.
Publicado: (2024)
Normalization Equivariance for Arbitrary Backbones, with Application to Image Denoising
por: Saied, Youssef, et al.
Publicado: (2026)
por: Saied, Youssef, et al.
Publicado: (2026)
Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model
por: Lin, Han, et al.
Publicado: (2024)
por: Lin, Han, et al.
Publicado: (2024)
Selective Classification Under Distribution Shifts
por: Liang, Hengyue, et al.
Publicado: (2024)
por: Liang, Hengyue, et al.
Publicado: (2024)
Understanding Pruning Regimes in Vision-Language Models Through Domain-Aware Layer Selection
por: Khaki, Saeed, et al.
Publicado: (2026)
por: Khaki, Saeed, et al.
Publicado: (2026)
Sparse Model Inversion: Efficient Inversion of Vision Transformers for Data-Free Applications
por: Hu, Zixuan, et al.
Publicado: (2025)
por: Hu, Zixuan, et al.
Publicado: (2025)
Efficient Vision-and-Language Pre-training with Text-Relevant Image Patch Selection
por: Ye, Wei, et al.
Publicado: (2024)
por: Ye, Wei, et al.
Publicado: (2024)
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
por: Hong, Yunqi, et al.
Publicado: (2025)
por: Hong, Yunqi, et al.
Publicado: (2025)
Sensitive Image Classification by Vision Transformers
por: He, Hanxian, et al.
Publicado: (2024)
por: He, Hanxian, et al.
Publicado: (2024)
SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation
por: Yoon, Jaehong, et al.
Publicado: (2024)
por: Yoon, Jaehong, et al.
Publicado: (2024)
Data-Driven Analysis of Intersectional Bias in Image Classification: A Framework with Bias-Weighted Augmentation
por: Yesmin, Farjana
Publicado: (2025)
por: Yesmin, Farjana
Publicado: (2025)
Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels
por: Moghaddam, Alireza Sedighi, et al.
Publicado: (2025)
por: Moghaddam, Alireza Sedighi, et al.
Publicado: (2025)
Convolutional Neural Nets vs Vision Transformers: A SpaceNet Case Study with Balanced vs Imbalanced Regimes
por: Gothi, Akshar
Publicado: (2025)
por: Gothi, Akshar
Publicado: (2025)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
por: Choi, Kanghyun, et al.
Publicado: (2024)
por: Choi, Kanghyun, et al.
Publicado: (2024)
TRIPS: Efficient Vision-and-Language Pre-training with Text-Relevant Image Patch Selection
por: Jiang, Chaoya, et al.
Publicado: (2023)
por: Jiang, Chaoya, et al.
Publicado: (2023)
FastVLM: Efficient Vision Encoding for Vision Language Models
por: Vasu, Pavan Kumar Anasosalu, et al.
Publicado: (2024)
por: Vasu, Pavan Kumar Anasosalu, et al.
Publicado: (2024)
Ejemplares similares
-
Hierarchical Classification for Automated Image Annotation of Coral Reef Benthic Structures
por: Blondin, Célia, et al.
Publicado: (2024) -
ViR: Towards Efficient Vision Retention Backbones
por: Hatamizadeh, Ali, et al.
Publicado: (2023) -
CPUBone: Efficient Vision Backbone Design for Devices with Low Parallelization Capabilities
por: Nottebaum, Moritz, et al.
Publicado: (2026) -
Vision-LSTM: xLSTM as Generic Vision Backbone
por: Alkin, Benedikt, et al.
Publicado: (2024) -
CoMViT: An Efficient Vision Backbone for Supervised Classification in Medical Imaging
por: Safdar, Aon, et al.
Publicado: (2025)