Saved in:
| Main Authors: | Kisel, Nikita, Volkov, Illia, Hanzelkova, Katerina, Janouskova, Klara, Matas, Jiri |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.00076 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025)
by: Volkov, Illia, et al.
Published: (2025)
Multimodal Large Language Models as Image Classifiers
by: Kisel, Nikita, et al.
Published: (2026)
by: Kisel, Nikita, et al.
Published: (2026)
Bringing the Context Back into Object Recognition, Robustly
by: Janouskova, Klara, et al.
Published: (2024)
by: Janouskova, Klara, et al.
Published: (2024)
Robust Context-Aware Object Recognition
by: Janouskova, Klara, et al.
Published: (2025)
by: Janouskova, Klara, et al.
Published: (2025)
Single Image Test-Time Adaptation for Segmentation
by: Janouskova, Klara, et al.
Published: (2023)
by: Janouskova, Klara, et al.
Published: (2023)
Koo-Fu CLIP: Closed-Form Adaptation of Vision-Language Models via Fukunaga-Koontz Linear Discriminant Analysis
by: Suchanek, Matej, et al.
Published: (2026)
by: Suchanek, Matej, et al.
Published: (2026)
FungiTastic: A multi-modal dataset and benchmark for image categorization
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
Speedrunning ImageNet Diffusion
by: Bhanded, Swayam
Published: (2025)
by: Bhanded, Swayam
Published: (2025)
Beyond ImageNet: Understanding Cross-Dataset Robustness of Lightweight Vision Models
by: Zhang, Weidong, et al.
Published: (2025)
by: Zhang, Weidong, et al.
Published: (2025)
Mitigating Overfitting in Medical Imaging: Self-Supervised Pretraining vs. ImageNet Transfer Learning for Dermatological Diagnosis
by: Matas, Iván, et al.
Published: (2025)
by: Matas, Iván, et al.
Published: (2025)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
by: Geigle, Gregor, et al.
Published: (2023)
by: Geigle, Gregor, et al.
Published: (2023)
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets
by: Amangeldi, Aidar, et al.
Published: (2025)
by: Amangeldi, Aidar, et al.
Published: (2025)
Fine-Grained ImageNet Classification in the Wild
by: Lymperaiou, Maria, et al.
Published: (2023)
by: Lymperaiou, Maria, et al.
Published: (2023)
ImageNet-Think-250K: A Large-Scale Synthetic Dataset for Multimodal Reasoning for Vision Language Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
Accessing Vision Foundation Models via ImageNet-1K
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
How far can we go with ImageNet for Text-to-Image generation?
by: Degeorge, L., et al.
Published: (2025)
by: Degeorge, L., et al.
Published: (2025)
What Makes ImageNet Look Unlike LAION
by: Shirali, Ali, et al.
Published: (2023)
by: Shirali, Ali, et al.
Published: (2023)
ImageNot: A contrast with ImageNet preserves model rankings
by: Salaudeen, Olawale, et al.
Published: (2024)
by: Salaudeen, Olawale, et al.
Published: (2024)
Detection, Pose Estimation and Segmentation for Multiple Bodies: Closing the Virtuous Circle
by: Purkrabek, Miroslav, et al.
Published: (2024)
by: Purkrabek, Miroslav, et al.
Published: (2024)
ProbPose: A Probabilistic Approach to 2D Human Pose Estimation
by: Purkrabek, Miroslav, et al.
Published: (2024)
by: Purkrabek, Miroslav, et al.
Published: (2024)
Video shutter angle estimation using optical flow and linear blur
by: Korcak, David, et al.
Published: (2023)
by: Korcak, David, et al.
Published: (2023)
Accurate Planar Tracking With Robust Re-Detection
by: Serych, Jonas, et al.
Published: (2026)
by: Serych, Jonas, et al.
Published: (2026)
Improving 2D Human Pose Estimation in Rare Camera Views with Synthetic Data
by: Purkrabek, Miroslav, et al.
Published: (2023)
by: Purkrabek, Miroslav, et al.
Published: (2023)
ImageNet-OOD: Deciphering Modern Out-of-Distribution Detection Algorithms
by: Yang, William, et al.
Published: (2023)
by: Yang, William, et al.
Published: (2023)
A New Dataset and a Distractor-Aware Architecture for Transparent Object Tracking
by: Lukezic, Alan, et al.
Published: (2024)
by: Lukezic, Alan, et al.
Published: (2024)
Going Beyond U-Net: Assessing Vision Transformers for Semantic Segmentation in Microscopy Image Analysis
by: Tsiporenko, Illia, et al.
Published: (2024)
by: Tsiporenko, Illia, et al.
Published: (2024)
Automated Classification of Model Errors on ImageNet
by: Peychev, Momchil, et al.
Published: (2023)
by: Peychev, Momchil, et al.
Published: (2023)
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
by: Muralidharan, Srikanth, et al.
Published: (2025)
by: Muralidharan, Srikanth, et al.
Published: (2025)
SOOD-ImageNet: a Large-Scale Dataset for Semantic Out-Of-Distribution Image Classification and Semantic Segmentation
by: Bacchin, Alberto, et al.
Published: (2024)
by: Bacchin, Alberto, et al.
Published: (2024)
Geometry aware 3D generation from in-the-wild images in ImageNet
by: Shen, Qijia, et al.
Published: (2024)
by: Shen, Qijia, et al.
Published: (2024)
G3DR: Generative 3D Reconstruction in ImageNet
by: Reddy, Pradyumna, et al.
Published: (2024)
by: Reddy, Pradyumna, et al.
Published: (2024)
SAM2RL: Towards Reinforcement Learning Memory Control in Segment Anything Model 2
by: Adamyan, Alen, et al.
Published: (2025)
by: Adamyan, Alen, et al.
Published: (2025)
Can Biases in ImageNet Models Explain Generalization?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
Toward Errorless Training ImageNet-1k
by: Deng, Bo, et al.
Published: (2025)
by: Deng, Bo, et al.
Published: (2025)
Comparative Performance of Finetuned ImageNet Pre-trained Models for Electronic Component Classification
by: Shao, Yidi, et al.
Published: (2025)
by: Shao, Yidi, et al.
Published: (2025)
Corner Cases: How Size and Position of Objects Challenge ImageNet-Trained Models
by: Fatima, Mishal, et al.
Published: (2025)
by: Fatima, Mishal, et al.
Published: (2025)
Unlocking ImageNet's Multi-Object Nature: Automated Large-Scale Multilabel Annotation
by: Chen, Junyu, et al.
Published: (2026)
by: Chen, Junyu, et al.
Published: (2026)
Do ImageNet-trained models learn shortcuts? The impact of frequency shortcuts on generalization
by: Wang, Shunxin, et al.
Published: (2025)
by: Wang, Shunxin, et al.
Published: (2025)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
by: Stevens, Samuel
Published: (2025)
by: Stevens, Samuel
Published: (2025)
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023)
by: Vishniakov, Kirill, et al.
Published: (2023)
Similar Items
-
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025) -
Multimodal Large Language Models as Image Classifiers
by: Kisel, Nikita, et al.
Published: (2026) -
Bringing the Context Back into Object Recognition, Robustly
by: Janouskova, Klara, et al.
Published: (2024) -
Robust Context-Aware Object Recognition
by: Janouskova, Klara, et al.
Published: (2025) -
Single Image Test-Time Adaptation for Segmentation
by: Janouskova, Klara, et al.
Published: (2023)