Beyond ImageNet: Understanding Cross-Dataset Robustness of Lightweight Vision Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Weidong, Ding, Pak Lun Kevin, Liu, Huan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accessing Vision Foundation Models via ImageNet-1K
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023)
by: Vishniakov, Kirill, et al.
Published: (2023)
ImageNet-Think-250K: A Large-Scale Synthetic Dataset for Multimodal Reasoning for Vision Language Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
ImageNet-Patch: A Dataset for Benchmarking Machine Learning Robustness against Adversarial Patches
by: Pintor, Maura, et al.
Published: (2022)
by: Pintor, Maura, et al.
Published: (2022)
Automated Classification of Model Errors on ImageNet
by: Peychev, Momchil, et al.
Published: (2023)
by: Peychev, Momchil, et al.
Published: (2023)
Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts
by: Crabbé, Jonathan, et al.
Published: (2023)
by: Crabbé, Jonathan, et al.
Published: (2023)
ImageNot: A contrast with ImageNet preserves model rankings
by: Salaudeen, Olawale, et al.
Published: (2024)
by: Salaudeen, Olawale, et al.
Published: (2024)
What Makes ImageNet Look Unlike LAION
by: Shirali, Ali, et al.
Published: (2023)
by: Shirali, Ali, et al.
Published: (2023)
SOOD-ImageNet: a Large-Scale Dataset for Semantic Out-Of-Distribution Image Classification and Semantic Segmentation
by: Bacchin, Alberto, et al.
Published: (2024)
by: Bacchin, Alberto, et al.
Published: (2024)
ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object
by: Zhang, Chenshuang, et al.
Published: (2024)
by: Zhang, Chenshuang, et al.
Published: (2024)
Can Biases in ImageNet Models Explain Generalization?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
Scaling Up Deep Clustering Methods Beyond ImageNet-1K
by: Adaloglou, Nikolas, et al.
Published: (2024)
by: Adaloglou, Nikolas, et al.
Published: (2024)
From MNIST to ImageNet: Understanding the Scalability Boundaries of Differentiable Logic Gate Networks
by: Brändle, Sven, et al.
Published: (2025)
by: Brändle, Sven, et al.
Published: (2025)
Flaws of ImageNet, Computer Vision's Favourite Dataset
by: Kisel, Nikita, et al.
Published: (2024)
by: Kisel, Nikita, et al.
Published: (2024)
Toward Errorless Training ImageNet-1k
by: Deng, Bo, et al.
Published: (2025)
by: Deng, Bo, et al.
Published: (2025)
Squeeze, Recover and Relabel: Dataset Condensation at ImageNet Scale From A New Perspective
by: Yin, Zeyuan, et al.
Published: (2023)
by: Yin, Zeyuan, et al.
Published: (2023)
Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets?
by: Ozbulak, Utku, et al.
Published: (2025)
by: Ozbulak, Utku, et al.
Published: (2025)
Swin-UMamba: Mamba-based UNet with ImageNet-based pretraining
by: Liu, Jiarun, et al.
Published: (2024)
by: Liu, Jiarun, et al.
Published: (2024)
Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion
by: Hoogeboom, Emiel, et al.
Published: (2024)
by: Hoogeboom, Emiel, et al.
Published: (2024)
Efficiera Residual Networks: Hardware-Friendly Fully Binary Weight with 2-bit Activation Model Achieves Practical ImageNet Accuracy
by: Takahashi, Shuntaro, et al.
Published: (2024)
by: Takahashi, Shuntaro, et al.
Published: (2024)
Spikformer V2: Join the High Accuracy Club on ImageNet with an SNN Ticket
by: Zhou, Zhaokun, et al.
Published: (2024)
by: Zhou, Zhaokun, et al.
Published: (2024)
Pre-training of Lightweight Vision Transformers on Small Datasets with Minimally Scaled Images
by: Tan, Jen Hong
Published: (2024)
by: Tan, Jen Hong
Published: (2024)
SteerVLM: Robust Model Control through Lightweight Activation Steering for Vision Language Models
by: Sivakumar, Anushka, et al.
Published: (2025)
by: Sivakumar, Anushka, et al.
Published: (2025)
ImageNet-trained CNNs are not biased towards texture: Revisiting feature reliance through controlled suppression
by: Burgert, Tom, et al.
Published: (2025)
by: Burgert, Tom, et al.
Published: (2025)
Speedrunning ImageNet Diffusion
by: Bhanded, Swayam
Published: (2025)
by: Bhanded, Swayam
Published: (2025)
A Lightweight Large Vision-language Model for Multimodal Medical Images
by: Alsinglawi, Belal, et al.
Published: (2025)
by: Alsinglawi, Belal, et al.
Published: (2025)
CrossFuse: Learning Infrared and Visible Image Fusion by Cross-Sensor Top-K Vision Alignment and Beyond
by: Shi, Yukai, et al.
Published: (2025)
by: Shi, Yukai, et al.
Published: (2025)
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers
by: Islam, Chashi Mahiul, et al.
Published: (2025)
by: Islam, Chashi Mahiul, et al.
Published: (2025)
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets
by: Amangeldi, Aidar, et al.
Published: (2025)
by: Amangeldi, Aidar, et al.
Published: (2025)
Beyond Off-the-Shelf Models: A Lightweight and Accessible Machine Learning Pipeline for Ecologists Working with Image Data
by: Chemery, Clare, et al.
Published: (2026)
by: Chemery, Clare, et al.
Published: (2026)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
by: Geigle, Gregor, et al.
Published: (2023)
by: Geigle, Gregor, et al.
Published: (2023)
Source Matters: Source Dataset Impact on Model Robustness in Medical Imaging
by: Juodelyte, Dovile, et al.
Published: (2024)
by: Juodelyte, Dovile, et al.
Published: (2024)
Foundation Model-oriented Robustness: Robust Image Model Evaluation with Pretrained Models
by: Zhang, Peiyan, et al.
Published: (2023)
by: Zhang, Peiyan, et al.
Published: (2023)
STAR-Net: An Interpretable Model-Aided Network for Remote Sensing Image Denoising
by: Liu, Jingjing, et al.
Published: (2025)
by: Liu, Jingjing, et al.
Published: (2025)
Beyond Training: Dynamic Token Merging for Zero-Shot Video Understanding
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
BioBench: A Blueprint to Move Beyond ImageNet for Scientific ML Benchmarks
by: Stevens, Samuel
Published: (2025)
by: Stevens, Samuel
Published: (2025)
Beyond Modality Collapse: Representations Blending for Multimodal Dataset Distillation
by: Zhang, Xin, et al.
Published: (2025)
by: Zhang, Xin, et al.
Published: (2025)
Towards Understanding How Knowledge Evolves in Large Vision-Language Models
by: Wang, Sudong, et al.
Published: (2025)
by: Wang, Sudong, et al.
Published: (2025)
Towards Robust Cross-Dataset Object Detection Generalization under Domain Specificity
by: Chakraborty, Ritabrata, et al.
Published: (2026)
by: Chakraborty, Ritabrata, et al.
Published: (2026)
Similar Items
-
Accessing Vision Foundation Models via ImageNet-1K
by: Zhang, Yitian, et al.
Published: (2024) -
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023) -
ImageNet-Think-250K: A Large-Scale Synthetic Dataset for Multimodal Reasoning for Vision Language Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025) -
ImageNet-Patch: A Dataset for Benchmarking Machine Learning Robustness against Adversarial Patches
by: Pintor, Maura, et al.
Published: (2022) -
Automated Classification of Model Errors on ImageNet
by: Peychev, Momchil, et al.
Published: (2023)