Benchmarking Feature Upsampling Methods for Vision Foundation Models using Interactive Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Havrylov, Volodymyr, Huang, Haiwen, Zhang, Dan, Geiger, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
by: Huang, Haiwen, et al.
Published: (2025)
by: Huang, Haiwen, et al.
Published: (2025)
Benchmarking the Fairness of Image Upsampling Methods
by: Laszkiewicz, Mike, et al.
Published: (2024)
by: Laszkiewicz, Mike, et al.
Published: (2024)
How to Benchmark Vision Foundation Models for Semantic Segmentation?
by: Kerssies, Tommie, et al.
Published: (2024)
by: Kerssies, Tommie, et al.
Published: (2024)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
by: Ma, Huan, et al.
Published: (2024)
by: Ma, Huan, et al.
Published: (2024)
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models
by: Mai, Zheda, et al.
Published: (2025)
by: Mai, Zheda, et al.
Published: (2025)
Hyperspectral Adapter for Semantic Segmentation with Vision Foundation Models
by: Hurtado, Juana Valeria, et al.
Published: (2025)
by: Hurtado, Juana Valeria, et al.
Published: (2025)
Renovating Names in Open-Vocabulary Segmentation Benchmarks
by: Huang, Haiwen, et al.
Published: (2024)
by: Huang, Haiwen, et al.
Published: (2024)
ConformalSAM: Unlocking the Potential of Foundational Segmentation Models in Semi-Supervised Semantic Segmentation with Conformal Prediction
by: Chen, Danhui, et al.
Published: (2025)
by: Chen, Danhui, et al.
Published: (2025)
Few-Shot Segmentation of Historical Maps via Linear Probing of Vision Foundation Models
by: Sterzinger, Rafael, et al.
Published: (2025)
by: Sterzinger, Rafael, et al.
Published: (2025)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024)
by: Chitta, Kashyap, et al.
Published: (2024)
A Foundation Model for General Moving Object Segmentation in Medical Images
by: Yan, Zhongnuo, et al.
Published: (2023)
by: Yan, Zhongnuo, et al.
Published: (2023)
Benchmarking the Robustness of Instance Segmentation Models
by: Dalva, Yusuf, et al.
Published: (2021)
by: Dalva, Yusuf, et al.
Published: (2021)
MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models
by: Shi, Yang, et al.
Published: (2026)
by: Shi, Yang, et al.
Published: (2026)
Domain-Aware Fine-Tuning of Foundation Models
by: Kaplan, Ugur Ali, et al.
Published: (2024)
by: Kaplan, Ugur Ali, et al.
Published: (2024)
Collaborating Foundation Models for Domain Generalized Semantic Segmentation
by: Benigmim, Yasser, et al.
Published: (2023)
by: Benigmim, Yasser, et al.
Published: (2023)
Accessing Vision Foundation Models via ImageNet-1K
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
by: Miyato, Takeru, et al.
Published: (2023)
by: Miyato, Takeru, et al.
Published: (2023)
CFM: Language-aligned Concept Foundation Model for Vision
by: Wittenmayer, Kai, et al.
Published: (2026)
by: Wittenmayer, Kai, et al.
Published: (2026)
GLoG-CSUnet: Enhancing Vision Transformers with Adaptable Radiomic Features for Medical Image Segmentation
by: Eghbali, Niloufar, et al.
Published: (2025)
by: Eghbali, Niloufar, et al.
Published: (2025)
ViTime: Foundation Model for Time Series Forecasting Powered by Vision Intelligence
by: Yang, Luoxiao, et al.
Published: (2024)
by: Yang, Luoxiao, et al.
Published: (2024)
First Place Solution to the ECCV 2024 BRAVO Challenge: Evaluating Robustness of Vision Foundation Models for Semantic Segmentation
by: Kerssies, Tommie, et al.
Published: (2024)
by: Kerssies, Tommie, et al.
Published: (2024)
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
by: Zhou, Andy, et al.
Published: (2023)
by: Zhou, Andy, et al.
Published: (2023)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
Post-pre-training for Modality Alignment in Vision-Language Foundation Models
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
Prompting Foundation Models for Zero-Shot Ship Instance Segmentation in SAR Imagery
by: Mansour, Islam, et al.
Published: (2026)
by: Mansour, Islam, et al.
Published: (2026)
PRIMA: Multi-Image Vision-Language Models for Reasoning Segmentation
by: Wahed, Muntasir, et al.
Published: (2024)
by: Wahed, Muntasir, et al.
Published: (2024)
ELMO: Enhanced Real-time LiDAR Motion Capture through Upsampling
by: Jang, Deok-Kyeong, et al.
Published: (2024)
by: Jang, Deok-Kyeong, et al.
Published: (2024)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
by: Pach, Mateusz, et al.
Published: (2025)
by: Pach, Mateusz, et al.
Published: (2025)
Navigating Data Scarcity using Foundation Models: A Benchmark of Few-Shot and Zero-Shot Learning Approaches in Medical Imaging
by: Woerner, Stefano, et al.
Published: (2024)
by: Woerner, Stefano, et al.
Published: (2024)
FunduSegmenter: Leveraging the RETFound Foundation Model for Joint Optic Disc and Optic Cup Segmentation in Retinal Fundus Images
by: Zhao, Zhenyi, et al.
Published: (2025)
by: Zhao, Zhenyi, et al.
Published: (2025)
Mars-Bench: A Benchmark for Evaluating Foundation Models for Mars Science Tasks
by: Purohit, Mirali, et al.
Published: (2025)
by: Purohit, Mirali, et al.
Published: (2025)
A Reasoning-Enabled Vision-Language Foundation Model for Chest X-ray Interpretation
by: Zhang, Yabin, et al.
Published: (2026)
by: Zhang, Yabin, et al.
Published: (2026)
Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets
by: Nezhurina, Marianna, et al.
Published: (2025)
by: Nezhurina, Marianna, et al.
Published: (2025)
Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection
by: Delibasoglu, Ibrahim
Published: (2026)
by: Delibasoglu, Ibrahim
Published: (2026)
Composition Vision-Language Understanding via Segment and Depth Anything Model
by: Huo, Mingxiao, et al.
Published: (2024)
by: Huo, Mingxiao, et al.
Published: (2024)
NAVSIM: Data-Driven Non-Reactive Autonomous Vehicle Simulation and Benchmarking
by: Dauner, Daniel, et al.
Published: (2024)
by: Dauner, Daniel, et al.
Published: (2024)
Decipher-MR: A Vision-Language Foundation Model for 3D MRI Representations
by: Yang, Zhijian, et al.
Published: (2025)
by: Yang, Zhijian, et al.
Published: (2025)
I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation
by: Sassoon, Jordan, et al.
Published: (2025)
by: Sassoon, Jordan, et al.
Published: (2025)
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks
by: Ramachandran, Rahul, et al.
Published: (2025)
by: Ramachandran, Rahul, et al.
Published: (2025)
Similar Items
-
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
by: Huang, Haiwen, et al.
Published: (2025) -
Benchmarking the Fairness of Image Upsampling Methods
by: Laszkiewicz, Mike, et al.
Published: (2024) -
How to Benchmark Vision Foundation Models for Semantic Segmentation?
by: Kerssies, Tommie, et al.
Published: (2024) -
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025) -
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
by: Ma, Huan, et al.
Published: (2024)