BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Baranwal, Aaditya, Yadav, Vishal, Rajora, Abhishek |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
por: Saha, Oindrila, et al.
Publicado: (2024)
por: Saha, Oindrila, et al.
Publicado: (2024)
MolSight: Molecular Property Prediction with Images
por: Baranwal, Aaditya, et al.
Publicado: (2026)
por: Baranwal, Aaditya, et al.
Publicado: (2026)
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
por: Hasanebrahimi, Afsaneh, et al.
Publicado: (2026)
por: Hasanebrahimi, Afsaneh, et al.
Publicado: (2026)
Vote-in-Context: Turning VLMs into Zero-Shot Rank Fusers
por: Eltahir, Mohamed, et al.
Publicado: (2025)
por: Eltahir, Mohamed, et al.
Publicado: (2025)
Re:Verse -- Can Your VLM Read a Manga?
por: Baranwal, Aaditya, et al.
Publicado: (2025)
por: Baranwal, Aaditya, et al.
Publicado: (2025)
Exploring Prompt Alignment with Clinical Factors in Zero-Shot Segmentation VLMs for NSCLC Tumor Segmentation
por: Pai, Suraj, et al.
Publicado: (2026)
por: Pai, Suraj, et al.
Publicado: (2026)
SynSpill: Improved Industrial Spill Detection With Synthetic Data
por: Baranwal, Aaditya, et al.
Publicado: (2025)
por: Baranwal, Aaditya, et al.
Publicado: (2025)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
por: Luo, Kun, et al.
Publicado: (2026)
por: Luo, Kun, et al.
Publicado: (2026)
Zero-Shot Monocular Scene Flow Estimation in the Wild
por: Liang, Yiqing, et al.
Publicado: (2025)
por: Liang, Yiqing, et al.
Publicado: (2025)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
por: Imam, Raza, et al.
Publicado: (2025)
por: Imam, Raza, et al.
Publicado: (2025)
Image Matching by Bare Homography
por: Bellavia, Fabio
Publicado: (2023)
por: Bellavia, Fabio
Publicado: (2023)
TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation
por: Wang, Yiyao, et al.
Publicado: (2026)
por: Wang, Yiyao, et al.
Publicado: (2026)
VROOM - Visual Reconstruction over Onboard Multiview
por: Yadav, Yajat, et al.
Publicado: (2025)
por: Yadav, Yajat, et al.
Publicado: (2025)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
ZDySS -- Zero-Shot Dynamic Scene Stylization using Gaussian Splatting
por: Saroha, Abhishek, et al.
Publicado: (2025)
por: Saroha, Abhishek, et al.
Publicado: (2025)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
por: Xu, Jiacong, et al.
Publicado: (2025)
por: Xu, Jiacong, et al.
Publicado: (2025)
TOMCAT: Test-time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning
por: Yan, Xudong, et al.
Publicado: (2025)
por: Yan, Xudong, et al.
Publicado: (2025)
Open-Pose 3D Zero-Shot Learning: Benchmark and Challenges
por: Zhao, Weiguang, et al.
Publicado: (2023)
por: Zhao, Weiguang, et al.
Publicado: (2023)
MAC: A Benchmark for Multiple Attributes Compositional Zero-Shot Learning
por: Xu, Shuo, et al.
Publicado: (2024)
por: Xu, Shuo, et al.
Publicado: (2024)
From Words to Wavelengths: VLMs for Few-Shot Multispectral Object Detection
por: Nkegoum, Manuel, et al.
Publicado: (2025)
por: Nkegoum, Manuel, et al.
Publicado: (2025)
From YOLO to VLMs: Advancing Zero-Shot and Few-Shot Detection of Wastewater Treatment Plants Using Satellite Imagery in MENA Region
por: Premarathna, Akila, et al.
Publicado: (2025)
por: Premarathna, Akila, et al.
Publicado: (2025)
Epsilon: Exploring Comprehensive Visual-Semantic Projection for Multi-Label Zero-Shot Learning
por: Liu, Ziming, et al.
Publicado: (2024)
por: Liu, Ziming, et al.
Publicado: (2024)
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
por: Silva-Rodríguez, Julio, et al.
Publicado: (2025)
por: Silva-Rodríguez, Julio, et al.
Publicado: (2025)
Will It Zero-Shot?: Predicting Zero-Shot Classification Performance For Arbitrary Queries
por: Robbins, Kevin, et al.
Publicado: (2026)
por: Robbins, Kevin, et al.
Publicado: (2026)
XDT-CXR: Investigating Cross-Disease Transferability in Zero-Shot Binary Classification of Chest X-Rays
por: Rahman, Umaima, et al.
Publicado: (2024)
por: Rahman, Umaima, et al.
Publicado: (2024)
ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models
por: Roberts, Jonathan, et al.
Publicado: (2025)
por: Roberts, Jonathan, et al.
Publicado: (2025)
DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs
por: Pan, Jiazhen, et al.
Publicado: (2026)
por: Pan, Jiazhen, et al.
Publicado: (2026)
Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
por: Yang, Yuchen, et al.
Publicado: (2026)
por: Yang, Yuchen, et al.
Publicado: (2026)
DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving
por: Vo, Hao, et al.
Publicado: (2026)
por: Vo, Hao, et al.
Publicado: (2026)
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
por: Sony, Redwan, et al.
Publicado: (2025)
por: Sony, Redwan, et al.
Publicado: (2025)
Few-Shot, Now for Real: Medical VLMs Adaptation without Balanced Sets or Validation
por: Silva-Rodríguez, Julio, et al.
Publicado: (2025)
por: Silva-Rodríguez, Julio, et al.
Publicado: (2025)
Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment
por: Chen, Jingkun, et al.
Publicado: (2026)
por: Chen, Jingkun, et al.
Publicado: (2026)
Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos
por: Han, Mingfei, et al.
Publicado: (2023)
por: Han, Mingfei, et al.
Publicado: (2023)
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
por: Fang, Zhengru, et al.
Publicado: (2026)
por: Fang, Zhengru, et al.
Publicado: (2026)
Benchmarking VLMs' Reasoning About Persuasive Atypical Images
por: Malakouti, Sina, et al.
Publicado: (2024)
por: Malakouti, Sina, et al.
Publicado: (2024)
WARM-CAT: Warm-Started Test-Time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning
por: Yan, Xudong, et al.
Publicado: (2026)
por: Yan, Xudong, et al.
Publicado: (2026)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
por: Guizilini, Vitor, et al.
Publicado: (2025)
por: Guizilini, Vitor, et al.
Publicado: (2025)
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
por: Yao, Manyi, et al.
Publicado: (2025)
por: Yao, Manyi, et al.
Publicado: (2025)
AnchoredDream: Zero-Shot 360° Indoor Scene Generation from a Single View via Geometric Grounding
por: Yao, Runmao, et al.
Publicado: (2026)
por: Yao, Runmao, et al.
Publicado: (2026)
Med-R2: An Adversarial Benchmark for Evidence-Grounded Reasoning in Medical VLMs
por: Ma, Wen, et al.
Publicado: (2026)
por: Ma, Wen, et al.
Publicado: (2026)
Ejemplares similares
-
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
por: Saha, Oindrila, et al.
Publicado: (2024) -
MolSight: Molecular Property Prediction with Images
por: Baranwal, Aaditya, et al.
Publicado: (2026) -
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
por: Hasanebrahimi, Afsaneh, et al.
Publicado: (2026) -
Vote-in-Context: Turning VLMs into Zero-Shot Rank Fusers
por: Eltahir, Mohamed, et al.
Publicado: (2025) -
Re:Verse -- Can Your VLM Read a Manga?
por: Baranwal, Aaditya, et al.
Publicado: (2025)