Gespeichert in:
| Hauptverfasser: | Baranwal, Aaditya, Yadav, Vishal, Rajora, Abhishek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.10528 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MolSight: Molecular Property Prediction with Images
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
von: Saha, Oindrila, et al.
Veröffentlicht: (2024)
von: Saha, Oindrila, et al.
Veröffentlicht: (2024)
Re:Verse -- Can Your VLM Read a Manga?
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
SynSpill: Improved Industrial Spill Detection With Synthetic Data
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025)
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
von: Hasanebrahimi, Afsaneh, et al.
Veröffentlicht: (2026)
von: Hasanebrahimi, Afsaneh, et al.
Veröffentlicht: (2026)
Vote-in-Context: Turning VLMs into Zero-Shot Rank Fusers
von: Eltahir, Mohamed, et al.
Veröffentlicht: (2025)
von: Eltahir, Mohamed, et al.
Veröffentlicht: (2025)
Exploring Prompt Alignment with Clinical Factors in Zero-Shot Segmentation VLMs for NSCLC Tumor Segmentation
von: Pai, Suraj, et al.
Veröffentlicht: (2026)
von: Pai, Suraj, et al.
Veröffentlicht: (2026)
VROOM - Visual Reconstruction over Onboard Multiview
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
von: Yadav, Yajat, et al.
Veröffentlicht: (2025)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
von: Luo, Kun, et al.
Veröffentlicht: (2026)
von: Luo, Kun, et al.
Veröffentlicht: (2026)
Zero-Shot Monocular Scene Flow Estimation in the Wild
von: Liang, Yiqing, et al.
Veröffentlicht: (2025)
von: Liang, Yiqing, et al.
Veröffentlicht: (2025)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
von: Imam, Raza, et al.
Veröffentlicht: (2025)
von: Imam, Raza, et al.
Veröffentlicht: (2025)
Image Matching by Bare Homography
von: Bellavia, Fabio
Veröffentlicht: (2023)
von: Bellavia, Fabio
Veröffentlicht: (2023)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
von: Xu, Jiacong, et al.
Veröffentlicht: (2025)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2025)
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2025)
TrajRAG: Retrieving Geometric-Semantic Experience for Zero-Shot Object Navigation
von: Wang, Yiyao, et al.
Veröffentlicht: (2026)
von: Wang, Yiyao, et al.
Veröffentlicht: (2026)
ZDySS -- Zero-Shot Dynamic Scene Stylization using Gaussian Splatting
von: Saroha, Abhishek, et al.
Veröffentlicht: (2025)
von: Saroha, Abhishek, et al.
Veröffentlicht: (2025)
From YOLO to VLMs: Advancing Zero-Shot and Few-Shot Detection of Wastewater Treatment Plants Using Satellite Imagery in MENA Region
von: Premarathna, Akila, et al.
Veröffentlicht: (2025)
von: Premarathna, Akila, et al.
Veröffentlicht: (2025)
From Words to Wavelengths: VLMs for Few-Shot Multispectral Object Detection
von: Nkegoum, Manuel, et al.
Veröffentlicht: (2025)
von: Nkegoum, Manuel, et al.
Veröffentlicht: (2025)
TOMCAT: Test-time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning
von: Yan, Xudong, et al.
Veröffentlicht: (2025)
von: Yan, Xudong, et al.
Veröffentlicht: (2025)
ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models
von: Roberts, Jonathan, et al.
Veröffentlicht: (2025)
von: Roberts, Jonathan, et al.
Veröffentlicht: (2025)
Open-Pose 3D Zero-Shot Learning: Benchmark and Challenges
von: Zhao, Weiguang, et al.
Veröffentlicht: (2023)
von: Zhao, Weiguang, et al.
Veröffentlicht: (2023)
MAC: A Benchmark for Multiple Attributes Compositional Zero-Shot Learning
von: Xu, Shuo, et al.
Veröffentlicht: (2024)
von: Xu, Shuo, et al.
Veröffentlicht: (2024)
XDT-CXR: Investigating Cross-Disease Transferability in Zero-Shot Binary Classification of Chest X-Rays
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
Epsilon: Exploring Comprehensive Visual-Semantic Projection for Multi-Label Zero-Shot Learning
von: Liu, Ziming, et al.
Veröffentlicht: (2024)
von: Liu, Ziming, et al.
Veröffentlicht: (2024)
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
von: Sony, Redwan, et al.
Veröffentlicht: (2025)
von: Sony, Redwan, et al.
Veröffentlicht: (2025)
Will It Zero-Shot?: Predicting Zero-Shot Classification Performance For Arbitrary Queries
von: Robbins, Kevin, et al.
Veröffentlicht: (2026)
von: Robbins, Kevin, et al.
Veröffentlicht: (2026)
DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs
von: Pan, Jiazhen, et al.
Veröffentlicht: (2026)
von: Pan, Jiazhen, et al.
Veröffentlicht: (2026)
Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports
von: Yang, Yuchen, et al.
Veröffentlicht: (2026)
von: Yang, Yuchen, et al.
Veröffentlicht: (2026)
DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving
von: Vo, Hao, et al.
Veröffentlicht: (2026)
von: Vo, Hao, et al.
Veröffentlicht: (2026)
Few-Shot, Now for Real: Medical VLMs Adaptation without Balanced Sets or Validation
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
Benchmarking VLMs' Reasoning About Persuasive Atypical Images
von: Malakouti, Sina, et al.
Veröffentlicht: (2024)
von: Malakouti, Sina, et al.
Veröffentlicht: (2024)
Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment
von: Chen, Jingkun, et al.
Veröffentlicht: (2026)
von: Chen, Jingkun, et al.
Veröffentlicht: (2026)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
von: Guizilini, Vitor, et al.
Veröffentlicht: (2025)
von: Guizilini, Vitor, et al.
Veröffentlicht: (2025)
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
von: Yao, Manyi, et al.
Veröffentlicht: (2025)
von: Yao, Manyi, et al.
Veröffentlicht: (2025)
CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
von: Fang, Zhengru, et al.
Veröffentlicht: (2026)
Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos
von: Han, Mingfei, et al.
Veröffentlicht: (2023)
von: Han, Mingfei, et al.
Veröffentlicht: (2023)
LISR: Learning Linear 3D Implicit Surface Representation Using Compactly Supported Radial Basis Functions
von: Pandey, Atharva, et al.
Veröffentlicht: (2024)
von: Pandey, Atharva, et al.
Veröffentlicht: (2024)
WARM-CAT: Warm-Started Test-Time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning
von: Yan, Xudong, et al.
Veröffentlicht: (2026)
von: Yan, Xudong, et al.
Veröffentlicht: (2026)
When Does RL Help Medical VLMs? Disentangling Vision, SFT, and RL Gains
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2026)
von: Jeddi, Ahmadreza, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MolSight: Molecular Property Prediction with Images
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026) -
Improved Zero-Shot Classification by Adapting VLMs with Text Descriptions
von: Saha, Oindrila, et al.
Veröffentlicht: (2024) -
Re:Verse -- Can Your VLM Read a Manga?
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025) -
SynSpill: Improved Industrial Spill Detection With Synthetic Data
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2025) -
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
von: Hasanebrahimi, Afsaneh, et al.
Veröffentlicht: (2026)