See It, Say It, Sorted: An Iterative Training-Free Framework for Visually-Grounded Multimodal Reasoning in LVLMs

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhang, Yongchang, Ma, Oliver, Liu, Tianyi, Zhou, Guangquan, Chen, Yang
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!