WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Pingyi, Zhu, Chenglu, Zheng, Sunyi, Li, Honglin, Yang, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
von: Chen, Pingyi, et al.
Veröffentlicht: (2023)
von: Chen, Pingyi, et al.
Veröffentlicht: (2023)
Attention-Challenging Multiple Instance Learning for Whole Slide Image Classification
von: Zhang, Yunlong, et al.
Veröffentlicht: (2023)
von: Zhang, Yunlong, et al.
Veröffentlicht: (2023)
Rethinking Transformer for Long Contextual Histopathology Whole Slide Image Analysis
von: Li, Honglin, et al.
Veröffentlicht: (2024)
von: Li, Honglin, et al.
Veröffentlicht: (2024)
VQA$^2$: Visual Question Answering for Video Quality Assessment
von: Jia, Ziheng, et al.
Veröffentlicht: (2024)
von: Jia, Ziheng, et al.
Veröffentlicht: (2024)
CoralVQA: A Large-Scale Visual Question Answering Dataset for Coral Reef Image Understanding
von: Han, Hongyong, et al.
Veröffentlicht: (2025)
von: Han, Hongyong, et al.
Veröffentlicht: (2025)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
von: Vu, Sinh Trong, et al.
Veröffentlicht: (2025)
von: Vu, Sinh Trong, et al.
Veröffentlicht: (2025)
SCRA-VQA: Summarized Caption-Rerank for Augmented Large Language Models in Visual Question Answering
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA
von: Yang, Chunze, et al.
Veröffentlicht: (2026)
von: Yang, Chunze, et al.
Veröffentlicht: (2026)
PathVQ: Reforming Computational Pathology Foundation Model for Whole Slide Image Analysis via Vector Quantization
von: Li, Honglin, et al.
Veröffentlicht: (2025)
von: Li, Honglin, et al.
Veröffentlicht: (2025)
MaS-VQA: A Mask-and-Select Framework for Knowledge-Based Visual Question Answering
von: Mao, Xianwei, et al.
Veröffentlicht: (2026)
von: Mao, Xianwei, et al.
Veröffentlicht: (2026)
VSA4VQA: Scaling a Vector Symbolic Architecture to Visual Question Answering on Natural Images
von: Penzkofer, Anna, et al.
Veröffentlicht: (2024)
von: Penzkofer, Anna, et al.
Veröffentlicht: (2024)
From Image to Language: A Critical Analysis of Visual Question Answering (VQA) Approaches, Challenges, and Opportunities
von: Ishmam, Md Farhan, et al.
Veröffentlicht: (2023)
von: Ishmam, Md Farhan, et al.
Veröffentlicht: (2023)
MGA-VQA: Secure and Interpretable Graph-Augmented Visual Question Answering with Memory-Guided Protection Against Unauthorized Knowledge Use
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2025)
M$^3$-VQA: A Benchmark for Multimodal, Multi-Entity, Multi-Hop Visual Question Answering
von: Ma, Jiatong, et al.
Veröffentlicht: (2026)
von: Ma, Jiatong, et al.
Veröffentlicht: (2026)
TableVQA-Bench: A Visual Question Answering Benchmark on Multiple Table Domains
von: Kim, Yoonsik, et al.
Veröffentlicht: (2024)
von: Kim, Yoonsik, et al.
Veröffentlicht: (2024)
Towards Effective and Efficient Context-aware Nucleus Detection in Histopathology Whole Slide Images
von: Shui, Zhongyi, et al.
Veröffentlicht: (2025)
von: Shui, Zhongyi, et al.
Veröffentlicht: (2025)
Object Attribute Matters in Visual Question Answering
von: Li, Peize, et al.
Veröffentlicht: (2023)
von: Li, Peize, et al.
Veröffentlicht: (2023)
AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification
von: Zhang, Yunlong, et al.
Veröffentlicht: (2024)
von: Zhang, Yunlong, et al.
Veröffentlicht: (2024)
AutoViVQA: A Large-Scale Automatically Constructed Dataset for Vietnamese Visual Question Answering
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
von: Tuong, Nguyen Anh, et al.
Veröffentlicht: (2026)
TAKT: Target-Aware Knowledge Transfer for Whole Slide Image Classification
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
SAC-MIL: Spatial-Aware Correlated Multiple Instance Learning for Histopathology Whole Slide Image Classification
von: Bai, Yu, et al.
Veröffentlicht: (2025)
von: Bai, Yu, et al.
Veröffentlicht: (2025)
PitVQA++: Vector Matrix-Low-Rank Adaptation for Open-Ended Visual Question Answering in Pituitary Surgery
von: He, Runlong, et al.
Veröffentlicht: (2025)
von: He, Runlong, et al.
Veröffentlicht: (2025)
RadImageNet-VQA: A Large-Scale CT and MRI Dataset for Radiologic Visual Question Answering
von: Butsanets, Léo, et al.
Veröffentlicht: (2025)
von: Butsanets, Léo, et al.
Veröffentlicht: (2025)
ProtoVQA: An Adaptable Prototypical Framework for Explainable Fine-Grained Visual Question Answering
von: Diao, Xingjian, et al.
Veröffentlicht: (2025)
von: Diao, Xingjian, et al.
Veröffentlicht: (2025)
Adversarial Attacks on VQA-NLE: Exposing and Alleviating Inconsistencies in Visual Question Answering Explanations
von: Yeh, Yahsin, et al.
Veröffentlicht: (2025)
von: Yeh, Yahsin, et al.
Veröffentlicht: (2025)
UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation
von: Ghosh, Shiv, et al.
Veröffentlicht: (2026)
von: Ghosh, Shiv, et al.
Veröffentlicht: (2026)
Multi-Sourced Compositional Generalization in Visual Question Answering
von: Li, Chuanhao, et al.
Veröffentlicht: (2025)
von: Li, Chuanhao, et al.
Veröffentlicht: (2025)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
von: Ahir, Param, et al.
Veröffentlicht: (2023)
von: Ahir, Param, et al.
Veröffentlicht: (2023)
Hypergraph Mamba for Efficient Whole Slide Image Understanding
von: Lu, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Lu, Jiaxuan, et al.
Veröffentlicht: (2025)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
von: Chen, Ying, et al.
Veröffentlicht: (2024)
von: Chen, Ying, et al.
Veröffentlicht: (2024)
Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts
von: Özdemir, Övgü, et al.
Veröffentlicht: (2024)
von: Özdemir, Övgü, et al.
Veröffentlicht: (2024)
TinyVQA: Compact Multimodal Deep Neural Network for Visual Question Answering on Resource-Constrained Devices
von: Rashid, Hasib-Al, et al.
Veröffentlicht: (2024)
von: Rashid, Hasib-Al, et al.
Veröffentlicht: (2024)
CPath-Omni: A Unified Multimodal Foundation Model for Patch and Whole Slide Image Analysis in Computational Pathology
von: Sun, Yuxuan, et al.
Veröffentlicht: (2024)
von: Sun, Yuxuan, et al.
Veröffentlicht: (2024)
Dynamic Clue Bottlenecks: Towards Interpretable-by-Design Visual Question Answering
von: Fu, Xingyu, et al.
Veröffentlicht: (2023)
von: Fu, Xingyu, et al.
Veröffentlicht: (2023)
LCV2: An Efficient Pretraining-Free Framework for Grounded Visual Question Answering
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
von: Chen, Yuhan, et al.
Veröffentlicht: (2024)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
PlantVillageVQA: A Visual Question Answering Dataset for Benchmarking Vision-Language Models in Plant Science
von: Sakib, Syed Nazmus, et al.
Veröffentlicht: (2025)
von: Sakib, Syed Nazmus, et al.
Veröffentlicht: (2025)
Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering
von: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Veröffentlicht: (2026)
von: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Veröffentlicht: (2026)
Large-scale cervical precancerous screening via AI-assisted cytology whole slide image analysis
von: Li, Honglin, et al.
Veröffentlicht: (2024)
von: Li, Honglin, et al.
Veröffentlicht: (2024)
MoReVQA: Exploring Modular Reasoning Models for Video Question Answering
von: Min, Juhong, et al.
Veröffentlicht: (2024)
von: Min, Juhong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
von: Chen, Pingyi, et al.
Veröffentlicht: (2023) -
Attention-Challenging Multiple Instance Learning for Whole Slide Image Classification
von: Zhang, Yunlong, et al.
Veröffentlicht: (2023) -
Rethinking Transformer for Long Contextual Histopathology Whole Slide Image Analysis
von: Li, Honglin, et al.
Veröffentlicht: (2024) -
VQA$^2$: Visual Question Answering for Video Quality Assessment
von: Jia, Ziheng, et al.
Veröffentlicht: (2024) -
CoralVQA: A Large-Scale Visual Question Answering Dataset for Coral Reef Image Understanding
von: Han, Hongyong, et al.
Veröffentlicht: (2025)