Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following
Fuente:
arXiv
Salvato in:
| Autori principali: | Kang, Myeongkyun, Kim, Soopil, Li, Xiaoxiao, Park, Sang Hyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Model Agnostic Preference Optimization for Medical Image Segmentation
di: Nam, Yunseong, et al.
Pubblicazione: (2025)
di: Nam, Yunseong, et al.
Pubblicazione: (2025)
Few Shot Part Segmentation Reveals Compositional Logic for Industrial Anomaly Detection
di: Kim, Soopil, et al.
Pubblicazione: (2023)
di: Kim, Soopil, et al.
Pubblicazione: (2023)
Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA
di: Jin, Ruinan, et al.
Pubblicazione: (2026)
di: Jin, Ruinan, et al.
Pubblicazione: (2026)
LoFi: Location-Aware Fine-Grained Representation Learning for Chest X-ray
di: Kang, Myeongkyun, et al.
Pubblicazione: (2026)
di: Kang, Myeongkyun, et al.
Pubblicazione: (2026)
Instruction-Following Evaluation of Large Vision-Language Models
di: Shiono, Daiki, et al.
Pubblicazione: (2025)
di: Shiono, Daiki, et al.
Pubblicazione: (2025)
Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning
di: Park, Dongmin, et al.
Pubblicazione: (2024)
di: Park, Dongmin, et al.
Pubblicazione: (2024)
InstructTA: Instruction-Tuned Targeted Attack for Large Vision-Language Models
di: Wang, Xunguang, et al.
Pubblicazione: (2023)
di: Wang, Xunguang, et al.
Pubblicazione: (2023)
Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
di: Zhang, Jinrui, et al.
Pubblicazione: (2024)
Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
di: Cao, Meng, et al.
Pubblicazione: (2024)
di: Cao, Meng, et al.
Pubblicazione: (2024)
Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning
di: Hu, Rui, et al.
Pubblicazione: (2024)
di: Hu, Rui, et al.
Pubblicazione: (2024)
MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models
di: Yan, Qiao, et al.
Pubblicazione: (2025)
di: Yan, Qiao, et al.
Pubblicazione: (2025)
ST-VLM: Kinematic Instruction Tuning for Spatio-Temporal Reasoning in Vision-Language Models
di: Ko, Dohwan, et al.
Pubblicazione: (2025)
di: Ko, Dohwan, et al.
Pubblicazione: (2025)
On the Evaluation and Refinement of Vision-Language Instruction Tuning Datasets
di: Liao, Ning, et al.
Pubblicazione: (2023)
di: Liao, Ning, et al.
Pubblicazione: (2023)
Do we Really Need Visual Instructions? Towards Visual Instruction-Free Fine-tuning for Large Vision-Language Models
di: Liu, Zikang, et al.
Pubblicazione: (2025)
di: Liu, Zikang, et al.
Pubblicazione: (2025)
InstructCV: Instruction-Tuned Text-to-Image Diffusion Models as Vision Generalists
di: Gan, Yulu, et al.
Pubblicazione: (2023)
di: Gan, Yulu, et al.
Pubblicazione: (2023)
Efficient Inference of Vision Instruction-Following Models with Elastic Cache
di: Liu, Zuyan, et al.
Pubblicazione: (2024)
di: Liu, Zuyan, et al.
Pubblicazione: (2024)
SkyEyeGPT: Unifying Remote Sensing Vision-Language Tasks via Instruction Tuning with Large Language Model
di: Zhan, Yang, et al.
Pubblicazione: (2024)
di: Zhan, Yang, et al.
Pubblicazione: (2024)
Instruction Tuning of Large Language Models for Tabular Data Generation-in One Day
di: Abdollahzadeh, Milad, et al.
Pubblicazione: (2025)
di: Abdollahzadeh, Milad, et al.
Pubblicazione: (2025)
Enhancing Model Performance: Another Approach to Vision-Language Instruction Tuning
di: Vedanshu, et al.
Pubblicazione: (2024)
di: Vedanshu, et al.
Pubblicazione: (2024)
Incentivizing Reasoning for Advanced Instruction-Following of Large Language Models
di: Qin, Yulei, et al.
Pubblicazione: (2025)
di: Qin, Yulei, et al.
Pubblicazione: (2025)
GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest
di: Zhang, Shilong, et al.
Pubblicazione: (2023)
di: Zhang, Shilong, et al.
Pubblicazione: (2023)
Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks
di: Kim, Jonghun, et al.
Pubblicazione: (2026)
di: Kim, Jonghun, et al.
Pubblicazione: (2026)
Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models
di: Li, Bin, et al.
Pubblicazione: (2025)
di: Li, Bin, et al.
Pubblicazione: (2025)
LayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding
di: Luo, Chuwei, et al.
Pubblicazione: (2024)
di: Luo, Chuwei, et al.
Pubblicazione: (2024)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
di: Yang, Shuai, et al.
Pubblicazione: (2025)
di: Yang, Shuai, et al.
Pubblicazione: (2025)
Double Visual Defense: Adversarial Pre-training and Instruction Tuning for Improving Vision-Language Model Robustness
di: Wang, Zeyu, et al.
Pubblicazione: (2025)
di: Wang, Zeyu, et al.
Pubblicazione: (2025)
From Generalist to Specialist: Adapting Vision Language Models via Task-Specific Visual Instruction Tuning
di: Bai, Yang, et al.
Pubblicazione: (2024)
di: Bai, Yang, et al.
Pubblicazione: (2024)
AVION: Aerial Vision-Language Instruction from Offline Teacher to Prompt-Tuned Network
di: Hu, Yu, et al.
Pubblicazione: (2026)
di: Hu, Yu, et al.
Pubblicazione: (2026)
SEAGULL: No-reference Image Quality Assessment for Regions of Interest via Vision-Language Instruction Tuning
di: Chen, Zewen, et al.
Pubblicazione: (2024)
di: Chen, Zewen, et al.
Pubblicazione: (2024)
Is 'Right' Right? Enhancing Object Orientation Understanding in Multimodal Large Language Models through Egocentric Instruction Tuning
di: Jung, Ji Hyeok, et al.
Pubblicazione: (2024)
di: Jung, Ji Hyeok, et al.
Pubblicazione: (2024)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
di: Kim, Dain, et al.
Pubblicazione: (2026)
di: Kim, Dain, et al.
Pubblicazione: (2026)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
di: Kang, Weitai, et al.
Pubblicazione: (2024)
di: Kang, Weitai, et al.
Pubblicazione: (2024)
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning
di: You, Zebin, et al.
Pubblicazione: (2025)
di: You, Zebin, et al.
Pubblicazione: (2025)
Text as Images: Can Multimodal Large Language Models Follow Printed Instructions in Pixels?
di: Li, Xiujun, et al.
Pubblicazione: (2023)
di: Li, Xiujun, et al.
Pubblicazione: (2023)
Prompter: Utilizing Large Language Model Prompting for a Data Efficient Embodied Instruction Following
di: Inoue, Yuki, et al.
Pubblicazione: (2022)
di: Inoue, Yuki, et al.
Pubblicazione: (2022)
Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models
di: Jin, Hyundong, et al.
Pubblicazione: (2025)
di: Jin, Hyundong, et al.
Pubblicazione: (2025)
Multimodal Instruction Tuning with Hybrid State Space Models
di: Zhou, Jianing, et al.
Pubblicazione: (2024)
di: Zhou, Jianing, et al.
Pubblicazione: (2024)
Deep Instruction Tuning for Segment Anything Model
di: Huang, Xiaorui, et al.
Pubblicazione: (2024)
di: Huang, Xiaorui, et al.
Pubblicazione: (2024)
Mixture of Cluster-conditional LoRA Experts for Vision-language Instruction Tuning
di: Gou, Yunhao, et al.
Pubblicazione: (2023)
di: Gou, Yunhao, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Model Agnostic Preference Optimization for Medical Image Segmentation
di: Nam, Yunseong, et al.
Pubblicazione: (2025) -
Few Shot Part Segmentation Reveals Compositional Logic for Industrial Anomaly Detection
di: Kim, Soopil, et al.
Pubblicazione: (2023) -
Verification Mirage: Mapping the Reliability Boundary of Self-Verification in Medical VQA
di: Jin, Ruinan, et al.
Pubblicazione: (2026) -
LoFi: Location-Aware Fine-Grained Representation Learning for Chest X-ray
di: Kang, Myeongkyun, et al.
Pubblicazione: (2026) -
Instruction-Following Evaluation of Large Vision-Language Models
di: Shiono, Daiki, et al.
Pubblicazione: (2025)