Evian: Towards Explainable Visual Instruction-tuning Data Auditing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jia, Zimu, Xu, Mingjie, Estornell, Andrew, Wei, Jiaheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild
von: Masry, Ahmed, et al.
Veröffentlicht: (2024)
von: Masry, Ahmed, et al.
Veröffentlicht: (2024)
HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data
von: Yu, Qifan, et al.
Veröffentlicht: (2023)
von: Yu, Qifan, et al.
Veröffentlicht: (2023)
Coherent Zero-Shot Visual Instruction Generation
von: Phung, Quynh, et al.
Veröffentlicht: (2024)
von: Phung, Quynh, et al.
Veröffentlicht: (2024)
Instruction-tuned Self-Questioning Framework for Multimodal Reasoning
von: Jang, You-Won, et al.
Veröffentlicht: (2025)
von: Jang, You-Won, et al.
Veröffentlicht: (2025)
MotIF: Motion Instruction Fine-tuning
von: Hwang, Minyoung, et al.
Veröffentlicht: (2024)
von: Hwang, Minyoung, et al.
Veröffentlicht: (2024)
StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs
von: Che, Chang, et al.
Veröffentlicht: (2026)
von: Che, Chang, et al.
Veröffentlicht: (2026)
VIGC: Visual Instruction Generation and Correction
von: Wang, Bin, et al.
Veröffentlicht: (2023)
von: Wang, Bin, et al.
Veröffentlicht: (2023)
CoralVQA: A Large-Scale Visual Question Answering Dataset for Coral Reef Image Understanding
von: Han, Hongyong, et al.
Veröffentlicht: (2025)
von: Han, Hongyong, et al.
Veröffentlicht: (2025)
Activating Visual Context and Commonsense Reasoning through Masked Prediction in VLMs
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning
von: Safaei, Bardia, et al.
Veröffentlicht: (2025)
von: Safaei, Bardia, et al.
Veröffentlicht: (2025)
VisualWebInstruct: Scaling up Multimodal Instruction Data through Web Search
von: Jia, Yiming, et al.
Veröffentlicht: (2025)
von: Jia, Yiming, et al.
Veröffentlicht: (2025)
Fine-tuning Vision Language Models with Graph-based Knowledge for Explainable Medical Image Analysis
von: Li, Chenjun, et al.
Veröffentlicht: (2025)
von: Li, Chenjun, et al.
Veröffentlicht: (2025)
Fake-in-Facext: Towards Fine-Grained Explainable DeepFake Analysis
von: Qin, Lixiong, et al.
Veröffentlicht: (2025)
von: Qin, Lixiong, et al.
Veröffentlicht: (2025)
Towards Visual-Prompt Temporal Answering Grounding in Medical Instructional Video
von: Li, Bin, et al.
Veröffentlicht: (2022)
von: Li, Bin, et al.
Veröffentlicht: (2022)
Supervised Fine-tuning in turn Improves Visual Foundation Models
von: Jiang, Xiaohu, et al.
Veröffentlicht: (2024)
von: Jiang, Xiaohu, et al.
Veröffentlicht: (2024)
LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning
von: Che, Chang, et al.
Veröffentlicht: (2025)
von: Che, Chang, et al.
Veröffentlicht: (2025)
InstrAct: Towards Action-Centric Understanding in Instructional Videos
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2026)
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2026)
ScalSelect: Scalable Training-Free Multimodal Data Selection for Efficient Visual Instruction Tuning
von: Wu, Changti, et al.
Veröffentlicht: (2026)
von: Wu, Changti, et al.
Veröffentlicht: (2026)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
B-AVIBench: Towards Evaluating the Robustness of Large Vision-Language Model on Black-box Adversarial Visual-Instructions
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation
von: Airinei, Daniel, et al.
Veröffentlicht: (2025)
von: Airinei, Daniel, et al.
Veröffentlicht: (2025)
Robotic Visual Instruction
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
von: Li, Yanbang, et al.
Veröffentlicht: (2025)
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy
von: Yang, Te, et al.
Veröffentlicht: (2024)
von: Yang, Te, et al.
Veröffentlicht: (2024)
TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types
von: Chen, Jiankang, et al.
Veröffentlicht: (2025)
von: Chen, Jiankang, et al.
Veröffentlicht: (2025)
In-Video Instructions: Visual Signals as Generative Control
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
Tokensome: Towards a Genetic Vision-Language GPT for Explainable and Cognitive Karyotyping
von: Zhang, Haoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Haoxi, et al.
Veröffentlicht: (2024)
Explainable Visual Anomaly Detection via Concept Bottleneck Models
von: Stropeni, Arianna, et al.
Veröffentlicht: (2025)
von: Stropeni, Arianna, et al.
Veröffentlicht: (2025)
Advancing Multimodal Large Language Models in Chart Question Answering with Visualization-Referenced Instruction Tuning
von: Zeng, Xingchen, et al.
Veröffentlicht: (2024)
von: Zeng, Xingchen, et al.
Veröffentlicht: (2024)
Parrot: Multilingual Visual Instruction Tuning
von: Sun, Hai-Long, et al.
Veröffentlicht: (2024)
von: Sun, Hai-Long, et al.
Veröffentlicht: (2024)
Membership Inference Test: Auditing Training Data in Object Classification Models
von: Mancera, Gonzalo, et al.
Veröffentlicht: (2026)
von: Mancera, Gonzalo, et al.
Veröffentlicht: (2026)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
Parameter-Free Fine-tuning via Redundancy Elimination for Vision Foundation Models
von: Long, Jiahuan, et al.
Veröffentlicht: (2025)
von: Long, Jiahuan, et al.
Veröffentlicht: (2025)
Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection
von: Tsigos, Konstantinos, et al.
Veröffentlicht: (2024)
von: Tsigos, Konstantinos, et al.
Veröffentlicht: (2024)
Towards Counterfactual and Contrastive Explainability and Transparency of DCNN Image Classifiers
von: Tariq, Syed Ali, et al.
Veröffentlicht: (2025)
von: Tariq, Syed Ali, et al.
Veröffentlicht: (2025)
Learning to Correction: Explainable Feedback Generation for Visual Commonsense Reasoning Distractor
von: Chen, Jiali, et al.
Veröffentlicht: (2024)
von: Chen, Jiali, et al.
Veröffentlicht: (2024)
Auditing and Mitigating Bias in Gender Classification Algorithms: A Data-Centric Approach
von: Bahiru, Tadesse K, et al.
Veröffentlicht: (2025)
von: Bahiru, Tadesse K, et al.
Veröffentlicht: (2025)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
von: Xie, Hongxia, et al.
Veröffentlicht: (2024)
von: Xie, Hongxia, et al.
Veröffentlicht: (2024)
On Thin Ice: Towards Explainable Conservation Monitoring via Attribution and Perturbations
von: Zhou, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhou, Jiayi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
von: Xu, Mingjie, et al.
Veröffentlicht: (2025) -
ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild
von: Masry, Ahmed, et al.
Veröffentlicht: (2024) -
HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data
von: Yu, Qifan, et al.
Veröffentlicht: (2023) -
Coherent Zero-Shot Visual Instruction Generation
von: Phung, Quynh, et al.
Veröffentlicht: (2024) -
Instruction-tuned Self-Questioning Framework for Multimodal Reasoning
von: Jang, You-Won, et al.
Veröffentlicht: (2025)