LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Stan, Gabriela Ben Melech, Aflalo, Estelle, Rohekar, Raanan Yehezkel, Bhiwandiwalla, Anahita, Tseng, Shao-Yen, Olson, Matthew Lyle, Gurwicz, Yaniv, Wu, Chenfei, Duan, Nan, Lal, Vasudev |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment
par: Rohekar, Raanan Y., et autres
Publié: (2024)
par: Rohekar, Raanan Y., et autres
Publié: (2024)
ClimDetect: A Benchmark Dataset for Climate Change Detection and Attribution
par: Yu, Sungduk, et autres
Publié: (2024)
par: Yu, Sungduk, et autres
Publié: (2024)
Learning from Reasoning Failures via Synthetic Data Generation
par: Stan, Gabriela Ben Melech, et autres
Publié: (2025)
par: Stan, Gabriela Ben Melech, et autres
Publié: (2025)
FastRM: An efficient and automatic explainability framework for multimodal generative models
par: Stan, Gabriela Ben-Melech, et autres
Publié: (2024)
par: Stan, Gabriela Ben-Melech, et autres
Publié: (2024)
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
par: Aflalo, Estelle, et autres
Publié: (2024)
par: Aflalo, Estelle, et autres
Publié: (2024)
Debias your Large Multi-Modal Model at Test-Time via Non-Contrastive Visual Attribute Steering
par: Ratzlaff, Neale, et autres
Publié: (2024)
par: Ratzlaff, Neale, et autres
Publié: (2024)
Why do LLaVA Vision-Language Models Reply to Images in English?
par: Hinck, Musashi, et autres
Publié: (2024)
par: Hinck, Musashi, et autres
Publié: (2024)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
par: Kundu, Souvik, et autres
Publié: (2025)
par: Kundu, Souvik, et autres
Publié: (2025)
Getting it Right: Improving Spatial Consistency in Text-to-Image Models
par: Chatterjee, Agneet, et autres
Publié: (2024)
par: Chatterjee, Agneet, et autres
Publié: (2024)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
par: Ratzlaff, Neale, et autres
Publié: (2024)
par: Ratzlaff, Neale, et autres
Publié: (2024)
L-MAGIC: Language Model Assisted Generation of Images with Coherence
par: Cai, Zhipeng, et autres
Publié: (2024)
par: Cai, Zhipeng, et autres
Publié: (2024)
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
par: Hinck, Musashi, et autres
Publié: (2024)
par: Hinck, Musashi, et autres
Publié: (2024)
Steering Large Language Models to Evaluate and Amplify Creativity
par: Olson, Matthew Lyle, et autres
Publié: (2024)
par: Olson, Matthew Lyle, et autres
Publié: (2024)
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
par: Howard, Phillip, et autres
Publié: (2023)
par: Howard, Phillip, et autres
Publié: (2023)
Probing the Representational Power of Sparse Autoencoders in Vision Models
par: Olson, Matthew Lyle, et autres
Publié: (2025)
par: Olson, Matthew Lyle, et autres
Publié: (2025)
Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders
par: Olson, Matthew Lyle, et autres
Publié: (2025)
par: Olson, Matthew Lyle, et autres
Publié: (2025)
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
par: Olson, Matthew Lyle, et autres
Publié: (2026)
par: Olson, Matthew Lyle, et autres
Publié: (2026)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
par: Xu, Xiao, et autres
Publié: (2022)
par: Xu, Xiao, et autres
Publié: (2022)
Quantifying and Enabling the Interpretability of CLIP-like Models
par: Madasu, Avinash, et autres
Publié: (2024)
par: Madasu, Avinash, et autres
Publié: (2024)
DPO Learning with LLMs-Judge Signal for Computer Use Agents
par: Luo, Man, et autres
Publié: (2025)
par: Luo, Man, et autres
Publié: (2025)
Uncovering Bias in Large Vision-Language Models at Scale with Counterfactuals
par: Howard, Phillip, et autres
Publié: (2024)
par: Howard, Phillip, et autres
Publié: (2024)
Uncovering Bias in Large Vision-Language Models with Counterfactuals
par: Howard, Phillip, et autres
Publié: (2024)
par: Howard, Phillip, et autres
Publié: (2024)
Symmetry Breaking in Transformers for Efficient and Interpretable Training
par: Silverstein, Eva, et autres
Publié: (2026)
par: Silverstein, Eva, et autres
Publié: (2026)
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
par: Madasu, Avinash, et autres
Publié: (2023)
par: Madasu, Avinash, et autres
Publié: (2023)
Probing Semantic Routing in Large Mixture-of-Expert Models
par: Olson, Matthew Lyle, et autres
Publié: (2025)
par: Olson, Matthew Lyle, et autres
Publié: (2025)
Más allá de la incertidumbre: lo inconcebible.
par: Yehezkel Dror
Publié: (2001)
par: Yehezkel Dror
Publié: (2001)
Quantum Gradient Class Activation Map for Model Interpretability
par: Lin, Hsin-Yi, et autres
Publié: (2024)
par: Lin, Hsin-Yi, et autres
Publié: (2024)
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
par: Rosenman, Shachar, et autres
Publié: (2023)
par: Rosenman, Shachar, et autres
Publié: (2023)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
par: Madasu, Avinash, et autres
Publié: (2025)
par: Madasu, Avinash, et autres
Publié: (2025)
Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
par: Madasu, Avinash, et autres
Publié: (2025)
par: Madasu, Avinash, et autres
Publié: (2025)
Improving LLM Interpretability and Performance via Guided Embedding Refinement for Sequential Recommendation
par: Jia, Nanshan, et autres
Publié: (2025)
par: Jia, Nanshan, et autres
Publié: (2025)
Robust and Interpretable Graph Neural Networks for Power Systems State Estimation
par: Yaniv, Arbel, et autres
Publié: (2026)
par: Yaniv, Arbel, et autres
Publié: (2026)
Data-Centric Interpretability for LLM-based Multi-Agent Reinforcement Learning
par: Yan, John, et autres
Publié: (2026)
par: Yan, John, et autres
Publié: (2026)
Interpreting Moment Matrix Blocks Spectra using Mutual Shadow Area
par: Brick, Yaniv, et autres
Publié: (2026)
par: Brick, Yaniv, et autres
Publié: (2026)
A simbologia da devoção: o retrato da fé demonstrado pelos ex-votos e a relação com a Igreja midiatizada
par: Ana Maria de Souza Melech
Publié: (2015)
par: Ana Maria de Souza Melech
Publié: (2015)
Mechanistic Interpretability of Emotion Inference in Large Language Models
par: Tak, Ala N., et autres
Publié: (2025)
par: Tak, Ala N., et autres
Publié: (2025)
“For the public benefit”: Data policy in platform markets
par: Sarit Markovich, et autres
Publié: (2024)
par: Sarit Markovich, et autres
Publié: (2024)
Regulating Platform Competition in Markets with Network Externalities: Will Predatory Pricing Restrictions Increase Social Welfare?*
par: Ohad Atad, et autres
Publié: (2024)
par: Ohad Atad, et autres
Publié: (2024)
Vertical Collusion to Exclude Product Improvement*
par: David Gilo, et autres
Publié: (2024)
par: David Gilo, et autres
Publié: (2024)
Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU
par: Spoczynski, Marcin, et autres
Publié: (2026)
par: Spoczynski, Marcin, et autres
Publié: (2026)
Documents similaires
-
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment
par: Rohekar, Raanan Y., et autres
Publié: (2024) -
ClimDetect: A Benchmark Dataset for Climate Change Detection and Attribution
par: Yu, Sungduk, et autres
Publié: (2024) -
Learning from Reasoning Failures via Synthetic Data Generation
par: Stan, Gabriela Ben Melech, et autres
Publié: (2025) -
FastRM: An efficient and automatic explainability framework for multimodal generative models
par: Stan, Gabriela Ben-Melech, et autres
Publié: (2024) -
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
par: Aflalo, Estelle, et autres
Publié: (2024)