FastRM: An efficient and automatic explainability framework for multimodal generative models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Stan, Gabriela Ben-Melech, Aflalo, Estelle, Luo, Man, Rosenman, Shachar, Le, Tiep, Paul, Sayak, Tseng, Shao-Yen, Lal, Vasudev |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
Learning from Reasoning Failures via Synthetic Data Generation
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2025)
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2025)
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
von: Rosenman, Shachar, et al.
Veröffentlicht: (2023)
von: Rosenman, Shachar, et al.
Veröffentlicht: (2023)
DPO Learning with LLMs-Judge Signal for Computer Use Agents
von: Luo, Man, et al.
Veröffentlicht: (2025)
von: Luo, Man, et al.
Veröffentlicht: (2025)
LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2024)
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2024)
Getting it Right: Improving Spatial Consistency in Text-to-Image Models
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2024)
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2024)
Debias your Large Multi-Modal Model at Test-Time via Non-Contrastive Visual Attribute Steering
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
L-MAGIC: Language Model Assisted Generation of Images with Coherence
von: Cai, Zhipeng, et al.
Veröffentlicht: (2024)
von: Cai, Zhipeng, et al.
Veröffentlicht: (2024)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
von: Xu, Xiao, et al.
Veröffentlicht: (2022)
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment
von: Rohekar, Raanan Y., et al.
Veröffentlicht: (2024)
von: Rohekar, Raanan Y., et al.
Veröffentlicht: (2024)
Scaling Knowledge Graph Construction through Synthetic Data Generation and Distillation
von: Choubey, Prafulla Kumar, et al.
Veröffentlicht: (2024)
von: Choubey, Prafulla Kumar, et al.
Veröffentlicht: (2024)
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
von: Hinck, Musashi, et al.
Veröffentlicht: (2024)
von: Hinck, Musashi, et al.
Veröffentlicht: (2024)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
Probing the Representational Power of Sparse Autoencoders in Vision Models
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
Steering Large Language Models to Evaluate and Amplify Creativity
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2024)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2024)
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
von: Howard, Phillip, et al.
Veröffentlicht: (2023)
von: Howard, Phillip, et al.
Veröffentlicht: (2023)
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2023)
von: Madasu, Avinash, et al.
Veröffentlicht: (2023)
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
Is Your Paper Being Reviewed by an LLM? Benchmarking AI Text Detection in Peer Review
von: Yu, Sungduk, et al.
Veröffentlicht: (2025)
von: Yu, Sungduk, et al.
Veröffentlicht: (2025)
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
Training-Free Mitigation of Language Reasoning Degradation After Multimodal Instruction Tuning
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU
von: Spoczynski, Marcin, et al.
Veröffentlicht: (2026)
von: Spoczynski, Marcin, et al.
Veröffentlicht: (2026)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
von: Kundu, Souvik, et al.
Veröffentlicht: (2025)
von: Kundu, Souvik, et al.
Veröffentlicht: (2025)
A simbologia da devoção: o retrato da fé demonstrado pelos ex-votos e a relação com a Igreja midiatizada
von: Ana Maria de Souza Melech
Veröffentlicht: (2015)
von: Ana Maria de Souza Melech
Veröffentlicht: (2015)
FA-Seg: A Fast and Accurate Diffusion-Based Method for Open-Vocabulary Segmentation
von: Che, Huy, et al.
Veröffentlicht: (2025)
von: Che, Huy, et al.
Veröffentlicht: (2025)
Moonflowers and efficient code sparsification
von: Lovett, Shachar, et al.
Veröffentlicht: (2026)
von: Lovett, Shachar, et al.
Veröffentlicht: (2026)
Quantifying and Enabling the Interpretability of CLIP-like Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2024)
von: Madasu, Avinash, et al.
Veröffentlicht: (2024)
SK-VQA: Synthetic Knowledge Generation at Scale for Training Context-Augmented Multimodal LLMs
von: Su, Xin, et al.
Veröffentlicht: (2024)
von: Su, Xin, et al.
Veröffentlicht: (2024)
An interpretable generative multimodal neuroimaging-genomics framework for decoding Alzheimer's disease
von: Dolci, Giorgio, et al.
Veröffentlicht: (2024)
von: Dolci, Giorgio, et al.
Veröffentlicht: (2024)
Tomographic electron flow in confined geometries: Beyond the dual-relaxation time approximation
von: Ben-Shachar, Nitay, et al.
Veröffentlicht: (2025)
von: Ben-Shachar, Nitay, et al.
Veröffentlicht: (2025)
A Foundation Model Approach for Fetal Stress Prediction During Labor From cardiotocography (CTG) recordings
von: Fridman, Naomi, et al.
Veröffentlicht: (2026)
von: Fridman, Naomi, et al.
Veröffentlicht: (2026)
Magnetotransport of tomographic electrons in a channel
von: Ben-Shachar, Nitay, et al.
Veröffentlicht: (2025)
von: Ben-Shachar, Nitay, et al.
Veröffentlicht: (2025)
Pension reforms : is there a tradeoff between efficiency and equity? Estelle James / James Estelle
von: James, Estelle
Veröffentlicht: (1997)
von: James, Estelle
Veröffentlicht: (1997)
Index-MSR: A high-efficiency multimodal fusion framework for speech recognition
von: Chen, Jinming, et al.
Veröffentlicht: (2025)
von: Chen, Jinming, et al.
Veröffentlicht: (2025)
Computer generated physical properties / Stan Bumble
von: Bumble, Stan
von: Bumble, Stan
A Bayesian framework for opinion dynamics models
von: Chen, Yen-Shao, et al.
Veröffentlicht: (2025)
von: Chen, Yen-Shao, et al.
Veröffentlicht: (2025)
Recovering Unobserved Network Links from Aggregated Relational Data: Discussions on Bayesian Latent Surface Modeling and Penalized Regression
von: Tseng, Yen-hsuan
Veröffentlicht: (2025)
von: Tseng, Yen-hsuan
Veröffentlicht: (2025)
Ähnliche Einträge
-
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024) -
Learning from Reasoning Failures via Synthetic Data Generation
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2025) -
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
von: Rosenman, Shachar, et al.
Veröffentlicht: (2023) -
DPO Learning with LLMs-Judge Signal for Computer Use Agents
von: Luo, Man, et al.
Veröffentlicht: (2025) -
LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2024)