Attention Head Entropy of LLMs Predicts Answer Correctness
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ostmeier, Sophie, Axelrod, Brian, Varma, Maya, Aali, Asad, Zhang, Yabin, Paschali, Magdalini, Koyejo, Sanmi, Langlotz, Curtis, Chaudhari, Akshay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LieRE: Lie Rotational Positional Encodings
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
Foundation Models in Radiology: What, How, When, Why and Why Not
von: Paschali, Magdalini, et al.
Veröffentlicht: (2024)
von: Paschali, Magdalini, et al.
Veröffentlicht: (2024)
TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
von: Varma, Maya, et al.
Veröffentlicht: (2025)
von: Varma, Maya, et al.
Veröffentlicht: (2025)
From Detection to Mitigation: Addressing Bias in Deep Learning Models for Chest X-Ray Diagnosis
von: Mottez, Clemence, et al.
Veröffentlicht: (2025)
von: Mottez, Clemence, et al.
Veröffentlicht: (2025)
MedVAL: Toward Expert-Level Medical Text Validation with Language Models
von: Aali, Asad, et al.
Veröffentlicht: (2025)
von: Aali, Asad, et al.
Veröffentlicht: (2025)
Spectral Graph Sample Weighting for Interpretable Sub-cohort Analysis in Predictive Models for Neuroimaging
von: Paschali, Magdalini, et al.
Veröffentlicht: (2024)
von: Paschali, Magdalini, et al.
Veröffentlicht: (2024)
RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models
von: Varma, Maya, et al.
Veröffentlicht: (2024)
von: Varma, Maya, et al.
Veröffentlicht: (2024)
Learning Generalizable 3D Medical Image Representations from Mask-Guided Self-Supervision
von: Gao, Yunhe, et al.
Veröffentlicht: (2026)
von: Gao, Yunhe, et al.
Veröffentlicht: (2026)
MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders
von: Varma, Maya, et al.
Veröffentlicht: (2025)
von: Varma, Maya, et al.
Veröffentlicht: (2025)
Structuring Radiology Reports: Challenging LLMs with Lightweight Models
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
von: Moll, Johannes, et al.
Veröffentlicht: (2025)
Activation Matters: Test-time Activated Negative Labels for OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
SOE: SO(3)-Equivariant 3D MRI Encoding
von: He, Shizhe, et al.
Veröffentlicht: (2024)
von: He, Shizhe, et al.
Veröffentlicht: (2024)
GREEN: Generative Radiology Report Evaluation and Error Notation
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024)
Toward expanding the scope of radiology report summarization to multiple anatomies and modalities
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
von: Chen, Zhihong, et al.
Veröffentlicht: (2022)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
Causally Inspired Regularization Enables Domain General Representations
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2024)
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2024)
Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes
von: Robertson, Zachary, et al.
Veröffentlicht: (2025)
von: Robertson, Zachary, et al.
Veröffentlicht: (2025)
Prompt Triage: Structured Optimization Enhances Vision-Language Model Performance on Medical Imaging Benchmarks
von: Singhvi, Arnav, et al.
Veröffentlicht: (2025)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2025)
Improving Performance, Robustness, and Fairness of Radiographic AI Models with Finely-Controllable Synthetic Data
von: Moroianu, Stefania L., et al.
Veröffentlicht: (2025)
von: Moroianu, Stefania L., et al.
Veröffentlicht: (2025)
In-Context Learning of Energy Functions
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2024)
A Framework for Objective-Driven Dynamical Stochastic Fields
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2025)
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2025)
Discovering Implicit Large Language Model Alignment Objectives
von: Chen, Edward, et al.
Veröffentlicht: (2026)
von: Chen, Edward, et al.
Veröffentlicht: (2026)
Distributional Machine Unlearning via Selective Data Removal
von: Allouah, Youssef, et al.
Veröffentlicht: (2025)
von: Allouah, Youssef, et al.
Veröffentlicht: (2025)
A Reasoning-Enabled Vision-Language Foundation Model for Chest X-ray Interpretation
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
Label Noise Robustness for Domain-Agnostic Fair Corrections via Nearest Neighbors Label Spreading
von: Stromberg, Nathan, et al.
Veröffentlicht: (2024)
von: Stromberg, Nathan, et al.
Veröffentlicht: (2024)
High-Dimensional Markov-switching Ordinary Differential Processes
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
von: Tsai, Katherine, et al.
Veröffentlicht: (2024)
Is Backpropagation Optimal? When Synthetic Gradients Improve Sample Efficiency
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2026)
von: Zhang, Yibo Jacky, et al.
Veröffentlicht: (2026)
Principled Federated Domain Adaptation: Gradient Projection and Auto-Weighting
von: Jiang, Enyi, et al.
Veröffentlicht: (2023)
von: Jiang, Enyi, et al.
Veröffentlicht: (2023)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
von: Vo, Truong, et al.
Veröffentlicht: (2025)
von: Vo, Truong, et al.
Veröffentlicht: (2025)
HiFA: High-fidelity Text-to-3D Generation with Advanced Diffusion Guidance
von: Zhu, Junzhe, et al.
Veröffentlicht: (2023)
von: Zhu, Junzhe, et al.
Veröffentlicht: (2023)
Splitwiser: Efficient LM inference with constrained resources
von: Aali, Asad, et al.
Veröffentlicht: (2025)
von: Aali, Asad, et al.
Veröffentlicht: (2025)
A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling
von: Wang, Chong, et al.
Veröffentlicht: (2026)
von: Wang, Chong, et al.
Veröffentlicht: (2026)
Reasoning Models Don't Just Think Longer, They Move Differently
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
Process Reward Models for Sentence-Level Verification of LVLM Radiology Reports
von: Thomas, Alois, et al.
Veröffentlicht: (2025)
von: Thomas, Alois, et al.
Veröffentlicht: (2025)
Structured Prompts Improve Evaluation of Language Models
von: Aali, Asad, et al.
Veröffentlicht: (2025)
von: Aali, Asad, et al.
Veröffentlicht: (2025)
Pretraining Scaling Laws for Generative Evaluations of Language Models
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
Automated detection of underdiagnosed medical conditions via opportunistic imaging
von: Aali, Asad, et al.
Veröffentlicht: (2024)
von: Aali, Asad, et al.
Veröffentlicht: (2024)
Are Domain Generalization Benchmarks with Accuracy on the Line Misspecified?
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2025)
von: Salaudeen, Olawale, et al.
Veröffentlicht: (2025)
Invariant Aggregator for Defending against Federated Backdoor Attacks
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2022)
von: Wang, Xiaoyang, et al.
Veröffentlicht: (2022)
RadDiff: Describing Differences in Radiology Image Sets with Natural Language
von: Shen, Xiaoxian, et al.
Veröffentlicht: (2026)
von: Shen, Xiaoxian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LieRE: Lie Rotational Positional Encodings
von: Ostmeier, Sophie, et al.
Veröffentlicht: (2024) -
Foundation Models in Radiology: What, How, When, Why and Why Not
von: Paschali, Magdalini, et al.
Veröffentlicht: (2024) -
TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
von: Varma, Maya, et al.
Veröffentlicht: (2025) -
From Detection to Mitigation: Addressing Bias in Deep Learning Models for Chest X-Ray Diagnosis
von: Mottez, Clemence, et al.
Veröffentlicht: (2025) -
MedVAL: Toward Expert-Level Medical Text Validation with Language Models
von: Aali, Asad, et al.
Veröffentlicht: (2025)