Investigating the Impact of Model Instability on Explanations and Uncertainty
Fuente:
arXiv
Guardado en:
| Autores principales: | Marjanović, Sara Vera, Augenstein, Isabelle, Lioma, Christina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
por: Marjanović, Sara Vera, et al.
Publicado: (2024)
por: Marjanović, Sara Vera, et al.
Publicado: (2024)
Joint Extraction and Classification of Danish Competences for Job Matching
por: Li, Qiuchi, et al.
Publicado: (2024)
por: Li, Qiuchi, et al.
Publicado: (2024)
Quantifying Gender Biases Towards Politicians on Reddit
por: Marjanovic, Sara, et al.
Publicado: (2021)
por: Marjanovic, Sara, et al.
Publicado: (2021)
A Reality Check on Context Utilisation for Retrieval-Augmented Generation
por: Hagström, Lovisa, et al.
Publicado: (2024)
por: Hagström, Lovisa, et al.
Publicado: (2024)
Semantic Sensitivities and Inconsistent Predictions: Measuring the Fragility of NLI Models
por: Arakelyan, Erik, et al.
Publicado: (2024)
por: Arakelyan, Erik, et al.
Publicado: (2024)
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization
por: Sun, Jingyi, et al.
Publicado: (2026)
por: Sun, Jingyi, et al.
Publicado: (2026)
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
por: Kaffee, Lucie-Aimée, et al.
Publicado: (2023)
por: Kaffee, Lucie-Aimée, et al.
Publicado: (2023)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
por: Mujahid, Zain Muhammad, et al.
Publicado: (2025)
por: Mujahid, Zain Muhammad, et al.
Publicado: (2025)
Multi-Step Knowledge Interaction Analysis via Rank-2 Subspace Disentanglement
por: Islam, Sekh Mainul, et al.
Publicado: (2025)
por: Islam, Sekh Mainul, et al.
Publicado: (2025)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
por: Wang, Qianli, et al.
Publicado: (2026)
por: Wang, Qianli, et al.
Publicado: (2026)
Revealing Fine-Grained Values and Opinions in Large Language Models
por: Wright, Dustin, et al.
Publicado: (2024)
por: Wright, Dustin, et al.
Publicado: (2024)
SynDARin: Synthesising Datasets for Automated Reasoning in Low-Resource Languages
por: Ghazaryan, Gayane, et al.
Publicado: (2024)
por: Ghazaryan, Gayane, et al.
Publicado: (2024)
Multi-Modal Framing Analysis of News
por: Arora, Arnav, et al.
Publicado: (2025)
por: Arora, Arnav, et al.
Publicado: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
por: Zambrano, Alejandra, et al.
Publicado: (2026)
por: Zambrano, Alejandra, et al.
Publicado: (2026)
Quantifying Uncertainty in Natural Language Explanations of Large Language Models for Question Answering
por: Li, Yangyi, et al.
Publicado: (2025)
por: Li, Yangyi, et al.
Publicado: (2025)
Aggregating Soft Labels from Crowd Annotations Improves Uncertainty Estimation Under Distribution Shift
por: Wright, Dustin, et al.
Publicado: (2022)
por: Wright, Dustin, et al.
Publicado: (2022)
FLARE: Faithful Logic-Aided Reasoning and Exploration
por: Arakelyan, Erik, et al.
Publicado: (2024)
por: Arakelyan, Erik, et al.
Publicado: (2024)
Decoding Uncertainty: The Impact of Decoding Strategies for Uncertainty Estimation in Large Language Models
por: Hashimoto, Wataru, et al.
Publicado: (2025)
por: Hashimoto, Wataru, et al.
Publicado: (2025)
Estimation of Concept Explanations Should be Uncertainty Aware
por: Piratla, Vihari, et al.
Publicado: (2023)
por: Piratla, Vihari, et al.
Publicado: (2023)
BiasGym: A Simple and Generalizable Framework for Analyzing and Removing Biases through Elicitation
por: Islam, Sekh Mainul, et al.
Publicado: (2025)
por: Islam, Sekh Mainul, et al.
Publicado: (2025)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
por: Wang, Qianli, et al.
Publicado: (2025)
por: Wang, Qianli, et al.
Publicado: (2025)
Cross-Refine: Improving Natural Language Explanation Generation by Learning in Tandem
por: Wang, Qianli, et al.
Publicado: (2024)
por: Wang, Qianli, et al.
Publicado: (2024)
Investigating the Impact of Data Selection Strategies on Language Model Performance
por: Gu, Jiayao, et al.
Publicado: (2025)
por: Gu, Jiayao, et al.
Publicado: (2025)
Understanding the Interplay between LLMs' Utilisation of Parametric and Contextual Knowledge: A keynote at ECIR 2025
por: Augenstein, Isabelle
Publicado: (2026)
por: Augenstein, Isabelle
Publicado: (2026)
Cascade-Aware Training of Language Models
por: Wang, Congchao, et al.
Publicado: (2024)
por: Wang, Congchao, et al.
Publicado: (2024)
Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation
por: Ranaldi, Federico, et al.
Publicado: (2024)
por: Ranaldi, Federico, et al.
Publicado: (2024)
FLEx: Language Modeling with Few-shot Language Explanations
por: Avsian, Adar, et al.
Publicado: (2026)
por: Avsian, Adar, et al.
Publicado: (2026)
Reasoning-Grounded Natural Language Explanations for Language Models
por: Cahlik, Vojtech, et al.
Publicado: (2025)
por: Cahlik, Vojtech, et al.
Publicado: (2025)
Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models
por: Sun, Jingyi, et al.
Publicado: (2025)
por: Sun, Jingyi, et al.
Publicado: (2025)
Explaining Sources of Uncertainty in Automated Fact-Checking
por: Sun, Jingyi, et al.
Publicado: (2025)
por: Sun, Jingyi, et al.
Publicado: (2025)
Epistemic Diversity and Knowledge Collapse in Large Language Models
por: Wright, Dustin, et al.
Publicado: (2025)
por: Wright, Dustin, et al.
Publicado: (2025)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
por: Sun, Jingyi, et al.
Publicado: (2024)
por: Sun, Jingyi, et al.
Publicado: (2024)
Harnessing LLMs Explanations to Boost Surrogate Models in Tabular Data Classification
por: Shi, Ruxue, et al.
Publicado: (2025)
por: Shi, Ruxue, et al.
Publicado: (2025)
Selective Explanations
por: Paes, Lucas Monteiro, et al.
Publicado: (2024)
por: Paes, Lucas Monteiro, et al.
Publicado: (2024)
LLMs for XAI: Future Directions for Explaining Explanations
por: Zytek, Alexandra, et al.
Publicado: (2024)
por: Zytek, Alexandra, et al.
Publicado: (2024)
Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations
por: Bhan, Milan, et al.
Publicado: (2024)
por: Bhan, Milan, et al.
Publicado: (2024)
Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation
por: Zhao, Haiyan, et al.
Publicado: (2026)
por: Zhao, Haiyan, et al.
Publicado: (2026)
Uncertainty in Semantic Language Modeling with PIXELS
por: Radu, Stefania, et al.
Publicado: (2025)
por: Radu, Stefania, et al.
Publicado: (2025)
Transparent Neighborhood Approximation for Text Classifier Explanation
por: Cai, Yi, et al.
Publicado: (2024)
por: Cai, Yi, et al.
Publicado: (2024)
Query-Focused Extractive Summarization for Sentiment Explanation
por: Moubtahij, Ahmed, et al.
Publicado: (2025)
por: Moubtahij, Ahmed, et al.
Publicado: (2025)
Ejemplares similares
-
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
por: Marjanović, Sara Vera, et al.
Publicado: (2024) -
Joint Extraction and Classification of Danish Competences for Job Matching
por: Li, Qiuchi, et al.
Publicado: (2024) -
Quantifying Gender Biases Towards Politicians on Reddit
por: Marjanovic, Sara, et al.
Publicado: (2021) -
A Reality Check on Context Utilisation for Retrieval-Augmented Generation
por: Hagström, Lovisa, et al.
Publicado: (2024) -
Semantic Sensitivities and Inconsistent Predictions: Measuring the Fragility of NLI Models
por: Arakelyan, Erik, et al.
Publicado: (2024)