Gespeichert in:
| Hauptverfasser: | Marjanović, Sara Vera, Augenstein, Isabelle, Lioma, Christina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.13006 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2024)
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2024)
A Reality Check on Context Utilisation for Retrieval-Augmented Generation
von: Hagström, Lovisa, et al.
Veröffentlicht: (2024)
von: Hagström, Lovisa, et al.
Veröffentlicht: (2024)
Quantifying Gender Biases Towards Politicians on Reddit
von: Marjanovic, Sara, et al.
Veröffentlicht: (2021)
von: Marjanovic, Sara, et al.
Veröffentlicht: (2021)
Joint Extraction and Classification of Danish Competences for Job Matching
von: Li, Qiuchi, et al.
Veröffentlicht: (2024)
von: Li, Qiuchi, et al.
Veröffentlicht: (2024)
Semantic Sensitivities and Inconsistent Predictions: Measuring the Fragility of NLI Models
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization
von: Sun, Jingyi, et al.
Veröffentlicht: (2026)
von: Sun, Jingyi, et al.
Veröffentlicht: (2026)
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
von: Kaffee, Lucie-Aimée, et al.
Veröffentlicht: (2023)
von: Kaffee, Lucie-Aimée, et al.
Veröffentlicht: (2023)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
Multi-Step Knowledge Interaction Analysis via Rank-2 Subspace Disentanglement
von: Islam, Sekh Mainul, et al.
Veröffentlicht: (2025)
von: Islam, Sekh Mainul, et al.
Veröffentlicht: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
von: Zambrano, Alejandra, et al.
Veröffentlicht: (2026)
SynDARin: Synthesising Datasets for Automated Reasoning in Low-Resource Languages
von: Ghazaryan, Gayane, et al.
Veröffentlicht: (2024)
von: Ghazaryan, Gayane, et al.
Veröffentlicht: (2024)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
Revealing Fine-Grained Values and Opinions in Large Language Models
von: Wright, Dustin, et al.
Veröffentlicht: (2024)
von: Wright, Dustin, et al.
Veröffentlicht: (2024)
Multi-Modal Framing Analysis of News
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
FLARE: Faithful Logic-Aided Reasoning and Exploration
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)
Aggregating Soft Labels from Crowd Annotations Improves Uncertainty Estimation Under Distribution Shift
von: Wright, Dustin, et al.
Veröffentlicht: (2022)
von: Wright, Dustin, et al.
Veröffentlicht: (2022)
Quantifying Uncertainty in Natural Language Explanations of Large Language Models for Question Answering
von: Li, Yangyi, et al.
Veröffentlicht: (2025)
von: Li, Yangyi, et al.
Veröffentlicht: (2025)
BiasGym: A Simple and Generalizable Framework for Analyzing and Removing Biases through Elicitation
von: Islam, Sekh Mainul, et al.
Veröffentlicht: (2025)
von: Islam, Sekh Mainul, et al.
Veröffentlicht: (2025)
Estimation of Concept Explanations Should be Uncertainty Aware
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
Decoding Uncertainty: The Impact of Decoding Strategies for Uncertainty Estimation in Large Language Models
von: Hashimoto, Wataru, et al.
Veröffentlicht: (2025)
von: Hashimoto, Wataru, et al.
Veröffentlicht: (2025)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
Understanding the Interplay between LLMs' Utilisation of Parametric and Contextual Knowledge: A keynote at ECIR 2025
von: Augenstein, Isabelle
Veröffentlicht: (2026)
von: Augenstein, Isabelle
Veröffentlicht: (2026)
Cross-Refine: Improving Natural Language Explanation Generation by Learning in Tandem
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
Epistemic Diversity and Knowledge Collapse in Large Language Models
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
von: Sun, Jingyi, et al.
Veröffentlicht: (2024)
Cascade-Aware Training of Language Models
von: Wang, Congchao, et al.
Veröffentlicht: (2024)
von: Wang, Congchao, et al.
Veröffentlicht: (2024)
Explaining Sources of Uncertainty in Automated Fact-Checking
von: Sun, Jingyi, et al.
Veröffentlicht: (2025)
von: Sun, Jingyi, et al.
Veröffentlicht: (2025)
Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models
von: Sun, Jingyi, et al.
Veröffentlicht: (2025)
von: Sun, Jingyi, et al.
Veröffentlicht: (2025)
Investigating the Impact of Data Selection Strategies on Language Model Performance
von: Gu, Jiayao, et al.
Veröffentlicht: (2025)
von: Gu, Jiayao, et al.
Veröffentlicht: (2025)
LLMs for XAI: Future Directions for Explaining Explanations
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation
von: Ranaldi, Federico, et al.
Veröffentlicht: (2024)
von: Ranaldi, Federico, et al.
Veröffentlicht: (2024)
FLEx: Language Modeling with Few-shot Language Explanations
von: Avsian, Adar, et al.
Veröffentlicht: (2026)
von: Avsian, Adar, et al.
Veröffentlicht: (2026)
Reasoning-Grounded Natural Language Explanations for Language Models
von: Cahlik, Vojtech, et al.
Veröffentlicht: (2025)
von: Cahlik, Vojtech, et al.
Veröffentlicht: (2025)
Selective Explanations
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
Harnessing LLMs Explanations to Boost Surrogate Models in Tabular Data Classification
von: Shi, Ruxue, et al.
Veröffentlicht: (2025)
von: Shi, Ruxue, et al.
Veröffentlicht: (2025)
Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations
von: Bhan, Milan, et al.
Veröffentlicht: (2024)
von: Bhan, Milan, et al.
Veröffentlicht: (2024)
Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation
von: Zhao, Haiyan, et al.
Veröffentlicht: (2026)
von: Zhao, Haiyan, et al.
Veröffentlicht: (2026)
Uncertainty in Semantic Language Modeling with PIXELS
von: Radu, Stefania, et al.
Veröffentlicht: (2025)
von: Radu, Stefania, et al.
Veröffentlicht: (2025)
Graph-Guided Textual Explanation Generation Framework
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2024)
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2024)
Factuality Challenges in the Era of Large Language Models
von: Augenstein, Isabelle, et al.
Veröffentlicht: (2023)
von: Augenstein, Isabelle, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
DYNAMICQA: Tracing Internal Knowledge Conflicts in Language Models
von: Marjanović, Sara Vera, et al.
Veröffentlicht: (2024) -
A Reality Check on Context Utilisation for Retrieval-Augmented Generation
von: Hagström, Lovisa, et al.
Veröffentlicht: (2024) -
Quantifying Gender Biases Towards Politicians on Reddit
von: Marjanovic, Sara, et al.
Veröffentlicht: (2021) -
Joint Extraction and Classification of Danish Competences for Job Matching
von: Li, Qiuchi, et al.
Veröffentlicht: (2024) -
Semantic Sensitivities and Inconsistent Predictions: Measuring the Fragility of NLI Models
von: Arakelyan, Erik, et al.
Veröffentlicht: (2024)