Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Zhixue, Aletras, Nikolaos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Incorporating Attribution Importance for Improving Faithfulness Metrics
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
Evaluating Monolingual and Multilingual Large Language Models for Greek Question Answering: The DemosQA Benchmark
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026)
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
von: Williams, Miles, et al.
Veröffentlicht: (2023)
von: Williams, Miles, et al.
Veröffentlicht: (2023)
An Empirical Study on Cross-lingual Vocabulary Adaptation for Efficient Language Model Inference
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
Adapting Chat Language Models Using Only Target Unlabeled Language Data
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text?
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2025)
FaithLM: Towards Faithful Explanations for Large Language Models
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Mitigating Content Effects on Reasoning in Language Models through Fine-Grained Activation Steering
von: Valentino, Marco, et al.
Veröffentlicht: (2025)
von: Valentino, Marco, et al.
Veröffentlicht: (2025)
Fine-tuning Large Language Models for Multigenerator, Multidomain, and Multilingual Machine-Generated Text Detection
von: Xiong, Feng, et al.
Veröffentlicht: (2024)
von: Xiong, Feng, et al.
Veröffentlicht: (2024)
The Scandinavian Embedding Benchmarks: Comprehensive Assessment of Multilingual and Monolingual Text Embedding
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
von: Enevoldsen, Kenneth, et al.
Veröffentlicht: (2024)
Monolingual or Multilingual Instruction Tuning: Which Makes a Better Alpaca
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models
von: Tan, Xingwei, et al.
Veröffentlicht: (2026)
von: Tan, Xingwei, et al.
Veröffentlicht: (2026)
The Probabilities Also Matter: A More Faithful Metric for Faithfulness of Free-Text Explanations in Large Language Models
von: Siegel, Noah Y., et al.
Veröffentlicht: (2024)
von: Siegel, Noah Y., et al.
Veröffentlicht: (2024)
Multilingual Information Retrieval with a Monolingual Knowledge Base
von: Zhuang, Yingying, et al.
Veröffentlicht: (2025)
von: Zhuang, Yingying, et al.
Veröffentlicht: (2025)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
Where does output diversity collapse in post-training?
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models
von: Sovrano, Francesco, et al.
Veröffentlicht: (2026)
von: Sovrano, Francesco, et al.
Veröffentlicht: (2026)
Fine-tuning Language Models for Recipe Generation: A Comparative Analysis and Benchmark Study
von: Vij, Anneketh, et al.
Veröffentlicht: (2025)
von: Vij, Anneketh, et al.
Veröffentlicht: (2025)
Progtuning: Progressive Fine-tuning Framework for Transformer-based Language Models
von: Ji, Xiaoshuang, et al.
Veröffentlicht: (2025)
von: Ji, Xiaoshuang, et al.
Veröffentlicht: (2025)
ExU: AI Models for Examining Multilingual Disinformation Narratives and Understanding their Spread
von: Vasilakes, Jake, et al.
Veröffentlicht: (2024)
von: Vasilakes, Jake, et al.
Veröffentlicht: (2024)
Introducing cosmosGPT: Monolingual Training for Turkish Language Models
von: Kesgin, H. Toprak, et al.
Veröffentlicht: (2024)
von: Kesgin, H. Toprak, et al.
Veröffentlicht: (2024)
Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
von: Matton, Katie, et al.
Veröffentlicht: (2025)
von: Matton, Katie, et al.
Veröffentlicht: (2025)
SPARSEFIT: Few-shot Prompting with Sparse Fine-tuning for Jointly Generating Predictions and Natural Language Explanations
von: Solano, Jesus, et al.
Veröffentlicht: (2023)
von: Solano, Jesus, et al.
Veröffentlicht: (2023)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
Bilingual Adaptation of Monolingual Foundation Models
von: Gosal, Gurpreet, et al.
Veröffentlicht: (2024)
von: Gosal, Gurpreet, et al.
Veröffentlicht: (2024)
Fine-tuning ChatGPT for Automatic Scoring of Written Scientific Explanations in Chinese
von: Yang, Jie, et al.
Veröffentlicht: (2025)
von: Yang, Jie, et al.
Veröffentlicht: (2025)
UrduLM: A Resource-Efficient Monolingual Urdu Language Model
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
von: Ali, Syed Muhammad, et al.
Veröffentlicht: (2026)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
von: Fragkathoulas, Christos, et al.
Veröffentlicht: (2024)
von: Fragkathoulas, Christos, et al.
Veröffentlicht: (2024)
Has this Fact been Edited? Detecting Knowledge Edits in Language Models
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
Semi-supervised Fine-tuning for Large Language Models
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
von: Manna, Supriya, et al.
Veröffentlicht: (2024)
von: Manna, Supriya, et al.
Veröffentlicht: (2024)
Faithful and Plausible Natural Language Explanations for Image Classification: A Pipeline Approach
von: Wojciechowski, Adam, et al.
Veröffentlicht: (2024)
von: Wojciechowski, Adam, et al.
Veröffentlicht: (2024)
CAuSE: Decoding Multimodal Classifiers using Faithful Natural Language Explanation
von: Bandyopadhyay, Dibyanayan, et al.
Veröffentlicht: (2025)
von: Bandyopadhyay, Dibyanayan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Incorporating Attribution Importance for Improving Faithfulness Metrics
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023) -
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023) -
Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025) -
Evaluating Monolingual and Multilingual Large Language Models for Greek Question Answering: The DemosQA Benchmark
von: Mastrokostas, Charalampos, et al.
Veröffentlicht: (2026) -
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)