Enhancing Fact Retrieval in PLMs through Truthfulness
Fuente:
arXiv
Guardado en:
| Autores principales: | Youssef, Paul, Schlötterer, Jörg, Seifert, Christin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Queen of England is not England's Queen: On the Lack of Factual Coherency in PLMs
por: Youssef, Paul, et al.
Publicado: (2024)
por: Youssef, Paul, et al.
Publicado: (2024)
Has this Fact been Edited? Detecting Knowledge Edits in Language Models
por: Youssef, Paul, et al.
Publicado: (2024)
por: Youssef, Paul, et al.
Publicado: (2024)
Persuasion Tokens for Editing Factual Knowledge in LLMs
por: Youssef, Paul, et al.
Publicado: (2026)
por: Youssef, Paul, et al.
Publicado: (2026)
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
por: Youssef, Paul, et al.
Publicado: (2024)
por: Youssef, Paul, et al.
Publicado: (2024)
Tracing and Reversing Edits in LLMs
por: Youssef, Paul, et al.
Publicado: (2025)
por: Youssef, Paul, et al.
Publicado: (2025)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
por: Nguyen, Van Bach, et al.
Publicado: (2024)
por: Nguyen, Van Bach, et al.
Publicado: (2024)
Position: Editing Large Language Models Poses Serious Safety Risks
por: Youssef, Paul, et al.
Publicado: (2025)
por: Youssef, Paul, et al.
Publicado: (2025)
Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation
por: Cheng, Yinjie, et al.
Publicado: (2025)
por: Cheng, Yinjie, et al.
Publicado: (2025)
Guiding LLMs to Generate High-Fidelity and High-Quality Counterfactual Explanations for Text Classification
por: Nguyen, Van Bach, et al.
Publicado: (2025)
por: Nguyen, Van Bach, et al.
Publicado: (2025)
A Second Look on BASS -- Boosting Abstractive Summarization with Unified Semantic Graphs -- A Replication Study
por: Koraş, Osman Alperen, et al.
Publicado: (2024)
por: Koraş, Osman Alperen, et al.
Publicado: (2024)
CEval: A Benchmark for Evaluating Counterfactual Text Generation
por: Nguyen, Van Bach, et al.
Publicado: (2024)
por: Nguyen, Van Bach, et al.
Publicado: (2024)
From Black Boxes to Conversations: Incorporating XAI in a Conversational Agent
por: Nguyen, Van Bach, et al.
Publicado: (2022)
por: Nguyen, Van Bach, et al.
Publicado: (2022)
Behavioral Analysis of Information Salience in Large Language Models
por: Trienes, Jan, et al.
Publicado: (2025)
por: Trienes, Jan, et al.
Publicado: (2025)
Marcel: A Lightweight and Open-Source Conversational Agent for University Student Support
por: Trienes, Jan, et al.
Publicado: (2025)
por: Trienes, Jan, et al.
Publicado: (2025)
An XAI-based Analysis of Shortcut Learning in Neural Networks
por: Le, Phuong Quynh, et al.
Publicado: (2025)
por: Le, Phuong Quynh, et al.
Publicado: (2025)
Is Last Layer Re-Training Truly Sufficient for Robustness to Spurious Correlations?
por: Le, Phuong Quynh, et al.
Publicado: (2023)
por: Le, Phuong Quynh, et al.
Publicado: (2023)
InfoLossQA: Characterizing and Recovering Information Loss in Text Simplification
por: Trienes, Jan, et al.
Publicado: (2024)
por: Trienes, Jan, et al.
Publicado: (2024)
QASE Enhanced PLMs: Improved Control in Text Generation for MRC
por: Ai, Lin, et al.
Publicado: (2024)
por: Ai, Lin, et al.
Publicado: (2024)
Towards Interpretable Deep Neural Networks for Tabular Data
por: Elhadri, Khawla, et al.
Publicado: (2025)
por: Elhadri, Khawla, et al.
Publicado: (2025)
XNNTab -- Interpretable Neural Networks for Tabular Data using Sparse Autoencoders
por: Elhadri, Khawla, et al.
Publicado: (2025)
por: Elhadri, Khawla, et al.
Publicado: (2025)
Investigating the Impact of Randomness on Reproducibility in Computer Vision: A Study on Applications in Civil Engineering and Medicine
por: Eryılmaz, Bahadır, et al.
Publicado: (2024)
por: Eryılmaz, Bahadır, et al.
Publicado: (2024)
Truth or Twist? Optimal Model Selection for Reliable Label Flipping Evaluation in LLM-based Counterfactuals
por: Wang, Qianli, et al.
Publicado: (2025)
por: Wang, Qianli, et al.
Publicado: (2025)
MLPs Compass: What is learned when MLPs are combined with PLMs?
por: Zhou, Li, et al.
Publicado: (2024)
por: Zhou, Li, et al.
Publicado: (2024)
Probing Critical Learning Dynamics of PLMs for Hate Speech Detection
por: Masud, Sarah, et al.
Publicado: (2024)
por: Masud, Sarah, et al.
Publicado: (2024)
Out of Spuriousity: Improving Robustness to Spurious Correlations without Group Annotations
por: Le, Phuong Quynh, et al.
Publicado: (2024)
por: Le, Phuong Quynh, et al.
Publicado: (2024)
Invariant Learning with Annotation-free Environments
por: Le, Phuong Quynh, et al.
Publicado: (2025)
por: Le, Phuong Quynh, et al.
Publicado: (2025)
Bridging the Gap: Transfer Learning from English PLMs to Malaysian English
por: Chanthran, Mohan Raj, et al.
Publicado: (2024)
por: Chanthran, Mohan Raj, et al.
Publicado: (2024)
DAdEE: Unsupervised Domain Adaptation in Early Exit PLMs
por: Bajpai, Divya Jyoti, et al.
Publicado: (2024)
por: Bajpai, Divya Jyoti, et al.
Publicado: (2024)
The Missing Parts: Augmenting Fact Verification with Half-Truth Detection
por: Tang, Yixuan, et al.
Publicado: (2025)
por: Tang, Yixuan, et al.
Publicado: (2025)
Explanation format does not matter; but explanations do -- An Eggsbert study on explaining Bayesian Optimisation tasks
por: Chakraborty, Tanmay, et al.
Publicado: (2025)
por: Chakraborty, Tanmay, et al.
Publicado: (2025)
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
The Impact of Annotator Personas on LLM Behavior Across the Perspectivism Spectrum
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
Patch-based Intuitive Multimodal Prototypes Network (PIMPNet) for Alzheimer's Disease classification
por: De Santi, Lisa Anita, et al.
Publicado: (2024)
por: De Santi, Lisa Anita, et al.
Publicado: (2024)
One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them
por: Holmov, Ali, et al.
Publicado: (2026)
por: Holmov, Ali, et al.
Publicado: (2026)
RulePrompt: Weakly Supervised Text Classification with Prompting PLMs and Self-Iterative Logical Rules
por: Li, Miaomiao, et al.
Publicado: (2024)
por: Li, Miaomiao, et al.
Publicado: (2024)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024)
por: Chatrath, Veronica, et al.
Publicado: (2024)
Efficient Unsupervised Shortcut Learning Detection and Mitigation in Transformers
por: Kuhn, Lukas, et al.
Publicado: (2025)
por: Kuhn, Lukas, et al.
Publicado: (2025)
Scaling Truth: The Confidence Paradox in AI Fact-Checking
por: Qazi, Ihsan A., et al.
Publicado: (2025)
por: Qazi, Ihsan A., et al.
Publicado: (2025)
Fact or Fiction? Improving Fact Verification with Knowledge Graphs through Simplified Subgraph Retrievals
por: Opsahl, Tobias A.
Publicado: (2024)
por: Opsahl, Tobias A.
Publicado: (2024)
Revisit Few-shot Intent Classification with PLMs: Direct Fine-tuning vs. Continual Pre-training
por: Zhang, Haode, et al.
Publicado: (2023)
por: Zhang, Haode, et al.
Publicado: (2023)
Ejemplares similares
-
The Queen of England is not England's Queen: On the Lack of Factual Coherency in PLMs
por: Youssef, Paul, et al.
Publicado: (2024) -
Has this Fact been Edited? Detecting Knowledge Edits in Language Models
por: Youssef, Paul, et al.
Publicado: (2024) -
Persuasion Tokens for Editing Factual Knowledge in LLMs
por: Youssef, Paul, et al.
Publicado: (2026) -
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
por: Youssef, Paul, et al.
Publicado: (2024) -
Tracing and Reversing Edits in LLMs
por: Youssef, Paul, et al.
Publicado: (2025)