Has this Fact been Edited? Detecting Knowledge Edits in Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Youssef, Paul, Zhao, Zhixue, Seifert, Christin, Schlötterer, Jörg |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation
von: Cheng, Yinjie, et al.
Veröffentlicht: (2025)
von: Cheng, Yinjie, et al.
Veröffentlicht: (2025)
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
Tracing and Reversing Edits in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2025)
von: Youssef, Paul, et al.
Veröffentlicht: (2025)
Position: Editing Large Language Models Poses Serious Safety Risks
von: Youssef, Paul, et al.
Veröffentlicht: (2025)
von: Youssef, Paul, et al.
Veröffentlicht: (2025)
Persuasion Tokens for Editing Factual Knowledge in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
Enhancing Fact Retrieval in PLMs through Truthfulness
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2024)
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2024)
CEval: A Benchmark for Evaluating Counterfactual Text Generation
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2024)
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2024)
From Black Boxes to Conversations: Incorporating XAI in a Conversational Agent
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2022)
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2022)
The Queen of England is not England's Queen: On the Lack of Factual Coherency in PLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
von: Youssef, Paul, et al.
Veröffentlicht: (2024)
AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
von: Vasilakes, Jake, et al.
Veröffentlicht: (2025)
Behavioral Analysis of Information Salience in Large Language Models
von: Trienes, Jan, et al.
Veröffentlicht: (2025)
von: Trienes, Jan, et al.
Veröffentlicht: (2025)
DeepEdit: Knowledge Editing as Decoding with Constraints
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
von: Wang, Yiwei, et al.
Veröffentlicht: (2024)
FLEKE: Federated Locate-then-Edit Knowledge Editing
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
StruEdit: Structured Outputs Enable the Fast and Accurate Knowledge Editing for Large Language Models
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
von: Bi, Baolong, et al.
Veröffentlicht: (2024)
Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
von: Yang, Linyao, et al.
Veröffentlicht: (2023)
von: Yang, Linyao, et al.
Veröffentlicht: (2023)
Guiding LLMs to Generate High-Fidelity and High-Quality Counterfactual Explanations for Text Classification
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2025)
von: Nguyen, Van Bach, et al.
Veröffentlicht: (2025)
Investigating the Impact of Randomness on Reproducibility in Computer Vision: A Study on Applications in Civil Engineering and Medicine
von: Eryılmaz, Bahadır, et al.
Veröffentlicht: (2024)
von: Eryılmaz, Bahadır, et al.
Veröffentlicht: (2024)
Out of Spuriousity: Improving Robustness to Spurious Correlations without Group Annotations
von: Le, Phuong Quynh, et al.
Veröffentlicht: (2024)
von: Le, Phuong Quynh, et al.
Veröffentlicht: (2024)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
A Second Look on BASS -- Boosting Abstractive Summarization with Unified Semantic Graphs -- A Replication Study
von: Koraş, Osman Alperen, et al.
Veröffentlicht: (2024)
von: Koraş, Osman Alperen, et al.
Veröffentlicht: (2024)
CKnowEdit: A New Chinese Knowledge Editing Dataset for Linguistics, Facts, and Logic Error Correction in LLMs
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
von: Fang, Jizhan, et al.
Veröffentlicht: (2024)
Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
InstructEdit: Instruction-based Knowledge Editing for Large Language Models
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
von: Zhang, Ningyu, et al.
Veröffentlicht: (2024)
ChartEditBench: Evaluating Grounded Multi-Turn Chart Editing in Multimodal Language Models
von: Kapadnis, Manav Nitin, et al.
Veröffentlicht: (2026)
von: Kapadnis, Manav Nitin, et al.
Veröffentlicht: (2026)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Context-Robust Knowledge Editing for Language Models
von: Park, Haewon, et al.
Veröffentlicht: (2025)
von: Park, Haewon, et al.
Veröffentlicht: (2025)
Resolving UnderEdit & OverEdit with Iterative & Neighbor-Assisted Model Editing
von: Baghel, Bhiman Kumar, et al.
Veröffentlicht: (2025)
von: Baghel, Bhiman Kumar, et al.
Veröffentlicht: (2025)
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
EasyEdit: An Easy-to-use Knowledge Editing Framework for Large Language Models
von: Wang, Peng, et al.
Veröffentlicht: (2023)
von: Wang, Peng, et al.
Veröffentlicht: (2023)
NeuralDB: Scaling Knowledge Editing in LLMs to 100,000 Facts with Neural KV Database
von: Fei, Weizhi, et al.
Veröffentlicht: (2025)
von: Fei, Weizhi, et al.
Veröffentlicht: (2025)
ChainEdit: Propagating Ripple Effects in LLM Knowledge Editing through Logical Rule-Guided Chains
von: Dong, Zilu, et al.
Veröffentlicht: (2025)
von: Dong, Zilu, et al.
Veröffentlicht: (2025)
UltraEdit: Training-, Subject-, and Memory-Free Lifelong Editing in Language Models
von: Gu, Xiaojie, et al.
Veröffentlicht: (2025)
von: Gu, Xiaojie, et al.
Veröffentlicht: (2025)
SUCEA: Reasoning-Intensive Retrieval for Adversarial Fact-checking through Claim Decomposition and Editing
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
Identifying Knowledge Editing Types in Large Language Models
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
Cross-Lingual Knowledge Editing in Large Language Models
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
Knowledge Editing for Large Language Models: A Survey
von: Wang, Song, et al.
Veröffentlicht: (2023)
von: Wang, Song, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation
von: Cheng, Yinjie, et al.
Veröffentlicht: (2025) -
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
von: Youssef, Paul, et al.
Veröffentlicht: (2024) -
Tracing and Reversing Edits in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2025) -
Position: Editing Large Language Models Poses Serious Safety Risks
von: Youssef, Paul, et al.
Veröffentlicht: (2025) -
Persuasion Tokens for Editing Factual Knowledge in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2026)