TAXI: Evaluating Categorical Knowledge Editing for Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Powell, Derek, Gerych, Walter, Hartvigsen, Thomas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
Efficient Knowledge Editing via Minimal Precomputation
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
Model Editing with Graph-Based External Memory
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
Lifelong Knowledge Editing requires Better Regularization
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
Norm Growth and Stability Challenges in Localized Sequential Knowledge Editing
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
Evaluating Temporal Consistency in Multi-Turn Language Models
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2026)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2026)
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
von: Thede, Lukas, et al.
Veröffentlicht: (2025)
von: Thede, Lukas, et al.
Veröffentlicht: (2025)
Continually Self-Improving Language Models for Bariatric Surgery Question--Answering
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
Can Large Language Models Predict Associations Among Human Attitudes?
von: Ma, Ana, et al.
Veröffentlicht: (2025)
von: Ma, Ana, et al.
Veröffentlicht: (2025)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
von: Jain, Devansh, et al.
Veröffentlicht: (2024)
von: Jain, Devansh, et al.
Veröffentlicht: (2024)
Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?
von: Seah, Natalie, et al.
Veröffentlicht: (2026)
von: Seah, Natalie, et al.
Veröffentlicht: (2026)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
Decoding Knowledge in Large Language Models: A Framework for Categorization and Comprehension
von: Fang, Yanbo, et al.
Veröffentlicht: (2025)
von: Fang, Yanbo, et al.
Veröffentlicht: (2025)
Stable Knowledge Editing in Large Language Models
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Inspecting and Editing Knowledge Representations in Language Models
von: Hernandez, Evan, et al.
Veröffentlicht: (2023)
von: Hernandez, Evan, et al.
Veröffentlicht: (2023)
BendVLM: Test-Time Debiasing of Vision-Language Embeddings
von: Gerych, Walter, et al.
Veröffentlicht: (2024)
von: Gerych, Walter, et al.
Veröffentlicht: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
Knowledge Graph Enhanced Large Language Model Editing
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
Understanding Language Model Circuits through Knowledge Editing
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
von: Ge, Huaizhi, et al.
Veröffentlicht: (2024)
Neighboring Perturbations of Knowledge Editing on Large Language Models
von: Ma, Jun-Yu, et al.
Veröffentlicht: (2024)
von: Ma, Jun-Yu, et al.
Veröffentlicht: (2024)
Benchmarking and Rethinking Knowledge Editing for Large Language Models
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
Disentangling Knowledge Representations for Large Language Model Editing
von: Zhang, Mengqi, et al.
Veröffentlicht: (2025)
von: Zhang, Mengqi, et al.
Veröffentlicht: (2025)
When Style Breaks Safety: Defending LLMs Against Superficial Style Alignment
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
Editing the Mind of Giants: An In-Depth Exploration of Pitfalls of Knowledge Editing in Large Language Models
von: Hsueh, Cheng-Hsun, et al.
Veröffentlicht: (2024)
von: Hsueh, Cheng-Hsun, et al.
Veröffentlicht: (2024)
Mapping from Meaning: Addressing the Miscalibration of Prompt-Sensitive Language Models
von: Cox, Kyle, et al.
Veröffentlicht: (2025)
von: Cox, Kyle, et al.
Veröffentlicht: (2025)
Knowledge in Superposition: Unveiling the Failures of Lifelong Knowledge Editing for Large Language Models
von: Hu, Chenhui, et al.
Veröffentlicht: (2024)
von: Hu, Chenhui, et al.
Veröffentlicht: (2024)
Editing Across Languages: A Survey of Multilingual Knowledge Editing
von: Durrani, Nadir, et al.
Veröffentlicht: (2025)
von: Durrani, Nadir, et al.
Veröffentlicht: (2025)
Knowledge Editing for Large Language Model with Knowledge Neuronal Ensemble
von: Li, Yongchang, et al.
Veröffentlicht: (2024)
von: Li, Yongchang, et al.
Veröffentlicht: (2024)
Context-Robust Knowledge Editing for Language Models
von: Park, Haewon, et al.
Veröffentlicht: (2025)
von: Park, Haewon, et al.
Veröffentlicht: (2025)
GeoEdit: Geometric Knowledge Editing for Large Language Models
von: Feng, Yujie, et al.
Veröffentlicht: (2025)
von: Feng, Yujie, et al.
Veröffentlicht: (2025)
Mechanistic Circuit-Based Knowledge Editing in Large Language Models
von: Zhao, Tianyi, et al.
Veröffentlicht: (2026)
von: Zhao, Tianyi, et al.
Veröffentlicht: (2026)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
von: Christ, Bryan R, et al.
Veröffentlicht: (2024)
von: Christ, Bryan R, et al.
Veröffentlicht: (2024)
Aligning Language Models with Real-time Knowledge Editing
von: Tang, Chenming, et al.
Veröffentlicht: (2025)
von: Tang, Chenming, et al.
Veröffentlicht: (2025)
MLaKE: Multilingual Knowledge Editing Benchmark for Large Language Models
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Identifying Knowledge Editing Types in Large Language Models
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
Cross-Lingual Knowledge Editing in Large Language Models
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
Knowledge Editing for Large Language Models: A Survey
von: Wang, Song, et al.
Veröffentlicht: (2023)
von: Wang, Song, et al.
Veröffentlicht: (2023)
Medical Large Language Model Benchmarks Should Prioritize Construct Validity
von: Alaa, Ahmed, et al.
Veröffentlicht: (2025)
von: Alaa, Ahmed, et al.
Veröffentlicht: (2025)
Language Models Still Struggle to Zero-shot Reason about Time Series
von: Merrill, Mike A., et al.
Veröffentlicht: (2024)
von: Merrill, Mike A., et al.
Veröffentlicht: (2024)
DIEKAE: Difference Injection for Efficient Knowledge Augmentation and Editing of Large Language Models
von: Galatolo, Alessio, et al.
Veröffentlicht: (2024)
von: Galatolo, Alessio, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024) -
Efficient Knowledge Editing via Minimal Precomputation
von: Gupta, Akshat, et al.
Veröffentlicht: (2025) -
Model Editing with Graph-Based External Memory
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025) -
Lifelong Knowledge Editing requires Better Regularization
von: Gupta, Akshat, et al.
Veröffentlicht: (2025) -
Norm Growth and Stability Challenges in Localized Sequential Knowledge Editing
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)