Resolving Lexical Bias in Model Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Rizwan, Hammad, Rosati, Domenic, Wu, Ga, Sajjad, Hassan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dependency Parsing is More Parameter-Efficient with Normalization
by: Gajo, Paolo, et al.
Published: (2025)
by: Gajo, Paolo, et al.
Published: (2025)
LLMs Underperform Graph-Based Parsers on Supervised Relation Extraction for Complex Graphs
by: Gajo, Paolo, et al.
Published: (2026)
by: Gajo, Paolo, et al.
Published: (2026)
Immunization against harmful fine-tuning attacks
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
Instance-Level Difficulty: A Missing Perspective in Machine Unlearning
by: Rizwan, Hammad, et al.
Published: (2024)
by: Rizwan, Hammad, et al.
Published: (2024)
Evaluating Defences against Unsafe Feedback in RLHF
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution
by: Haider, Muhammad Umair, et al.
Published: (2025)
by: Haider, Muhammad Umair, et al.
Published: (2025)
Semantic Consistency for Assuring Reliability of Large Language Models
by: Raj, Harsh, et al.
Published: (2023)
by: Raj, Harsh, et al.
Published: (2023)
Long-form evaluation of model editing
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
Improving Consistency in Large Language Models through Chain of Guidance
by: Raj, Harsh, et al.
Published: (2025)
by: Raj, Harsh, et al.
Published: (2025)
Representation Noising: A Defence Mechanism Against Harmful Finetuning
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
VISLA Benchmark: Evaluating Embedding Sensitivity to Semantic and Lexical Alterations
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Consistency in Language Models: Current Landscape, Challenges, and Future Directions
by: Novikova, Jekaterina, et al.
Published: (2025)
by: Novikova, Jekaterina, et al.
Published: (2025)
Sensitivity of Generative VLMs to Semantically and Lexically Altered Prompts
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
An LLM-Enhanced Adversarial Editing System for Lexical Simplification
by: Tan, Keren, et al.
Published: (2024)
by: Tan, Keren, et al.
Published: (2024)
SUGARCREPE++ Dataset: Vision-Language Model Sensitivity to Semantic and Lexical Alterations
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Data-centric Prediction Explanation via Kernelized Stein Discrepancy
by: Sarvmaili, Mahtab, et al.
Published: (2024)
by: Sarvmaili, Mahtab, et al.
Published: (2024)
Discovering Salient Neurons in Deep NLP Models
by: Durrani, Nadir, et al.
Published: (2022)
by: Durrani, Nadir, et al.
Published: (2022)
Understanding Syntactic Generalization in Structure-inducing Language Models
by: Arps, David, et al.
Published: (2025)
by: Arps, David, et al.
Published: (2025)
Digital Linguistic Bias in Spanish: Evidence from Lexical Variation in LLMs
by: Kawasaki, Yoshifumi
Published: (2026)
by: Kawasaki, Yoshifumi
Published: (2026)
Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning
by: Ma, Chuang, et al.
Published: (2026)
by: Ma, Chuang, et al.
Published: (2026)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Resolving Transcription Ambiguity in Spanish: A Hybrid Acoustic-Lexical System for Punctuation Restoration
by: Zhu, Xiliang, et al.
Published: (2024)
by: Zhu, Xiliang, et al.
Published: (2024)
Rebuilding ROME : Resolving Model Collapse during Sequential Model Editing
by: Gupta, Akshat, et al.
Published: (2024)
by: Gupta, Akshat, et al.
Published: (2024)
UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
by: Adeeba, Farah, et al.
Published: (2025)
by: Adeeba, Farah, et al.
Published: (2025)
Multilingual Nonce Dependency Treebanks: Understanding how Language Models represent and process syntactic structure
by: Arps, David, et al.
Published: (2023)
by: Arps, David, et al.
Published: (2023)
"Flex Tape Can't Fix That": Bias and Misinformation in Edited Language Models
by: Halevy, Karina, et al.
Published: (2024)
by: Halevy, Karina, et al.
Published: (2024)
Quantifying the Capabilities of LLMs across Scale and Precision
by: Badshah, Sher, et al.
Published: (2024)
by: Badshah, Sher, et al.
Published: (2024)
Interpreting the Effects of Quantization on LLMs
by: Singh, Manpreet, et al.
Published: (2025)
by: Singh, Manpreet, et al.
Published: (2025)
Resolving Editing-Unlearning Conflicts: A Knowledge Codebook Framework for Large Language Model Updating
by: Zhang, Binchi, et al.
Published: (2025)
by: Zhang, Binchi, et al.
Published: (2025)
Latent Concept-based Explanation of NLP Models
by: Yu, Xuemin, et al.
Published: (2024)
by: Yu, Xuemin, et al.
Published: (2024)
Crowdsourcing Lexical Diversity
by: Khalilia, Hadi, et al.
Published: (2024)
by: Khalilia, Hadi, et al.
Published: (2024)
Dispersion Measures as Predictors of Lexical Decision Time, Word Familiarity, and Lexical Complexity
by: Nohejl, Adam, et al.
Published: (2025)
by: Nohejl, Adam, et al.
Published: (2025)
Tackling Social Bias against the Poor: A Dataset and Taxonomy on Aporophobia
by: Curto, Georgina, et al.
Published: (2025)
by: Curto, Georgina, et al.
Published: (2025)
TALE: A Tool-Augmented Framework for Reference-Free Evaluation of Large Language Models
by: Badshah, Sher, et al.
Published: (2025)
by: Badshah, Sher, et al.
Published: (2025)
Quantifying Label-Induced Bias in Large Language Model Self- and Cross-Evaluations
by: Saraf, Muskan, et al.
Published: (2025)
by: Saraf, Muskan, et al.
Published: (2025)
Large Language Model Bias Mitigation from the Perspective of Knowledge Editing
by: Chen, Ruizhe, et al.
Published: (2024)
by: Chen, Ruizhe, et al.
Published: (2024)
Lexically Grounded Subword Segmentation
by: Libovický, Jindřich, et al.
Published: (2024)
by: Libovický, Jindřich, et al.
Published: (2024)
Automatic Lexical Simplification for Turkish
by: Uluslu, Ahmet Yavuz
Published: (2022)
by: Uluslu, Ahmet Yavuz
Published: (2022)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution
by: Kamp, Jonathan, et al.
Published: (2025)
by: Kamp, Jonathan, et al.
Published: (2025)
Similar Items
-
Dependency Parsing is More Parameter-Efficient with Normalization
by: Gajo, Paolo, et al.
Published: (2025) -
LLMs Underperform Graph-Based Parsers on Supervised Relation Extraction for Complex Graphs
by: Gajo, Paolo, et al.
Published: (2026) -
Immunization against harmful fine-tuning attacks
by: Rosati, Domenic, et al.
Published: (2024) -
Instance-Level Difficulty: A Missing Perspective in Machine Unlearning
by: Rizwan, Hammad, et al.
Published: (2024) -
Evaluating Defences against Unsafe Feedback in RLHF
by: Rosati, Domenic, et al.
Published: (2024)