Making Language Models Robust Against Negation
Fuente:
arXiv
Saved in:
| Main Authors: | Rezaei, MohammadHossein, Blanco, Eduardo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Paraphrasing in Affirmative Terms Improves Negation Understanding
by: Rezaei, MohammadHossein, et al.
Published: (2024)
by: Rezaei, MohammadHossein, et al.
Published: (2024)
Commonsense Knowledge with Negation: A Resource to Enhance Negation Understanding
by: Wang, Zijie, et al.
Published: (2026)
by: Wang, Zijie, et al.
Published: (2026)
AI in Mental Health: Emotional and Sentiment Analysis of Large Language Models' Responses to Depression, Anxiety, and Stress Queries
by: VarastehNezhad, Arya, et al.
Published: (2025)
by: VarastehNezhad, Arya, et al.
Published: (2025)
PersoDPO: Scalable Preference Optimization for Instruction-Adherent, Persona-Grounded Dialogue via Multi-LLM Evaluation
by: Afzoon, Saleh, et al.
Published: (2026)
by: Afzoon, Saleh, et al.
Published: (2026)
EgoNormia: Benchmarking Physical Social Norm Understanding
by: Rezaei, MohammadHossein, et al.
Published: (2025)
by: Rezaei, MohammadHossein, et al.
Published: (2025)
Online Rubrics Elicitation from Pairwise Comparisons
by: Rezaei, MohammadHossein, et al.
Published: (2025)
by: Rezaei, MohammadHossein, et al.
Published: (2025)
Benchmarking Gaslighting Negation Attacks Against Multimodal Large Language Models
by: Zhu, Bin, et al.
Published: (2025)
by: Zhu, Bin, et al.
Published: (2025)
Robustness of Large Language Models Against Adversarial Attacks
by: Tao, Yiyi, et al.
Published: (2024)
by: Tao, Yiyi, et al.
Published: (2024)
Algorithmic Thinking Theory
by: Bateni, MohammadHossein, et al.
Published: (2025)
by: Bateni, MohammadHossein, et al.
Published: (2025)
Evaluating Robustness of Large Language Models Against Multilingual Typographical Errors
by: Zhao, Raoyuan, et al.
Published: (2025)
by: Zhao, Raoyuan, et al.
Published: (2025)
Task-Agnostic Language Model Watermarking via High Entropy Passthrough Layers
by: Masrani, Vaden, et al.
Published: (2024)
by: Masrani, Vaden, et al.
Published: (2024)
Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks
by: Mu, Lin, et al.
Published: (2025)
by: Mu, Lin, et al.
Published: (2025)
RAG Makes Guardrails Unsafe? Investigating Robustness of Guardrails under RAG-style Contexts
by: She, Yining, et al.
Published: (2025)
by: She, Yining, et al.
Published: (2025)
Can Large Language Models Make Everyone Happy?
by: Naseem, Usman, et al.
Published: (2026)
by: Naseem, Usman, et al.
Published: (2026)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
by: Paul, Indraneil, et al.
Published: (2024)
by: Paul, Indraneil, et al.
Published: (2024)
Counterfactual Cultural Cues Reduce Medical QA Accuracy in LLMs: Identifier vs Context Effects
by: Rezaei, Amirhossein Haji Mohammad, et al.
Published: (2026)
by: Rezaei, Amirhossein Haji Mohammad, et al.
Published: (2026)
How Language Models Process Negation
by: Zhou, Zhejian, et al.
Published: (2026)
by: Zhou, Zhejian, et al.
Published: (2026)
Robust Prompt Optimization for Large Language Models Against Distribution Shifts
by: Li, Moxin, et al.
Published: (2023)
by: Li, Moxin, et al.
Published: (2023)
Making Retrieval-Augmented Language Models Robust to Irrelevant Context
by: Yoran, Ori, et al.
Published: (2023)
by: Yoran, Ori, et al.
Published: (2023)
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025)
by: Zhang, Zongmeng, et al.
Published: (2025)
Advancing Retrieval-Augmented Generation for Persian: Development of Language Models, Comprehensive Benchmarks, and Best Practices for Optimization
by: Hosseinbeigi, Sara Bourbour, et al.
Published: (2025)
by: Hosseinbeigi, Sara Bourbour, et al.
Published: (2025)
STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
by: Quamar, Mohammad Atif, et al.
Published: (2025)
by: Quamar, Mohammad Atif, et al.
Published: (2025)
Temporal Preferences in Language Models for Long-Horizon Assistance
by: Mazyaki, Ali, et al.
Published: (2025)
by: Mazyaki, Ali, et al.
Published: (2025)
Outcome-Constrained Large Language Models for Countering Hate Speech
by: Hong, Lingzi, et al.
Published: (2024)
by: Hong, Lingzi, et al.
Published: (2024)
Enhance Robustness of Language Models Against Variation Attack through Graph Integration
by: Xiong, Zi, et al.
Published: (2024)
by: Xiong, Zi, et al.
Published: (2024)
Camouflage is all you need: Evaluating and Enhancing Language Model Robustness Against Camouflage Adversarial Attacks
by: Huertas-García, Álvaro, et al.
Published: (2024)
by: Huertas-García, Álvaro, et al.
Published: (2024)
Disparities In Negation Understanding Across Languages In Vision-Language Models
by: Moraitaki, Charikleia, et al.
Published: (2026)
by: Moraitaki, Charikleia, et al.
Published: (2026)
Learning Robust Negation Text Representations
by: Truong, Thinh Hung, et al.
Published: (2025)
by: Truong, Thinh Hung, et al.
Published: (2025)
Mitigating Catastrophic Forgetting in Continual Learning through Model Growth
by: Süalp, Ege, et al.
Published: (2025)
by: Süalp, Ege, et al.
Published: (2025)
NegativePrompt: Leveraging Psychology for Large Language Models Enhancement via Negative Emotional Stimuli
by: Wang, Xu, et al.
Published: (2024)
by: Wang, Xu, et al.
Published: (2024)
Making Old Kurdish Publications Processable by Augmenting Available Optical Character Recognition Engines
by: Yaseen, Blnd, et al.
Published: (2024)
by: Yaseen, Blnd, et al.
Published: (2024)
Consistent Joint Decision-Making with Heterogeneous Learning Models
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
by: Faghihi, Hossein Rajaby, et al.
Published: (2024)
Negative-Prompt-driven Alignment for Generative Language Model
by: Qiao, Shiqi, et al.
Published: (2024)
by: Qiao, Shiqi, et al.
Published: (2024)
Vendi-RAG: Adaptively Trading-Off Diversity And Quality Significantly Improves Retrieval Augmented Generation With LLMs
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
Negation Blindness in Large Language Models: Unveiling the NO Syndrome in Image Generation
by: Nadeem, Mohammad, et al.
Published: (2024)
by: Nadeem, Mohammad, et al.
Published: (2024)
Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks
by: Zhou, Andy, et al.
Published: (2024)
by: Zhou, Andy, et al.
Published: (2024)
Negation: A Pink Elephant in the Large Language Models' Room?
by: Vrabcová, Tereza, et al.
Published: (2025)
by: Vrabcová, Tereza, et al.
Published: (2025)
Vision-Language Models Do Not Understand Negation
by: Alhamoud, Kumail, et al.
Published: (2025)
by: Alhamoud, Kumail, et al.
Published: (2025)
What Makes Language Models Good-enough?
by: Asami, Daiki, et al.
Published: (2024)
by: Asami, Daiki, et al.
Published: (2024)
Correcting Negative Bias in Large Language Models through Negative Attention Score Alignment
by: Yu, Sangwon, et al.
Published: (2024)
by: Yu, Sangwon, et al.
Published: (2024)
Similar Items
-
Paraphrasing in Affirmative Terms Improves Negation Understanding
by: Rezaei, MohammadHossein, et al.
Published: (2024) -
Commonsense Knowledge with Negation: A Resource to Enhance Negation Understanding
by: Wang, Zijie, et al.
Published: (2026) -
AI in Mental Health: Emotional and Sentiment Analysis of Large Language Models' Responses to Depression, Anxiety, and Stress Queries
by: VarastehNezhad, Arya, et al.
Published: (2025) -
PersoDPO: Scalable Preference Optimization for Instruction-Adherent, Persona-Grounded Dialogue via Multi-LLM Evaluation
by: Afzoon, Saleh, et al.
Published: (2026) -
EgoNormia: Benchmarking Physical Social Norm Understanding
by: Rezaei, MohammadHossein, et al.
Published: (2025)