We Politely Insist: Your LLM Must Learn the Persian Art of Taarof
Fuente:
arXiv
Saved in:
| Main Authors: | Sadr, Nikta Gohari, Heidariasl, Sahar, Megerdoomian, Karine, Seyyed-Kalantari, Laleh, Emami, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Which Words Matter Most in Zero-Shot Prompts?
by: Sadr, Nikta Gohari, et al.
Published: (2025)
by: Sadr, Nikta Gohari, et al.
Published: (2025)
Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
by: Madhusudan, Sangmitra, et al.
Published: (2025)
by: Madhusudan, Sangmitra, et al.
Published: (2025)
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
by: Kohankhaki, Farnaz, et al.
Published: (2024)
by: Kohankhaki, Farnaz, et al.
Published: (2024)
Soft-prompt Tuning for Large Language Models to Evaluate Bias
by: Tian, Jacob-Junqi, et al.
Published: (2023)
by: Tian, Jacob-Junqi, et al.
Published: (2023)
Khayyam Challenge (PersianMMLU): Is Your LLM Truly Wise to The Persian Language?
by: Ghahroodi, Omid, et al.
Published: (2024)
by: Ghahroodi, Omid, et al.
Published: (2024)
Can We Afford The Perfect Prompt? Balancing Cost and Accuracy with the Economical Prompting Index
by: McDonald, Tyler, et al.
Published: (2024)
by: McDonald, Tyler, et al.
Published: (2024)
Memory Dial: A Training Framework for Controllable Memorization in Language Models
by: Zhang, Xiangbo, et al.
Published: (2026)
by: Zhang, Xiangbo, et al.
Published: (2026)
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries
by: Sun, Jing Han, et al.
Published: (2024)
by: Sun, Jing Han, et al.
Published: (2024)
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
by: Zahraei, Pardis Sadat, et al.
Published: (2025)
Personality Matters: User Traits Predict LLM Preferences in Multi-Turn Collaborative Tasks
by: Yunusov, Sarfaroz, et al.
Published: (2025)
by: Yunusov, Sarfaroz, et al.
Published: (2025)
To Build Our Future, We Must Know Our Past: Contextualizing Paradigm Shifts in Natural Language Processing
by: Gururaja, Sireesh, et al.
Published: (2023)
by: Gururaja, Sireesh, et al.
Published: (2023)
Stop Overvaluing Multi-Agent Debate -- We Must Rethink Evaluation and Embrace Model Heterogeneity
by: Zhang, Hangfan, et al.
Published: (2025)
by: Zhang, Hangfan, et al.
Published: (2025)
Trace-of-Thought Prompting: Investigating Prompt-Based Knowledge Distillation Through Question Decomposition
by: McDonald, Tyler, et al.
Published: (2025)
by: McDonald, Tyler, et al.
Published: (2025)
Common to Whom? Regional Cultural Commonsense and LLM Bias in India
by: Madhusudan, Sangmitra, et al.
Published: (2026)
by: Madhusudan, Sangmitra, et al.
Published: (2026)
Deep Learning-based Sentiment Analysis in Persian Language
by: Heydari, Mohammad, et al.
Published: (2024)
by: Heydari, Mohammad, et al.
Published: (2024)
KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches
by: Yuan, Jiayi, et al.
Published: (2024)
by: Yuan, Jiayi, et al.
Published: (2024)
Opportunities for Persian Digital Humanities Research with Artificial Intelligence Language Models; Case Study: Forough Farrokhzad
by: Meymandi, Arash Rasti, et al.
Published: (2024)
by: Meymandi, Arash Rasti, et al.
Published: (2024)
Should We Respect LLMs? A Cross-Lingual Study on the Influence of Prompt Politeness on LLM Performance
by: Yin, Ziqi, et al.
Published: (2024)
by: Yin, Ziqi, et al.
Published: (2024)
MirrorStories: Reflecting Diversity through Personalized Narrative Generation with Large Language Models
by: Yunusov, Sarfaroz, et al.
Published: (2024)
by: Yunusov, Sarfaroz, et al.
Published: (2024)
Your Model is Overconfident, and Other Lies We Tell Ourselves
by: Mickus, Timothee, et al.
Published: (2025)
by: Mickus, Timothee, et al.
Published: (2025)
The Dog the Cat Chased Stumped the Model: Measuring When Language Models Abandon Structure for Shortcuts
by: Madhusudan, Sangmitra, et al.
Published: (2025)
by: Madhusudan, Sangmitra, et al.
Published: (2025)
DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training
by: Pan, Ziwen, et al.
Published: (2026)
by: Pan, Ziwen, et al.
Published: (2026)
WSC+: Enhancing The Winograd Schema Challenge Using Tree-of-Experts
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
by: Zahraei, Pardis Sadat, et al.
Published: (2024)
ELAB: Extensive LLM Alignment Benchmark in Persian Language
by: Pourbahman, Zahra, et al.
Published: (2025)
by: Pourbahman, Zahra, et al.
Published: (2025)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
by: Kumar, Abhishek, et al.
Published: (2024)
by: Kumar, Abhishek, et al.
Published: (2024)
HamRaz: A Culture-Based Persian Conversation Dataset for Person-Centered Therapy Using LLM Agents
by: Abbasi, Mohammad Amin, et al.
Published: (2025)
by: Abbasi, Mohammad Amin, et al.
Published: (2025)
Winning with Less for Low Resource Languages: Advantage of Cross-Lingual English_Persian Argument Mining Model over LLM Augmentation
by: Jahan, Ali, et al.
Published: (2025)
by: Jahan, Ali, et al.
Published: (2025)
PersianMind: A Cross-Lingual Persian-English Large Language Model
by: Rostami, Pedram, et al.
Published: (2024)
by: Rostami, Pedram, et al.
Published: (2024)
"We Demand Justice!": Towards Social Context Grounding of Political Texts
by: Pujari, Rajkumar, et al.
Published: (2023)
by: Pujari, Rajkumar, et al.
Published: (2023)
Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk
by: Chen, Zichen, et al.
Published: (2025)
by: Chen, Zichen, et al.
Published: (2025)
Unmasking the Factual-Conceptual Gap in Persian Language Models
by: Sakhaeirad, Alireza, et al.
Published: (2026)
by: Sakhaeirad, Alireza, et al.
Published: (2026)
Calibrated Language Models Must Hallucinate
by: Kalai, Adam Tauman, et al.
Published: (2023)
by: Kalai, Adam Tauman, et al.
Published: (2023)
Telecom Language Models: Must They Be Large?
by: Piovesan, Nicola, et al.
Published: (2024)
by: Piovesan, Nicola, et al.
Published: (2024)
TALE: A Tool-Augmented Framework for Reference-Free Evaluation of Large Language Models
by: Badshah, Sher, et al.
Published: (2025)
by: Badshah, Sher, et al.
Published: (2025)
Mind Your Tone: Investigating How Prompt Politeness Affects LLM Accuracy (short paper)
by: Dobariya, Om, et al.
Published: (2025)
by: Dobariya, Om, et al.
Published: (2025)
TookaBERT: A Step Forward for Persian NLU
by: SadraeiJavaheri, MohammadAli, et al.
Published: (2024)
by: SadraeiJavaheri, MohammadAli, et al.
Published: (2024)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
by: Morabito, Robert, et al.
Published: (2024)
by: Morabito, Robert, et al.
Published: (2024)
Can LLMs Separate Instructions From Data? And What Do We Even Mean By That?
by: Zverev, Egor, et al.
Published: (2024)
by: Zverev, Egor, et al.
Published: (2024)
A Computational Approach to Language Contact -- A Case Study of Persian
by: Basirat, Ali, et al.
Published: (2026)
by: Basirat, Ali, et al.
Published: (2026)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
by: Abdallah, Abdelrahman, et al.
Published: (2025)
by: Abdallah, Abdelrahman, et al.
Published: (2025)
Similar Items
-
Which Words Matter Most in Zero-Shot Prompts?
by: Sadr, Nikta Gohari, et al.
Published: (2025) -
Fine-Tuned LLMs are "Time Capsules" for Tracking Societal Bias Through Books
by: Madhusudan, Sangmitra, et al.
Published: (2025) -
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
by: Kohankhaki, Farnaz, et al.
Published: (2024) -
Soft-prompt Tuning for Large Language Models to Evaluate Bias
by: Tian, Jacob-Junqi, et al.
Published: (2023) -
Khayyam Challenge (PersianMMLU): Is Your LLM Truly Wise to The Persian Language?
by: Ghahroodi, Omid, et al.
Published: (2024)