Women, Infamous, and Exotic Beings: A Comparative Study of Honorific Usages in Wikipedia and LLMs for Bengali and Hindi
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Mukherjee, Sourabrata, Mehta, Atharva, Saha, Sougata, Arora, Akhil, Choudhury, Monojit |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
par: Saha, Sougata, et autres
Publié: (2025)
par: Saha, Sougata, et autres
Publié: (2025)
To Generate or Discriminate? Methodological Considerations for Measuring Cultural Alignment in LLMs
par: Pandey, Saurabh Kumar, et autres
Publié: (2026)
par: Pandey, Saurabh Kumar, et autres
Publié: (2026)
Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?
par: Saha, Sougata, et autres
Publié: (2025)
par: Saha, Sougata, et autres
Publié: (2025)
Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness
par: Saha, Sougata, et autres
Publié: (2025)
par: Saha, Sougata, et autres
Publié: (2025)
Missing Melodies: AI Music Generation and its "Nearly" Complete Omission of the Global South
par: Mehta, Atharva, et autres
Publié: (2024)
par: Mehta, Atharva, et autres
Publié: (2024)
Exploring Adapter Design Tradeoffs for Low Resource Music Generation
par: Mehta, Atharva, et autres
Publié: (2025)
par: Mehta, Atharva, et autres
Publié: (2025)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
par: Das, Mithun, et autres
Publié: (2024)
par: Das, Mithun, et autres
Publié: (2024)
Text Detoxification as Style Transfer in English and Hindi
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
All that is English may be Hindi: Enhancing language identification through automatic ranking of likeliness of word borrowing in social media
par: Patro, Jasabanta, et autres
Publié: (2017)
par: Patro, Jasabanta, et autres
Publié: (2017)
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks
par: Rao, Abhinav, et autres
Publié: (2023)
par: Rao, Abhinav, et autres
Publié: (2023)
Music for All: Representational Bias and Cross-Cultural Adaptability of Music Generation Models
par: Mehta, Atharva, et autres
Publié: (2025)
par: Mehta, Atharva, et autres
Publié: (2025)
On the effective transfer of knowledge from English to Hindi Wikipedia
par: Das, Paramita, et autres
Publié: (2024)
par: Das, Paramita, et autres
Publié: (2024)
Text Style Transfer: An Introductory Overview
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
[WIP] Jailbreak Paradox: The Achilles' Heel of LLMs
par: Rao, Abhinav, et autres
Publié: (2024)
par: Rao, Abhinav, et autres
Publié: (2024)
Modeling Romanized Hindi and Bengali: Dataset Creation and Multilingual LLM Integration
par: Gharami, Kanchon, et autres
Publié: (2025)
par: Gharami, Kanchon, et autres
Publié: (2025)
BERTopic for Topic Modeling of Hindi Short Texts: A Comparative Study
par: Mutsaddi, Atharva, et autres
Publié: (2025)
par: Mutsaddi, Atharva, et autres
Publié: (2025)
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test
par: Khandelwal, Aditi, et autres
Publié: (2024)
par: Khandelwal, Aditi, et autres
Publié: (2024)
Are Large Language Models Actually Good at Text Style Transfer?
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
A Survey of Text Style Transfer: Applications and Ethical Implications
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
par: Mukherjee, Sourabrata, et autres
Publié: (2024)
Do Language Models Understand Honorific Systems in Javanese?
par: Farhansyah, Mohammad Rifqi, et autres
Publié: (2025)
par: Farhansyah, Mohammad Rifqi, et autres
Publié: (2025)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
par: Agarwal, Utkarsh, et autres
Publié: (2024)
par: Agarwal, Utkarsh, et autres
Publié: (2024)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
par: Dhamaskar, Mohammed Amaan, et autres
Publié: (2025)
par: Dhamaskar, Mohammed Amaan, et autres
Publié: (2025)
Entity Insertion in Multilingual Linked Corpora: The Case of Wikipedia
par: Feith, Tomás, et autres
Publié: (2024)
par: Feith, Tomás, et autres
Publié: (2024)
Consolidating Strategies for Countering Hate Speech Using Persuasive Dialogues
par: Saha, Sougata, et autres
Publié: (2024)
par: Saha, Sougata, et autres
Publié: (2024)
Benchmarking Hindi LLMs: A New Suite of Datasets and a Comparative Analysis
par: Kamath, Anusha, et autres
Publié: (2025)
par: Kamath, Anusha, et autres
Publié: (2025)
Automatic Speech Recognition for Hindi
par: Saha, Anish, et autres
Publié: (2024)
par: Saha, Anish, et autres
Publié: (2024)
Towards Measuring and Modeling "Culture" in LLMs: A Survey
par: Adilazuarda, Muhammad Farid, et autres
Publié: (2024)
par: Adilazuarda, Muhammad Farid, et autres
Publié: (2024)
Building Benchmarks from the Ground Up: Community-Centered Evaluation of LLMs in Healthcare Chatbot Settings
par: Hamna, Hamna, et autres
Publié: (2025)
par: Hamna, Hamna, et autres
Publié: (2025)
Evaluating Subword Tokenization Techniques for Bengali: A Benchmark Study with BengaliBPE
par: Patwary, Firoj Ahmmed, et autres
Publié: (2025)
par: Patwary, Firoj Ahmmed, et autres
Publié: (2025)
Steering Conversational Large Language Models for Long Emotional Support Conversations
par: Madani, Navid, et autres
Publié: (2024)
par: Madani, Navid, et autres
Publié: (2024)
Evaluating Text Style Transfer Evaluation: Are There Any Reliable Metrics?
par: Mukherjee, Sourabrata, et autres
Publié: (2025)
par: Mukherjee, Sourabrata, et autres
Publié: (2025)
Evaluating Large Language Models for Health-related Queries with Presuppositions
par: Kaur, Navreet, et autres
Publié: (2023)
par: Kaur, Navreet, et autres
Publié: (2023)
Polite on the Surface, Wrong in Practice: A Curated Dataset for Fixing Honorific Failures in Multilingual Bangla Generation
par: Shuvo, Md. Asaduzzaman, et autres
Publié: (2026)
par: Shuvo, Md. Asaduzzaman, et autres
Publié: (2026)
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting
par: Mukherjee, Sagnik, et autres
Publié: (2024)
par: Mukherjee, Sagnik, et autres
Publié: (2024)
HiMed: Incentivizing Hindi Reasoning in Medical LLMs
par: Jiang, Dingfeng, et autres
Publié: (2026)
par: Jiang, Dingfeng, et autres
Publié: (2026)
The Honorific Effect: Exploring the Impact of Japanese Linguistic Formalities on AI-Generated Physics Explanations
par: Sato, Keisuke
Publié: (2024)
par: Sato, Keisuke
Publié: (2024)
Sacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
par: Atif, Farah, et autres
Publié: (2025)
par: Atif, Farah, et autres
Publié: (2025)
BNLI: A Linguistically-Refined Bengali Dataset for Natural Language Inference
par: Haque, Farah Binta, et autres
Publié: (2025)
par: Haque, Farah Binta, et autres
Publié: (2025)
Parameter-Efficient Fine-Tuning for Low-Resource Languages: A Comparative Study of LLMs for Bengali Hate Speech Detection
par: Islam, Akif, et autres
Publié: (2025)
par: Islam, Akif, et autres
Publié: (2025)
Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models
par: Mittal, Avni, et autres
Publié: (2026)
par: Mittal, Avni, et autres
Publié: (2026)
Documents similaires
-
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
par: Saha, Sougata, et autres
Publié: (2025) -
To Generate or Discriminate? Methodological Considerations for Measuring Cultural Alignment in LLMs
par: Pandey, Saurabh Kumar, et autres
Publié: (2026) -
Reading between the Lines: Can LLMs Identify Cross-Cultural Communication Gaps?
par: Saha, Sougata, et autres
Publié: (2025) -
Meta-Cultural Competence: Climbing the Right Hill of Cultural Awareness
par: Saha, Sougata, et autres
Publié: (2025) -
Missing Melodies: AI Music Generation and its "Nearly" Complete Omission of the Global South
par: Mehta, Atharva, et autres
Publié: (2024)