Designing LLMs for cultural sensitivity: Evidence from English-Japanese translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tenzer, Helene, Abidi, Oumnia, Feuerriegel, Stefan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
NARRA-Gym for Evaluating Interactive Narrative Agents
por: Huang, Yue, et al.
Publicado: (2026)
por: Huang, Yue, et al.
Publicado: (2026)
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
por: Arita, Takaya, et al.
Publicado: (2025)
por: Arita, Takaya, et al.
Publicado: (2025)
Your Students Don't Use LLMs Like You Wish They Did
por: Kobler, Sebastian, et al.
Publicado: (2026)
por: Kobler, Sebastian, et al.
Publicado: (2026)
Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
por: Ivey, Jonathan, et al.
Publicado: (2024)
por: Ivey, Jonathan, et al.
Publicado: (2024)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
por: Badawi, Abeer, et al.
Publicado: (2025)
por: Badawi, Abeer, et al.
Publicado: (2025)
Designing Computational Tools for Exploring Causal Relationships in Qualitative Data
por: Meng, Han, et al.
Publicado: (2026)
por: Meng, Han, et al.
Publicado: (2026)
"Ownership, Not Just Happy Talk": Co-Designing a Participatory Large Language Model for Journalism
por: Tseng, Emily, et al.
Publicado: (2025)
por: Tseng, Emily, et al.
Publicado: (2025)
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
por: Liu, Houjiang, et al.
Publicado: (2023)
por: Liu, Houjiang, et al.
Publicado: (2023)
Design and consensus content validity of the questionnaire for b-learning education: A 2-Tuple Fuzzy Linguistic Delphi based Decision Support Tool
por: Montes, Rosana, et al.
Publicado: (2024)
por: Montes, Rosana, et al.
Publicado: (2024)
Bottom-Up Perspectives on AI Governance: Insights from User Reviews of AI Products
por: Pasch, Stefan
Publicado: (2025)
por: Pasch, Stefan
Publicado: (2025)
First Contact with Dark Patterns and Deceptive Designs in Chinese and Japanese Free-to-Play Mobile Games
por: Zhang, Gloria Xiaodan, et al.
Publicado: (2025)
por: Zhang, Gloria Xiaodan, et al.
Publicado: (2025)
LLM Content Moderation and User Satisfaction: Evidence from Response Refusals in Chatbot Arena
por: Pasch, Stefan
Publicado: (2025)
por: Pasch, Stefan
Publicado: (2025)
Overreliance on AI in Information-seeking from Video Content
por: Møller, Anders Giovanni, et al.
Publicado: (2026)
por: Møller, Anders Giovanni, et al.
Publicado: (2026)
Evidence of conceptual mastery in the application of rules by Large Language Models
por: Nunes, José Luiz, et al.
Publicado: (2025)
por: Nunes, José Luiz, et al.
Publicado: (2025)
Can LLMs Reason About Trust?: A Pilot Study
por: Debnath, Anushka, et al.
Publicado: (2025)
por: Debnath, Anushka, et al.
Publicado: (2025)
The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness
por: Subedi, Krishna
Publicado: (2025)
por: Subedi, Krishna
Publicado: (2025)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
por: Ranjan, Rajesh, et al.
Publicado: (2024)
por: Ranjan, Rajesh, et al.
Publicado: (2024)
Evidence of a log scaling law for political persuasion with large language models
por: Hackenburg, Kobi, et al.
Publicado: (2024)
por: Hackenburg, Kobi, et al.
Publicado: (2024)
Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms
por: Ashkinaze, Joshua, et al.
Publicado: (2024)
por: Ashkinaze, Joshua, et al.
Publicado: (2024)
From Script to Stage: Automating Experimental Design for Social Simulations with LLMs
por: Guo, Yuwei, et al.
Publicado: (2025)
por: Guo, Yuwei, et al.
Publicado: (2025)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
por: Dong, Wenhan, et al.
Publicado: (2025)
por: Dong, Wenhan, et al.
Publicado: (2025)
Red Teaming LLMs as Socio-Technical Practice: From Exploration and Data Creation to Evaluation
por: Garcia, Adriana Alvarado, et al.
Publicado: (2026)
por: Garcia, Adriana Alvarado, et al.
Publicado: (2026)
From "Help" to Helpful: A Hierarchical Assessment of LLMs in Mental e-Health Applications
por: Steigerwald, Philipp, et al.
Publicado: (2026)
por: Steigerwald, Philipp, et al.
Publicado: (2026)
Contextualizing Recommendation Explanations with LLMs: A User Study
por: Feng, Yuanjun, et al.
Publicado: (2025)
por: Feng, Yuanjun, et al.
Publicado: (2025)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
por: Zhou, Yuhang, et al.
Publicado: (2025)
por: Zhou, Yuhang, et al.
Publicado: (2025)
What is Stigma Attributed to? A Theory-Grounded, Expert-Annotated Interview Corpus for Demystifying Mental-Health Stigma
por: Meng, Han, et al.
Publicado: (2025)
por: Meng, Han, et al.
Publicado: (2025)
Epistemological Fault Lines Between Human and Artificial Intelligence
por: Quattrociocchi, Walter, et al.
Publicado: (2025)
por: Quattrociocchi, Walter, et al.
Publicado: (2025)
"Would You Want an AI Tutor?" Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom
por: Fuligni, Caterina, et al.
Publicado: (2025)
por: Fuligni, Caterina, et al.
Publicado: (2025)
When Algorithms Meet Artists: Semantic Compression of Artists' Concerns in the Public AI-Art Debate
por: Mukherjee-Gandhi, Ariya, et al.
Publicado: (2025)
por: Mukherjee-Gandhi, Ariya, et al.
Publicado: (2025)
A Systematic Review on the Evaluation of Large Language Models in Theory of Mind Tasks
por: Sarıtaş, Karahan, et al.
Publicado: (2025)
por: Sarıtaş, Karahan, et al.
Publicado: (2025)
How do datasets, developers, and models affect biases in a low-resourced language?: The Case of the Bengali Language
por: Das, Dipto, et al.
Publicado: (2025)
por: Das, Dipto, et al.
Publicado: (2025)
Human Capital Visualization using Speech Amount during Meetings
por: Hashimoto, Ekai, et al.
Publicado: (2025)
por: Hashimoto, Ekai, et al.
Publicado: (2025)
Enhancing Mathematics Learning for Hard-of-Hearing Students Through Real-Time Palestinian Sign Language Recognition: A New Dataset
por: Khandaqji, Fidaa, et al.
Publicado: (2025)
por: Khandaqji, Fidaa, et al.
Publicado: (2025)
Longitudinal Monitoring of LLM Content Moderation of Social Issues
por: Dai, Yunlang, et al.
Publicado: (2025)
por: Dai, Yunlang, et al.
Publicado: (2025)
Cooperative Speech, Semantic Competence, and AI
por: Almotahari, Mahrad
Publicado: (2025)
por: Almotahari, Mahrad
Publicado: (2025)
AI as a deliberative partner fosters intercultural empathy for Americans but fails for Latin American participants
por: Villanueva, Isabel, et al.
Publicado: (2025)
por: Villanueva, Isabel, et al.
Publicado: (2025)
The Impact and Feasibility of Self-Confidence Shaping for AI-Assisted Decision-Making
por: Takayanagi, Takehiro, et al.
Publicado: (2025)
por: Takayanagi, Takehiro, et al.
Publicado: (2025)
Deconstructing Depression Stigma: Integrating AI-driven Data Collection and Analysis with Causal Knowledge Graphs
por: Meng, Han, et al.
Publicado: (2025)
por: Meng, Han, et al.
Publicado: (2025)
Voices of Freelance Professional Writers on AI: Limitations, Expectations, and Fears
por: Ivanova, Anastasiia, et al.
Publicado: (2025)
por: Ivanova, Anastasiia, et al.
Publicado: (2025)
DiMA: An LLM-Powered Ride-Hailing Assistant at DiDi
por: Ning, Yansong, et al.
Publicado: (2025)
por: Ning, Yansong, et al.
Publicado: (2025)
Ejemplares similares
-
NARRA-Gym for Evaluating Interactive Narrative Agents
por: Huang, Yue, et al.
Publicado: (2026) -
Assessing LLMs in Art Contexts: Critique Generation and Theory of Mind Evaluation
por: Arita, Takaya, et al.
Publicado: (2025) -
Your Students Don't Use LLMs Like You Wish They Did
por: Kobler, Sebastian, et al.
Publicado: (2026) -
Real or Robotic? Assessing Whether LLMs Accurately Simulate Qualities of Human Responses in Dialogue
por: Ivey, Jonathan, et al.
Publicado: (2024) -
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
por: Badawi, Abeer, et al.
Publicado: (2025)