OLA: Output Language Alignment in Code-Switched LLM Interactions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Oh, Juhyun, Yoo, Haneul, Haznitrama, Faiz Ghifari, Oh, Alice |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can LLM Generate Culturally Relevant Commonsense QA Data? Case Study in Indonesian and Sundanese
von: Putri, Rifki Afina, et al.
Veröffentlicht: (2024)
von: Putri, Rifki Afina, et al.
Veröffentlicht: (2024)
A Neuropsychologically Grounded Evaluation of LLM Cognitive Abilities
von: Haznitrama, Faiz Ghifari, et al.
Veröffentlicht: (2026)
von: Haznitrama, Faiz Ghifari, et al.
Veröffentlicht: (2026)
Code-Switching In-Context Learning for Cross-Lingual Transfer of Large Language Models
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
von: Yoo, Haneul, et al.
Veröffentlicht: (2025)
Code-Switching Curriculum Learning for Multilingual Transfer in LLMs
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
Survey of Cultural Awareness in Language Models: Text and Beyond
von: Pawar, Siddhesh, et al.
Veröffentlicht: (2024)
von: Pawar, Siddhesh, et al.
Veröffentlicht: (2024)
On the Effect of Uncertainty on Layer-wise Inference Dynamics
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
von: Kim, Sunwoo, et al.
Veröffentlicht: (2025)
Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
Not What, But How: A Communicative Audit of LLM Response Framing
von: Pawar, Siddhesh Milind, et al.
Veröffentlicht: (2026)
von: Pawar, Siddhesh Milind, et al.
Veröffentlicht: (2026)
Multi-FAct: Assessing Factuality of Multilingual LLMs using FActScore
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
von: Shafayat, Sheikh, et al.
Veröffentlicht: (2024)
CLIcK: A Benchmark Dataset of Cultural and Linguistic Intelligence in Korean
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
von: Kim, Eunsu, et al.
Veröffentlicht: (2024)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
von: Yoo, Haneul, et al.
Veröffentlicht: (2024)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
Shared Heritage, Distinct Writing: Rethinking Resource Selection for East Asian Historical Documents
von: Song, Seyoung, et al.
Veröffentlicht: (2024)
von: Song, Seyoung, et al.
Veröffentlicht: (2024)
HERITAGE: An End-to-End Web Platform for Processing Korean Historical Documents in Hanja
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
BenchHub: A Unified Benchmark Suite for Holistic and Customizable LLM Evaluation
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
Uncovering Factor Level Preferences to Improve Human-Model Alignment
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
von: Oh, Juhyun, et al.
Veröffentlicht: (2024)
Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
RECIPE4U: Student-ChatGPT Interaction Dataset in EFL Writing Education
von: Han, Jieun, et al.
Veröffentlicht: (2024)
von: Han, Jieun, et al.
Veröffentlicht: (2024)
FINEST: Improving LLM Responses to Sensitive Topics Through Fine-Grained Evaluation
von: Oh, Juhyun, et al.
Veröffentlicht: (2026)
von: Oh, Juhyun, et al.
Veröffentlicht: (2026)
Culture is Everywhere: A Call for Intentionally Cultural Evaluation
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)
LLM-as-a-tutor in EFL Writing Education: Focusing on Evaluation of Student-LLM Interaction
von: Han, Jieun, et al.
Veröffentlicht: (2023)
von: Han, Jieun, et al.
Veröffentlicht: (2023)
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
von: Jin, Jiho, et al.
Veröffentlicht: (2026)
von: Jin, Jiho, et al.
Veröffentlicht: (2026)
RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
Translating Hanja Historical Documents to Contemporary Korean and English
von: Son, Juhee, et al.
Veröffentlicht: (2022)
von: Son, Juhee, et al.
Veröffentlicht: (2022)
ChEDDAR: Student-ChatGPT Dialogue in EFL Writing Education
von: Han, Jieun, et al.
Veröffentlicht: (2023)
von: Han, Jieun, et al.
Veröffentlicht: (2023)
Open Korean Historical Corpus: A Millennia-Scale Diachronic Collection of Public Domain Texts
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
One-Topic-Doesn't-Fit-All: Transcreating Reading Comprehension Test for Personalized Learning
von: Han, Jieun, et al.
Veröffentlicht: (2025)
von: Han, Jieun, et al.
Veröffentlicht: (2025)
CHOP: Integrating ChatGPT into EFL Oral Presentation Practice
von: Cha, Jungyoub, et al.
Veröffentlicht: (2024)
von: Cha, Jungyoub, et al.
Veröffentlicht: (2024)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
von: Kim, Jiseon, et al.
Veröffentlicht: (2025)
PapersPlease: A Benchmark for Evaluating Motivational Values of Large Language Models Based on ERG Theory
von: Myung, Junho, et al.
Veröffentlicht: (2025)
von: Myung, Junho, et al.
Veröffentlicht: (2025)
GECKO: Generative Language Model for English, Code and Korean
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
MAQA: Evaluating Uncertainty Quantification in LLMs Regarding Data Uncertainty
von: Yang, Yongjin, et al.
Veröffentlicht: (2024)
von: Yang, Yongjin, et al.
Veröffentlicht: (2024)
Can Large Language Models Understand, Reason About, and Generate Code-Switched Text?
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2026)
Does Alignment Tuning Really Break LLMs' Internal Confidence?
von: Oh, Hongseok, et al.
Veröffentlicht: (2024)
von: Oh, Hongseok, et al.
Veröffentlicht: (2024)
LLM-C3MOD: A Human-LLM Collaborative System for Cross-Cultural Hate Speech Moderation
von: Park, Junyeong, et al.
Veröffentlicht: (2025)
von: Park, Junyeong, et al.
Veröffentlicht: (2025)
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
von: Kim, Woojin, et al.
Veröffentlicht: (2026)
von: Kim, Woojin, et al.
Veröffentlicht: (2026)
OrthoRank: Token Selection via Sink Token Orthogonality for Efficient LLM inference
von: Shin, Seungjun, et al.
Veröffentlicht: (2025)
von: Shin, Seungjun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Can LLM Generate Culturally Relevant Commonsense QA Data? Case Study in Indonesian and Sundanese
von: Putri, Rifki Afina, et al.
Veröffentlicht: (2024) -
A Neuropsychologically Grounded Evaluation of LLM Cognitive Abilities
von: Haznitrama, Faiz Ghifari, et al.
Veröffentlicht: (2026) -
Code-Switching In-Context Learning for Cross-Lingual Transfer of Large Language Models
von: Yoo, Haneul, et al.
Veröffentlicht: (2025) -
Code-Switching Curriculum Learning for Multilingual Transfer in LLMs
von: Yoo, Haneul, et al.
Veröffentlicht: (2024) -
Flex-TravelPlanner: A Benchmark for Flexible Planning with Language Agents
von: Oh, Juhyun, et al.
Veröffentlicht: (2025)