Evaluating Large language models on Understanding Korean indirect Speech acts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Koo, Youngeun, Lee, Jiwoo, Park, Dojun, Park, Seohyun, Lee, Sungeun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
MultiPragEval: Multilingual Pragmatic Evaluation of Large Language Models
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
Multi-Dimensional Machine Translation Evaluation: Model Evaluation and Resource for Korean
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings
von: Lee, Jonghyun, et al.
Veröffentlicht: (2025)
von: Lee, Jonghyun, et al.
Veröffentlicht: (2025)
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
von: Hahm, Sungeun, et al.
Veröffentlicht: (2025)
von: Hahm, Sungeun, et al.
Veröffentlicht: (2025)
KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs
von: Kim, Haechan, et al.
Veröffentlicht: (2026)
von: Kim, Haechan, et al.
Veröffentlicht: (2026)
OLKAVS: An Open Large-Scale Korean Audio-Visual Speech Dataset
von: Park, Jeongkyun, et al.
Veröffentlicht: (2023)
von: Park, Jeongkyun, et al.
Veröffentlicht: (2023)
Ko-MuSR: A Multistep Soft Reasoning Benchmark for LLMs Capable of Understanding Korean
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
Making Qwen3 Think in Korean with Reinforcement Learning
von: Lee, Jungyup, et al.
Veröffentlicht: (2025)
von: Lee, Jungyup, et al.
Veröffentlicht: (2025)
Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
Unlocking Korean Verbs: A User-Friendly Exploration into the Verb Lexicon
von: Song, Seohyun, et al.
Veröffentlicht: (2024)
von: Song, Seohyun, et al.
Veröffentlicht: (2024)
Evaluating the Consistency of LLM Evaluators
von: Lee, Noah, et al.
Veröffentlicht: (2024)
von: Lee, Noah, et al.
Veröffentlicht: (2024)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
KMMLU: Measuring Massive Multitask Language Understanding in Korean
von: Son, Guijin, et al.
Veröffentlicht: (2024)
von: Son, Guijin, et al.
Veröffentlicht: (2024)
Constituency Structure over Eojeol in Korean Treebanks
von: Park, Jungyeul, et al.
Veröffentlicht: (2025)
von: Park, Jungyeul, et al.
Veröffentlicht: (2025)
Thunder-LLM: Efficiently Adapting LLMs to Korean with Minimal Resources
von: Kim, Jinpyo, et al.
Veröffentlicht: (2025)
von: Kim, Jinpyo, et al.
Veröffentlicht: (2025)
SDS KoPub VDR: A Benchmark Dataset for Visual Document Retrieval in Korean Public Documents
von: Lee, Jaehoon, et al.
Veröffentlicht: (2025)
von: Lee, Jaehoon, et al.
Veröffentlicht: (2025)
KoDialogBench: Evaluating Conversational Understanding of Language Models with Korean Dialogue Benchmark
von: Jang, Seongbo, et al.
Veröffentlicht: (2024)
von: Jang, Seongbo, et al.
Veröffentlicht: (2024)
Steering LLMs toward Korean Local Speech: Iterative Refinement Framework for Faithful Dialect Translation
von: Park, Keunhyeung, et al.
Veröffentlicht: (2025)
von: Park, Keunhyeung, et al.
Veröffentlicht: (2025)
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
von: Kim, Nayeon, et al.
Veröffentlicht: (2025)
von: Kim, Nayeon, et al.
Veröffentlicht: (2025)
Evaluating Multimodal Generative AI with Korean Educational Standards
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
How language models extrapolate outside the training data: A case study in Textualized Gridworld
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
Open Ko-LLM Leaderboard2: Bridging Foundational and Practical Evaluation for Korean LLMs
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
KFinEval-Pilot: A Comprehensive Benchmark Suite for Korean Financial Language Understanding
von: Hwang, Bokwang, et al.
Veröffentlicht: (2025)
von: Hwang, Bokwang, et al.
Veröffentlicht: (2025)
ODPG: Outfitting Diffusion with Pose Guided Condition
von: Lee, Seohyun, et al.
Veröffentlicht: (2025)
von: Lee, Seohyun, et al.
Veröffentlicht: (2025)
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
von: Kim, Eunsu, et al.
Veröffentlicht: (2025)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
von: Kim, Dain, et al.
Veröffentlicht: (2026)
von: Kim, Dain, et al.
Veröffentlicht: (2026)
KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters
von: Kim, SungHo, et al.
Veröffentlicht: (2026)
von: Kim, SungHo, et al.
Veröffentlicht: (2026)
Inappropriate Pause Detection In Dysarthric Speech Using Large-Scale Speech Recognition
von: Lee, Jeehyun, et al.
Veröffentlicht: (2024)
von: Lee, Jeehyun, et al.
Veröffentlicht: (2024)
Enhancing Korean Dependency Parsing with Morphosyntactic Features
von: Park, Jungyeul, et al.
Veröffentlicht: (2025)
von: Park, Jungyeul, et al.
Veröffentlicht: (2025)
Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding
von: Jung, Sungmok, et al.
Veröffentlicht: (2026)
von: Jung, Sungmok, et al.
Veröffentlicht: (2026)
Learning from Negative Samples in Biomedical Generative Entity Linking
von: Kim, Chanhwi, et al.
Veröffentlicht: (2024)
von: Kim, Chanhwi, et al.
Veröffentlicht: (2024)
Thunder-Tok: Minimizing Tokens per Word in Tokenizing Korean Texts for Generative Language Models
von: Cho, Gyeongje, et al.
Veröffentlicht: (2025)
von: Cho, Gyeongje, et al.
Veröffentlicht: (2025)
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models
von: Kim, Yeeun, et al.
Veröffentlicht: (2024)
von: Kim, Yeeun, et al.
Veröffentlicht: (2024)
KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness
von: Kim, Jinyoung, et al.
Veröffentlicht: (2026)
von: Kim, Jinyoung, et al.
Veröffentlicht: (2026)
LifeTox: Unveiling Implicit Toxicity in Life Advice
von: Kim, Minbeom, et al.
Veröffentlicht: (2023)
von: Kim, Minbeom, et al.
Veröffentlicht: (2023)
Automata-based constraints for language model decoding
von: Koo, Terry, et al.
Veröffentlicht: (2024)
von: Koo, Terry, et al.
Veröffentlicht: (2024)
Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean
von: Choi, ChangSu, et al.
Veröffentlicht: (2024)
von: Choi, ChangSu, et al.
Veröffentlicht: (2024)
HiKE: Hierarchical Evaluation Framework for Korean-English Code-Switching Speech Recognition
von: Paik, Gio, et al.
Veröffentlicht: (2025)
von: Paik, Gio, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
von: Park, Dojun, et al.
Veröffentlicht: (2024) -
MultiPragEval: Multilingual Pragmatic Evaluation of Large Language Models
von: Park, Dojun, et al.
Veröffentlicht: (2024) -
Multi-Dimensional Machine Translation Evaluation: Model Evaluation and Resource for Korean
von: Park, Dojun, et al.
Veröffentlicht: (2024) -
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings
von: Lee, Jonghyun, et al.
Veröffentlicht: (2025) -
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
von: Hahm, Sungeun, et al.
Veröffentlicht: (2025)