KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, SungHo, Park, Juhyeong, Kim, Yeachan, Lee, SangKeun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SCRIPT: A Subcharacter Compositional Representation Injection Module for Korean Pre-Trained Language Models
von: Kim, SungHo, et al.
Veröffentlicht: (2026)
von: Kim, SungHo, et al.
Veröffentlicht: (2026)
Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean
von: Kim, SungHo, et al.
Veröffentlicht: (2025)
von: Kim, SungHo, et al.
Veröffentlicht: (2025)
Incorporating Domain Knowledge into Materials Tokenization
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
von: Oh, Yerim, et al.
Veröffentlicht: (2025)
Zero-shot Commonsense Reasoning over Machine Imagination
von: Park, Hyuntae, et al.
Veröffentlicht: (2024)
von: Park, Hyuntae, et al.
Veröffentlicht: (2024)
MELT: Materials-aware Continued Pre-training for Language Model Adaptation to Materials Science
von: Kim, Junho, et al.
Veröffentlicht: (2024)
von: Kim, Junho, et al.
Veröffentlicht: (2024)
Enhancing Zero-shot Commonsense Reasoning by Integrating Visual Knowledge via Machine Imagination
von: Park, Hyuntae, et al.
Veröffentlicht: (2026)
von: Park, Hyuntae, et al.
Veröffentlicht: (2026)
Bridging the Gap Between Molecule and Textual Descriptions via Substructure-aware Alignment
von: Park, Hyuntae, et al.
Veröffentlicht: (2025)
von: Park, Hyuntae, et al.
Veröffentlicht: (2025)
Mentor-KD: Making Small Language Models Better Multi-step Reasoners
von: Lee, Hojae, et al.
Veröffentlicht: (2024)
von: Lee, Hojae, et al.
Veröffentlicht: (2024)
CleaR: Towards Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Label Learning
von: Kim, Yeachan, et al.
Veröffentlicht: (2024)
von: Kim, Yeachan, et al.
Veröffentlicht: (2024)
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
von: Kim, Nayeon, et al.
Veröffentlicht: (2025)
von: Kim, Nayeon, et al.
Veröffentlicht: (2025)
C2A: Client-Customized Adaptation for Parameter-Efficient Federated Learning
von: Kim, Yeachan, et al.
Veröffentlicht: (2024)
von: Kim, Yeachan, et al.
Veröffentlicht: (2024)
Open Ko-LLM Leaderboard2: Bridging Foundational and Practical Evaluation for Korean LLMs
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
Obfuscation Rules for Detecting and Detoxifying Korean Toxicity
von: Lee, Yejin, et al.
Veröffentlicht: (2025)
von: Lee, Yejin, et al.
Veröffentlicht: (2025)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean
von: Choi, ChangSu, et al.
Veröffentlicht: (2024)
von: Choi, ChangSu, et al.
Veröffentlicht: (2024)
SLM-Based Agentic AI with P-C-G: Optimized for Korean Tool Use
von: Jeon, Changhyun, et al.
Veröffentlicht: (2025)
von: Jeon, Changhyun, et al.
Veröffentlicht: (2025)
KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs
von: Kim, Haechan, et al.
Veröffentlicht: (2026)
von: Kim, Haechan, et al.
Veröffentlicht: (2026)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
MolTRES: Improving Chemical Language Representation Learning for Molecular Property Prediction
von: Park, Jun-Hyung, et al.
Veröffentlicht: (2024)
von: Park, Jun-Hyung, et al.
Veröffentlicht: (2024)
DIVE: Towards Descriptive and Diverse Visual Commonsense Generation
von: Park, Jun-Hyung, et al.
Veröffentlicht: (2024)
von: Park, Jun-Hyung, et al.
Veröffentlicht: (2024)
Evaluating Multimodal Generative AI with Korean Educational Standards
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
KatFishNet: Detecting LLM-Generated Korean Text through Linguistic Feature Analysis
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
von: Park, Shinwoo, et al.
Veröffentlicht: (2025)
KoCoSa: Korean Context-aware Sarcasm Detection Dataset
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
von: Kim, Yumin, et al.
Veröffentlicht: (2024)
UKTA: Unified Korean Text Analyzer
von: Ahn, Seokho, et al.
Veröffentlicht: (2025)
von: Ahn, Seokho, et al.
Veröffentlicht: (2025)
Ko-MuSR: A Multistep Soft Reasoning Benchmark for LLMs Capable of Understanding Korean
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
von: Kim, Kyuhee, et al.
Veröffentlicht: (2025)
von: Kim, Kyuhee, et al.
Veröffentlicht: (2025)
REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning
von: Lee, Seungmin, et al.
Veröffentlicht: (2026)
von: Lee, Seungmin, et al.
Veröffentlicht: (2026)
KoBBQ: Korean Bias Benchmark for Question Answering
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
von: Jin, Jiho, et al.
Veröffentlicht: (2023)
Leveraging Large Language Models for Active Merchant Non-player Characters
von: Kim, Byungjun, et al.
Veröffentlicht: (2024)
von: Kim, Byungjun, et al.
Veröffentlicht: (2024)
From KMMLU-Redux to KMMLU-Pro: A Professional Korean Benchmark Suite for LLM Evaluation
von: Hong, Seokhee, et al.
Veröffentlicht: (2025)
von: Hong, Seokhee, et al.
Veröffentlicht: (2025)
GECKO: Generative Language Model for English, Code and Korean
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
von: Oh, Sungwoo, et al.
Veröffentlicht: (2024)
SoftEDA: Rethinking Rule-Based Data Augmentation with Soft Labels
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
Theme-Explanation Structure for Table Summarization using Large Language Models: A Case Study on Korean Tabular Data
von: Kwack, TaeYoon, et al.
Veröffentlicht: (2025)
von: Kwack, TaeYoon, et al.
Veröffentlicht: (2025)
Acceleration of Grokking in Learning Arithmetic Operations via Kolmogorov-Arnold Representation
von: Park, Yeachan, et al.
Veröffentlicht: (2024)
von: Park, Yeachan, et al.
Veröffentlicht: (2024)
Language-Agnostic Suicidal Risk Detection Using Large Language Models
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
Expanding Foundational Language Capabilities in Open-Source LLMs through a Korean Case Study
von: Lim, Junghwan, et al.
Veröffentlicht: (2025)
von: Lim, Junghwan, et al.
Veröffentlicht: (2025)
Building Resource-Constrained Language Agents: A Korean Case Study on Chemical Toxicity Information
von: Cho, Hojun, et al.
Veröffentlicht: (2025)
von: Cho, Hojun, et al.
Veröffentlicht: (2025)
KFinEval-Pilot: A Comprehensive Benchmark Suite for Korean Financial Language Understanding
von: Hwang, Bokwang, et al.
Veröffentlicht: (2025)
von: Hwang, Bokwang, et al.
Veröffentlicht: (2025)
KMI: A Dataset of Korean Motivational Interviewing Dialogues for Psychotherapy
von: Kim, Hyunjong, et al.
Veröffentlicht: (2025)
von: Kim, Hyunjong, et al.
Veröffentlicht: (2025)
KoACD: The First Korean Adolescent Dataset for Cognitive Distortion Analysis via Role-Switching Multi-LLM Negotiation
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
von: Kim, JunSeo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SCRIPT: A Subcharacter Compositional Representation Injection Module for Korean Pre-Trained Language Models
von: Kim, SungHo, et al.
Veröffentlicht: (2026) -
Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean
von: Kim, SungHo, et al.
Veröffentlicht: (2025) -
Incorporating Domain Knowledge into Materials Tokenization
von: Oh, Yerim, et al.
Veröffentlicht: (2025) -
Zero-shot Commonsense Reasoning over Machine Imagination
von: Park, Hyuntae, et al.
Veröffentlicht: (2024) -
MELT: Materials-aware Continued Pre-training for Language Model Adaptation to Materials Science
von: Kim, Junho, et al.
Veröffentlicht: (2024)