Gespeichert in:
| Hauptverfasser: | Kim, Dongjun, Kim, Minhyuk, Chun, YongChan, Park, Chanjun, Lim, Heuiseok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2412.07113 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Automatic Term Extraction with Large Language Models via Syntactic Retrieval
von: Chun, Yongchan, et al.
Veröffentlicht: (2025)
von: Chun, Yongchan, et al.
Veröffentlicht: (2025)
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
LANGSAE EDITING: Improving Multilingual Information Retrieval via Post-hoc Language Identity Removal
von: Kim, Dongjun, et al.
Veröffentlicht: (2026)
von: Kim, Dongjun, et al.
Veröffentlicht: (2026)
ChatLang-8: An LLM-Based Synthetic Data Generation Framework for Grammatical Error Correction
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
CharacterGPT: A Persona Reconstruction Framework for Role-Playing Agents
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
Understanding LLM Development Through Longitudinal Study: Insights from the Open Ko-LLM Leaderboard
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
Who speaks like a style of Vitamin: Towards Syntax-Aware DialogueSummarization using Multi-task Learning
von: Lee, Seolhwa, et al.
Veröffentlicht: (2021)
von: Lee, Seolhwa, et al.
Veröffentlicht: (2021)
Can Code-Switched Texts Activate a Knowledge Switch in LLMs? A Case Study on English-Korean Code-Switching
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2024)
Toward Practical Automatic Speech Recognition and Post-Processing: a Call for Explainable Error Benchmark Guideline
von: Koo, Seonmin, et al.
Veröffentlicht: (2024)
von: Koo, Seonmin, et al.
Veröffentlicht: (2024)
Alternative Speech: Complementary Method to Counter-Narrative for Better Discourse
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
CoME: An Unlearning-based Approach to Conflict-free Model Editing
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
CLEAR: Cross-Lingual Enhancement in Alignment via Reverse-training
von: Lee, Seungyoon, et al.
Veröffentlicht: (2026)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2026)
InstaTrans: An Instruction-Aware Translation Framework for Non-English Instruction Datasets
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Metric Calculating Benchmark: Code-Verifiable Complicate Instruction Following Benchmark for Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
Open Ko-LLM Leaderboard2: Bridging Foundational and Practical Evaluation for Korean LLMs
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
TORSO: Template-Oriented Reasoning Towards General Tasks
von: Kim, Minhyuk, et al.
Veröffentlicht: (2025)
von: Kim, Minhyuk, et al.
Veröffentlicht: (2025)
Sensory-Aware Sequential Recommendation via Review-Distilled Representations
von: Yoon, Yeo Chan, et al.
Veröffentlicht: (2026)
von: Yoon, Yeo Chan, et al.
Veröffentlicht: (2026)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
Mind the Blind Spots: A Focus-Level Evaluation Framework for LLM Reviews
von: Shin, Hyungyu, et al.
Veröffentlicht: (2025)
von: Shin, Hyungyu, et al.
Veröffentlicht: (2025)
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
Evalverse: Unified and Accessible Library for Large Language Model Evaluation
von: Kim, Jihoo, et al.
Veröffentlicht: (2024)
von: Kim, Jihoo, et al.
Veröffentlicht: (2024)
Representing the Under-Represented: Cultural and Core Capability Benchmarks for Developing Thai Large Language Models
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
Don't Judge Code by Its Cover: Exploring Biases in LLM Judges for Code Evaluation
von: Moon, Jiwon, et al.
Veröffentlicht: (2025)
von: Moon, Jiwon, et al.
Veröffentlicht: (2025)
Dataverse: Open-Source ETL (Extract, Transform, Load) Pipeline for Large Language Models
von: Park, Hyunbyung, et al.
Veröffentlicht: (2024)
von: Park, Hyunbyung, et al.
Veröffentlicht: (2024)
MCS-SQL: Leveraging Multiple Prompts and Multiple-Choice Selection For Text-to-SQL Generation
von: Lee, Dongjun, et al.
Veröffentlicht: (2024)
von: Lee, Dongjun, et al.
Veröffentlicht: (2024)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
sDPO: Don't Use Your Data All at Once
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
SAAS: Solving Ability Amplification Strategy for Enhanced Mathematical Reasoning in Large Language Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
Assessing the Answerability of Queries in Retrieval-Augmented Code Generation
von: Kim, Geonmin, et al.
Veröffentlicht: (2024)
von: Kim, Geonmin, et al.
Veröffentlicht: (2024)
LP Data Pipeline: Lightweight, Purpose-driven Data Pipeline for Large Language Models
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Rethinking KenLM: Good and Bad Model Ensembles for Efficient Text Quality Filtering in Large Web Corpora
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Enhancing Automatic Term Extraction with Large Language Models via Syntactic Retrieval
von: Chun, Yongchan, et al.
Veröffentlicht: (2025) -
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
von: Kim, Dongjun, et al.
Veröffentlicht: (2025) -
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025) -
LANGSAE EDITING: Improving Multilingual Information Retrieval via Post-hoc Language Identity Removal
von: Kim, Dongjun, et al.
Veröffentlicht: (2026) -
ChatLang-8: An LLM-Based Synthetic Data Generation Framework for Grammatical Error Correction
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)