Translation of Multifaceted Data without Re-Training of Machine Translation Systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Moon, Hyeonseok, Lee, Seungyoon, Hong, Seongtae, Lee, Seungjun, Park, Chanjun, Lim, Heuiseok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
von: Lee, Seungyoon, et al.
Veröffentlicht: (2025)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2025)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
Alternative Speech: Complementary Method to Counter-Narrative for Better Discourse
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
Find the Intention of Instruction: Comprehensive Evaluation of Instruction Understanding for Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Cross-Lingual Optimization for Language Transfer in Large Language Models
von: Lee, Jungseob, et al.
Veröffentlicht: (2025)
von: Lee, Jungseob, et al.
Veröffentlicht: (2025)
Call for Rigor in Reporting Quality of Instruction Tuning Data
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
ChatLang-8: An LLM-Based Synthetic Data Generation Framework for Grammatical Error Correction
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
CharacterGPT: A Persona Reconstruction Framework for Role-Playing Agents
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
The Impact of Negated Text on Hallucination with Large Language Models
von: Seo, Jaehyung, et al.
Veröffentlicht: (2025)
von: Seo, Jaehyung, et al.
Veröffentlicht: (2025)
Metric Calculating Benchmark: Code-Verifiable Complicate Instruction Following Benchmark for Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
CoME: An Unlearning-based Approach to Conflict-free Model Editing
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
InstaTrans: An Instruction-Aware Translation Framework for Non-English Instruction Datasets
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Revise: A Framework for Revising OCRed text in Practical Information Systems with Data Contamination Strategy
von: Shim, Gyuho, et al.
Veröffentlicht: (2026)
von: Shim, Gyuho, et al.
Veröffentlicht: (2026)
Debate Only When Necessary: Adaptive Multiagent Collaboration for Efficient LLM Reasoning
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
CLEAR: Cross-Lingual Enhancement in Alignment via Reverse-training
von: Lee, Seungyoon, et al.
Veröffentlicht: (2026)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2026)
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
Post-hoc Utterance Refining Method by Entity Mining for Faithful Knowledge Grounded Conversations
von: Jang, Yoonna, et al.
Veröffentlicht: (2024)
von: Jang, Yoonna, et al.
Veröffentlicht: (2024)
LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
TORSO: Template-Oriented Reasoning Towards General Tasks
von: Kim, Minhyuk, et al.
Veröffentlicht: (2025)
von: Kim, Minhyuk, et al.
Veröffentlicht: (2025)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
Toward Practical Automatic Speech Recognition and Post-Processing: a Call for Explainable Error Benchmark Guideline
von: Koo, Seonmin, et al.
Veröffentlicht: (2024)
von: Koo, Seonmin, et al.
Veröffentlicht: (2024)
Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
Improving Korean-English Cross-Lingual Retrieval: A Data-Centric Study of Language Composition and Model Merging
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Efficient Technical Term Translation: A Knowledge Distillation Approach for Parenthetical Terminology Translation
von: Myung, Jiyoon, et al.
Veröffentlicht: (2024)
von: Myung, Jiyoon, et al.
Veröffentlicht: (2024)
LP Data Pipeline: Lightweight, Purpose-driven Data Pipeline for Large Language Models
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
Understanding LLM Development Through Longitudinal Study: Insights from the Open Ko-LLM Leaderboard
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
QE-EBM: Using Quality Estimators as Energy Loss for Machine Translation
von: Yoo, Gahyun, et al.
Veröffentlicht: (2024)
von: Yoo, Gahyun, et al.
Veröffentlicht: (2024)
Syntriever: How to Train Your Retriever with Synthetic Data from LLMs
von: Kim, Minsang, et al.
Veröffentlicht: (2025)
von: Kim, Minsang, et al.
Veröffentlicht: (2025)
CANTONMT: Investigating Back-Translation and Model-Switch Mechanisms for Cantonese-English Neural Machine Translation
von: Hong, Kung Yin, et al.
Veröffentlicht: (2024)
von: Hong, Kung Yin, et al.
Veröffentlicht: (2024)
RL from Teacher-Model Refinement: Gradual Imitation Learning for Machine Translation
von: Lee, Dongyub Jude, et al.
Veröffentlicht: (2025)
von: Lee, Dongyub Jude, et al.
Veröffentlicht: (2025)
Who speaks like a style of Vitamin: Towards Syntax-Aware DialogueSummarization using Multi-task Learning
von: Lee, Seolhwa, et al.
Veröffentlicht: (2021)
von: Lee, Seolhwa, et al.
Veröffentlicht: (2021)
Beyond Hard Negatives: The Importance of Score Distribution in Knowledge Distillation for Dense Retrieval
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
MLAIRE: Multilingual Language-Aware Information Retrieval Evaluation Protocal
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
Representing the Under-Represented: Cultural and Core Capability Benchmarks for Developing Thai Large Language Models
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
von: Kim, Dahyun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
von: Lee, Seungyoon, et al.
Veröffentlicht: (2025) -
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
von: Jung, Dahyun, et al.
Veröffentlicht: (2025) -
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025) -
Alternative Speech: Complementary Method to Counter-Narrative for Better Discourse
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024) -
Find the Intention of Instruction: Comprehensive Evaluation of Instruction Understanding for Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)