Call for Rigor in Reporting Quality of Instruction Tuning Data
Fuente:
arXiv
Salvato in:
| Autori principali: | Moon, Hyeonseok, Seo, Jaehyung, Lim, Heuiseok |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Impact of Negated Text on Hallucination with Large Language Models
di: Seo, Jaehyung, et al.
Pubblicazione: (2025)
di: Seo, Jaehyung, et al.
Pubblicazione: (2025)
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025)
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025)
Metric Calculating Benchmark: Code-Verifiable Complicate Instruction Following Benchmark for Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025)
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025)
Find the Intention of Instruction: Comprehensive Evaluation of Instruction Understanding for Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
di: Lee, Seungyoon, et al.
Pubblicazione: (2025)
di: Lee, Seungyoon, et al.
Pubblicazione: (2025)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
di: Park, Chanhee, et al.
Pubblicazione: (2025)
di: Park, Chanhee, et al.
Pubblicazione: (2025)
No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
di: Jung, Jimin, et al.
Pubblicazione: (2026)
di: Jung, Jimin, et al.
Pubblicazione: (2026)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
CoME: An Unlearning-based Approach to Conflict-free Model Editing
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
di: Jung, Dahyun, et al.
Pubblicazione: (2025)
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
Toward Practical Automatic Speech Recognition and Post-Processing: a Call for Explainable Error Benchmark Guideline
di: Koo, Seonmin, et al.
Pubblicazione: (2024)
di: Koo, Seonmin, et al.
Pubblicazione: (2024)
LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model
di: Jang, Youngjoon, et al.
Pubblicazione: (2026)
di: Jang, Youngjoon, et al.
Pubblicazione: (2026)
Post-hoc Utterance Refining Method by Entity Mining for Faithful Knowledge Grounded Conversations
di: Jang, Yoonna, et al.
Pubblicazione: (2024)
di: Jang, Yoonna, et al.
Pubblicazione: (2024)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses
di: Eo, Sugyeong, et al.
Pubblicazione: (2026)
di: Eo, Sugyeong, et al.
Pubblicazione: (2026)
ChatLang-8: An LLM-Based Synthetic Data Generation Framework for Grammatical Error Correction
di: Park, Jeiyoon, et al.
Pubblicazione: (2024)
di: Park, Jeiyoon, et al.
Pubblicazione: (2024)
CharacterGPT: A Persona Reconstruction Framework for Role-Playing Agents
di: Park, Jeiyoon, et al.
Pubblicazione: (2024)
di: Park, Jeiyoon, et al.
Pubblicazione: (2024)
Cross-Lingual Optimization for Language Transfer in Large Language Models
di: Lee, Jungseob, et al.
Pubblicazione: (2025)
di: Lee, Jungseob, et al.
Pubblicazione: (2025)
M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
di: Park, Jeiyoon, et al.
Pubblicazione: (2022)
di: Park, Jeiyoon, et al.
Pubblicazione: (2022)
EchoQA: A Large Collection of Instruction Tuning Data for Echocardiogram Reports
di: Moukheiber, Lama, et al.
Pubblicazione: (2025)
di: Moukheiber, Lama, et al.
Pubblicazione: (2025)
Alternative Speech: Complementary Method to Counter-Narrative for Better Discourse
di: Lee, Seungyoon, et al.
Pubblicazione: (2024)
di: Lee, Seungyoon, et al.
Pubblicazione: (2024)
SelectLLM: Can LLMs Select Important Instructions to Annotate?
di: Parkar, Ritik Sachin, et al.
Pubblicazione: (2024)
di: Parkar, Ritik Sachin, et al.
Pubblicazione: (2024)
Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation
di: Yoon, Jeongho, et al.
Pubblicazione: (2026)
di: Yoon, Jeongho, et al.
Pubblicazione: (2026)
Debate Only When Necessary: Adaptive Multiagent Collaboration for Efficient LLM Reasoning
di: Eo, Sugyeong, et al.
Pubblicazione: (2025)
di: Eo, Sugyeong, et al.
Pubblicazione: (2025)
Neuro-RIT: Neuron-Guided Instruction Tuning for Robust Retrieval-Augmented Language Model
di: Kim, Jaemin, et al.
Pubblicazione: (2026)
di: Kim, Jaemin, et al.
Pubblicazione: (2026)
Data Selection for Multi-turn Dialogue Instruction Tuning
di: Li, Bo, et al.
Pubblicazione: (2026)
di: Li, Bo, et al.
Pubblicazione: (2026)
Gap-K%: Measuring Top-1 Prediction Gap for Detecting Pretraining Data
di: Kwak, Minseo, et al.
Pubblicazione: (2026)
di: Kwak, Minseo, et al.
Pubblicazione: (2026)
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems
di: Jang, Youngjoon, et al.
Pubblicazione: (2025)
di: Jang, Youngjoon, et al.
Pubblicazione: (2025)
HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
Evidential Transformation Network: Turning Pretrained Models into Evidential Models for Post-hoc Uncertainty Estimation
di: Chun, Yongchan, et al.
Pubblicazione: (2026)
di: Chun, Yongchan, et al.
Pubblicazione: (2026)
Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
di: Lee, Hyunji, et al.
Pubblicazione: (2025)
di: Lee, Hyunji, et al.
Pubblicazione: (2025)
Instruction Tuning With Loss Over Instructions
di: Shi, Zhengyan, et al.
Pubblicazione: (2024)
di: Shi, Zhengyan, et al.
Pubblicazione: (2024)
EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context
di: Koo, Hamin, et al.
Pubblicazione: (2025)
di: Koo, Hamin, et al.
Pubblicazione: (2025)
TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA
di: Jung, Chanjoo, et al.
Pubblicazione: (2025)
di: Jung, Chanjoo, et al.
Pubblicazione: (2025)
Not All Documents Are What You Need for Extracting Instruction Tuning Data
di: Zhang, Chi, et al.
Pubblicazione: (2025)
di: Zhang, Chi, et al.
Pubblicazione: (2025)
Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning
di: Zhou, Hang, et al.
Pubblicazione: (2024)
di: Zhou, Hang, et al.
Pubblicazione: (2024)
Rethinking Table Instruction Tuning
di: Deng, Naihao, et al.
Pubblicazione: (2025)
di: Deng, Naihao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Impact of Negated Text on Hallucination with Large Language Models
di: Seo, Jaehyung, et al.
Pubblicazione: (2025) -
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025) -
Metric Calculating Benchmark: Code-Verifiable Complicate Instruction Following Benchmark for Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2025) -
Find the Intention of Instruction: Comprehensive Evaluation of Instruction Understanding for Large Language Models
di: Moon, Hyeonseok, et al.
Pubblicazione: (2024) -
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer
di: Lee, Seungyoon, et al.
Pubblicazione: (2025)