Diversify and Conquer: Diversity-Centric Data Selection with Iterative Refinement
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Simon, Chen, Liangyu, Ahmadian, Sara, Fadaee, Marzieh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
by: Aakanksha, et al.
Published: (2024)
by: Aakanksha, et al.
Published: (2024)
A Post-trainer's Guide to Multilingual Training Data: Uncovering Cross-lingual Transfer Dynamics
by: Shimabucoro, Luisa, et al.
Published: (2025)
by: Shimabucoro, Luisa, et al.
Published: (2025)
LLM See, LLM Do: Guiding Data Generation to Target Non-Differentiable Objectives
by: Shimabucoro, Luísa, et al.
Published: (2024)
by: Shimabucoro, Luísa, et al.
Published: (2024)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
by: Rajaee, Sara, et al.
Published: (2026)
by: Rajaee, Sara, et al.
Published: (2026)
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
by: Aakanksha, et al.
Published: (2024)
by: Aakanksha, et al.
Published: (2024)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
by: Kreutzer, Julia, et al.
Published: (2025)
by: Kreutzer, Julia, et al.
Published: (2025)
SimMerge: Learning to Select Merge Operators from Similarity Signals
by: Bolton, Oliver, et al.
Published: (2026)
by: Bolton, Oliver, et al.
Published: (2026)
Verification Limits Code LLM Training
by: Gureja, Srishti, et al.
Published: (2025)
by: Gureja, Srishti, et al.
Published: (2025)
Iterative Translation Refinement with Large Language Models
by: Chen, Pinzhen, et al.
Published: (2023)
by: Chen, Pinzhen, et al.
Published: (2023)
NeoBabel: A Multilingual Open Tower for Visual Generation
by: Derakhshani, Mohammad Mahdi, et al.
Published: (2025)
by: Derakhshani, Mohammad Mahdi, et al.
Published: (2025)
Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
by: Tsai, Yu-Che, et al.
Published: (2025)
by: Tsai, Yu-Che, et al.
Published: (2025)
Spontaneous Reward Hacking in Iterative Self-Refinement
by: Pan, Jane, et al.
Published: (2024)
by: Pan, Jane, et al.
Published: (2024)
In-Context Learning with Iterative Demonstration Selection
by: Qin, Chengwei, et al.
Published: (2023)
by: Qin, Chengwei, et al.
Published: (2023)
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation
by: Zhou, Changzhi, et al.
Published: (2025)
by: Zhou, Changzhi, et al.
Published: (2025)
Piece of Table: A Divide-and-Conquer Approach for Selecting Subtables in Table Question Answering
by: Lee, Wonjin, et al.
Published: (2024)
by: Lee, Wonjin, et al.
Published: (2024)
Boosting LLM via Learning from Data Iteratively and Selectively
by: Jia, Qi, et al.
Published: (2024)
by: Jia, Qi, et al.
Published: (2024)
AIR: Complex Instruction Generation via Automatic Iterative Refinement
by: Liu, Wei, et al.
Published: (2025)
by: Liu, Wei, et al.
Published: (2025)
Evolutionary Guided Decoding: Iterative Value Refinement for LLMs
by: Liu, Zhenhua, et al.
Published: (2025)
by: Liu, Zhenhua, et al.
Published: (2025)
Diversifying the Mixture-of-Experts Representation for Language Models with Orthogonal Optimizer
by: Liu, Boan, et al.
Published: (2023)
by: Liu, Boan, et al.
Published: (2023)
Iterative Critique-Refine Framework for Enhancing LLM Personalization
by: Maram, Durga Prasad, et al.
Published: (2025)
by: Maram, Durga Prasad, et al.
Published: (2025)
SEG:Seeds-Enhanced Iterative Refinement Graph Neural Network for Entity Alignment
by: Ai, Wei, et al.
Published: (2024)
by: Ai, Wei, et al.
Published: (2024)
FLAIRR-TS -- Forecasting LLM-Agents with Iterative Refinement and Retrieval for Time Series
by: Jalori, Gunjan, et al.
Published: (2025)
by: Jalori, Gunjan, et al.
Published: (2025)
SimpleStrat: Diversifying Language Model Generation with Stratification
by: Wong, Justin, et al.
Published: (2024)
by: Wong, Justin, et al.
Published: (2024)
Diversifying Question Generation over Knowledge Base via External Natural Questions
by: Guo, Shasha, et al.
Published: (2023)
by: Guo, Shasha, et al.
Published: (2023)
LLM driven Text-to-Table Generation through Sub-Tasks Guidance and Iterative Refinement
by: C, Rajmohan, et al.
Published: (2025)
by: C, Rajmohan, et al.
Published: (2025)
Iterative Prompt Refinement for Dyslexia-Friendly Text Summarization Using GPT-4o
by: Bhojwani, Samay, et al.
Published: (2026)
by: Bhojwani, Samay, et al.
Published: (2026)
STRIVE: A Think & Improve Approach with Iterative Refinement for Enhancing Question Quality Estimation
by: Deroy, Aniket, et al.
Published: (2025)
by: Deroy, Aniket, et al.
Published: (2025)
Nova: An Iterative Planning and Search Approach to Enhance Novelty and Diversity of LLM Generated Ideas
by: Hu, Xiang, et al.
Published: (2024)
by: Hu, Xiang, et al.
Published: (2024)
M-RewardBench: Evaluating Reward Models in Multilingual Settings
by: Gureja, Srishti, et al.
Published: (2024)
by: Gureja, Srishti, et al.
Published: (2024)
Iterative Experience Refinement of Software-Developing Agents
by: Qian, Chen, et al.
Published: (2024)
by: Qian, Chen, et al.
Published: (2024)
SCIR: A Self-Correcting Iterative Refinement Framework for Enhanced Information Extraction Based on Schema
by: Fang, Yushen, et al.
Published: (2025)
by: Fang, Yushen, et al.
Published: (2025)
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement
by: Ye, Liqin, et al.
Published: (2025)
by: Ye, Liqin, et al.
Published: (2025)
Learning from the Best, Differently: A Diversity-Driven Rethinking on Data Selection
by: He, Hongyi, et al.
Published: (2025)
by: He, Hongyi, et al.
Published: (2025)
Enhancing LLM Character-Level Manipulation via Divide and Conquer
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment
by: Wang, Zekun Moore, et al.
Published: (2024)
by: Wang, Zekun Moore, et al.
Published: (2024)
WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems
by: Huerta, Juan M.
Published: (2026)
by: Huerta, Juan M.
Published: (2026)
Merge and Conquer: Instructing Multilingual Models by Adding Target Language Weights
by: Valero, Eneko, et al.
Published: (2026)
by: Valero, Eneko, et al.
Published: (2026)
Shifting AI Efficiency From Model-Centric to Data-Centric Compression
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
Self-DC: When to Reason and When to Act? Self Divide-and-Conquer for Compositional Unknown Questions
by: Wang, Hongru, et al.
Published: (2024)
by: Wang, Hongru, et al.
Published: (2024)
Guiding and Diversifying LLM-Based Story Generation via Answer Set Programming
by: Wang, Phoebe J., et al.
Published: (2024)
by: Wang, Phoebe J., et al.
Published: (2024)
Similar Items
-
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
by: Aakanksha, et al.
Published: (2024) -
A Post-trainer's Guide to Multilingual Training Data: Uncovering Cross-lingual Transfer Dynamics
by: Shimabucoro, Luisa, et al.
Published: (2025) -
LLM See, LLM Do: Guiding Data Generation to Target Non-Differentiable Objectives
by: Shimabucoro, Luísa, et al.
Published: (2024) -
Unlocking Reasoning Capability on Machine Translation in Large Language Models
by: Rajaee, Sara, et al.
Published: (2026) -
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
by: Aakanksha, et al.
Published: (2024)