Model Performance-Guided Evaluation Data Selection for Effective Prompt Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Ximing, Wang, Shaowei, Lin, Dayi, Hassan, Ahmed E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Framework for Real-time Safeguarding the Text Generation of Large Language Model
by: Dong, Ximing, et al.
Published: (2024)
by: Dong, Ximing, et al.
Published: (2024)
PromptExp: Multi-granularity Prompt Explanation of Large Language Models
by: Dong, Ximing, et al.
Published: (2024)
by: Dong, Ximing, et al.
Published: (2024)
Beyond Tokens: Semantic-Aware Speculative Decoding for Efficient Inference by Probing Internal States
by: Dong, Ximing, et al.
Published: (2026)
by: Dong, Ximing, et al.
Published: (2026)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
by: Ravichander, Abhilasha, et al.
Published: (2025)
by: Ravichander, Abhilasha, et al.
Published: (2025)
Submodular Evaluation Subset Selection in Automatic Prompt Optimization
by: Nian, Jinming, et al.
Published: (2026)
by: Nian, Jinming, et al.
Published: (2026)
Evaluating the Effectiveness of Black-Box Prompt Optimization as the Scale of LLMs Continues to Grow
by: Zhou, Ziyu, et al.
Published: (2025)
by: Zhou, Ziyu, et al.
Published: (2025)
When Elo Lies: Hidden Biases in Codeforces-Based Evaluation of Large Language Models
by: Zheng, Shenyu, et al.
Published: (2026)
by: Zheng, Shenyu, et al.
Published: (2026)
Language Model Meets Prototypes: Towards Interpretable Text Classification Models through Prototypical Networks
by: Wen, Ximing
Published: (2024)
by: Wen, Ximing
Published: (2024)
Efficient and Accurate Prompt Optimization: the Benefit of Memory in Exemplar-Guided Reflection
by: Yan, Cilin, et al.
Published: (2024)
by: Yan, Cilin, et al.
Published: (2024)
Beyond One-Size-Fits-All: Inversion Learning for Highly Effective NLG Evaluation Prompts
by: Hong, Hanhua, et al.
Published: (2025)
by: Hong, Hanhua, et al.
Published: (2025)
Prompt Space Optimizing Few-shot Reasoning Success with Large Language Models
by: Shi, Fobo, et al.
Published: (2023)
by: Shi, Fobo, et al.
Published: (2023)
Evaluating the Effectiveness of the Foundational Models for Q&A Classification in Mental Health care
by: Alhuzali, Hassan, et al.
Published: (2024)
by: Alhuzali, Hassan, et al.
Published: (2024)
A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection
by: Wen, Ximing, et al.
Published: (2025)
by: Wen, Ximing, et al.
Published: (2025)
From Quantity to Quality: Boosting LLM Performance with Self-Guided Data Selection for Instruction Tuning
by: Li, Ming, et al.
Published: (2023)
by: Li, Ming, et al.
Published: (2023)
Data Selection via Optimal Control for Language Models
by: Gu, Yuxian, et al.
Published: (2024)
by: Gu, Yuxian, et al.
Published: (2024)
Effective In-Context Example Selection through Data Compression
by: Sun, Zhongxiang, et al.
Published: (2024)
by: Sun, Zhongxiang, et al.
Published: (2024)
Effectiveness of Prompt Optimization in NL2SQL Systems
by: Gurajada, Sairam, et al.
Published: (2025)
by: Gurajada, Sairam, et al.
Published: (2025)
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
by: Fu, Jinlan, et al.
Published: (2024)
by: Fu, Jinlan, et al.
Published: (2024)
Language Model Prompt Selection via Simulation Optimization
by: Zhang, Haoting, et al.
Published: (2024)
by: Zhang, Haoting, et al.
Published: (2024)
Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
by: Cheng, Jiale, et al.
Published: (2023)
by: Cheng, Jiale, et al.
Published: (2023)
Can Prompts Rewind Time for LLMs? Evaluating the Effectiveness of Prompted Knowledge Cutoffs
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
Creating Arabic LLM Prompts at Scale
by: El-Sheikh, Abdelrahman, et al.
Published: (2024)
by: El-Sheikh, Abdelrahman, et al.
Published: (2024)
Structure Guided Prompt: Instructing Large Language Model in Multi-Step Reasoning by Exploring Graph Structure of the Text
by: Cheng, Kewei, et al.
Published: (2024)
by: Cheng, Kewei, et al.
Published: (2024)
PEML: Parameter-efficient Multi-Task Learning with Optimized Continuous Prompts
by: Chowdhury, Anjir Ahmed, et al.
Published: (2026)
by: Chowdhury, Anjir Ahmed, et al.
Published: (2026)
On the Step Length Confounding in LLM Reasoning Data Selection
by: Wang, Bing, et al.
Published: (2026)
by: Wang, Bing, et al.
Published: (2026)
Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads
by: He, Xingyang, et al.
Published: (2025)
by: He, Xingyang, et al.
Published: (2025)
Error Taxonomy-Guided Prompt Optimization
by: Singh, Mayank, et al.
Published: (2026)
by: Singh, Mayank, et al.
Published: (2026)
MAPO: Boosting Large Language Model Performance with Model-Adaptive Prompt Optimization
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization
by: Liu, Yuanye, et al.
Published: (2025)
by: Liu, Yuanye, et al.
Published: (2025)
Beyond Words: Evaluating Large Language Models in Transportation Planning
by: Ying, Shaowei, et al.
Published: (2024)
by: Ying, Shaowei, et al.
Published: (2024)
Improving Probability-based Prompt Selection Through Unified Evaluation and Analysis
by: Yang, Sohee, et al.
Published: (2023)
by: Yang, Sohee, et al.
Published: (2023)
The Art of Asking: Multilingual Prompt Optimization for Synthetic Data
by: Mora, David, et al.
Published: (2025)
by: Mora, David, et al.
Published: (2025)
Stability-Aware Prompt Optimization for Clinical Data Abstraction
by: Kolbeinsson, Arinbjörn, et al.
Published: (2026)
by: Kolbeinsson, Arinbjörn, et al.
Published: (2026)
An Evaluation of Large Language Models on Text Summarization Tasks Using Prompt Engineering Techniques
by: Aly, Walid Mohamed, et al.
Published: (2025)
by: Aly, Walid Mohamed, et al.
Published: (2025)
Do Physicians Know How to Prompt? The Need for Automatic Prompt Optimization Help in Clinical Note Generation
by: Yao, Zonghai, et al.
Published: (2023)
by: Yao, Zonghai, et al.
Published: (2023)
Targeted Efficient Fine-tuning: Optimizing Parameter Updates with Data-Driven Sample Selection
by: Dong, Ming, et al.
Published: (2024)
by: Dong, Ming, et al.
Published: (2024)
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization
by: Abrar, Ajwad, et al.
Published: (2025)
by: Abrar, Ajwad, et al.
Published: (2025)
Data Swarms: Optimizable Generation of Synthetic Evaluation Data
by: Feng, Shangbin, et al.
Published: (2025)
by: Feng, Shangbin, et al.
Published: (2025)
Benchmarking Large Language Model Uncertainty for Prompt Optimization
by: Guo, Pei-Fu, et al.
Published: (2024)
by: Guo, Pei-Fu, et al.
Published: (2024)
Influence Guided Context Selection for Effective Retrieval-Augmented Generation
by: Deng, Jiale, et al.
Published: (2025)
by: Deng, Jiale, et al.
Published: (2025)
Similar Items
-
A Framework for Real-time Safeguarding the Text Generation of Large Language Model
by: Dong, Ximing, et al.
Published: (2024) -
PromptExp: Multi-granularity Prompt Explanation of Large Language Models
by: Dong, Ximing, et al.
Published: (2024) -
Beyond Tokens: Semantic-Aware Speculative Decoding for Efficient Inference by Probing Internal States
by: Dong, Ximing, et al.
Published: (2026) -
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
by: Ravichander, Abhilasha, et al.
Published: (2025) -
Submodular Evaluation Subset Selection in Automatic Prompt Optimization
by: Nian, Jinming, et al.
Published: (2026)