Learning to Select In-Context Demonstration Preferred by Large Language Model
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Zheng, Lan, Shaocheng, Song, Lei, Bian, Jiang, Li, Yexin, Ren, Kan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Large Language Models Explore by Latent Distilling
par: Zeng, Yuanhao, et autres
Publié: (2026)
par: Zeng, Yuanhao, et autres
Publié: (2026)
MLCopilot: Unleashing the Power of Large Language Models in Solving Machine Learning Tasks
par: Zhang, Lei, et autres
Publié: (2023)
par: Zhang, Lei, et autres
Publié: (2023)
Holdout-Loss-Based Data Selection for LLM Finetuning via In-Context Learning
par: Zhang, Ling, et autres
Publié: (2025)
par: Zhang, Ling, et autres
Publié: (2025)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
par: Xu, Qiushui, et autres
Publié: (2025)
par: Xu, Qiushui, et autres
Publié: (2025)
CAE: Repurposing the Critic as an Explorer in Deep Reinforcement Learning
par: Li, Yexin
Publié: (2025)
par: Li, Yexin
Publié: (2025)
Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model
par: Tu, Songjun, et autres
Publié: (2024)
par: Tu, Songjun, et autres
Publié: (2024)
Linear-Time Demonstration Selection for In-Context Learning via Gradient Estimation
par: Zhang, Ziniu, et autres
Publié: (2025)
par: Zhang, Ziniu, et autres
Publié: (2025)
Unraveling the Mechanics of Learning-Based Demonstration Selection for In-Context Learning
par: Liu, Hui, et autres
Publié: (2024)
par: Liu, Hui, et autres
Publié: (2024)
Are Large Language Models In-Context Graph Learners?
par: Li, Jintang, et autres
Publié: (2025)
par: Li, Jintang, et autres
Publié: (2025)
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning
par: Wang, Xinyi, et autres
Publié: (2023)
par: Wang, Xinyi, et autres
Publié: (2023)
Self-Rewarding PPO: Aligning Large Language Models with Demonstrations Only
par: Zhang, Qingru, et autres
Publié: (2025)
par: Zhang, Qingru, et autres
Publié: (2025)
In-Context Reward Adaptation for Robust Preference Modeling
par: Sun, Zhenyu, et autres
Publié: (2026)
par: Sun, Zhenyu, et autres
Publié: (2026)
Human Alignment of Large Language Models through Online Preference Optimisation
par: Calandriello, Daniele, et autres
Publié: (2024)
par: Calandriello, Daniele, et autres
Publié: (2024)
Active Preference Learning for Large Language Models
par: Muldrew, William, et autres
Publié: (2024)
par: Muldrew, William, et autres
Publié: (2024)
Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning
par: Shu, Dong, et autres
Publié: (2024)
par: Shu, Dong, et autres
Publié: (2024)
Enhancing Large Language Models for Time-Series Forecasting via Vector-Injected In-Context Learning
par: Zhang, Jianqi, et autres
Publié: (2026)
par: Zhang, Jianqi, et autres
Publié: (2026)
Process In-Context Learning: Enhancing Mathematical Reasoning via Dynamic Demonstration Insertion
par: Gao, Ang, et autres
Publié: (2026)
par: Gao, Ang, et autres
Publié: (2026)
Understanding the Dynamics of Demonstration Conflict in In-Context Learning
par: Jiao, Difan, et autres
Publié: (2026)
par: Jiao, Difan, et autres
Publié: (2026)
Context-Aware Assistant Selection for Improved Inference Acceleration with Large Language Models
par: Huang, Jerry, et autres
Publié: (2024)
par: Huang, Jerry, et autres
Publié: (2024)
Large Language Model-Powered Evolutionary Code Optimization on a Phylogenetic Tree
par: Zhao, Leyi, et autres
Publié: (2026)
par: Zhao, Leyi, et autres
Publié: (2026)
Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning
par: Wang, Xubin, et autres
Publié: (2026)
par: Wang, Xubin, et autres
Publié: (2026)
Demonstration of DB-GPT: Next Generation Data Interaction System Empowered by Large Language Models
par: Xue, Siqiao, et autres
Publié: (2024)
par: Xue, Siqiao, et autres
Publié: (2024)
LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations
par: Ruoss, Anian, et autres
Publié: (2024)
par: Ruoss, Anian, et autres
Publié: (2024)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
par: Sun, Hao, et autres
Publié: (2024)
par: Sun, Hao, et autres
Publié: (2024)
Pareto-Optimal Learning from Preferences with Hidden Context
par: Bahlous-Boldi, Ryan, et autres
Publié: (2024)
par: Bahlous-Boldi, Ryan, et autres
Publié: (2024)
FastGAS: Fast Graph-based Annotation Selection for In-Context Learning
par: Chen, Zihan, et autres
Publié: (2024)
par: Chen, Zihan, et autres
Publié: (2024)
Uncovering Emergent Physics Representations Learned In-Context by Large Language Models
par: Song, Yeongwoo, et autres
Publié: (2025)
par: Song, Yeongwoo, et autres
Publié: (2025)
Improving Context-Aware Preference Modeling for Language Models
par: Pitis, Silviu, et autres
Publié: (2024)
par: Pitis, Silviu, et autres
Publié: (2024)
The Curse of Depth in Large Language Models
par: Sun, Wenfang, et autres
Publié: (2025)
par: Sun, Wenfang, et autres
Publié: (2025)
Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization
par: Nguyen, Thanh Thi, et autres
Publié: (2025)
par: Nguyen, Thanh Thi, et autres
Publié: (2025)
Position: Rethinking Post-Hoc Search-Based Neural Approaches for Solving Large-Scale Traveling Salesman Problems
par: Xia, Yifan, et autres
Publié: (2024)
par: Xia, Yifan, et autres
Publié: (2024)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
par: Yue, Sheng, et autres
Publié: (2024)
par: Yue, Sheng, et autres
Publié: (2024)
Vision-Language Model Selection and Reuse for Downstream Adaptation
par: Tan, Hao-Zhe, et autres
Publié: (2025)
par: Tan, Hao-Zhe, et autres
Publié: (2025)
Reinforcement Learning with Promising Tokens for Large Language Models
par: Pang, Jing-Cheng, et autres
Publié: (2026)
par: Pang, Jing-Cheng, et autres
Publié: (2026)
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations
par: Yang, Zhicheng, et autres
Publié: (2023)
par: Yang, Zhicheng, et autres
Publié: (2023)
When Correct Demonstrations Hurt: Rethinking the Role of Exemplars in In-Context Learning
par: Qiu, Chenghao, et autres
Publié: (2026)
par: Qiu, Chenghao, et autres
Publié: (2026)
A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions
par: Yu, Zhiyin, et autres
Publié: (2026)
par: Yu, Zhiyin, et autres
Publié: (2026)
Fusion of Multiscale Features Via Centralized Sparse-attention Network for EEG Decoding
par: Cai, Xiangrui, et autres
Publié: (2025)
par: Cai, Xiangrui, et autres
Publié: (2025)
AsynDBT: Asynchronous Distributed Bilevel Tuning for efficient In-Context Learning with Large Language Models
par: Ma, Hui, et autres
Publié: (2026)
par: Ma, Hui, et autres
Publié: (2026)
Large Language Models for Sequential Decision-Making: Improving In-Context Learning via Supervised Fine-Tuning
par: Zhang, Minmin, et autres
Publié: (2026)
par: Zhang, Minmin, et autres
Publié: (2026)
Documents similaires
-
Large Language Models Explore by Latent Distilling
par: Zeng, Yuanhao, et autres
Publié: (2026) -
MLCopilot: Unleashing the Power of Large Language Models in Solving Machine Learning Tasks
par: Zhang, Lei, et autres
Publié: (2023) -
Holdout-Loss-Based Data Selection for LLM Finetuning via In-Context Learning
par: Zhang, Ling, et autres
Publié: (2025) -
In-Context Compositional Q-Learning for Offline Reinforcement Learning
par: Xu, Qiushui, et autres
Publié: (2025) -
CAE: Repurposing the Critic as an Explorer in Deep Reinforcement Learning
par: Li, Yexin
Publié: (2025)