Linear-Time Demonstration Selection for In-Context Learning via Gradient Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Ziniu, Zhang, Zhenshuo, Li, Dongyue, Wang, Lu, Dy, Jennifer, Zhang, Hongyang R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Efficient Estimation of Kernel Surrogate Models for Task Attribution
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2026)
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2026)
Scalable Multi-Objective and Meta Reinforcement Learning via Gradient Estimation
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2025)
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2025)
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
di: Li, Dongyue, et al.
Pubblicazione: (2025)
di: Li, Dongyue, et al.
Pubblicazione: (2025)
Scalable Multitask Learning Using Gradient-based Estimation of Task Affinity
di: Li, Dongyue, et al.
Pubblicazione: (2024)
di: Li, Dongyue, et al.
Pubblicazione: (2024)
Scalable Fine-tuning from Multiple Data Sources: A First-Order Approximation Approach
di: Li, Dongyue, et al.
Pubblicazione: (2024)
di: Li, Dongyue, et al.
Pubblicazione: (2024)
Efficient Ensemble for Fine-tuning Language Models on Multiple Datasets
di: Li, Dongyue, et al.
Pubblicazione: (2025)
di: Li, Dongyue, et al.
Pubblicazione: (2025)
Unraveling the Mechanics of Learning-Based Demonstration Selection for In-Context Learning
di: Liu, Hui, et al.
Pubblicazione: (2024)
di: Liu, Hui, et al.
Pubblicazione: (2024)
Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning
di: Wang, Xubin, et al.
Pubblicazione: (2026)
di: Wang, Xubin, et al.
Pubblicazione: (2026)
Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
DemoShapley: Valuation of Demonstrations for In-Context Learning
di: Xie, Shan, et al.
Pubblicazione: (2024)
di: Xie, Shan, et al.
Pubblicazione: (2024)
Enhancing In-Context Learning via Implicit Demonstration Augmentation
di: Zhou, Xiaoling, et al.
Pubblicazione: (2024)
di: Zhou, Xiaoling, et al.
Pubblicazione: (2024)
A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms
di: Li, Yingru, et al.
Pubblicazione: (2025)
di: Li, Yingru, et al.
Pubblicazione: (2025)
Learning to Select In-Context Demonstration Preferred by Large Language Model
di: Zhang, Zheng, et al.
Pubblicazione: (2025)
di: Zhang, Zheng, et al.
Pubblicazione: (2025)
On the Robustness of Transformers against Context Hijacking for Linear Classification
di: Li, Tianle, et al.
Pubblicazione: (2025)
di: Li, Tianle, et al.
Pubblicazione: (2025)
TAGCOS: Task-agnostic Gradient Clustered Coreset Selection for Instruction Tuning Data
di: Zhang, Jipeng, et al.
Pubblicazione: (2024)
di: Zhang, Jipeng, et al.
Pubblicazione: (2024)
Process In-Context Learning: Enhancing Mathematical Reasoning via Dynamic Demonstration Insertion
di: Gao, Ang, et al.
Pubblicazione: (2026)
di: Gao, Ang, et al.
Pubblicazione: (2026)
Large Language Models Are Latent Variable Models: Explaining and Finding Good Demonstrations for In-Context Learning
di: Wang, Xinyi, et al.
Pubblicazione: (2023)
di: Wang, Xinyi, et al.
Pubblicazione: (2023)
Influential Language Data Selection via Gradient Trajectory Pursuit
di: Deng, Zhiwei, et al.
Pubblicazione: (2024)
di: Deng, Zhiwei, et al.
Pubblicazione: (2024)
MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
di: MiniCPM Team, et al.
Pubblicazione: (2026)
di: MiniCPM Team, et al.
Pubblicazione: (2026)
SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models
di: Wang, Xiaoxuan, et al.
Pubblicazione: (2023)
di: Wang, Xiaoxuan, et al.
Pubblicazione: (2023)
From Introspection to Best Practices: Principled Analysis of Demonstrations in Multimodal In-Context Learning
di: Xu, Nan, et al.
Pubblicazione: (2024)
di: Xu, Nan, et al.
Pubblicazione: (2024)
Multi-Layer Transformers Gradient Can be Approximated in Almost Linear Time
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
di: Liang, Yingyu, et al.
Pubblicazione: (2024)
Do pretrained Transformers Learn In-Context by Gradient Descent?
di: Shen, Lingfeng, et al.
Pubblicazione: (2023)
di: Shen, Lingfeng, et al.
Pubblicazione: (2023)
LEC: Linear Expectation Constraints for Selection-Conditioned Risk Control in Selective Prediction and Routing Systems
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
di: Li, Ziniu, et al.
Pubblicazione: (2025)
di: Li, Ziniu, et al.
Pubblicazione: (2025)
Affinity and Diversity: A Unified Metric for Demonstration Selection via Internal Representations
di: Kato, Mariko, et al.
Pubblicazione: (2025)
di: Kato, Mariko, et al.
Pubblicazione: (2025)
IntPro: A Proxy Agent for Context-Aware Intent Understanding via Retrieval-conditioned Inference
di: Liu, Guanming, et al.
Pubblicazione: (2026)
di: Liu, Guanming, et al.
Pubblicazione: (2026)
Hybrid DQN-TD3 Reinforcement Learning for Autonomous Navigation in Dynamic Environments
di: He, Xiaoyi, et al.
Pubblicazione: (2025)
di: He, Xiaoyi, et al.
Pubblicazione: (2025)
Self-Rewarding PPO: Aligning Large Language Models with Demonstrations Only
di: Zhang, Qingru, et al.
Pubblicazione: (2025)
di: Zhang, Qingru, et al.
Pubblicazione: (2025)
MarginSel : Max-Margin Demonstration Selection for LLMs
di: Ambati, Rajeev Bhatt, et al.
Pubblicazione: (2025)
di: Ambati, Rajeev Bhatt, et al.
Pubblicazione: (2025)
Learn More, Forget Less: A Gradient-Aware Data Selection Approach for LLM
di: Liu, Yibai, et al.
Pubblicazione: (2025)
di: Liu, Yibai, et al.
Pubblicazione: (2025)
CoT-ICL Lab: A Synthetic Framework for Studying Chain-of-Thought Learning from In-Context Demonstrations
di: Kothapalli, Vignesh, et al.
Pubblicazione: (2025)
di: Kothapalli, Vignesh, et al.
Pubblicazione: (2025)
In-Context Learning with Iterative Demonstration Selection
di: Qin, Chengwei, et al.
Pubblicazione: (2023)
di: Qin, Chengwei, et al.
Pubblicazione: (2023)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
di: Wei, Zeming, et al.
Pubblicazione: (2023)
di: Wei, Zeming, et al.
Pubblicazione: (2023)
Demonstration Selection for In-Context Learning via Reinforcement Learning
di: Wang, Xubin, et al.
Pubblicazione: (2024)
di: Wang, Xubin, et al.
Pubblicazione: (2024)
Scaling In-Context Online Learning Capability of LLMs via Cross-Episode Meta-RL
di: Lin, Xiaofeng, et al.
Pubblicazione: (2026)
di: Lin, Xiaofeng, et al.
Pubblicazione: (2026)
CL4KGE: A Curriculum Learning Method for Knowledge Graph Embedding
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Language Control Diffusion: Efficiently Scaling through Space, Time, and Tasks
di: Zhang, Edwin, et al.
Pubblicazione: (2022)
di: Zhang, Edwin, et al.
Pubblicazione: (2022)
Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long Contexts
di: Chen, Yingfa, et al.
Pubblicazione: (2026)
di: Chen, Yingfa, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Efficient Estimation of Kernel Surrogate Models for Task Attribution
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2026) -
Scalable Multi-Objective and Meta Reinforcement Learning via Gradient Estimation
di: Zhang, Zhenshuo, et al.
Pubblicazione: (2025) -
Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning
di: Li, Dongyue, et al.
Pubblicazione: (2025) -
Scalable Multitask Learning Using Gradient-based Estimation of Task Affinity
di: Li, Dongyue, et al.
Pubblicazione: (2024) -
Scalable Fine-tuning from Multiple Data Sources: A First-Order Approximation Approach
di: Li, Dongyue, et al.
Pubblicazione: (2024)