Prompt replay: speeding up grpo with on-policy reuse of high-signal prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Baroian, Andrei, Berger, Rutger |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Supervised Fine-Tuning or In-Context Learning? Evaluating LLMs for Clinical NER
by: Baroian, Andrei
Published: (2025)
by: Baroian, Andrei
Published: (2025)
Crown, Frame, Reverse: Layer-Wise Scaling Variants for LLM Pre-Training
by: Baroian, Andrei, et al.
Published: (2025)
by: Baroian, Andrei, et al.
Published: (2025)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
by: Dong, Harry, et al.
Published: (2024)
by: Dong, Harry, et al.
Published: (2024)
An empirical study of task and feature correlations in the reuse of pre-trained models
by: Mohamud, Jama Hussein, et al.
Published: (2025)
by: Mohamud, Jama Hussein, et al.
Published: (2025)
Intent-based Prompt Calibration: Enhancing prompt optimization with synthetic boundary cases
by: Levi, Elad, et al.
Published: (2024)
by: Levi, Elad, et al.
Published: (2024)
Pre-Forgettable Models: Prompt Learning as a Native Mechanism for Unlearning
by: Hendrix, Rutger, et al.
Published: (2025)
by: Hendrix, Rutger, et al.
Published: (2025)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
FoGE: Fock Space inspired encoding for graph prompting
by: Chytas, Sotirios Panagiotis, et al.
Published: (2025)
by: Chytas, Sotirios Panagiotis, et al.
Published: (2025)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
by: Sclar, Melanie, et al.
Published: (2023)
by: Sclar, Melanie, et al.
Published: (2023)
Efficient multi-prompt evaluation of LLMs
by: Polo, Felipe Maia, et al.
Published: (2024)
by: Polo, Felipe Maia, et al.
Published: (2024)
It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt
by: Bladon, Stuart, et al.
Published: (2026)
by: Bladon, Stuart, et al.
Published: (2026)
P2DT: Mitigating Forgetting in task-incremental Learning with progressive prompt Decision Transformer
by: Wang, Zhiyuan, et al.
Published: (2024)
by: Wang, Zhiyuan, et al.
Published: (2024)
Brittlebench: Quantifying LLM robustness via prompt sensitivity
by: Romanou, Angelika, et al.
Published: (2026)
by: Romanou, Angelika, et al.
Published: (2026)
Evil twins are not that evil: Qualitative insights into machine-generated prompts
by: Rakotonirina, Nathanaël Carraz, et al.
Published: (2024)
by: Rakotonirina, Nathanaël Carraz, et al.
Published: (2024)
Soft-prompt Tuning for Large Language Models to Evaluate Bias
by: Tian, Jacob-Junqi, et al.
Published: (2023)
by: Tian, Jacob-Junqi, et al.
Published: (2023)
Structured prompt interrogation and recursive extraction of semantics (SPIRES): A method for populating knowledge bases using zero-shot learning
by: Caufield, J. Harry, et al.
Published: (2023)
by: Caufield, J. Harry, et al.
Published: (2023)
Constructive Symbolic Reinforcement Learning via Intuitionistic Logic and Goal-Chaining Inference
by: Patrascu, Andrei T.
Published: (2025)
by: Patrascu, Andrei T.
Published: (2025)
Expected Possession Value of Control and Duel Actions for Soccer Player's Skills Estimation
by: Shelopugin, Andrei
Published: (2024)
by: Shelopugin, Andrei
Published: (2024)
Proximal Policy Optimization with Adaptive Exploration
by: Lixandru, Andrei
Published: (2024)
by: Lixandru, Andrei
Published: (2024)
(GG) MoE vs. MLP on Tabular Data
by: Chernov, Andrei
Published: (2025)
by: Chernov, Andrei
Published: (2025)
GreenTEA: Gradient Descent with Topic-modeling and Evolutionary Auto-prompting
by: Dong, Zheng, et al.
Published: (2025)
by: Dong, Zheng, et al.
Published: (2025)
StateAct: Enhancing LLM Base Agents via Self-prompting and State-tracking
by: Rozanov, Nikolai, et al.
Published: (2024)
by: Rozanov, Nikolai, et al.
Published: (2024)
PromptAudit: Auditing Prompt Sensitivity in LLM-Based Vulnerability Detection
by: Camarato, Steffen J., et al.
Published: (2026)
by: Camarato, Steffen J., et al.
Published: (2026)
Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
by: Wu, Zhaomin, et al.
Published: (2025)
by: Wu, Zhaomin, et al.
Published: (2025)
Automating Computational Design with Generative AI
by: Ploennigs, Joern, et al.
Published: (2023)
by: Ploennigs, Joern, et al.
Published: (2023)
Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting
by: Wang, Zesen, et al.
Published: (2025)
by: Wang, Zesen, et al.
Published: (2025)
PromptWise: Online Learning for Cost-Aware Prompt Assignment in Generative Models
by: Hu, Xiaoyan, et al.
Published: (2025)
by: Hu, Xiaoyan, et al.
Published: (2025)
The Empirical Impact of Reducing Symmetries on the Performance of Deep Ensembles and MoE
by: Chernov, Andrei, et al.
Published: (2025)
by: Chernov, Andrei, et al.
Published: (2025)
Rubric-based On-policy Distillation
by: Fang, Junfeng, et al.
Published: (2026)
by: Fang, Junfeng, et al.
Published: (2026)
UniPrompt-CL: Sustainable Continual Learning in Medical AI with Unified Prompt Pools
by: Oh, Gyutae, et al.
Published: (2025)
by: Oh, Gyutae, et al.
Published: (2025)
Prompt Tuning Strikes Back: Customizing Foundation Models with Low-Rank Prompt Adaptation
by: Jain, Abhinav, et al.
Published: (2024)
by: Jain, Abhinav, et al.
Published: (2024)
Prompt Optimization with Human Feedback
by: Lin, Xiaoqiang, et al.
Published: (2024)
by: Lin, Xiaoqiang, et al.
Published: (2024)
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering
by: Agrawal, Aryan
Published: (2024)
by: Agrawal, Aryan
Published: (2024)
DAGPrompT: Pushing the Limits of Graph Prompting with a Distribution-aware Graph Prompt Tuning Approach
by: Chen, Qin, et al.
Published: (2025)
by: Chen, Qin, et al.
Published: (2025)
PromptTSS: A Prompting-Based Approach for Interactive Multi-Granularity Time Series Segmentation
by: Chang, Ching, et al.
Published: (2025)
by: Chang, Ching, et al.
Published: (2025)
ChordPrompt: Orchestrating Cross-Modal Prompt Synergy for Multi-Domain Incremental Learning in CLIP
by: Wang, Zhiyuan, et al.
Published: (2025)
by: Wang, Zhiyuan, et al.
Published: (2025)
How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks
by: Stupariu, Teodor-Mihai, et al.
Published: (2026)
by: Stupariu, Teodor-Mihai, et al.
Published: (2026)
Markov flow policy -- deep MC
by: Soffair, Nitsan, et al.
Published: (2024)
by: Soffair, Nitsan, et al.
Published: (2024)
Dual policy as self-model for planning
by: Yoo, Jaesung, et al.
Published: (2023)
by: Yoo, Jaesung, et al.
Published: (2023)
Semantic Geometry for policy-constrained interpretation
by: Phadke, Nikit
Published: (2025)
by: Phadke, Nikit
Published: (2025)
Similar Items
-
Supervised Fine-Tuning or In-Context Learning? Evaluating LLMs for Clinical NER
by: Baroian, Andrei
Published: (2025) -
Crown, Frame, Reverse: Layer-Wise Scaling Variants for LLM Pre-Training
by: Baroian, Andrei, et al.
Published: (2025) -
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
by: Dong, Harry, et al.
Published: (2024) -
An empirical study of task and feature correlations in the reuse of pre-trained models
by: Mohamud, Jama Hussein, et al.
Published: (2025) -
Intent-based Prompt Calibration: Enhancing prompt optimization with synthetic boundary cases
by: Levi, Elad, et al.
Published: (2024)