Discrete Prompt Compression with Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Jung, Hoyoun, Kim, Kyung-Joong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
IPCGRL: Language-Instructed Reinforcement Learning for Procedural Level Generation
por: Baek, In-Chang, et al.
Publicado: (2025)
por: Baek, In-Chang, et al.
Publicado: (2025)
Reinforcement Learning for Optimizing RAG for Domain Chatbots
por: Kulkarni, Mandar, et al.
Publicado: (2024)
por: Kulkarni, Mandar, et al.
Publicado: (2024)
PRL: Prompts from Reinforcement Learning
por: Batorski, Paweł, et al.
Publicado: (2025)
por: Batorski, Paweł, et al.
Publicado: (2025)
MENTOR: A Reinforcement Learning Framework for Enabling Tool Use in Small Models via Teacher-Optimized Rewards
por: Choi, ChangSu, et al.
Publicado: (2025)
por: Choi, ChangSu, et al.
Publicado: (2025)
Conditional [MASK] Discrete Diffusion Language Model
por: Koh, Hyukhun, et al.
Publicado: (2024)
por: Koh, Hyukhun, et al.
Publicado: (2024)
LLMs Position Themselves as More Rational Than Humans: Emergence of AI Self-Awareness Measured Through Game Theory
por: Kim, Kyung-Hoon
Publicado: (2025)
por: Kim, Kyung-Hoon
Publicado: (2025)
ACoRN: Noise-Robust Abstractive Compression in Retrieval-Augmented Language Models
por: Kim, Singon, et al.
Publicado: (2025)
por: Kim, Singon, et al.
Publicado: (2025)
GRL-Prompt: Towards Knowledge Graph based Prompt Optimization via Reinforcement Learning
por: Liu, Yuze, et al.
Publicado: (2024)
por: Liu, Yuze, et al.
Publicado: (2024)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
por: Khodabandeh, Mahdi, et al.
Publicado: (2026)
por: Khodabandeh, Mahdi, et al.
Publicado: (2026)
Learning to Compress Prompt in Natural Language Formats
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
Parse Trees Guided LLM Prompt Compression
por: Mao, Wenhao, et al.
Publicado: (2024)
por: Mao, Wenhao, et al.
Publicado: (2024)
EFPC: Towards Efficient and Flexible Prompt Compression
por: Cao, Yun-Hao, et al.
Publicado: (2025)
por: Cao, Yun-Hao, et al.
Publicado: (2025)
ICPC: In-context Prompt Compression with Faster Inference
por: Yu, Ziyang, et al.
Publicado: (2025)
por: Yu, Ziyang, et al.
Publicado: (2025)
Efficient and Effective Prompt Tuning via Prompt Decomposition and Compressed Outer Product
por: Lan, Pengxiang, et al.
Publicado: (2025)
por: Lan, Pengxiang, et al.
Publicado: (2025)
Self-Compression of Chain-of-Thought via Multi-Agent Reinforcement Learning
por: Chen, Yiqun, et al.
Publicado: (2026)
por: Chen, Yiqun, et al.
Publicado: (2026)
SCOPE: A Generative Approach for LLM Prompt Compression
por: Zhang, Tinghui, et al.
Publicado: (2025)
por: Zhang, Tinghui, et al.
Publicado: (2025)
An Empirical Study on Prompt Compression for Large Language Models
por: Zhang, Zheng, et al.
Publicado: (2025)
por: Zhang, Zheng, et al.
Publicado: (2025)
PRewrite: Prompt Rewriting with Reinforcement Learning
por: Kong, Weize, et al.
Publicado: (2024)
por: Kong, Weize, et al.
Publicado: (2024)
Dynamic Compressing Prompts for Efficient Inference of Large Language Models
por: Hu, Jinwu, et al.
Publicado: (2025)
por: Hu, Jinwu, et al.
Publicado: (2025)
No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping
por: Le, Thanh-Long V., et al.
Publicado: (2025)
por: Le, Thanh-Long V., et al.
Publicado: (2025)
PIS: Linking Importance Sampling and Attention Mechanisms for Efficient Prompt Compression
por: Chen, Lizhe, et al.
Publicado: (2025)
por: Chen, Lizhe, et al.
Publicado: (2025)
Prior Prompt Engineering for Reinforcement Fine-Tuning
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
por: Chen, Zhuoen, et al.
Publicado: (2026)
por: Chen, Zhuoen, et al.
Publicado: (2026)
Prompt-SAW: Leveraging Relation-Aware Graphs for Textual Prompt Compression
por: Ali, Muhammad Asif, et al.
Publicado: (2024)
por: Ali, Muhammad Asif, et al.
Publicado: (2024)
Contextual Reinforcement in Multimodal Token Compression for Large Language Models
por: Piero, Naderdel, et al.
Publicado: (2025)
por: Piero, Naderdel, et al.
Publicado: (2025)
Harmful Prompt Laundering: Jailbreaking LLMs with Abductive Styles and Symbolic Encoding
por: Joo, Seongho, et al.
Publicado: (2025)
por: Joo, Seongho, et al.
Publicado: (2025)
PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions
por: Kyung, Daeun, et al.
Publicado: (2025)
por: Kyung, Daeun, et al.
Publicado: (2025)
Prompt Compression in Diffusion Large Language Models: Evaluating LLMLingua-2 on LLaDA
por: Huang, Sterling, et al.
Publicado: (2026)
por: Huang, Sterling, et al.
Publicado: (2026)
PMoE: Progressive Mixture of Experts with Asymmetric Transformer for Continual Learning
por: Jung, Min Jae, et al.
Publicado: (2024)
por: Jung, Min Jae, et al.
Publicado: (2024)
Natural Language Declarative Prompting (NLD-P): A Modular Governance Method for Prompt Design Under Model Drift
por: Kim, Hyunwoo, et al.
Publicado: (2026)
por: Kim, Hyunwoo, et al.
Publicado: (2026)
Noise-Robust Abstractive Compression in Retrieval-Augmented Language Models
por: Kim, Singon
Publicado: (2025)
por: Kim, Singon
Publicado: (2025)
Prism: Spectral Parameter Sharing for Multi-Agent Reinforcement Learning
por: Kim, Kyungbeom, et al.
Publicado: (2026)
por: Kim, Kyungbeom, et al.
Publicado: (2026)
Beyond Learning: A Training-Free Alternative to Model Adaptation
por: Yoon, Namkyung, et al.
Publicado: (2026)
por: Yoon, Namkyung, et al.
Publicado: (2026)
Reinforcement Learning for Chain of Thought Compression with One-Domain-to-All Generalization
por: Li, Hanyu, et al.
Publicado: (2025)
por: Li, Hanyu, et al.
Publicado: (2025)
ConsPrompt: Exploiting Contrastive Samples for Fewshot Prompt Learning
por: Weng, Jinta, et al.
Publicado: (2022)
por: Weng, Jinta, et al.
Publicado: (2022)
KatFishNet: Detecting LLM-Generated Korean Text through Linguistic Feature Analysis
por: Park, Shinwoo, et al.
Publicado: (2025)
por: Park, Shinwoo, et al.
Publicado: (2025)
Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog Generation
por: Park, Jinyoung, et al.
Publicado: (2024)
por: Park, Jinyoung, et al.
Publicado: (2024)
Contrastive and Consistency Learning for Neural Noisy-Channel Model in Spoken Language Understanding
por: Kim, Suyoung, et al.
Publicado: (2024)
por: Kim, Suyoung, et al.
Publicado: (2024)
Can Separators Improve Chain-of-Thought Prompting?
por: Park, Yoonjeong, et al.
Publicado: (2024)
por: Park, Yoonjeong, et al.
Publicado: (2024)
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification
por: Yun, Jungmin, et al.
Publicado: (2024)
por: Yun, Jungmin, et al.
Publicado: (2024)
Ejemplares similares
-
IPCGRL: Language-Instructed Reinforcement Learning for Procedural Level Generation
por: Baek, In-Chang, et al.
Publicado: (2025) -
Reinforcement Learning for Optimizing RAG for Domain Chatbots
por: Kulkarni, Mandar, et al.
Publicado: (2024) -
PRL: Prompts from Reinforcement Learning
por: Batorski, Paweł, et al.
Publicado: (2025) -
MENTOR: A Reinforcement Learning Framework for Enabling Tool Use in Small Models via Teacher-Optimized Rewards
por: Choi, ChangSu, et al.
Publicado: (2025) -
Conditional [MASK] Discrete Diffusion Language Model
por: Koh, Hyukhun, et al.
Publicado: (2024)