StablePrompt: Automatic Prompt Tuning using Reinforcement Learning for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Kwon, Minchan, Kim, Gaeun, Kim, Jongsuk, Lee, Haeil, Kim, Junmo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Preference Distillation via Value based Reinforcement Learning
di: Kwon, Minchan, et al.
Pubblicazione: (2025)
di: Kwon, Minchan, et al.
Pubblicazione: (2025)
FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition
di: Kim, Jongsuk, et al.
Pubblicazione: (2025)
di: Kim, Jongsuk, et al.
Pubblicazione: (2025)
Bayesian Multi-Task Transfer Learning for Soft Prompt Tuning
di: Lee, Haeju, et al.
Pubblicazione: (2024)
di: Lee, Haeju, et al.
Pubblicazione: (2024)
FxSearcher: gradient-free text-driven audio transformation
di: Ki, Hojoon, et al.
Pubblicazione: (2025)
di: Ki, Hojoon, et al.
Pubblicazione: (2025)
AVCap: Leveraging Audio-Visual Features as Text Tokens for Captioning
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
di: Lee, Hansang, et al.
Pubblicazione: (2017)
di: Lee, Hansang, et al.
Pubblicazione: (2017)
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
di: Kwon, Minchan, et al.
Pubblicazione: (2026)
di: Kwon, Minchan, et al.
Pubblicazione: (2026)
The Effects of Mixed Sample Data Augmentation are Class Dependent
di: Lee, Haeil, et al.
Pubblicazione: (2023)
di: Lee, Haeil, et al.
Pubblicazione: (2023)
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
di: Lee, Hansang, et al.
Pubblicazione: (2022)
di: Lee, Hansang, et al.
Pubblicazione: (2022)
GFlowPO: Generative Flow Network as a Language Model Prompt Optimizer
di: Cho, Junmo, et al.
Pubblicazione: (2026)
di: Cho, Junmo, et al.
Pubblicazione: (2026)
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
di: Lee, Haeil, et al.
Pubblicazione: (2024)
di: Lee, Haeil, et al.
Pubblicazione: (2024)
Inspecting Explainability of Transformer Models with Additional Statistical Information
di: Nguyen, Hoang C., et al.
Pubblicazione: (2023)
di: Nguyen, Hoang C., et al.
Pubblicazione: (2023)
Hard Prompts Made Interpretable: Sparse Entropy Regularization for Prompt Tuning with RL
di: Choi, Yunseon, et al.
Pubblicazione: (2024)
di: Choi, Yunseon, et al.
Pubblicazione: (2024)
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
di: Kim, Minseo, et al.
Pubblicazione: (2026)
di: Kim, Minseo, et al.
Pubblicazione: (2026)
Query-Aware Learnable Graph Pooling Tokens as Prompt for Large Language Models
di: Kim, Wooyoung, et al.
Pubblicazione: (2025)
di: Kim, Wooyoung, et al.
Pubblicazione: (2025)
EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined Criteria
di: Kim, Tae Soo, et al.
Pubblicazione: (2023)
di: Kim, Tae Soo, et al.
Pubblicazione: (2023)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
di: Lee, Hansang, et al.
Pubblicazione: (2022)
di: Lee, Hansang, et al.
Pubblicazione: (2022)
Learning to Reduce: Optimal Representations of Structured Data in Prompting Large Language Models
di: Lee, Younghun, et al.
Pubblicazione: (2024)
di: Lee, Younghun, et al.
Pubblicazione: (2024)
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning
di: Kim, Gyeongman, et al.
Pubblicazione: (2024)
di: Kim, Gyeongman, et al.
Pubblicazione: (2024)
Discrete Prompt Compression with Reinforcement Learning
di: Jung, Hoyoun, et al.
Pubblicazione: (2023)
di: Jung, Hoyoun, et al.
Pubblicazione: (2023)
DETAIL Matters: Measuring the Impact of Prompt Specificity on Reasoning in Large Language Models
di: Kim, Olivia
Pubblicazione: (2025)
di: Kim, Olivia
Pubblicazione: (2025)
Comparison Reveals Commonality: Customized Image Generation through Contrastive Inversion
di: Kim, Minseo, et al.
Pubblicazione: (2025)
di: Kim, Minseo, et al.
Pubblicazione: (2025)
Counterfactual-Consistency Prompting for Relative Temporal Understanding in Large Language Models
di: Kim, Jongho, et al.
Pubblicazione: (2025)
di: Kim, Jongho, et al.
Pubblicazione: (2025)
Automatic Prompt Selection for Large Language Models
di: Do, Viet-Tung, et al.
Pubblicazione: (2024)
di: Do, Viet-Tung, et al.
Pubblicazione: (2024)
MedRep: Medical Concept Representation for General Electronic Health Record Foundation Models
di: Kim, Junmo, et al.
Pubblicazione: (2025)
di: Kim, Junmo, et al.
Pubblicazione: (2025)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
di: Kim, Minchan, et al.
Pubblicazione: (2024)
di: Kim, Minchan, et al.
Pubblicazione: (2024)
Natural Language Declarative Prompting (NLD-P): A Modular Governance Method for Prompt Design Under Model Drift
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
di: Kim, Hyunwoo, et al.
Pubblicazione: (2026)
Automatic Summarization of Doctor-Patient Encounter Dialogues Using Large Language Model through Prompt Tuning
di: Lyu, Mengxian, et al.
Pubblicazione: (2024)
di: Lyu, Mengxian, et al.
Pubblicazione: (2024)
BAPO: Base-Anchored Preference Optimization for Overcoming Forgetting in Large Language Models Personalization
di: Lee, Gihun, et al.
Pubblicazione: (2024)
di: Lee, Gihun, et al.
Pubblicazione: (2024)
Bayesian Principles Improve Prompt Learning In Vision-Language Models
di: Kim, Mingyu, et al.
Pubblicazione: (2025)
di: Kim, Mingyu, et al.
Pubblicazione: (2025)
Instructive Decoding: Instruction-Tuned Large Language Models are Self-Refiner from Noisy Instructions
di: Kim, Taehyeon, et al.
Pubblicazione: (2023)
di: Kim, Taehyeon, et al.
Pubblicazione: (2023)
RePrompt: Planning by Automatic Prompt Engineering for Large Language Models Agents
di: Chen, Weizhe, et al.
Pubblicazione: (2024)
di: Chen, Weizhe, et al.
Pubblicazione: (2024)
Beyond Prompts: Learning from Human Communication for Enhanced AI Intent Alignment
di: Kim, Yoonsu, et al.
Pubblicazione: (2024)
di: Kim, Yoonsu, et al.
Pubblicazione: (2024)
PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses
di: Hong, Minki, et al.
Pubblicazione: (2026)
di: Hong, Minki, et al.
Pubblicazione: (2026)
IAPT: Instruction-Aware Prompt Tuning for Large Language Models
di: Zhu, Wei, et al.
Pubblicazione: (2024)
di: Zhu, Wei, et al.
Pubblicazione: (2024)
DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
di: Jung, Sunghee, et al.
Pubblicazione: (2025)
di: Jung, Sunghee, et al.
Pubblicazione: (2025)
OWQ: Outlier-Aware Weight Quantization for Efficient Fine-Tuning and Inference of Large Language Models
di: Lee, Changhun, et al.
Pubblicazione: (2023)
di: Lee, Changhun, et al.
Pubblicazione: (2023)
Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning
di: Liu, Wenjin, et al.
Pubblicazione: (2025)
di: Liu, Wenjin, et al.
Pubblicazione: (2025)
Promptomatix: An Automatic Prompt Optimization Framework for Large Language Models
di: Murthy, Rithesh, et al.
Pubblicazione: (2025)
di: Murthy, Rithesh, et al.
Pubblicazione: (2025)
TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs
di: Wanna, Selma, et al.
Pubblicazione: (2024)
di: Wanna, Selma, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Preference Distillation via Value based Reinforcement Learning
di: Kwon, Minchan, et al.
Pubblicazione: (2025) -
FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition
di: Kim, Jongsuk, et al.
Pubblicazione: (2025) -
Bayesian Multi-Task Transfer Learning for Soft Prompt Tuning
di: Lee, Haeju, et al.
Pubblicazione: (2024) -
FxSearcher: gradient-free text-driven audio transformation
di: Ki, Hojoon, et al.
Pubblicazione: (2025) -
AVCap: Leveraging Audio-Visual Features as Text Tokens for Captioning
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)