Prompt Optimization via Adversarial In-Context Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Do, Xuan Long, Zhao, Yiran, Brown, Hannah, Xie, Yuxi, Zhao, James Xu, Chen, Nancy F., Kawaguchi, Kenji, Shieh, Michael, He, Junxian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
von: Brown, Hannah, et al.
Veröffentlicht: (2024)
In-Context Reinforcement Learning for Tool Use in Large Language Models
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026)
Single Character Perturbations Break LLM Alignment
von: Lin, Leon, et al.
Veröffentlicht: (2024)
von: Lin, Leon, et al.
Veröffentlicht: (2024)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
von: Gan, Esther, et al.
Veröffentlicht: (2024)
von: Gan, Esther, et al.
Veröffentlicht: (2024)
Advancing Adversarial Suffix Transfer Learning on Aligned Large Language Models
von: Liu, Hongfu, et al.
Veröffentlicht: (2024)
von: Liu, Hongfu, et al.
Veröffentlicht: (2024)
Beyond In-Context Learning: Aligning Long-form Generation of Large Language Models via Task-Inherent Attribute Guidelines
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning
von: Long, Do Xuan, et al.
Veröffentlicht: (2023)
von: Long, Do Xuan, et al.
Veröffentlicht: (2023)
Multi-expert Prompting Improves Reliability, Safety, and Usefulness of Large Language Models
von: Long, Do Xuan, et al.
Veröffentlicht: (2024)
von: Long, Do Xuan, et al.
Veröffentlicht: (2024)
What Makes a Good Natural Language Prompt?
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
von: Long, Do Xuan, et al.
Veröffentlicht: (2025)
ImageEdit-R1: Boosting Multi-Agent Image Editing via Reinforcement Learning
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
AdaMergeX: Cross-Lingual Transfer with Large Language Models via Adaptive Adapter Merging
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
Gradually Compacting Large Language Models for Reasoning Like a Boiling Frog
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
How do Large Language Models Handle Multilingualism?
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
The Emergence of Abstract Thought in Large Language Models Beyond Any Language
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards
von: Chen, Guanzheng, et al.
Veröffentlicht: (2026)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2026)
InstructCoder: Instruction Tuning Large Language Models for Code Editing
von: Li, Kaixin, et al.
Veröffentlicht: (2023)
von: Li, Kaixin, et al.
Veröffentlicht: (2023)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
von: Long, Do Xuan, et al.
Veröffentlicht: (2024)
von: Long, Do Xuan, et al.
Veröffentlicht: (2024)
Unnatural Languages Are Not Bugs but Features for LLMs
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
von: Duan, Keyu, et al.
Veröffentlicht: (2025)
Defending Jailbreak Prompts via In-Context Adversarial Game
von: Zhou, Yujun, et al.
Veröffentlicht: (2024)
von: Zhou, Yujun, et al.
Veröffentlicht: (2024)
Context-aware Prompt Tuning: Advancing In-Context Learning with Adversarial Methods
von: Blau, Tsachi, et al.
Veröffentlicht: (2024)
von: Blau, Tsachi, et al.
Veröffentlicht: (2024)
LongPO: Long Context Self-Evolution of Large Language Models through Short-to-Long Preference Optimization
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
Exact Conversion of In-Context Learning to Model Weights in Linearized-Attention Transformers
von: Chen, Brian K, et al.
Veröffentlicht: (2024)
von: Chen, Brian K, et al.
Veröffentlicht: (2024)
LOCA-bench: Benchmarking Language Agents Under Controllable and Extreme Context Growth
von: Zeng, Weihao, et al.
Veröffentlicht: (2026)
von: Zeng, Weihao, et al.
Veröffentlicht: (2026)
Joint-Optimized Unsupervised Adversarial Domain Adaptation in Remote Sensing Segmentation with Prompted Foundation Model
von: Lyu, Shuchang, et al.
Veröffentlicht: (2024)
von: Lyu, Shuchang, et al.
Veröffentlicht: (2024)
Timelike Entanglement Entropy in Higher Curvature Gravity
von: Zhao, Zi-Xuan, et al.
Veröffentlicht: (2025)
von: Zhao, Zi-Xuan, et al.
Veröffentlicht: (2025)
Dynamics as Prompts: In-Context Learning for Sim-to-Real System Identifications
von: Zhang, Xilun, et al.
Veröffentlicht: (2024)
von: Zhang, Xilun, et al.
Veröffentlicht: (2024)
Strong identifiability and parameter learning in regression with heterogeneous response
von: Do, Dat, et al.
Veröffentlicht: (2022)
von: Do, Dat, et al.
Veröffentlicht: (2022)
Can AI Be as Creative as Humans?
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Entanglement and Pseudo Entanglement Dynamics versus Fusion in CFT
von: He, Song, et al.
Veröffentlicht: (2023)
von: He, Song, et al.
Veröffentlicht: (2023)
V-DPO: Mitigating Hallucination in Large Vision Language Models via Vision-Guided Direct Preference Optimization
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
von: Xie, Yuxi, et al.
Veröffentlicht: (2024)
Design and Evaluation of Deep Learning-Based Dual-Spectrum Image Fusion Methods
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
Power and sample size calculations for testing linear combinations of group means under variance heterogeneity with applications to meta and moderation analyses
von: Gwowen Shieh
Veröffentlicht: (2015)
von: Gwowen Shieh
Veröffentlicht: (2015)
On using a pilot sample variance for sample size determination in the detection of differences between two means: Power consideration
von: Gwowen Shieh
Veröffentlicht: (2013)
von: Gwowen Shieh
Veröffentlicht: (2013)
Improving Fairness of Large Language Model-Based ICU Mortality Prediction via Case-Based Prompting
von: Zhang, Gangxiong, et al.
Veröffentlicht: (2025)
von: Zhang, Gangxiong, et al.
Veröffentlicht: (2025)
Mirage or Method? How Model-Task Alignment Induces Divergent RL Conclusions
von: Wu, Haoze, et al.
Veröffentlicht: (2025)
von: Wu, Haoze, et al.
Veröffentlicht: (2025)
Bi-Level Prompt Optimization for Multimodal LLM-as-a-Judge
von: Pan, Bo, et al.
Veröffentlicht: (2026)
von: Pan, Bo, et al.
Veröffentlicht: (2026)
Joint Semantic Token Selection and Prompt Optimization for Interpretable Prompt Learning
von: Wang, Yating, et al.
Veröffentlicht: (2026)
von: Wang, Yating, et al.
Veröffentlicht: (2026)
RAPID: Long-Context Inference with Retrieval-Augmented Speculative Decoding
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
von: Chen, Guanzheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling
von: Zhao, Yiran, et al.
Veröffentlicht: (2024) -
Self-Evaluation as a Defense Against Adversarial Attacks on LLMs
von: Brown, Hannah, et al.
Veröffentlicht: (2024) -
In-Context Reinforcement Learning for Tool Use in Large Language Models
von: Ye, Yaoqi, et al.
Veröffentlicht: (2026) -
Single Character Perturbations Break LLM Alignment
von: Lin, Leon, et al.
Veröffentlicht: (2024) -
Reasoning Robustness of LLMs to Adversarial Typographical Errors
von: Gan, Esther, et al.
Veröffentlicht: (2024)