SEE: Strategic Exploration and Exploitation for Cohesive In-Context Prompt Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Cui, Wendi, Li, Zhuohang, Sun, Hao, Lopez, Damien, Das, Kamalika, Malin, Bradley, Kumar, Sricharan, Zhang, Jiaxin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Survey of Automatic Prompt Optimization with Instruction-focused Heuristic-based Search Algorithm
di: Cui, Wendi, et al.
Pubblicazione: (2025)
di: Cui, Wendi, et al.
Pubblicazione: (2025)
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
di: Cui, Wendi, et al.
Pubblicazione: (2024)
di: Cui, Wendi, et al.
Pubblicazione: (2024)
SCE: Scalable Consistency Ensembles Make Blackbox Large Language Model Generation More Reliable
di: Zhang, Jiaxin, et al.
Pubblicazione: (2025)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2025)
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
di: Zhang, Jiaxin, et al.
Pubblicazione: (2023)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2023)
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
di: Li, Zhuohang, et al.
Pubblicazione: (2024)
di: Li, Zhuohang, et al.
Pubblicazione: (2024)
Synthetic Knowledge Ingestion: Towards Knowledge Refinement and Injection for Enhancing Large Language Models
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
Survival of the Safest: Towards Secure Prompt Optimization through Interleaved Multi-Objective Evolution
di: Sinha, Ankita, et al.
Pubblicazione: (2024)
di: Sinha, Ankita, et al.
Pubblicazione: (2024)
Discriminant Distance-Aware Representation on Deterministic Uncertainty Quantification Methods
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
Towards Statistical Factuality Guarantee for Large Vision-Language Models
di: Li, Zhuohang, et al.
Pubblicazione: (2025)
di: Li, Zhuohang, et al.
Pubblicazione: (2025)
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
Customizing Language Model Responses with Contrastive In-Context Learning
di: Gao, Xiang, et al.
Pubblicazione: (2024)
di: Gao, Xiang, et al.
Pubblicazione: (2024)
From Passive Metric to Active Signal: The Evolving Role of Uncertainty Quantification in Large Language Models
di: Zhang, Jiaxin, et al.
Pubblicazione: (2026)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2026)
Learning to Search Effective Example Sequences for In-Context Learning
di: Gao, Xiang, et al.
Pubblicazione: (2025)
di: Gao, Xiang, et al.
Pubblicazione: (2025)
SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models
di: Gao, Xiang, et al.
Pubblicazione: (2024)
di: Gao, Xiang, et al.
Pubblicazione: (2024)
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff
di: Tang, Hao, et al.
Pubblicazione: (2024)
di: Tang, Hao, et al.
Pubblicazione: (2024)
StraGo: Harnessing Strategic Guidance for Prompt Optimization
di: Wu, Yurong, et al.
Pubblicazione: (2024)
di: Wu, Yurong, et al.
Pubblicazione: (2024)
General Exploratory Bonus for Optimistic Exploration in RLHF
di: Li, Wendi, et al.
Pubblicazione: (2025)
di: Li, Wendi, et al.
Pubblicazione: (2025)
HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control
di: Yao, Xincheng, et al.
Pubblicazione: (2026)
di: Yao, Xincheng, et al.
Pubblicazione: (2026)
Prompt Exploration with Prompt Regression
di: Feffer, Michael, et al.
Pubblicazione: (2024)
di: Feffer, Michael, et al.
Pubblicazione: (2024)
Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization
di: Loiseau, Gabriel, et al.
Pubblicazione: (2026)
di: Loiseau, Gabriel, et al.
Pubblicazione: (2026)
SEE: Sememe Entanglement Encoding for Transformer-bases Models Compression
di: Zhang, Jing, et al.
Pubblicazione: (2024)
di: Zhang, Jing, et al.
Pubblicazione: (2024)
Semantic-Space Exploration and Exploitation in RLVR for LLM Reasoning
di: Huang, Fanding, et al.
Pubblicazione: (2025)
di: Huang, Fanding, et al.
Pubblicazione: (2025)
Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration
di: Huang, Langlin, et al.
Pubblicazione: (2026)
di: Huang, Langlin, et al.
Pubblicazione: (2026)
SEE: Continual Fine-tuning with Sequential Ensemble of Experts
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
A Prompt-Based Knowledge Graph Foundation Model for Universal In-Context Reasoning
di: Cui, Yuanning, et al.
Pubblicazione: (2024)
di: Cui, Yuanning, et al.
Pubblicazione: (2024)
ExpLang: Improved Exploration and Exploitation in LLM Reasoning with On-Policy Thinking Language Selection
di: Gao, Changjiang, et al.
Pubblicazione: (2026)
di: Gao, Changjiang, et al.
Pubblicazione: (2026)
Token-weighted Direct Preference Optimization with Attention
di: Huang, Chengyu, et al.
Pubblicazione: (2026)
di: Huang, Chengyu, et al.
Pubblicazione: (2026)
LLMsPark: A Benchmark for Evaluating Large Language Models in Strategic Gaming Contexts
di: Chen, Junhao, et al.
Pubblicazione: (2025)
di: Chen, Junhao, et al.
Pubblicazione: (2025)
Disentangling Exploration of Large Language Models by Optimal Exploitation
di: Grams, Tim, et al.
Pubblicazione: (2025)
di: Grams, Tim, et al.
Pubblicazione: (2025)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
di: Baidya, Avinash, et al.
Pubblicazione: (2025)
di: Baidya, Avinash, et al.
Pubblicazione: (2025)
Token Hidden Reward: Steering Exploration-Exploitation in Group Relative Deep Reinforcement Learning
di: Deng, Wenlong, et al.
Pubblicazione: (2025)
di: Deng, Wenlong, et al.
Pubblicazione: (2025)
Exploration vs Exploitation: Rethinking RLVR through Clipping, Entropy, and Spurious Reward
di: Chen, Peter, et al.
Pubblicazione: (2025)
di: Chen, Peter, et al.
Pubblicazione: (2025)
Dual-Track CoT: Budget-Aware Stepwise Guidance for Small LMs
di: Chatterjee, Sagnik, et al.
Pubblicazione: (2026)
di: Chatterjee, Sagnik, et al.
Pubblicazione: (2026)
Human-Interpretable Adversarial Prompt Attack on Large Language Models with Situational Context
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
Balancing Exploration and Exploitation in LLM using Soft RLLF for Enhanced Negation Understanding
di: Nguyen, Ha-Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Ha-Thanh, et al.
Pubblicazione: (2024)
Prompt Optimization via Adversarial In-Context Learning
di: Do, Xuan Long, et al.
Pubblicazione: (2023)
di: Do, Xuan Long, et al.
Pubblicazione: (2023)
Dist2ill: Distributional Distillation for One-Pass Uncertainty Estimation in Large Language Models
di: Zhao, Yicong, et al.
Pubblicazione: (2025)
di: Zhao, Yicong, et al.
Pubblicazione: (2025)
Attention Overflow: Language Model Input Blur during Long-Context Missing Items Recommendation
di: Sileo, Damien
Pubblicazione: (2024)
di: Sileo, Damien
Pubblicazione: (2024)
Logic Haystacks: Probing LLMs Long-Context Logical Reasoning (Without Easily Identifiable Unrelated Padding)
di: Sileo, Damien
Pubblicazione: (2025)
di: Sileo, Damien
Pubblicazione: (2025)
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
di: Shen, Guangyu, et al.
Pubblicazione: (2024)
di: Shen, Guangyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Survey of Automatic Prompt Optimization with Instruction-focused Heuristic-based Search Algorithm
di: Cui, Wendi, et al.
Pubblicazione: (2025) -
DCR-Consistency: Divide-Conquer-Reasoning for Consistency Evaluation and Improvement of Large Language Models
di: Cui, Wendi, et al.
Pubblicazione: (2024) -
SCE: Scalable Consistency Ensembles Make Blackbox Large Language Model Generation More Reliable
di: Zhang, Jiaxin, et al.
Pubblicazione: (2025) -
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
di: Zhang, Jiaxin, et al.
Pubblicazione: (2023) -
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
di: Li, Zhuohang, et al.
Pubblicazione: (2024)