Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Xiaoqiang, Wu, Zhaoxuan, Dai, Zhongxiang, Hu, Wenyang, Shu, Yao, Ng, See-Kiong, Jaillet, Patrick, Low, Bryan Kian Hsiang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
by: Wu, Zhaoxuan, et al.
Published: (2024)
by: Wu, Zhaoxuan, et al.
Published: (2024)
Prompt Optimization with Human Feedback
by: Lin, Xiaoqiang, et al.
Published: (2024)
by: Lin, Xiaoqiang, et al.
Published: (2024)
Localized Zeroth-Order Prompt Optimization
by: Hu, Wenyang, et al.
Published: (2024)
by: Hu, Wenyang, et al.
Published: (2024)
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
by: Lin, Xiaoqiang, et al.
Published: (2025)
by: Lin, Xiaoqiang, et al.
Published: (2025)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
by: Verma, Arun, et al.
Published: (2024)
by: Verma, Arun, et al.
Published: (2024)
Ferret: Federated Full-Parameter Tuning at Scale for Large Language Models
by: Shu, Yao, et al.
Published: (2024)
by: Shu, Yao, et al.
Published: (2024)
TRACE: TRansformer-based Attribution using Contrastive Embeddings in LLMs
by: Wang, Cheng, et al.
Published: (2024)
by: Wang, Cheng, et al.
Published: (2024)
Source Attribution for Large Language Model-Generated Data
by: Wang, Jingtan, et al.
Published: (2023)
by: Wang, Jingtan, et al.
Published: (2023)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
Active Human Feedback Collection via Neural Contextual Dueling Bandits
by: Verma, Arun, et al.
Published: (2025)
by: Verma, Arun, et al.
Published: (2025)
Robustifying and Boosting Training-Free Neural Architecture Search
by: He, Zhenfeng, et al.
Published: (2024)
by: He, Zhenfeng, et al.
Published: (2024)
PIED: Physics-Informed Experimental Design for Inverse Problems
by: Hemachandra, Apivich, et al.
Published: (2025)
by: Hemachandra, Apivich, et al.
Published: (2025)
PINNACLE: PINN Adaptive ColLocation and Experimental points selection
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
by: Lau, Gregory Kang Ruey, et al.
Published: (2024)
On Newton's Method to Unlearn Neural Networks
by: Bui, Nhung, et al.
Published: (2024)
by: Bui, Nhung, et al.
Published: (2024)
Adjusted Expected Improvement for Cumulative Regret Minimization in Noisy Bayesian Optimization
by: Hu, Shouri, et al.
Published: (2022)
by: Hu, Shouri, et al.
Published: (2022)
Paid with Models: Optimal Contract Design for Collaborative Machine Learning
by: Wang, Bingchen, et al.
Published: (2024)
by: Wang, Bingchen, et al.
Published: (2024)
DUPRE: Data Utility Prediction for Efficient Data Valuation
by: Pham, Kieu Thao Nguyen, et al.
Published: (2025)
by: Pham, Kieu Thao Nguyen, et al.
Published: (2025)
Understanding Domain Generalization: A Noise Robustness Perspective
by: Qiao, Rui, et al.
Published: (2024)
by: Qiao, Rui, et al.
Published: (2024)
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
by: Vijayan, Nithia, et al.
Published: (2024)
by: Vijayan, Nithia, et al.
Published: (2024)
Decentralized Sum-of-Nonconvex Optimization
by: Liu, Zhuanghua, et al.
Published: (2024)
by: Liu, Zhuanghua, et al.
Published: (2024)
Group-robust Sample Reweighting for Subpopulation Shifts via Influence Functions
by: Qiao, Rui, et al.
Published: (2025)
by: Qiao, Rui, et al.
Published: (2025)
PC-MoE: Memory-Efficient and Privacy-Preserving Collaborative Training for Mixture-of-Experts LLMs
by: Zhang, Ze Yu, et al.
Published: (2025)
by: Zhang, Ze Yu, et al.
Published: (2025)
REFRAG: Rethinking RAG based Decoding
by: Lin, Xiaoqiang, et al.
Published: (2025)
by: Lin, Xiaoqiang, et al.
Published: (2025)
TETRIS: Optimal Draft Token Selection for Batch Speculative Decoding
by: Wu, Zhaoxuan, et al.
Published: (2025)
by: Wu, Zhaoxuan, et al.
Published: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
by: Zhao, James Xu, et al.
Published: (2025)
by: Zhao, James Xu, et al.
Published: (2025)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
by: Wang, Jingtan, et al.
Published: (2024)
by: Wang, Jingtan, et al.
Published: (2024)
Broaden your SCOPE! Efficient Multi-turn Conversation Planning for LLMs with Semantic Space
by: Chen, Zhiliang, et al.
Published: (2025)
by: Chen, Zhiliang, et al.
Published: (2025)
Incentives in Private Collaborative Machine Learning
by: Sim, Rachael Hwee Ling, et al.
Published: (2024)
by: Sim, Rachael Hwee Ling, et al.
Published: (2024)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
by: Goodge, Adam, et al.
Published: (2025)
by: Goodge, Adam, et al.
Published: (2025)
Global-to-Local Support Spectrums for Language Model Explainability
by: Agussurja, Lucas, et al.
Published: (2024)
by: Agussurja, Lucas, et al.
Published: (2024)
TreeGrad-Ranker: Feature Ranking via $O(L)$-Time Gradients for Decision Trees
by: Li, Weida, et al.
Published: (2026)
by: Li, Weida, et al.
Published: (2026)
Provably Adaptive Linear Approximation for the Shapley Value and Beyond
by: Li, Weida, et al.
Published: (2026)
by: Li, Weida, et al.
Published: (2026)
Incremental Quasi-Newton Methods with Faster Superlinear Convergence Rates
by: Liu, Zhuanghua, et al.
Published: (2024)
by: Liu, Zhuanghua, et al.
Published: (2024)
Uncovering Scaling Laws for Large Language Models via Inverse Problems
by: Verma, Arun, et al.
Published: (2025)
by: Verma, Arun, et al.
Published: (2025)
DETAIL: Task DEmonsTration Attribution for Interpretable In-context Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
MineDraft: A Framework for Batch Parallel Speculative Decoding
by: Tang, Zhenwei, et al.
Published: (2026)
by: Tang, Zhenwei, et al.
Published: (2026)
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
by: Liu, Renyang, et al.
Published: (2025)
by: Liu, Renyang, et al.
Published: (2025)
THERE IS NO LANGUAGE INSTINCT
by: Geoffrey Sampson
Published: (2007)
by: Geoffrey Sampson
Published: (2007)
Data-Centric AI in the Age of Large Language Models
by: Xu, Xinyi, et al.
Published: (2024)
by: Xu, Xinyi, et al.
Published: (2024)
Similar Items
-
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
by: Wu, Zhaoxuan, et al.
Published: (2024) -
Prompt Optimization with Human Feedback
by: Lin, Xiaoqiang, et al.
Published: (2024) -
Localized Zeroth-Order Prompt Optimization
by: Hu, Wenyang, et al.
Published: (2024) -
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
by: Lin, Xiaoqiang, et al.
Published: (2025) -
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
by: Verma, Arun, et al.
Published: (2024)