Dependency Structure Search Bayesian Optimization for Decision Making Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rajpal, Mohit, Tran, Lac Gia, Zhang, Yehong, Low, Bryan Kian Hsiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
BoLT: A Benchmark to Democratize Black-box Optimization Research for Expensive LLM Tasks
von: Chew, Ruth Wan Theng, et al.
Veröffentlicht: (2026)
von: Chew, Ruth Wan Theng, et al.
Veröffentlicht: (2026)
Keep Everyone Happy: Online Fair Division of Numerous Items with Few Copies
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
COBRA: Contextual Bandit Algorithm for Ensuring Truthful Strategic Agents
von: Verma, Arun, et al.
Veröffentlicht: (2025)
von: Verma, Arun, et al.
Veröffentlicht: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks
von: Chen, Zhiliang, et al.
Veröffentlicht: (2025)
von: Chen, Zhiliang, et al.
Veröffentlicht: (2025)
Fine-tuning Language Models with Generative Adversarial Reward Modelling
von: Yu, Zhang Ze, et al.
Veröffentlicht: (2023)
von: Yu, Zhang Ze, et al.
Veröffentlicht: (2023)
Prompt Optimization with Human Feedback
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2024)
BarrierSteer: LLM Safety via Learning Barrier Steering
von: Tran, Thanh Q., et al.
Veröffentlicht: (2026)
von: Tran, Thanh Q., et al.
Veröffentlicht: (2026)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
von: Wang, Jingtan, et al.
Veröffentlicht: (2024)
von: Wang, Jingtan, et al.
Veröffentlicht: (2024)
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
Data value estimation on private gradients
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
DeRDaVa: Deletion-Robust Data Valuation for Machine Learning
von: Tian, Xiao, et al.
Veröffentlicht: (2023)
von: Tian, Xiao, et al.
Veröffentlicht: (2023)
INO-SGD: Addressing Utility Imbalance under Individualized Differential Privacy
von: Tian, Xiao, et al.
Veröffentlicht: (2026)
von: Tian, Xiao, et al.
Veröffentlicht: (2026)
Ferret: Federated Full-Parameter Tuning at Scale for Large Language Models
von: Shu, Yao, et al.
Veröffentlicht: (2024)
von: Shu, Yao, et al.
Veröffentlicht: (2024)
Paid with Models: Optimal Contract Design for Collaborative Machine Learning
von: Wang, Bingchen, et al.
Veröffentlicht: (2024)
von: Wang, Bingchen, et al.
Veröffentlicht: (2024)
REFRAG: Rethinking RAG based Decoding
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2025)
On Newton's Method to Unlearn Neural Networks
von: Bui, Nhung, et al.
Veröffentlicht: (2024)
von: Bui, Nhung, et al.
Veröffentlicht: (2024)
Source Attribution for Large Language Model-Generated Data
von: Wang, Jingtan, et al.
Veröffentlicht: (2023)
von: Wang, Jingtan, et al.
Veröffentlicht: (2023)
DETAIL: Task DEmonsTration Attribution for Interpretable In-context Learning
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
von: Zhou, Zijian, et al.
Veröffentlicht: (2024)
Is Data Shapley Not Better than Random in Data Selection? Ask NASH
von: Tian, Xiao, et al.
Veröffentlicht: (2026)
von: Tian, Xiao, et al.
Veröffentlicht: (2026)
How Hard Can It Be? Hardness-Aware Multi-Objective Unlearning
von: Chen, Jiangwei, et al.
Veröffentlicht: (2026)
von: Chen, Jiangwei, et al.
Veröffentlicht: (2026)
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
von: Wu, Zhaoxuan, et al.
Veröffentlicht: (2024)
von: Wu, Zhaoxuan, et al.
Veröffentlicht: (2024)
Decentralized Sum-of-Nonconvex Optimization
von: Liu, Zhuanghua, et al.
Veröffentlicht: (2024)
von: Liu, Zhuanghua, et al.
Veröffentlicht: (2024)
BILBO: BILevel Bayesian Optimization
von: Chew, Ruth Wan Theng, et al.
Veröffentlicht: (2025)
von: Chew, Ruth Wan Theng, et al.
Veröffentlicht: (2025)
Group-robust Sample Reweighting for Subpopulation Shifts via Influence Functions
von: Qiao, Rui, et al.
Veröffentlicht: (2025)
von: Qiao, Rui, et al.
Veröffentlicht: (2025)
Uncertainty Quantification for Multimodal Large Language Models with Incoherence-adjusted Semantic Volume
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2026)
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2026)
TreeGrad-Ranker: Feature Ranking via $O(L)$-Time Gradients for Decision Trees
von: Li, Weida, et al.
Veröffentlicht: (2026)
von: Li, Weida, et al.
Veröffentlicht: (2026)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
Decision Making in Non-Stationary Environments with Policy-Augmented Search
von: Pettet, Ava, et al.
Veröffentlicht: (2024)
von: Pettet, Ava, et al.
Veröffentlicht: (2024)
PIED: Physics-Informed Experimental Design for Inverse Problems
von: Hemachandra, Apivich, et al.
Veröffentlicht: (2025)
von: Hemachandra, Apivich, et al.
Veröffentlicht: (2025)
PINNACLE: PINN Adaptive ColLocation and Experimental points selection
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
von: Lau, Gregory Kang Ruey, et al.
Veröffentlicht: (2024)
Bayesian Decision Making around Experts
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
von: Ornia, Daniel Jarne, et al.
Veröffentlicht: (2025)
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
The Chicken and Egg Dilemma: Co-optimizing Data and Model Configurations for LLMs
von: Chen, Zhiliang, et al.
Veröffentlicht: (2026)
von: Chen, Zhiliang, et al.
Veröffentlicht: (2026)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2023)
von: Lin, Xiaoqiang, et al.
Veröffentlicht: (2023)
Robustifying and Boosting Training-Free Neural Architecture Search
von: He, Zhenfeng, et al.
Veröffentlicht: (2024)
von: He, Zhenfeng, et al.
Veröffentlicht: (2024)
Hierarchical Decision Making Based on Structural Information Principles
von: Zeng, Xianghua, et al.
Veröffentlicht: (2024)
von: Zeng, Xianghua, et al.
Veröffentlicht: (2024)
Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
Generative Models in Decision Making: A Survey
von: Shao, Xinyu, et al.
Veröffentlicht: (2025)
von: Shao, Xinyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024) -
BoLT: A Benchmark to Democratize Black-box Optimization Research for Expensive LLM Tasks
von: Chew, Ruth Wan Theng, et al.
Veröffentlicht: (2026) -
Keep Everyone Happy: Online Fair Division of Numerous Items with Few Copies
von: Verma, Arun, et al.
Veröffentlicht: (2024) -
COBRA: Contextual Bandit Algorithm for Ensuring Truthful Strategic Agents
von: Verma, Arun, et al.
Veröffentlicht: (2025) -
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)