LILO: Bayesian Optimization with Natural Language Feedback
Fuente:
arXiv
Guardado en:
| Autores principales: | Kobalczyk, Katarzyna, Lin, Zhiyuan Jerry, Letham, Benjamin, Zhao, Zhuokai, Balandat, Maximilian, Bakshy, Eytan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Embedding by Elicitation: Dynamic Representations for Bayesian Optimization of System Prompts
por: Lin, Zhiyuan Jerry, et al.
Publicado: (2026)
por: Lin, Zhiyuan Jerry, et al.
Publicado: (2026)
BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability
por: Daulton, Samuel, et al.
Publicado: (2026)
por: Daulton, Samuel, et al.
Publicado: (2026)
Experimenting, Fast and Slow: Bayesian Optimization of Long-term Outcomes with Online Experiments
por: Feng, Qing, et al.
Publicado: (2025)
por: Feng, Qing, et al.
Publicado: (2025)
Joint Composite Latent Space Bayesian Optimization
por: Maus, Natalie, et al.
Publicado: (2023)
por: Maus, Natalie, et al.
Publicado: (2023)
LILO: Learning Interpretable Libraries by Compressing and Documenting Code
por: Grand, Gabriel, et al.
Publicado: (2023)
por: Grand, Gabriel, et al.
Publicado: (2023)
Active Learning for Derivative-Based Global Sensitivity Analysis with Gaussian Processes
por: Belakaria, Syrine, et al.
Publicado: (2024)
por: Belakaria, Syrine, et al.
Publicado: (2024)
Robust Gaussian Processes via Relevance Pursuit
por: Ament, Sebastian, et al.
Publicado: (2024)
por: Ament, Sebastian, et al.
Publicado: (2024)
Informed Initialization for Bayesian Optimization and Active Learning
por: Hvarfner, Carl, et al.
Publicado: (2025)
por: Hvarfner, Carl, et al.
Publicado: (2025)
Active Task Disambiguation with LLMs
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
Unexpected Improvements to Expected Improvement for Bayesian Optimization
por: Ament, Sebastian, et al.
Publicado: (2023)
por: Ament, Sebastian, et al.
Publicado: (2023)
Scaling Gaussian Processes for Learning Curve Prediction via Latent Kronecker Structure
por: Lin, Jihao Andreas, et al.
Publicado: (2024)
por: Lin, Jihao Andreas, et al.
Publicado: (2024)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
por: Namkoong, Hongseok, et al.
Publicado: (2020)
por: Namkoong, Hongseok, et al.
Publicado: (2020)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
por: Cui, Ganqu, et al.
Publicado: (2023)
por: Cui, Ganqu, et al.
Publicado: (2023)
Preference Learning for AI Alignment: a Causal Perspective
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
LaFFi: Leveraging Hybrid Natural Language Feedback for Fine-tuning Language Models
por: Li, Qianxi, et al.
Publicado: (2023)
por: Li, Qianxi, et al.
Publicado: (2023)
Natural Language Reinforcement Learning
por: Feng, Xidong, et al.
Publicado: (2024)
por: Feng, Xidong, et al.
Publicado: (2024)
Discovery of Hidden Miscalibration Regimes
por: Kobalczyk, Katarzyna, et al.
Publicado: (2026)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2026)
Language-Guided Tuning: Enhancing Numeric Optimization with Textual Feedback
por: Lu, Yuxing, et al.
Publicado: (2025)
por: Lu, Yuxing, et al.
Publicado: (2025)
Improving Code Generation by Training with Natural Language Feedback
por: Chen, Angelica, et al.
Publicado: (2023)
por: Chen, Angelica, et al.
Publicado: (2023)
The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data
por: Pouplin, Thomas, et al.
Publicado: (2024)
por: Pouplin, Thomas, et al.
Publicado: (2024)
Training Language Models with Language Feedback at Scale
por: Scheurer, Jérémy, et al.
Publicado: (2023)
por: Scheurer, Jérémy, et al.
Publicado: (2023)
TARo: Token-level Adaptive Routing for LLM Test-time Alignment
por: Rai, Arushi, et al.
Publicado: (2026)
por: Rai, Arushi, et al.
Publicado: (2026)
Token-Level LLM Collaboration via FusionRoute
por: Xiong, Nuoya, et al.
Publicado: (2026)
por: Xiong, Nuoya, et al.
Publicado: (2026)
Eliciting Numerical Predictive Distributions of LLMs Without Autoregression
por: Piskorz, Julianna, et al.
Publicado: (2026)
por: Piskorz, Julianna, et al.
Publicado: (2026)
Research on Optimization of Natural Language Processing Model Based on Multimodal Deep Learning
por: Sun, Dan, et al.
Publicado: (2024)
por: Sun, Dan, et al.
Publicado: (2024)
SELF: Self-Evolution with Language Feedback
por: Lu, Jianqiao, et al.
Publicado: (2023)
por: Lu, Jianqiao, et al.
Publicado: (2023)
Enhancing AI Assisted Writing with One-Shot Implicit Negative Feedback
por: Towle, Benjamin, et al.
Publicado: (2024)
por: Towle, Benjamin, et al.
Publicado: (2024)
Solving General Natural-Language-Description Optimization Problems with Large Language Models
por: Zhang, Jihai, et al.
Publicado: (2024)
por: Zhang, Jihai, et al.
Publicado: (2024)
Interactive Training: Feedback-Driven Neural Network Optimization
por: Zhang, Wentao, et al.
Publicado: (2025)
por: Zhang, Wentao, et al.
Publicado: (2025)
Language Models Can Learn from Verbal Feedback Without Scalar Rewards
por: Luo, Renjie, et al.
Publicado: (2025)
por: Luo, Renjie, et al.
Publicado: (2025)
How Well Can a Long Sequence Model Model Long Sequences? Comparing Architechtural Inductive Biases on Long-Context Abilities
por: Huang, Jerry
Publicado: (2024)
por: Huang, Jerry
Publicado: (2024)
Policy Improvement using Language Feedback Models
por: Zhong, Victor, et al.
Publicado: (2024)
por: Zhong, Victor, et al.
Publicado: (2024)
Towards Aligning Language Models with Textual Feedback
por: Lloret, Saüc Abadal, et al.
Publicado: (2024)
por: Lloret, Saüc Abadal, et al.
Publicado: (2024)
An Analysis of Hyper-Parameter Optimization Methods for Retrieval Augmented Generation
por: Orbach, Matan, et al.
Publicado: (2025)
por: Orbach, Matan, et al.
Publicado: (2025)
Personalized Language Modeling from Personalized Human Feedback
por: Li, Xinyu, et al.
Publicado: (2024)
por: Li, Xinyu, et al.
Publicado: (2024)
Reasoning Elicitation in Language Models via Counterfactual Feedback
por: Hüyük, Alihan, et al.
Publicado: (2024)
por: Hüyük, Alihan, et al.
Publicado: (2024)
SIPDO: Closed-Loop Prompt Optimization via Synthetic Data Feedback
por: Yu, Yaoning, et al.
Publicado: (2025)
por: Yu, Yaoning, et al.
Publicado: (2025)
Natural Language Reinforcement Learning
por: Feng, Xidong, et al.
Publicado: (2024)
por: Feng, Xidong, et al.
Publicado: (2024)
On Uncertainty In Natural Language Processing
por: Ulmer, Dennis
Publicado: (2024)
por: Ulmer, Dennis
Publicado: (2024)
DISCO Balances the Scales: Adaptive Domain- and Difficulty-Aware Reinforcement Learning on Imbalanced Data
por: Zhou, Yuhang, et al.
Publicado: (2025)
por: Zhou, Yuhang, et al.
Publicado: (2025)
Ejemplares similares
-
Embedding by Elicitation: Dynamic Representations for Bayesian Optimization of System Prompts
por: Lin, Zhiyuan Jerry, et al.
Publicado: (2026) -
BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability
por: Daulton, Samuel, et al.
Publicado: (2026) -
Experimenting, Fast and Slow: Bayesian Optimization of Long-term Outcomes with Online Experiments
por: Feng, Qing, et al.
Publicado: (2025) -
Joint Composite Latent Space Bayesian Optimization
por: Maus, Natalie, et al.
Publicado: (2023) -
LILO: Learning Interpretable Libraries by Compressing and Documenting Code
por: Grand, Gabriel, et al.
Publicado: (2023)