Large language models as uncertainty-calibrated optimizers for experimental discovery
Fuente:
arXiv
Guardado en:
| Autores principales: | Ranković, Bojana, Griffiths, Ryan-Rhys, Schwaller, Philippe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Gradient Guided Hypotheses: A unified solution to enable machine learning models on scarce and noisy data regimes
por: Neves, Paulo, et al.
Publicado: (2024)
por: Neves, Paulo, et al.
Publicado: (2024)
Lookup multivariate Kolmogorov-Arnold Networks
por: Pozdnyakov, Sergey, et al.
Publicado: (2025)
por: Pozdnyakov, Sergey, et al.
Publicado: (2025)
SynthStrategy: Extracting and Formalizing Latent Strategic Insights from LLMs in Organic Chemistry
por: Armstrong, Daniel, et al.
Publicado: (2025)
por: Armstrong, Daniel, et al.
Publicado: (2025)
Aviary: training language agents on challenging scientific tasks
por: Narayanan, Siddharth, et al.
Publicado: (2024)
por: Narayanan, Siddharth, et al.
Publicado: (2024)
Quantifying construct validity in large language model evaluations
por: Kearns, Ryan Othniel
Publicado: (2026)
por: Kearns, Ryan Othniel
Publicado: (2026)
A calibration test for evaluating set-based epistemic uncertainty representations
por: Jürgens, Mira, et al.
Publicado: (2025)
por: Jürgens, Mira, et al.
Publicado: (2025)
Trust the uncertain teacher: distilling dark knowledge via calibrated uncertainty
por: Kim, Jeonghyun, et al.
Publicado: (2026)
por: Kim, Jeonghyun, et al.
Publicado: (2026)
Neural Nonmyopic Bayesian Optimization in Dynamic Cost Settings
por: Truong, Sang T., et al.
Publicado: (2026)
por: Truong, Sang T., et al.
Publicado: (2026)
Pretraining with random noise for uncertainty calibration
por: Cheon, Jeonghwan, et al.
Publicado: (2024)
por: Cheon, Jeonghwan, et al.
Publicado: (2024)
Perovskite-R1: a domain-specialized large language model for intelligent discovery of precursor additives and experimental design
por: Wang, Xin-De, et al.
Publicado: (2025)
por: Wang, Xin-De, et al.
Publicado: (2025)
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks
por: Mei, Taiyuan, et al.
Publicado: (2024)
por: Mei, Taiyuan, et al.
Publicado: (2024)
Text-guided multi-property molecular optimization with a diffusion language model
por: Xiong, Yida, et al.
Publicado: (2024)
por: Xiong, Yida, et al.
Publicado: (2024)
Towards a Theory of AI Personhood
por: Ward, Francis Rhys
Publicado: (2025)
por: Ward, Francis Rhys
Publicado: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
por: Park, Ji Won, et al.
Publicado: (2025)
por: Park, Ji Won, et al.
Publicado: (2025)
Toward Efficient Exploration by Large Language Model Agents
por: Arumugam, Dilip, et al.
Publicado: (2025)
por: Arumugam, Dilip, et al.
Publicado: (2025)
Continuous-Time Analysis of Adaptive Optimization and Normalization
por: Gould, Rhys, et al.
Publicado: (2024)
por: Gould, Rhys, et al.
Publicado: (2024)
How do Large Language Models Navigate Conflicts between Honesty and Helpfulness?
por: Liu, Ryan, et al.
Publicado: (2024)
por: Liu, Ryan, et al.
Publicado: (2024)
Prior-informed optimization of treatment recommendation via bandit algorithms trained on large language model-processed historical records
por: Nessari, Saman, et al.
Publicado: (2025)
por: Nessari, Saman, et al.
Publicado: (2025)
Failure to Mix: Large language models struggle to answer according to desired probability distributions
por: Yang, Ivy Yuqian, et al.
Publicado: (2025)
por: Yang, Ivy Yuqian, et al.
Publicado: (2025)
Leveraging AI to optimize website structure discovery during Penetration Testing
por: Antonelli, Diego, et al.
Publicado: (2021)
por: Antonelli, Diego, et al.
Publicado: (2021)
Latent label distribution grid representation for modeling uncertainty
por: Sun, ShuNing, et al.
Publicado: (2025)
por: Sun, ShuNing, et al.
Publicado: (2025)
Are Large Language Models Sensitive to the Motives Behind Communication?
por: Wu, Addison J., et al.
Publicado: (2025)
por: Wu, Addison J., et al.
Publicado: (2025)
De-identification of clinical free text using natural language processing: A systematic review of current approaches
por: Kovačević, Aleksandar, et al.
Publicado: (2023)
por: Kovačević, Aleksandar, et al.
Publicado: (2023)
Performance of AI agents based on reasoning language models on ALD process optimization tasks
por: Yanguas-Gil, Angel
Publicado: (2026)
por: Yanguas-Gil, Angel
Publicado: (2026)
Real-time system optimal traffic routing under uncertainties -- Can physics models boost reinforcement learning?
por: Ke, Zemian, et al.
Publicado: (2024)
por: Ke, Zemian, et al.
Publicado: (2024)
How predictable is language model benchmark performance?
por: Owen, David
Publicado: (2024)
por: Owen, David
Publicado: (2024)
Pre-trained knowledge elevates large language models beyond traditional chemical reaction optimizers
por: MacKnight, Robert, et al.
Publicado: (2025)
por: MacKnight, Robert, et al.
Publicado: (2025)
Incorporating uncertainty quantification into travel mode choice modeling: a Bayesian neural network (BNN) approach and an uncertainty-guided active survey framework
por: Zheng, Shuwen, et al.
Publicado: (2024)
por: Zheng, Shuwen, et al.
Publicado: (2024)
Are large language models superhuman chemists?
por: Mirza, Adrian, et al.
Publicado: (2024)
por: Mirza, Adrian, et al.
Publicado: (2024)
Alignment faking in large language models
por: Greenblatt, Ryan, et al.
Publicado: (2024)
por: Greenblatt, Ryan, et al.
Publicado: (2024)
Applying sparse autoencoders to unlearn knowledge in language models
por: Farrell, Eoin, et al.
Publicado: (2024)
por: Farrell, Eoin, et al.
Publicado: (2024)
Auxiliary task discovery through generate-and-test
por: Rafiee, Banafsheh, et al.
Publicado: (2022)
por: Rafiee, Banafsheh, et al.
Publicado: (2022)
Time series causal discovery with variable lags
por: Petrungaro, Bruno, et al.
Publicado: (2026)
por: Petrungaro, Bruno, et al.
Publicado: (2026)
Nonparametric Distribution Regression Re-calibration
por: Jung, Ádám, et al.
Publicado: (2026)
por: Jung, Ádám, et al.
Publicado: (2026)
Fine-tuning LLaMA 2 interference: a comparative study of language implementations for optimal efficiency
por: Hossain, Sazzad, et al.
Publicado: (2025)
por: Hossain, Sazzad, et al.
Publicado: (2025)
Block removal for large language models through constrained binary optimization
por: Jansen, David, et al.
Publicado: (2026)
por: Jansen, David, et al.
Publicado: (2026)
Measuring multi-calibration
por: Guy, Ido, et al.
Publicado: (2025)
por: Guy, Ido, et al.
Publicado: (2025)
The language of time: a language model perspective on time-series foundation models
por: Xie, Yi, et al.
Publicado: (2025)
por: Xie, Yi, et al.
Publicado: (2025)
IDEQ: an improved diffusion model for the TSP
por: Basson, Mickael, et al.
Publicado: (2024)
por: Basson, Mickael, et al.
Publicado: (2024)
Extending confidence calibration to generalised measures of variation
por: Thompson, Andrew, et al.
Publicado: (2026)
por: Thompson, Andrew, et al.
Publicado: (2026)
Ejemplares similares
-
Gradient Guided Hypotheses: A unified solution to enable machine learning models on scarce and noisy data regimes
por: Neves, Paulo, et al.
Publicado: (2024) -
Lookup multivariate Kolmogorov-Arnold Networks
por: Pozdnyakov, Sergey, et al.
Publicado: (2025) -
SynthStrategy: Extracting and Formalizing Latent Strategic Insights from LLMs in Organic Chemistry
por: Armstrong, Daniel, et al.
Publicado: (2025) -
Aviary: training language agents on challenging scientific tasks
por: Narayanan, Siddharth, et al.
Publicado: (2024) -
Quantifying construct validity in large language model evaluations
por: Kearns, Ryan Othniel
Publicado: (2026)