Prior-informed optimization of treatment recommendation via bandit algorithms trained on large language model-processed historical records
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Nessari, Saman, Bozorgi-Amiri, Ali |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Benchmarking large language models for biomedical natural language processing applications and recommendations
par: Chen, Qingyu, et autres
Publié: (2023)
par: Chen, Qingyu, et autres
Publié: (2023)
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks
par: Mei, Taiyuan, et autres
Publié: (2024)
par: Mei, Taiyuan, et autres
Publié: (2024)
Adversarial bandit optimization for approximately linear functions
par: Cheng, Zhuoyu, et autres
Publié: (2025)
par: Cheng, Zhuoyu, et autres
Publié: (2025)
Pre-trained knowledge elevates large language models beyond traditional chemical reaction optimizers
par: MacKnight, Robert, et autres
Publié: (2025)
par: MacKnight, Robert, et autres
Publié: (2025)
Spectral bandits
par: Kocák, Tomáš, et autres
Publié: (2026)
par: Kocák, Tomáš, et autres
Publié: (2026)
Post-training makes large language models less human-like
par: Binz, Marcel, et autres
Publié: (2026)
par: Binz, Marcel, et autres
Publié: (2026)
Linear bandits with polylogarithmic minimax regret
par: Lumbreras, Josep, et autres
Publié: (2024)
par: Lumbreras, Josep, et autres
Publié: (2024)
Context information can be more important than reasoning for time series forecasting with a large language model
par: Yang, Janghoon
Publié: (2025)
par: Yang, Janghoon
Publié: (2025)
Quantifying construct validity in large language model evaluations
par: Kearns, Ryan Othniel
Publié: (2026)
par: Kearns, Ryan Othniel
Publié: (2026)
Representation in large language models
par: Yetman, Cameron
Publié: (2025)
par: Yetman, Cameron
Publié: (2025)
Multiview graph dual-attention deep learning and contrastive learning for multi-criteria recommender systems
par: Forouzandeh, Saman, et autres
Publié: (2025)
par: Forouzandeh, Saman, et autres
Publié: (2025)
Reinforcement learning with combinatorial actions for coupled restless bandits
par: Xu, Lily, et autres
Publié: (2025)
par: Xu, Lily, et autres
Publié: (2025)
Block removal for large language models through constrained binary optimization
par: Jansen, David, et autres
Publié: (2026)
par: Jansen, David, et autres
Publié: (2026)
Performance of AI agents based on reasoning language models on ALD process optimization tasks
par: Yanguas-Gil, Angel
Publié: (2026)
par: Yanguas-Gil, Angel
Publié: (2026)
Alignment faking in large language models
par: Greenblatt, Ryan, et autres
Publié: (2024)
par: Greenblatt, Ryan, et autres
Publié: (2024)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
par: Lin, Xiaoqiang, et autres
Publié: (2023)
par: Lin, Xiaoqiang, et autres
Publié: (2023)
Large language models as uncertainty-calibrated optimizers for experimental discovery
par: Ranković, Bojana, et autres
Publié: (2025)
par: Ranković, Bojana, et autres
Publié: (2025)
Long-form factuality in large language models
par: Wei, Jerry, et autres
Publié: (2024)
par: Wei, Jerry, et autres
Publié: (2024)
Can large language models explore in-context?
par: Krishnamurthy, Akshay, et autres
Publié: (2024)
par: Krishnamurthy, Akshay, et autres
Publié: (2024)
A federated large language model for long-term time series forecasting
par: Abdel-Sater, Raed, et autres
Publié: (2024)
par: Abdel-Sater, Raed, et autres
Publié: (2024)
Pretraining large language models with MXFP4 on Native FP4 Hardware
par: Cim, Musa, et autres
Publié: (2026)
par: Cim, Musa, et autres
Publié: (2026)
Logic-informed reinforcement learning for cross-domain optimization of large-scale cyber-physical systems
par: Wan, Guangxi, et autres
Publié: (2025)
par: Wan, Guangxi, et autres
Publié: (2025)
From Simulation to Enaction: Post-trained language models recognize and react to their own generations
par: G., Asvin, et autres
Publié: (2026)
par: G., Asvin, et autres
Publié: (2026)
Text-guided multi-property molecular optimization with a diffusion language model
par: Xiong, Yida, et autres
Publié: (2024)
par: Xiong, Yida, et autres
Publié: (2024)
AI-AI Bias: large language models favor communications generated by large language models
par: Laurito, Walter, et autres
Publié: (2024)
par: Laurito, Walter, et autres
Publié: (2024)
Relational In-Context Learning via Synthetic Pre-training with Structural Prior
par: Wang, Yanbo, et autres
Publié: (2026)
par: Wang, Yanbo, et autres
Publié: (2026)
Beyond designer's knowledge: Generating materials design hypotheses via large language models
par: Liu, Quanliang, et autres
Publié: (2024)
par: Liu, Quanliang, et autres
Publié: (2024)
Comprehensive benchmarking of large language models for RNA secondary structure prediction
par: Zablocki, L. I., et autres
Publié: (2024)
par: Zablocki, L. I., et autres
Publié: (2024)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
par: Byers, Neil, et autres
Publié: (2025)
par: Byers, Neil, et autres
Publié: (2025)
Uncovering mesa-optimization algorithms in Transformers
par: von Oswald, Johannes, et autres
Publié: (2023)
par: von Oswald, Johannes, et autres
Publié: (2023)
Benchmarking pre-trained text embedding models in aligning built asset information
par: Shahinmoghadam, Mehrzad, et autres
Publié: (2024)
par: Shahinmoghadam, Mehrzad, et autres
Publié: (2024)
Are large language models superhuman chemists?
par: Mirza, Adrian, et autres
Publié: (2024)
par: Mirza, Adrian, et autres
Publié: (2024)
Safety challenges of AI in medicine in the era of large language models
par: Wang, Xiaoye, et autres
Publié: (2024)
par: Wang, Xiaoye, et autres
Publié: (2024)
Inducing anxiety in large language models can induce bias
par: Coda-Forno, Julian, et autres
Publié: (2023)
par: Coda-Forno, Julian, et autres
Publié: (2023)
Using large language models for embodied planning introduces systematic safety risks
par: Zhang, Tao, et autres
Publié: (2026)
par: Zhang, Tao, et autres
Publié: (2026)
Fine-grained large-scale content recommendations for MSX sellers
par: Singh, Manpreet, et autres
Publié: (2024)
par: Singh, Manpreet, et autres
Publié: (2024)
Data filtering methods for training language models
par: Shevchenko, Egor, et autres
Publié: (2026)
par: Shevchenko, Egor, et autres
Publié: (2026)
Functional multi-armed bandit and the best function identification problems
par: Dorn, Yuriy, et autres
Publié: (2025)
par: Dorn, Yuriy, et autres
Publié: (2025)
Exploring the limits of strong membership inference attacks on large language models
par: Hayes, Jamie, et autres
Publié: (2025)
par: Hayes, Jamie, et autres
Publié: (2025)
Can large language models replace humans in the systematic review process? Evaluating GPT-4's efficacy in screening and extracting data from peer-reviewed and grey literature in multiple languages
par: Khraisha, Qusai, et autres
Publié: (2023)
par: Khraisha, Qusai, et autres
Publié: (2023)
Documents similaires
-
Benchmarking large language models for biomedical natural language processing applications and recommendations
par: Chen, Qingyu, et autres
Publié: (2023) -
Efficiency optimization of large-scale language models based on deep learning in natural language processing tasks
par: Mei, Taiyuan, et autres
Publié: (2024) -
Adversarial bandit optimization for approximately linear functions
par: Cheng, Zhuoyu, et autres
Publié: (2025) -
Pre-trained knowledge elevates large language models beyond traditional chemical reaction optimizers
par: MacKnight, Robert, et autres
Publié: (2025) -
Spectral bandits
par: Kocák, Tomáš, et autres
Publié: (2026)