DM-Bench: Benchmarking LLMs for Personalized Decision Making in Diabetes Management
Fuente:
arXiv
Saved in:
| Main Authors: | Cardei, Maria Ana, Lamp, Josephine, Derdzinski, Mark, Bhatia, Karan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Building Interpretable Models for Moral Decision-Making
by: Goel, Mayank, et al.
Published: (2026)
by: Goel, Mayank, et al.
Published: (2026)
Algorithmic Fairness in AI Surrogates for End-of-Life Decision-Making
by: Ahmad, Muhammad Aurangzeb
Published: (2025)
by: Ahmad, Muhammad Aurangzeb
Published: (2025)
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
by: Alamdari, Parand A., et al.
Published: (2023)
by: Alamdari, Parand A., et al.
Published: (2023)
Decision Making with Differential Privacy under a Fairness Lens
by: Fioretto, Ferdinando, et al.
Published: (2021)
by: Fioretto, Ferdinando, et al.
Published: (2021)
Integrating Reason-Based Moral Decision-Making in the Reinforcement Learning Architecture
by: Dargasz, Lisa
Published: (2025)
by: Dargasz, Lisa
Published: (2025)
NYU CTF Bench: A Scalable Open-Source Benchmark Dataset for Evaluating LLMs in Offensive Security
by: Shao, Minghao, et al.
Published: (2024)
by: Shao, Minghao, et al.
Published: (2024)
Reinforced Sequential Decision-Making for Sepsis Treatment: The POSNEGDM Framework with Mortality Classifier and Transformer
by: Tamboli, Dipesh, et al.
Published: (2024)
by: Tamboli, Dipesh, et al.
Published: (2024)
Language Agents as Digital Representatives in Collective Decision-Making
by: Jarrett, Daniel, et al.
Published: (2025)
by: Jarrett, Daniel, et al.
Published: (2025)
HypoBench: Towards Systematic and Principled Benchmarking for Hypothesis Generation
by: Liu, Haokun, et al.
Published: (2025)
by: Liu, Haokun, et al.
Published: (2025)
Decision-Making Behavior Evaluation Framework for LLMs under Uncertain Context
by: Jia, Jingru, et al.
Published: (2024)
by: Jia, Jingru, et al.
Published: (2024)
Learning Personalized Decision Support Policies
by: Bhatt, Umang, et al.
Published: (2023)
by: Bhatt, Umang, et al.
Published: (2023)
Implementation of Big Data Analytics for Diabetes Management: Needs Assessment in the Rwanda Healthcare System
by: Majyambere, Silas, et al.
Published: (2026)
by: Majyambere, Silas, et al.
Published: (2026)
From Perceptions to Decisions: Wildfire Evacuation Decision Prediction with Behavioral Theory-informed LLMs
by: Chen, Ruxiao, et al.
Published: (2025)
by: Chen, Ruxiao, et al.
Published: (2025)
Personalized Decision Modeling: Utility Optimization or Textualized-Symbolic Reasoning
by: Zhao, Yibo, et al.
Published: (2025)
by: Zhao, Yibo, et al.
Published: (2025)
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
by: Hu, Tiancheng, et al.
Published: (2025)
by: Hu, Tiancheng, et al.
Published: (2025)
BiasFreeBench: a Benchmark for Mitigating Bias in Large Language Model Responses
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
A Frank System for Co-Evolutionary Hybrid Decision-Making
by: Mazzoni, Federico, et al.
Published: (2025)
by: Mazzoni, Federico, et al.
Published: (2025)
The Personality Illusion: Revealing Dissociation Between Self-Reports & Behavior in LLMs
by: Han, Pengrui, et al.
Published: (2025)
by: Han, Pengrui, et al.
Published: (2025)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
Computational Basis of LLM's Decision Making in Social Simulation
by: Ma, Ji
Published: (2025)
by: Ma, Ji
Published: (2025)
TokenPowerBench: Benchmarking the Power Consumption of LLM Inference
by: Niu, Chenxu, et al.
Published: (2025)
by: Niu, Chenxu, et al.
Published: (2025)
WARBENCH: A Comprehensive Benchmark for Evaluating LLMs in Military Decision-Making
by: Li, Zongjie, et al.
Published: (2026)
by: Li, Zongjie, et al.
Published: (2026)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
by: Li, Manling, et al.
Published: (2024)
by: Li, Manling, et al.
Published: (2024)
Artificial Intelligence Should Genuinely Support Clinical Reasoning and Decision Making To Bridge the Translational Gap
by: Sokol, Kacper, et al.
Published: (2025)
by: Sokol, Kacper, et al.
Published: (2025)
New-Onset Diabetes Assessment Using Artificial Intelligence-Enhanced Electrocardiography
by: Zhang, Hao, et al.
Published: (2022)
by: Zhang, Hao, et al.
Published: (2022)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
CAMEL-Bench: A Comprehensive Arabic LMM Benchmark
by: Ghaboura, Sara, et al.
Published: (2024)
by: Ghaboura, Sara, et al.
Published: (2024)
Two Types of AI Existential Risk: Decisive and Accumulative
by: Kasirzadeh, Atoosa
Published: (2024)
by: Kasirzadeh, Atoosa
Published: (2024)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
by: Marius, Dumitran Adrian, et al.
Published: (2025)
by: Marius, Dumitran Adrian, et al.
Published: (2025)
NudgeRank: Digital Algorithmic Nudging for Personalized Health
by: Chiam, Jodi, et al.
Published: (2024)
by: Chiam, Jodi, et al.
Published: (2024)
Safeguarding Autonomy: a Focus on Machine Learning Decision Systems
by: Subías-Beltrán, Paula, et al.
Published: (2025)
by: Subías-Beltrán, Paula, et al.
Published: (2025)
Inducing Group Fairness in Prompt-Based Language Model Decisions
by: Atwood, James, et al.
Published: (2024)
by: Atwood, James, et al.
Published: (2024)
Deprecating Benchmarks: Criteria and Framework
by: Joaquin, Ayrton San, et al.
Published: (2025)
by: Joaquin, Ayrton San, et al.
Published: (2025)
Explainable AI Systems Must Be Contestable: Here's How to Make It Happen
by: Moreira, Catarina, et al.
Published: (2025)
by: Moreira, Catarina, et al.
Published: (2025)
Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems
by: Wilms, Mieke, et al.
Published: (2026)
by: Wilms, Mieke, et al.
Published: (2026)
AI-powered Digital Framework for Personalized Economical Quality Learning at Scale
by: VatandoustMohammadieh, Mrzieh, et al.
Published: (2024)
by: VatandoustMohammadieh, Mrzieh, et al.
Published: (2024)
EigenBench: A Comparative Behavioral Measure of Value Alignment
by: Chang, Jonathn, et al.
Published: (2025)
by: Chang, Jonathn, et al.
Published: (2025)
Comparative Analysis of Multi-Agent Reinforcement Learning Policies for Crop Planning Decision Support
by: Mahajan, Anubha, et al.
Published: (2024)
by: Mahajan, Anubha, et al.
Published: (2024)
Personalized Knowledge Tracing through Student Representation Reconstruction and Class Imbalance Mitigation
by: Chen, Zhiyu, et al.
Published: (2024)
by: Chen, Zhiyu, et al.
Published: (2024)
Enhancing Deep Knowledge Tracing via Diffusion Models for Personalized Adaptive Learning
by: Kuo, Ming, et al.
Published: (2024)
by: Kuo, Ming, et al.
Published: (2024)
Similar Items
-
Building Interpretable Models for Moral Decision-Making
by: Goel, Mayank, et al.
Published: (2026) -
Algorithmic Fairness in AI Surrogates for End-of-Life Decision-Making
by: Ahmad, Muhammad Aurangzeb
Published: (2025) -
Remembering to Be Fair: Non-Markovian Fairness in Sequential Decision Making
by: Alamdari, Parand A., et al.
Published: (2023) -
Decision Making with Differential Privacy under a Fairness Lens
by: Fioretto, Ferdinando, et al.
Published: (2021) -
Integrating Reason-Based Moral Decision-Making in the Reinforcement Learning Architecture
by: Dargasz, Lisa
Published: (2025)