Health-SCORE: Towards Scalable Rubrics for Improving Health-LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhichao, Janghorbani, Sepehr, Zhang, Dongxu, Han, Jun, Qian, Qian, Ressler II, Andrew, Lyng, Gregory D., Batra, Sanjit Singh, Tillman, Robert E. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast and Effective On-policy Distillation from Reasoning Prefixes
by: Zhang, Dongxu, et al.
Published: (2026)
by: Zhang, Dongxu, et al.
Published: (2026)
MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction
by: Yang, Zhichao, et al.
Published: (2026)
by: Yang, Zhichao, et al.
Published: (2026)
ESTAR: Early-Stopping Token-Aware Reasoning For Efficient Inference
by: Wang, Junda, et al.
Published: (2026)
by: Wang, Junda, et al.
Published: (2026)
MedFabric and EtHER: A Data-Centric Framework for Word-Level Fabrication Generation and Detection in Medical LLMs
by: Kwok, Tung Sum Thomas, et al.
Published: (2026)
by: Kwok, Tung Sum Thomas, et al.
Published: (2026)
Imputation of Unknown Missingness in Sparse Electronic Health Records
by: Han, Jun, et al.
Published: (2026)
by: Han, Jun, et al.
Published: (2026)
POET: Protocol Optimization via Eligibility Tuning
by: Das, Trisha, et al.
Published: (2026)
by: Das, Trisha, et al.
Published: (2026)
Health Education and Promotion for Minorities, No. 88-20. Current Bibliographies in Medicine.
by: Tillman, Peggie S.
Published: (1988)
by: Tillman, Peggie S.
Published: (1988)
OpenRubrics: Towards Scalable Synthetic Rubric Generation for Reward Modeling and LLM Alignment
by: Liu, Tianci, et al.
Published: (2025)
by: Liu, Tianci, et al.
Published: (2025)
Guided Discrete Diffusion for Electronic Health Record Generation
by: Han, Jun, et al.
Published: (2024)
by: Han, Jun, et al.
Published: (2024)
Organic aerosol-water interactions
by: Prisle, Nønne Lyng
Published: (2009)
by: Prisle, Nønne Lyng
Published: (2009)
Beyond the Rubric: Cultural Misalignment in LLM Benchmarks for Sexual and Reproductive Health
by: Dey, Sumon Kanti, et al.
Published: (2025)
by: Dey, Sumon Kanti, et al.
Published: (2025)
La patata original
by: Ressler, Katherine
Published: (2006)
by: Ressler, Katherine
Published: (2006)
Impact of Cyberchondria on Unverified Health Information Sharing: A Moderated Mediation Approach
by: Qian Xiao, et al.
Published: (2025)
by: Qian Xiao, et al.
Published: (2025)
Top Down versus Bottom Up: The Social Construction of the Health Literacy Movement
by: Huber, Jeffrey T., et al.
Published: (2012)
by: Huber, Jeffrey T., et al.
Published: (2012)
HealthBench: Evaluating Large Language Models Towards Improved Human Health
by: Arora, Rahul K., et al.
Published: (2025)
by: Arora, Rahul K., et al.
Published: (2025)
METHOD: Modular Efficient Transformer for Health Outcome Discovery
by: Qian, Linglong, et al.
Published: (2025)
by: Qian, Linglong, et al.
Published: (2025)
Improving Learning Through Assessment Rubrics
by: Gonsalves, Chahna, et al.
Published: (2024)
by: Gonsalves, Chahna, et al.
Published: (2024)
The Interprofessional Collaborator Assessment Rubric: A Construct Validity Study Among Allied Health Professionals
by: Justin Weppner
Published: (2025)
by: Justin Weppner
Published: (2025)
EHRMamba: Towards Generalizable and Scalable Foundation Models for Electronic Health Records
by: Fallahpour, Adibvafa, et al.
Published: (2024)
by: Fallahpour, Adibvafa, et al.
Published: (2024)
Healthful Diet and Nutritional Food as a Preventive and Interventional Paradigm in the Face of Microplastic and Nanoplastic Crisis
by: Hongkang Zhu, et al.
Published: (2025)
by: Hongkang Zhu, et al.
Published: (2025)
ArgenSCORE versus EuroSCORE II
by: Raúl A. Borracci
Published: (2014)
by: Raúl A. Borracci
Published: (2014)
ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning
by: Li, Xiaoyuan, et al.
Published: (2026)
by: Li, Xiaoyuan, et al.
Published: (2026)
SynthAgent: A Multi-Agent LLM Framework for Realistic Patient Simulation -- A Case Study in Obesity with Mental Health Comorbidities
by: Aghaee, Arman, et al.
Published: (2026)
by: Aghaee, Arman, et al.
Published: (2026)
RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
by: Li, Gaotang, et al.
Published: (2026)
by: Li, Gaotang, et al.
Published: (2026)
Learning to Judge: LLMs Designing and Applying Evaluation Rubrics
by: Siro, Clemencia, et al.
Published: (2026)
by: Siro, Clemencia, et al.
Published: (2026)
Perceptions of homelessness: Is there variation across medical careers and specialties?
by: David Bronstein, et al.
Published: (2024)
by: David Bronstein, et al.
Published: (2024)
Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation
by: Jang, Won Seok, et al.
Published: (2025)
by: Jang, Won Seok, et al.
Published: (2025)
Attribution of responsibility by Spanish and English speakers: How native language affects our social judgments
by: Richard Tillman
Published: (2013)
by: Richard Tillman
Published: (2013)
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
by: Dhole, Kaustubh D., et al.
Published: (2026)
by: Dhole, Kaustubh D., et al.
Published: (2026)
Research on Ecological and Environmental Water Requirement Threshold for Urban Rivers Based on River Health
by: QianQian Cao, et al.
Published: (2026)
by: QianQian Cao, et al.
Published: (2026)
Mapping Health Pathways: A Network Analysis for Improved Illness Prediction
by: Ankur Kumar Singhal, et al.
Published: (2024)
by: Ankur Kumar Singhal, et al.
Published: (2024)
Shock in Adult‐Onset Still’s Disease Complicated by Macrophage Activation Syndrome: Clinical Characteristics and Prognosis of 14 Patients
by: Dongxu Li, et al.
Published: (2025)
by: Dongxu Li, et al.
Published: (2025)
A Scalable FPGA Architecture for Quantum Computing Simulation
by: Belfore II, Lee A.
Published: (2024)
by: Belfore II, Lee A.
Published: (2024)
Ensemble BERT for Medication Event Classification on Electronic Health Records (EHRs)
by: Sarker, Shouvon, et al.
Published: (2025)
by: Sarker, Shouvon, et al.
Published: (2025)
Engagement-Optimized Care: When LLMs become Mental Health Infrastructure
by: Vecchione, Briana, et al.
Published: (2026)
by: Vecchione, Briana, et al.
Published: (2026)
Rubrics.
by: Callison, Daniel
Published: (2000)
by: Callison, Daniel
Published: (2000)
Kidney Beans ( Phaseolus vulgaris L.): Composition, Health Benefits, and Health‐Oriented Processing Strategies
by: Yixuan Yan, et al.
Published: (2026)
by: Yixuan Yan, et al.
Published: (2026)
An Efficient Rubric-based Generative Verifier for Search-Augmented LLMs
by: Ma, Linyue, et al.
Published: (2025)
by: Ma, Linyue, et al.
Published: (2025)
Device Context Protocol: A Compact, Safety-First Architecture for LLM-Driven Control of Constrained Devices
by: Yang, Dongxu
Published: (2026)
by: Yang, Dongxu
Published: (2026)
A Study of the Access to the Scholarly Record From a Hospital Health Science Core Collection.
by: Williams, James F., II, et al.
Published: (1970)
by: Williams, James F., II, et al.
Published: (1970)
Similar Items
-
Fast and Effective On-policy Distillation from Reasoning Prefixes
by: Zhang, Dongxu, et al.
Published: (2026) -
MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction
by: Yang, Zhichao, et al.
Published: (2026) -
ESTAR: Early-Stopping Token-Aware Reasoning For Efficient Inference
by: Wang, Junda, et al.
Published: (2026) -
MedFabric and EtHER: A Data-Centric Framework for Word-Level Fabrication Generation and Detection in Medical LLMs
by: Kwok, Tung Sum Thomas, et al.
Published: (2026) -
Imputation of Unknown Missingness in Sparse Electronic Health Records
by: Han, Jun, et al.
Published: (2026)