Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?
Fuente:
arXiv
Salvato in:
| Autori principali: | Seah, Natalie, Bitterman, Danielle S., Spiegel, Daphna, Hartvigsen, Thomas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
KScope: A Framework for Characterizing the Knowledge Status of Language Models
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
di: Xiao, Yuxin, et al.
Pubblicazione: (2025)
Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
di: Chen, Shan, et al.
Pubblicazione: (2024)
di: Chen, Shan, et al.
Pubblicazione: (2024)
Sparse Autoencoder Features for Classifications and Transferability
di: Gallifant, Jack, et al.
Pubblicazione: (2025)
di: Gallifant, Jack, et al.
Pubblicazione: (2025)
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
di: Gallifant, Jack, et al.
Pubblicazione: (2024)
di: Gallifant, Jack, et al.
Pubblicazione: (2024)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
MedBrowseComp: Benchmarking Medical Deep Research and Computer Use
di: Chen, Shan, et al.
Pubblicazione: (2025)
di: Chen, Shan, et al.
Pubblicazione: (2025)
TAXI: Evaluating Categorical Knowledge Editing for Language Models
di: Powell, Derek, et al.
Pubblicazione: (2024)
di: Powell, Derek, et al.
Pubblicazione: (2024)
When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
di: Qi, Jirui, et al.
Pubblicazione: (2025)
di: Qi, Jirui, et al.
Pubblicazione: (2025)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
Continually Self-Improving Language Models for Bariatric Surgery Question--Answering
di: Atri, Yash Kumar, et al.
Pubblicazione: (2025)
di: Atri, Yash Kumar, et al.
Pubblicazione: (2025)
ClinicalBench: Can LLMs Beat Traditional ML Models in Clinical Prediction?
di: Chen, Canyu, et al.
Pubblicazione: (2024)
di: Chen, Canyu, et al.
Pubblicazione: (2024)
Large Language Models to Identify Social Determinants of Health in Electronic Health Records
di: Guevara, Marco, et al.
Pubblicazione: (2023)
di: Guevara, Marco, et al.
Pubblicazione: (2023)
Model Editing with Graph-Based External Memory
di: Atri, Yash Kumar, et al.
Pubblicazione: (2025)
di: Atri, Yash Kumar, et al.
Pubblicazione: (2025)
Can Large Language Models Identify Authorship?
di: Huang, Baixiang, et al.
Pubblicazione: (2024)
di: Huang, Baixiang, et al.
Pubblicazione: (2024)
Evaluating Temporal Consistency in Multi-Turn Language Models
di: Atri, Yash Kumar, et al.
Pubblicazione: (2026)
di: Atri, Yash Kumar, et al.
Pubblicazione: (2026)
Improving Clinical NLP Performance through Language Model-Generated Synthetic Clinical Data
di: Chen, Shan, et al.
Pubblicazione: (2024)
di: Chen, Shan, et al.
Pubblicazione: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
di: Christ, Bryan R., et al.
Pubblicazione: (2024)
di: Christ, Bryan R., et al.
Pubblicazione: (2024)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
di: Jain, Devansh, et al.
Pubblicazione: (2024)
di: Jain, Devansh, et al.
Pubblicazione: (2024)
Labeling Free-text Data using Language Model Ensembles
di: Qiu, Jiaxing, et al.
Pubblicazione: (2025)
di: Qiu, Jiaxing, et al.
Pubblicazione: (2025)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
Medical Large Language Model Benchmarks Should Prioritize Construct Validity
di: Alaa, Ahmed, et al.
Pubblicazione: (2025)
di: Alaa, Ahmed, et al.
Pubblicazione: (2025)
Language Models Still Struggle to Zero-shot Reason about Time Series
di: Merrill, Mike A., et al.
Pubblicazione: (2024)
di: Merrill, Mike A., et al.
Pubblicazione: (2024)
Can Large Language Models Identify Implicit Suicidal Ideation? An Empirical Evaluation
di: Li, Tong, et al.
Pubblicazione: (2025)
di: Li, Tong, et al.
Pubblicazione: (2025)
A Large Language Model Pipeline for Breast Cancer Oncology
di: Pool, Tristen, et al.
Pubblicazione: (2024)
di: Pool, Tristen, et al.
Pubblicazione: (2024)
Decoding the Rule Book: Extracting Hidden Moderation Criteria from Reddit Communities
di: Kim, Youngwoo, et al.
Pubblicazione: (2025)
di: Kim, Youngwoo, et al.
Pubblicazione: (2025)
SpecDiff-2: Scaling Diffusion Drafter Alignment For Faster Speculative Decoding
di: Sandler, Jameson, et al.
Pubblicazione: (2025)
di: Sandler, Jameson, et al.
Pubblicazione: (2025)
PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages
di: Kumar, Priyanshu, et al.
Pubblicazione: (2025)
di: Kumar, Priyanshu, et al.
Pubblicazione: (2025)
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?
di: Gao, Yanjun, et al.
Pubblicazione: (2024)
di: Gao, Yanjun, et al.
Pubblicazione: (2024)
Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
di: Chen, Shan, et al.
Pubblicazione: (2023)
di: Chen, Shan, et al.
Pubblicazione: (2023)
Automatically Finding and Validating Unexpected Side-Effects of Interventions on Language Models
di: Pope, Quintin, et al.
Pubblicazione: (2026)
di: Pope, Quintin, et al.
Pubblicazione: (2026)
Can Humans Identify Domains?
di: Barrett, Maria, et al.
Pubblicazione: (2024)
di: Barrett, Maria, et al.
Pubblicazione: (2024)
Efficient Knowledge Editing via Minimal Precomputation
di: Gupta, Akshat, et al.
Pubblicazione: (2025)
di: Gupta, Akshat, et al.
Pubblicazione: (2025)
Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
di: Gao, Yanjun, et al.
Pubblicazione: (2024)
di: Gao, Yanjun, et al.
Pubblicazione: (2024)
Development of Application-Specific Large Language Models to Facilitate Research Ethics Review
di: Mann, Sebastian Porsdam, et al.
Pubblicazione: (2025)
di: Mann, Sebastian Porsdam, et al.
Pubblicazione: (2025)
Negation: A Pink Elephant in the Large Language Models' Room?
di: Vrabcová, Tereza, et al.
Pubblicazione: (2025)
di: Vrabcová, Tereza, et al.
Pubblicazione: (2025)
"Why" Has the Least Side Effect on Model Editing
di: Pan, Tsung-Hsuan, et al.
Pubblicazione: (2024)
di: Pan, Tsung-Hsuan, et al.
Pubblicazione: (2024)
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
di: Hu, Yujia, et al.
Pubblicazione: (2024)
di: Hu, Yujia, et al.
Pubblicazione: (2024)
Steering Without Side Effects: Improving Post-Deployment Control of Language Models
di: Stickland, Asa Cooper, et al.
Pubblicazione: (2024)
di: Stickland, Asa Cooper, et al.
Pubblicazione: (2024)
Proof of Time: A Benchmark for Evaluating Scientific Idea Judgments
di: Ye, Bingyang, et al.
Pubblicazione: (2026)
di: Ye, Bingyang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
KScope: A Framework for Characterizing the Knowledge Status of Language Models
di: Xiao, Yuxin, et al.
Pubblicazione: (2025) -
Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
di: Chen, Shan, et al.
Pubblicazione: (2024) -
Sparse Autoencoder Features for Classifications and Transferability
di: Gallifant, Jack, et al.
Pubblicazione: (2025) -
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
di: Gallifant, Jack, et al.
Pubblicazione: (2024) -
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)