Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?
Fuente:
arXiv
Guardado en:
| Autores principales: | Seah, Natalie, Bitterman, Danielle S., Spiegel, Daphna, Hartvigsen, Thomas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
KScope: A Framework for Characterizing the Knowledge Status of Language Models
por: Xiao, Yuxin, et al.
Publicado: (2025)
por: Xiao, Yuxin, et al.
Publicado: (2025)
Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
por: Chen, Shan, et al.
Publicado: (2024)
por: Chen, Shan, et al.
Publicado: (2024)
Sparse Autoencoder Features for Classifications and Transferability
por: Gallifant, Jack, et al.
Publicado: (2025)
por: Gallifant, Jack, et al.
Publicado: (2025)
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
por: Gallifant, Jack, et al.
Publicado: (2024)
por: Gallifant, Jack, et al.
Publicado: (2024)
Identifying Implicit Social Biases in Vision-Language Models
por: Hamidieh, Kimia, et al.
Publicado: (2024)
por: Hamidieh, Kimia, et al.
Publicado: (2024)
MedBrowseComp: Benchmarking Medical Deep Research and Computer Use
por: Chen, Shan, et al.
Publicado: (2025)
por: Chen, Shan, et al.
Publicado: (2025)
TAXI: Evaluating Categorical Knowledge Editing for Language Models
por: Powell, Derek, et al.
Publicado: (2024)
por: Powell, Derek, et al.
Publicado: (2024)
When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
por: Qi, Jirui, et al.
Publicado: (2025)
por: Qi, Jirui, et al.
Publicado: (2025)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
por: Sasse, Kuleen, et al.
Publicado: (2024)
por: Sasse, Kuleen, et al.
Publicado: (2024)
Continually Self-Improving Language Models for Bariatric Surgery Question--Answering
por: Atri, Yash Kumar, et al.
Publicado: (2025)
por: Atri, Yash Kumar, et al.
Publicado: (2025)
ClinicalBench: Can LLMs Beat Traditional ML Models in Clinical Prediction?
por: Chen, Canyu, et al.
Publicado: (2024)
por: Chen, Canyu, et al.
Publicado: (2024)
Large Language Models to Identify Social Determinants of Health in Electronic Health Records
por: Guevara, Marco, et al.
Publicado: (2023)
por: Guevara, Marco, et al.
Publicado: (2023)
Model Editing with Graph-Based External Memory
por: Atri, Yash Kumar, et al.
Publicado: (2025)
por: Atri, Yash Kumar, et al.
Publicado: (2025)
Can Large Language Models Identify Authorship?
por: Huang, Baixiang, et al.
Publicado: (2024)
por: Huang, Baixiang, et al.
Publicado: (2024)
Evaluating Temporal Consistency in Multi-Turn Language Models
por: Atri, Yash Kumar, et al.
Publicado: (2026)
por: Atri, Yash Kumar, et al.
Publicado: (2026)
Improving Clinical NLP Performance through Language Model-Generated Synthetic Clinical Data
por: Chen, Shan, et al.
Publicado: (2024)
por: Chen, Shan, et al.
Publicado: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
por: Christ, Bryan R., et al.
Publicado: (2024)
por: Christ, Bryan R., et al.
Publicado: (2024)
PolygloToxicityPrompts: Multilingual Evaluation of Neural Toxic Degeneration in Large Language Models
por: Jain, Devansh, et al.
Publicado: (2024)
por: Jain, Devansh, et al.
Publicado: (2024)
Labeling Free-text Data using Language Model Ensembles
por: Qiu, Jiaxing, et al.
Publicado: (2025)
por: Qiu, Jiaxing, et al.
Publicado: (2025)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
por: Christ, Bryan R, et al.
Publicado: (2024)
por: Christ, Bryan R, et al.
Publicado: (2024)
Medical Large Language Model Benchmarks Should Prioritize Construct Validity
por: Alaa, Ahmed, et al.
Publicado: (2025)
por: Alaa, Ahmed, et al.
Publicado: (2025)
Language Models Still Struggle to Zero-shot Reason about Time Series
por: Merrill, Mike A., et al.
Publicado: (2024)
por: Merrill, Mike A., et al.
Publicado: (2024)
Can Large Language Models Identify Implicit Suicidal Ideation? An Empirical Evaluation
por: Li, Tong, et al.
Publicado: (2025)
por: Li, Tong, et al.
Publicado: (2025)
A Large Language Model Pipeline for Breast Cancer Oncology
por: Pool, Tristen, et al.
Publicado: (2024)
por: Pool, Tristen, et al.
Publicado: (2024)
Decoding the Rule Book: Extracting Hidden Moderation Criteria from Reddit Communities
por: Kim, Youngwoo, et al.
Publicado: (2025)
por: Kim, Youngwoo, et al.
Publicado: (2025)
SpecDiff-2: Scaling Diffusion Drafter Alignment For Faster Speculative Decoding
por: Sandler, Jameson, et al.
Publicado: (2025)
por: Sandler, Jameson, et al.
Publicado: (2025)
PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages
por: Kumar, Priyanshu, et al.
Publicado: (2025)
por: Kumar, Priyanshu, et al.
Publicado: (2025)
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?
por: Gao, Yanjun, et al.
Publicado: (2024)
por: Gao, Yanjun, et al.
Publicado: (2024)
Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
por: Chen, Shan, et al.
Publicado: (2023)
por: Chen, Shan, et al.
Publicado: (2023)
Automatically Finding and Validating Unexpected Side-Effects of Interventions on Language Models
por: Pope, Quintin, et al.
Publicado: (2026)
por: Pope, Quintin, et al.
Publicado: (2026)
Can Humans Identify Domains?
por: Barrett, Maria, et al.
Publicado: (2024)
por: Barrett, Maria, et al.
Publicado: (2024)
Efficient Knowledge Editing via Minimal Precomputation
por: Gupta, Akshat, et al.
Publicado: (2025)
por: Gupta, Akshat, et al.
Publicado: (2025)
Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications
por: Liu, Mingxuan, et al.
Publicado: (2025)
por: Liu, Mingxuan, et al.
Publicado: (2025)
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
por: Gao, Yanjun, et al.
Publicado: (2024)
por: Gao, Yanjun, et al.
Publicado: (2024)
Development of Application-Specific Large Language Models to Facilitate Research Ethics Review
por: Mann, Sebastian Porsdam, et al.
Publicado: (2025)
por: Mann, Sebastian Porsdam, et al.
Publicado: (2025)
Negation: A Pink Elephant in the Large Language Models' Room?
por: Vrabcová, Tereza, et al.
Publicado: (2025)
por: Vrabcová, Tereza, et al.
Publicado: (2025)
"Why" Has the Least Side Effect on Model Editing
por: Pan, Tsung-Hsuan, et al.
Publicado: (2024)
por: Pan, Tsung-Hsuan, et al.
Publicado: (2024)
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
por: Hu, Yujia, et al.
Publicado: (2024)
por: Hu, Yujia, et al.
Publicado: (2024)
Steering Without Side Effects: Improving Post-Deployment Control of Language Models
por: Stickland, Asa Cooper, et al.
Publicado: (2024)
por: Stickland, Asa Cooper, et al.
Publicado: (2024)
Proof of Time: A Benchmark for Evaluating Scientific Idea Judgments
por: Ye, Bingyang, et al.
Publicado: (2026)
por: Ye, Bingyang, et al.
Publicado: (2026)
Ejemplares similares
-
KScope: A Framework for Characterizing the Knowledge Status of Language Models
por: Xiao, Yuxin, et al.
Publicado: (2025) -
Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
por: Chen, Shan, et al.
Publicado: (2024) -
Sparse Autoencoder Features for Classifications and Transferability
por: Gallifant, Jack, et al.
Publicado: (2025) -
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
por: Gallifant, Jack, et al.
Publicado: (2024) -
Identifying Implicit Social Biases in Vision-Language Models
por: Hamidieh, Kimia, et al.
Publicado: (2024)