An Evaluation Benchmark for Adverse Drug Event Prediction from Clinical Trial Results
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yazdani, Anthony, Bornet, Alban, Khlebnikov, Philipp, Zhang, Boya, Rouhizadeh, Hossein, Amini, Poorya, Teodoro, Douglas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HealthContradict: Evaluating Biomedical Knowledge Conflicts in Language Models
von: Zhang, Boya, et al.
Veröffentlicht: (2025)
von: Zhang, Boya, et al.
Veröffentlicht: (2025)
GLiNER-BioMed: A Suite of Efficient Models for Open Biomedical Named Entity Recognition
von: Yazdani, Anthony, et al.
Veröffentlicht: (2025)
von: Yazdani, Anthony, et al.
Veröffentlicht: (2025)
ICU-TSB: A Benchmark for Temporal Patient Representation Learning for Unsupervised Stratification into Patient Cohorts
von: Proios, Dimitrios, et al.
Veröffentlicht: (2025)
von: Proios, Dimitrios, et al.
Veröffentlicht: (2025)
MultiADE: A Multi-domain Benchmark for Adverse Drug Event Extraction
von: Dai, Xiang, et al.
Veröffentlicht: (2024)
von: Dai, Xiang, et al.
Veröffentlicht: (2024)
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
von: MacPhail, Dorothea, et al.
Veröffentlicht: (2024)
von: MacPhail, Dorothea, et al.
Veröffentlicht: (2024)
GASCADE: Grouped Summarization of Adverse Drug Event for Enhanced Cancer Pharmacovigilance
von: Jamil, Sofia, et al.
Veröffentlicht: (2025)
von: Jamil, Sofia, et al.
Veröffentlicht: (2025)
MKG-Rank: Enhancing Large Language Models with Knowledge Graph for Multilingual Medical Question Answering
von: Li, Feiyang, et al.
Veröffentlicht: (2025)
von: Li, Feiyang, et al.
Veröffentlicht: (2025)
Knowledge-augmented Graph Neural Networks with Concept-aware Attention for Adverse Drug Event Detection
von: Ji, Shaoxiong, et al.
Veröffentlicht: (2023)
von: Ji, Shaoxiong, et al.
Veröffentlicht: (2023)
Retrieval-Augmented Generation in Biomedicine: A Survey of Technologies, Datasets, and Clinical Applications
von: He, Jiawei, et al.
Veröffentlicht: (2025)
von: He, Jiawei, et al.
Veröffentlicht: (2025)
Automatically Labeling Clinical Trial Outcomes: A Large-Scale Benchmark for Drug Development
von: Gao, Chufan, et al.
Veröffentlicht: (2024)
von: Gao, Chufan, et al.
Veröffentlicht: (2024)
Detection of Adverse Drug Events in Dutch clinical free text documents using Transformer Models: benchmark study
von: Murphy, Rachel M., et al.
Veröffentlicht: (2025)
von: Murphy, Rachel M., et al.
Veröffentlicht: (2025)
From Bench to Bedside: A Review of Clinical Trials in Drug Discovery and Development
von: Wang, Tianyang, et al.
Veröffentlicht: (2024)
von: Wang, Tianyang, et al.
Veröffentlicht: (2024)
Enhancing Adverse Drug Event Detection with Multimodal Dataset: Corpus Creation and Model Development
von: Sahoo, Pranab, et al.
Veröffentlicht: (2024)
von: Sahoo, Pranab, et al.
Veröffentlicht: (2024)
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity
von: Abdous, Mohammad, et al.
Veröffentlicht: (2023)
von: Abdous, Mohammad, et al.
Veröffentlicht: (2023)
MADE: A Living Benchmark for Multi-Label Text Classification with Uncertainty Quantification of Medical Device Adverse Events
von: Agarwal, Raunak, et al.
Veröffentlicht: (2026)
von: Agarwal, Raunak, et al.
Veröffentlicht: (2026)
LLM-Augmented Therapy Normalization and Aspect-Based Sentiment Analysis for Treatment-Resistant Depression on Reddit
von: Zhu, Yuxin, et al.
Veröffentlicht: (2026)
von: Zhu, Yuxin, et al.
Veröffentlicht: (2026)
Early Risk Stratification of Dosing Errors in Clinical Trials Using Machine Learning
von: Hêche, Félicien, et al.
Veröffentlicht: (2026)
von: Hêche, Félicien, et al.
Veröffentlicht: (2026)
RAGPPI: RAG Benchmark for Protein-Protein Interactions in Drug Discovery
von: Jeon, Youngseung, et al.
Veröffentlicht: (2025)
von: Jeon, Youngseung, et al.
Veröffentlicht: (2025)
CTBench: A Comprehensive Benchmark for Evaluating Language Model Capabilities in Clinical Trial Design
von: Neehal, Nafis, et al.
Veröffentlicht: (2024)
von: Neehal, Nafis, et al.
Veröffentlicht: (2024)
Evaluation of Clinical Trials Reporting Quality using Large Language Models
von: Laï-king, Mathieu, et al.
Veröffentlicht: (2025)
von: Laï-king, Mathieu, et al.
Veröffentlicht: (2025)
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
von: Gallifant, Jack, et al.
Veröffentlicht: (2024)
von: Gallifant, Jack, et al.
Veröffentlicht: (2024)
CLARITY: A Framework and Benchmark for Conversational Language Ambiguity and Unanswerability in Interactive NL2SQL Systems
von: Sarwar, Tabinda, et al.
Veröffentlicht: (2026)
von: Sarwar, Tabinda, et al.
Veröffentlicht: (2026)
Design, Results and Industry Implications of the World's First Insurance Large Language Model Evaluation Benchmark
von: Zhou, Hua, et al.
Veröffentlicht: (2025)
von: Zhou, Hua, et al.
Veröffentlicht: (2025)
Automated PRO-CTCAE Symptom Selection based on Prior Adverse Event Profiles
von: Vandenhende, Francois, et al.
Veröffentlicht: (2025)
von: Vandenhende, Francois, et al.
Veröffentlicht: (2025)
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events
von: Li, Yiming, et al.
Veröffentlicht: (2023)
von: Li, Yiming, et al.
Veröffentlicht: (2023)
CURE: A Multimodal Benchmark for Clinical Understanding and Retrieval Evaluation
von: Gu, Yannian, et al.
Veröffentlicht: (2026)
von: Gu, Yannian, et al.
Veröffentlicht: (2026)
Learning to Describe for Predicting Zero-shot Drug-Drug Interactions
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
Multimodal Clinical Trial Outcome Prediction with Large Language Models
von: Zheng, Wenhao, et al.
Veröffentlicht: (2024)
von: Zheng, Wenhao, et al.
Veröffentlicht: (2024)
Adverse Event Extraction from Discharge Summaries: A New Dataset, Annotation Scheme, and Initial Findings
von: Guellil, Imane, et al.
Veröffentlicht: (2025)
von: Guellil, Imane, et al.
Veröffentlicht: (2025)
FDARxBench: Benchmarking Regulatory and Clinical Reasoning on FDA Generic Drug Assessment
von: Xiong, Betty, et al.
Veröffentlicht: (2026)
von: Xiong, Betty, et al.
Veröffentlicht: (2026)
A Cocktail-Party Benchmark: Multi-Modal dataset and Comparative Evaluation Results
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
$\texttt{AMEND++}$: Benchmarking Eligibility Criteria Amendments in Clinical Trials
von: Das, Trisha, et al.
Veröffentlicht: (2026)
von: Das, Trisha, et al.
Veröffentlicht: (2026)
CTP-LLM: Clinical Trial Phase Transition Prediction Using Large Language Models
von: Reinisch, Michael, et al.
Veröffentlicht: (2024)
von: Reinisch, Michael, et al.
Veröffentlicht: (2024)
TrialEnroll: Predicting Clinical Trial Enrollment Success with Deep & Cross Network and Large Language Models
von: Yue, Ling, et al.
Veröffentlicht: (2024)
von: Yue, Ling, et al.
Veröffentlicht: (2024)
DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering
von: Wang, Qing, et al.
Veröffentlicht: (2026)
von: Wang, Qing, et al.
Veröffentlicht: (2026)
EvidenceOutcomes: a Dataset of Clinical Trial Publications with Clinically Meaningful Outcomes
von: Zhou, Yiliang, et al.
Veröffentlicht: (2025)
von: Zhou, Yiliang, et al.
Veröffentlicht: (2025)
SciEvent: Benchmarking Multi-domain Scientific Event Extraction
von: Dong, Bofu, et al.
Veröffentlicht: (2025)
von: Dong, Bofu, et al.
Veröffentlicht: (2025)
Towards Reproducible LLM Evaluation: Quantifying Uncertainty in LLM Benchmark Scores
von: Blackwell, Robert E., et al.
Veröffentlicht: (2024)
von: Blackwell, Robert E., et al.
Veröffentlicht: (2024)
Improving Extraction of Clinical Event Contextual Properties from Electronic Health Records: A Comparative Study
von: Agarwal, Shubham, et al.
Veröffentlicht: (2024)
von: Agarwal, Shubham, et al.
Veröffentlicht: (2024)
Smart Trial: Evaluating the Use of Large Language Models for Recruiting Clinical Trial Participants via Social Media
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2025)
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HealthContradict: Evaluating Biomedical Knowledge Conflicts in Language Models
von: Zhang, Boya, et al.
Veröffentlicht: (2025) -
GLiNER-BioMed: A Suite of Efficient Models for Open Biomedical Named Entity Recognition
von: Yazdani, Anthony, et al.
Veröffentlicht: (2025) -
ICU-TSB: A Benchmark for Temporal Patient Representation Learning for Unsupervised Stratification into Patient Cohorts
von: Proios, Dimitrios, et al.
Veröffentlicht: (2025) -
MultiADE: A Multi-domain Benchmark for Adverse Drug Event Extraction
von: Dai, Xiang, et al.
Veröffentlicht: (2024) -
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
von: MacPhail, Dorothea, et al.
Veröffentlicht: (2024)