Robust or Suggestible? Exploring Non-Clinical Induction in LLM Drug-Safety Decisions
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Siying, Zhang, Shisheng, Bala, Indu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Analysis of Voluntarily Reported Data Post Mesh Implantation for Detecting Public Emotion and Identifying Concern Reports
di: Bala, Indu, et al.
Pubblicazione: (2025)
di: Bala, Indu, et al.
Pubblicazione: (2025)
Ensembling LLM-Induced Decision Trees for Explainable and Robust Error Detection
di: Wang, Mengqi, et al.
Pubblicazione: (2025)
di: Wang, Mengqi, et al.
Pubblicazione: (2025)
LocalSUG: City-Preference-Enhanced LLM for Query Suggestion in Local-Life Services
di: Chen, Jinwen, et al.
Pubblicazione: (2026)
di: Chen, Jinwen, et al.
Pubblicazione: (2026)
How Value Induction Reshapes LLM Behaviour
di: Arora, Arnav, et al.
Pubblicazione: (2026)
di: Arora, Arnav, et al.
Pubblicazione: (2026)
Dialog Flow Induction for Constrainable LLM-Based Chatbots
di: Agrawal, Stuti, et al.
Pubblicazione: (2024)
di: Agrawal, Stuti, et al.
Pubblicazione: (2024)
In the LLM era, Word Sense Induction remains unsolved
di: Mosolova, Anna, et al.
Pubblicazione: (2026)
di: Mosolova, Anna, et al.
Pubblicazione: (2026)
Active Learning for Robust and Representative LLM Generation in Safety-Critical Scenarios
di: Hassan, Sabit, et al.
Pubblicazione: (2024)
di: Hassan, Sabit, et al.
Pubblicazione: (2024)
Evaluation Drift in LLM Personality Induction: Are We Moving the Goalpost?
di: Rajput, Prateek, et al.
Pubblicazione: (2026)
di: Rajput, Prateek, et al.
Pubblicazione: (2026)
Asking Again and Again: Exploring LLM Robustness to Repeated Questions
di: Shaier, Sagi, et al.
Pubblicazione: (2024)
di: Shaier, Sagi, et al.
Pubblicazione: (2024)
Agent-SafetyBench: Evaluating the Safety of LLM Agents
di: Zhang, Zhexin, et al.
Pubblicazione: (2024)
di: Zhang, Zhexin, et al.
Pubblicazione: (2024)
DriveSafe: A Framework for Risk Detection and Safety Suggestions in Driving Scenarios
di: Artham, Sainithin, et al.
Pubblicazione: (2026)
di: Artham, Sainithin, et al.
Pubblicazione: (2026)
SafetyFlow: An Agent-Flow System for Automated LLM Safety Benchmarking
di: Zhu, Xiangyang, et al.
Pubblicazione: (2025)
di: Zhu, Xiangyang, et al.
Pubblicazione: (2025)
Exploring Safety Alignment Evaluation of LLMs in Chinese Mental Health Dialogues via LLM-as-Judge
di: Cai, Yunna, et al.
Pubblicazione: (2025)
di: Cai, Yunna, et al.
Pubblicazione: (2025)
Sequential LLM Framework for Fashion Recommendation
di: Liu, Han, et al.
Pubblicazione: (2024)
di: Liu, Han, et al.
Pubblicazione: (2024)
ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory
di: Ge, Zhuohan, et al.
Pubblicazione: (2026)
di: Ge, Zhuohan, et al.
Pubblicazione: (2026)
The Better Angels of Machine Personality: How Personality Relates to LLM Safety
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
Evolutionary Computation and Large Language Models: A Survey of Methods, Synergies, and Applications
di: Chauhan, Dikshit, et al.
Pubblicazione: (2025)
di: Chauhan, Dikshit, et al.
Pubblicazione: (2025)
MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications
di: Kanithi, Praveenkumar, et al.
Pubblicazione: (2024)
di: Kanithi, Praveenkumar, et al.
Pubblicazione: (2024)
HalluCounter: Reference-free LLM Hallucination Detection in the Wild!
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
di: Urlana, Ashok, et al.
Pubblicazione: (2025)
FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios
di: Hou, Yutao, et al.
Pubblicazione: (2026)
di: Hou, Yutao, et al.
Pubblicazione: (2026)
From Bench to Bedside: A Review of Clinical Trials in Drug Discovery and Development
di: Wang, Tianyang, et al.
Pubblicazione: (2024)
di: Wang, Tianyang, et al.
Pubblicazione: (2024)
Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties
di: Ong, Jasmine Chiat Ling, et al.
Pubblicazione: (2024)
di: Ong, Jasmine Chiat Ling, et al.
Pubblicazione: (2024)
"Oh LLM, I'm Asking Thee, Please Give Me a Decision Tree": Zero-Shot Decision Tree Induction and Embedding with Large Language Models
di: Knauer, Ricardo, et al.
Pubblicazione: (2024)
di: Knauer, Ricardo, et al.
Pubblicazione: (2024)
METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues
di: Yang, Haofu, et al.
Pubblicazione: (2026)
di: Yang, Haofu, et al.
Pubblicazione: (2026)
LiSA: Lifelong Safety Adaptation via Conservative Policy Induction
di: Kim, Minbeom, et al.
Pubblicazione: (2026)
di: Kim, Minbeom, et al.
Pubblicazione: (2026)
Knowledge Graph Assisted Automatic Sports News Writing
di: Cao, Yang, et al.
Pubblicazione: (2024)
di: Cao, Yang, et al.
Pubblicazione: (2024)
Taxonomy of Comprehensive Safety for Clinical Agents
di: Seo, Jean, et al.
Pubblicazione: (2025)
di: Seo, Jean, et al.
Pubblicazione: (2025)
A Hybrid Supervised-LLM Pipeline for Actionable Suggestion Mining in Unstructured Customer Reviews
di: Trivedi, Aakash, et al.
Pubblicazione: (2026)
di: Trivedi, Aakash, et al.
Pubblicazione: (2026)
Evil Geniuses: Delving into the Safety of LLM-based Agents
di: Tian, Yu, et al.
Pubblicazione: (2023)
di: Tian, Yu, et al.
Pubblicazione: (2023)
GraSP: Graph-Structured Skill Compositions for LLM Agents
di: Xia, Tianle, et al.
Pubblicazione: (2026)
di: Xia, Tianle, et al.
Pubblicazione: (2026)
SAFENLIDB: A Privacy-Preserving Safety Alignment Framework for LLM-based Natural Language Database Interfaces
di: Liu, Ruiheng, et al.
Pubblicazione: (2025)
di: Liu, Ruiheng, et al.
Pubblicazione: (2025)
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM
di: Zhang, Chi, et al.
Pubblicazione: (2025)
di: Zhang, Chi, et al.
Pubblicazione: (2025)
A Coin Flip for Safety: LLM Judges Fail to Reliably Measure Adversarial Robustness
di: Schwinn, Leo, et al.
Pubblicazione: (2026)
di: Schwinn, Leo, et al.
Pubblicazione: (2026)
Leveraging Robust Optimization for LLM Alignment under Distribution Shifts
di: Zhu, Mingye, et al.
Pubblicazione: (2025)
di: Zhu, Mingye, et al.
Pubblicazione: (2025)
An Evaluation Benchmark for Adverse Drug Event Prediction from Clinical Trial Results
di: Yazdani, Anthony, et al.
Pubblicazione: (2024)
di: Yazdani, Anthony, et al.
Pubblicazione: (2024)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
di: Wu, Ya, et al.
Pubblicazione: (2025)
di: Wu, Ya, et al.
Pubblicazione: (2025)
CARES: Comprehensive Evaluation of Safety and Adversarial Robustness in Medical LLMs
di: Chen, Sijia, et al.
Pubblicazione: (2025)
di: Chen, Sijia, et al.
Pubblicazione: (2025)
FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents
di: Jia, Haoxuan, et al.
Pubblicazione: (2026)
di: Jia, Haoxuan, et al.
Pubblicazione: (2026)
LOGOS: LLM-driven End-to-End Grounded Theory Development and Schema Induction for Qualitative Research
di: Pi, Xinyu, et al.
Pubblicazione: (2025)
di: Pi, Xinyu, et al.
Pubblicazione: (2025)
STRUX: An LLM for Decision-Making with Structured Explanations
di: Lu, Yiming, et al.
Pubblicazione: (2024)
di: Lu, Yiming, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Analysis of Voluntarily Reported Data Post Mesh Implantation for Detecting Public Emotion and Identifying Concern Reports
di: Bala, Indu, et al.
Pubblicazione: (2025) -
Ensembling LLM-Induced Decision Trees for Explainable and Robust Error Detection
di: Wang, Mengqi, et al.
Pubblicazione: (2025) -
LocalSUG: City-Preference-Enhanced LLM for Query Suggestion in Local-Life Services
di: Chen, Jinwen, et al.
Pubblicazione: (2026) -
How Value Induction Reshapes LLM Behaviour
di: Arora, Arnav, et al.
Pubblicazione: (2026) -
Dialog Flow Induction for Constrainable LLM-Based Chatbots
di: Agrawal, Stuti, et al.
Pubblicazione: (2024)