Saved in:
| Main Authors: | Rosenbaum, Gabriel R., Jiang, Lavender Yao, Sheth, Ivaxi, Stryker, Jaden, Alyakin, Anton, Alber, Daniel Alexander, Goff, Nicolas K., Kwon, Young Joon Fred, Markert, John, Nasir-Moin, Mustafa, Niehues, Jan Moritz, Sangwon, Karl L., Yang, Eunice, Oermann, Eric Karl |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.10982 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedMobile: A mobile-sized language model with clinical capabilities
by: Vishwanath, Krithik, et al.
Published: (2024)
by: Vishwanath, Krithik, et al.
Published: (2024)
It is Too Many Options: Pitfalls of Multiple-Choice Questions in Generative AI and Medical Education
by: Singh, Shrutika, et al.
Published: (2025)
by: Singh, Shrutika, et al.
Published: (2025)
BPQA Dataset: Evaluating How Well Language Models Leverage Blood Pressures to Answer Biomedical Questions
by: Hang, Chi, et al.
Published: (2025)
by: Hang, Chi, et al.
Published: (2025)
Evaluating the performance and fragility of large language models on the self-assessment for neurological surgeons
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
Generalist Large Language Models Outperform Clinical Tools on Medical Benchmarks
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
Generalist Foundation Models Are Not Clinical Enough for Hospital Operations
by: Jiang, Lavender Y., et al.
Published: (2025)
by: Jiang, Lavender Y., et al.
Published: (2025)
Medical large language models are easily distracted
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
ProtocolLLM: RTL Benchmark for SystemVerilog Generation of Communication Protocols
by: Sheth, Arnav, et al.
Published: (2025)
by: Sheth, Arnav, et al.
Published: (2025)
Refining Packing and Shuffling Strategies for Enhanced Performance in Generative Language Models
by: Chen, Yanbing, et al.
Published: (2024)
by: Chen, Yanbing, et al.
Published: (2024)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
by: Rahman, Salman, et al.
Published: (2024)
by: Rahman, Salman, et al.
Published: (2024)
CausalGraph2LLM: Evaluating LLMs for Causal Queries
by: Sheth, Ivaxi, et al.
Published: (2024)
by: Sheth, Ivaxi, et al.
Published: (2024)
Context-Aware Reasoning On Parametric Knowledge for Inferring Causal Variables
by: Sheth, Ivaxi, et al.
Published: (2024)
by: Sheth, Ivaxi, et al.
Published: (2024)
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
by: Dua, Radhika, et al.
Published: (2025)
by: Dua, Radhika, et al.
Published: (2025)
CNS-Obsidian: A Neurosurgical Vision-Language Model Built From Scientific Publications
by: Alyakin, Anton, et al.
Published: (2025)
by: Alyakin, Anton, et al.
Published: (2025)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
Large Language Models Predict Functional Outcomes after Acute Ischemic Stroke
by: Kapoor, Anjali K., et al.
Published: (2026)
by: Kapoor, Anjali K., et al.
Published: (2026)
On the Relationship Between the Choice of Representation and In-Context Learning
by: Marinescu, Ioana, et al.
Published: (2025)
by: Marinescu, Ioana, et al.
Published: (2025)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
by: Jiang, Lavender Y., et al.
Published: (2026)
by: Jiang, Lavender Y., et al.
Published: (2026)
"All You Need" is Not All You Need for a Paper Title: On the Origins of a Scientific Meme
by: Alyakin, Anton
Published: (2025)
by: Alyakin, Anton
Published: (2025)
IV Co-Scientist: Multi-Agent LLM Framework for Causal Instrumental Variable Discovery
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
by: Labroo, Arya, et al.
Published: (2026)
by: Labroo, Arya, et al.
Published: (2026)
LLM Task Interference: An Initial Study on the Impact of Task-Switch in Conversational History
by: Gupta, Akash, et al.
Published: (2024)
by: Gupta, Akash, et al.
Published: (2024)
Safety Must Precede the Deployment of Open-Ended AI
by: Sheth, Ivaxi, et al.
Published: (2025)
by: Sheth, Ivaxi, et al.
Published: (2025)
Zwei Weltkriege überleben und alt werden
by: Karl, Fred
Published: (2025)
by: Karl, Fred
Published: (2025)
A spinor proof of the classification of stable minimal surfaces in $\mathbb{R}^3$
by: Stryker, Douglas
Published: (2026)
by: Stryker, Douglas
Published: (2026)
Stable 2-systole bounds in positive scalar curvature
by: Stryker, Douglas
Published: (2026)
by: Stryker, Douglas
Published: (2026)
Min-max construction of prescribed mean curvature hypersurfaces in noncompact manifolds
by: Stryker, Douglas
Published: (2024)
by: Stryker, Douglas
Published: (2024)
Lebensgeschichten alter Eltern kognitiv beeinträchtigter Menschen
by: Oermann, Lisa
Published: (2023)
by: Oermann, Lisa
Published: (2023)
PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?
by: Pulipaka, Sidharth, et al.
Published: (2026)
by: Pulipaka, Sidharth, et al.
Published: (2026)
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
by: Binkyte, Ruta, et al.
Published: (2025)
by: Binkyte, Ruta, et al.
Published: (2025)
Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution
by: Binkyte, Ruta, et al.
Published: (2026)
by: Binkyte, Ruta, et al.
Published: (2026)
Survey on AI Ethics: A Socio-technical Perspective
by: Mbiazi, Dave, et al.
Published: (2023)
by: Mbiazi, Dave, et al.
Published: (2023)
Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews
by: Vasu, Sai Suresh Macharla, et al.
Published: (2025)
by: Vasu, Sai Suresh Macharla, et al.
Published: (2025)
Shearing approach to gauge-invariant Trotterization
by: Stryker, Jesse R.
Published: (2021)
by: Stryker, Jesse R.
Published: (2021)
New Trends in English Education: Selected Addresses Delivered at the Conference on English Education (4th, Carnegie Institute of Technology, March 31, April 1, 2, 1966).
by: Stryker, David, Ed.
Published: (1966)
by: Stryker, David, Ed.
Published: (1966)
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
by: Pulipaka, Sidharth, et al.
Published: (2026)
by: Pulipaka, Sidharth, et al.
Published: (2026)
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
Gottsche, Karl M. Nov. 3, 1869
by: Gottsche, Karl Moritz
Published: (1869)
by: Gottsche, Karl Moritz
Published: (1869)
Nachträge zur Revision der Turbellarien
by: Diesing, Karl Moritz
Published: (1863)
by: Diesing, Karl Moritz
Published: (1863)
Optimal regularity for minimizers of the prescribed mean curvature functional over isotopies
by: Sarnataro, Lorenzo, et al.
Published: (2023)
by: Sarnataro, Lorenzo, et al.
Published: (2023)
Similar Items
-
MedMobile: A mobile-sized language model with clinical capabilities
by: Vishwanath, Krithik, et al.
Published: (2024) -
It is Too Many Options: Pitfalls of Multiple-Choice Questions in Generative AI and Medical Education
by: Singh, Shrutika, et al.
Published: (2025) -
BPQA Dataset: Evaluating How Well Language Models Leverage Blood Pressures to Answer Biomedical Questions
by: Hang, Chi, et al.
Published: (2025) -
Evaluating the performance and fragility of large language models on the self-assessment for neurological surgeons
by: Vishwanath, Krithik, et al.
Published: (2025) -
Generalist Large Language Models Outperform Clinical Tools on Medical Benchmarks
by: Vishwanath, Krithik, et al.
Published: (2025)