What Does Neuro Mean to Cardio? Investigating the Role of Clinical Specialty Data in Medical LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Yan, Xinlan, Wu, Di, Lei, Yibin, Monz, Christof, Calixto, Iacer |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Calibrating Translation Decoding with Quality Estimation on LLMs
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Mind the Gap: Benchmarking LLM Uncertainty and Calibration with Specialty-Aware Clinical QA and Reasoning-Based Behavioural Features
by: Testoni, Alberto, et al.
Published: (2025)
by: Testoni, Alberto, et al.
Published: (2025)
Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA
by: Testoni, Alberto, et al.
Published: (2026)
by: Testoni, Alberto, et al.
Published: (2026)
How to Learn in a Noisy World? Self-Correcting the Real-World Data Noise in Machine Translation
by: Meng, Yan, et al.
Published: (2024)
by: Meng, Yan, et al.
Published: (2024)
Differentially Private De-identification of Dutch Clinical Notes: A Comparative Evaluation
by: Miranda, Michele, et al.
Published: (2026)
by: Miranda, Michele, et al.
Published: (2026)
Beyond Shared Vocabulary: Increasing Representational Word Similarities across Languages for Multilingual Machine Translation
by: Wu, Di, et al.
Published: (2023)
by: Wu, Di, et al.
Published: (2023)
Disentangling the Roles of Target-Side Transfer and Regularization in Multilingual Machine Translation
by: Meng, Yan, et al.
Published: (2024)
by: Meng, Yan, et al.
Published: (2024)
Please Translate Again: Two Simple Experiments on Whether Human-Like Reasoning Helps Translation
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Neuron Specialization: Leveraging intrinsic task modularity for multilingual machine translation
by: Tan, Shaomu, et al.
Published: (2024)
by: Tan, Shaomu, et al.
Published: (2024)
MedPath: Multi-Domain Cross-Vocabulary Hierarchical Paths for Biomedical Entity Linking
by: Mishra, Nishant, et al.
Published: (2025)
by: Mishra, Nishant, et al.
Published: (2025)
Can LLMs Really Learn to Translate a Low-Resource Language from One Grammar Book?
by: Aycock, Seth, et al.
Published: (2024)
by: Aycock, Seth, et al.
Published: (2024)
How Far Can 100 Samples Go? Unlocking Overall Zero-Shot Multilingual Translation via Tiny Multi-Parallel Data
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Evaluating Linguistic Capabilities of Multimodal LLMs in the Lens of Few-Shot Learning
by: Dogan, Mustafa, et al.
Published: (2024)
by: Dogan, Mustafa, et al.
Published: (2024)
DeVisE: Behavioral Testing of Medical Large Language Models
by: Tagliabue, Camila Zurdo, et al.
Published: (2025)
by: Tagliabue, Camila Zurdo, et al.
Published: (2025)
Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling
by: Tan, Shaomu, et al.
Published: (2025)
by: Tan, Shaomu, et al.
Published: (2025)
The Effect of Language Diversity When Fine-Tuning Large Language Models for Translation
by: Stap, David, et al.
Published: (2025)
by: Stap, David, et al.
Published: (2025)
IKUN for WMT24 General MT Task: LLMs Are here for Multilingual Machine Translation
by: Liao, Baohao, et al.
Published: (2024)
by: Liao, Baohao, et al.
Published: (2024)
Is It a Free Lunch for Removing Outliers during Pretraining?
by: Liao, Baohao, et al.
Published: (2024)
by: Liao, Baohao, et al.
Published: (2024)
Do Language Models Reason Across Languages?
by: Meng, Yan, et al.
Published: (2026)
by: Meng, Yan, et al.
Published: (2026)
AnyMatch -- Efficient Zero-Shot Entity Matching with a Small Language Model
by: Zhang, Zeyu, et al.
Published: (2024)
by: Zhang, Zeyu, et al.
Published: (2024)
DistillNote: Toward a Functional Evaluation Framework of LLM-Generated Clinical Note Summaries
by: Boll, Heloisa Oss, et al.
Published: (2025)
by: Boll, Heloisa Oss, et al.
Published: (2025)
FewMMBench: A Benchmark for Multimodal Few-Shot Learning
by: Dogan, Mustafa, et al.
Published: (2026)
by: Dogan, Mustafa, et al.
Published: (2026)
Analyzing the Evaluation of Cross-Lingual Knowledge Transfer in Multilingual Language Models
by: Rajaee, Sara, et al.
Published: (2024)
by: Rajaee, Sara, et al.
Published: (2024)
3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability
by: Liao, Baohao, et al.
Published: (2024)
by: Liao, Baohao, et al.
Published: (2024)
When Does Meaning Backfire? Investigating the Role of AMRs in NLI
by: Min, Junghyun, et al.
Published: (2025)
by: Min, Junghyun, et al.
Published: (2025)
Does Less Hallucination Mean Less Creativity? An Empirical Investigation in LLMs
by: Banerjee, Mohor, et al.
Published: (2025)
by: Banerjee, Mohor, et al.
Published: (2025)
When Contextual Inference Fails: Cancelability in Interactive Instruction Following
by: Bila, Natalia, et al.
Published: (2026)
by: Bila, Natalia, et al.
Published: (2026)
LLMs-Healthcare : Current Applications and Challenges of Large Language Models in various Medical Specialties
by: Mumtaz, Ummara, et al.
Published: (2023)
by: Mumtaz, Ummara, et al.
Published: (2023)
Communicating with Speakers and Listeners of Different Pragmatic Levels
by: Naszadi, Kata, et al.
Published: (2024)
by: Naszadi, Kata, et al.
Published: (2024)
The SIFo Benchmark: Investigating the Sequential Instruction Following Ability of Large Language Models
by: Chen, Xinyi, et al.
Published: (2024)
by: Chen, Xinyi, et al.
Published: (2024)
SOLVE-Med: Specialized Orchestration for Leading Vertical Experts across Medical Specialties
by: Di Marino, Roberta, et al.
Published: (2025)
by: Di Marino, Roberta, et al.
Published: (2025)
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
by: Rajaee, Sara, et al.
Published: (2025)
by: Rajaee, Sara, et al.
Published: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
by: Choenni, Rochelle, et al.
Published: (2024)
by: Choenni, Rochelle, et al.
Published: (2024)
ApiQ: Finetuning of 2-Bit Quantized Large Language Model
by: Liao, Baohao, et al.
Published: (2024)
by: Liao, Baohao, et al.
Published: (2024)
Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties
by: Ong, Jasmine Chiat Ling, et al.
Published: (2024)
by: Ong, Jasmine Chiat Ling, et al.
Published: (2024)
Remedy-R: Generative Reasoning for Machine Translation Evaluation without Error Annotations
by: Tan, Shaomu, et al.
Published: (2025)
by: Tan, Shaomu, et al.
Published: (2025)
The Fine-Tuning Paradox: Boosting Translation Quality Without Sacrificing LLM Abilities
by: Stap, David, et al.
Published: (2024)
by: Stap, David, et al.
Published: (2024)
On the Limits of Model Merging for Multilinguality in Pre-Training
by: Aycock, Seth, et al.
Published: (2026)
by: Aycock, Seth, et al.
Published: (2026)
Specialty-Specific Medical Language Model for Immune-Mediated Diseases
by: Kocaman, Veysel, et al.
Published: (2026)
by: Kocaman, Veysel, et al.
Published: (2026)
VALSE: A Task-Independent Benchmark for Vision and Language Models Centered on Linguistic Phenomena
by: Parcalabescu, Letitia, et al.
Published: (2021)
by: Parcalabescu, Letitia, et al.
Published: (2021)
Similar Items
-
Calibrating Translation Decoding with Quality Estimation on LLMs
by: Wu, Di, et al.
Published: (2025) -
Mind the Gap: Benchmarking LLM Uncertainty and Calibration with Specialty-Aware Clinical QA and Reasoning-Based Behavioural Features
by: Testoni, Alberto, et al.
Published: (2025) -
Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA
by: Testoni, Alberto, et al.
Published: (2026) -
How to Learn in a Noisy World? Self-Correcting the Real-World Data Noise in Machine Translation
by: Meng, Yan, et al.
Published: (2024) -
Differentially Private De-identification of Dutch Clinical Notes: A Comparative Evaluation
by: Miranda, Michele, et al.
Published: (2026)