Is Biomedical Specialization Still Worth It? Insights from Domain-Adaptive Language Modelling with a New French Health Corpus
Fuente:
arXiv
Saved in:
| Main Authors: | Mannion, Aidan, Macaire, Cécile, Violle, Armand, Ohayon, Stéphane, Tannier, Xavier, Schwab, Didier, Goeuriot, Lorraine, Portet, François |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UMLS-KGI-BERT: Data-Centric Knowledge Integration in Transformers for Biomedical Entity Recognition
by: Mannion, Aidan, et al.
Published: (2023)
by: Mannion, Aidan, et al.
Published: (2023)
LongBEL: Long-Context and Document-Consistent Biomedical Entity Linking
by: Remaki, Adam, et al.
Published: (2026)
by: Remaki, Adam, et al.
Published: (2026)
Few-shot clinical entity recognition in English, French and Spanish: masked language models outperform generative model prompting
by: Naguib, Marco, et al.
Published: (2024)
by: Naguib, Marco, et al.
Published: (2024)
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
by: Le, Phuong-Hang, et al.
Published: (2026)
by: Le, Phuong-Hang, et al.
Published: (2026)
A Benchmark Evaluation of Clinical Named Entity Recognition in French
by: Bannour, Nesrine, et al.
Published: (2024)
by: Bannour, Nesrine, et al.
Published: (2024)
Merging Continual Pretraining Models for Domain-Specialized LLMs: A Case Study in Finance
by: Ueda, Kentaro, et al.
Published: (2025)
by: Ueda, Kentaro, et al.
Published: (2025)
SynCABEL: Synthetic Contextualized Augmentation for Biomedical Entity Linking
by: Remaki, Adam, et al.
Published: (2026)
by: Remaki, Adam, et al.
Published: (2026)
BenCSSmark: Making the Social Sciences Count in LLM Research
by: Chatelain, Arnault, et al.
Published: (2026)
by: Chatelain, Arnault, et al.
Published: (2026)
PARHAF, a human-authored corpus of clinical reports for fictitious patients in French
by: Tannier, Xavier, et al.
Published: (2026)
by: Tannier, Xavier, et al.
Published: (2026)
Clinical trial cohort selection using Large Language Models on n2c2 Challenges
by: Tai, Chi-en Amy, et al.
Published: (2025)
by: Tai, Chi-en Amy, et al.
Published: (2025)
Interactions par franchissement grâce a un système de suivi du regard
by: Riou, Sébastien, et al.
Published: (2025)
by: Riou, Sébastien, et al.
Published: (2025)
Reflexiones Sociocríticas sobre la alteridad. El Otro modelo hegemónico y el tercer interpretante. Ejemplo de caso: entre imagen improbable y figura banalizada. El árabe en la cultura francesa
by: Monique Carcaud-Macaire
Published: (2008)
by: Monique Carcaud-Macaire
Published: (2008)
How Much is Brain Data Worth for Machine Learning?
by: Lewis, Lane, et al.
Published: (2026)
by: Lewis, Lane, et al.
Published: (2026)
mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
by: Servan, Christophe, et al.
Published: (2024)
by: Servan, Christophe, et al.
Published: (2024)
Strategies for improving low resource speech to text translation relying on pre-trained ASR models
by: Kesiraju, Santosh, et al.
Published: (2023)
by: Kesiraju, Santosh, et al.
Published: (2023)
DrBenchmark: A Large Language Understanding Evaluation Benchmark for French Biomedical Domain
by: Labrak, Yanis, et al.
Published: (2024)
by: Labrak, Yanis, et al.
Published: (2024)
Biological Templates for Gold Nanocluster Assembly: Design and Biomedical Applications
by: Zeineb Ayed, et al.
Published: (2026)
by: Zeineb Ayed, et al.
Published: (2026)
CFPR accent metadata
by: FABRE, Diandra, et al.
Published: (2026)
by: FABRE, Diandra, et al.
Published: (2026)
PSentScore: Evaluating Sentiment Polarity in Dialogue Summarization
by: Zhou, Yongxin, et al.
Published: (2023)
by: Zhou, Yongxin, et al.
Published: (2023)
Mechanistic Interpretability as Statistical Estimation: A Variance Analysis
by: Méloux, Maxime, et al.
Published: (2025)
by: Méloux, Maxime, et al.
Published: (2025)
Can GPT models Follow Human Summarization Guidelines? A Study for Targeted Communication Goals
by: Zhou, Yongxin, et al.
Published: (2023)
by: Zhou, Yongxin, et al.
Published: (2023)
Transformer-based Models to Deal with Heterogeneous Environments in Human Activity Recognition
by: EK, Sannara, et al.
Published: (2022)
by: EK, Sannara, et al.
Published: (2022)
Automated Clinical Report Generation for Remote Cognitive Remediation: Comparing Knowledge-Engineered Templates and LLMs in Low-Resource Settings
by: Zhou, Yongxin, et al.
Published: (2026)
by: Zhou, Yongxin, et al.
Published: (2026)
The multiple dimensions of intraspecific variation in seed dispersal
by: Catharina Y. Utami, et al.
Published: (2026)
by: Catharina Y. Utami, et al.
Published: (2026)
A French Version of the OLDI Seed Corpus
by: Marmonier, Malik, et al.
Published: (2025)
by: Marmonier, Malik, et al.
Published: (2025)
FFSTC: Fongbe to French Speech Translation Corpus
by: Kponou, D. Fortune, et al.
Published: (2024)
by: Kponou, D. Fortune, et al.
Published: (2024)
Diagnostic ultrasound in small animal practice / Paddy Mannion
by: Mannion, Paddy
by: Mannion, Paddy
Comparative, Arab, and European Studies: Still a French Exceptionalism?
by: Leca, Jean
Published: (2009)
by: Leca, Jean
Published: (2009)
Cause and Effect Extraction from Biomedical Corpus
by: Sindhuja Gopalan
Published: (2017)
by: Sindhuja Gopalan
Published: (2017)
Prediction of amputation risk of patients with diabetic foot using classification algorithms: A clinical study from a tertiary center
by: Denizhan Demirkol, et al.
Published: (2024)
by: Denizhan Demirkol, et al.
Published: (2024)
MedMeta: A Benchmark for LLMs in Synthesizing Meta-Analysis Conclusion from Medical Studies
by: Ha, Huy Hoang, et al.
Published: (2026)
by: Ha, Huy Hoang, et al.
Published: (2026)
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
by: Parcollet, Titouan, et al.
Published: (2023)
by: Parcollet, Titouan, et al.
Published: (2023)
Critical evaluation of reference charge radii and applications in mirror nuclei
by: Ohayon, Ben
Published: (2024)
by: Ohayon, Ben
Published: (2024)
Iniciação científica: uma metodologia de avaliação
by: Pierre Ohayon
Published: (2007)
by: Pierre Ohayon
Published: (2007)
Recent vegetation shifts in the French Alps with winners outnumbering losers
by: Romain Goury, et al.
Published: (2025)
by: Romain Goury, et al.
Published: (2025)
The Development of Specialized Biomedical Information.
by: Ginn, David S.
Published: (1993)
by: Ginn, David S.
Published: (1993)
TempPerturb-Eval: On the Joint Effects of Internal Temperature and External Perturbations in RAG Robustness
by: Zhou, Yongxin, et al.
Published: (2025)
by: Zhou, Yongxin, et al.
Published: (2025)
Prompt engineering paradigms for medical applications: scoping review and recommendations for better practices
by: Zaghir, Jamil, et al.
Published: (2024)
by: Zaghir, Jamil, et al.
Published: (2024)
Development of the user-friendly decision aid Rule-based Evaluation and Support Tool (REST) for optimizing the resources of an information extraction task
by: Bazin, Guillaume, et al.
Published: (2025)
by: Bazin, Guillaume, et al.
Published: (2025)
From Individual to Stand Performance in Hybrids: Challenging the Optimal Parental Genetic Distance
by: Catharina Y. Utami, et al.
Published: (2025)
by: Catharina Y. Utami, et al.
Published: (2025)
Similar Items
-
UMLS-KGI-BERT: Data-Centric Knowledge Integration in Transformers for Biomedical Entity Recognition
by: Mannion, Aidan, et al.
Published: (2023) -
LongBEL: Long-Context and Document-Consistent Biomedical Entity Linking
by: Remaki, Adam, et al.
Published: (2026) -
Few-shot clinical entity recognition in English, French and Spanish: masked language models outperform generative model prompting
by: Naguib, Marco, et al.
Published: (2024) -
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
by: Le, Phuong-Hang, et al.
Published: (2026) -
A Benchmark Evaluation of Clinical Named Entity Recognition in French
by: Bannour, Nesrine, et al.
Published: (2024)