Evaluation of LLMs in Medical Text Summarization: The Role of Vocabulary Adaptation in High OOV Settings
Fuente:
arXiv
Saved in:
| Main Authors: | Balde, Gunjan, Roy, Soumyadeep, Mondal, Mainack, Ganguly, Niloy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MEDVOC: Vocabulary Adaptation for Fine-tuning Pre-trained Language Models on Medical Text Summarization
by: Balde, Gunjan, et al.
Published: (2024)
by: Balde, Gunjan, et al.
Published: (2024)
Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization
by: Balde, Gunjan, et al.
Published: (2026)
by: Balde, Gunjan, et al.
Published: (2026)
Adaptive BPE Tokenization for Enhanced Vocabulary Adaptation in Finetuning Pretrained Language Models
by: Balde, Gunjan, et al.
Published: (2024)
by: Balde, Gunjan, et al.
Published: (2024)
Unlocking Efficiency: Adaptive Masking for Gene Transformer Models
by: Roy, Soumyadeep, et al.
Published: (2024)
by: Roy, Soumyadeep, et al.
Published: (2024)
Beyond Accuracy: Investigating Error Types in GPT-4 Responses to USMLE Questions
by: Roy, Soumyadeep, et al.
Published: (2024)
by: Roy, Soumyadeep, et al.
Published: (2024)
Evaluating Subword Tokenization: Alien Subword Composition and OOV Generalization Challenge
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
by: Batsuren, Khuyagbaatar, et al.
Published: (2024)
Order-Based Pre-training Strategies for Procedural Text Understanding
by: Nandy, Abhilash, et al.
Published: (2024)
by: Nandy, Abhilash, et al.
Published: (2024)
Leveraging the Power of LLMs: A Fine-Tuning Approach for High-Quality Aspect-Based Summarization
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
by: Nag, Arijit, et al.
Published: (2024)
by: Nag, Arijit, et al.
Published: (2024)
Long Dialog Summarization: An Analysis
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
Off-The-Shelf Image-to-Image Models Are All You Need To Defeat Image Protection Schemes
by: Pleimling, Xavier, et al.
Published: (2026)
by: Pleimling, Xavier, et al.
Published: (2026)
Efficient Continual Pre-training of LLMs for Low-resource Languages
by: Nag, Arijit, et al.
Published: (2024)
by: Nag, Arijit, et al.
Published: (2024)
Instruction-Guided Bullet Point Summarization of Long Financial Earnings Call Transcripts
by: Khatuya, Subhendu, et al.
Published: (2024)
by: Khatuya, Subhendu, et al.
Published: (2024)
On the Role of Summary Content Units in Text Summarization Evaluation
by: Nawrath, Marcel, et al.
Published: (2024)
by: Nawrath, Marcel, et al.
Published: (2024)
One Arrow, Many Targets: Probing LLMs for Multi-Attribute Controllable Text Summarization
by: Roy, Tathagato, et al.
Published: (2024)
by: Roy, Tathagato, et al.
Published: (2024)
On The Persona-based Summarization of Domain-Specific Documents
by: Mullick, Ankan, et al.
Published: (2024)
by: Mullick, Ankan, et al.
Published: (2024)
AdaptEval: Evaluating Large Language Models on Domain Adaptation for Text Summarization
by: Afzal, Anum, et al.
Published: (2024)
by: Afzal, Anum, et al.
Published: (2024)
Less Is More? Selective Visual Attention to High-Importance Regions for Multimodal Radiology Summarization
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2026)
by: Naznin, Mst. Fahmida Sultana, et al.
Published: (2026)
Evaluating LLMs and Pre-trained Models for Text Summarization Across Diverse Datasets
by: Rehman, Tohida, et al.
Published: (2025)
by: Rehman, Tohida, et al.
Published: (2025)
Fact-Controlled Diagnosis of Hallucinations in Medical Text Summarization
by: BN, Suhas, et al.
Published: (2025)
by: BN, Suhas, et al.
Published: (2025)
Navigating Nuance: In Quest for Political Truth
by: Sar, Soumyadeep, et al.
Published: (2025)
by: Sar, Soumyadeep, et al.
Published: (2025)
Vocab Diet: Reshaping the Vocabulary of LLMs via Vector Arithmetic
by: Reif, Yuval, et al.
Published: (2025)
by: Reif, Yuval, et al.
Published: (2025)
Multi-Dimensional Evaluation of Text Summarization with In-Context Learning
by: Jain, Sameer, et al.
Published: (2023)
by: Jain, Sameer, et al.
Published: (2023)
Extrinsically-Focused Evaluation of Omissions in Medical Summarization
by: Schumacher, Elliot, et al.
Published: (2023)
by: Schumacher, Elliot, et al.
Published: (2023)
A LongFormer-Based Framework for Accurate and Efficient Medical Text Summarization
by: Sun, Dan, et al.
Published: (2025)
by: Sun, Dan, et al.
Published: (2025)
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
by: Moroni, Luca, et al.
Published: (2025)
by: Moroni, Luca, et al.
Published: (2025)
Program of Thoughts for Financial Reasoning: Leveraging Dynamic In-Context Examples and Generative Retrieval
by: Khatuya, Subhendu, et al.
Published: (2025)
by: Khatuya, Subhendu, et al.
Published: (2025)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
by: Casola, Silvia, et al.
Published: (2025)
by: Casola, Silvia, et al.
Published: (2025)
Reliability Gated Multi-Teacher Distillation for Low Resource Abstractive Summarization
by: Sumit, Dipto, et al.
Published: (2026)
by: Sumit, Dipto, et al.
Published: (2026)
CCSBench: Evaluating Compositional Controllability in LLMs for Scientific Document Summarization
by: Ding, Yixi, et al.
Published: (2024)
by: Ding, Yixi, et al.
Published: (2024)
Adapted Large Language Models Can Outperform Medical Experts in Clinical Text Summarization
by: Van Veen, Dave, et al.
Published: (2023)
by: Van Veen, Dave, et al.
Published: (2023)
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence
by: Joseph, Sebastian Antony, et al.
Published: (2024)
by: Joseph, Sebastian Antony, et al.
Published: (2024)
How Robust are the Tabular QA Models for Scientific Tables? A Study using Customized Dataset
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
Abstractive Text Summarization for Bangla Language Using NLP and Machine Learning Approaches
by: Miazee, Asif Ahammad, et al.
Published: (2025)
by: Miazee, Asif Ahammad, et al.
Published: (2025)
Multi-LLM Text Summarization
by: Fang, Jiangnan, et al.
Published: (2024)
by: Fang, Jiangnan, et al.
Published: (2024)
Comparative Analysis of Open-Source Language Models in Summarizing Medical Text Data
by: Chen, Yuhao, et al.
Published: (2024)
by: Chen, Yuhao, et al.
Published: (2024)
IDALC: A Semi-Supervised Framework for Intent Detection and Active Learning based Correction
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
Towards Sustainable NLP: Insights from Benchmarking Inference Energy in Large Language Models
by: Poddar, Soham, et al.
Published: (2025)
by: Poddar, Soham, et al.
Published: (2025)
Gavel: Agent Meets Checklist for Evaluating LLMs on Long-Context Legal Summarization
by: Dou, Yao, et al.
Published: (2026)
by: Dou, Yao, et al.
Published: (2026)
QAPyramid: Fine-grained Evaluation of Content Selection for Text Summarization
by: Zhang, Shiyue, et al.
Published: (2024)
by: Zhang, Shiyue, et al.
Published: (2024)
Similar Items
-
MEDVOC: Vocabulary Adaptation for Fine-tuning Pre-trained Language Models on Medical Text Summarization
by: Balde, Gunjan, et al.
Published: (2024) -
Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization
by: Balde, Gunjan, et al.
Published: (2026) -
Adaptive BPE Tokenization for Enhanced Vocabulary Adaptation in Finetuning Pretrained Language Models
by: Balde, Gunjan, et al.
Published: (2024) -
Unlocking Efficiency: Adaptive Masking for Gene Transformer Models
by: Roy, Soumyadeep, et al.
Published: (2024) -
Beyond Accuracy: Investigating Error Types in GPT-4 Responses to USMLE Questions
by: Roy, Soumyadeep, et al.
Published: (2024)