Compressing Language Models for Specialized Domains
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Williams, Miles, Chrysostomou, George, Jeronymo, Vitor, Aletras, Nikolaos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-calibration for Language Model Quantization and Pruning
von: Williams, Miles, et al.
Veröffentlicht: (2024)
von: Williams, Miles, et al.
Veröffentlicht: (2024)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
von: Williams, Miles, et al.
Veröffentlicht: (2023)
von: Williams, Miles, et al.
Veröffentlicht: (2023)
On the Impact of Calibration Data in Post-training Quantization and Pruning
von: Williams, Miles, et al.
Veröffentlicht: (2023)
von: Williams, Miles, et al.
Veröffentlicht: (2023)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
How Private are Language Models in Abstractive Summarization?
von: Hughes, Anthony, et al.
Veröffentlicht: (2024)
von: Hughes, Anthony, et al.
Veröffentlicht: (2024)
Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2026)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2026)
An Empirical Study on Cross-lingual Vocabulary Adaptation for Efficient Language Model Inference
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
Deconstructing Attention: Investigating Design Principles for Effective Language Modeling
von: Xue, Huiyin, et al.
Veröffentlicht: (2025)
von: Xue, Huiyin, et al.
Veröffentlicht: (2025)
Adapting Chat Language Models Using Only Target Unlabeled Language Data
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
Fundamental Reasoning Paradigms Induce Out-of-Domain Generalization in Language Models
von: Cao, Mingzi, et al.
Veröffentlicht: (2026)
von: Cao, Mingzi, et al.
Veröffentlicht: (2026)
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text?
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
Incorporating Attribution Importance for Improving Faithfulness Metrics
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2023)
Fine-Tuning on Noisy Instructions: Effects on Generalization and Performance
von: Alajrami, Ahmed, et al.
Veröffentlicht: (2025)
von: Alajrami, Ahmed, et al.
Veröffentlicht: (2025)
Progressive Depth Up-scaling via Optimal Transport
von: Cao, Mingzi, et al.
Veröffentlicht: (2025)
von: Cao, Mingzi, et al.
Veröffentlicht: (2025)
Enhancing Logical Reasoning in Language Models via Symbolically-Guided Monte Carlo Process Supervision
von: Tan, Xingwei, et al.
Veröffentlicht: (2025)
von: Tan, Xingwei, et al.
Veröffentlicht: (2025)
Examining the Limitations of Computational Rumor Detection Models Trained on Static Datasets
von: Mu, Yida, et al.
Veröffentlicht: (2023)
von: Mu, Yida, et al.
Veröffentlicht: (2023)
Mitigating Catastrophic Forgetting in Target Language Adaptation of LLMs via Source-Shielded Updates
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2025)
PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHing
von: Hughes, Anthony, et al.
Veröffentlicht: (2025)
von: Hughes, Anthony, et al.
Veröffentlicht: (2025)
GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2025)
Enhancing Data Quality through Simple De-duplication: Navigating Responsible Computational Social Science Research
von: Mu, Yida, et al.
Veröffentlicht: (2024)
von: Mu, Yida, et al.
Veröffentlicht: (2024)
Where does output diversity collapse in post-training?
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
von: Karouzos, Constantinos, et al.
Veröffentlicht: (2026)
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2026)
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2026)
Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models
von: Tan, Xingwei, et al.
Veröffentlicht: (2026)
von: Tan, Xingwei, et al.
Veröffentlicht: (2026)
Can Confidence Estimates Decide When Chain-of-Thought Is Necessary for LLMs?
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
We Need to Talk About Classification Evaluation Metrics in NLP
von: Vickers, Peter, et al.
Veröffentlicht: (2024)
von: Vickers, Peter, et al.
Veröffentlicht: (2024)
Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
von: Lewis-Lim, Samuel, et al.
Veröffentlicht: (2025)
Navigating Prompt Complexity for Zero-Shot Classification: A Study of Large Language Models in Computational Social Science
von: Mu, Yida, et al.
Veröffentlicht: (2023)
von: Mu, Yida, et al.
Veröffentlicht: (2023)
Improving Multimodal Classification of Social Media Posts by Leveraging Image-Text Auxiliary Tasks
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2023)
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2023)
Who is bragging more online? A large scale analysis of bragging in social media
von: Jin, Mali, et al.
Veröffentlicht: (2024)
von: Jin, Mali, et al.
Veröffentlicht: (2024)
Boundary-targeted Membership Inference Attacks on Safety Classifiers
von: Hughes, Anthony, et al.
Veröffentlicht: (2026)
von: Hughes, Anthony, et al.
Veröffentlicht: (2026)
Optimal Splitting of Language Models from Mixtures to Specialized Domains
von: Seto, Skyler, et al.
Veröffentlicht: (2026)
von: Seto, Skyler, et al.
Veröffentlicht: (2026)
Very High Level Programming Languages (e.g., SNOBOL, COMIT) in the Special Librarian's Future
von: Libbey, Miles A.
Veröffentlicht: (1975)
von: Libbey, Miles A.
Veröffentlicht: (1975)
Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
von: Zhu, Jie, et al.
Veröffentlicht: (2025)
Language Model Adaptation to Specialized Domains through Selective Masking based on Genre and Topical Characteristics
von: Belfathi, Anas, et al.
Veröffentlicht: (2024)
von: Belfathi, Anas, et al.
Veröffentlicht: (2024)
Domain Specialization as the Key to Make Large Language Models Disruptive: A Comprehensive Survey
von: Ling, Chen, et al.
Veröffentlicht: (2023)
von: Ling, Chen, et al.
Veröffentlicht: (2023)
LegalMidm: Use-Case-Driven Legal Domain Specialization for Korean Large Language Model
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2026)
Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?
von: Doll, Niclas, et al.
Veröffentlicht: (2026)
von: Doll, Niclas, et al.
Veröffentlicht: (2026)
Is Biomedical Specialization Still Worth It? Insights from Domain-Adaptive Language Modelling with a New French Health Corpus
von: Mannion, Aidan, et al.
Veröffentlicht: (2026)
von: Mannion, Aidan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Self-calibration for Language Model Quantization and Pruning
von: Williams, Miles, et al.
Veröffentlicht: (2024) -
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023) -
Vocabulary-level Memory Efficiency for Language Model Fine-tuning
von: Williams, Miles, et al.
Veröffentlicht: (2023) -
On the Impact of Calibration Data in Post-training Quantization and Pruning
von: Williams, Miles, et al.
Veröffentlicht: (2023) -
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)