Prune or Retrain: Optimizing the Vocabulary of Multilingual Models for Estonian
Fuente:
arXiv
Guardado en:
| Autores principales: | Dorkin, Aleksei, Purason, Taido, Sirts, Kairit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Comparison of Current Approaches to Lemmatization: A Case Study in Estonian
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
por: Dorkin, Aleksei, et al.
Publicado: (2026)
por: Dorkin, Aleksei, et al.
Publicado: (2026)
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
por: Dorkin, Aleksei, et al.
Publicado: (2025)
por: Dorkin, Aleksei, et al.
Publicado: (2025)
Sõnajaht: Definition Embeddings and Semantic Search for Reverse Dictionary Creation
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
TartuNLP @ SIGTYP 2024 Shared Task: Adapting XLM-RoBERTa for Ancient and Historical Languages
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
TartuNLP @ AXOLOTL-24: Leveraging Classifier Output for New Sense Detection in Lexical Semantics
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
TartuNLP at EvaLatin 2024: Emotion Polarity Detection
por: Dorkin, Aleksei, et al.
Publicado: (2024)
por: Dorkin, Aleksei, et al.
Publicado: (2024)
Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation
por: Ojastu, Marii, et al.
Publicado: (2025)
por: Ojastu, Marii, et al.
Publicado: (2025)
Creation of the Estonian Subjectivity Dataset: Assessing the Degree of Subjectivity on a Scale
por: Gailit, Karl Gustav, et al.
Publicado: (2025)
por: Gailit, Karl Gustav, et al.
Publicado: (2025)
Exploratory Study into Relations between Cognitive Distortions and Emotional Appraisals
por: Agarwal, Navneet, et al.
Publicado: (2025)
por: Agarwal, Navneet, et al.
Publicado: (2025)
Context is Important in Depressive Language: A Study of the Interaction Between the Sentiments and Linguistic Markers in Reddit Discussions
por: Sharma, Neha, et al.
Publicado: (2024)
por: Sharma, Neha, et al.
Publicado: (2024)
Exploring Profiles of Cognitive Distortions Associated with Mental Health Disorders
por: Anikejeva, Alina, et al.
Publicado: (2026)
por: Anikejeva, Alina, et al.
Publicado: (2026)
LLMs for Extremely Low-Resource Finno-Ugric Languages
por: Purason, Taido, et al.
Publicado: (2024)
por: Purason, Taido, et al.
Publicado: (2024)
Teaching Old Tokenizers New Words: Efficient Tokenizer Adaptation for Pre-trained Models
por: Purason, Taido, et al.
Publicado: (2025)
por: Purason, Taido, et al.
Publicado: (2025)
Your Model Is Not Predicting Depression Well And That Is Why: A Case Study of PRIMATE Dataset
por: Milintsevich, Kirill, et al.
Publicado: (2024)
por: Milintsevich, Kirill, et al.
Publicado: (2024)
Teaching Llama a New Language Through Cross-Lingual Knowledge Transfer
por: Kuulmets, Hele-Andra, et al.
Publicado: (2024)
por: Kuulmets, Hele-Andra, et al.
Publicado: (2024)
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings
por: Ruder, Deniss, et al.
Publicado: (2025)
por: Ruder, Deniss, et al.
Publicado: (2025)
Evaluating Lexicon Incorporation for Depression Symptom Estimation
por: Milintsevich, Kirill, et al.
Publicado: (2024)
por: Milintsevich, Kirill, et al.
Publicado: (2024)
To Err Is Human, but Llamas Can Learn It Too
por: Luhtaru, Agnes, et al.
Publicado: (2024)
por: Luhtaru, Agnes, et al.
Publicado: (2024)
Pruning Foundation Models for High Accuracy without Retraining
por: Zhao, Pu, et al.
Publicado: (2024)
por: Zhao, Pu, et al.
Publicado: (2024)
Olica: Efficient Structured Pruning of Large Language Models without Retraining
por: He, Jiujun, et al.
Publicado: (2025)
por: He, Jiujun, et al.
Publicado: (2025)
Estonian Native Large Language Model Benchmark
por: Lillepalu, Helena Grete, et al.
Publicado: (2025)
por: Lillepalu, Helena Grete, et al.
Publicado: (2025)
Pruning Multilingual Large Language Models for Multilingual Inference
por: Kim, Hwichan, et al.
Publicado: (2024)
por: Kim, Hwichan, et al.
Publicado: (2024)
Z-Pruner: Post-Training Pruning of Large Language Models for Efficiency without Retraining
por: Bhuiyan, Samiul Basir, et al.
Publicado: (2025)
por: Bhuiyan, Samiul Basir, et al.
Publicado: (2025)
Shortened LLaMA: Depth Pruning for Large Language Models with Comparison of Retraining Methods
por: Kim, Bo-Kyeong, et al.
Publicado: (2024)
por: Kim, Bo-Kyeong, et al.
Publicado: (2024)
Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
por: Park, Seungcheol, et al.
Publicado: (2023)
por: Park, Seungcheol, et al.
Publicado: (2023)
Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs
por: Zhou, Yixiao, et al.
Publicado: (2025)
por: Zhou, Yixiao, et al.
Publicado: (2025)
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets
por: Barbu, Eduard, et al.
Publicado: (2025)
por: Barbu, Eduard, et al.
Publicado: (2025)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
por: Kallini, Julie, et al.
Publicado: (2025)
por: Kallini, Julie, et al.
Publicado: (2025)
An experimental and computational study of an Estonian single-person word naming
por: Lõo, Kaidi, et al.
Publicado: (2025)
por: Lõo, Kaidi, et al.
Publicado: (2025)
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
por: Fedorchenko, Artem, et al.
Publicado: (2025)
por: Fedorchenko, Artem, et al.
Publicado: (2025)
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
por: Sildam, Tiia, et al.
Publicado: (2024)
por: Sildam, Tiia, et al.
Publicado: (2024)
VOCABTRIM: Vocabulary Pruning for Efficient Speculative Decoding in LLMs
por: Goel, Raghavv, et al.
Publicado: (2025)
por: Goel, Raghavv, et al.
Publicado: (2025)
The Impact of Vocabulary Overlaps on Knowledge Transfer in Multilingual Machine Translation
por: Itkonen, Oona, et al.
Publicado: (2026)
por: Itkonen, Oona, et al.
Publicado: (2026)
Efficient and Effective Vocabulary Expansion Towards Multilingual Large Language Models
por: Kim, Seungduk, et al.
Publicado: (2024)
por: Kim, Seungduk, et al.
Publicado: (2024)
Dynamic Vocabulary Pruning in Early-Exit LLMs
por: Vincenti, Jort, et al.
Publicado: (2024)
por: Vincenti, Jort, et al.
Publicado: (2024)
How Vocabulary Sharing Facilitates Multilingualism in LLaMA?
por: Yuan, Fei, et al.
Publicado: (2023)
por: Yuan, Fei, et al.
Publicado: (2023)
Autocorrect for Estonian texts: final report from project EKTB25
por: Luhtaru, Agnes, et al.
Publicado: (2024)
por: Luhtaru, Agnes, et al.
Publicado: (2024)
Reuse, Don't Retrain: A Recipe for Continued Pretraining of Language Models
por: Parmar, Jupinder, et al.
Publicado: (2024)
por: Parmar, Jupinder, et al.
Publicado: (2024)
Ejemplares similares
-
Comparison of Current Approaches to Lemmatization: A Case Study in Estonian
por: Dorkin, Aleksei, et al.
Publicado: (2024) -
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
por: Dorkin, Aleksei, et al.
Publicado: (2024) -
EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
por: Dorkin, Aleksei, et al.
Publicado: (2026) -
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
por: Dorkin, Aleksei, et al.
Publicado: (2025) -
Sõnajaht: Definition Embeddings and Semantic Search for Reverse Dictionary Creation
por: Dorkin, Aleksei, et al.
Publicado: (2024)