Vi-Mistral-X: Building a Vietnamese Language Model with Advanced Continual Pre-training
Fuente:
arXiv
Guardado en:
| Autor principal: | Vo, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model
por: Dat, Phan Tran Minh, et al.
Publicado: (2025)
por: Dat, Phan Tran Minh, et al.
Publicado: (2025)
ViX-Ray: A Vietnamese Chest X-Ray Dataset for Vision-Language Models
por: Nguyen, Duy Vu Minh, et al.
Publicado: (2026)
por: Nguyen, Duy Vu Minh, et al.
Publicado: (2026)
Efficient Continual Pre-training for Building Domain Specific Large Language Models
por: Xie, Yong, et al.
Publicado: (2023)
por: Xie, Yong, et al.
Publicado: (2023)
PhoGPT: Generative Pre-training for Vietnamese
por: Nguyen, Dat Quoc, et al.
Publicado: (2023)
por: Nguyen, Dat Quoc, et al.
Publicado: (2023)
ViTHSD: Exploiting Hatred by Targets for Hate Speech Detection on Vietnamese Social Media Texts
por: Vo, Cuong Nhat, et al.
Publicado: (2024)
por: Vo, Cuong Nhat, et al.
Publicado: (2024)
ViMedCSS: A Vietnamese Medical Code-Switching Speech Dataset & Benchmark
por: Nguyen, Tung X., et al.
Publicado: (2026)
por: Nguyen, Tung X., et al.
Publicado: (2026)
ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models
por: Nguyen, Trong-Hieu, et al.
Publicado: (2024)
por: Nguyen, Trong-Hieu, et al.
Publicado: (2024)
ViHERMES: A Graph-Grounded Multihop Question Answering Benchmark and System for Vietnamese Healthcare Regulations
por: Nguyen, Long S. T., et al.
Publicado: (2026)
por: Nguyen, Long S. T., et al.
Publicado: (2026)
Examining Forgetting in Continual Pre-training of Aligned Large Language Models
por: Li, Chen-An, et al.
Publicado: (2024)
por: Li, Chen-An, et al.
Publicado: (2024)
GreenMind: A Next-Generation Vietnamese Large Language Model for Structured and Logical Reasoning
por: Tung, Luu Quy, et al.
Publicado: (2025)
por: Tung, Luu Quy, et al.
Publicado: (2025)
Combining Entropy and Matrix Nuclear Norm for Enhanced Evaluation of Language Models
por: Vo, James
Publicado: (2024)
por: Vo, James
Publicado: (2024)
ViLegalNLI: Natural Language Inference for Vietnamese Legal Texts
por: Duong, Nhung Thi-Hong, et al.
Publicado: (2026)
por: Duong, Nhung Thi-Hong, et al.
Publicado: (2026)
ViBidirectionMT-Eval: Machine Translation for Vietnamese-Chinese and Vietnamese-Lao language pair
por: Tran, Hong-Viet, et al.
Publicado: (2025)
por: Tran, Hong-Viet, et al.
Publicado: (2025)
Transformer Layer Injection: A Novel Approach for Efficient Upscaling of Large Language Models
por: Vo, James
Publicado: (2024)
por: Vo, James
Publicado: (2024)
Stable Language Model Pre-training by Reducing Embedding Variability
por: Chung, Woojin, et al.
Publicado: (2024)
por: Chung, Woojin, et al.
Publicado: (2024)
ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances
por: Do, Huy Ba, et al.
Publicado: (2025)
por: Do, Huy Ba, et al.
Publicado: (2025)
ViSpeechFormer: A Phonemic Approach for Vietnamese Automatic Speech Recognition
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
por: Nguyen, Khoa Anh, et al.
Publicado: (2026)
LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training
por: Zhu, Tong, et al.
Publicado: (2024)
por: Zhu, Tong, et al.
Publicado: (2024)
Vintern-1B: An Efficient Multimodal Large Language Model for Vietnamese
por: Doan, Khang T., et al.
Publicado: (2024)
por: Doan, Khang T., et al.
Publicado: (2024)
Improving Vietnamese-English Medical Machine Translation
por: Vo, Nhu, et al.
Publicado: (2024)
por: Vo, Nhu, et al.
Publicado: (2024)
Large Malaysian Language Model Based on Mistral for Enhanced Local Language Understanding
por: Zolkepli, Husein, et al.
Publicado: (2024)
por: Zolkepli, Husein, et al.
Publicado: (2024)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
por: Nguyen, Anh Thi-Hoang, et al.
Publicado: (2026)
por: Nguyen, Anh Thi-Hoang, et al.
Publicado: (2026)
Construction of Domain-specified Japanese Large Language Model for Finance through Continual Pre-training
por: Hirano, Masanori, et al.
Publicado: (2024)
por: Hirano, Masanori, et al.
Publicado: (2024)
ViWikiFC: Fact-Checking for Vietnamese Wikipedia-Based Textual Knowledge Source
por: Le, Hung Tuan, et al.
Publicado: (2024)
por: Le, Hung Tuan, et al.
Publicado: (2024)
ViLexNorm: A Lexical Normalization Corpus for Vietnamese Social Media Text
por: Nguyen, Thanh-Nhi, et al.
Publicado: (2024)
por: Nguyen, Thanh-Nhi, et al.
Publicado: (2024)
FreeTxt-Vi: A Benchmarked Vietnamese-English Toolkit for Segmentation, Sentiment, and Summarisation
por: Huy, Hung Nguyen, et al.
Publicado: (2026)
por: Huy, Hung Nguyen, et al.
Publicado: (2026)
ViMQ: A Vietnamese Medical Question Dataset for Healthcare Dialogue System Development
por: Huy, Ta Duc, et al.
Publicado: (2023)
por: Huy, Ta Duc, et al.
Publicado: (2023)
ViMMRC 2.0 -- Enhancing Machine Reading Comprehension on Vietnamese Literature Text
por: Luu, Son T., et al.
Publicado: (2023)
por: Luu, Son T., et al.
Publicado: (2023)
ViHateT5: Enhancing Hate Speech Detection in Vietnamese With A Unified Text-to-Text Transformer Model
por: Nguyen, Luan Thanh
Publicado: (2024)
por: Nguyen, Luan Thanh
Publicado: (2024)
Efficient Continual Pre-training of LLMs for Low-resource Languages
por: Nag, Arijit, et al.
Publicado: (2024)
por: Nag, Arijit, et al.
Publicado: (2024)
Simple and Scalable Strategies to Continually Pre-train Large Language Models
por: Ibrahim, Adam, et al.
Publicado: (2024)
por: Ibrahim, Adam, et al.
Publicado: (2024)
Scaling Agents via Continual Pre-training
por: Su, Liangcai, et al.
Publicado: (2025)
por: Su, Liangcai, et al.
Publicado: (2025)
ViMultiChoice: Toward a Method That Gives Explanation for Multiple-Choice Reading Comprehension in Vietnamese
por: Cao, Trung Tien, et al.
Publicado: (2026)
por: Cao, Trung Tien, et al.
Publicado: (2026)
AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages
por: Yu, Hao, et al.
Publicado: (2026)
por: Yu, Hao, et al.
Publicado: (2026)
MELT: Materials-aware Continued Pre-training for Language Model Adaptation to Materials Science
por: Kim, Junho, et al.
Publicado: (2024)
por: Kim, Junho, et al.
Publicado: (2024)
Making Pre-trained Language Models Better Continual Few-Shot Relation Extractors
por: Ma, Shengkun, et al.
Publicado: (2024)
por: Ma, Shengkun, et al.
Publicado: (2024)
Development of Cognitive Intelligence in Pre-trained Language Models
por: Shah, Raj Sanjay, et al.
Publicado: (2024)
por: Shah, Raj Sanjay, et al.
Publicado: (2024)
ViGoEmotions: A Benchmark Dataset For Fine-grained Emotion Detection on Vietnamese Texts
por: Tran, Hung Quang, et al.
Publicado: (2026)
por: Tran, Hung Quang, et al.
Publicado: (2026)
Metadata Conditioning Accelerates Language Model Pre-training
por: Gao, Tianyu, et al.
Publicado: (2025)
por: Gao, Tianyu, et al.
Publicado: (2025)
Evaluating Discourse Cohesion in Pre-trained Language Models
por: He, Jie, et al.
Publicado: (2025)
por: He, Jie, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model
por: Dat, Phan Tran Minh, et al.
Publicado: (2025) -
ViX-Ray: A Vietnamese Chest X-Ray Dataset for Vision-Language Models
por: Nguyen, Duy Vu Minh, et al.
Publicado: (2026) -
Efficient Continual Pre-training for Building Domain Specific Large Language Models
por: Xie, Yong, et al.
Publicado: (2023) -
PhoGPT: Generative Pre-training for Vietnamese
por: Nguyen, Dat Quoc, et al.
Publicado: (2023) -
ViTHSD: Exploiting Hatred by Targets for Hate Speech Detection on Vietnamese Social Media Texts
por: Vo, Cuong Nhat, et al.
Publicado: (2024)