Rethinking LLM Language Adaptation: A Case Study on Chinese Mixtral
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cui, Yiming, Yao, Xin |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations
par: Li, Yiming, et autres
Publié: (2024)
par: Li, Yiming, et autres
Publié: (2024)
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events
par: Li, Yiming, et autres
Publié: (2023)
par: Li, Yiming, et autres
Publié: (2023)
Evaluating Large Language Models on Multimodal Chemistry Olympiad Exams
par: Cui, Yiming, et autres
Publié: (2025)
par: Cui, Yiming, et autres
Publié: (2025)
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources
par: Li, Yiming, et autres
Publié: (2024)
par: Li, Yiming, et autres
Publié: (2024)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
par: Du, Xinrun, et autres
Publié: (2024)
par: Du, Xinrun, et autres
Publié: (2024)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
par: Khan, Eeham, et autres
Publié: (2025)
par: Khan, Eeham, et autres
Publié: (2025)
Can LLM Substitute Human Labeling? A Case Study of Fine-grained Chinese Address Entity Recognition Dataset for UAV Delivery
par: Yao, Yuxuan, et autres
Publié: (2024)
par: Yao, Yuxuan, et autres
Publié: (2024)
ChartHal: A Fine-grained Framework Evaluating Hallucination of Large Vision Language Models in Chart Understanding
par: Wang, Xingqi, et autres
Publié: (2025)
par: Wang, Xingqi, et autres
Publié: (2025)
ORBIT: Cost-Effective Dataset Curation for Large Language Model Domain Adaptation with an Astronomy Case Study
par: Modesitt, Eric, et autres
Publié: (2024)
par: Modesitt, Eric, et autres
Publié: (2024)
MERaLiON-TextLLM: Cross-Lingual Understanding of Large Language Models in Chinese, Indonesian, Malay, and Singlish
par: Huang, Xin, et autres
Publié: (2024)
par: Huang, Xin, et autres
Publié: (2024)
Rethinking the Outlier Distribution in Large Language Models: An In-depth Study
par: Raman, Rahul, et autres
Publié: (2025)
par: Raman, Rahul, et autres
Publié: (2025)
Rethinking Toxicity Evaluation in Large Language Models: A Multi-Label Perspective
par: Kou, Zhiqiang, et autres
Publié: (2025)
par: Kou, Zhiqiang, et autres
Publié: (2025)
JailBench: A Comprehensive Chinese Security Assessment Benchmark for Large Language Models
par: Liu, Shuyi, et autres
Publié: (2025)
par: Liu, Shuyi, et autres
Publié: (2025)
Qibo: A Large Language Model for Traditional Chinese Medicine
par: Zhang, Heyi, et autres
Publié: (2024)
par: Zhang, Heyi, et autres
Publié: (2024)
Mixtral of Experts
par: Jiang, Albert Q., et autres
Publié: (2024)
par: Jiang, Albert Q., et autres
Publié: (2024)
Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry
par: Li, Zhuochun, et autres
Publié: (2026)
par: Li, Zhuochun, et autres
Publié: (2026)
Banishing LLM Hallucinations Requires Rethinking Generalization
par: Li, Johnny, et autres
Publié: (2024)
par: Li, Johnny, et autres
Publié: (2024)
Rethinking Human Preference Evaluation of LLM Rationales
par: Li, Ziang, et autres
Publié: (2025)
par: Li, Ziang, et autres
Publié: (2025)
Generative LLM Powered Conversational AI Application for Personalized Risk Assessment: A Case Study in COVID-19
par: Roshani, Mohammad Amin, et autres
Publié: (2024)
par: Roshani, Mohammad Amin, et autres
Publié: (2024)
Flexora: Flexible Low Rank Adaptation for Large Language Models
par: Wei, Chenxing, et autres
Publié: (2024)
par: Wei, Chenxing, et autres
Publié: (2024)
FineMath: A Fine-Grained Mathematical Evaluation Benchmark for Chinese Large Language Models
par: Liu, Yan, et autres
Publié: (2024)
par: Liu, Yan, et autres
Publié: (2024)
Rethinking Text-based Protein Understanding: Retrieval or LLM?
par: Wu, Juntong, et autres
Publié: (2025)
par: Wu, Juntong, et autres
Publié: (2025)
SafeLLM: Domain-Specific Safety Monitoring for Large Language Models: A Case Study of Offshore Wind Maintenance
par: Walker, Connor, et autres
Publié: (2024)
par: Walker, Connor, et autres
Publié: (2024)
Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language Capabilities
par: Fujii, Kazuki, et autres
Publié: (2024)
par: Fujii, Kazuki, et autres
Publié: (2024)
An Empirical Investigation of Domain Adaptation Ability for Chinese Spelling Check Models
par: Wang, Xi, et autres
Publié: (2024)
par: Wang, Xi, et autres
Publié: (2024)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
par: Liu, Xiang, et autres
Publié: (2025)
par: Liu, Xiang, et autres
Publié: (2025)
Repurposing Annotation Guidelines to Instruct LLM Annotators: A Case Study
par: Kim, Kon Woo, et autres
Publié: (2025)
par: Kim, Kon Woo, et autres
Publié: (2025)
Who Benchmarks the Benchmarks? A Case Study of LLM Evaluation in Icelandic
par: Ingimundarson, Finnur Ágúst, et autres
Publié: (2026)
par: Ingimundarson, Finnur Ágúst, et autres
Publié: (2026)
CMoralEval: A Moral Evaluation Benchmark for Chinese Large Language Models
par: Yu, Linhao, et autres
Publié: (2024)
par: Yu, Linhao, et autres
Publié: (2024)
EduEval: A Hierarchical Cognitive Benchmark for Evaluating Large Language Models in Chinese Education
par: Ma, Guoqing, et autres
Publié: (2025)
par: Ma, Guoqing, et autres
Publié: (2025)
Rethinking Adapter Placement: A Dominant Adaptation Module Perspective
par: Zhang, Suoxin, et autres
Publié: (2026)
par: Zhang, Suoxin, et autres
Publié: (2026)
LawGPT: A Chinese Legal Knowledge-Enhanced Large Language Model
par: Zhou, Zhi, et autres
Publié: (2024)
par: Zhou, Zhi, et autres
Publié: (2024)
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
par: Yang, Hang, et autres
Publié: (2024)
par: Yang, Hang, et autres
Publié: (2024)
Rethinking the Bounds of LLM Reasoning: Are Multi-Agent Discussions the Key?
par: Wang, Qineng, et autres
Publié: (2024)
par: Wang, Qineng, et autres
Publié: (2024)
TCMBench: A Comprehensive Benchmark for Evaluating Large Language Models in Traditional Chinese Medicine
par: Yue, Wenjing, et autres
Publié: (2024)
par: Yue, Wenjing, et autres
Publié: (2024)
LLM-DER:A Named Entity Recognition Method Based on Large Language Models for Chinese Coal Chemical Domain
par: Xiao, Le, et autres
Publié: (2024)
par: Xiao, Le, et autres
Publié: (2024)
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction
par: Bailis, Suma, et autres
Publié: (2024)
par: Bailis, Suma, et autres
Publié: (2024)
CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models
par: LI, Yizhi, et autres
Publié: (2024)
par: LI, Yizhi, et autres
Publié: (2024)
Rhyme-aware Chinese lyric generator based on GPT
par: Yuan, Yixiao, et autres
Publié: (2024)
par: Yuan, Yixiao, et autres
Publié: (2024)
CTourLLM: Enhancing LLMs with Chinese Tourism Knowledge
par: Wei, Qikai, et autres
Publié: (2024)
par: Wei, Qikai, et autres
Publié: (2024)
Documents similaires
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations
par: Li, Yiming, et autres
Publié: (2024) -
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events
par: Li, Yiming, et autres
Publié: (2023) -
Evaluating Large Language Models on Multimodal Chemistry Olympiad Exams
par: Cui, Yiming, et autres
Publié: (2025) -
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources
par: Li, Yiming, et autres
Publié: (2024) -
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
par: Du, Xinrun, et autres
Publié: (2024)