Building a Strong Instruction Language Model for a Less-Resourced Language
Fuente:
arXiv
Guardado en:
| Autores principales: | Vreš, Domen, Arčon, Tjaša, Petrič, Timotej, Vajda, Dario, Robnik-Šikonja, Marko, Bajec, Iztok Lebar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving LLMs for Machine Translation Using Synthetic Preference Data
por: Vajda, Dario, et al.
Publicado: (2025)
por: Vajda, Dario, et al.
Publicado: (2025)
Generative Model for Less-Resourced Language with 1 billion parameters
por: Vreš, Domen, et al.
Publicado: (2024)
por: Vreš, Domen, et al.
Publicado: (2024)
Sarcasm Detection in a Less-Resourced Language
por: Đoković, Lazar, et al.
Publicado: (2024)
por: Đoković, Lazar, et al.
Publicado: (2024)
Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
por: Arčon, Tjaša, et al.
Publicado: (2026)
por: Arčon, Tjaša, et al.
Publicado: (2026)
Large language models for folktale type automation based on motifs: Cinderella case study
por: Arčon, Tjaša, et al.
Publicado: (2025)
por: Arčon, Tjaša, et al.
Publicado: (2025)
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
por: Klemen, Matej, et al.
Publicado: (2025)
por: Klemen, Matej, et al.
Publicado: (2025)
Measuring Catastrophic Forgetting in Cross-Lingual Transfer Paradigms: Exploring Tuning Strategies
por: Koloski, Boshko, et al.
Publicado: (2023)
por: Koloski, Boshko, et al.
Publicado: (2023)
TT-XAI: Trustworthy Clinical Text Explanations via Keyword Distillation and LLM Reasoning
por: Miok, Kristian, et al.
Publicado: (2025)
por: Miok, Kristian, et al.
Publicado: (2025)
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
por: Pranjić, Marko, et al.
Publicado: (2026)
por: Pranjić, Marko, et al.
Publicado: (2026)
Retrieval-augmented code completion for local projects using large language models
por: Hostnik, Marko, et al.
Publicado: (2024)
por: Hostnik, Marko, et al.
Publicado: (2024)
Review of Natural Language Processing in Pharmacology
por: Trajanov, Dimitar, et al.
Publicado: (2022)
por: Trajanov, Dimitar, et al.
Publicado: (2022)
Solving Word-Sense Disambiguation and Word-Sense Induction with Dictionary Examples
por: Škvorc, Tadej, et al.
Publicado: (2025)
por: Škvorc, Tadej, et al.
Publicado: (2025)
QFS-Composer: Query-focused summarization pipeline for less resourced languages
por: Đuranović, Vuk, et al.
Publicado: (2026)
por: Đuranović, Vuk, et al.
Publicado: (2026)
Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning
por: Vajda, Dario
Publicado: (2026)
por: Vajda, Dario
Publicado: (2026)
Real-time News Story Identification
por: Škvorc, Tadej, et al.
Publicado: (2025)
por: Škvorc, Tadej, et al.
Publicado: (2025)
Challenges in Explaining Pretrained Clinical Text Classifiers
por: Miok, Kristian, et al.
Publicado: (2026)
por: Miok, Kristian, et al.
Publicado: (2026)
Are Compressed Language Models Less Subgroup Robust?
por: Gee, Leonidas, et al.
Publicado: (2024)
por: Gee, Leonidas, et al.
Publicado: (2024)
Neural spell-checker: Beyond words with synthetic data generation
por: Klemen, Matej, et al.
Publicado: (2024)
por: Klemen, Matej, et al.
Publicado: (2024)
Less is KEN: a Universal and Simple Non-Parametric Pruning Algorithm for Large Language Models
por: Mastromattei, Michele, et al.
Publicado: (2024)
por: Mastromattei, Michele, et al.
Publicado: (2024)
LEGO: Language Model Building Blocks
por: Bhansali, Shrenik, et al.
Publicado: (2024)
por: Bhansali, Shrenik, et al.
Publicado: (2024)
Pay Less Attention to Function Words for Free Robustness of Vision-Language Models
por: Tian, Qiwei, et al.
Publicado: (2025)
por: Tian, Qiwei, et al.
Publicado: (2025)
Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
Less but Better: Parameter-Efficient Fine-Tuning of Large Language Models for Personality Detection
por: Shen, Lingzhi, et al.
Publicado: (2025)
por: Shen, Lingzhi, et al.
Publicado: (2025)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2024)
por: Köksal, Abdullatif, et al.
Publicado: (2024)
Less is More: Local Intrinsic Dimensions of Contextual Language Models
por: Ruppik, Benjamin Matthias, et al.
Publicado: (2025)
por: Ruppik, Benjamin Matthias, et al.
Publicado: (2025)
Evaluating Large Language Models at Evaluating Instruction Following
por: Zeng, Zhiyuan, et al.
Publicado: (2023)
por: Zeng, Zhiyuan, et al.
Publicado: (2023)
Compositional Instruction Following with Language Models and Reinforcement Learning
por: Cohen, Vanya, et al.
Publicado: (2025)
por: Cohen, Vanya, et al.
Publicado: (2025)
Shuttle Between the Instructions and the Parameters of Large Language Models
por: Sun, Wangtao, et al.
Publicado: (2025)
por: Sun, Wangtao, et al.
Publicado: (2025)
KInITVeraAI at SemEval-2023 Task 3: Simple yet Powerful Multilingual Fine-Tuning for Persuasion Techniques Detection
por: Hromadka, Timo, et al.
Publicado: (2023)
por: Hromadka, Timo, et al.
Publicado: (2023)
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
por: Zhang, Zhicheng, et al.
Publicado: (2025)
por: Zhang, Zhicheng, et al.
Publicado: (2025)
Zero-to-Strong Generalization: Eliciting Strong Capabilities of Large Language Models Iteratively without Gold Labels
por: Liu, Chaoqun, et al.
Publicado: (2024)
por: Liu, Chaoqun, et al.
Publicado: (2024)
Investigating the Multilingual Calibration Effects of Language Model Instruction-Tuning
por: Huang, Jerry, et al.
Publicado: (2026)
por: Huang, Jerry, et al.
Publicado: (2026)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
por: Cao, Yihan, et al.
Publicado: (2023)
por: Cao, Yihan, et al.
Publicado: (2023)
Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
por: Chen, Zixiang, et al.
Publicado: (2024)
por: Chen, Zixiang, et al.
Publicado: (2024)
AF Adapter: Continual Pretraining for Building Chinese Biomedical Language Model
por: Yan, Yongyu, et al.
Publicado: (2022)
por: Yan, Yongyu, et al.
Publicado: (2022)
On the Generalizability of "Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals"
por: Dotsinski, Asen, et al.
Publicado: (2025)
por: Dotsinski, Asen, et al.
Publicado: (2025)
Large Language Models for Math Education in Low-Resource Languages: A Study in Sinhala and Tamil
por: Kishanthan, Sukumar, et al.
Publicado: (2026)
por: Kishanthan, Sukumar, et al.
Publicado: (2026)
Large Language Models can be Strong Self-Detoxifiers
por: Ko, Ching-Yun, et al.
Publicado: (2024)
por: Ko, Ching-Yun, et al.
Publicado: (2024)
Generalizing Large Language Model Usability Across Resource-Constrained
por: Tsai, Yun-Da
Publicado: (2025)
por: Tsai, Yun-Da
Publicado: (2025)
Kakugo: Distillation of Low-Resource Languages into Small Language Models
por: Devine, Peter, et al.
Publicado: (2026)
por: Devine, Peter, et al.
Publicado: (2026)
Ejemplares similares
-
Improving LLMs for Machine Translation Using Synthetic Preference Data
por: Vajda, Dario, et al.
Publicado: (2025) -
Generative Model for Less-Resourced Language with 1 billion parameters
por: Vreš, Domen, et al.
Publicado: (2024) -
Sarcasm Detection in a Less-Resourced Language
por: Đoković, Lazar, et al.
Publicado: (2024) -
Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
por: Arčon, Tjaša, et al.
Publicado: (2026) -
Large language models for folktale type automation based on motifs: Cinderella case study
por: Arčon, Tjaša, et al.
Publicado: (2025)