ChocoLlama: Lessons Learned From Teaching Llamas Dutch
Fuente:
arXiv
Saved in:
| Main Authors: | Meeus, Matthieu, Rathé, Anthony, Remy, François, Delobelle, Pieter, Decorte, Jens-Joris, Demeester, Thomas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SkillMatch: Evaluating Self-supervised Learning of Skill Relatedness
by: Decorte, Jens-Joris, et al.
Published: (2024)
by: Decorte, Jens-Joris, et al.
Published: (2024)
Trans-Tokenization and Cross-lingual Vocabulary Transfers: Language Adaptation of LLMs for Low-Resource NLP
by: Remy, François, et al.
Published: (2024)
by: Remy, François, et al.
Published: (2024)
On the Biased Assessment of Expert Finding Systems
by: Decorte, Jens-Joris, et al.
Published: (2024)
by: Decorte, Jens-Joris, et al.
Published: (2024)
Efficient Text Encoders for Labor Market Analysis
by: Decorte, Jens-Joris, et al.
Published: (2025)
by: Decorte, Jens-Joris, et al.
Published: (2025)
Unified Work Embeddings: Contrastive Learning of a Bidirectional Multi-task Ranker
by: De Lange, Matthias, et al.
Published: (2025)
by: De Lange, Matthias, et al.
Published: (2025)
Llama See, Llama Do: A Mechanistic Perspective on Contextual Entrainment and Distraction in LLMs
by: Niu, Jingcheng, et al.
Published: (2025)
by: Niu, Jingcheng, et al.
Published: (2025)
MGH Radiology Llama: A Llama 3 70B Model for Radiology
by: Shi, Yucheng, et al.
Published: (2024)
by: Shi, Yucheng, et al.
Published: (2024)
To Err Is Human, but Llamas Can Learn It Too
by: Luhtaru, Agnes, et al.
Published: (2024)
by: Luhtaru, Agnes, et al.
Published: (2024)
BadLlama: cheaply removing safety fine-tuning from Llama 2-Chat 13B
by: Gade, Pranav, et al.
Published: (2023)
by: Gade, Pranav, et al.
Published: (2023)
Llama-Mob: Instruction-Tuning Llama-3-8B Excels in City-Scale Mobility Prediction
by: Tang, Peizhi, et al.
Published: (2024)
by: Tang, Peizhi, et al.
Published: (2024)
Teaching Llama a New Language Through Cross-Lingual Knowledge Transfer
by: Kuulmets, Hele-Andra, et al.
Published: (2024)
by: Kuulmets, Hele-Andra, et al.
Published: (2024)
Multilingual JobBERT for Cross-Lingual Job Title Matching
by: Decorte, Jens-Joris, et al.
Published: (2025)
by: Decorte, Jens-Joris, et al.
Published: (2025)
The Llama 3 Herd of Models
by: Grattafiori, Aaron, et al.
Published: (2024)
by: Grattafiori, Aaron, et al.
Published: (2024)
Code Llama: Open Foundation Models for Code
by: Rozière, Baptiste, et al.
Published: (2023)
by: Rozière, Baptiste, et al.
Published: (2023)
Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders
by: He, Zhengfu, et al.
Published: (2024)
by: He, Zhengfu, et al.
Published: (2024)
Feedback Indicators: The Alignment between Llama and a Teacher in Language Learning
by: Rüdian, Sylvio, et al.
Published: (2025)
by: Rüdian, Sylvio, et al.
Published: (2025)
Extending Llama-3's Context Ten-Fold Overnight
by: Zhang, Peitian, et al.
Published: (2024)
by: Zhang, Peitian, et al.
Published: (2024)
SinLlama -- A Large Language Model for Sinhala
by: Aravinda, H. W. K., et al.
Published: (2025)
by: Aravinda, H. W. K., et al.
Published: (2025)
Efficient Speculative Decoding for Llama at Scale: Challenges and Solutions
by: Tang, Bangsheng, et al.
Published: (2025)
by: Tang, Bangsheng, et al.
Published: (2025)
Llama-Nemotron: Efficient Reasoning Models
by: Bercovich, Akhiad, et al.
Published: (2025)
by: Bercovich, Akhiad, et al.
Published: (2025)
OneLove beyond the field -- A few-shot pipeline for topic and sentiment analysis during the FIFA World Cup in Qatar
by: Rauchegger, Christoph, et al.
Published: (2024)
by: Rauchegger, Christoph, et al.
Published: (2024)
MobiLlama: Towards Accurate and Lightweight Fully Transparent GPT
by: Thawakar, Omkar, et al.
Published: (2024)
by: Thawakar, Omkar, et al.
Published: (2024)
In-Context Learning for Extreme Multi-Label Classification
by: D'Oosterlinck, Karel, et al.
Published: (2024)
by: D'Oosterlinck, Karel, et al.
Published: (2024)
Llama-Mimi: Exploring the Limits of Flattened Speech Language Modeling
by: Sugiura, Issa, et al.
Published: (2025)
by: Sugiura, Issa, et al.
Published: (2025)
Do Llamas Work in English? On the Latent Language of Multilingual Transformers
by: Wendler, Chris, et al.
Published: (2024)
by: Wendler, Chris, et al.
Published: (2024)
TinyLlama: An Open-Source Small Language Model
by: Zhang, Peiyuan, et al.
Published: (2024)
by: Zhang, Peiyuan, et al.
Published: (2024)
Zebra-Llama: Towards Extremely Efficient Hybrid Models
by: Yang, Mingyu, et al.
Published: (2025)
by: Yang, Mingyu, et al.
Published: (2025)
BanglaLlama: LLaMA for Bangla Language
by: Zehady, Abdullah Khan, et al.
Published: (2024)
by: Zehady, Abdullah Khan, et al.
Published: (2024)
Open Llama2 Model for the Lithuanian Language
by: Nakvosas, Artūras, et al.
Published: (2024)
by: Nakvosas, Artūras, et al.
Published: (2024)
Chinese-Vicuna: A Chinese Instruction-following Llama-based Model
by: Fan, Chenghao, et al.
Published: (2025)
by: Fan, Chenghao, et al.
Published: (2025)
Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness
by: Feng, Xincan, et al.
Published: (2024)
by: Feng, Xincan, et al.
Published: (2024)
Steering Llama 2 via Contrastive Activation Addition
by: Panickssery, Nina, et al.
Published: (2023)
by: Panickssery, Nina, et al.
Published: (2023)
Fearful Falcons and Angry Llamas: Emotion Category Annotations of Arguments by Humans and LLMs
by: Greschner, Lynn, et al.
Published: (2024)
by: Greschner, Lynn, et al.
Published: (2024)
Llama meets EU: Investigating the European Political Spectrum through the Lens of LLMs
by: Chalkidis, Ilias, et al.
Published: (2024)
by: Chalkidis, Ilias, et al.
Published: (2024)
Extending LLMs to New Languages: A Case Study of Llama and Persian Adaptation
by: Sani, Samin Mahdizadeh, et al.
Published: (2024)
by: Sani, Samin Mahdizadeh, et al.
Published: (2024)
Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection
by: Chen, Tianxiang, et al.
Published: (2024)
by: Chen, Tianxiang, et al.
Published: (2024)
Llama2Vec: Unsupervised Adaptation of Large Language Models for Dense Retrieval
by: Liu, Zheng, et al.
Published: (2023)
by: Liu, Zheng, et al.
Published: (2023)
TableLlama: Towards Open Large Generalist Models for Tables
by: Zhang, Tianshu, et al.
Published: (2023)
by: Zhang, Tianshu, et al.
Published: (2023)
Investigating Bias Representations in Llama 2 Chat via Activation Steering
by: Lu, Dawn, et al.
Published: (2024)
by: Lu, Dawn, et al.
Published: (2024)
LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
by: Zheng, Yaowei, et al.
Published: (2024)
by: Zheng, Yaowei, et al.
Published: (2024)
Similar Items
-
SkillMatch: Evaluating Self-supervised Learning of Skill Relatedness
by: Decorte, Jens-Joris, et al.
Published: (2024) -
Trans-Tokenization and Cross-lingual Vocabulary Transfers: Language Adaptation of LLMs for Low-Resource NLP
by: Remy, François, et al.
Published: (2024) -
On the Biased Assessment of Expert Finding Systems
by: Decorte, Jens-Joris, et al.
Published: (2024) -
Efficient Text Encoders for Labor Market Analysis
by: Decorte, Jens-Joris, et al.
Published: (2025) -
Unified Work Embeddings: Contrastive Learning of a Bidirectional Multi-task Ranker
by: De Lange, Matthias, et al.
Published: (2025)