Comparing Knowledge Injection Methods for LLMs in a Low-Resource Regime
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abonizio, Hugo, Almeida, Thales, Lotufo, Roberto, Nogueira, Rodrigo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines
von: Abonizio, Hugo, et al.
Veröffentlicht: (2026)
von: Abonizio, Hugo, et al.
Veröffentlicht: (2026)
Sabiá-2: A New Generation of Portuguese Large Language Models
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2024)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2024)
PoETa v2: Toward More Robust Evaluation of Large Language Models in Portuguese
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
Measuring Cross-lingual Transfer in Bytes
von: de Souza, Leandro Rodrigues, et al.
Veröffentlicht: (2024)
von: de Souza, Leandro Rodrigues, et al.
Veröffentlicht: (2024)
TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
MLissard: Multilingual Long and Simple Sequential Reasoning Benchmarks
von: Bueno, Mirelle, et al.
Veröffentlicht: (2024)
von: Bueno, Mirelle, et al.
Veröffentlicht: (2024)
Building High-Quality Datasets for Portuguese LLMs: From Common Crawl Snapshots to Industrial-Grade Corpora
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
von: Pires, Ramon, et al.
Veröffentlicht: (2026)
von: Pires, Ramon, et al.
Veröffentlicht: (2026)
LLM-Based Persuasion Enables Guardrail Override in Frontier LLMs
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
Prosa: Rubric-Based Evaluation of LLMs on Real User Chats in Brazilian Portuguese
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2026)
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2026)
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs
von: Ferraretto, Fernando, et al.
Veröffentlicht: (2024)
von: Ferraretto, Fernando, et al.
Veröffentlicht: (2024)
Lissard: Long and Simple Sequential Reasoning Datasets
von: Bueno, Mirelle, et al.
Veröffentlicht: (2024)
von: Bueno, Mirelle, et al.
Veröffentlicht: (2024)
Sabiá-3 Technical Report
von: Abonizio, Hugo, et al.
Veröffentlicht: (2024)
von: Abonizio, Hugo, et al.
Veröffentlicht: (2024)
SurveySum: A Dataset for Summarizing Multiple Scientific Articles into a Survey Section
von: Fernandes, Leandro Carísio, et al.
Veröffentlicht: (2024)
von: Fernandes, Leandro Carísio, et al.
Veröffentlicht: (2024)
ptt5-v2: A Closer Look at Continued Pretraining of T5 Models for the Portuguese Language
von: Piau, Marcos, et al.
Veröffentlicht: (2024)
von: Piau, Marcos, et al.
Veröffentlicht: (2024)
Synthetic Rewriting as a Quality Multiplier: Evidence from Portuguese Continued Pretraining
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2026)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2026)
CAPITU: A Benchmark for Evaluating Instruction-Following in Brazilian Portuguese with Literary Context
von: Bonás, Giovana Kerche, et al.
Veröffentlicht: (2026)
von: Bonás, Giovana Kerche, et al.
Veröffentlicht: (2026)
MARCA: A Checklist-Based Benchmark for Multilingual Web Search
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2026)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2026)
Sabiá-4 Technical Report
von: Laitz, Thiago, et al.
Veröffentlicht: (2026)
von: Laitz, Thiago, et al.
Veröffentlicht: (2026)
Curió-Edu 7B: Examining Data Selection Impacts in LLM Continued Pretraining
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
Measuring Opinion Bias and Sycophancy via LLM-based Persuasion
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
Fine-Tuning or Retrieval? Comparing Knowledge Injection in LLMs
von: Ovadia, Oded, et al.
Veröffentlicht: (2023)
von: Ovadia, Oded, et al.
Veröffentlicht: (2023)
Automated Clinical Report Generation for Remote Cognitive Remediation: Comparing Knowledge-Engineered Templates and LLMs in Low-Resource Settings
von: Zhou, Yongxin, et al.
Veröffentlicht: (2026)
von: Zhou, Yongxin, et al.
Veröffentlicht: (2026)
Memorization and Knowledge Injection in Gated LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
BRoverbs -- Measuring how much LLMs understand Portuguese proverbs
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)
Check-Eval: A Checklist-based Approach for Evaluating Text Quality
von: Pereira, Jayr, et al.
Veröffentlicht: (2024)
von: Pereira, Jayr, et al.
Veröffentlicht: (2024)
Comparative Analysis of Different Efficient Fine Tuning Methods of Large Language Models (LLMs) in Low-Resource Setting
von: Srinivasan, Krishna Prasad Varadarajan, et al.
Veröffentlicht: (2024)
von: Srinivasan, Krishna Prasad Varadarajan, et al.
Veröffentlicht: (2024)
Efficient Knowledge Injection in LLMs via Self-Distillation
von: Kujanpää, Kalle, et al.
Veröffentlicht: (2024)
von: Kujanpää, Kalle, et al.
Veröffentlicht: (2024)
LinGO: A Linguistic Graph Optimization Framework with LLMs for Interpreting Intents of Online Uncivil Discourse
von: Zhang, Yuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuan, et al.
Veröffentlicht: (2026)
Mixture of Experts for Low-Resource LLMs
von: Joseph, Ori Bar, et al.
Veröffentlicht: (2026)
von: Joseph, Ori Bar, et al.
Veröffentlicht: (2026)
Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
The interplay between domain specialization and model size
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2025)
von: Junior, Roseval Malaquias, et al.
Veröffentlicht: (2025)
Diffusion-Inspired Masked Fine-Tuning for Knowledge Injection in Autoregressive LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
LLM Probe: Evaluating LLMs for Low-Resource Languages
von: Teklehaymanot, Hailay Kidu, et al.
Veröffentlicht: (2026)
von: Teklehaymanot, Hailay Kidu, et al.
Veröffentlicht: (2026)
LLMs for Extremely Low-Resource Finno-Ugric Languages
von: Purason, Taido, et al.
Veröffentlicht: (2024)
von: Purason, Taido, et al.
Veröffentlicht: (2024)
INACIA: Integrating Large Language Models in Brazilian Audit Courts: Opportunities and Challenges
von: Pereira, Jayr, et al.
Veröffentlicht: (2024)
von: Pereira, Jayr, et al.
Veröffentlicht: (2024)
Automatic Legal Writing Evaluation of LLMs
von: Pires, Ramon, et al.
Veröffentlicht: (2025)
von: Pires, Ramon, et al.
Veröffentlicht: (2025)
Comparative Analysis of Tokenization Algorithms for Low-Resource Language Dzongkha
von: Wangchuk, Tandin, et al.
Veröffentlicht: (2025)
von: Wangchuk, Tandin, et al.
Veröffentlicht: (2025)
Compensating for Data with Reasoning: Low-Resource Machine Translation with LLMs
von: Frontull, Samuel, et al.
Veröffentlicht: (2025)
von: Frontull, Samuel, et al.
Veröffentlicht: (2025)
Vuyko Mistral: Adapting LLMs for Low-Resource Dialectal Translation
von: Kyslyi, Roman, et al.
Veröffentlicht: (2025)
von: Kyslyi, Roman, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines
von: Abonizio, Hugo, et al.
Veröffentlicht: (2026) -
Sabiá-2: A New Generation of Portuguese Large Language Models
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2024) -
PoETa v2: Toward More Robust Evaluation of Large Language Models in Portuguese
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025) -
Measuring Cross-lingual Transfer in Bytes
von: de Souza, Leandro Rodrigues, et al.
Veröffentlicht: (2024) -
TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time
von: Almeida, Thales Sales, et al.
Veröffentlicht: (2025)