The interplay between domain specialization and model size
Fuente:
arXiv
Saved in:
| Main Authors: | Junior, Roseval Malaquias, Pires, Ramon, Almeida, Thales Sales, Sakiyama, Kenzo, Romero, Roseli A. F., Nogueira, Rodrigo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Juru: Legal Brazilian Large Language Model from Reputable Sources
by: Junior, Roseval Malaquias, et al.
Published: (2024)
by: Junior, Roseval Malaquias, et al.
Published: (2024)
Automatic Legal Writing Evaluation of LLMs
by: Pires, Ramon, et al.
Published: (2025)
by: Pires, Ramon, et al.
Published: (2025)
Sabiá-3 Technical Report
by: Abonizio, Hugo, et al.
Published: (2024)
by: Abonizio, Hugo, et al.
Published: (2024)
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
by: Pires, Ramon, et al.
Published: (2026)
by: Pires, Ramon, et al.
Published: (2026)
Prosa: Rubric-Based Evaluation of LLMs on Real User Chats in Brazilian Portuguese
by: Junior, Roseval Malaquias, et al.
Published: (2026)
by: Junior, Roseval Malaquias, et al.
Published: (2026)
Sabiá-2: A New Generation of Portuguese Large Language Models
by: Almeida, Thales Sales, et al.
Published: (2024)
by: Almeida, Thales Sales, et al.
Published: (2024)
CAPITU: A Benchmark for Evaluating Instruction-Following in Brazilian Portuguese with Literary Context
by: Bonás, Giovana Kerche, et al.
Published: (2026)
by: Bonás, Giovana Kerche, et al.
Published: (2026)
MARCA: A Checklist-Based Benchmark for Multilingual Web Search
by: Almeida, Thales Sales, et al.
Published: (2026)
by: Almeida, Thales Sales, et al.
Published: (2026)
Sabiá-4 Technical Report
by: Laitz, Thiago, et al.
Published: (2026)
by: Laitz, Thiago, et al.
Published: (2026)
Measuring Opinion Bias and Sycophancy via LLM-based Persuasion
by: Nogueira, Rodrigo, et al.
Published: (2026)
by: Nogueira, Rodrigo, et al.
Published: (2026)
LLM-Based Persuasion Enables Guardrail Override in Frontier LLMs
by: Nogueira, Rodrigo, et al.
Published: (2026)
by: Nogueira, Rodrigo, et al.
Published: (2026)
TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
BLUEX Revisited: Enhancing Benchmark Coverage with Automatic Captioning
by: Santos, João Guilherme Alves, et al.
Published: (2025)
by: Santos, João Guilherme Alves, et al.
Published: (2025)
PoETa v2: Toward More Robust Evaluation of Large Language Models in Portuguese
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
Synthetic Rewriting as a Quality Multiplier: Evidence from Portuguese Continued Pretraining
by: Almeida, Thales Sales, et al.
Published: (2026)
by: Almeida, Thales Sales, et al.
Published: (2026)
Building High-Quality Datasets for Portuguese LLMs: From Common Crawl Snapshots to Industrial-Grade Corpora
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
Curió-Edu 7B: Examining Data Selection Impacts in LLM Continued Pretraining
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
Explainable LightGBM Approach for Predicting Myocardial Infarction Mortality
by: Vicente, Ana Letícia Garcez, et al.
Published: (2024)
by: Vicente, Ana Letícia Garcez, et al.
Published: (2024)
Lissard: Long and Simple Sequential Reasoning Datasets
by: Bueno, Mirelle, et al.
Published: (2024)
by: Bueno, Mirelle, et al.
Published: (2024)
Generalist embedding models are better at short-context clinical semantic search than specialized embedding models
by: Excoffier, Jean-Baptiste, et al.
Published: (2024)
by: Excoffier, Jean-Baptiste, et al.
Published: (2024)
Measuring Cross-lingual Transfer in Bytes
by: de Souza, Leandro Rodrigues, et al.
Published: (2024)
by: de Souza, Leandro Rodrigues, et al.
Published: (2024)
Large language models in healthcare and medical domain: A review
by: Nazi, Zabir Al, et al.
Published: (2023)
by: Nazi, Zabir Al, et al.
Published: (2023)
ptt5-v2: A Closer Look at Continued Pretraining of T5 Models for the Portuguese Language
by: Piau, Marcos, et al.
Published: (2024)
by: Piau, Marcos, et al.
Published: (2024)
An overview of domain-specific foundation model: key technologies, applications and challenges
by: Chen, Haolong, et al.
Published: (2024)
by: Chen, Haolong, et al.
Published: (2024)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
by: Barboule, Camille, et al.
Published: (2024)
by: Barboule, Camille, et al.
Published: (2024)
SER Evals: In-domain and Out-of-domain Benchmarking for Speech Emotion Recognition
by: Osman, Mohamed, et al.
Published: (2024)
by: Osman, Mohamed, et al.
Published: (2024)
Towards Fairness Assessment of Dutch Hate Speech Detection
by: Bauer, Julie, et al.
Published: (2025)
by: Bauer, Julie, et al.
Published: (2025)
StatLLaMA: Multi-Stage training for domain-optimized statistical large language models
by: Zeng, Jing-Yi, et al.
Published: (2025)
by: Zeng, Jing-Yi, et al.
Published: (2025)
ExaRanker-Open: Synthetic Explanation for IR using Open-Source LLMs
by: Ferraretto, Fernando, et al.
Published: (2024)
by: Ferraretto, Fernando, et al.
Published: (2024)
Predictive Authoring for Brazilian Portuguese Augmentative and Alternative Communication
by: Pereira, Jayr, et al.
Published: (2023)
by: Pereira, Jayr, et al.
Published: (2023)
SteuerLLM: Local specialized large language model for German tax law analysis
by: Wind, Sebastian, et al.
Published: (2026)
by: Wind, Sebastian, et al.
Published: (2026)
DepressLLM: Interpretable domain-adapted language model for depression detection from real-world narratives
by: Moon, Sehwan, et al.
Published: (2025)
by: Moon, Sehwan, et al.
Published: (2025)
Emission-GPT: A domain-specific language model agent for knowledge retrieval, emission inventory and data analysis
by: Ye, Jiashu, et al.
Published: (2025)
by: Ye, Jiashu, et al.
Published: (2025)
SAMGPT: Text-free Graph Foundation Model for Multi-domain Pre-training and Cross-domain Adaptation
by: Yu, Xingtong, et al.
Published: (2025)
by: Yu, Xingtong, et al.
Published: (2025)
Cross-domain Chinese Sentence Pattern Parsing
by: Yu, Jingsi, et al.
Published: (2024)
by: Yu, Jingsi, et al.
Published: (2024)
In-domain SSL pre-training and streaming ASR
by: Duret, Jarod, et al.
Published: (2025)
by: Duret, Jarod, et al.
Published: (2025)
Comparing Knowledge Injection Methods for LLMs in a Low-Resource Regime
by: Abonizio, Hugo, et al.
Published: (2025)
by: Abonizio, Hugo, et al.
Published: (2025)
Matching domain experts by training from scratch on domain knowledge
by: Luo, Xiaoliang, et al.
Published: (2024)
by: Luo, Xiaoliang, et al.
Published: (2024)
COGNET-MD, an evaluation framework and dataset for Large Language Model benchmarks in the medical domain
by: Panagoulias, Dimitrios P., et al.
Published: (2024)
by: Panagoulias, Dimitrios P., et al.
Published: (2024)
MOSAIC: Masked Objective with Selective Adaptation for In-domain Contrastive Learning
by: Pavlova, Vera, et al.
Published: (2025)
by: Pavlova, Vera, et al.
Published: (2025)
Similar Items
-
Juru: Legal Brazilian Large Language Model from Reputable Sources
by: Junior, Roseval Malaquias, et al.
Published: (2024) -
Automatic Legal Writing Evaluation of LLMs
by: Pires, Ramon, et al.
Published: (2025) -
Sabiá-3 Technical Report
by: Abonizio, Hugo, et al.
Published: (2024) -
Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks
by: Pires, Ramon, et al.
Published: (2026) -
Prosa: Rubric-Based Evaluation of LLMs on Real User Chats in Brazilian Portuguese
by: Junior, Roseval Malaquias, et al.
Published: (2026)