Hala Technical Report: Building Arabic-Centric Instruction & Translation Models at Scale
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hammoud, Hasan Abed Al Kader, Zbeeb, Mohammad, Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
par: Zbeeb, Mohammad, et autres
Publié: (2025)
par: Zbeeb, Mohammad, et autres
Publié: (2025)
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
DiffCLIP: Differential Attention Meets CLIP
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
par: Shairah, Harethah Abu, et autres
Publié: (2025)
par: Shairah, Harethah Abu, et autres
Publié: (2025)
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection
par: Shairah, Harethah Abu, et autres
Publié: (2025)
par: Shairah, Harethah Abu, et autres
Publié: (2025)
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
par: Zbeeb, Mohammad, et autres
Publié: (2025)
par: Zbeeb, Mohammad, et autres
Publié: (2025)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
On the Importance of Pretraining Data Alignment for Atomic Property Prediction
par: Ghunaim, Yasir, et autres
Publié: (2025)
par: Ghunaim, Yasir, et autres
Publié: (2025)
TAPS: Task Aware Proposal Distributions for Speculative Sampling
par: Zbib, Mohamad, et autres
Publié: (2026)
par: Zbib, Mohamad, et autres
Publié: (2026)
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
par: Alssum, Lama, et autres
Publié: (2025)
par: Alssum, Lama, et autres
Publié: (2025)
On Pretraining Data Diversity for Self-Supervised Learning
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2024)
QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation
par: Slim, Ali, et autres
Publié: (2026)
par: Slim, Ali, et autres
Publié: (2026)
Building Large-Scale English-Romanian Literary Translation Resources with Open Models
par: Nadas, Mihai, et autres
Publié: (2025)
par: Nadas, Mihai, et autres
Publié: (2025)
Benchmarking the Medical Understanding and Reasoning of Large Language Models in Arabic Healthcare Tasks
par: AlDahoul, Nouar, et autres
Publié: (2025)
par: AlDahoul, Nouar, et autres
Publié: (2025)
On Translating Technical Terminology: A Translation Workflow for Machine-Translated Acronyms
par: Yue, Richard, et autres
Publié: (2024)
par: Yue, Richard, et autres
Publié: (2024)
Exploring the Landscape for Generative Sequence Models for Specialized Data Synthesis
par: Zbeeb, Mohammad, et autres
Publié: (2024)
par: Zbeeb, Mohammad, et autres
Publié: (2024)
Forget Less, Retain More: A Lightweight Regularizer for Rehearsal-Based Continual Learning
par: Alssum, Lama, et autres
Publié: (2025)
par: Alssum, Lama, et autres
Publié: (2025)
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
par: Hinojosa, Carlos, et autres
Publié: (2026)
par: Hinojosa, Carlos, et autres
Publié: (2026)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
par: AlDahoul, Nouar, et autres
Publié: (2025)
par: AlDahoul, Nouar, et autres
Publié: (2025)
Multiple Noises in Diffusion Model for Semi-Supervised Multi-Domain Translation
par: Mayet, Tsiry, et autres
Publié: (2023)
par: Mayet, Tsiry, et autres
Publié: (2023)
Yi-Lightning Technical Report
par: Wake, Alan, et autres
Publié: (2024)
par: Wake, Alan, et autres
Publié: (2024)
Better Alignment with Instruction Back-and-Forth Translation
par: Nguyen, Thao, et autres
Publié: (2024)
par: Nguyen, Thao, et autres
Publié: (2024)
Nemotron-4 15B Technical Report
par: Parmar, Jupinder, et autres
Publié: (2024)
par: Parmar, Jupinder, et autres
Publié: (2024)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
par: AlDahoul, Nouar, et autres
Publié: (2025)
par: AlDahoul, Nouar, et autres
Publié: (2025)
Technical Report: Small Language Model for Japanese Clinical and Medicine
par: Watanabe, Shogo
Publié: (2024)
par: Watanabe, Shogo
Publié: (2024)
Arabic Little STT: Arabic Children Speech Recognition Dataset
par: Alkadri, Mouhand, et autres
Publié: (2025)
par: Alkadri, Mouhand, et autres
Publié: (2025)
PLaMo 2 Technical Report
par: Networks, Preferred, et autres
Publié: (2025)
par: Networks, Preferred, et autres
Publié: (2025)
The Zamba2 Suite: Technical Report
par: Glorioso, Paolo, et autres
Publié: (2024)
par: Glorioso, Paolo, et autres
Publié: (2024)
Trillion 7B Technical Report
par: Han, Sungjun, et autres
Publié: (2025)
par: Han, Sungjun, et autres
Publié: (2025)
Nemotron-4 340B Technical Report
par: Nvidia, et autres
Publié: (2024)
par: Nvidia, et autres
Publié: (2024)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
par: Tang, Zhengyang, et autres
Publié: (2024)
par: Tang, Zhengyang, et autres
Publié: (2024)
Skywork Open Reasoner 1 Technical Report
par: He, Jujie, et autres
Publié: (2025)
par: He, Jujie, et autres
Publié: (2025)
EuroLLM-22B: Technical Report
par: Ramos, Miguel Moura, et autres
Publié: (2026)
par: Ramos, Miguel Moura, et autres
Publié: (2026)
EuroLLM-9B: Technical Report
par: Martins, Pedro Henrique, et autres
Publié: (2025)
par: Martins, Pedro Henrique, et autres
Publié: (2025)
ALHD: A Large-Scale and Multigenre Benchmark Dataset for Arabic LLM-Generated Text Detection
par: Khairallah, Ali, et autres
Publié: (2025)
par: Khairallah, Ali, et autres
Publié: (2025)
Nile-Chat: Egyptian Language Models for Arabic and Latin Scripts
par: Shang, Guokan, et autres
Publié: (2025)
par: Shang, Guokan, et autres
Publié: (2025)
ArzEn-LLM: Code-Switched Egyptian Arabic-English Translation and Speech Recognition Using LLMs
par: Heakl, Ahmed, et autres
Publié: (2024)
par: Heakl, Ahmed, et autres
Publié: (2024)
Diffusion Language Models Can Perform Many Tasks with Scaling and Instruction-Finetuning
par: Ye, Jiasheng, et autres
Publié: (2023)
par: Ye, Jiasheng, et autres
Publié: (2023)
Documents similaires
-
AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models
par: Zbeeb, Mohammad, et autres
Publié: (2025) -
Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025) -
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025) -
DiffCLIP: Differential Attention Meets CLIP
par: Hammoud, Hasan Abed Al Kader, et autres
Publié: (2025) -
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
par: Shairah, Harethah Abu, et autres
Publié: (2025)