MRL Parsing Without Tears: The Case of Hebrew
Fuente:
arXiv
Saved in:
| Main Authors: | Shmidman, Shaltiel, Shmidman, Avi, Koppel, Moshe, Tsarfaty, Reut |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NeoDictaBERT: Pushing the Frontier of BERT models for Hebrew
by: Shmidman, Shaltiel, et al.
Published: (2025)
by: Shmidman, Shaltiel, et al.
Published: (2025)
Do Pretrained Contextual Language Models Distinguish between Hebrew Homograph Analyses?
by: Shmidman, Avi, et al.
Published: (2024)
by: Shmidman, Avi, et al.
Published: (2024)
Dicta-LM 3.0: Advancing The Frontier of Hebrew Sovereign LLMs
by: Shmidman, Shaltiel, et al.
Published: (2026)
by: Shmidman, Shaltiel, et al.
Published: (2026)
Adapting LLMs to Hebrew: Unveiling DictaLM 2.0 with Enhanced Vocabulary and Instruction Capabilities
by: Shmidman, Shaltiel, et al.
Published: (2024)
by: Shmidman, Shaltiel, et al.
Published: (2024)
MsBERT: A New Model for the Reconstruction of Lacunae in Hebrew Manuscripts
by: Shmidman, Avi, et al.
Published: (2025)
by: Shmidman, Avi, et al.
Published: (2025)
Splintering Nonconcatenative Languages for Better Tokenization
by: Gazit, Bar, et al.
Published: (2025)
by: Gazit, Bar, et al.
Published: (2025)
Learning to Reason: Training LLMs with GPT-OSS or DeepSeek R1 Reasoning Traces
by: Shmidman, Shaltiel, et al.
Published: (2025)
by: Shmidman, Shaltiel, et al.
Published: (2025)
Diversity Over Quantity: A Lesson From Few Shot Relation Classification
by: Cohen, Amir DN, et al.
Published: (2024)
by: Cohen, Amir DN, et al.
Published: (2024)
Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text
by: Greenfeld, Refael Shaked, et al.
Published: (2026)
by: Greenfeld, Refael Shaked, et al.
Published: (2026)
A Truly Joint Neural Architecture for Segmentation and Parsing
by: Levi, Danit Yshaayahu, et al.
Published: (2024)
by: Levi, Danit Yshaayahu, et al.
Published: (2024)
HeQ: a Large and Diverse Hebrew Reading Comprehension Benchmark
by: Cohen, Amir DN, et al.
Published: (2025)
by: Cohen, Amir DN, et al.
Published: (2025)
HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model
by: Kayzer, Noam, et al.
Published: (2026)
by: Kayzer, Noam, et al.
Published: (2026)
HeSum: a Novel Dataset for Abstractive Text Summarization in Hebrew
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
A Novel Computational and Modeling Foundation for Automatic Coherence Assessment
by: Maimon, Aviya, et al.
Published: (2023)
by: Maimon, Aviya, et al.
Published: (2023)
Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)
by: Manevich, Avshalom, et al.
Published: (2024)
by: Manevich, Avshalom, et al.
Published: (2024)
LLMs' Reading Comprehension Is Affected by Parametric Knowledge and Struggles with Hypothetical Statements
by: Basmov, Victoria, et al.
Published: (2024)
by: Basmov, Victoria, et al.
Published: (2024)
Simple Linguistic Inferences of Large Language Models (LLMs): Blind Spots and Blinds
by: Basmov, Victoria, et al.
Published: (2023)
by: Basmov, Victoria, et al.
Published: (2023)
NoviCode: Generating Programs from Natural Language Utterances by Novices
by: Mordechai, Asaf Achi, et al.
Published: (2024)
by: Mordechai, Asaf Achi, et al.
Published: (2024)
Beyond N-Grams: Rethinking Evaluation Metrics and Strategies for Multilingual Abstractive Summarization
by: Mondshine, Itai, et al.
Published: (2025)
by: Mondshine, Itai, et al.
Published: (2025)
Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs
by: Mondshine, Itai, et al.
Published: (2025)
by: Mondshine, Itai, et al.
Published: (2025)
To MRL or not to MRL: Text Embeddings are Robust to Truncation Without Matryoshka Learning, Except In Heavy Truncation Scenarios
by: Takeshita, Sotaro, et al.
Published: (2026)
by: Takeshita, Sotaro, et al.
Published: (2026)
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
by: Pyatkin, Valentina, et al.
Published: (2024)
by: Pyatkin, Valentina, et al.
Published: (2024)
Unpacking Tokenization: Evaluating Text Compression and its Correlation with Model Performance
by: Goldman, Omer, et al.
Published: (2024)
by: Goldman, Omer, et al.
Published: (2024)
IQ Test for LLMs: An Evaluation Framework for Uncovering Core Skills in LLMs
by: Maimon, Aviya, et al.
Published: (2025)
by: Maimon, Aviya, et al.
Published: (2025)
Into the Unknown: Generating Geospatial Descriptions for New Environments
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
Where Do We Go from Here? Multi-scale Allocentric Relational Inference from Natural Spatial Descriptions
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
by: Paz-Argaman, Tzuf, et al.
Published: (2024)
Effective QA-driven Annotation of Predicate-Argument Relations Across Languages
by: Davidov, Jonathan, et al.
Published: (2026)
by: Davidov, Jonathan, et al.
Published: (2026)
Is It Really Long Context if All You Need Is Retrieval? Towards Genuinely Difficult Long Context NLP
by: Goldman, Omer, et al.
Published: (2024)
by: Goldman, Omer, et al.
Published: (2024)
Multilingual Instruction Tuning With Just a Pinch of Multilinguality
by: Shaham, Uri, et al.
Published: (2024)
by: Shaham, Uri, et al.
Published: (2024)
Breaking the Language Barrier: Can Direct Inference Outperform Pre-Translation in Multilingual LLM Applications?
by: Intrator, Yotam, et al.
Published: (2024)
by: Intrator, Yotam, et al.
Published: (2024)
MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents
by: Wolfson, Tomer, et al.
Published: (2025)
by: Wolfson, Tomer, et al.
Published: (2025)
Training a Huggingface Model on AWS Sagemaker (Without Tears)
by: Tan, Liling
Published: (2025)
by: Tan, Liling
Published: (2025)
Modelling the Morphology of Verbal Paradigms: A Case Study in the Tokenization of Turkish and Hebrew
by: Samo, Giuseppe, et al.
Published: (2026)
by: Samo, Giuseppe, et al.
Published: (2026)
SandboxAQ's submission to MRL 2024 Shared Task on Multi-lingual Multi-task Information Retrieval
by: Tourni, Isidora Chara, et al.
Published: (2024)
by: Tourni, Isidora Chara, et al.
Published: (2024)
Culturally Grounded Physical Commonsense Reasoning in Italian and English: A Submission to the MRL 2025 Shared Task
by: De Santis, Marco, et al.
Published: (2025)
by: De Santis, Marco, et al.
Published: (2025)
Prefix Parsing is Just Parsing
by: Pasti, Clemente, et al.
Published: (2026)
by: Pasti, Clemente, et al.
Published: (2026)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
MenakBERT -- Hebrew Diacriticizer
by: Cohen, Ido, et al.
Published: (2024)
by: Cohen, Ido, et al.
Published: (2024)
Hebrew Diacritics Restoration using Visual Representation
by: Elboher, Yair, et al.
Published: (2025)
by: Elboher, Yair, et al.
Published: (2025)
Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs
by: Mor-Lan, Guy, et al.
Published: (2026)
by: Mor-Lan, Guy, et al.
Published: (2026)
Similar Items
-
NeoDictaBERT: Pushing the Frontier of BERT models for Hebrew
by: Shmidman, Shaltiel, et al.
Published: (2025) -
Do Pretrained Contextual Language Models Distinguish between Hebrew Homograph Analyses?
by: Shmidman, Avi, et al.
Published: (2024) -
Dicta-LM 3.0: Advancing The Frontier of Hebrew Sovereign LLMs
by: Shmidman, Shaltiel, et al.
Published: (2026) -
Adapting LLMs to Hebrew: Unveiling DictaLM 2.0 with Enhanced Vocabulary and Instruction Capabilities
by: Shmidman, Shaltiel, et al.
Published: (2024) -
MsBERT: A New Model for the Reconstruction of Lacunae in Hebrew Manuscripts
by: Shmidman, Avi, et al.
Published: (2025)