Saved in:
| Main Authors: | Shapira, Ori, Chazan, Shlomo E., Cohen, Amir DN |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2502.13645 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Overlooked Role of Graded Relevance Thresholds in Multilingual Dense Retrieval
by: Wullach, Tomer, et al.
Published: (2026)
by: Wullach, Tomer, et al.
Published: (2026)
Information Types in Product Reviews
by: Shapira, Ori, et al.
Published: (2025)
by: Shapira, Ori, et al.
Published: (2025)
A Unifying Scheme for Extractive Content Selection Tasks
by: Amar, Shmuel, et al.
Published: (2025)
by: Amar, Shmuel, et al.
Published: (2025)
Dicta-LM 3.0: Advancing The Frontier of Hebrew Sovereign LLMs
by: Shmidman, Shaltiel, et al.
Published: (2026)
by: Shmidman, Shaltiel, et al.
Published: (2026)
Adapting LLMs to Hebrew: Unveiling DictaLM 2.0 with Enhanced Vocabulary and Instruction Capabilities
by: Shmidman, Shaltiel, et al.
Published: (2024)
by: Shmidman, Shaltiel, et al.
Published: (2024)
HeQ: a Large and Diverse Hebrew Reading Comprehension Benchmark
by: Cohen, Amir DN, et al.
Published: (2025)
by: Cohen, Amir DN, et al.
Published: (2025)
Diversity Over Quantity: A Lesson From Few Shot Relation Classification
by: Cohen, Amir DN, et al.
Published: (2024)
by: Cohen, Amir DN, et al.
Published: (2024)
SEAM: A Stochastic Benchmark for Multi-Document Tasks
by: Lior, Gili, et al.
Published: (2024)
by: Lior, Gili, et al.
Published: (2024)
Quality Matters: Evaluating Synthetic Data for Tool-Using LLMs
by: Iskander, Shadi, et al.
Published: (2024)
by: Iskander, Shadi, et al.
Published: (2024)
IQ Test for LLMs: An Evaluation Framework for Uncovering Core Skills in LLMs
by: Maimon, Aviya, et al.
Published: (2025)
by: Maimon, Aviya, et al.
Published: (2025)
Multi-Review Fusion-in-Context
by: Slobodkin, Aviv, et al.
Published: (2024)
by: Slobodkin, Aviv, et al.
Published: (2024)
HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model
by: Kayzer, Noam, et al.
Published: (2026)
by: Kayzer, Noam, et al.
Published: (2026)
Description-Based Text Similarity
by: Ravfogel, Shauli, et al.
Published: (2023)
by: Ravfogel, Shauli, et al.
Published: (2023)
Consensus or Conflict? Fine-Grained Evaluation of Conflicting Answers in Question-Answering
by: Nachshoni, Eviatar, et al.
Published: (2025)
by: Nachshoni, Eviatar, et al.
Published: (2025)
The Power of Summary-Source Alignments
by: Ernst, Ori, et al.
Published: (2024)
by: Ernst, Ori, et al.
Published: (2024)
The Degree of Language Diacriticity and Its Effect on Tasks
by: Cohen, Adi, et al.
Published: (2026)
by: Cohen, Adi, et al.
Published: (2026)
Scaling Laws for Downstream Task Performance of Large Language Models
by: Isik, Berivan, et al.
Published: (2024)
by: Isik, Berivan, et al.
Published: (2024)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
by: Li, Miaomiao, et al.
Published: (2025)
by: Li, Miaomiao, et al.
Published: (2025)
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks
by: Suganthan, Paul, et al.
Published: (2025)
by: Suganthan, Paul, et al.
Published: (2025)
Optimising Language Models for Downstream Tasks: A Post-Training Perspective
by: Shi, Zhengyan
Published: (2025)
by: Shi, Zhengyan
Published: (2025)
Instability in Downstream Task Performance During LLM Pretraining
by: Nishida, Yuto, et al.
Published: (2025)
by: Nishida, Yuto, et al.
Published: (2025)
Measuring Hong Kong Massive Multi-Task Language Understanding
by: Cao, Chuxue, et al.
Published: (2025)
by: Cao, Chuxue, et al.
Published: (2025)
An Evaluation of Sindhi Word Embedding in Semantic Analogies and Downstream Tasks
by: Ali, Wazir, et al.
Published: (2024)
by: Ali, Wazir, et al.
Published: (2024)
Edit Distances and Their Applications to Downstream Tasks in Research and Commercial Contexts
by: Carmo, Félix do, et al.
Published: (2024)
by: Carmo, Félix do, et al.
Published: (2024)
Making Retrieval-Augmented Language Models Robust to Irrelevant Context
by: Yoran, Ori, et al.
Published: (2023)
by: Yoran, Ori, et al.
Published: (2023)
TransformerRanker: A Tool for Efficiently Finding the Best-Suited Language Models for Downstream Classification Tasks
by: Garbas, Lukas, et al.
Published: (2024)
by: Garbas, Lukas, et al.
Published: (2024)
VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models
by: Zhang, Hanling, et al.
Published: (2025)
by: Zhang, Hanling, et al.
Published: (2025)
Applying Intrinsic Debiasing on Downstream Tasks: Challenges and Considerations for Machine Translation
by: Iluz, Bar, et al.
Published: (2024)
by: Iluz, Bar, et al.
Published: (2024)
Learning to Rewrite Prompts for Bootstrapping LLMs on Downstream Tasks
by: Zhou, Qinhao, et al.
Published: (2025)
by: Zhou, Qinhao, et al.
Published: (2025)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
by: Dymkiewicz, Kajetan, et al.
Published: (2025)
by: Dymkiewicz, Kajetan, et al.
Published: (2025)
Scaling Laws Are Unreliable for Downstream Tasks: A Reality Check
by: Lourie, Nicholas, et al.
Published: (2025)
by: Lourie, Nicholas, et al.
Published: (2025)
Can Character-based Language Models Improve Downstream Task Performance in Low-Resource and Noisy Language Scenarios?
by: Riabi, Arij, et al.
Published: (2021)
by: Riabi, Arij, et al.
Published: (2021)
DFPE: A Diverse Fingerprint Ensemble for Enhancing LLM Performance
by: Cohen, Seffi, et al.
Published: (2025)
by: Cohen, Seffi, et al.
Published: (2025)
Forget What You Know about LLMs Evaluations -- LLMs are Like a Chameleon
by: Cohen-Inger, Nurit, et al.
Published: (2025)
by: Cohen-Inger, Nurit, et al.
Published: (2025)
Not-Just-Scaling Laws: Towards a Better Understanding of the Downstream Impact of Language Model Design Decisions
by: Liu, Emmy, et al.
Published: (2025)
by: Liu, Emmy, et al.
Published: (2025)
Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks
by: Chen, Hao, et al.
Published: (2023)
by: Chen, Hao, et al.
Published: (2023)
FedEval-LLM: Federated Evaluation of Large Language Models on Downstream Tasks with Collective Wisdom
by: He, Yuanqin, et al.
Published: (2024)
by: He, Yuanqin, et al.
Published: (2024)
Measuring Bias or Measuring the Task: Understanding the Brittle Nature of LLM Gender Biases
by: Gao, Bufan, et al.
Published: (2025)
by: Gao, Bufan, et al.
Published: (2025)
CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks
by: Suharitdamrong, Wish, et al.
Published: (2026)
by: Suharitdamrong, Wish, et al.
Published: (2026)
Spectral or spatial? Leveraging both for speaker extraction in challenging data conditions
by: Eisenberg, Aviad, et al.
Published: (2025)
by: Eisenberg, Aviad, et al.
Published: (2025)
Similar Items
-
The Overlooked Role of Graded Relevance Thresholds in Multilingual Dense Retrieval
by: Wullach, Tomer, et al.
Published: (2026) -
Information Types in Product Reviews
by: Shapira, Ori, et al.
Published: (2025) -
A Unifying Scheme for Extractive Content Selection Tasks
by: Amar, Shmuel, et al.
Published: (2025) -
Dicta-LM 3.0: Advancing The Frontier of Hebrew Sovereign LLMs
by: Shmidman, Shaltiel, et al.
Published: (2026) -
Adapting LLMs to Hebrew: Unveiling DictaLM 2.0 with Enhanced Vocabulary and Instruction Capabilities
by: Shmidman, Shaltiel, et al.
Published: (2024)