Linguini: A benchmark for language-agnostic linguistic reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Sánchez, Eduardo, Alastruey, Belen, Ropers, Christophe, Stenetorp, Pontus, Artetxe, Mikel, Costa-jussà, Marta R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Gender-specific Machine Translation with Large Language Models
di: Sánchez, Eduardo, et al.
Pubblicazione: (2023)
di: Sánchez, Eduardo, et al.
Pubblicazione: (2023)
Translate, then Detect: Leveraging Machine Translation for Cross-Lingual Toxicity Classification
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
di: Bell, Samuel J., et al.
Pubblicazione: (2025)
Unveiling the Role of Pretraining in Direct Speech Translation
di: Alastruey, Belen, et al.
Pubblicazione: (2024)
di: Alastruey, Belen, et al.
Pubblicazione: (2024)
SpeechAlign: a Framework for Speech Translation Alignment Evaluation
di: Alastruey, Belen, et al.
Pubblicazione: (2023)
di: Alastruey, Belen, et al.
Pubblicazione: (2023)
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Improving Language Plasticity via Pretraining with Active Forgetting
di: Chen, Yihong, et al.
Pubblicazione: (2023)
di: Chen, Yihong, et al.
Pubblicazione: (2023)
2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Interference Matrix: Quantifying Cross-Lingual Interference in Transformer Encoders
di: Alastruey, Belen, et al.
Pubblicazione: (2025)
di: Alastruey, Belen, et al.
Pubblicazione: (2025)
Quantifying Generative Media Bias with a Corpus of Real-world and Generated News Articles
di: Trhlik, Filip, et al.
Pubblicazione: (2024)
di: Trhlik, Filip, et al.
Pubblicazione: (2024)
Towards Massive Multilingual Holistic Bias
di: Tan, Xiaoqing Ellen, et al.
Pubblicazione: (2024)
di: Tan, Xiaoqing Ellen, et al.
Pubblicazione: (2024)
WiCkeD: A Simple Method to Make Multiple Choice Benchmarks More Challenging
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
Cross-lingual Self-Consistency for Multilingual Reasoning with Language Models
di: Elhady, Ahmed, et al.
Pubblicazione: (2026)
di: Elhady, Ahmed, et al.
Pubblicazione: (2026)
On the Similarity of Circuits across Languages: a Case Study on the Subject-verb Agreement Task
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
di: The Omnilingual MT Team, et al.
Pubblicazione: (2025)
di: The Omnilingual MT Team, et al.
Pubblicazione: (2025)
Emergent Abilities of Large Language Models under Continued Pretraining for Language Adaptation
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
di: Elhady, Ahmed, et al.
Pubblicazione: (2025)
On the Role of Speech Data in Reducing Toxicity Detection Bias
di: Bell, Samuel J., et al.
Pubblicazione: (2024)
di: Bell, Samuel J., et al.
Pubblicazione: (2024)
Improving Language and Modality Transfer in Translation by Character-level Modeling
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2025)
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2025)
Large Concept Models: Language Modeling in a Sentence Representation Space
di: LCM team, et al.
Pubblicazione: (2024)
di: LCM team, et al.
Pubblicazione: (2024)
Lost in Inference: Rediscovering the Role of Natural Language Inference for Large Language Models
di: Madaan, Lovish, et al.
Pubblicazione: (2024)
di: Madaan, Lovish, et al.
Pubblicazione: (2024)
Investigating the interaction of linguistic and mathematical reasoning in language models using multilingual number puzzles
di: Bhattacharya, Antara Raaghavi, et al.
Pubblicazione: (2025)
di: Bhattacharya, Antara Raaghavi, et al.
Pubblicazione: (2025)
Using Natural Language Explanations to Improve Robustness of In-context Learning
di: He, Xuanli, et al.
Pubblicazione: (2023)
di: He, Xuanli, et al.
Pubblicazione: (2023)
A Primer on the Inner Workings of Transformer-based Language Models
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
Strings from the Library of Babel: Random Sampling as a Strong Baseline for Prompt Optimisation
di: Lu, Yao, et al.
Pubblicazione: (2023)
di: Lu, Yao, et al.
Pubblicazione: (2023)
MuTox: Universal MUltilingual Audio-based TOXicity Dataset and Zero-shot Detector
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Drawing Conclusions from Draws: Rethinking Preference Semantics in Arena-Style LLM Evaluation
di: Tang, Raphael, et al.
Pubblicazione: (2025)
di: Tang, Raphael, et al.
Pubblicazione: (2025)
Jet Expansions of Residual Computation
di: Chen, Yihong, et al.
Pubblicazione: (2024)
di: Chen, Yihong, et al.
Pubblicazione: (2024)
LCFO: Long Context and Long Form Output Dataset and Benchmarking
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)
Towards Red Teaming in Multimodal and Multilingual Translation
di: Ropers, Christophe, et al.
Pubblicazione: (2024)
di: Ropers, Christophe, et al.
Pubblicazione: (2024)
Multilingual Pretraining Using a Large Corpus Machine-Translated from a Single Source Language
di: Wang, Jiayi, et al.
Pubblicazione: (2024)
di: Wang, Jiayi, et al.
Pubblicazione: (2024)
The Role of Mixed-Language Documents for Multilingual Large Language Model Pretraining
di: Shao, Jiandong, et al.
Pubblicazione: (2026)
di: Shao, Jiandong, et al.
Pubblicazione: (2026)
Pushing the Limits of Zero-shot End-to-End Speech Translation
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2024)
di: Tsiamas, Ioannis, et al.
Pubblicazione: (2024)
A conclusive remark on linguistic theorizing and language modeling
di: Chesi, Cristiano
Pubblicazione: (2025)
di: Chesi, Cristiano
Pubblicazione: (2025)
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
di: Benkirane, Kenza, et al.
Pubblicazione: (2024)
Large language models and linguistic intentionality
di: Grindrod, Jumbly
Pubblicazione: (2024)
di: Grindrod, Jumbly
Pubblicazione: (2024)
Warmup Generations: A Task-Agnostic Approach for Guiding Sequence-to-Sequence Learning with Unsupervised Initial State Generation
di: Li, Senyu, et al.
Pubblicazione: (2025)
di: Li, Senyu, et al.
Pubblicazione: (2025)
Multilingual Language Model Pretraining using Machine-translated Data
di: Wang, Jiayi, et al.
Pubblicazione: (2025)
di: Wang, Jiayi, et al.
Pubblicazione: (2025)
Prompting language influences diagnostic reasoning and accuracy of large language models
di: Bazoge, Adrien, et al.
Pubblicazione: (2026)
di: Bazoge, Adrien, et al.
Pubblicazione: (2026)
Neural networks for abstraction and reasoning: Towards broad generalization in machines
di: Bober-Irizar, Mikel, et al.
Pubblicazione: (2024)
di: Bober-Irizar, Mikel, et al.
Pubblicazione: (2024)
Omnilingual SONAR: Cross-Lingual and Cross-Modal Sentence Embeddings Bridging Massively Multilingual Text and Speech
di: Omnilingual SONAR Team, et al.
Pubblicazione: (2026)
di: Omnilingual SONAR Team, et al.
Pubblicazione: (2026)
Do language models accommodate their users? A study of linguistic convergence
di: Blevins, Terra, et al.
Pubblicazione: (2025)
di: Blevins, Terra, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Gender-specific Machine Translation with Large Language Models
di: Sánchez, Eduardo, et al.
Pubblicazione: (2023) -
Translate, then Detect: Leveraging Machine Translation for Cross-Lingual Toxicity Classification
di: Bell, Samuel J., et al.
Pubblicazione: (2025) -
Unveiling the Role of Pretraining in Direct Speech Translation
di: Alastruey, Belen, et al.
Pubblicazione: (2024) -
SpeechAlign: a Framework for Speech Translation Alignment Evaluation
di: Alastruey, Belen, et al.
Pubblicazione: (2023) -
Y-NQ: English-Yorùbá Evaluation dataset for Open-Book Reading Comprehension and Text Generation
di: Costa-jussà, Marta R., et al.
Pubblicazione: (2024)