Evaluating Multilingual Long-Context Models for Retrieval and Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Agrawal, Ameeta, Dang, Andy, Nezhad, Sina Bagheri, Pokharel, Rhitabrat, Scheinberg, Russell |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Data Quantity: Key Factors Driving Performance in Multilingual Language Models
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
MTQ-Eval: Multilingual Text Quality Evaluation for Language Models
by: Pokharel, Rhitabrat, et al.
Published: (2025)
by: Pokharel, Rhitabrat, et al.
Published: (2025)
The Impact of Model Scaling on Seen and Unseen Language Performance
by: Pokharel, Rhitabrat, et al.
Published: (2025)
by: Pokharel, Rhitabrat, et al.
Published: (2025)
Enhancing Large Language Models with Neurosymbolic Reasoning for Multilingual Tasks
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
Exploring the Maze of Multilingual Modeling
by: Nezhad, Sina Bagheri, et al.
Published: (2023)
by: Nezhad, Sina Bagheri, et al.
Published: (2023)
What Drives Performance in Multilingual Language Models?
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
Cross-Lingual Activation Steering for Multilingual Language Models
by: Pokharel, Rhitabrat, et al.
Published: (2026)
by: Pokharel, Rhitabrat, et al.
Published: (2026)
CAPO: Confidence Aware Preference Optimization Learning for Multilingual Preferences
by: Pokharel, Rhitabrat, et al.
Published: (2025)
by: Pokharel, Rhitabrat, et al.
Published: (2025)
SymCode: A Neurosymbolic Approach to Mathematical Reasoning via Verifiable Code Generation
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
by: Nezhad, Sina Bagheri, et al.
Published: (2025)
From Policy to Logic for Efficient and Interpretable Coverage Assessment
by: Pokharel, Rhitabrat, et al.
Published: (2026)
by: Pokharel, Rhitabrat, et al.
Published: (2026)
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
by: Nezhad, Sina Bagheri, et al.
Published: (2024)
Multilingual Relative Clause Attachment Ambiguity Resolution in Large Language Models
by: Lee, So Young, et al.
Published: (2025)
by: Lee, So Young, et al.
Published: (2025)
Who Relies More on World Knowledge and Bias for Syntactic Ambiguity Resolution: Humans or LLMs?
by: Lee, So Young, et al.
Published: (2025)
by: Lee, So Young, et al.
Published: (2025)
Explain-then-Process: Using Grammar Prompting to Enhance Grammatical Acceptability Judgments
by: Scheinberg, Russell, et al.
Published: (2025)
by: Scheinberg, Russell, et al.
Published: (2025)
Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs
by: Shore, Amber, et al.
Published: (2025)
by: Shore, Amber, et al.
Published: (2025)
Toward Culturally Grounded Natural Language Processing
by: Nezhad, Sina Bagheri
Published: (2026)
by: Nezhad, Sina Bagheri
Published: (2026)
No-Worse Context-Aware Decoding: Preventing Neutral Regression in Context-Conditioned Generation
by: Tao, Yufei, et al.
Published: (2026)
by: Tao, Yufei, et al.
Published: (2026)
Making a Long Story Short in Conversation Modeling
by: Tao, Yufei, et al.
Published: (2024)
by: Tao, Yufei, et al.
Published: (2024)
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
by: Olabisi, Olubusayo, et al.
Published: (2024)
by: Olabisi, Olubusayo, et al.
Published: (2024)
Reliable Classroom AI via Neuro-Symbolic Multimodal Reasoning
by: Nezhad, Sina Bagheri
Published: (2026)
by: Nezhad, Sina Bagheri
Published: (2026)
When Context Leads but Parametric Memory Follows in Large Language Models
by: Tao, Yufei, et al.
Published: (2024)
by: Tao, Yufei, et al.
Published: (2024)
MFTCXplain: A Multilingual Benchmark Dataset for Evaluating the Moral Reasoning of LLMs through Multi-hop Hate Speech Explanation
by: Trager, Jackson, et al.
Published: (2025)
by: Trager, Jackson, et al.
Published: (2025)
ThreadSumm: Summarization of Nested Discourse Threads Using Tree of Thoughts
by: Olabisi, Olubusayo, et al.
Published: (2026)
by: Olabisi, Olubusayo, et al.
Published: (2026)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
by: Tao, Yufei, et al.
Published: (2025)
by: Tao, Yufei, et al.
Published: (2025)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
by: Wang, Yiming, et al.
Published: (2025)
by: Wang, Yiming, et al.
Published: (2025)
ChatGPT Role-play Dataset: Analysis of User Motives and Model Naturalness
by: Tao, Yufei, et al.
Published: (2024)
by: Tao, Yufei, et al.
Published: (2024)
Are Language Models Borrowing-Blind? A Multilingual Evaluation of Loanword Identification across 10 Languages
by: Silva, Mérilin Sousa, et al.
Published: (2025)
by: Silva, Mérilin Sousa, et al.
Published: (2025)
mGTE: Generalized Long-Context Text Representation and Reranking Models for Multilingual Text Retrieval
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
Knowledge-Aware Conversation Derailment Forecasting Using Graph Convolutional Networks
by: Altarawneh, Enas, et al.
Published: (2024)
by: Altarawneh, Enas, et al.
Published: (2024)
Predicting Evoked Emotions in Conversations
by: Altarawneh, Enas, et al.
Published: (2023)
by: Altarawneh, Enas, et al.
Published: (2023)
Narrating Causal Graphs with Large Language Models
by: Phatak, Atharva, et al.
Published: (2024)
by: Phatak, Atharva, et al.
Published: (2024)
Mitigation of Gender and Ethnicity Bias in AI-Generated Stories through Model Explanations
by: Dimgba, Martha O., et al.
Published: (2025)
by: Dimgba, Martha O., et al.
Published: (2025)
Grounding Long-Context Reasoning with Contextual Normalization for Retrieval-Augmented Generation
by: Chen, Jiamin, et al.
Published: (2025)
by: Chen, Jiamin, et al.
Published: (2025)
Facilitating Long Context Understanding via Supervised Chain-of-Thought Reasoning
by: Lin, Jingyang, et al.
Published: (2025)
by: Lin, Jingyang, et al.
Published: (2025)
Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs
by: Myers, Skatje, et al.
Published: (2025)
by: Myers, Skatje, et al.
Published: (2025)
Query-Focused Retrieval Heads Improve Long-Context Reasoning and Re-ranking
by: Zhang, Wuwei, et al.
Published: (2025)
by: Zhang, Wuwei, et al.
Published: (2025)
Beyond Topical Similarity: Contrastive Evidence Retrieval with Interpretable Attention Alignment in RAG
by: Vargas, Francielle, et al.
Published: (2026)
by: Vargas, Francielle, et al.
Published: (2026)
DetectiveQA: Evaluating Long-Context Reasoning on Detective Novels
by: Xu, Zhe, et al.
Published: (2024)
by: Xu, Zhe, et al.
Published: (2024)
Improving Multilingual Retrieval-Augmented Language Models through Dialectic Reasoning Argumentations
by: Ranaldi, Leonardo, et al.
Published: (2025)
by: Ranaldi, Leonardo, et al.
Published: (2025)
Towards Inducing Long-Context Abilities in Multilingual Neural Machine Translation Models
by: Gumma, Varun, et al.
Published: (2024)
by: Gumma, Varun, et al.
Published: (2024)
Similar Items
-
Beyond Data Quantity: Key Factors Driving Performance in Multilingual Language Models
by: Nezhad, Sina Bagheri, et al.
Published: (2024) -
MTQ-Eval: Multilingual Text Quality Evaluation for Language Models
by: Pokharel, Rhitabrat, et al.
Published: (2025) -
The Impact of Model Scaling on Seen and Unseen Language Performance
by: Pokharel, Rhitabrat, et al.
Published: (2025) -
Enhancing Large Language Models with Neurosymbolic Reasoning for Multilingual Tasks
by: Nezhad, Sina Bagheri, et al.
Published: (2025) -
Exploring the Maze of Multilingual Modeling
by: Nezhad, Sina Bagheri, et al.
Published: (2023)