Multilingual LLMs Struggle to Link Orthography and Semantics in Bilingual Word Processing
Fuente:
arXiv
Saved in:
| Main Authors: | Tanwar, Eshaan, Oke, Gayatri, Chakraborty, Tanmoy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding the Effects of Domain Finetuning on LLMs
by: Tanwar, Eshaan, et al.
Published: (2025)
by: Tanwar, Eshaan, et al.
Published: (2025)
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks
by: Chatterjee, Anwoy, et al.
Published: (2024)
by: Chatterjee, Anwoy, et al.
Published: (2024)
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
by: Tanwar, Eshaan, et al.
Published: (2025)
by: Tanwar, Eshaan, et al.
Published: (2025)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
by: Bajpai, Ashutosh, et al.
Published: (2024)
by: Bajpai, Ashutosh, et al.
Published: (2024)
Multilingual Test-Time Scaling via Initial Thought Transfer
by: Bajpai, Prasoon, et al.
Published: (2025)
by: Bajpai, Prasoon, et al.
Published: (2025)
Decoding Memes: Benchmarking Narrative Role Classification across Multilingual and Multimodal Models
by: Sharma, Shivam, et al.
Published: (2025)
by: Sharma, Shivam, et al.
Published: (2025)
Crowdsourcing Piedmontese to Test LLMs on Non-Standard Orthography
by: Vico, Gianluca, et al.
Published: (2026)
by: Vico, Gianluca, et al.
Published: (2026)
Multilingual Needle in a Haystack: Investigating Long-Context Behavior of Multilingual Large Language Models
by: Hengle, Amey, et al.
Published: (2024)
by: Hengle, Amey, et al.
Published: (2024)
Position: Enough of Scaling LLMs! Lets Focus on Downscaling
by: Goel, Yash, et al.
Published: (2025)
by: Goel, Yash, et al.
Published: (2025)
Value-Guided KV Compression for LLMs via Approximated CUR Decomposition
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
by: Bajpai, Ashutosh, et al.
Published: (2025)
by: Bajpai, Ashutosh, et al.
Published: (2025)
Normalized Orthography for Tunisian Arabic
by: Turki, Houcemeddine, et al.
Published: (2024)
by: Turki, Houcemeddine, et al.
Published: (2024)
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
From Images to Words: Efficient Cross-Modal Knowledge Distillation to Language Models from Black-box Teachers
by: Sengupta, Ayan, et al.
Published: (2026)
by: Sengupta, Ayan, et al.
Published: (2026)
Better To Ask in English? Evaluating Factual Accuracy of Multilingual LLMs in English and Low-Resource Languages
by: Rohera, Pritika, et al.
Published: (2025)
by: Rohera, Pritika, et al.
Published: (2025)
Multilingual Language Models Encode Script Over Linguistic Structure
by: Verma, Aastha A K, et al.
Published: (2026)
by: Verma, Aastha A K, et al.
Published: (2026)
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators
by: Bajpai, Prasoon, et al.
Published: (2024)
by: Bajpai, Prasoon, et al.
Published: (2024)
LAMA-UT: Language Agnostic Multilingual ASR through Orthography Unification and Language-Specific Transliteration
by: Lee, Sangmin, et al.
Published: (2024)
by: Lee, Sangmin, et al.
Published: (2024)
Harmonizing Code-mixed Conversations: Personality-assisted Code-mixed Response Generation in Dialogues
by: Kumar, Shivani, et al.
Published: (2024)
by: Kumar, Shivani, et al.
Published: (2024)
DATASHI: A Parallel English-Tashlhiyt Corpus for Orthography Normalization and Low-Resource Language Processing
by: Monir, Nasser-Eddine, et al.
Published: (2026)
by: Monir, Nasser-Eddine, et al.
Published: (2026)
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacks
by: Hengle, Amey, et al.
Published: (2025)
by: Hengle, Amey, et al.
Published: (2025)
Hate Personified: Investigating the role of LLMs in content moderation
by: Masud, Sarah, et al.
Published: (2024)
by: Masud, Sarah, et al.
Published: (2024)
Adaptative Bilingual Aligning Using Multilingual Sentence Embedding
by: Kraif, Olivier
Published: (2024)
by: Kraif, Olivier
Published: (2024)
Using LLMs for Multilingual Clinical Entity Linking to ICD-10
by: Vassileva, Sylvia, et al.
Published: (2025)
by: Vassileva, Sylvia, et al.
Published: (2025)
Exploring Representational Disparities Between Multilingual and Bilingual Translation Models
by: Verma, Neha, et al.
Published: (2023)
by: Verma, Neha, et al.
Published: (2023)
Bilingual Word Level Language Identification for Omotic Languages
by: Yigezu, Mesay Gemeda, et al.
Published: (2025)
by: Yigezu, Mesay Gemeda, et al.
Published: (2025)
Rank, Chunk and Expand: Lineage-Oriented Reasoning for Taxonomy Expansion
by: Mishra, Sahil, et al.
Published: (2025)
by: Mishra, Sahil, et al.
Published: (2025)
The Art of Scaling Test-Time Compute for Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025)
by: Agarwal, Aradhye, et al.
Published: (2025)
First Finish Search: Efficient Test-Time Scaling in Large Language Models
by: Agarwal, Aradhye, et al.
Published: (2025)
by: Agarwal, Aradhye, et al.
Published: (2025)
Innocence in the Crossfire: Roles of Skip Connections in Jailbreaking Visual Language Models
by: Nandi, Palash, et al.
Published: (2025)
by: Nandi, Palash, et al.
Published: (2025)
Compression Laws for Large Language Models
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Counterspeech the ultimate shield! Multi-Conditioned Counterspeech Generation through Attributed Prefix Learning
by: Kumar, Aswini, et al.
Published: (2025)
by: Kumar, Aswini, et al.
Published: (2025)
You Only Prune Once: Designing Calibration-Free Model Compression With Policy Learning
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Markovian ODE-guided scoring can assess the quality of offline reasoning traces in language models
by: Nandi, Arghodeep, et al.
Published: (2026)
by: Nandi, Arghodeep, et al.
Published: (2026)
Information Anxiety in Large Language Models
by: Bajpai, Prasoon, et al.
Published: (2024)
by: Bajpai, Prasoon, et al.
Published: (2024)
Exposing Long-Tail Safety Failures in Large Language Models through Efficient Diverse Response Sampling
by: Hajra, Suvadeep, et al.
Published: (2026)
by: Hajra, Suvadeep, et al.
Published: (2026)
SAFE-MEME: Structured Reasoning Framework for Robust Hate Speech Detection in Memes
by: Nandi, Palash, et al.
Published: (2024)
by: Nandi, Palash, et al.
Published: (2024)
Focal Inferential Infusion Coupled with Tractable Density Discrimination for Implicit Hate Detection
by: Masud, Sarah, et al.
Published: (2023)
by: Masud, Sarah, et al.
Published: (2023)
LombardoGraphia: Automatic Classification of Lombard Orthography Variants
by: Signoroni, Edoardo, et al.
Published: (2026)
by: Signoroni, Edoardo, et al.
Published: (2026)
A computational system to handle the orthographic layer of tajwid in contemporary Quranic Orthography
by: Martínez, Alicia González
Published: (2025)
by: Martínez, Alicia González
Published: (2025)
Similar Items
-
Understanding the Effects of Domain Finetuning on LLMs
by: Tanwar, Eshaan, et al.
Published: (2025) -
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks
by: Chatterjee, Anwoy, et al.
Published: (2024) -
Do You Know About My Nation? Investigating Multilingual Language Models' Cultural Literacy Through Factual Knowledge
by: Tanwar, Eshaan, et al.
Published: (2025) -
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
by: Bajpai, Ashutosh, et al.
Published: (2024) -
Multilingual Test-Time Scaling via Initial Thought Transfer
by: Bajpai, Prasoon, et al.
Published: (2025)