Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Karim, Aabid, Karim, Abdul, Lohana, Bhoomika, Keon, Matt, Singh, Jaswinder, Sattar, Abdul |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Galton's Law of Mediocrity: Why Large Language Models Regress to the Mean and Fail at Creativity in Advertising
par: Keon, Matt, et autres
Publié: (2025)
par: Keon, Matt, et autres
Publié: (2025)
When Intelligence Fails: An Empirical Study on Why LLMs Struggle with Password Cracking
par: Rehman, Mohammad Abdul, et autres
Publié: (2025)
par: Rehman, Mohammad Abdul, et autres
Publié: (2025)
FTT-GRU: A Hybrid Fast Temporal Transformer with GRU for Remaining Useful Life Prediction
par: Chirukiri, Varun Teja, et autres
Publié: (2025)
par: Chirukiri, Varun Teja, et autres
Publié: (2025)
FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization
par: Karim, Mohammed Asad, et autres
Publié: (2026)
par: Karim, Mohammed Asad, et autres
Publié: (2026)
Learning Bug Context for PyTorch-to-JAX Translation with LLMs
par: Phan, Hung, et autres
Publié: (2025)
par: Phan, Hung, et autres
Publié: (2025)
Dynamic Vocabulary Pruning in Early-Exit LLMs
par: Vincenti, Jort, et autres
Publié: (2024)
par: Vincenti, Jort, et autres
Publié: (2024)
TaoBench: Do Automated Theorem Prover LLMs Generalize Beyond MathLib?
par: Taylor, Alexander K, et autres
Publié: (2026)
par: Taylor, Alexander K, et autres
Publié: (2026)
The Depth Delusion: Why Transformers Should Be Wider, Not Deeper
par: Fahim, Md Muhtasim Munif, et autres
Publié: (2026)
par: Fahim, Md Muhtasim Munif, et autres
Publié: (2026)
Pairwise Difference Learning for Classification
par: Belaid, Mohamed Karim, et autres
Publié: (2024)
par: Belaid, Mohamed Karim, et autres
Publié: (2024)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
par: Malek, Alan, et autres
Publié: (2025)
par: Malek, Alan, et autres
Publié: (2025)
The Struggles of LLMs in Cross-lingual Code Clone Detection
par: Moumoula, Micheline Bénédicte, et autres
Publié: (2024)
par: Moumoula, Micheline Bénédicte, et autres
Publié: (2024)
LEMMA: Learning from Errors for MatheMatical Advancement in LLMs
par: Pan, Zhuoshi, et autres
Publié: (2025)
par: Pan, Zhuoshi, et autres
Publié: (2025)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
par: Zhang, Yuanhe, et autres
Publié: (2025)
par: Zhang, Yuanhe, et autres
Publié: (2025)
Lost and Found in Translation: Variational Diagnostics for Neural Codebook Channels
par: Hayashi, Yusuke
Publié: (2026)
par: Hayashi, Yusuke
Publié: (2026)
The Struggle Between Continuation and Refusal: A Mechanistic Analysis of the Continuation-Triggered Jailbreak in LLMs
par: Deng, Yonghong, et autres
Publié: (2026)
par: Deng, Yonghong, et autres
Publié: (2026)
Multi-Armed Bandits-Based Optimization of Decision Trees
par: Shanto, Hasibul Karim, et autres
Publié: (2025)
par: Shanto, Hasibul Karim, et autres
Publié: (2025)
Dreaming of Many Worlds: Learning Contextual World Models Aids Zero-Shot Generalization
par: Prasanna, Sai, et autres
Publié: (2024)
par: Prasanna, Sai, et autres
Publié: (2024)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
par: Jiang, Dongwei, et autres
Publié: (2024)
par: Jiang, Dongwei, et autres
Publié: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
par: Vakili, Sattar, et autres
Publié: (2024)
par: Vakili, Sattar, et autres
Publié: (2024)
Kernelized Reinforcement Learning with Order Optimal Regret Bounds
par: Vakili, Sattar, et autres
Publié: (2023)
par: Vakili, Sattar, et autres
Publié: (2023)
AMLNet: A Knowledge-Based Multi-Agent Framework to Generate and Detect Realistic Money Laundering Transactions
par: Huda, Sabin, et autres
Publié: (2025)
par: Huda, Sabin, et autres
Publié: (2025)
Empowering Bengali Education with AI: Solving Bengali Math Word Problems through Transformer Models
par: Era, Jalisha Jashim, et autres
Publié: (2025)
par: Era, Jalisha Jashim, et autres
Publié: (2025)
Safe Learning Under Irreversible Dynamics via Asking for Help
par: Plaut, Benjamin, et autres
Publié: (2025)
par: Plaut, Benjamin, et autres
Publié: (2025)
CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming
par: Chiu, Yu Ying, et autres
Publié: (2024)
par: Chiu, Yu Ying, et autres
Publié: (2024)
Why Do Open-Source LLMs Struggle with Data Analysis? A Systematic Empirical Study
par: Zhu, Yuqi, et autres
Publié: (2025)
par: Zhu, Yuqi, et autres
Publié: (2025)
Do LLMs Encode Functional Importance of Reasoning Tokens?
par: Singh, Janvijay, et autres
Publié: (2026)
par: Singh, Janvijay, et autres
Publié: (2026)
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
par: Wahab, Abdul, et autres
Publié: (2026)
par: Wahab, Abdul, et autres
Publié: (2026)
Efficiently Deploying LLMs with Controlled Risk
par: Zellinger, Michael J., et autres
Publié: (2024)
par: Zellinger, Michael J., et autres
Publié: (2024)
BertaQA: How Much Do Language Models Know About Local Culture?
par: Etxaniz, Julen, et autres
Publié: (2024)
par: Etxaniz, Julen, et autres
Publié: (2024)
Enhancing Frame Detection with Retrieval Augmented Generation
par: Diallo, Papa Abdou Karim Karou, et autres
Publié: (2025)
par: Diallo, Papa Abdou Karim Karou, et autres
Publié: (2025)
CultureLLM: Incorporating Cultural Differences into Large Language Models
par: Li, Cheng, et autres
Publié: (2024)
par: Li, Cheng, et autres
Publié: (2024)
ABS: Enforcing Constraint Satisfaction On Generated Sequences Via Automata-Guided Beam Search
par: Collura, Vincenzo, et autres
Publié: (2025)
par: Collura, Vincenzo, et autres
Publié: (2025)
Constraint-Guided Prediction Refinement via Deterministic Diffusion Trajectories
par: Dogoulis, Pantelis, et autres
Publié: (2025)
par: Dogoulis, Pantelis, et autres
Publié: (2025)
X-REFINE: XAI-based RElevance input-Filtering and archItecture fiNe-tuning for channel Estimation
par: Gizzini, Abdul Karim, et autres
Publié: (2026)
par: Gizzini, Abdul Karim, et autres
Publié: (2026)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
par: Petrov, Ivo, et autres
Publié: (2025)
par: Petrov, Ivo, et autres
Publié: (2025)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
par: Chen, Zui, et autres
Publié: (2024)
par: Chen, Zui, et autres
Publié: (2024)
Leveraging Intermediate Representations of Time Series Foundation Models for Anomaly Detection
par: Han, Chan Sik, et autres
Publié: (2025)
par: Han, Chan Sik, et autres
Publié: (2025)
eMargin: Revisiting Contrastive Learning with Margin-Based Separation
par: Shamba, Abdul-Kazeem, et autres
Publié: (2025)
par: Shamba, Abdul-Kazeem, et autres
Publié: (2025)
Machine Learning Risk Intelligence for Green Hydrogen Investment: Insights for Duqm R3 Auction
par: Nwafor, Obumneme, et autres
Publié: (2025)
par: Nwafor, Obumneme, et autres
Publié: (2025)
Generative Modeling with Flow-Guided Density Ratio Learning
par: Heng, Alvin, et autres
Publié: (2023)
par: Heng, Alvin, et autres
Publié: (2023)
Documents similaires
-
Galton's Law of Mediocrity: Why Large Language Models Regress to the Mean and Fail at Creativity in Advertising
par: Keon, Matt, et autres
Publié: (2025) -
When Intelligence Fails: An Empirical Study on Why LLMs Struggle with Password Cracking
par: Rehman, Mohammad Abdul, et autres
Publié: (2025) -
FTT-GRU: A Hybrid Fast Temporal Transformer with GRU for Remaining Useful Life Prediction
par: Chirukiri, Varun Teja, et autres
Publié: (2025) -
FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization
par: Karim, Mohammed Asad, et autres
Publié: (2026) -
Learning Bug Context for PyTorch-to-JAX Translation with LLMs
par: Phan, Hung, et autres
Publié: (2025)