All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Dan, Mo, Guozhao, Shi, Yafei, Zhang, Cheng, Zheng, Bo, Cao, Boxi, Chen, Xuanang, Lu, Yaojie, Lin, Hongyu, He, Ben, Han, Xianpei, Sun, Le |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models
par: Yan, Xinru, et autres
Publié: (2026)
par: Yan, Xinru, et autres
Publié: (2026)
The Life Cycle of Knowledge in Big Language Models: A Survey
par: Cao, Boxi, et autres
Publié: (2023)
par: Cao, Boxi, et autres
Publié: (2023)
LiveMCPBench: Can Agents Navigate an Ocean of MCP Tools?
par: Mo, Guozhao, et autres
Publié: (2025)
par: Mo, Guozhao, et autres
Publié: (2025)
DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generation
par: Zheng, Hao, et autres
Publié: (2026)
par: Zheng, Hao, et autres
Publié: (2026)
Beyond Correctness: Benchmarking Multi-dimensional Code Generation for Large Language Models
par: Zheng, Jiasheng, et autres
Publié: (2024)
par: Zheng, Jiasheng, et autres
Publié: (2024)
CoCoNUTS: Concentrating on Content while Neglecting Uninformative Textual Styles for AI-Generated Peer Review Detection
par: Chen, Yihan, et autres
Publié: (2025)
par: Chen, Yihan, et autres
Publié: (2025)
LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents
par: Peng, Xiaoxuan, et autres
Publié: (2026)
par: Peng, Xiaoxuan, et autres
Publié: (2026)
ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch
par: Chen, Jiawei, et autres
Publié: (2025)
par: Chen, Jiawei, et autres
Publié: (2025)
Multi-Facet Counterfactual Learning for Content Quality Evaluation
par: Zheng, Jiasheng, et autres
Publié: (2024)
par: Zheng, Jiasheng, et autres
Publié: (2024)
Combinatorial Synthesis: Scaling Code RLVR via Atomic Decomposition and Recombination
par: Zheng, Jiasheng, et autres
Publié: (2026)
par: Zheng, Jiasheng, et autres
Publié: (2026)
Executing Natural Language-Described Algorithms with Large Language Models: An Investigation
par: Zheng, Xin, et autres
Publié: (2024)
par: Zheng, Xin, et autres
Publié: (2024)
Beyond Isolated Dots: Benchmarking Structured Table Construction as Deep Knowledge Extraction
par: Zhong, Tianyun, et autres
Publié: (2025)
par: Zhong, Tianyun, et autres
Publié: (2025)
CRUXEval-X: A Benchmark for Multilingual Code Reasoning, Understanding and Execution
par: Xu, Ruiyang, et autres
Publié: (2024)
par: Xu, Ruiyang, et autres
Publié: (2024)
When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models
par: Mao, Yingzhi, et autres
Publié: (2025)
par: Mao, Yingzhi, et autres
Publié: (2025)
Rule or Story, Which is a Better Commonsense Expression for Talking with Large Language Models?
par: Bian, Ning, et autres
Publié: (2024)
par: Bian, Ning, et autres
Publié: (2024)
Across Programming Language Silos: A Study on Cross-Lingual Retrieval-augmented Code Generation
par: Zhu, Qiming, et autres
Publié: (2025)
par: Zhu, Qiming, et autres
Publié: (2025)
ScaleBox: Enabling High-Fidelity and Scalable Code Verification for Large Language Models
par: Zheng, Jiasheng, et autres
Publié: (2026)
par: Zheng, Jiasheng, et autres
Publié: (2026)
READoc: A Unified Benchmark for Realistic Document Structured Extraction
par: Li, Zichao, et autres
Publié: (2024)
par: Li, Zichao, et autres
Publié: (2024)
DeepRAG: Thinking to Retrieve Step by Step for Large Language Models
par: Guan, Xinyan, et autres
Publié: (2025)
par: Guan, Xinyan, et autres
Publié: (2025)
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization
par: Li, Zhuoqun, et autres
Publié: (2024)
par: Li, Zhuoqun, et autres
Publié: (2024)
SAISA: Towards Multimodal Large Language Models with Both Training and Inference Efficiency
par: Yuan, Qianhao, et autres
Publié: (2025)
par: Yuan, Qianhao, et autres
Publié: (2025)
AI-Salesman: Towards Reliable Large Language Model Driven Telemarketing
par: Zhang, Qingyu, et autres
Publié: (2025)
par: Zhang, Qingyu, et autres
Publié: (2025)
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation
par: Pan, Ruotong, et autres
Publié: (2024)
par: Pan, Ruotong, et autres
Publié: (2024)
StructEval: Deepen and Broaden Large Language Model Assessment via Structured Evaluation
par: Cao, Boxi, et autres
Publié: (2024)
par: Cao, Boxi, et autres
Publié: (2024)
Spiral of Silence: How is Large Language Model Killing Information Retrieval? -- A Case Study on Open Domain Question Answering
par: Chen, Xiaoyang, et autres
Publié: (2024)
par: Chen, Xiaoyang, et autres
Publié: (2024)
Meta-Cognitive Analysis: Evaluating Declarative and Procedural Knowledge in Datasets and Large Language Models
par: Li, Zhuoqun, et autres
Publié: (2024)
par: Li, Zhuoqun, et autres
Publié: (2024)
The Rise and Down of Babel Tower: Investigating the Evolution Process of Multilingual Code Large Language Model
par: Chen, Jiawei, et autres
Publié: (2024)
par: Chen, Jiawei, et autres
Publié: (2024)
PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
par: Li, Zhuoqun, et autres
Publié: (2025)
par: Li, Zhuoqun, et autres
Publié: (2025)
Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
par: Xu, Ruoxi, et autres
Publié: (2025)
par: Xu, Ruoxi, et autres
Publié: (2025)
ChatGPT is a Knowledgeable but Inexperienced Solver: An Investigation of Commonsense Problem in Large Language Models
par: Bian, Ning, et autres
Publié: (2023)
par: Bian, Ning, et autres
Publié: (2023)
P^2O: Joint Policy and Prompt Optimization
par: Lu, Xinyu, et autres
Publié: (2026)
par: Lu, Xinyu, et autres
Publié: (2026)
Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards
par: Ma, Zhengzhao, et autres
Publié: (2026)
par: Ma, Zhengzhao, et autres
Publié: (2026)
Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards
par: Ren, Mengjie, et autres
Publié: (2026)
par: Ren, Mengjie, et autres
Publié: (2026)
Towards Universal Dense Blocking for Entity Resolution
par: Wang, Tianshu, et autres
Publié: (2024)
par: Wang, Tianshu, et autres
Publié: (2024)
Critic-CoT: Boosting the reasoning abilities of large language model via Chain-of-thoughts Critic
par: Zheng, Xin, et autres
Publié: (2024)
par: Zheng, Xin, et autres
Publié: (2024)
Match, Compare, or Select? An Investigation of Large Language Models for Entity Matching
par: Wang, Tianshu, et autres
Publié: (2024)
par: Wang, Tianshu, et autres
Publié: (2024)
Seg2Act: Global Context-aware Action Generation for Document Logical Structuring
par: Li, Zichao, et autres
Publié: (2024)
par: Li, Zichao, et autres
Publié: (2024)
DeepSolution: Boosting Complex Engineering Solution Design via Tree-based Exploration and Bi-point Thinking
par: Li, Zhuoqun, et autres
Publié: (2025)
par: Li, Zhuoqun, et autres
Publié: (2025)
LiveFMBench: Unveiling the Power and Limits of Agentic Workflows in Specification Generation
par: Xu, Dong, et autres
Publié: (2026)
par: Xu, Dong, et autres
Publié: (2026)
Coupled Variational Reinforcement Learning for Language Model General Reasoning
par: Wen, Xueru, et autres
Publié: (2025)
par: Wen, Xueru, et autres
Publié: (2025)
Documents similaires
-
Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models
par: Yan, Xinru, et autres
Publié: (2026) -
The Life Cycle of Knowledge in Big Language Models: A Survey
par: Cao, Boxi, et autres
Publié: (2023) -
LiveMCPBench: Can Agents Navigate an Ocean of MCP Tools?
par: Mo, Guozhao, et autres
Publié: (2025) -
DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generation
par: Zheng, Hao, et autres
Publié: (2026) -
Beyond Correctness: Benchmarking Multi-dimensional Code Generation for Large Language Models
par: Zheng, Jiasheng, et autres
Publié: (2024)