Learning to Route Languages for Multilingual Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Geyang, Wakaki, Hiromi, Mitsufuji, Yuki, Ritter, Alan, Xu, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CARE: Multilingual Human Preference Learning for Cultural Awareness
by: Guo, Geyang, et al.
Published: (2025)
by: Guo, Geyang, et al.
Published: (2025)
DiffuCOMET: Contextual Commonsense Knowledge Diffusion
by: Gao, Silin, et al.
Published: (2024)
by: Gao, Silin, et al.
Published: (2024)
DeepResonance: Enhancing Multimodal Music Understanding via Music-centric Multi-way Instruction Tuning
by: Mao, Zhuoyuan, et al.
Published: (2025)
by: Mao, Zhuoyuan, et al.
Published: (2025)
MCA: Modality Composition Awareness for Robust Composed Multimodal Retrieval
by: Wu, Qiyu, et al.
Published: (2025)
by: Wu, Qiyu, et al.
Published: (2025)
Cross-Modal Learning for Music-to-Music-Video Description Generation
by: Mao, Zhuoyuan, et al.
Published: (2025)
by: Mao, Zhuoyuan, et al.
Published: (2025)
ComperDial: Commonsense Persona-grounded Dialogue Dataset and Benchmark
by: Wakaki, Hiromi, et al.
Published: (2024)
by: Wakaki, Hiromi, et al.
Published: (2024)
TED: Turn Emphasis with Dialogue Feature Attention for Emotion Recognition in Conversation
by: Ono, Junya, et al.
Published: (2025)
by: Ono, Junya, et al.
Published: (2025)
OpenMU: Your Swiss Army Knife for Music Understanding
by: Zhao, Mengjie, et al.
Published: (2024)
by: Zhao, Mengjie, et al.
Published: (2024)
VinaBench: Benchmark for Faithful and Consistent Visual Narratives
by: Gao, Silin, et al.
Published: (2025)
by: Gao, Silin, et al.
Published: (2025)
Meta-Tuning LLMs to Leverage Lexical Knowledge for Generalizable Language Style Understanding
by: Guo, Ruohao, et al.
Published: (2023)
by: Guo, Ruohao, et al.
Published: (2023)
NEO-BENCH: Evaluating Robustness of Large Language Models with Neologisms
by: Zheng, Jonathan, et al.
Published: (2024)
by: Zheng, Jonathan, et al.
Published: (2024)
How to Protect Yourself from 5G Radiation? Investigating LLM Responses to Implicit Misinformation
by: Guo, Ruohao, et al.
Published: (2025)
by: Guo, Ruohao, et al.
Published: (2025)
Investigating and Alleviating Harm Amplification in LLM Interactions
by: Guo, Ruohao, et al.
Published: (2026)
by: Guo, Ruohao, et al.
Published: (2026)
Using Natural Language Inference to Improve Persona Extraction from Dialogue in a New Domain
by: DeLucia, Alexandra, et al.
Published: (2024)
by: DeLucia, Alexandra, et al.
Published: (2024)
Preference Optimization for Reasoning with Pseudo Feedback
by: Jiao, Fangkai, et al.
Published: (2024)
by: Jiao, Fangkai, et al.
Published: (2024)
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning
by: Xie, Zhouhang, et al.
Published: (2024)
by: Xie, Zhouhang, et al.
Published: (2024)
Tree-based Dialogue Reinforced Policy Optimization for Red-Teaming Attacks
by: Guo, Ruohao, et al.
Published: (2025)
by: Guo, Ruohao, et al.
Published: (2025)
Demystifying MaskGIT Sampler and Beyond: Adaptive Order Selection in Masked Diffusion
by: Hayakawa, Satoshi, et al.
Published: (2025)
by: Hayakawa, Satoshi, et al.
Published: (2025)
Distillation of Discrete Diffusion through Dimensional Correlations
by: Hayakawa, Satoshi, et al.
Published: (2024)
by: Hayakawa, Satoshi, et al.
Published: (2024)
No One Fits All: From Fixed Prompting to Learned Routing in Multilingual LLMs
by: Wu, Wei-Chi, et al.
Published: (2026)
by: Wu, Wei-Chi, et al.
Published: (2026)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
by: Huang, Xu, et al.
Published: (2026)
by: Huang, Xu, et al.
Published: (2026)
Tabular Data Understanding with LLMs: A Survey of Recent Advances and Challenges
by: Wu, Xiaofeng, et al.
Published: (2025)
by: Wu, Xiaofeng, et al.
Published: (2025)
What are Foundation Models Cooking in the Post-Soviet World?
by: Lavrouk, Anton, et al.
Published: (2025)
by: Lavrouk, Anton, et al.
Published: (2025)
Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages
by: Naous, Tarek, et al.
Published: (2025)
by: Naous, Tarek, et al.
Published: (2025)
Safe and Scalable Web Agent Learning via Recreated Websites
by: Chae, Hyungjoo, et al.
Published: (2026)
by: Chae, Hyungjoo, et al.
Published: (2026)
Language Models can Self-Improve at State-Value Estimation for Better Search
by: Mendes, Ethan, et al.
Published: (2025)
by: Mendes, Ethan, et al.
Published: (2025)
Having Beer after Prayer? Measuring Cultural Bias in Large Language Models
by: Naous, Tarek, et al.
Published: (2023)
by: Naous, Tarek, et al.
Published: (2023)
Granular Privacy Control for Geolocation with Vision Language Models
by: Mendes, Ethan, et al.
Published: (2024)
by: Mendes, Ethan, et al.
Published: (2024)
Frustratingly Easy Label Projection for Cross-lingual Transfer
by: Chen, Yang, et al.
Published: (2022)
by: Chen, Yang, et al.
Published: (2022)
Probabilistic Reasoning with LLMs for k-anonymity Estimation
by: Zheng, Jonathan, et al.
Published: (2025)
by: Zheng, Jonathan, et al.
Published: (2025)
Reducing Privacy Risks in Online Self-Disclosures with Language Models
by: Dou, Yao, et al.
Published: (2023)
by: Dou, Yao, et al.
Published: (2023)
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
by: Zheng, Kening, et al.
Published: (2026)
by: Zheng, Kening, et al.
Published: (2026)
Constrained Decoding for Cross-lingual Label Projection
by: Le, Duong Minh, et al.
Published: (2024)
by: Le, Duong Minh, et al.
Published: (2024)
Theoretical Refinement of CLIP by Utilizing Linear Structure of Optimal Similarity
by: Yoshida, Naoki, et al.
Published: (2025)
by: Yoshida, Naoki, et al.
Published: (2025)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
by: Zhou, Hao, et al.
Published: (2024)
by: Zhou, Hao, et al.
Published: (2024)
MAPO: Advancing Multilingual Reasoning through Multilingual Alignment-as-Preference Optimization
by: She, Shuaijie, et al.
Published: (2024)
by: She, Shuaijie, et al.
Published: (2024)
Beyond Imitation: Leveraging Fine-grained Quality Signals for Alignment
by: Guo, Geyang, et al.
Published: (2023)
by: Guo, Geyang, et al.
Published: (2023)
Optimizing Multilingual LLMs via Federated Learning: A Study of Client Language Composition
by: Sant, Aleix, et al.
Published: (2026)
by: Sant, Aleix, et al.
Published: (2026)
Multilingual Routing in Mixture-of-Experts
by: Bandarkar, Lucas, et al.
Published: (2025)
by: Bandarkar, Lucas, et al.
Published: (2025)
Language Steering for Multilingual In-Context Learning
by: Kirtane, Neeraja, et al.
Published: (2026)
by: Kirtane, Neeraja, et al.
Published: (2026)
Similar Items
-
CARE: Multilingual Human Preference Learning for Cultural Awareness
by: Guo, Geyang, et al.
Published: (2025) -
DiffuCOMET: Contextual Commonsense Knowledge Diffusion
by: Gao, Silin, et al.
Published: (2024) -
DeepResonance: Enhancing Multimodal Music Understanding via Music-centric Multi-way Instruction Tuning
by: Mao, Zhuoyuan, et al.
Published: (2025) -
MCA: Modality Composition Awareness for Robust Composed Multimodal Retrieval
by: Wu, Qiyu, et al.
Published: (2025) -
Cross-Modal Learning for Music-to-Music-Video Description Generation
by: Mao, Zhuoyuan, et al.
Published: (2025)