MiCRo: Mixture Modeling and Context-aware Routing for Personalized Preference Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Jingyan, Yao, Jiarui, Yang, Rui, Sun, Yifan, Luo, Feng, Pan, Rui, Zhang, Tong, Zhao, Han |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Diverse Human Preference Learning through Principal Component Analysis
by: Luo, Feng, et al.
Published: (2025)
by: Luo, Feng, et al.
Published: (2025)
Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
by: Yang, Rui, et al.
Published: (2024)
by: Yang, Rui, et al.
Published: (2024)
Self-supervised Attribute-aware Dynamic Preference Ranking Alignment
by: Yang, Hongyu, et al.
Published: (2025)
by: Yang, Hongyu, et al.
Published: (2025)
Personality-aware Human-centric Multimodal Reasoning: A New Task, Dataset and Baselines
by: Zhu, Yaochen, et al.
Published: (2023)
by: Zhu, Yaochen, et al.
Published: (2023)
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Personalized Adaptation via In-Context Preference Learning
by: Lau, Allison, et al.
Published: (2024)
by: Lau, Allison, et al.
Published: (2024)
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
by: Lin, Hongzhan, et al.
Published: (2024)
by: Lin, Hongzhan, et al.
Published: (2024)
Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
by: Wang, Haoxiang, et al.
Published: (2024)
by: Wang, Haoxiang, et al.
Published: (2024)
Preference-Aware Rubric Learning for Personalized Evaluation
by: Qiu, Yilun, et al.
Published: (2026)
by: Qiu, Yilun, et al.
Published: (2026)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
by: Xu, Haolei, et al.
Published: (2026)
by: Xu, Haolei, et al.
Published: (2026)
Maximum Score Routing For Mixture-of-Experts
by: Dong, Bowen, et al.
Published: (2025)
by: Dong, Bowen, et al.
Published: (2025)
CRoCoDiL: Continuous and Robust Conditioned Diffusion for Language
by: Uziel, Roy, et al.
Published: (2026)
by: Uziel, Roy, et al.
Published: (2026)
Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
by: Wang, Haoxiang, et al.
Published: (2024)
by: Wang, Haoxiang, et al.
Published: (2024)
From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning
by: zhang, Ranxu, et al.
Published: (2026)
by: zhang, Ranxu, et al.
Published: (2026)
ACIL: Auto Chain of Thoughts for In-Context Learning
by: Chu, Rui
Published: (2026)
by: Chu, Rui
Published: (2026)
Strengthening Multimodal Large Language Model with Bootstrapped Preference Optimization
by: Pi, Renjie, et al.
Published: (2024)
by: Pi, Renjie, et al.
Published: (2024)
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation
by: Pan, Ruotong, et al.
Published: (2024)
by: Pan, Ruotong, et al.
Published: (2024)
CRoP: Context-wise Robust Static Human-Sensing Personalization
by: Kaur, Sawinder, et al.
Published: (2024)
by: Kaur, Sawinder, et al.
Published: (2024)
Route to Reason: Adaptive Routing for LLM and Reasoning Strategy Selection
by: Pan, Zhihong, et al.
Published: (2025)
by: Pan, Zhihong, et al.
Published: (2025)
Towards Generalizable Implicit In-Context Learning with Attention Routing
by: Li, Jiaqian, et al.
Published: (2025)
by: Li, Jiaqian, et al.
Published: (2025)
Extended Inductive Reasoning for Personalized Preference Inference from Behavioral Signals
by: Li, Jia-Nan, et al.
Published: (2025)
by: Li, Jia-Nan, et al.
Published: (2025)
MPMA: Preference Manipulation Attack Against Model Context Protocol
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
LATTE: Forecasting Peer Anchored Preference Trajectories for Personalized LLM Generation
by: Li, Jinze, et al.
Published: (2026)
by: Li, Jinze, et al.
Published: (2026)
Routing-Free Mixture-of-Experts
by: Liu, Yilun, et al.
Published: (2026)
by: Liu, Yilun, et al.
Published: (2026)
RouteLLM: Learning to Route LLMs with Preference Data
by: Ong, Isaac, et al.
Published: (2024)
by: Ong, Isaac, et al.
Published: (2024)
Ask, Answer, and Detect: Role-Playing LLMs for Personality Detection with Question-Conditioned Mixture-of-Experts
by: Lyu, Yifan, et al.
Published: (2025)
by: Lyu, Yifan, et al.
Published: (2025)
Understanding and Leveraging the Expert Specialization of Context Faithfulness in Mixture-of-Experts LLMs
by: Bai, Jun, et al.
Published: (2025)
by: Bai, Jun, et al.
Published: (2025)
ReMiT: RL-Guided Mid-Training for Iterative LLM Evolution
by: Huang, Junjie, et al.
Published: (2026)
by: Huang, Junjie, et al.
Published: (2026)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
by: Zhou, Hao, et al.
Published: (2024)
by: Zhou, Hao, et al.
Published: (2024)
Efficient Preference-based Reinforcement Learning via Aligned Experience Estimation
by: Bai, Fengshuo, et al.
Published: (2024)
by: Bai, Fengshuo, et al.
Published: (2024)
FaST: Feature-aware Sampling and Tuning for Personalized Preference Alignment with Limited Data
by: Thonet, Thibaut, et al.
Published: (2025)
by: Thonet, Thibaut, et al.
Published: (2025)
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts
by: Jawahar, Ganesh, et al.
Published: (2023)
by: Jawahar, Ganesh, et al.
Published: (2023)
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference
by: Han, Yang, et al.
Published: (2024)
by: Han, Yang, et al.
Published: (2024)
RedAgent: Red Teaming Large Language Models with Context-aware Autonomous Language Agent
by: Xu, Huiyu, et al.
Published: (2024)
by: Xu, Huiyu, et al.
Published: (2024)
Multilingual Routing in Mixture-of-Experts
by: Bandarkar, Lucas, et al.
Published: (2025)
by: Bandarkar, Lucas, et al.
Published: (2025)
PersonaMail: Learning and Adapting Personal Communication Preferences for Context-Aware Email Writing
by: Yao, Rui, et al.
Published: (2026)
by: Yao, Rui, et al.
Published: (2026)
Membership Inference Attacks Against In-Context Learning
by: Wen, Rui, et al.
Published: (2024)
by: Wen, Rui, et al.
Published: (2024)
Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling
by: Ran, Junfeng, et al.
Published: (2025)
by: Ran, Junfeng, et al.
Published: (2025)
Guided Profile Generation Improves Personalization with LLMs
by: Zhang, Jiarui
Published: (2024)
by: Zhang, Jiarui
Published: (2024)
Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
by: Zhao, Siyan, et al.
Published: (2025)
by: Zhao, Siyan, et al.
Published: (2025)
Similar Items
-
Rethinking Diverse Human Preference Learning through Principal Component Analysis
by: Luo, Feng, et al.
Published: (2025) -
Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
by: Yang, Rui, et al.
Published: (2024) -
Self-supervised Attribute-aware Dynamic Preference Ranking Alignment
by: Yang, Hongyu, et al.
Published: (2025) -
Personality-aware Human-centric Multimodal Reasoning: A New Task, Dataset and Baselines
by: Zhu, Yaochen, et al.
Published: (2023) -
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
by: Wang, Zihan, et al.
Published: (2025)