Beyond Augmented-Action Surrogates for Multi-Expert Learning-to-Defer
Fuente:
arXiv
Saved in:
| Main Authors: | Montreuil, Yannis, Carlier, Axel, Ng, Lai Xing, Ooi, Wei Tsang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning-to-Defer with Expert-Conditional Advice
by: Montreuil, Yannis, et al.
Published: (2026)
by: Montreuil, Yannis, et al.
Published: (2026)
One-Stage Top-$k$ Learning-to-Defer: Score-Based Surrogates with Theoretical Guarantees
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
Why Ask One When You Can Ask $k$? Learning-to-Defer to the Top-$k$ Experts
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
Adversarial Robustness in Two-Stage Learning-to-Defer: Algorithms and Guarantees
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
Online Learning-to-Defer with Varying Experts
by: Duy, Dang Hoang, et al.
Published: (2026)
by: Duy, Dang Hoang, et al.
Published: (2026)
Adversarial Robustness in One-Stage Learning-to-Defer
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
A Two-Stage Learning-to-Defer Approach for Multi-Task Learning
by: Montreuil, Yannis, et al.
Published: (2024)
by: Montreuil, Yannis, et al.
Published: (2024)
Learning-to-Defer in Non-Stationary Time Series via Switching State-Space Models
by: Montreuil, Yannis, et al.
Published: (2026)
by: Montreuil, Yannis, et al.
Published: (2026)
Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees
by: Montreuil, Yannis, et al.
Published: (2024)
by: Montreuil, Yannis, et al.
Published: (2024)
When More Experts Hurt: Underfitting in Multi-Expert Learning to Defer
by: Liu, Shuqi, et al.
Published: (2026)
by: Liu, Shuqi, et al.
Published: (2026)
Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate Experts
by: Pugnana, Andrea, et al.
Published: (2025)
by: Pugnana, Andrea, et al.
Published: (2025)
Principled Approaches for Learning to Defer with Multiple Experts
by: Mao, Anqi, et al.
Published: (2023)
by: Mao, Anqi, et al.
Published: (2023)
Learning to Defer for Causal Discovery with Imperfect Experts
by: Clivio, Oscar, et al.
Published: (2025)
by: Clivio, Oscar, et al.
Published: (2025)
Cost-Sensitive Learning to Defer to Multiple Experts with Workload Constraints
by: Alves, Jean V., et al.
Published: (2024)
by: Alves, Jean V., et al.
Published: (2024)
Mastering Multiple-Expert Routing: Realizable $H$-Consistency and Strong Guarantees for Learning to Defer
by: Mao, Anqi, et al.
Published: (2025)
by: Mao, Anqi, et al.
Published: (2025)
Expert Proximity as Surrogate Rewards for Single Demonstration Imitation Learning
by: Chiang, Chia-Cheng, et al.
Published: (2024)
by: Chiang, Chia-Cheng, et al.
Published: (2024)
Deferred is Better: A Framework for Multi-Granularity Deferred Interaction of Heterogeneous Features
by: Xu, Yi, et al.
Published: (2026)
by: Xu, Yi, et al.
Published: (2026)
Learning to Partially Defer for Sequences
by: Rayan, Sahana, et al.
Published: (2025)
by: Rayan, Sahana, et al.
Published: (2025)
RADAR: Recall Augmentation through Deferred Asynchronous Retrieval
by: Jaspal, Amit, et al.
Published: (2025)
by: Jaspal, Amit, et al.
Published: (2025)
Beyond-Expert Performance with Limited Demonstrations: Efficient Imitation Learning with Double Exploration
by: Zhao, Heyang, et al.
Published: (2025)
by: Zhao, Heyang, et al.
Published: (2025)
Enhanced Parcel Arrival Forecasting for Logistic Hubs: An Ensemble Deep Learning Approach
by: Pan, Xinyue, et al.
Published: (2026)
by: Pan, Xinyue, et al.
Published: (2026)
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Learning to Defer: A Survey
by: Strong, Joshua, et al.
Published: (2025)
by: Strong, Joshua, et al.
Published: (2025)
No Need for Learning to Defer? A Training Free Deferral Framework to Multiple Experts through Conformal Prediction
by: Bary, Tim, et al.
Published: (2025)
by: Bary, Tim, et al.
Published: (2025)
Beyond Non-Expert Demonstrations: Outcome-Driven Action Constraint for Offline Reinforcement Learning
by: Jiang, Ke, et al.
Published: (2025)
by: Jiang, Ke, et al.
Published: (2025)
Learning to Defer to a Population: A Meta-Learning Approach
by: Tailor, Dharmesh, et al.
Published: (2024)
by: Tailor, Dharmesh, et al.
Published: (2024)
Density-Ratio Losses for Post-Hoc Learning to Defer
by: Soen, Alexander, et al.
Published: (2026)
by: Soen, Alexander, et al.
Published: (2026)
Fatigue-Aware Learning to Defer via Constrained Optimisation
by: Zhang, Zheng, et al.
Published: (2026)
by: Zhang, Zheng, et al.
Published: (2026)
A Unifying Post-Processing Framework for Multi-Objective Learn-to-Defer Problems
by: Charusaie, Mohammad-Amin, et al.
Published: (2024)
by: Charusaie, Mohammad-Amin, et al.
Published: (2024)
Continual Learning Beyond Experience Rehearsal and Full Model Surrogates
by: Bhat, Prashant, et al.
Published: (2025)
by: Bhat, Prashant, et al.
Published: (2025)
Learn to Change the World: Multi-level Reinforcement Learning with Model-Changing Actions
by: Lu, Ziqing, et al.
Published: (2025)
by: Lu, Ziqing, et al.
Published: (2025)
Deep Learning Meets Queue-Reactive: A Framework for Realistic Limit Order Book Simulation
by: Bodor, Hamza, et al.
Published: (2025)
by: Bodor, Hamza, et al.
Published: (2025)
Realizable $H$-Consistent and Bayes-Consistent Loss Functions for Learning to Defer
by: Mao, Anqi, et al.
Published: (2024)
by: Mao, Anqi, et al.
Published: (2024)
MAST: A Multi-fidelity Augmented Surrogate model via Spatial Trust-weighting
by: Nasr, Ahmed Mohamed Eisa, et al.
Published: (2026)
by: Nasr, Ahmed Mohamed Eisa, et al.
Published: (2026)
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
A Continuous Encoding-Based Representation for Efficient Multi-Fidelity Multi-Objective Neural Architecture Search
by: Wei, Zhao, et al.
Published: (2025)
by: Wei, Zhao, et al.
Published: (2025)
Structure Learning with Continuous Optimization: A Sober Look and Beyond
by: Ng, Ignavier, et al.
Published: (2023)
by: Ng, Ignavier, et al.
Published: (2023)
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
by: Pu, Yuanhao, et al.
Published: (2026)
by: Pu, Yuanhao, et al.
Published: (2026)
When to Accept Automated Predictions and When to Defer to Human Judgment?
by: Sikar, Daniel, et al.
Published: (2024)
by: Sikar, Daniel, et al.
Published: (2024)
Electric Vehicle Charging Load Forecasting: An Experimental Comparison of Machine Learning Methods
by: Kyriakopoulos, Iason, et al.
Published: (2025)
by: Kyriakopoulos, Iason, et al.
Published: (2025)
Similar Items
-
Learning-to-Defer with Expert-Conditional Advice
by: Montreuil, Yannis, et al.
Published: (2026) -
One-Stage Top-$k$ Learning-to-Defer: Score-Based Surrogates with Theoretical Guarantees
by: Montreuil, Yannis, et al.
Published: (2025) -
Why Ask One When You Can Ask $k$? Learning-to-Defer to the Top-$k$ Experts
by: Montreuil, Yannis, et al.
Published: (2025) -
Adversarial Robustness in Two-Stage Learning-to-Defer: Algorithms and Guarantees
by: Montreuil, Yannis, et al.
Published: (2025) -
Online Learning-to-Defer with Varying Experts
by: Duy, Dang Hoang, et al.
Published: (2026)