YRC-Bench: A Benchmark for Learning to Coordinate with Experts
Fuente:
arXiv
Guardado en:
| Autores principales: | Danesh, Mohamad H., Khanh, Nguyen X., Trinh, Tu, Plaut, Benjamin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Getting By Goal Misgeneralization With a Little Help From a Mentor
por: Trinh, Tu, et al.
Publicado: (2024)
por: Trinh, Tu, et al.
Publicado: (2024)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
por: Plaut, Benjamin, et al.
Publicado: (2024)
por: Plaut, Benjamin, et al.
Publicado: (2024)
Safety Training Persists Through Helpfulness Optimization in LLM Agents
por: Plaut, Benjamin
Publicado: (2026)
por: Plaut, Benjamin
Publicado: (2026)
Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards
por: Liaw, Sarah, et al.
Publicado: (2025)
por: Liaw, Sarah, et al.
Publicado: (2025)
On Linear Mode Connectivity of Mixture-of-Experts Architectures
por: Tran, Viet-Hoang, et al.
Publicado: (2025)
por: Tran, Viet-Hoang, et al.
Publicado: (2025)
Avoiding Catastrophe in Online Learning by Asking for Help
por: Plaut, Benjamin, et al.
Publicado: (2024)
por: Plaut, Benjamin, et al.
Publicado: (2024)
Safe Learning Under Irreversible Dynamics via Asking for Help
por: Plaut, Benjamin, et al.
Publicado: (2025)
por: Plaut, Benjamin, et al.
Publicado: (2025)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
por: Azran, Guy, et al.
Publicado: (2023)
por: Azran, Guy, et al.
Publicado: (2023)
Towards Layer-Wise Personalized Federated Learning: Adaptive Layer Disentanglement via Conflicting Gradients
por: Nguyen, Minh Duong, et al.
Publicado: (2024)
por: Nguyen, Minh Duong, et al.
Publicado: (2024)
Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning
por: Trinh, Tu, et al.
Publicado: (2022)
por: Trinh, Tu, et al.
Publicado: (2022)
Drift Q-Learning
por: Houssaini, Anas, et al.
Publicado: (2026)
por: Houssaini, Anas, et al.
Publicado: (2026)
Language Models are Bounded Pragmatic Speakers: Understanding RLHF from a Bayesian Cognitive Modeling Perspective
por: Nguyen, Khanh
Publicado: (2023)
por: Nguyen, Khanh
Publicado: (2023)
Taming the Tail in Class-Conditional GANs: Knowledge Sharing via Unconditional Training at Lower Resolutions
por: Khorram, Saeed, et al.
Publicado: (2024)
por: Khorram, Saeed, et al.
Publicado: (2024)
Safe Domain Randomization via Uncertainty-Aware Out-of-Distribution Detection and Policy Adaptation
por: Danesh, Mohamad H., et al.
Publicado: (2025)
por: Danesh, Mohamad H., et al.
Publicado: (2025)
GraphAllocBench: A Flexible Benchmark for Preference-Conditioned Multi-Objective Policy Learning
por: Jiang, Zhiheng, et al.
Publicado: (2026)
por: Jiang, Zhiheng, et al.
Publicado: (2026)
HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization
por: Tran, Trinh, et al.
Publicado: (2026)
por: Tran, Trinh, et al.
Publicado: (2026)
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
por: Nguyen, Dung V., et al.
Publicado: (2025)
por: Nguyen, Dung V., et al.
Publicado: (2025)
ML-Driven Approaches to Combat Medicare Fraud: Advances in Class Imbalance Solutions, Feature Engineering, Adaptive Learning, and Business Impact
por: Farahmandazad, Dorsa, et al.
Publicado: (2025)
por: Farahmandazad, Dorsa, et al.
Publicado: (2025)
DHG-Bench: A Comprehensive Benchmark for Deep Hypergraph Learning
por: Li, Fan, et al.
Publicado: (2025)
por: Li, Fan, et al.
Publicado: (2025)
TopoBench: A Framework for Benchmarking Topological Deep Learning
por: Telyatnikov, Lev, et al.
Publicado: (2024)
por: Telyatnikov, Lev, et al.
Publicado: (2024)
Mixture of Experts Meets Prompt-Based Continual Learning
por: Le, Minh, et al.
Publicado: (2024)
por: Le, Minh, et al.
Publicado: (2024)
CausalBench: A Comprehensive Benchmark for Causal Learning Capability of LLMs
por: Zhou, Yu, et al.
Publicado: (2024)
por: Zhou, Yu, et al.
Publicado: (2024)
CAMEx: Curvature-aware Merging of Experts
por: Nguyen, Dung V., et al.
Publicado: (2025)
por: Nguyen, Dung V., et al.
Publicado: (2025)
Toward Cost-efficient Adaptive Clinical Trials in Knee Osteoarthritis with Reinforcement Learning
por: Nguyen, Khanh, et al.
Publicado: (2024)
por: Nguyen, Khanh, et al.
Publicado: (2024)
CardBench: A Benchmark for Learned Cardinality Estimation in Relational Databases
por: Chronis, Yannis, et al.
Publicado: (2024)
por: Chronis, Yannis, et al.
Publicado: (2024)
ExtractBench: A Benchmark and Evaluation Methodology for Complex Structured Extraction
por: Ferguson, Nick, et al.
Publicado: (2026)
por: Ferguson, Nick, et al.
Publicado: (2026)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
por: Wu, Baoyuan, et al.
Publicado: (2024)
por: Wu, Baoyuan, et al.
Publicado: (2024)
RouterBench: A Benchmark for Multi-LLM Routing System
por: Hu, Qitian Jason, et al.
Publicado: (2024)
por: Hu, Qitian Jason, et al.
Publicado: (2024)
SMMILE: An Expert-Driven Benchmark for Multimodal Medical In-Context Learning
por: Rieff, Melanie, et al.
Publicado: (2025)
por: Rieff, Melanie, et al.
Publicado: (2025)
Multiphysics Bench: Benchmarking and Investigating Scientific Machine Learning for Multiphysics PDEs
por: Yang, Changfan, et al.
Publicado: (2025)
por: Yang, Changfan, et al.
Publicado: (2025)
Introducing CausalBench: A Flexible Benchmark Framework for Causal Analysis and Machine Learning
por: Kapkiç, Ahmet, et al.
Publicado: (2024)
por: Kapkiç, Ahmet, et al.
Publicado: (2024)
LLP-Bench: A Large Scale Tabular Benchmark for Learning from Label Proportions
por: Brahmbhatt, Anand, et al.
Publicado: (2023)
por: Brahmbhatt, Anand, et al.
Publicado: (2023)
Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety
por: Bonagiri, Vamshi Krishna, et al.
Publicado: (2025)
por: Bonagiri, Vamshi Krishna, et al.
Publicado: (2025)
Do Chatbot LLMs Talk Too Much? The YapBench Benchmark
por: Borisov, Vadim, et al.
Publicado: (2026)
por: Borisov, Vadim, et al.
Publicado: (2026)
Contractive Diffusion Policies: Robust Action Diffusion via Contractive Score-Based Sampling with Differential Equations
por: Abyaneh, Amin, et al.
Publicado: (2026)
por: Abyaneh, Amin, et al.
Publicado: (2026)
RelBench: A Benchmark for Deep Learning on Relational Databases
por: Robinson, Joshua, et al.
Publicado: (2024)
por: Robinson, Joshua, et al.
Publicado: (2024)
On Parameter Estimation in Deviated Gaussian Mixture of Experts
por: Nguyen, Huy, et al.
Publicado: (2024)
por: Nguyen, Huy, et al.
Publicado: (2024)
Adam Exploits $\ell_\infty$-geometry of Loss Landscape via Coordinate-wise Adaptivity
por: Xie, Shuo, et al.
Publicado: (2024)
por: Xie, Shuo, et al.
Publicado: (2024)
SemBench: A Benchmark for Semantic Query Processing Engines
por: Lao, Jiale, et al.
Publicado: (2025)
por: Lao, Jiale, et al.
Publicado: (2025)
PerturBench: Benchmarking Machine Learning Models for Cellular Perturbation Analysis
por: Wu, Yan, et al.
Publicado: (2024)
por: Wu, Yan, et al.
Publicado: (2024)
Ejemplares similares
-
Getting By Goal Misgeneralization With a Little Help From a Mentor
por: Trinh, Tu, et al.
Publicado: (2024) -
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
por: Plaut, Benjamin, et al.
Publicado: (2024) -
Safety Training Persists Through Helpfulness Optimization in LLM Agents
por: Plaut, Benjamin
Publicado: (2026) -
Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards
por: Liaw, Sarah, et al.
Publicado: (2025) -
On Linear Mode Connectivity of Mixture-of-Experts Architectures
por: Tran, Viet-Hoang, et al.
Publicado: (2025)