Saved in:
| Main Authors: | Hudson, Benjamin, Charlin, Laurent, Frejinger, Emma |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.17139 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Arc travel time and path choice model estimation subsumed
by: Mohammadpour, Sobhan, et al.
Published: (2022)
by: Mohammadpour, Sobhan, et al.
Published: (2022)
A Survey of Contextual Optimization Methods for Decision Making under Uncertainty
by: Sadana, Utsav, et al.
Published: (2023)
by: Sadana, Utsav, et al.
Published: (2023)
Decoupling regularization from the action space
by: Mohammadpour, Sobhan, et al.
Published: (2024)
by: Mohammadpour, Sobhan, et al.
Published: (2024)
Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization
by: Penaloza, Emiliano, et al.
Published: (2025)
by: Penaloza, Emiliano, et al.
Published: (2025)
Maximum entropy GFlowNets with soft Q-learning
by: Mohammadpour, Sobhan, et al.
Published: (2023)
by: Mohammadpour, Sobhan, et al.
Published: (2023)
Integrating Present and Past in Unsupervised Continual Learning
by: Zhang, Yipeng, et al.
Published: (2024)
by: Zhang, Yipeng, et al.
Published: (2024)
Affective Music Recommendation: A Rollout-Based World Model for Offline Preference Optimization
by: Chan, Audrey, et al.
Published: (2026)
by: Chan, Audrey, et al.
Published: (2026)
Self-Supervised Learning from Structural Invariance
by: Zhang, Yipeng, et al.
Published: (2026)
by: Zhang, Yipeng, et al.
Published: (2026)
TEARS: Textual Representations for Scrutable Recommendations
by: Penaloza, Emiliano, et al.
Published: (2024)
by: Penaloza, Emiliano, et al.
Published: (2024)
Bayesian learning of Causal Structure and Mechanisms with GFlowNets and Variational Bayes
by: Nishikawa-Toomey, Mizu, et al.
Published: (2022)
by: Nishikawa-Toomey, Mizu, et al.
Published: (2022)
Active Learning for Stochastic Contextual Linear Bandits
by: Brunskill, Emma, et al.
Published: (2026)
by: Brunskill, Emma, et al.
Published: (2026)
Relative Explanations for Contextual Problems with Endogenous Uncertainty: An Application to Competitive Facility Location
by: Ramírez-Ayerbe, Jasone, et al.
Published: (2025)
by: Ramírez-Ayerbe, Jasone, et al.
Published: (2025)
Discovering Data Structures: Nearest Neighbor Search and Beyond
by: Salemohamed, Omar, et al.
Published: (2024)
by: Salemohamed, Omar, et al.
Published: (2024)
Privileged Information Distillation for Language Models
by: Penaloza, Emiliano, et al.
Published: (2026)
by: Penaloza, Emiliano, et al.
Published: (2026)
Contextual Online Uncertainty-Aware Preference Learning for Human Feedback
by: Lu, Nan, et al.
Published: (2025)
by: Lu, Nan, et al.
Published: (2025)
Contextualized Hybrid Ensemble Q-learning: Learning Fast with Control Priors
by: Cramer, Emma, et al.
Published: (2024)
by: Cramer, Emma, et al.
Published: (2024)
Learning Parametric Distributions from Samples and Preferences
by: Jourdan, Marc, et al.
Published: (2025)
by: Jourdan, Marc, et al.
Published: (2025)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
by: Shen, Yi, et al.
Published: (2023)
by: Shen, Yi, et al.
Published: (2023)
Contextual Bandits for Unbounded Context Distributions
by: Zhao, Puning, et al.
Published: (2024)
by: Zhao, Puning, et al.
Published: (2024)
Contextual Preference Collaborative Measure Framework Based on Belief System
by: Yu, Hang, et al.
Published: (2025)
by: Yu, Hang, et al.
Published: (2025)
High Probability Bound for Cross-Learning Contextual Bandits with Unknown Context Distributions
by: Huang, Ruiyuan, et al.
Published: (2024)
by: Huang, Ruiyuan, et al.
Published: (2024)
Towards Modular LLMs by Building and Reusing a Library of LoRAs
by: Ostapenko, Oleksiy, et al.
Published: (2024)
by: Ostapenko, Oleksiy, et al.
Published: (2024)
Distributed Direct Preference Optimization
by: Jiang, Zhanhong
Published: (2026)
by: Jiang, Zhanhong
Published: (2026)
Generalizing Reward Modeling for Out-of-Distribution Preference Learning
by: Jia, Chen
Published: (2024)
by: Jia, Chen
Published: (2024)
Offline Contextual Bandit with Counterfactual Sample Identification
by: Gilotte, Alexandre, et al.
Published: (2025)
by: Gilotte, Alexandre, et al.
Published: (2025)
Multi-Type Preference Learning: Empowering Preference-Based Reinforcement Learning with Equal Preferences
by: Liu, Ziang, et al.
Published: (2024)
by: Liu, Ziang, et al.
Published: (2024)
Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
by: Siththaranjan, Anand, et al.
Published: (2023)
by: Siththaranjan, Anand, et al.
Published: (2023)
Joint Scoring Rules: Zero-Sum Competition Avoids Performative Prediction
by: Hudson, Rubi
Published: (2024)
by: Hudson, Rubi
Published: (2024)
Structure Detection for Contextual Reinforcement Learning
by: Zhou, Tianyue, et al.
Published: (2026)
by: Zhou, Tianyue, et al.
Published: (2026)
Optimizing the Landscape of LLM Embeddings with Dynamic Exploratory Graph Analysis for Generative Psychometrics: A Monte Carlo Study
by: Golino, Hudson
Published: (2026)
by: Golino, Hudson
Published: (2026)
A model-free approach for solving choice-based competitive facility location problems using simulation and submodularity
by: Legault, Robin, et al.
Published: (2022)
by: Legault, Robin, et al.
Published: (2022)
Fisher Random Walk: Automatic Debiasing Contextual Preference Inference for Large Language Model Evaluation
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
Model-Based Transfer Learning for Contextual Reinforcement Learning
by: Cho, Jung-Hoon, et al.
Published: (2024)
by: Cho, Jung-Hoon, et al.
Published: (2024)
Distributional Preference Alignment of LLMs via Optimal Transport
by: Melnyk, Igor, et al.
Published: (2024)
by: Melnyk, Igor, et al.
Published: (2024)
Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems
by: Carr, Jonathan Colaço, et al.
Published: (2026)
by: Carr, Jonathan Colaço, et al.
Published: (2026)
Distributed GNEP Algorithms without Multiplier Sharing and Applications to Multi-Robot Coordination and Contextual Bandit-Based Active Learning
by: Yin, Shao-An
Published: (2026)
by: Yin, Shao-An
Published: (2026)
Contextual Intelligence The Next Leap for Reinforcement Learning
by: Biedenkapp, André
Published: (2026)
by: Biedenkapp, André
Published: (2026)
Contextual Learning for Anomaly Detection in Tabular Data
by: King, Spencer, et al.
Published: (2025)
by: King, Spencer, et al.
Published: (2025)
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Similar Items
-
Arc travel time and path choice model estimation subsumed
by: Mohammadpour, Sobhan, et al.
Published: (2022) -
A Survey of Contextual Optimization Methods for Decision Making under Uncertainty
by: Sadana, Utsav, et al.
Published: (2023) -
Decoupling regularization from the action space
by: Mohammadpour, Sobhan, et al.
Published: (2024) -
Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization
by: Penaloza, Emiliano, et al.
Published: (2025) -
Maximum entropy GFlowNets with soft Q-learning
by: Mohammadpour, Sobhan, et al.
Published: (2023)