Reduced-Rank Multi-objective Policy Learning and Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Nwankwo, Ezinne, Jordan, Michael I., Zhou, Angela |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Batch-Adaptive Causal Annotations
by: Nwankwo, Ezinne, et al.
Published: (2025)
by: Nwankwo, Ezinne, et al.
Published: (2025)
Auditing LLMs for Algorithmic Fairness in Casenote-Augmented Tabular Prediction
by: Lee, Xiao Qi, et al.
Published: (2026)
by: Lee, Xiao Qi, et al.
Published: (2026)
Differentiating Policies for Non-Myopic Bayesian Optimization
by: Nwankwo, Darian, et al.
Published: (2024)
by: Nwankwo, Darian, et al.
Published: (2024)
Data-Driven Influence Functions for Optimization-Based Causal Inference
by: Jordan, Michael I., et al.
Published: (2022)
by: Jordan, Michael I., et al.
Published: (2022)
Safety vs. Performance: How Multi-Objective Learning Reduces Barriers to Market Entry
by: Jagadeesan, Meena, et al.
Published: (2024)
by: Jagadeesan, Meena, et al.
Published: (2024)
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
An Optimistic Algorithm for Online Convex Optimization with Adversarial Constraints
by: Lekeufack, Jordan, et al.
Published: (2024)
by: Lekeufack, Jordan, et al.
Published: (2024)
Pessimistic Off-Policy Optimization for Learning to Rank
by: Cief, Matej, et al.
Published: (2022)
by: Cief, Matej, et al.
Published: (2022)
Adaptive Coverage Policies in Conformal Prediction
by: Gauthier, Etienne, et al.
Published: (2025)
by: Gauthier, Etienne, et al.
Published: (2025)
A Gentle Introduction to Gradient-Based Optimization and Variational Inequalities for Machine Learning
by: Wadia, Neha S., et al.
Published: (2023)
by: Wadia, Neha S., et al.
Published: (2023)
Cooperative Multi-Agent Deep Reinforcement Learning in Content Ranking Optimization
by: Qin, Zhou, et al.
Published: (2024)
by: Qin, Zhou, et al.
Published: (2024)
Mind the Gap: Optimal and Equitable Encouragement Policies
by: Zhou, Angela
Published: (2023)
by: Zhou, Angela
Published: (2023)
GOPO: Policy Optimization using Ranked Rewards
by: Choi, Kyuseong, et al.
Published: (2026)
by: Choi, Kyuseong, et al.
Published: (2026)
Locally Adaptive Multi-Objective Learning
by: Kaur, Jivat Neet, et al.
Published: (2026)
by: Kaur, Jivat Neet, et al.
Published: (2026)
F-GRPO: Factorized Group-Relative Policy Optimization for Unified Candidate Generation and Ranking
by: Surana, Rohan, et al.
Published: (2026)
by: Surana, Rohan, et al.
Published: (2026)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Balancing Multi-modal Sensor Learning via Multi-objective Optimization
by: Fernando, Heshan, et al.
Published: (2025)
by: Fernando, Heshan, et al.
Published: (2025)
Matrix Low-Rank Trust Region Policy Optimization
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Accelerated First-Order Optimization under Nonlinear Constraints
by: Muehlebach, Michael, et al.
Published: (2023)
by: Muehlebach, Michael, et al.
Published: (2023)
Transfer Learning for High-dimensional Reduced Rank Time Series Models
by: Safikhani, Mingliang Ma Abolfazl
Published: (2025)
by: Safikhani, Mingliang Ma Abolfazl
Published: (2025)
Model-Based Learning of Near-Optimal Finite-Window Policies in POMDPs
by: Jordan, Philip, et al.
Published: (2026)
by: Jordan, Philip, et al.
Published: (2026)
Multi-Agent Reinforcement Learning for Unmanned Aerial Vehicle Coordination by Multi-Critic Policy Gradient Optimization
by: Alon, Yoav, et al.
Published: (2020)
by: Alon, Yoav, et al.
Published: (2020)
On Constraints in First-Order Optimization: A View from Non-Smooth Dynamical Systems
by: Muehlebach, Michael, et al.
Published: (2021)
by: Muehlebach, Michael, et al.
Published: (2021)
RewardRank: Optimizing True Learning-to-Rank Utility
by: Bhatt, Gaurav, et al.
Published: (2025)
by: Bhatt, Gaurav, et al.
Published: (2025)
Learning Fair Ranking Policies via Differentiable Optimization of Ordered Weighted Averages
by: Dinh, My H., et al.
Published: (2024)
by: Dinh, My H., et al.
Published: (2024)
Statistical Collusion by Collectives on Learning Platforms
by: Gauthier, Etienne, et al.
Published: (2025)
by: Gauthier, Etienne, et al.
Published: (2025)
Towards Off-Policy Reinforcement Learning for Ranking Policies with Human Feedback
by: Xiao, Teng, et al.
Published: (2024)
by: Xiao, Teng, et al.
Published: (2024)
Optimizing adaptive sampling via Policy Ranking
by: Nadeem, Hassan, et al.
Published: (2024)
by: Nadeem, Hassan, et al.
Published: (2024)
Off-Policy Evaluation for Ranking Policies under Deterministic Logging Policies
by: Tanaka, Koichi, et al.
Published: (2026)
by: Tanaka, Koichi, et al.
Published: (2026)
A Variance-Reduced Cubic-Regularized Newton for Policy Optimization
by: Sun, Cheng, et al.
Published: (2025)
by: Sun, Cheng, et al.
Published: (2025)
The Rank-Reduced Kalman Filter: Approximate Dynamical-Low-Rank Filtering In High Dimensions
by: Schmidt, Jonathan, et al.
Published: (2023)
by: Schmidt, Jonathan, et al.
Published: (2023)
ECPO: Evidence-Coupled Policy Optimization for Evidence-Certified Candidate Ranking
by: Hu, Miaobo, et al.
Published: (2026)
by: Hu, Miaobo, et al.
Published: (2026)
A Study of Plasticity Loss in On-Policy Deep Reinforcement Learning
by: Juliani, Arthur, et al.
Published: (2024)
by: Juliani, Arthur, et al.
Published: (2024)
Extreme Value Policy Optimization for Safe Reinforcement Learning
by: Gao, Shiqing, et al.
Published: (2026)
by: Gao, Shiqing, et al.
Published: (2026)
HCPO: Hierarchical Conductor-Based Policy Optimization in Multi-Agent Reinforcement Learning
by: Liu, Zejiao, et al.
Published: (2025)
by: Liu, Zejiao, et al.
Published: (2025)
A resource-constrained stochastic scheduling algorithm for homeless street outreach and gleaning edible food
by: Artman, Conor M., et al.
Published: (2024)
by: Artman, Conor M., et al.
Published: (2024)
Two-Timescale Gradient Descent Ascent Algorithms for Nonconvex Minimax Optimization
by: Lin, Tianyi, et al.
Published: (2024)
by: Lin, Tianyi, et al.
Published: (2024)
Low Rank Learning for Offline Query Optimization
by: Yi, Zixuan, et al.
Published: (2025)
by: Yi, Zixuan, et al.
Published: (2025)
Multi-CATE: Multi-Accurate Conditional Average Treatment Effect Estimation Robust to Unknown Covariate Shifts
by: Kern, Christoph, et al.
Published: (2024)
by: Kern, Christoph, et al.
Published: (2024)
Multi-Objective and Mixed-Reward Reinforcement Learning via Reward-Decorrelated Policy Optimization
by: Bai, Yang, et al.
Published: (2026)
by: Bai, Yang, et al.
Published: (2026)
Similar Items
-
Batch-Adaptive Causal Annotations
by: Nwankwo, Ezinne, et al.
Published: (2025) -
Auditing LLMs for Algorithmic Fairness in Casenote-Augmented Tabular Prediction
by: Lee, Xiao Qi, et al.
Published: (2026) -
Differentiating Policies for Non-Myopic Bayesian Optimization
by: Nwankwo, Darian, et al.
Published: (2024) -
Data-Driven Influence Functions for Optimization-Based Causal Inference
by: Jordan, Michael I., et al.
Published: (2022) -
Safety vs. Performance: How Multi-Objective Learning Reduces Barriers to Market Entry
by: Jagadeesan, Meena, et al.
Published: (2024)