What is the Right Notion of Distance between Predict-then-Optimize Tasks?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rodriguez-Diaz, Paula, Kong, Lingkai, Wang, Kai, Alvarez-Melis, David, Tambe, Milind |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
von: Verma, Shresth, et al.
Veröffentlicht: (2024)
von: Verma, Shresth, et al.
Veröffentlicht: (2024)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
von: Wang, Haichuan, et al.
Veröffentlicht: (2026)
von: Wang, Haichuan, et al.
Veröffentlicht: (2026)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
von: Shah, Sanket, et al.
Veröffentlicht: (2023)
von: Shah, Sanket, et al.
Veröffentlicht: (2023)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
von: Biswas, Arpita, et al.
Veröffentlicht: (2023)
von: Biswas, Arpita, et al.
Veröffentlicht: (2023)
Rule-Bottleneck Reinforcement Learning: Joint Explanation and Decision Optimization for Resource Allocation with Language Agents
von: Tec, Mauricio, et al.
Veröffentlicht: (2025)
von: Tec, Mauricio, et al.
Veröffentlicht: (2025)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
von: Shah, Sanket, et al.
Veröffentlicht: (2024)
von: Shah, Sanket, et al.
Veröffentlicht: (2024)
Reinforcement learning with combinatorial actions for coupled restless bandits
von: Xu, Lily, et al.
Veröffentlicht: (2025)
von: Xu, Lily, et al.
Veröffentlicht: (2025)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
von: Verma, Shresth, et al.
Veröffentlicht: (2026)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
von: Martinson, Sarah, et al.
Veröffentlicht: (2025)
von: Martinson, Sarah, et al.
Veröffentlicht: (2025)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
Learning Graph Structures and Uncertainty for Accurate and Calibrated Time-series Forecasting
von: Kamarthi, Harshavardhan, et al.
Veröffentlicht: (2024)
von: Kamarthi, Harshavardhan, et al.
Veröffentlicht: (2024)
Contrasting local and global modeling with machine learning and satellite data: A case study estimating tree canopy height in African savannas
von: Rolf, Esther, et al.
Veröffentlicht: (2024)
von: Rolf, Esther, et al.
Veröffentlicht: (2024)
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
Evaluating the Effectiveness of Index-Based Treatment Allocation
von: Boehmer, Niclas, et al.
Veröffentlicht: (2024)
von: Boehmer, Niclas, et al.
Veröffentlicht: (2024)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Ambreesh, et al.
Veröffentlicht: (2025)
IGDA: Interactive Graph Discovery through Large Language Model Agents
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
von: Havrilla, Alex, et al.
Veröffentlicht: (2025)
Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order Information
von: Ai, Rui, et al.
Veröffentlicht: (2025)
von: Ai, Rui, et al.
Veröffentlicht: (2025)
Two Birds with One Stone: Enhancing Uncertainty Quantification and Interpretability with Graph Functional Neural Process
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
Trivialized Momentum Facilitates Diffusion Generative Modeling on Lie Groups
von: Zhu, Yuchen, et al.
Veröffentlicht: (2024)
von: Zhu, Yuchen, et al.
Veröffentlicht: (2024)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
Improving Health Information Access in the World's Largest Maternal Mobile Health Program via Bandit Algorithms
von: Lalan, Arshika, et al.
Veröffentlicht: (2024)
von: Lalan, Arshika, et al.
Veröffentlicht: (2024)
Can Interpretation Predict Behavior on Unseen Data?
von: Li, Victoria R., et al.
Veröffentlicht: (2025)
von: Li, Victoria R., et al.
Veröffentlicht: (2025)
LLM Advertisement based on Neuron Auctions
von: Yun, Peiran, et al.
Veröffentlicht: (2026)
von: Yun, Peiran, et al.
Veröffentlicht: (2026)
Transcendence: Generative Models Can Outperform The Experts That Train Them
von: Zhang, Edwin, et al.
Veröffentlicht: (2024)
von: Zhang, Edwin, et al.
Veröffentlicht: (2024)
Diffusion Models as Constrained Samplers for Optimization with Unknown Constraints
von: Kong, Lingkai, et al.
Veröffentlicht: (2024)
von: Kong, Lingkai, et al.
Veröffentlicht: (2024)
Policy-Embedded Graph Expansion: Networked HIV Testing with Diffusion-Driven Network Samples
von: Kangaslahti, Akseli, et al.
Veröffentlicht: (2026)
von: Kangaslahti, Akseli, et al.
Veröffentlicht: (2026)
Foundations of Artificial Intelligence Frameworks: Notion and Limits of AGI
von: Bui, Khanh Gia
Veröffentlicht: (2025)
von: Bui, Khanh Gia
Veröffentlicht: (2025)
An Explainable Multi-Task Similarity Measure: Integrating Accumulated Local Effects and Weighted Fréchet Distance
von: Hidalgo, Pablo, et al.
Veröffentlicht: (2026)
von: Hidalgo, Pablo, et al.
Veröffentlicht: (2026)
Robust Optimization with Diffusion Models for Green Security
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
OMNI: Open-endedness via Models of human Notions of Interestingness
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
DF2: Distribution-Free Decision-Focused Learning
von: Kong, Lingkai, et al.
Veröffentlicht: (2023)
von: Kong, Lingkai, et al.
Veröffentlicht: (2023)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing
von: Choo, Davin, et al.
Veröffentlicht: (2025)
von: Choo, Davin, et al.
Veröffentlicht: (2025)
Application-Driven Innovation in Machine Learning
von: Rolnick, David, et al.
Veröffentlicht: (2024)
von: Rolnick, David, et al.
Veröffentlicht: (2024)
Efficient Ensemble Selection from Binary and Pairwise Feedback
von: Neoh, Tzeh Yuan, et al.
Veröffentlicht: (2026)
von: Neoh, Tzeh Yuan, et al.
Veröffentlicht: (2026)
Social Environment Design
von: Zhang, Edwin, et al.
Veröffentlicht: (2024)
von: Zhang, Edwin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
von: Kong, Lingkai, et al.
Veröffentlicht: (2025) -
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
von: Verma, Shresth, et al.
Veröffentlicht: (2024) -
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
von: Wang, Haichuan, et al.
Veröffentlicht: (2026) -
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
von: Shah, Sanket, et al.
Veröffentlicht: (2023) -
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)