Safe Online Bid Optimization with Return on Investment and Budget Constraints
Fuente:
arXiv
Guardado en:
| Autores principales: | Castiglioni, Matteo, Nuara, Alessandro, Romano, Giulia, Spadaro, Giorgio, Trovò, Francesco, Gatti, Nicola |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
No-Regret Learning Under Adversarial Resource Constraints: A Spending Plan Is All You Need!
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025)
Learning Adversarial MDPs with Stochastic Hard Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
Multi-Armed Bandits With Best-Action Queries
por: Bacchiocchi, Francesco, et al.
Publicado: (2026)
por: Bacchiocchi, Francesco, et al.
Publicado: (2026)
Safe Offline Reinforcement Learning with Real-Time Budget Constraints
por: Lin, Qian, et al.
Publicado: (2023)
por: Lin, Qian, et al.
Publicado: (2023)
Toward Optimal Regret in Robust Pricing: Decoupling Corruption and Time
por: Kalupahana, Kalana, et al.
Publicado: (2026)
por: Kalupahana, Kalana, et al.
Publicado: (2026)
Optimal Strong Regret and Violation in Constrained MDPs via Policy Optimization
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
Truly Adapting to Adversarial Constraints in Constrained MABs
por: Stradi, Francesco Emanuele, et al.
Publicado: (2026)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2026)
Beyond Hard Constraints: Budget-Conditioned Reachability For Safe Offline Reinforcement Learning
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
LeakSealer: A Semisupervised Defense for LLMs Against Prompt Injection and Leakage Attacks
por: Panebianco, Francesco, et al.
Publicado: (2025)
por: Panebianco, Francesco, et al.
Publicado: (2025)
Learning Constrained Markov Decision Processes With Non-stationary Rewards and Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
Online Learning under Budget and ROI Constraints via Weak Adaptivity
por: Castiglioni, Matteo, et al.
Publicado: (2023)
por: Castiglioni, Matteo, et al.
Publicado: (2023)
Online Optimization for Offline Safe Reinforcement Learning
por: Chemingui, Yassine, et al.
Publicado: (2025)
por: Chemingui, Yassine, et al.
Publicado: (2025)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
por: Yao, Yihang, et al.
Publicado: (2023)
por: Yao, Yihang, et al.
Publicado: (2023)
Best-of-Both-Worlds Policy Optimization for CMDPs with Bandit Feedback
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024)
Learning Concave Bid Shading Strategies in Online Auctions via Measure-valued Proximal Optimization
por: Nodozi, Iman, et al.
Publicado: (2025)
por: Nodozi, Iman, et al.
Publicado: (2025)
Information-Theoretic Safe Bayesian Optimization
por: Bottero, Alessandro G., et al.
Publicado: (2024)
por: Bottero, Alessandro G., et al.
Publicado: (2024)
MicroFlow: An Efficient Rust-Based Inference Engine for TinyML
por: Carnelos, Matteo, et al.
Publicado: (2024)
por: Carnelos, Matteo, et al.
Publicado: (2024)
VAO: Validation-Aligned Optimization for Cross-Task Generative Auto-Bidding
por: Lv, Yiqin, et al.
Publicado: (2025)
por: Lv, Yiqin, et al.
Publicado: (2025)
A Best-of-Both-Worlds Algorithm for Constrained MDPs with Long-Term Constraints
por: Germano, Jacopo, et al.
Publicado: (2023)
por: Germano, Jacopo, et al.
Publicado: (2023)
Delivery Optimized Discovery in Behavioral User Segmentation under Budget Constraint
por: Chopra, Harshita, et al.
Publicado: (2024)
por: Chopra, Harshita, et al.
Publicado: (2024)
Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond
por: Bacchiocchi, Francesco, et al.
Publicado: (2025)
por: Bacchiocchi, Francesco, et al.
Publicado: (2025)
Learning Optimal Contracts: How to Exploit Small Action Spaces
por: Bacchiocchi, Francesco, et al.
Publicado: (2023)
por: Bacchiocchi, Francesco, et al.
Publicado: (2023)
Behaviour Policy Optimization: Provably Lower Variance Return Estimates for Off-Policy Reinforcement Learning
por: Goodall, Alexander W., et al.
Publicado: (2025)
por: Goodall, Alexander W., et al.
Publicado: (2025)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
por: Anisimov, Maksim, et al.
Publicado: (2026)
por: Anisimov, Maksim, et al.
Publicado: (2026)
A Survey of Constraint Formulations in Safe Reinforcement Learning
por: Wachi, Akifumi, et al.
Publicado: (2024)
por: Wachi, Akifumi, et al.
Publicado: (2024)
SMLE: Safe Machine Learning via Embedded Overapproximation
por: Francobaldi, Matteo, et al.
Publicado: (2024)
por: Francobaldi, Matteo, et al.
Publicado: (2024)
$(ε, u)$-Adaptive Regret Minimization in Heavy-Tailed Bandits
por: Genalti, Gianmarco, et al.
Publicado: (2023)
por: Genalti, Gianmarco, et al.
Publicado: (2023)
Joint Continual Learning of Local Language Models and Cloud Offloading Decisions with Budget Constraints
por: Chen, Evan, et al.
Publicado: (2026)
por: Chen, Evan, et al.
Publicado: (2026)
Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning
por: Chemingui, Yassine, et al.
Publicado: (2024)
por: Chemingui, Yassine, et al.
Publicado: (2024)
SB-TRPO: Towards Safe Reinforcement Learning with Hard Constraints
por: Wagner, Dominik, et al.
Publicado: (2025)
por: Wagner, Dominik, et al.
Publicado: (2025)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
por: Low, Siow Meng, et al.
Publicado: (2024)
por: Low, Siow Meng, et al.
Publicado: (2024)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
por: Pandit, Kartik, et al.
Publicado: (2025)
por: Pandit, Kartik, et al.
Publicado: (2025)
Robust Shielding for Safe Reinforcement Learning
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
por: Court, Edwin Hamel-De le, et al.
Publicado: (2026)
Automating the loop in traffic incident management on highway
por: Cercola, Matteo, et al.
Publicado: (2025)
por: Cercola, Matteo, et al.
Publicado: (2025)
Data-Dependent Regret Bounds for Constrained MABs
por: Genalti, Gianmarco, et al.
Publicado: (2025)
por: Genalti, Gianmarco, et al.
Publicado: (2025)
Moments Matter:Stabilizing Policy Optimization using Return Distributions
por: Jabs, Dennis, et al.
Publicado: (2026)
por: Jabs, Dennis, et al.
Publicado: (2026)
Adaptive Budget Optimization for Multichannel Advertising Using Combinatorial Bandits
por: Gangopadhyay, Briti, et al.
Publicado: (2025)
por: Gangopadhyay, Briti, et al.
Publicado: (2025)
Gym4ReaL: A Suite for Benchmarking Real-World Reinforcement Learning
por: Salaorni, Davide, et al.
Publicado: (2025)
por: Salaorni, Davide, et al.
Publicado: (2025)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
por: Bozkurt, Alper Kamil, et al.
Publicado: (2026)
por: Bozkurt, Alper Kamil, et al.
Publicado: (2026)
Ejemplares similares
-
No-Regret Learning Under Adversarial Resource Constraints: A Spending Plan Is All You Need!
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025) -
Beyond Slater's Condition in Online CMDPs with Stochastic and Adversarial Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2025) -
Learning Adversarial MDPs with Stochastic Hard Constraints
por: Stradi, Francesco Emanuele, et al.
Publicado: (2024) -
Multi-Armed Bandits With Best-Action Queries
por: Bacchiocchi, Francesco, et al.
Publicado: (2026) -
Safe Offline Reinforcement Learning with Real-Time Budget Constraints
por: Lin, Qian, et al.
Publicado: (2023)