HALO: Hindsight-Augmented Learning for Online Auto-Bidding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Pusen, Cao, Chenglong, Zhou, Xinyu, You, Jirong, Xu, Linhe, Xu, Feifan, Yuan, Shuo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Ensemble Shape Calibration: Multi-Field Post-hoc Calibration in Online Advertising
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
von: Yang, Shuai, et al.
Veröffentlicht: (2024)
Learning-Augmented Online Bidding in Stochastic Settings
von: Angelopoulos, Spyros, et al.
Veröffentlicht: (2025)
von: Angelopoulos, Spyros, et al.
Veröffentlicht: (2025)
An Adaptable Budget Planner for Enhancing Budget-Constrained Auto-Bidding in Online Advertising
von: Duan, Zhijian, et al.
Veröffentlicht: (2025)
von: Duan, Zhijian, et al.
Veröffentlicht: (2025)
Translating Flow to Policy via Hindsight Online Imitation
von: Zheng, Yitian, et al.
Veröffentlicht: (2025)
von: Zheng, Yitian, et al.
Veröffentlicht: (2025)
Bidding-Aware Retrieval for Multi-Stage Consistency in Online Advertising
von: Liu, Bin, et al.
Veröffentlicht: (2025)
von: Liu, Bin, et al.
Veröffentlicht: (2025)
VAO: Validation-Aligned Optimization for Cross-Task Generative Auto-Bidding
von: Lv, Yiqin, et al.
Veröffentlicht: (2025)
von: Lv, Yiqin, et al.
Veröffentlicht: (2025)
Optimal Return-to-Go Guided Decision Transformer for Auto-Bidding in Advertisement
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
von: Jiang, Hao, et al.
Veröffentlicht: (2025)
Generative Auto-Bidding with Value-Guided Explorations
von: Gao, Jingtong, et al.
Veröffentlicht: (2025)
von: Gao, Jingtong, et al.
Veröffentlicht: (2025)
Expert-Guided Diffusion Planner for Auto-Bidding
von: Peng, Yunshan, et al.
Veröffentlicht: (2025)
von: Peng, Yunshan, et al.
Veröffentlicht: (2025)
EBaReT: Expert-guided Bag Reward Transformer for Auto Bidding
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions
von: Hu, Zihao, et al.
Veröffentlicht: (2026)
von: Hu, Zihao, et al.
Veröffentlicht: (2026)
CHIP: Adaptive Compliance for Humanoid Control through Hindsight Perturbation
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
Online Bidding under RoS Constraints without Knowing the Value
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
Multi-agent Auto-Bidding with Latent Graph Diffusion Models
von: Huh, Dom, et al.
Veröffentlicht: (2025)
von: Huh, Dom, et al.
Veröffentlicht: (2025)
Maximum Entropy Hindsight Experience Replay
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
Q-Regularized Generative Auto-Bidding: From Suboptimal Trajectories to Optimal Policies
von: Zhang, Mingming, et al.
Veröffentlicht: (2026)
von: Zhang, Mingming, et al.
Veröffentlicht: (2026)
HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation
von: Zhang, Wenjing, et al.
Veröffentlicht: (2026)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2026)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2024)
von: Gao, Chen-Xiao, et al.
Veröffentlicht: (2024)
Dynamical-VAE-based Hindsight to Learn the Causal Dynamics of Factored-POMDPs
von: Han, Chao, et al.
Veröffentlicht: (2024)
von: Han, Chao, et al.
Veröffentlicht: (2024)
Provable Interactive Learning with Hindsight Instruction Feedback
von: Misra, Dipendra, et al.
Veröffentlicht: (2024)
von: Misra, Dipendra, et al.
Veröffentlicht: (2024)
Online Budgeted Matching with General Bids
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
von: Yang, Jianyi, et al.
Veröffentlicht: (2024)
Reinforcement Learning Based Bidding Framework with High-dimensional Bids in Power Markets
von: Liu, Jinyu, et al.
Veröffentlicht: (2024)
von: Liu, Jinyu, et al.
Veröffentlicht: (2024)
Safe Online Bid Optimization with Return on Investment and Budget Constraints
von: Castiglioni, Matteo, et al.
Veröffentlicht: (2022)
von: Castiglioni, Matteo, et al.
Veröffentlicht: (2022)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
Provably Efficient Partially Observable Risk-Sensitive Reinforcement Learning with Hindsight Observation
von: Zhang, Tonghe, et al.
Veröffentlicht: (2024)
von: Zhang, Tonghe, et al.
Veröffentlicht: (2024)
HiBid: A Cross-Channel Constrained Bidding System with Budget Allocation by Hierarchical Offline Deep Reinforcement Learning
von: Wang, Hao, et al.
Veröffentlicht: (2023)
von: Wang, Hao, et al.
Veröffentlicht: (2023)
Optimizing Search Advertising Strategies: Integrating Reinforcement Learning with Generalized Second-Price Auctions for Enhanced Ad Ranking and Bidding
von: Zhou, Chang, et al.
Veröffentlicht: (2024)
von: Zhou, Chang, et al.
Veröffentlicht: (2024)
Learning to Bid in Non-Stationary Repeated First-Price Auctions
von: Hu, Zihao, et al.
Veröffentlicht: (2025)
von: Hu, Zihao, et al.
Veröffentlicht: (2025)
Learning to Mitigate Externalities: the Coase Theorem with Hindsight Rationality
von: Scheid, Antoine, et al.
Veröffentlicht: (2024)
von: Scheid, Antoine, et al.
Veröffentlicht: (2024)
Hindsight Experience Replay Accelerates Proximal Policy Optimization
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
von: Romio, Gabriel, et al.
Veröffentlicht: (2026)
PIPER: Primitive-Informed Preference-based Hierarchical Reinforcement Learning via Hindsight Relabeling
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
von: Singh, Utsav, et al.
Veröffentlicht: (2024)
GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning
von: Lei, Xing, et al.
Veröffentlicht: (2025)
von: Lei, Xing, et al.
Veröffentlicht: (2025)
Hindsight PRIORs for Reward Learning from Human Preferences
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
Transductive and Learning-Augmented Online Regression
von: Raman, Vinod, et al.
Veröffentlicht: (2025)
von: Raman, Vinod, et al.
Veröffentlicht: (2025)
Generative Auto-Bidding in Large-Scale Competitive Auctions via Diffusion Completer-Aligner
von: Li, Yewen, et al.
Veröffentlicht: (2025)
von: Li, Yewen, et al.
Veröffentlicht: (2025)
Preliminary Tests of the Anticipatory Classifier System with Hindsight Experience Replay
von: Unold, Olgierd, et al.
Veröffentlicht: (2026)
von: Unold, Olgierd, et al.
Veröffentlicht: (2026)
HISR: Hindsight Information Modulated Segmental Process Rewards For Multi-turn Agentic Reinforcement Learning
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
von: Lu, Zhicong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Deep Ensemble Shape Calibration: Multi-Field Post-hoc Calibration in Online Advertising
von: Yang, Shuai, et al.
Veröffentlicht: (2024) -
Learning-Augmented Online Bidding in Stochastic Settings
von: Angelopoulos, Spyros, et al.
Veröffentlicht: (2025) -
An Adaptable Budget Planner for Enhancing Budget-Constrained Auto-Bidding in Online Advertising
von: Duan, Zhijian, et al.
Veröffentlicht: (2025) -
Translating Flow to Policy via Hindsight Online Imitation
von: Zheng, Yitian, et al.
Veröffentlicht: (2025) -
Bidding-Aware Retrieval for Multi-Stage Consistency in Online Advertising
von: Liu, Bin, et al.
Veröffentlicht: (2025)