Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
Fuente:
arXiv
Saved in:
| Main Authors: | Mou, Zhiyu, Lv, Yiqin, Xu, Miao, Wang, Qi, Mao, Yixiu, Chen, Jinghao, Ye, Qichen, Li, Chao, Bai, Rongquan, Yu, Chuan, Xu, Jian, Zheng, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Permutation Equivariant Model-based Offline Reinforcement Learning for Auto-bidding
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
Large-Scale Auto-bidding with Nash Equilibrium Constraints
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
VAO: Validation-Aligned Optimization for Cross-Task Generative Auto-Bidding
by: Lv, Yiqin, et al.
Published: (2025)
by: Lv, Yiqin, et al.
Published: (2025)
GAS: Generative Auto-bidding with Post-training Search
by: Li, Yewen, et al.
Published: (2024)
by: Li, Yewen, et al.
Published: (2024)
AIGB: Generative Auto-bidding via Conditional Diffusion Modeling
by: Guo, Jiayan, et al.
Published: (2024)
by: Guo, Jiayan, et al.
Published: (2024)
HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Auction Mechanisms with Organic Traffic
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
Trajectory-wise Iterative Reinforcement Learning Framework for Auto-bidding
by: Li, Haoming, et al.
Published: (2024)
by: Li, Haoming, et al.
Published: (2024)
Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments
by: Qu, Yun, et al.
Published: (2025)
by: Qu, Yun, et al.
Published: (2025)
Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation
by: Wang, Cheems, et al.
Published: (2024)
by: Wang, Cheems, et al.
Published: (2024)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
by: Mao, Yixiu, et al.
Published: (2025)
by: Mao, Yixiu, et al.
Published: (2025)
Doubly Mild Generalization for Offline Reinforcement Learning
by: Mao, Yixiu, et al.
Published: (2024)
by: Mao, Yixiu, et al.
Published: (2024)
Model Predictive Task Sampling for Efficient and Robust Adaptation
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
by: Mao, Yixiu, et al.
Published: (2024)
by: Mao, Yixiu, et al.
Published: (2024)
Lightweight Auto-bidding based on Traffic Prediction in Live Advertising
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
BAT: Benchmark for Auto-bidding Task
by: Khirianova, Alexandra, et al.
Published: (2025)
by: Khirianova, Alexandra, et al.
Published: (2025)
Incentive Compatibility in the Auto-bidding World
by: Alimohammadi, Yeganeh, et al.
Published: (2023)
by: Alimohammadi, Yeganeh, et al.
Published: (2023)
On the Existence and Nonexistence of Splitter Sets
by: Yuan, Zhiyu, et al.
Published: (2025)
by: Yuan, Zhiyu, et al.
Published: (2025)
Constraint-Aware Generative Auto-bidding via Pareto-Prioritized Regret Optimization
by: Wu, Binglin, et al.
Published: (2026)
by: Wu, Binglin, et al.
Published: (2026)
Auto-bidding and Auctions in Online Advertising: A Survey
by: Aggarwal, Gagan, et al.
Published: (2024)
by: Aggarwal, Gagan, et al.
Published: (2024)
MEBS: Multi-task End-to-end Bid Shading for Multi-slot Display Advertising
by: Gong, Zhen, et al.
Published: (2024)
by: Gong, Zhen, et al.
Published: (2024)
Efficiency of Non-Truthful Auctions in Auto-bidding with Budget Constraints
by: Liaw, Christopher, et al.
Published: (2023)
by: Liaw, Christopher, et al.
Published: (2023)
Auto-bidding under Return-on-Spend Constraints with Uncertainty Quantification
by: Han, Jiale, et al.
Published: (2025)
by: Han, Jiale, et al.
Published: (2025)
Risk-Averse and Optimistic Advertiser Incentive Compatibility in Auto-bidding
by: Liaw, Christopher, et al.
Published: (2025)
by: Liaw, Christopher, et al.
Published: (2025)
RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning
by: Mao, Yixiu, et al.
Published: (2026)
by: Mao, Yixiu, et al.
Published: (2026)
OPRIDE: Offline Preference-based Reinforcement Learning via In-Dataset Exploration
by: Yang, Yiqin, et al.
Published: (2026)
by: Yang, Yiqin, et al.
Published: (2026)
Fewer May Be Better: Enhancing Offline Reinforcement Learning with Reduced Dataset
by: Yang, Yiqin, et al.
Published: (2025)
by: Yang, Yiqin, et al.
Published: (2025)
Optimal Boost Design for Auto-bidding Mechanism with Publisher Quality Constraints
by: Yan, Huanyu, et al.
Published: (2025)
by: Yan, Huanyu, et al.
Published: (2025)
CROP: Conservative Reward for Model-based Offline Policy Optimization
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
Teaching LLM to be Persuasive: Reward-Enhanced Policy Optimization for Alignment from Heterogeneous Rewards
by: Zeng, Xia, et al.
Published: (2025)
by: Zeng, Xia, et al.
Published: (2025)
Data-Enabled Policy and Value Iteration for Continuous-Time Linear Quadratic Output Feedback Control
by: Xie, Jun, et al.
Published: (2026)
by: Xie, Jun, et al.
Published: (2026)
Beyond Advertising: Mechanism Design for Platform-Wide Marketing Service "QuanZhanTui"
by: Li, Ningyuan, et al.
Published: (2025)
by: Li, Ningyuan, et al.
Published: (2025)
An Adaptable Budget Planner for Enhancing Budget-Constrained Auto-Bidding in Online Advertising
by: Duan, Zhijian, et al.
Published: (2025)
by: Duan, Zhijian, et al.
Published: (2025)
Auto-bidding in real-time auctions via Oracle Imitation Learning (OIL)
by: Chiappa, Alberto Silvio, et al.
Published: (2024)
by: Chiappa, Alberto Silvio, et al.
Published: (2024)
Proportional Dynamics in Linear Fisher Markets with Auto-bidding: Convergence, Incentives and Fairness
by: Li, Juncheng, et al.
Published: (2024)
by: Li, Juncheng, et al.
Published: (2024)
Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning
by: Huang, Zhiyu, et al.
Published: (2024)
by: Huang, Zhiyu, et al.
Published: (2024)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
by: Ding, Ruiyi, et al.
Published: (2026)
by: Ding, Ruiyi, et al.
Published: (2026)
Bayesian Design Principles for Offline-to-Online Reinforcement Learning
by: Hu, Hao, et al.
Published: (2024)
by: Hu, Hao, et al.
Published: (2024)
A multimodal personality prediction framework based on adaptive graph transformer network and multi‐task learning
by: Rongquan Wang, et al.
Published: (2025)
by: Rongquan Wang, et al.
Published: (2025)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
by: Jing, Tan, et al.
Published: (2025)
by: Jing, Tan, et al.
Published: (2025)
Similar Items
-
Permutation Equivariant Model-based Offline Reinforcement Learning for Auto-bidding
by: Mou, Zhiyu, et al.
Published: (2025) -
Large-Scale Auto-bidding with Nash Equilibrium Constraints
by: Mou, Zhiyu, et al.
Published: (2025) -
VAO: Validation-Aligned Optimization for Cross-Task Generative Auto-Bidding
by: Lv, Yiqin, et al.
Published: (2025) -
GAS: Generative Auto-bidding with Post-training Search
by: Li, Yewen, et al.
Published: (2024) -
AIGB: Generative Auto-bidding via Conditional Diffusion Modeling
by: Guo, Jiayan, et al.
Published: (2024)