Auto-bidding in real-time auctions via Oracle Imitation Learning (OIL)
Fuente:
arXiv
Saved in:
| Main Authors: | Chiappa, Alberto Silvio, Gangopadhyay, Briti, Wang, Zhao, Takamatsu, Shingo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Budget Optimization for Multichannel Advertising Using Combinatorial Bandits
by: Gangopadhyay, Briti, et al.
Published: (2025)
by: Gangopadhyay, Briti, et al.
Published: (2025)
Integrating Domain Knowledge for handling Limited Data in Offline RL
by: Gangopadhyay, Briti, et al.
Published: (2024)
by: Gangopadhyay, Briti, et al.
Published: (2024)
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Talk Structurally, Act Hierarchically: A Collaborative Framework for LLM Multi-Agent Systems
by: Wang, Zhao, et al.
Published: (2025)
by: Wang, Zhao, et al.
Published: (2025)
Forecasting Clicks in Digital Advertising: Multimodal Inputs and Interpretable Outputs
by: Gangopadhyay, Briti, et al.
Published: (2025)
by: Gangopadhyay, Briti, et al.
Published: (2025)
OKG: On-the-Fly Keyword Generation in Sponsored Search Advertising
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
KINESIS: Motion Imitation for Human Musculoskeletal Locomotion
by: Simos, Merkourios, et al.
Published: (2025)
by: Simos, Merkourios, et al.
Published: (2025)
Permutation Equivariant Model-based Offline Reinforcement Learning for Auto-bidding
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
BAT: Benchmark for Auto-bidding Task
by: Khirianova, Alexandra, et al.
Published: (2025)
by: Khirianova, Alexandra, et al.
Published: (2025)
Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
AIGB: Generative Auto-bidding via Conditional Diffusion Modeling
by: Guo, Jiayan, et al.
Published: (2024)
by: Guo, Jiayan, et al.
Published: (2024)
Large-Scale Auto-bidding with Nash Equilibrium Constraints
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
Trajectory-wise Iterative Reinforcement Learning Framework for Auto-bidding
by: Li, Haoming, et al.
Published: (2024)
by: Li, Haoming, et al.
Published: (2024)
Arnold: a generalist muscle transformer policy
by: Chiappa, Alberto Silvio, et al.
Published: (2025)
by: Chiappa, Alberto Silvio, et al.
Published: (2025)
Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery
by: Chen, Tingting, et al.
Published: (2025)
by: Chen, Tingting, et al.
Published: (2025)
Imitation Learning as Return Distribution Matching
by: Lazzati, Filippo, et al.
Published: (2025)
by: Lazzati, Filippo, et al.
Published: (2025)
Adversarial Imitation Learning via Boosting
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
Imitation Learning via Focused Satisficing
by: Shah, Rushit N., et al.
Published: (2025)
by: Shah, Rushit N., et al.
Published: (2025)
Boolean Satisfiability via Imitation Learning
by: Zhang, Zewei, et al.
Published: (2025)
by: Zhang, Zewei, et al.
Published: (2025)
Reinforcement Learning via Implicit Imitation Guidance
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
Self-Improvement Imitation with Biologically Guided Search for Protein Design Under Oracle Budgets
by: Khanna, Ashima, et al.
Published: (2026)
by: Khanna, Ashima, et al.
Published: (2026)
Generalization Capability for Imitation Learning
by: Wang, Yixiao
Published: (2025)
by: Wang, Yixiao
Published: (2025)
Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles
by: Tang, Zhiwei, et al.
Published: (2023)
by: Tang, Zhiwei, et al.
Published: (2023)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
by: Fan, Jiangdong, et al.
Published: (2024)
by: Fan, Jiangdong, et al.
Published: (2024)
Imitation Bootstrapped Reinforcement Learning
by: Hu, Hengyuan, et al.
Published: (2023)
by: Hu, Hengyuan, et al.
Published: (2023)
Quantifying Generalisation in Imitation Learning
by: Gavenski, Nathan, et al.
Published: (2025)
by: Gavenski, Nathan, et al.
Published: (2025)
Diverse Policies Recovering via Pointwise Mutual Information Weighted Imitation Learning
by: Yang, Hanlin, et al.
Published: (2024)
by: Yang, Hanlin, et al.
Published: (2024)
Is Efficient PAC Learning Possible with an Oracle That Responds 'Yes' or 'No'?
by: Daskalakis, Constantinos, et al.
Published: (2024)
by: Daskalakis, Constantinos, et al.
Published: (2024)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
by: Bobrin, Maksim, et al.
Published: (2024)
by: Bobrin, Maksim, et al.
Published: (2024)
OMS: On-the-fly, Multi-Objective, Self-Reflective Ad Keyword Generation via LLM Agent
by: Chen, Bowen, et al.
Published: (2025)
by: Chen, Bowen, et al.
Published: (2025)
Offline Imitation Learning Through Graph Search and Retrieval
by: Yin, Zhao-Heng, et al.
Published: (2024)
by: Yin, Zhao-Heng, et al.
Published: (2024)
Cross-Domain Imitation Learning via Optimal Transport
by: Fickinger, Arnaud, et al.
Published: (2021)
by: Fickinger, Arnaud, et al.
Published: (2021)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
OIL-AD: An Anomaly Detection Framework for Sequential Decision Sequences
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
Denoising-based Contractive Imitation Learning
by: Shen, Macheng, et al.
Published: (2025)
by: Shen, Macheng, et al.
Published: (2025)
Noise-Guided Transport for Imitation Learning
by: Blondé, Lionel, et al.
Published: (2025)
by: Blondé, Lionel, et al.
Published: (2025)
Sample-efficient Adversarial Imitation Learning
by: Jung, Dahuin, et al.
Published: (2023)
by: Jung, Dahuin, et al.
Published: (2023)
Imitation Learning for Multi-turn LM Agents via On-policy Expert Corrections
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
by: Namkoong, Hongseok, et al.
Published: (2020)
by: Namkoong, Hongseok, et al.
Published: (2020)
Similar Items
-
Adaptive Budget Optimization for Multichannel Advertising Using Combinatorial Bandits
by: Gangopadhyay, Briti, et al.
Published: (2025) -
Integrating Domain Knowledge for handling Limited Data in Offline RL
by: Gangopadhyay, Briti, et al.
Published: (2024) -
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024) -
Talk Structurally, Act Hierarchically: A Collaborative Framework for LLM Multi-Agent Systems
by: Wang, Zhao, et al.
Published: (2025) -
Forecasting Clicks in Digital Advertising: Multimodal Inputs and Interpretable Outputs
by: Gangopadhyay, Briti, et al.
Published: (2025)