Improving Generative Ad Text on Facebook using Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Jiang, Daniel R., Nikulkov, Alex, Chen, Yu-Chia, Bai, Yang, Zhu, Zheqing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling
por: Nikulkov, Alex
Publicado: (2026)
por: Nikulkov, Alex
Publicado: (2026)
Pearl: A Production-ready Reinforcement Learning Agent
por: Zhu, Zheqing, et al.
Publicado: (2023)
por: Zhu, Zheqing, et al.
Publicado: (2023)
Uncertainty of Joint Neural Contextual Bandit
por: Guo, Hongbo, et al.
Publicado: (2024)
por: Guo, Hongbo, et al.
Publicado: (2024)
Exploiting Structure in Offline Multi-Agent RL: The Benefits of Low Interaction Rank
por: Zhan, Wenhao, et al.
Publicado: (2024)
por: Zhan, Wenhao, et al.
Publicado: (2024)
Session-Level Dynamic Ad Load Optimization using Offline Robust Reinforcement Learning
por: Liu, Tao, et al.
Publicado: (2025)
por: Liu, Tao, et al.
Publicado: (2025)
Aligned Multi Objective Optimization
por: Efroni, Yonathan, et al.
Publicado: (2025)
por: Efroni, Yonathan, et al.
Publicado: (2025)
Seldonian Reinforcement Learning for Ad Hoc Teamwork
por: Zorzi, Edoardo, et al.
Publicado: (2025)
por: Zorzi, Edoardo, et al.
Publicado: (2025)
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence
por: Zhu, Lingwei, et al.
Publicado: (2023)
por: Zhu, Lingwei, et al.
Publicado: (2023)
Optimizing Search Advertising Strategies: Integrating Reinforcement Learning with Generalized Second-Price Auctions for Enhanced Ad Ranking and Bidding
por: Zhou, Chang, et al.
Publicado: (2024)
por: Zhou, Chang, et al.
Publicado: (2024)
Deep Reinforcement Learning for Ranking Utility Tuning in the Ad Recommender System at Pinterest
por: Yang, Xiao, et al.
Publicado: (2025)
por: Yang, Xiao, et al.
Publicado: (2025)
Towards Robust Offline-to-Online Reinforcement Learning via Uncertainty and Smoothness
por: Wen, Xiaoyu, et al.
Publicado: (2023)
por: Wen, Xiaoyu, et al.
Publicado: (2023)
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation
por: Zhao, Runze, et al.
Publicado: (2025)
por: Zhao, Runze, et al.
Publicado: (2025)
IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control
por: Chitnis, Rohan, et al.
Publicado: (2023)
por: Chitnis, Rohan, et al.
Publicado: (2023)
Improving the Effectiveness of Potential-Based Reward Shaping in Reinforcement Learning
por: Müller, Henrik, et al.
Publicado: (2025)
por: Müller, Henrik, et al.
Publicado: (2025)
Learning Personalized Ad Impact via Contextual Reinforcement Learning under Delayed Rewards
por: Cheng, Yuwei, et al.
Publicado: (2025)
por: Cheng, Yuwei, et al.
Publicado: (2025)
Is Inverse Reinforcement Learning Harder than Standard Reinforcement Learning? A Theoretical Perspective
por: Zhao, Lei, et al.
Publicado: (2023)
por: Zhao, Lei, et al.
Publicado: (2023)
Improving Reinforcement Learning Sample-Efficiency using Local Approximation
por: Prashant, Mohit, et al.
Publicado: (2025)
por: Prashant, Mohit, et al.
Publicado: (2025)
From Reward-Free Representations to Preferences: Rethinking Offline Preference-Based Reinforcement Learning
por: Yang, Jun-Jie, et al.
Publicado: (2026)
por: Yang, Jun-Jie, et al.
Publicado: (2026)
Any-step Dynamics Model Improves Future Predictions for Online and Offline Reinforcement Learning
por: Lin, Haoxin, et al.
Publicado: (2024)
por: Lin, Haoxin, et al.
Publicado: (2024)
LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning
por: Wu, Yuhao, et al.
Publicado: (2025)
por: Wu, Yuhao, et al.
Publicado: (2025)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
por: Wang, Changhong, et al.
Publicado: (2024)
por: Wang, Changhong, et al.
Publicado: (2024)
Improved Off-policy Reinforcement Learning in Biological Sequence Design
por: Kim, Hyeonah, et al.
Publicado: (2024)
por: Kim, Hyeonah, et al.
Publicado: (2024)
Expanding the Capabilities of Reinforcement Learning via Text Feedback
por: Song, Yuda, et al.
Publicado: (2026)
por: Song, Yuda, et al.
Publicado: (2026)
Instance Selection for Dynamic Algorithm Configuration with Reinforcement Learning: Improving Generalization
por: Benjamins, Carolin, et al.
Publicado: (2024)
por: Benjamins, Carolin, et al.
Publicado: (2024)
Improving the Data-efficiency of Reinforcement Learning by Warm-starting with LLM
por: Duong, Thang, et al.
Publicado: (2025)
por: Duong, Thang, et al.
Publicado: (2025)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
por: Liu, Xu-Hui, et al.
Publicado: (2024)
por: Liu, Xu-Hui, et al.
Publicado: (2024)
DRIVE: Data Curation Best Practices for Reinforcement Learning with Verifiable Reward in Competitive Code Generation
por: Zhu, Speed, et al.
Publicado: (2025)
por: Zhu, Speed, et al.
Publicado: (2025)
Improving Sample Efficiency of Reinforcement Learning with Background Knowledge from Large Language Models
por: Zhang, Fuxiang, et al.
Publicado: (2024)
por: Zhang, Fuxiang, et al.
Publicado: (2024)
Provable Risk-Sensitive Distributional Reinforcement Learning with General Function Approximation
por: Chen, Yu, et al.
Publicado: (2024)
por: Chen, Yu, et al.
Publicado: (2024)
On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification
por: Wu, Yongliang, et al.
Publicado: (2025)
por: Wu, Yongliang, et al.
Publicado: (2025)
Data Acquisition for Improving Model Fairness using Reinforcement Learning
por: Hasan, Jahid, et al.
Publicado: (2024)
por: Hasan, Jahid, et al.
Publicado: (2024)
Class-Balanced and Reinforced Active Learning on Graphs
por: Yu, Chengcheng, et al.
Publicado: (2024)
por: Yu, Chengcheng, et al.
Publicado: (2024)
Adaptive Testing Environment Generation for Connected and Automated Vehicles with Dense Reinforcement Learning
por: Yang, Jingxuan, et al.
Publicado: (2024)
por: Yang, Jingxuan, et al.
Publicado: (2024)
Generalization in Reinforcement Learning for Radio Access Networks
por: Demirel, Burak, et al.
Publicado: (2025)
por: Demirel, Burak, et al.
Publicado: (2025)
Experiential Reinforcement Learning
por: Shi, Taiwei, et al.
Publicado: (2026)
por: Shi, Taiwei, et al.
Publicado: (2026)
The Generalization Gap in Offline Reinforcement Learning
por: Mediratta, Ishita, et al.
Publicado: (2023)
por: Mediratta, Ishita, et al.
Publicado: (2023)
Residuals-based Offline Reinforcement Learning
por: Zhu, Qing, et al.
Publicado: (2026)
por: Zhu, Qing, et al.
Publicado: (2026)
Contrastive UCB: Provably Efficient Contrastive Self-Supervised Learning in Online Reinforcement Learning
por: Qiu, Shuang, et al.
Publicado: (2022)
por: Qiu, Shuang, et al.
Publicado: (2022)
Conditional Sequence Modeling for Safe Reinforcement Learning
por: Bai, Wensong, et al.
Publicado: (2026)
por: Bai, Wensong, et al.
Publicado: (2026)
Do Agents Dream of Electric Sheep?: Improving Generalization in Reinforcement Learning through Generative Learning
por: Franceschelli, Giorgio, et al.
Publicado: (2024)
por: Franceschelli, Giorgio, et al.
Publicado: (2024)
Ejemplares similares
-
Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling
por: Nikulkov, Alex
Publicado: (2026) -
Pearl: A Production-ready Reinforcement Learning Agent
por: Zhu, Zheqing, et al.
Publicado: (2023) -
Uncertainty of Joint Neural Contextual Bandit
por: Guo, Hongbo, et al.
Publicado: (2024) -
Exploiting Structure in Offline Multi-Agent RL: The Benefits of Low Interaction Rank
por: Zhan, Wenhao, et al.
Publicado: (2024) -
Session-Level Dynamic Ad Load Optimization using Offline Robust Reinforcement Learning
por: Liu, Tao, et al.
Publicado: (2025)