Guardado en:
| Autores principales: | Shabadi, Guruprerana, Mallik, Kaushik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.02151 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Risk-Sensitive Agent Compositions
por: Shabadi, Guruprerana, et al.
Publicado: (2025)
por: Shabadi, Guruprerana, et al.
Publicado: (2025)
Programmatic Reinforcement Learning: Navigating Gridworlds
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
por: Shabadi, Guruprerana, et al.
Publicado: (2024)
Do We Need Frontier Models to Verify Mathematical Proofs?
por: Naik, Aaditya, et al.
Publicado: (2026)
por: Naik, Aaditya, et al.
Publicado: (2026)
Optimization Modulo Integer Linear-Exponential Programs
por: Hitarth, S, et al.
Publicado: (2025)
por: Hitarth, S, et al.
Publicado: (2025)
Auction-Based Scheduling
por: Avni, Guy, et al.
Publicado: (2023)
por: Avni, Guy, et al.
Publicado: (2023)
Learning in Budgeted Auctions with Spacing Objectives
por: Fikioris, Giannis, et al.
Publicado: (2024)
por: Fikioris, Giannis, et al.
Publicado: (2024)
Automated Deterministic Auction Design with Objective Decomposition
por: Duan, Zhijian, et al.
Publicado: (2024)
por: Duan, Zhijian, et al.
Publicado: (2024)
Evolving Diffusion and Flow Matching Policies for Online Reinforcement Learning
por: Zhang, Chubin, et al.
Publicado: (2025)
por: Zhang, Chubin, et al.
Publicado: (2025)
Glitches in Decision Tree Ensemble Models
por: Chandra, Satyankar, et al.
Publicado: (2025)
por: Chandra, Satyankar, et al.
Publicado: (2025)
Monitoring of Static Fairness
por: Henzinger, Thomas A., et al.
Publicado: (2025)
por: Henzinger, Thomas A., et al.
Publicado: (2025)
Online Adaptation for Enhancing Imitation Learning Policies
por: Malato, Federico, et al.
Publicado: (2024)
por: Malato, Federico, et al.
Publicado: (2024)
PolicyEvolve: Evolving Programmatic Policies by LLMs for multi-player games via Population-Based Training
por: Lv, Mingrui, et al.
Publicado: (2025)
por: Lv, Mingrui, et al.
Publicado: (2025)
Efficient Dynamic Shielding for Parametric Safety Specifications
por: Corsi, Davide, et al.
Publicado: (2025)
por: Corsi, Davide, et al.
Publicado: (2025)
Co-Evolving Policy Distillation
por: Gu, Naibin, et al.
Publicado: (2026)
por: Gu, Naibin, et al.
Publicado: (2026)
Online SLA Decomposition: Enabling Real-Time Adaptation to Evolving Network Systems
por: Hsu, Cyril Shih-Huan, et al.
Publicado: (2024)
por: Hsu, Cyril Shih-Huan, et al.
Publicado: (2024)
Breaking Determinism: Stochastic Modeling for Reliable Off-Policy Evaluation in Ad Auctions
por: Yeom, Hongseon, et al.
Publicado: (2025)
por: Yeom, Hongseon, et al.
Publicado: (2025)
Online Combinatorial Allocations and Auctions with Few Samples
por: Dütting, Paul, et al.
Publicado: (2024)
por: Dütting, Paul, et al.
Publicado: (2024)
Effective Policy Learning for Multi-Agent Online Coordination Beyond Submodular Objectives
por: Zhang, Qixin, et al.
Publicado: (2025)
por: Zhang, Qixin, et al.
Publicado: (2025)
OASIS: Online Activation Subspace Learning for Memory-Efficient Training
por: Choudhary, Sakshi, et al.
Publicado: (2026)
por: Choudhary, Sakshi, et al.
Publicado: (2026)
Off-Policy Evaluation and Counterfactual Methods in Dynamic Auction Environments
por: Guha, Ritam, et al.
Publicado: (2025)
por: Guha, Ritam, et al.
Publicado: (2025)
Channel Estimation by Infinite Width Convolutional Networks
por: Mallik, Mohammed, et al.
Publicado: (2025)
por: Mallik, Mohammed, et al.
Publicado: (2025)
Online Causal Inference for Advertising in Real-Time Bidding Auctions
por: Waisman, Caio, et al.
Publicado: (2019)
por: Waisman, Caio, et al.
Publicado: (2019)
Improved Online Learning Algorithms for CTR Prediction in Ad Auctions
por: Feng, Zhe, et al.
Publicado: (2024)
por: Feng, Zhe, et al.
Publicado: (2024)
Optimizing Online Advertising with Multi-Armed Bandits: Mitigating the Cold Start Problem under Auction Dynamics
por: Soboleva, Anastasiia, et al.
Publicado: (2025)
por: Soboleva, Anastasiia, et al.
Publicado: (2025)
Evolving Restricted Boltzmann Machine-Kohonen Network for Online Clustering
por: Senthilnath, J., et al.
Publicado: (2024)
por: Senthilnath, J., et al.
Publicado: (2024)
Flow-Based Policy for Online Reinforcement Learning
por: Lv, Lei, et al.
Publicado: (2025)
por: Lv, Lei, et al.
Publicado: (2025)
MANGO: Meta-Adaptive Network Gradient Optimization for Online Continual Learning
por: Awasthi, Ankita, et al.
Publicado: (2026)
por: Awasthi, Ankita, et al.
Publicado: (2026)
ELENA: Epigenetic Learning through Evolved Neural Adaptation
por: Kriuk, Boris, et al.
Publicado: (2025)
por: Kriuk, Boris, et al.
Publicado: (2025)
Multi-Objective $\textit{min-max}$ Online Convex Optimization
por: Vaze, Rahul, et al.
Publicado: (2025)
por: Vaze, Rahul, et al.
Publicado: (2025)
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
por: Cho, Brian, et al.
Publicado: (2025)
por: Cho, Brian, et al.
Publicado: (2025)
Causal and Federated Multimodal Learning for Cardiovascular Risk Prediction under Heterogeneous Populations
por: Kaushik, Rohit, et al.
Publicado: (2026)
por: Kaushik, Rohit, et al.
Publicado: (2026)
Learning Control Policies for Variable Objectives from Offline Data
por: Weber, Marc, et al.
Publicado: (2023)
por: Weber, Marc, et al.
Publicado: (2023)
Online Auction Design Using Distribution-Free Uncertainty Quantification with Applications to E-Commerce
por: Han, Jiale, et al.
Publicado: (2024)
por: Han, Jiale, et al.
Publicado: (2024)
Information-Consistent Language Model Recommendations through Group Relative Policy Optimization
por: Prabhune, Sonal, et al.
Publicado: (2025)
por: Prabhune, Sonal, et al.
Publicado: (2025)
Evolved Sample Weights for Bias Mitigation: Effectiveness Depends on the Fairness Objective
por: Saini, Anil K., et al.
Publicado: (2025)
por: Saini, Anil K., et al.
Publicado: (2025)
Online Feature Updates Improve Online (Generalized) Label Shift Adaptation
por: Wu, Ruihan, et al.
Publicado: (2024)
por: Wu, Ruihan, et al.
Publicado: (2024)
An Offline Adaptation Framework for Constrained Multi-Objective Reinforcement Learning
por: Lin, Qian, et al.
Publicado: (2024)
por: Lin, Qian, et al.
Publicado: (2024)
One-Way Policy Optimization for Self-Evolving LLMs
por: Yang, Shuo, et al.
Publicado: (2026)
por: Yang, Shuo, et al.
Publicado: (2026)
Robust Multi-Objective Preference Alignment with Online DPO
por: Gupta, Raghav, et al.
Publicado: (2025)
por: Gupta, Raghav, et al.
Publicado: (2025)
InvEvolve: Evolving White-Box Inventory Policies via Large Language Models with Performance Guarantees
por: Huang, Chenyu, et al.
Publicado: (2026)
por: Huang, Chenyu, et al.
Publicado: (2026)
Ejemplares similares
-
Risk-Sensitive Agent Compositions
por: Shabadi, Guruprerana, et al.
Publicado: (2025) -
Programmatic Reinforcement Learning: Navigating Gridworlds
por: Shabadi, Guruprerana, et al.
Publicado: (2024) -
Do We Need Frontier Models to Verify Mathematical Proofs?
por: Naik, Aaditya, et al.
Publicado: (2026) -
Optimization Modulo Integer Linear-Exponential Programs
por: Hitarth, S, et al.
Publicado: (2025) -
Auction-Based Scheduling
por: Avni, Guy, et al.
Publicado: (2023)