Robust Optimization with Diffusion Models for Green Security
Fuente:
arXiv
Guardado en:
| Autores principales: | Kong, Lingkai, Wang, Haichuan, Pan, Yuqi, Kim, Cheol Woo, Song, Mingxiao, Nguyen, Alayna, Wang, Tonghan, Xu, Haifeng, Tambe, Milind |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Generative AI for Social Impact
por: Kong, Lingkai, et al.
Publicado: (2026)
por: Kong, Lingkai, et al.
Publicado: (2026)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
por: Kong, Lingkai, et al.
Publicado: (2025)
por: Kong, Lingkai, et al.
Publicado: (2025)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
por: Martinson, Sarah, et al.
Publicado: (2025)
por: Martinson, Sarah, et al.
Publicado: (2025)
Lightweight Robust Direct Preference Optimization
por: Kim, Cheol Woo, et al.
Publicado: (2025)
por: Kim, Cheol Woo, et al.
Publicado: (2025)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
por: Wang, Haichuan, et al.
Publicado: (2026)
por: Wang, Haichuan, et al.
Publicado: (2026)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
por: Kong, Lingkai, et al.
Publicado: (2025)
por: Kong, Lingkai, et al.
Publicado: (2025)
Preference Robustness for DPO with Applications to Public Health
por: Kim, Cheol Woo, et al.
Publicado: (2025)
por: Kim, Cheol Woo, et al.
Publicado: (2025)
Navigating the Social Welfare Frontier: Portfolios for Multi-objective Reinforcement Learning
por: Kim, Cheol Woo, et al.
Publicado: (2025)
por: Kim, Cheol Woo, et al.
Publicado: (2025)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
por: Verma, Shresth, et al.
Publicado: (2024)
por: Verma, Shresth, et al.
Publicado: (2024)
Incentive-Aware AI Safety via Strategic Resource Allocation: A Stackelberg Security Games Perspective
por: Kim, Cheol Woo, et al.
Publicado: (2026)
por: Kim, Cheol Woo, et al.
Publicado: (2026)
LLM Active Alignment: A Nash Equilibrium Perspective
por: Wang, Tonghan, et al.
Publicado: (2026)
por: Wang, Tonghan, et al.
Publicado: (2026)
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing
por: Choo, Davin, et al.
Publicado: (2025)
por: Choo, Davin, et al.
Publicado: (2025)
On Diffusion Models for Multi-Agent Partial Observability: Shared Attractors, Error Bounds, and Composite Flow
por: Wang, Tonghan, et al.
Publicado: (2024)
por: Wang, Tonghan, et al.
Publicado: (2024)
What is the Right Notion of Distance between Predict-then-Optimize Tasks?
por: Rodriguez-Diaz, Paula, et al.
Publicado: (2024)
por: Rodriguez-Diaz, Paula, et al.
Publicado: (2024)
Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions
por: Kong, Lingkai, et al.
Publicado: (2026)
por: Kong, Lingkai, et al.
Publicado: (2026)
Adaptive Multi-Round Allocation with Stochastic Arrivals
por: Pan, Yuqi, et al.
Publicado: (2026)
por: Pan, Yuqi, et al.
Publicado: (2026)
Escape Sensing Games: Detection-vs-Evasion in Security Applications
por: Boehmer, Niclas, et al.
Publicado: (2024)
por: Boehmer, Niclas, et al.
Publicado: (2024)
Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order Information
por: Ai, Rui, et al.
Publicado: (2025)
por: Ai, Rui, et al.
Publicado: (2025)
Rule-Bottleneck Reinforcement Learning: Joint Explanation and Decision Optimization for Resource Allocation with Language Agents
por: Tec, Mauricio, et al.
Publicado: (2025)
por: Tec, Mauricio, et al.
Publicado: (2025)
The Bandit Whisperer: Communication Learning for Restless Bandits
por: Zhao, Yunfan, et al.
Publicado: (2024)
por: Zhao, Yunfan, et al.
Publicado: (2024)
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
por: Xiong, Guojun, et al.
Publicado: (2025)
por: Xiong, Guojun, et al.
Publicado: (2025)
Policy-Embedded Graph Expansion: Networked HIV Testing with Diffusion-Driven Network Samples
por: Kangaslahti, Akseli, et al.
Publicado: (2026)
por: Kangaslahti, Akseli, et al.
Publicado: (2026)
Dual-Mandate Patrols: Multi-Armed Bandits for Green Security
por: Xu, Lily, et al.
Publicado: (2020)
por: Xu, Lily, et al.
Publicado: (2020)
Learning to Persuade a Biased Receiver
por: Pan, Yuqi, et al.
Publicado: (2026)
por: Pan, Yuqi, et al.
Publicado: (2026)
Many Preferences, Few Policies: Towards Scalable Language Model Personalization
por: Kim, Cheol Woo, et al.
Publicado: (2026)
por: Kim, Cheol Woo, et al.
Publicado: (2026)
How LLMs Are Persuaded: A Few Attention Heads, Rerouted
por: Sun, Xiangkun, et al.
Publicado: (2026)
por: Sun, Xiangkun, et al.
Publicado: (2026)
Diffusion-DFL: Decision-focused Diffusion Models for Stochastic Optimization
por: Zhao, Zihao, et al.
Publicado: (2025)
por: Zhao, Zihao, et al.
Publicado: (2025)
Network-Based Interventions for HIV Prevention via Cascade-Aware Suppression of Transmission
por: Kangaslahti, Akseli, et al.
Publicado: (2026)
por: Kangaslahti, Akseli, et al.
Publicado: (2026)
Preliminary Study of the Impact of AI-Based Interventions on Health and Behavioral Outcomes in Maternal Health Programs
por: Dasgupta, Arpan, et al.
Publicado: (2024)
por: Dasgupta, Arpan, et al.
Publicado: (2024)
NaiAD: Initiate Data-Driven Research for LLM Advertising
por: Zhang, Yihang, et al.
Publicado: (2026)
por: Zhang, Yihang, et al.
Publicado: (2026)
Auditing Work: Exploring the New York City algorithmic bias audit regime
por: Groves, Lara, et al.
Publicado: (2024)
por: Groves, Lara, et al.
Publicado: (2024)
LLM Advertisement based on Neuron Auctions
por: Yun, Peiran, et al.
Publicado: (2026)
por: Yun, Peiran, et al.
Publicado: (2026)
On Sequential Fault-Intolerant Process Planning
por: Kaczmarczyk, Andrzej, et al.
Publicado: (2025)
por: Kaczmarczyk, Andrzej, et al.
Publicado: (2025)
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
por: Xiong, Guojun, et al.
Publicado: (2025)
por: Xiong, Guojun, et al.
Publicado: (2025)
A Machine Learning Approach to Two-Stage Adaptive Robust Optimization
por: Bertsimas, Dimitris, et al.
Publicado: (2023)
por: Bertsimas, Dimitris, et al.
Publicado: (2023)
Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub
por: Hu, Haichuan, et al.
Publicado: (2026)
por: Hu, Haichuan, et al.
Publicado: (2026)
Quantitative Convergences of Lie Group Momentum Optimizers
por: Kong, Lingkai, et al.
Publicado: (2024)
por: Kong, Lingkai, et al.
Publicado: (2024)
Translate Policy to Language: Flow Matching Generated Rewards for LLM Explanations
por: Yang, Xinyi, et al.
Publicado: (2025)
por: Yang, Xinyi, et al.
Publicado: (2025)
The Publication Choice Problem
por: Wang, Haichuan, et al.
Publicado: (2025)
por: Wang, Haichuan, et al.
Publicado: (2025)
BundleFlow: Deep Menus for Combinatorial Auctions by Diffusion-Based Optimization
por: Wang, Tonghan, et al.
Publicado: (2025)
por: Wang, Tonghan, et al.
Publicado: (2025)
Ejemplares similares
-
Generative AI for Social Impact
por: Kong, Lingkai, et al.
Publicado: (2026) -
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
por: Kong, Lingkai, et al.
Publicado: (2025) -
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
por: Martinson, Sarah, et al.
Publicado: (2025) -
Lightweight Robust Direct Preference Optimization
por: Kim, Cheol Woo, et al.
Publicado: (2025) -
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
por: Wang, Haichuan, et al.
Publicado: (2026)