LLM Active Alignment: A Nash Equilibrium Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Tonghan, Pan, Yuqi, Yang, Xinyi, Jiang, Yanchen, Tambe, Milind, Parkes, David C. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GemNet: Menu-Based, Strategy-Proof Multi-Bidder Auctions Through Deep Learning
by: Wang, Tonghan, et al.
Published: (2024)
by: Wang, Tonghan, et al.
Published: (2024)
On Diffusion Models for Multi-Agent Partial Observability: Shared Attractors, Error Bounds, and Composite Flow
by: Wang, Tonghan, et al.
Published: (2024)
by: Wang, Tonghan, et al.
Published: (2024)
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing
by: Choo, Davin, et al.
Published: (2025)
by: Choo, Davin, et al.
Published: (2025)
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
by: Xiong, Guojun, et al.
Published: (2025)
by: Xiong, Guojun, et al.
Published: (2025)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
by: Wang, Haichuan, et al.
Published: (2026)
by: Wang, Haichuan, et al.
Published: (2026)
Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order Information
by: Ai, Rui, et al.
Published: (2025)
by: Ai, Rui, et al.
Published: (2025)
Robust Optimization with Diffusion Models for Green Security
by: Kong, Lingkai, et al.
Published: (2025)
by: Kong, Lingkai, et al.
Published: (2025)
Social Environment Design
by: Zhang, Edwin, et al.
Published: (2024)
by: Zhang, Edwin, et al.
Published: (2024)
Multi-Sender Persuasion: A Computational Perspective
by: Hossain, Safwan, et al.
Published: (2024)
by: Hossain, Safwan, et al.
Published: (2024)
BundleFlow: Deep Menus for Combinatorial Auctions by Diffusion-Based Optimization
by: Wang, Tonghan, et al.
Published: (2025)
by: Wang, Tonghan, et al.
Published: (2025)
Adaptive Multi-Round Allocation with Stochastic Arrivals
by: Pan, Yuqi, et al.
Published: (2026)
by: Pan, Yuqi, et al.
Published: (2026)
Incentive-Aware AI Safety via Strategic Resource Allocation: A Stackelberg Security Games Perspective
by: Kim, Cheol Woo, et al.
Published: (2026)
by: Kim, Cheol Woo, et al.
Published: (2026)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024)
by: Verma, Shresth, et al.
Published: (2024)
Embeddings for Preferences, Not Semantics
by: Blair, Carter, et al.
Published: (2026)
by: Blair, Carter, et al.
Published: (2026)
LLM-Powered Preference Elicitation in Combinatorial Assignment
by: Soumalias, Ermis, et al.
Published: (2025)
by: Soumalias, Ermis, et al.
Published: (2025)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
by: Martinson, Sarah, et al.
Published: (2025)
by: Martinson, Sarah, et al.
Published: (2025)
Towards Foundation-model-based Multiagent System to Accelerate AI for Social Impact
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Combining Diverse Information for Coordinated Action: Stochastic Bandit Algorithms for Heterogeneous Agents
by: Gordon, Lucia, et al.
Published: (2024)
by: Gordon, Lucia, et al.
Published: (2024)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
by: Shah, Sanket, et al.
Published: (2023)
by: Shah, Sanket, et al.
Published: (2023)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
by: Shah, Sanket, et al.
Published: (2024)
by: Shah, Sanket, et al.
Published: (2024)
Health Facility Location in Ethiopia: Leveraging LLMs to Integrate Expert Knowledge into Algorithmic Planning
by: Trabelsi, Yohai, et al.
Published: (2026)
by: Trabelsi, Yohai, et al.
Published: (2026)
On Sequential Fault-Intolerant Process Planning
by: Kaczmarczyk, Andrzej, et al.
Published: (2025)
by: Kaczmarczyk, Andrzej, et al.
Published: (2025)
Rule-Bottleneck Reinforcement Learning: Joint Explanation and Decision Optimization for Resource Allocation with Language Agents
by: Tec, Mauricio, et al.
Published: (2025)
by: Tec, Mauricio, et al.
Published: (2025)
What is the Right Notion of Distance between Predict-then-Optimize Tasks?
by: Rodriguez-Diaz, Paula, et al.
Published: (2024)
by: Rodriguez-Diaz, Paula, et al.
Published: (2024)
Reinforcement learning with combinatorial actions for coupled restless bandits
by: Xu, Lily, et al.
Published: (2025)
by: Xu, Lily, et al.
Published: (2025)
Contrasting local and global modeling with machine learning and satellite data: A case study estimating tree canopy height in African savannas
by: Rolf, Esther, et al.
Published: (2024)
by: Rolf, Esther, et al.
Published: (2024)
Online Allocation with Unknown Shared Supply
by: Neoh, Tzeh Yuan, et al.
Published: (2026)
by: Neoh, Tzeh Yuan, et al.
Published: (2026)
Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health
by: Verma, Shresth, et al.
Published: (2026)
by: Verma, Shresth, et al.
Published: (2026)
Beyond Listenership: AI-Predicted Interventions Drive Improvements in Maternal Health Behaviours
by: Dasgupta, Arpan, et al.
Published: (2025)
by: Dasgupta, Arpan, et al.
Published: (2025)
Large-Scale Auto-bidding with Nash Equilibrium Constraints
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
by: Jain, Gauri, et al.
Published: (2024)
by: Jain, Gauri, et al.
Published: (2024)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
Incentive-Aligned Multi-Source LLM Summaries
by: Jiang, Yanchen, et al.
Published: (2025)
by: Jiang, Yanchen, et al.
Published: (2025)
NaiAD: Initiate Data-Driven Research for LLM Advertising
by: Zhang, Yihang, et al.
Published: (2026)
by: Zhang, Yihang, et al.
Published: (2026)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
by: Behari, Nikhil, et al.
Published: (2024)
by: Behari, Nikhil, et al.
Published: (2024)
Nash CoT: Multi-Path Inference with Preference Equilibrium
by: Zhang, Ziqi, et al.
Published: (2024)
by: Zhang, Ziqi, et al.
Published: (2024)
Learning from Synthetic Labs: Language Models as Auction Participants
by: Shah, Anand, et al.
Published: (2025)
by: Shah, Anand, et al.
Published: (2025)
Evaluating the Effectiveness of Index-Based Treatment Allocation
by: Boehmer, Niclas, et al.
Published: (2024)
by: Boehmer, Niclas, et al.
Published: (2024)
IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment
by: Ming, Chenlin, et al.
Published: (2025)
by: Ming, Chenlin, et al.
Published: (2025)
Similar Items
-
GemNet: Menu-Based, Strategy-Proof Multi-Bidder Auctions Through Deep Learning
by: Wang, Tonghan, et al.
Published: (2024) -
On Diffusion Models for Multi-Agent Partial Observability: Shared Attractors, Error Bounds, and Composite Flow
by: Wang, Tonghan, et al.
Published: (2024) -
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing
by: Choo, Davin, et al.
Published: (2025) -
Composite Flow Matching for Reinforcement Learning with Shifted-Dynamics Data
by: Kong, Lingkai, et al.
Published: (2025) -
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
by: Xiong, Guojun, et al.
Published: (2025)