Language Alignment via Nash-learning and Adaptive feedback
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Azarafrooz, Ari, Faal, Farshid |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Nash CoT: Multi-Path Inference with Preference Equilibrium
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
Asymptotic Universal Alignment: A New Alignment Framework via Test-Time Scaling
von: Cai, Yang, et al.
Veröffentlicht: (2026)
von: Cai, Yang, et al.
Veröffentlicht: (2026)
Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms
von: Azarafrooz, Ari
Veröffentlicht: (2026)
von: Azarafrooz, Ari
Veröffentlicht: (2026)
On the Fundamental Impossibility of Hallucination Control in Large Language Models
von: Karpowicz, Michał P.
Veröffentlicht: (2025)
von: Karpowicz, Michał P.
Veröffentlicht: (2025)
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models
von: Goel, Naman
Veröffentlicht: (2023)
von: Goel, Naman
Veröffentlicht: (2023)
Large-Scale Auto-bidding with Nash Equilibrium Constraints
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
von: Mou, Zhiyu, et al.
Veröffentlicht: (2025)
Ad Auctions for LLMs via Retrieval Augmented Generation
von: Hajiaghayi, MohammadTaghi, et al.
Veröffentlicht: (2024)
von: Hajiaghayi, MohammadTaghi, et al.
Veröffentlicht: (2024)
Representative Social Choice: From Learning Theory to AI Alignment
von: Qiu, Tianyi
Veröffentlicht: (2024)
von: Qiu, Tianyi
Veröffentlicht: (2024)
Accelerating Nash Equilibrium Convergence in Monte Carlo Settings Through Counterfactual Value Based Fictitious Play
von: Qi, Ju, et al.
Veröffentlicht: (2023)
von: Qi, Ju, et al.
Veröffentlicht: (2023)
COMAL: A Convergent Meta-Algorithm for Aligning LLMs with General Preferences
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
Multi-Head Attention Is a Multi-Player Game
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2026)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2026)
The Limits of Preference Data for Post-Training
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
Truthfulness Despite Weak Supervision: Evaluating and Training LLMs Using Peer Prediction
von: Qiu, Tianyi Alex, et al.
Veröffentlicht: (2026)
von: Qiu, Tianyi Alex, et al.
Veröffentlicht: (2026)
PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations
von: Lei, Yingjie
Veröffentlicht: (2026)
von: Lei, Yingjie
Veröffentlicht: (2026)
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2026)
von: Miceli-Barone, Antonio Valerio, et al.
Veröffentlicht: (2026)
GLEE: A Unified Framework and Benchmark for Language-based Economic Environments
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
Clone-Robust AI Alignment
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
von: Procaccia, Ariel D., et al.
Veröffentlicht: (2025)
Nash Learning from Human Feedback
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
Axioms for AI Alignment from Human Feedback
von: Ge, Luise, et al.
Veröffentlicht: (2024)
von: Ge, Luise, et al.
Veröffentlicht: (2024)
ElicitationGPT: Text Elicitation Mechanisms via Language Models
von: Wu, Yifan, et al.
Veröffentlicht: (2024)
von: Wu, Yifan, et al.
Veröffentlicht: (2024)
Incentivizing Truthful Language Models via Peer Elicitation Games
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
Adaptive Contracts for Cost-Effective AI Delegation
von: Saig, Eden, et al.
Veröffentlicht: (2026)
von: Saig, Eden, et al.
Veröffentlicht: (2026)
Fusion-PSRO: Nash Policy Fusion for Policy Space Response Oracles
von: Lian, Jiesong, et al.
Veröffentlicht: (2024)
von: Lian, Jiesong, et al.
Veröffentlicht: (2024)
Offline Learning of Nash Stable Coalition Structures with Possibly Overlapping Coalitions
von: Cohen, Saar
Veröffentlicht: (2026)
von: Cohen, Saar
Veröffentlicht: (2026)
Predicting human decisions with behavioral theories and machine learning
von: Plonsky, Ori, et al.
Veröffentlicht: (2019)
von: Plonsky, Ori, et al.
Veröffentlicht: (2019)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
von: Zaman, Muhammad Aneeq uz, et al.
Veröffentlicht: (2024)
von: Zaman, Muhammad Aneeq uz, et al.
Veröffentlicht: (2024)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Transfer learning of state-based potential games for process optimization in decentralized manufacturing systems
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
von: Li, Zun, et al.
Veröffentlicht: (2023)
von: Li, Zun, et al.
Veröffentlicht: (2023)
The Battling Influencers Game: Nash Equilibria Structure of a Potential Game and Implications to Value Alignment
von: Wu, Young, et al.
Veröffentlicht: (2025)
von: Wu, Young, et al.
Veröffentlicht: (2025)
Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems
von: Chai, Rui
Veröffentlicht: (2026)
von: Chai, Rui
Veröffentlicht: (2026)
Nash Equilibria via Stochastic Eigendecomposition
von: Gemp, Ian
Veröffentlicht: (2024)
von: Gemp, Ian
Veröffentlicht: (2024)
In-Context Credit Assignment via the Core
von: Harris, Keegan, et al.
Veröffentlicht: (2026)
von: Harris, Keegan, et al.
Veröffentlicht: (2026)
SpinGPT: A Large-Language-Model Approach to Playing Poker Correctly
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
Incentivizing Quality Text Generation via Statistical Contracts
von: Saig, Eden, et al.
Veröffentlicht: (2024)
von: Saig, Eden, et al.
Veröffentlicht: (2024)
Ranking Abuse via Strategic Pairwise Data Perturbations
von: Yao, Junyi, et al.
Veröffentlicht: (2026)
von: Yao, Junyi, et al.
Veröffentlicht: (2026)
Human Choice Prediction in Language-based Persuasion Games: Simulation-based Off-Policy Evaluation
von: Shapira, Eilam, et al.
Veröffentlicht: (2023)
von: Shapira, Eilam, et al.
Veröffentlicht: (2023)
Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024) -
Nash CoT: Multi-Path Inference with Preference Equilibrium
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024) -
Asymptotic Universal Alignment: A New Alignment Framework via Test-Time Scaling
von: Cai, Yang, et al.
Veröffentlicht: (2026) -
Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms
von: Azarafrooz, Ari
Veröffentlicht: (2026) -
On the Fundamental Impossibility of Hallucination Control in Large Language Models
von: Karpowicz, Michał P.
Veröffentlicht: (2025)