Asymptotic Universal Alignment: A New Alignment Framework via Test-Time Scaling
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Yang, Zheng, Weiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
COMAL: A Convergent Meta-Algorithm for Aligning LLMs with General Preferences
by: Liu, Yixin, et al.
Published: (2024)
by: Liu, Yixin, et al.
Published: (2024)
Language Alignment via Nash-learning and Adaptive feedback
by: Azarafrooz, Ari, et al.
Published: (2024)
by: Azarafrooz, Ari, et al.
Published: (2024)
Representative Social Choice: From Learning Theory to AI Alignment
by: Qiu, Tianyi
Published: (2024)
by: Qiu, Tianyi
Published: (2024)
Clone-Robust AI Alignment
by: Procaccia, Ariel D., et al.
Published: (2025)
by: Procaccia, Ariel D., et al.
Published: (2025)
Axioms for AI Alignment from Human Feedback
by: Ge, Luise, et al.
Published: (2024)
by: Ge, Luise, et al.
Published: (2024)
Ad Auctions for LLMs via Retrieval Augmented Generation
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
by: Zhang, Yuheng, et al.
Published: (2024)
by: Zhang, Yuheng, et al.
Published: (2024)
Socially-Weighted Alignment: A Game-Theoretic Framework for Multi-Agent LLM Systems
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems
by: Chai, Rui
Published: (2026)
by: Chai, Rui
Published: (2026)
GLEE: A Unified Framework and Benchmark for Language-based Economic Environments
by: Shapira, Eilam, et al.
Published: (2024)
by: Shapira, Eilam, et al.
Published: (2024)
Test-Time Compute Games
by: Velasco, Ander Artola, et al.
Published: (2026)
by: Velasco, Ander Artola, et al.
Published: (2026)
Large-Scale Auto-bidding with Nash Equilibrium Constraints
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
Multi-Head Attention Is a Multi-Player Game
by: Chakrabarti, Kushal, et al.
Published: (2026)
by: Chakrabarti, Kushal, et al.
Published: (2026)
Truthfulness Despite Weak Supervision: Evaluating and Training LLMs Using Peer Prediction
by: Qiu, Tianyi Alex, et al.
Published: (2026)
by: Qiu, Tianyi Alex, et al.
Published: (2026)
PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations
by: Lei, Yingjie
Published: (2026)
by: Lei, Yingjie
Published: (2026)
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2026)
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2026)
The Limits of Preference Data for Post-Training
by: Zhao, Eric, et al.
Published: (2025)
by: Zhao, Eric, et al.
Published: (2025)
On the Fundamental Impossibility of Hallucination Control in Large Language Models
by: Karpowicz, Michał P.
Published: (2025)
by: Karpowicz, Michał P.
Published: (2025)
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models
by: Goel, Naman
Published: (2023)
by: Goel, Naman
Published: (2023)
Nash CoT: Multi-Path Inference with Preference Equilibrium
by: Zhang, Ziqi, et al.
Published: (2024)
by: Zhang, Ziqi, et al.
Published: (2024)
Proximal Regret and Proximal Correlated Equilibria: A New Tractable Solution Concept for Online Learning and Games
by: Cai, Yang, et al.
Published: (2025)
by: Cai, Yang, et al.
Published: (2025)
CERN for AI: A Theoretical Framework for Autonomous Simulation-Based Artificial Intelligence Testing and Alignment
by: Bojic, Ljubisa, et al.
Published: (2023)
by: Bojic, Ljubisa, et al.
Published: (2023)
Ranking Abuse via Strategic Pairwise Data Perturbations
by: Yao, Junyi, et al.
Published: (2026)
by: Yao, Junyi, et al.
Published: (2026)
Is Online Linear Optimization Sufficient for Strategic Robustness?
by: Cai, Yang, et al.
Published: (2026)
by: Cai, Yang, et al.
Published: (2026)
Distributive Fairness in Large Language Models: Evaluating Alignment with Human Values
by: Hosseini, Hadi, et al.
Published: (2025)
by: Hosseini, Hadi, et al.
Published: (2025)
Mechanism-Based Intelligence (MBI): Differentiable Incentives for Rational Coordination and Guaranteed Alignment in Multi-Agent Systems
by: Grassi, Stefano
Published: (2025)
by: Grassi, Stefano
Published: (2025)
Online Test Synthesis From Requirements: Enhancing Reinforcement Learning with Game Theory
by: Sankur, Ocan, et al.
Published: (2024)
by: Sankur, Ocan, et al.
Published: (2024)
A Framework for Adversarial Analysis of Decision Support Systems Prior to Deployment
by: Bissey, Brett, et al.
Published: (2025)
by: Bissey, Brett, et al.
Published: (2025)
A New Lower Bound for the Random Offerer Mechanism in Bilateral Trade using AI-Guided Evolutionary Search
by: Cai, Yang, et al.
Published: (2026)
by: Cai, Yang, et al.
Published: (2026)
Intrinsic Barriers and Practical Pathways for Human-AI Alignment: An Agreement-Based Complexity Analysis
by: Nayebi, Aran
Published: (2025)
by: Nayebi, Aran
Published: (2025)
Meta-Computing Enhanced Federated Learning in IIoT: Satisfaction-Aware Incentive Scheme via DRL-Based Stackelberg Game
by: Li, Xiaohuan, et al.
Published: (2025)
by: Li, Xiaohuan, et al.
Published: (2025)
Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
Real-Time Parallel Counterfactual Regret Minimization
by: Li, Boning, et al.
Published: (2026)
by: Li, Boning, et al.
Published: (2026)
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
On Tractable $Φ$-Equilibria in Non-Concave Games
by: Cai, Yang, et al.
Published: (2024)
by: Cai, Yang, et al.
Published: (2024)
GTAlign: Game-Theoretic Alignment of LLM Assistants for Social Welfare
by: Zhu, Siqi, et al.
Published: (2025)
by: Zhu, Siqi, et al.
Published: (2025)
The Battling Influencers Game: Nash Equilibria Structure of a Potential Game and Implications to Value Alignment
by: Wu, Young, et al.
Published: (2025)
by: Wu, Young, et al.
Published: (2025)
On Corrigibility and Alignment in Multi Agent Games
by: Dable-Heath, Edmund, et al.
Published: (2025)
by: Dable-Heath, Edmund, et al.
Published: (2025)
GroupSegment-SHAP: Shapley Value Explanations with Group-Segment Players for Multivariate Time Series
by: Kim, Jinwoong, et al.
Published: (2026)
by: Kim, Jinwoong, et al.
Published: (2026)
In-Context Credit Assignment via the Core
by: Harris, Keegan, et al.
Published: (2026)
by: Harris, Keegan, et al.
Published: (2026)
Similar Items
-
COMAL: A Convergent Meta-Algorithm for Aligning LLMs with General Preferences
by: Liu, Yixin, et al.
Published: (2024) -
Language Alignment via Nash-learning and Adaptive feedback
by: Azarafrooz, Ari, et al.
Published: (2024) -
Representative Social Choice: From Learning Theory to AI Alignment
by: Qiu, Tianyi
Published: (2024) -
Clone-Robust AI Alignment
by: Procaccia, Ariel D., et al.
Published: (2025) -
Axioms for AI Alignment from Human Feedback
by: Ge, Luise, et al.
Published: (2024)