Next-Token Prediction and Regret Minimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Mohri, Mehryar, Sanford, Clayton, Schneider, Jon, Vodrahalli, Kiran, Wu, Yifan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Distributional Alignment Games for Answer-Level Fine-Tuning
por: Mohri, Mehryar, et al.
Publicado: (2026)
por: Mohri, Mehryar, et al.
Publicado: (2026)
Online Learning with Bounded Recall
por: Schneider, Jon, et al.
Publicado: (2022)
por: Schneider, Jon, et al.
Publicado: (2022)
High-Dimensional Calibration from Swap Regret
por: Fishelson, Maxwell, et al.
Publicado: (2025)
por: Fishelson, Maxwell, et al.
Publicado: (2025)
Swap Regret and Correlated Equilibria Beyond Normal-Form Games
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2025)
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2025)
Deep (Predictive) Discounted Counterfactual Regret Minimization
por: Xu, Hang, et al.
Publicado: (2025)
por: Xu, Hang, et al.
Publicado: (2025)
Efficient Opportunistic Approachability
por: Marinov, Teodor Vanislavov, et al.
Publicado: (2026)
por: Marinov, Teodor Vanislavov, et al.
Publicado: (2026)
Real-Time Parallel Counterfactual Regret Minimization
por: Li, Boning, et al.
Publicado: (2026)
por: Li, Boning, et al.
Publicado: (2026)
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
por: Xu, Hang, et al.
Publicado: (2024)
por: Xu, Hang, et al.
Publicado: (2024)
Solving Pasur Using GPU-Accelerated Counterfactual Regret Minimization
por: Baghal, Sina
Publicado: (2025)
por: Baghal, Sina
Publicado: (2025)
Regret Minimization and Convergence to Equilibria in General-sum Markov Games
por: Erez, Liad, et al.
Publicado: (2022)
por: Erez, Liad, et al.
Publicado: (2022)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
por: Jaafari, Zakaria El
Publicado: (2025)
por: Jaafari, Zakaria El
Publicado: (2025)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
por: Li, Yingru, et al.
Publicado: (2024)
por: Li, Yingru, et al.
Publicado: (2024)
ElicitationGPT: Text Elicitation Mechanisms via Language Models
por: Wu, Yifan, et al.
Publicado: (2024)
por: Wu, Yifan, et al.
Publicado: (2024)
Parallelizing Counterfactual Regret Minimization
por: Kim, Juho, et al.
Publicado: (2026)
por: Kim, Juho, et al.
Publicado: (2026)
Learning in Markov Games with Adaptive Adversaries: Policy Regret, Fundamental Barriers, and Efficient Algorithms
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
por: Park, Chanwoo, et al.
Publicado: (2024)
por: Park, Chanwoo, et al.
Publicado: (2024)
From External to Swap Regret 2.0: An Efficient Reduction and Oblivious Adversary for Large Action Spaces
por: Dagan, Yuval, et al.
Publicado: (2023)
por: Dagan, Yuval, et al.
Publicado: (2023)
Regret Minimization in Stackelberg Games with Side Information
por: Harris, Keegan, et al.
Publicado: (2024)
por: Harris, Keegan, et al.
Publicado: (2024)
Generative Social Choice: The Next Generation
por: Boehmer, Niclas, et al.
Publicado: (2025)
por: Boehmer, Niclas, et al.
Publicado: (2025)
Full Swap Regret and Discretized Calibration
por: Fishelson, Maxwell, et al.
Publicado: (2025)
por: Fishelson, Maxwell, et al.
Publicado: (2025)
GPU-Accelerated Counterfactual Regret Minimization
por: Kim, Juho
Publicado: (2024)
por: Kim, Juho
Publicado: (2024)
Regret Minimization in Bilateral Trade With Perturbed Markets
por: Lunghi, Anna, et al.
Publicado: (2026)
por: Lunghi, Anna, et al.
Publicado: (2026)
Meta-Learning in Self-Play Regret Minimization
por: Sychrovský, David, et al.
Publicado: (2025)
por: Sychrovský, David, et al.
Publicado: (2025)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
por: Zhang, Yuheng, et al.
Publicado: (2024)
por: Zhang, Yuheng, et al.
Publicado: (2024)
Selling Joint Ads: A Regret Minimization Perspective
por: Aggarwal, Gagan, et al.
Publicado: (2024)
por: Aggarwal, Gagan, et al.
Publicado: (2024)
Efficient Last-Iterate Convergence in Regret Minimization via Adaptive Reward Transformation
por: Ren, Hang, et al.
Publicado: (2025)
por: Ren, Hang, et al.
Publicado: (2025)
Regret Minimization for Piecewise Linear Rewards: Contracts, Auctions, and Beyond
por: Bacchiocchi, Francesco, et al.
Publicado: (2025)
por: Bacchiocchi, Francesco, et al.
Publicado: (2025)
Computational Lower Bounds for Regret Minimization in Normal-Form Games
por: Anagnostides, Ioannis, et al.
Publicado: (2024)
por: Anagnostides, Ioannis, et al.
Publicado: (2024)
Predicting human decisions with behavioral theories and machine learning
por: Plonsky, Ori, et al.
Publicado: (2019)
por: Plonsky, Ori, et al.
Publicado: (2019)
Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives
por: Velasco, Ander Artola, et al.
Publicado: (2025)
por: Velasco, Ander Artola, et al.
Publicado: (2025)
Regret Bounds for Competitive Resource Allocation with Endogenous Costs
por: Chai, Rui
Publicado: (2026)
por: Chai, Rui
Publicado: (2026)
Comparing Uniform Price and Discriminatory Multi-Unit Auctions through Regret Minimization
por: Potfer, Marius, et al.
Publicado: (2025)
por: Potfer, Marius, et al.
Publicado: (2025)
Human Choice Prediction in Language-based Persuasion Games: Simulation-based Off-Policy Evaluation
por: Shapira, Eilam, et al.
Publicado: (2023)
por: Shapira, Eilam, et al.
Publicado: (2023)
ElementaryNet: A Non-Strategic Neural Network for Predicting Human Behavior in Normal-Form Games
por: d'Eon, Greg, et al.
Publicado: (2025)
por: d'Eon, Greg, et al.
Publicado: (2025)
The Relationship between No-Regret Learning and Online Conformal Prediction
por: Ramalingam, Ramya, et al.
Publicado: (2025)
por: Ramalingam, Ramya, et al.
Publicado: (2025)
On the Limitations and Possibilities of Nash Regret Minimization in Zero-Sum Matrix Games under Noisy Feedback
por: Maiti, Arnab, et al.
Publicado: (2023)
por: Maiti, Arnab, et al.
Publicado: (2023)
Contracting with a Learning Agent
por: Guruganesh, Guru, et al.
Publicado: (2024)
por: Guruganesh, Guru, et al.
Publicado: (2024)
Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers
por: Sun, Haoran, et al.
Publicado: (2025)
por: Sun, Haoran, et al.
Publicado: (2025)
Axioms for AI Alignment from Human Feedback
por: Ge, Luise, et al.
Publicado: (2024)
por: Ge, Luise, et al.
Publicado: (2024)
Learning not to Regret
por: Sychrovský, David, et al.
Publicado: (2023)
por: Sychrovský, David, et al.
Publicado: (2023)
Ejemplares similares
-
Distributional Alignment Games for Answer-Level Fine-Tuning
por: Mohri, Mehryar, et al.
Publicado: (2026) -
Online Learning with Bounded Recall
por: Schneider, Jon, et al.
Publicado: (2022) -
High-Dimensional Calibration from Swap Regret
por: Fishelson, Maxwell, et al.
Publicado: (2025) -
Swap Regret and Correlated Equilibria Beyond Normal-Form Games
por: Arunachaleswaran, Eshwar Ram, et al.
Publicado: (2025) -
Deep (Predictive) Discounted Counterfactual Regret Minimization
por: Xu, Hang, et al.
Publicado: (2025)