Online Learning from Strategic Human Feedback in LLM Fine-Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Hao, Shugang, Duan, Lingjie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Save Mobile Crowdsourcing from Cheap-talk: A Game Theoretic Learning Approach
by: Hao, Shugang, et al.
Published: (2023)
by: Hao, Shugang, et al.
Published: (2023)
Human-in-the-loop Learning for Dynamic Congestion Games
by: Li, Hongbo, et al.
Published: (2024)
by: Li, Hongbo, et al.
Published: (2024)
To Analyze and Regulate Human-in-the-loop Learning for Congestion Games
by: Li, Hongbo, et al.
Published: (2025)
by: Li, Hongbo, et al.
Published: (2025)
Competitive Multi-armed Bandit Games for Resource Sharing
by: Li, Hongbo, et al.
Published: (2025)
by: Li, Hongbo, et al.
Published: (2025)
Approximating Human Strategic Reasoning with LLM-Enhanced Recursive Reasoners Leveraging Multi-agent Hypergames
by: Trencsenyi, Vince, et al.
Published: (2025)
by: Trencsenyi, Vince, et al.
Published: (2025)
Do Large Language Models Learn Human-Like Strategic Preferences?
by: Roberts, Jesse, et al.
Published: (2024)
by: Roberts, Jesse, et al.
Published: (2024)
To Optimize Human-in-the-loop Learning in Repeated Routing Games
by: Li, Hongbo, et al.
Published: (2024)
by: Li, Hongbo, et al.
Published: (2024)
Governing AI Forgetting: Auditing for Machine Unlearning Compliance
by: Lin, Qinqi, et al.
Published: (2026)
by: Lin, Qinqi, et al.
Published: (2026)
The Disparate Effects of Partial Information in Bayesian Strategic Learning
by: Avasarala, Srikanth, et al.
Published: (2025)
by: Avasarala, Srikanth, et al.
Published: (2025)
Everyone Contributes! Incentivizing Strategic Cooperation in Multi-LLM Systems via Sequential Public Goods Games
by: Liang, Yunhao, et al.
Published: (2025)
by: Liang, Yunhao, et al.
Published: (2025)
VickreyFeedback: Cost-efficient Data Construction for Reinforcement Learning from Human Feedback
by: Zhang, Guoxi, et al.
Published: (2024)
by: Zhang, Guoxi, et al.
Published: (2024)
How Far Can LLMs Emulate Human Behavior?: A Strategic Analysis via the Buy-and-Sell Negotiation Game
by: Jeon, Mingyu, et al.
Published: (2025)
by: Jeon, Mingyu, et al.
Published: (2025)
Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments
by: Chen, Yixin, et al.
Published: (2023)
by: Chen, Yixin, et al.
Published: (2023)
Distributed Learning for Dynamic Congestion Games
by: Li, Hongbo, et al.
Published: (2024)
by: Li, Hongbo, et al.
Published: (2024)
Axioms for AI Alignment from Human Feedback
by: Ge, Luise, et al.
Published: (2024)
by: Ge, Luise, et al.
Published: (2024)
Strategic Resource Selection with Homophilic Agents
by: Harder, Jonathan Gadea, et al.
Published: (2023)
by: Harder, Jonathan Gadea, et al.
Published: (2023)
Nash Learning from Human Feedback
by: Munos, Rémi, et al.
Published: (2023)
by: Munos, Rémi, et al.
Published: (2023)
Conversation Games and a Strategic View of the Turing Test
by: Aryan, Kaveh
Published: (2025)
by: Aryan, Kaveh
Published: (2025)
AI Testing Should Account for Sophisticated Strategic Behaviour
by: Kovarik, Vojtech, et al.
Published: (2025)
by: Kovarik, Vojtech, et al.
Published: (2025)
Two-Stage Facility Location Games with Strategic Clients and Facilities
by: Krogmann, Simon, et al.
Published: (2021)
by: Krogmann, Simon, et al.
Published: (2021)
On the Complexity of Winner Determination and Strategic Control in Conditional Approval Voting
by: Markakis, Evangelos, et al.
Published: (2022)
by: Markakis, Evangelos, et al.
Published: (2022)
Strategic Facility Location with Clients that Minimize Total Waiting Time
by: Krogmann, Simon, et al.
Published: (2022)
by: Krogmann, Simon, et al.
Published: (2022)
Computing Voting Rules with Improvement Feedback
by: Micha, Evi, et al.
Published: (2025)
by: Micha, Evi, et al.
Published: (2025)
ALYMPICS: LLM Agents Meet Game Theory -- Exploring Strategic Decision-Making with AI Agents
by: Mao, Shaoguang, et al.
Published: (2023)
by: Mao, Shaoguang, et al.
Published: (2023)
Strategic Bidding in 6G Spectrum Auctions with Large Language Models
by: Lotfi, Ismail, et al.
Published: (2026)
by: Lotfi, Ismail, et al.
Published: (2026)
LLMs as Strategic Agents: Beliefs, Best Response Behavior, and Emergent Heuristics
by: de Fortuny, Enric Junque, et al.
Published: (2025)
by: de Fortuny, Enric Junque, et al.
Published: (2025)
Fine-Tuning Games: Bargaining and Adaptation for General-Purpose Models
by: Laufer, Benjamin, et al.
Published: (2023)
by: Laufer, Benjamin, et al.
Published: (2023)
Strategic Tradeoffs Between Humans and AI in Multi-Agent Bargaining
by: Qian, Crystal, et al.
Published: (2025)
by: Qian, Crystal, et al.
Published: (2025)
TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs
by: Wang, Haochuan, et al.
Published: (2024)
by: Wang, Haochuan, et al.
Published: (2024)
Beyond Right to be Forgotten: Managing Heterogeneity Side Effects Through Strategic Incentives
by: Shao, Jiaqi, et al.
Published: (2024)
by: Shao, Jiaqi, et al.
Published: (2024)
Beyond Nash Equilibrium: Bounded Rationality of LLMs and humans in Strategic Decision-making
by: Zheng, Kehan, et al.
Published: (2025)
by: Zheng, Kehan, et al.
Published: (2025)
Online Housing Market
by: Lesca, Julien
Published: (2025)
by: Lesca, Julien
Published: (2025)
Learning in Strategic Queuing Systems with Small Buffers
by: Abel, Ariana, et al.
Published: (2025)
by: Abel, Ariana, et al.
Published: (2025)
Do LLM Agents Have Regret? A Case Study in Online Learning and Games
by: Park, Chanwoo, et al.
Published: (2024)
by: Park, Chanwoo, et al.
Published: (2024)
Online Fair Division with Additional Information
by: Neoh, Tzeh Yuan, et al.
Published: (2025)
by: Neoh, Tzeh Yuan, et al.
Published: (2025)
Truthful Aggregation of LLMs with an Application to Online Advertising
by: Soumalias, Ermis, et al.
Published: (2024)
by: Soumalias, Ermis, et al.
Published: (2024)
Governing Strategic Dynamics: Equilibrium Stabilization via Divergence-Driven Control
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Multi-Agent Strategic Games with LLMs
by: Chupilkin, Maxim
Published: (2026)
by: Chupilkin, Maxim
Published: (2026)
Towards Strategic Persuasion with Language Models
by: Cheng, Zirui, et al.
Published: (2025)
by: Cheng, Zirui, et al.
Published: (2025)
Personalized Pricing Through Strategic User Profiling in Social Networks
by: Lin, Qinqi, et al.
Published: (2025)
by: Lin, Qinqi, et al.
Published: (2025)
Similar Items
-
To Save Mobile Crowdsourcing from Cheap-talk: A Game Theoretic Learning Approach
by: Hao, Shugang, et al.
Published: (2023) -
Human-in-the-loop Learning for Dynamic Congestion Games
by: Li, Hongbo, et al.
Published: (2024) -
To Analyze and Regulate Human-in-the-loop Learning for Congestion Games
by: Li, Hongbo, et al.
Published: (2025) -
Competitive Multi-armed Bandit Games for Resource Sharing
by: Li, Hongbo, et al.
Published: (2025) -
Approximating Human Strategic Reasoning with LLM-Enhanced Recursive Reasoners Leveraging Multi-agent Hypergames
by: Trencsenyi, Vince, et al.
Published: (2025)