LLM Agents for Bargaining with Utility-based Feedback
Fuente:
arXiv
Saved in:
| Main Author: | Oh, Jihwan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Belief Entrenchment to Robust Reasoning in LLM Agents
by: Oh, Jihwan, et al.
Published: (2025)
by: Oh, Jihwan, et al.
Published: (2025)
Diffusion-based Episodes Augmentation for Offline Multi-Agent Reinforcement Learning
by: Oh, Jihwan, et al.
Published: (2024)
by: Oh, Jihwan, et al.
Published: (2024)
MERIT Feedback Elicits Better Bargaining in LLM Negotiators
by: Oh, Jihwan, et al.
Published: (2026)
by: Oh, Jihwan, et al.
Published: (2026)
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
by: Hwang, Taehyun, et al.
Published: (2026)
by: Hwang, Taehyun, et al.
Published: (2026)
Robust and Consistent Ski Rental with Distributional Advice
by: Kim, Jihwan, et al.
Published: (2026)
by: Kim, Jihwan, et al.
Published: (2026)
DP-Muon: Differentially Private Optimization via Matrix-Orthogonalized Momentum
by: Kim, Jihwan, et al.
Published: (2026)
by: Kim, Jihwan, et al.
Published: (2026)
Combinatorial Reinforcement Learning with Preference Feedback
by: Lee, Joongkyu, et al.
Published: (2025)
by: Lee, Joongkyu, et al.
Published: (2025)
Queueing Matching Bandits with Preference Feedback
by: Kim, Jung-hun, et al.
Published: (2024)
by: Kim, Jung-hun, et al.
Published: (2024)
Representation Signatures and Risk-Feedback Alignment in LLM Trading Agents
by: Xue, Weicheng
Published: (2026)
by: Xue, Weicheng
Published: (2026)
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
Fairness-Aware Meta-Learning via Nash Bargaining
by: Zeng, Yi, et al.
Published: (2024)
by: Zeng, Yi, et al.
Published: (2024)
Preference Alignment with Flow Matching
by: Kim, Minu, et al.
Published: (2024)
by: Kim, Minu, et al.
Published: (2024)
Dynamic Assortment Selection and Pricing with Censored Preference Feedback
by: Kim, Jung-hun, et al.
Published: (2025)
by: Kim, Jung-hun, et al.
Published: (2025)
LLM Routing with Dueling Feedback
by: Chiang, Chao-Kai, et al.
Published: (2025)
by: Chiang, Chao-Kai, et al.
Published: (2025)
TVCACHE: A Stateful Tool-Value Cache for Post-Training LLM Agents
by: Kumar, Abhishek Vijaya, et al.
Published: (2026)
by: Kumar, Abhishek Vijaya, et al.
Published: (2026)
A Bargaining-based Approach for Feature Trading in Vertical Federated Learning
by: Cui, Yue, et al.
Published: (2024)
by: Cui, Yue, et al.
Published: (2024)
RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents
by: Aouali, Imad, et al.
Published: (2026)
by: Aouali, Imad, et al.
Published: (2026)
IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization
by: Chen, Zixuan, et al.
Published: (2026)
by: Chen, Zixuan, et al.
Published: (2026)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024)
by: Choudhury, Sanjiban, et al.
Published: (2024)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
by: Jiralerspong, Thomas, et al.
Published: (2026)
by: Jiralerspong, Thomas, et al.
Published: (2026)
PathFinder: MCTS and LLM Feedback-based Path Selection for Multi-Hop Question Answering
by: Maram, Durga Prasad, et al.
Published: (2025)
by: Maram, Durga Prasad, et al.
Published: (2025)
Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation
by: Kim, Jinuk, et al.
Published: (2026)
by: Kim, Jinuk, et al.
Published: (2026)
STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning
by: Zhang, Junjie, et al.
Published: (2026)
by: Zhang, Junjie, et al.
Published: (2026)
Can Graph Learning Improve Planning in LLM-based Agents?
by: Wu, Xixi, et al.
Published: (2024)
by: Wu, Xixi, et al.
Published: (2024)
Personalized Auto-Grading and Feedback System for Constructive Geometry Tasks Using Large Language Models on an Online Math Platform
by: Lee, Yong Oh, et al.
Published: (2025)
by: Lee, Yong Oh, et al.
Published: (2025)
Breaking the Barrier: Enhanced Utility and Robustness in Smoothed DRL Agents
by: Sun, Chung-En, et al.
Published: (2024)
by: Sun, Chung-En, et al.
Published: (2024)
Partial Inverse Design of High-Performance Concrete Using Cooperative Neural Networks for Constraint-Aware Mix Generation
by: Nugraha, Agung, et al.
Published: (2025)
by: Nugraha, Agung, et al.
Published: (2025)
Projectable Models: One-Shot Generation of Small Specialized Transformers from Large Ones
by: Zhmoginov, Andrey, et al.
Published: (2025)
by: Zhmoginov, Andrey, et al.
Published: (2025)
Lusifer: LLM-based User SImulated Feedback Environment for online Recommender systems
by: Ebrat, Danial, et al.
Published: (2024)
by: Ebrat, Danial, et al.
Published: (2024)
Utility-based Dueling Bandits as a Partial Monitoring Game
by: Gajane, Pratik, et al.
Published: (2015)
by: Gajane, Pratik, et al.
Published: (2015)
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2026)
by: Miceli-Barone, Antonio Valerio, et al.
Published: (2026)
Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization
by: Barj, Houda Nait El, et al.
Published: (2024)
by: Barj, Houda Nait El, et al.
Published: (2024)
Preserving Privacy and Utility in LLM-Based Product Recommendations
by: Khezresmaeilzadeh, Tina, et al.
Published: (2025)
by: Khezresmaeilzadeh, Tina, et al.
Published: (2025)
Aligning Dense Retrievers with LLM Utility via Distillation
by: Sandhu, Rajinder, et al.
Published: (2026)
by: Sandhu, Rajinder, et al.
Published: (2026)
$C^2$: Scalable Auto-Feedback for LLM-based Chart Generation
by: Koh, Woosung, et al.
Published: (2024)
by: Koh, Woosung, et al.
Published: (2024)
Last-Iterate Convergence of No-Regret Learning for Equilibria in Bargaining Games
by: Kamp, Serafina, et al.
Published: (2025)
by: Kamp, Serafina, et al.
Published: (2025)
Maximin Relative Improvement: Fair Learning as a Bargaining Problem
by: Han, Jiwoo, et al.
Published: (2026)
by: Han, Jiwoo, et al.
Published: (2026)
OSNIP: Breaking the Privacy-Utility-Efficiency Trilemma in LLM Inference via Obfuscated Semantic Null Space
by: Cao, Zhiyuan, et al.
Published: (2026)
by: Cao, Zhiyuan, et al.
Published: (2026)
Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation
by: Tripathi, Tuhina, et al.
Published: (2025)
by: Tripathi, Tuhina, et al.
Published: (2025)
In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration
by: Choi, Youngbin, et al.
Published: (2025)
by: Choi, Youngbin, et al.
Published: (2025)
Similar Items
-
From Belief Entrenchment to Robust Reasoning in LLM Agents
by: Oh, Jihwan, et al.
Published: (2025) -
Diffusion-based Episodes Augmentation for Offline Multi-Agent Reinforcement Learning
by: Oh, Jihwan, et al.
Published: (2024) -
MERIT Feedback Elicits Better Bargaining in LLM Negotiators
by: Oh, Jihwan, et al.
Published: (2026) -
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
by: Hwang, Taehyun, et al.
Published: (2026) -
Robust and Consistent Ski Rental with Distributional Advice
by: Kim, Jihwan, et al.
Published: (2026)