MERIT Feedback Elicits Better Bargaining in LLM Negotiators
Fuente:
arXiv
Salvato in:
| Autori principali: | Oh, Jihwan, Aghazada, Murad, Shin, Yooju, Yun, Se-Young, Kim, Taehyeon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
LLM Agents for Bargaining with Utility-based Feedback
di: Oh, Jihwan
Pubblicazione: (2025)
di: Oh, Jihwan
Pubblicazione: (2025)
Learning to Verify Summary Facts with Fine-Grained LLM Feedback
di: Oh, Jihwan, et al.
Pubblicazione: (2024)
di: Oh, Jihwan, et al.
Pubblicazione: (2024)
LLM Rationalis? Measuring Bargaining Capabilities of AI Negotiators
di: Shah, Cheril, et al.
Pubblicazione: (2025)
di: Shah, Cheril, et al.
Pubblicazione: (2025)
$C^2$: Scalable Auto-Feedback for LLM-based Chart Generation
di: Koh, Woosung, et al.
Pubblicazione: (2024)
di: Koh, Woosung, et al.
Pubblicazione: (2024)
The Language of Bargaining: Linguistic Effects in LLM Negotiations
di: Sinha, Stuti, et al.
Pubblicazione: (2026)
di: Sinha, Stuti, et al.
Pubblicazione: (2026)
Learning to Summarize from LLM-generated Feedback
di: Song, Hwanjun, et al.
Pubblicazione: (2024)
di: Song, Hwanjun, et al.
Pubblicazione: (2024)
VarDrop: Enhancing Training Efficiency by Reducing Variate Redundancy in Periodic Time Series Forecasting
di: Kang, Junhyeok, et al.
Pubblicazione: (2025)
di: Kang, Junhyeok, et al.
Pubblicazione: (2025)
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
di: Yun, Taewon, et al.
Pubblicazione: (2025)
di: Yun, Taewon, et al.
Pubblicazione: (2025)
Self-Training Elicits Concise Reasoning in Large Language Models
di: Munkhbat, Tergel, et al.
Pubblicazione: (2025)
di: Munkhbat, Tergel, et al.
Pubblicazione: (2025)
Multi-Drafter Speculative Decoding with Alignment Feedback
di: Kim, Taehyeon, et al.
Pubblicazione: (2026)
di: Kim, Taehyeon, et al.
Pubblicazione: (2026)
Predicting LLM Reasoning Performance with Small Proxy Model
di: Koh, Woosung, et al.
Pubblicazione: (2025)
di: Koh, Woosung, et al.
Pubblicazione: (2025)
Executable Code Actions Elicit Better LLM Agents
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
AdaSTaR: Adaptive Data Sampling for Training Self-Taught Reasoners
di: Koh, Woosung, et al.
Pubblicazione: (2025)
di: Koh, Woosung, et al.
Pubblicazione: (2025)
DistiLLM: Towards Streamlined Distillation for Large Language Models
di: Ko, Jongwoo, et al.
Pubblicazione: (2024)
di: Ko, Jongwoo, et al.
Pubblicazione: (2024)
Talking with Tables for Better LLM Factual Data Interactions
di: Oh, Jio, et al.
Pubblicazione: (2024)
di: Oh, Jio, et al.
Pubblicazione: (2024)
Block Transformer: Global-to-Local Language Modeling for Fast Inference
di: Ho, Namgyu, et al.
Pubblicazione: (2024)
di: Ho, Namgyu, et al.
Pubblicazione: (2024)
Automated Filtering of Human Feedback Data for Aligning Text-to-Image Diffusion Models
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
The MERIT Dataset: Modelling and Efficiently Rendering Interpretable Transcripts
di: de Rodrigo, I., et al.
Pubblicazione: (2024)
di: de Rodrigo, I., et al.
Pubblicazione: (2024)
FlickerFusion: Intra-trajectory Domain Generalizing Multi-Agent RL
di: Koh, Woosung, et al.
Pubblicazione: (2024)
di: Koh, Woosung, et al.
Pubblicazione: (2024)
What is the Alignment Objective of GRPO?
di: Vojnovic, Milan, et al.
Pubblicazione: (2025)
di: Vojnovic, Milan, et al.
Pubblicazione: (2025)
FishBargain: An LLM-Empowered Bargaining Agent for Online Fleamarket Platform Sellers
di: Kong, Dexin, et al.
Pubblicazione: (2025)
di: Kong, Dexin, et al.
Pubblicazione: (2025)
VorTEX: Various overlap ratio for Target speech EXtraction
di: Oh, Ro-hoon, et al.
Pubblicazione: (2026)
di: Oh, Ro-hoon, et al.
Pubblicazione: (2026)
Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs
di: Song, Woomin, et al.
Pubblicazione: (2024)
di: Song, Woomin, et al.
Pubblicazione: (2024)
MERIT: Multi-domain Efficient RAW Image Translation
di: Huang, Wenjun, et al.
Pubblicazione: (2026)
di: Huang, Wenjun, et al.
Pubblicazione: (2026)
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback
di: Yang, Yongjin, et al.
Pubblicazione: (2025)
di: Yang, Yongjin, et al.
Pubblicazione: (2025)
FedSOL: Stabilized Orthogonal Learning with Proximal Restrictions in Federated Learning
di: Lee, Gihun, et al.
Pubblicazione: (2023)
di: Lee, Gihun, et al.
Pubblicazione: (2023)
Scaling Laws of SignSGD in Linear Regression: When Does It Outperform SGD?
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
di: Kim, Jihwan, et al.
Pubblicazione: (2026)
MERIT: Memory-Enhanced Retrieval for Interpretable Knowledge Tracing
di: Li, Runze, et al.
Pubblicazione: (2026)
di: Li, Runze, et al.
Pubblicazione: (2026)
Revisiting Early-Learning Regularization When Federated Learning Meets Noisy Labels
di: Kim, Taehyeon, et al.
Pubblicazione: (2024)
di: Kim, Taehyeon, et al.
Pubblicazione: (2024)
Flex-Judge: Text-Only Reasoning Unleashes Zero-Shot Multimodal Evaluators
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
di: Ko, Jongwoo, et al.
Pubblicazione: (2025)
Universal Time-Series Representation Learning: A Survey
di: Trirat, Patara, et al.
Pubblicazione: (2024)
di: Trirat, Patara, et al.
Pubblicazione: (2024)
Position on LLM-Assisted Peer Review: Addressing Reviewer Gap through Mentoring and Feedback
di: Yun, JungMin, et al.
Pubblicazione: (2026)
di: Yun, JungMin, et al.
Pubblicazione: (2026)
BAPO: Base-Anchored Preference Optimization for Overcoming Forgetting in Large Language Models Personalization
di: Lee, Gihun, et al.
Pubblicazione: (2024)
di: Lee, Gihun, et al.
Pubblicazione: (2024)
Eliciting Rational Initial Weights in Gradual Argumentation
di: Oren, Nir, et al.
Pubblicazione: (2025)
di: Oren, Nir, et al.
Pubblicazione: (2025)
From Belief Entrenchment to Robust Reasoning in LLM Agents
di: Oh, Jihwan, et al.
Pubblicazione: (2025)
di: Oh, Jihwan, et al.
Pubblicazione: (2025)
Contextual Linear Bandits under Noisy Features: Towards Bayesian Oracles
di: Kim, Jung-hun, et al.
Pubblicazione: (2017)
di: Kim, Jung-hun, et al.
Pubblicazione: (2017)
Evaluating Multi-Turn Bargain Skills in LLM-Based Seller Agent
di: Wang, Issue Yishu, et al.
Pubblicazione: (2025)
di: Wang, Issue Yishu, et al.
Pubblicazione: (2025)
Synergistic Integration of Coordinate Network and Tensorial Feature for Improving Neural Radiance Fields from Sparse Inputs
di: Kim, Mingyu, et al.
Pubblicazione: (2024)
di: Kim, Mingyu, et al.
Pubblicazione: (2024)
Verbal Process Supervision Elicits Better Coding Agents
di: Chen, Hao-Yuan, et al.
Pubblicazione: (2025)
di: Chen, Hao-Yuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment
di: Oh, Jihwan, et al.
Pubblicazione: (2026) -
LLM Agents for Bargaining with Utility-based Feedback
di: Oh, Jihwan
Pubblicazione: (2025) -
Learning to Verify Summary Facts with Fine-Grained LLM Feedback
di: Oh, Jihwan, et al.
Pubblicazione: (2024) -
LLM Rationalis? Measuring Bargaining Capabilities of AI Negotiators
di: Shah, Cheril, et al.
Pubblicazione: (2025) -
$C^2$: Scalable Auto-Feedback for LLM-based Chart Generation
di: Koh, Woosung, et al.
Pubblicazione: (2024)