Using LLMs to Model the Beliefs and Preferences of Targeted Populations
Fuente:
arXiv
Salvato in:
| Autori principali: | Namikoshi, Keiichi, Filipowicz, Alex, Shamma, David A., Iliev, Rumen, Hogan, Candice L., Arechiga, Nikos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Leveraging Language Models and Bandit Algorithms to Drive Adoption of Battery-Electric Vehicles
di: Namikoshi, Keiichi, et al.
Pubblicazione: (2024)
di: Namikoshi, Keiichi, et al.
Pubblicazione: (2024)
ConjointNet: Enhancing Conjoint Analysis for Preference Prediction with Representation Learning
di: Zhang, Yanxia, et al.
Pubblicazione: (2025)
di: Zhang, Yanxia, et al.
Pubblicazione: (2025)
On LLM Wizards: Identifying Large Language Models' Behaviors for Wizard of Oz Experiments
di: Fang, Jingchao, et al.
Pubblicazione: (2024)
di: Fang, Jingchao, et al.
Pubblicazione: (2024)
The Belief State Transformer
di: Hu, Edward S., et al.
Pubblicazione: (2024)
di: Hu, Edward S., et al.
Pubblicazione: (2024)
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
di: Dugan, Owen, et al.
Pubblicazione: (2024)
di: Dugan, Owen, et al.
Pubblicazione: (2024)
Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
di: Wang, Haoxiang, et al.
Pubblicazione: (2024)
di: Wang, Haoxiang, et al.
Pubblicazione: (2024)
RouteLLM: Learning to Route LLMs with Preference Data
di: Ong, Isaac, et al.
Pubblicazione: (2024)
di: Ong, Isaac, et al.
Pubblicazione: (2024)
Targeted Visualization of the Backbone of Encoder LLMs
di: Roberts, Isaac, et al.
Pubblicazione: (2024)
di: Roberts, Isaac, et al.
Pubblicazione: (2024)
Active Preference Learning for Large Language Models
di: Muldrew, William, et al.
Pubblicazione: (2024)
di: Muldrew, William, et al.
Pubblicazione: (2024)
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization
di: Kawakami, Wataru, et al.
Pubblicazione: (2025)
di: Kawakami, Wataru, et al.
Pubblicazione: (2025)
DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
di: Chiu, Yu Ying, et al.
Pubblicazione: (2024)
di: Chiu, Yu Ying, et al.
Pubblicazione: (2024)
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
di: Boppana, Siddharth, et al.
Pubblicazione: (2026)
di: Boppana, Siddharth, et al.
Pubblicazione: (2026)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
di: Lai, Xin, et al.
Pubblicazione: (2024)
di: Lai, Xin, et al.
Pubblicazione: (2024)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
MPPO: Multi Pair-wise Preference Optimization for LLMs with Arbitrary Negative Samples
di: Xie, Shuo, et al.
Pubblicazione: (2024)
di: Xie, Shuo, et al.
Pubblicazione: (2024)
RLHF Can Speak Many Languages: Unlocking Multilingual Preference Optimization for LLMs
di: Dang, John, et al.
Pubblicazione: (2024)
di: Dang, John, et al.
Pubblicazione: (2024)
Learning to Represent Individual Differences for Choice Decision Making
di: Chen, Yan-Ying, et al.
Pubblicazione: (2025)
di: Chen, Yan-Ying, et al.
Pubblicazione: (2025)
Tool Preferences in Agentic LLMs are Unreliable
di: Faghih, Kazem, et al.
Pubblicazione: (2025)
di: Faghih, Kazem, et al.
Pubblicazione: (2025)
Course-Correction: Safety Alignment Using Synthetic Preferences
di: Xu, Rongwu, et al.
Pubblicazione: (2024)
di: Xu, Rongwu, et al.
Pubblicazione: (2024)
When Should Models Change Their Minds? Contextual Belief Management in Large Language Models
di: Xu, Haoming, et al.
Pubblicazione: (2026)
di: Xu, Haoming, et al.
Pubblicazione: (2026)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
di: Misra, Dipendra, et al.
Pubblicazione: (2026)
di: Misra, Dipendra, et al.
Pubblicazione: (2026)
When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure
di: Xiao, Boyu, et al.
Pubblicazione: (2026)
di: Xiao, Boyu, et al.
Pubblicazione: (2026)
Improving Context-Aware Preference Modeling for Language Models
di: Pitis, Silviu, et al.
Pubblicazione: (2024)
di: Pitis, Silviu, et al.
Pubblicazione: (2024)
Do LLMs Recognize Your Latent Preferences? A Benchmark for Latent Information Discovery in Personalized Interaction
di: Tsaknakis, Ioannis, et al.
Pubblicazione: (2025)
di: Tsaknakis, Ioannis, et al.
Pubblicazione: (2025)
Approximating Human Preferences Using a Multi-Judge Learned System
di: Sprejer, Eitán, et al.
Pubblicazione: (2025)
di: Sprejer, Eitán, et al.
Pubblicazione: (2025)
On the Role of Preference Variance in Preference Optimization
di: Guo, Jiacheng, et al.
Pubblicazione: (2025)
di: Guo, Jiacheng, et al.
Pubblicazione: (2025)
Theory of Space: Can Foundation Models Construct Spatial Beliefs through Active Exploration?
di: Zhang, Pingyue, et al.
Pubblicazione: (2026)
di: Zhang, Pingyue, et al.
Pubblicazione: (2026)
Preference Poisoning Attacks on Reward Model Learning
di: Wu, Junlin, et al.
Pubblicazione: (2024)
di: Wu, Junlin, et al.
Pubblicazione: (2024)
Aligners: Decoupling LLMs and Alignment
di: Ngweta, Lilian, et al.
Pubblicazione: (2024)
di: Ngweta, Lilian, et al.
Pubblicazione: (2024)
ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs
di: Wang, Zige, et al.
Pubblicazione: (2025)
di: Wang, Zige, et al.
Pubblicazione: (2025)
Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
di: Baumann, Joachim, et al.
Pubblicazione: (2025)
di: Baumann, Joachim, et al.
Pubblicazione: (2025)
ROPO: Robust Preference Optimization for Large Language Models
di: Liang, Xize, et al.
Pubblicazione: (2024)
di: Liang, Xize, et al.
Pubblicazione: (2024)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
di: Imai, Saki, et al.
Pubblicazione: (2026)
di: Imai, Saki, et al.
Pubblicazione: (2026)
Accelerated Preference Optimization for Large Language Model Alignment
di: He, Jiafan, et al.
Pubblicazione: (2024)
di: He, Jiafan, et al.
Pubblicazione: (2024)
Self-Play Preference Optimization for Language Model Alignment
di: Wu, Yue, et al.
Pubblicazione: (2024)
di: Wu, Yue, et al.
Pubblicazione: (2024)
Preference Learning Algorithms Do Not Learn Preference Rankings
di: Chen, Angelica, et al.
Pubblicazione: (2024)
di: Chen, Angelica, et al.
Pubblicazione: (2024)
Geometric-Averaged Preference Optimization for Soft Preference Labels
di: Furuta, Hiroki, et al.
Pubblicazione: (2024)
di: Furuta, Hiroki, et al.
Pubblicazione: (2024)
Adaptive Margin RLHF via Preference over Preferences
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
di: Chittepu, Yaswanth, et al.
Pubblicazione: (2025)
HelpSteer2-Preference: Complementing Ratings with Preferences
di: Wang, Zhilin, et al.
Pubblicazione: (2024)
di: Wang, Zhilin, et al.
Pubblicazione: (2024)
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning
di: Driss, Brahim, et al.
Pubblicazione: (2025)
di: Driss, Brahim, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Leveraging Language Models and Bandit Algorithms to Drive Adoption of Battery-Electric Vehicles
di: Namikoshi, Keiichi, et al.
Pubblicazione: (2024) -
ConjointNet: Enhancing Conjoint Analysis for Preference Prediction with Representation Learning
di: Zhang, Yanxia, et al.
Pubblicazione: (2025) -
On LLM Wizards: Identifying Large Language Models' Behaviors for Wizard of Oz Experiments
di: Fang, Jingchao, et al.
Pubblicazione: (2024) -
The Belief State Transformer
di: Hu, Edward S., et al.
Pubblicazione: (2024) -
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
di: Dugan, Owen, et al.
Pubblicazione: (2024)