FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Anikait, Hsu, Sheryl, Hsu, Kyle, Mitchell, Eric, Ermon, Stefano, Hashimoto, Tatsunori, Sharma, Archit, Finn, Chelsea |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
di: Rafailov, Rafael, et al.
Pubblicazione: (2023)
di: Rafailov, Rafael, et al.
Pubblicazione: (2023)
RLVF: Learning from Verbal Feedback without Overgeneralization
di: Stephan, Moritz, et al.
Pubblicazione: (2024)
di: Stephan, Moritz, et al.
Pubblicazione: (2024)
Personalized Preference Fine-tuning of Diffusion Models
di: Dang, Meihua, et al.
Pubblicazione: (2025)
di: Dang, Meihua, et al.
Pubblicazione: (2025)
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
di: Hsu, Sheryl, et al.
Pubblicazione: (2024)
di: Hsu, Sheryl, et al.
Pubblicazione: (2024)
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
di: Manvi, Rohin, et al.
Pubblicazione: (2024)
di: Manvi, Rohin, et al.
Pubblicazione: (2024)
Disentangling Length from Quality in Direct Preference Optimization
di: Park, Ryan, et al.
Pubblicazione: (2024)
di: Park, Ryan, et al.
Pubblicazione: (2024)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
di: Si, Chenglei, et al.
Pubblicazione: (2024)
di: Si, Chenglei, et al.
Pubblicazione: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025)
di: Si, Chenglei, et al.
Pubblicazione: (2025)
Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data
di: Tajwar, Fahim, et al.
Pubblicazione: (2024)
di: Tajwar, Fahim, et al.
Pubblicazione: (2024)
Towards Aligning Personalized Conversational Recommendation Agents with Users' Privacy Preferences
di: Zhang, Shuning, et al.
Pubblicazione: (2025)
di: Zhang, Shuning, et al.
Pubblicazione: (2025)
Personality Matters: User Traits Predict LLM Preferences in Multi-Turn Collaborative Tasks
di: Yunusov, Sarfaroz, et al.
Pubblicazione: (2025)
di: Yunusov, Sarfaroz, et al.
Pubblicazione: (2025)
Exploring the Role of Interaction Data to Empower End-User Decision-Making In UI Personalization
di: Alves, Sérgio, et al.
Pubblicazione: (2026)
di: Alves, Sérgio, et al.
Pubblicazione: (2026)
Citizen-Led Personalization of User Interfaces: Investigating How People Customize Interfaces for Themselves and Others
di: Alves, Sérgio, et al.
Pubblicazione: (2024)
di: Alves, Sérgio, et al.
Pubblicazione: (2024)
The Third-Party Access Effect: An Overlooked Challenge in Secondary Use of Educational Real-World Data
di: Ito, Hibiki, et al.
Pubblicazione: (2026)
di: Ito, Hibiki, et al.
Pubblicazione: (2026)
Users Mispredict Their Own Preferences for AI Writing Assistance
di: Lai, Vivian, et al.
Pubblicazione: (2026)
di: Lai, Vivian, et al.
Pubblicazione: (2026)
Test-Time Alignment via Hypothesis Reweighting
di: Lee, Yoonho, et al.
Pubblicazione: (2024)
di: Lee, Yoonho, et al.
Pubblicazione: (2024)
Navigation beyond Wayfinding: Robots Collaborating with Visually Impaired Users for Environmental Interactions
di: Cai, Shaojun, et al.
Pubblicazione: (2026)
di: Cai, Shaojun, et al.
Pubblicazione: (2026)
LiteraryTaste: A Preference Dataset for Creative Writing Personalization
di: Chung, John Joon Young, et al.
Pubblicazione: (2025)
di: Chung, John Joon Young, et al.
Pubblicazione: (2025)
Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation
di: Wei, Tianjun, et al.
Pubblicazione: (2025)
di: Wei, Tianjun, et al.
Pubblicazione: (2025)
Learning to Plan with Personalized Preferences
di: Xu, Manjie, et al.
Pubblicazione: (2025)
di: Xu, Manjie, et al.
Pubblicazione: (2025)
LearnAct: Few-Shot Mobile GUI Agent with a Unified Demonstration Benchmark
di: Liu, Guangyi, et al.
Pubblicazione: (2025)
di: Liu, Guangyi, et al.
Pubblicazione: (2025)
A Critical Evaluation of AI Feedback for Aligning Large Language Models
di: Sharma, Archit, et al.
Pubblicazione: (2024)
di: Sharma, Archit, et al.
Pubblicazione: (2024)
PrefIx: Understand and Adapt to User Preference in Human-Agent Interaction
di: Li, Jialin, et al.
Pubblicazione: (2026)
di: Li, Jialin, et al.
Pubblicazione: (2026)
User Characteristics in Explainable AI: The Rabbit Hole of Personalization?
di: Nimmo, Robert, et al.
Pubblicazione: (2024)
di: Nimmo, Robert, et al.
Pubblicazione: (2024)
IP-Dialog: Evaluating Implicit Personalization in Dialogue Systems with Synthetic Data
di: Peng, Bo, et al.
Pubblicazione: (2025)
di: Peng, Bo, et al.
Pubblicazione: (2025)
GraspR: A Computational Model of Spatial User Preferences for Adaptive Grasp UI Design
di: Caetano, Arthur, et al.
Pubblicazione: (2025)
di: Caetano, Arthur, et al.
Pubblicazione: (2025)
Fairness-Aware Few-Shot Learning for Audio-Visual Stress Detection
di: Shelke, Anushka Sanjay, et al.
Pubblicazione: (2025)
di: Shelke, Anushka Sanjay, et al.
Pubblicazione: (2025)
Exploring Personality-Driven Personalization in XAI: Enhancing User Trust in Gameplay
di: Li, Zhaoxin, et al.
Pubblicazione: (2024)
di: Li, Zhaoxin, et al.
Pubblicazione: (2024)
The Impact of Different Virtual Work Environments on Flow, Performance, User Emotions, and Preferences
di: Kiluk, Alicja, et al.
Pubblicazione: (2023)
di: Kiluk, Alicja, et al.
Pubblicazione: (2023)
MYCloth: Towards Intelligent and Interactive Online T-Shirt Customization based on User's Preference
di: Liu, Yexin, et al.
Pubblicazione: (2024)
di: Liu, Yexin, et al.
Pubblicazione: (2024)
Demonstrating PilotAR: A Tool to Assist Wizard-of-Oz Pilot Studies with OHMD
di: Janaka, Nuwan, et al.
Pubblicazione: (2024)
di: Janaka, Nuwan, et al.
Pubblicazione: (2024)
Preserving Sense of Agency: User Preferences for Robot Autonomy and User Control across Household Tasks
di: Yang, Claire, et al.
Pubblicazione: (2025)
di: Yang, Claire, et al.
Pubblicazione: (2025)
Adapting to the User: A Systematic Review of Personalized Interaction in VR
di: Li, Tangyao, et al.
Pubblicazione: (2025)
di: Li, Tangyao, et al.
Pubblicazione: (2025)
Promptimizer: User-Led Prompt Optimization for Personal Content Classification
di: Wang, Leijie, et al.
Pubblicazione: (2025)
di: Wang, Leijie, et al.
Pubblicazione: (2025)
Unveiling the Inter-Related Preferences of Crowdworkers: Implications for Personalized and Flexible Platform Design
di: Dutta, Senjuti, et al.
Pubblicazione: (2024)
di: Dutta, Senjuti, et al.
Pubblicazione: (2024)
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
"Actually I Can Count My Blessings": User-Centered Design of an Application to Promote Gratitude Among Young Adults
di: Bhattacharjee, Ananya, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Ananya, et al.
Pubblicazione: (2024)
Personalized Real-time Jargon Support for Online Meetings
di: Song, Yifan, et al.
Pubblicazione: (2025)
di: Song, Yifan, et al.
Pubblicazione: (2025)
CrowdGenUI: Aligning LLM-Based UI Generation with Crowdsourced User Preferences
di: Liu, Yimeng, et al.
Pubblicazione: (2024)
di: Liu, Yimeng, et al.
Pubblicazione: (2024)
Context-KG: Context-Aware Knowledge Graph Visualization with User Preferences and Ontological Guidance
di: Perera, Rumali, et al.
Pubblicazione: (2026)
di: Perera, Rumali, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
di: Rafailov, Rafael, et al.
Pubblicazione: (2023) -
RLVF: Learning from Verbal Feedback without Overgeneralization
di: Stephan, Moritz, et al.
Pubblicazione: (2024) -
Personalized Preference Fine-tuning of Diffusion Models
di: Dang, Meihua, et al.
Pubblicazione: (2025) -
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
di: Hsu, Sheryl, et al.
Pubblicazione: (2024) -
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
di: Manvi, Rohin, et al.
Pubblicazione: (2024)