PREDICT: Preference Reasoning by Evaluating Decomposed preferences Inferred from Candidate Trajectories
Fuente:
arXiv
Saved in:
| Main Authors: | Aroca-Ouellette, Stephane, Mackraz, Natalie, Theobald, Barry-John, Metcalf, Katherine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Aligning LLMs by Predicting Preferences from User Writing Samples
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024)
by: Metcalf, Katherine, et al.
Published: (2024)
Eyes on the Game: Deciphering Implicit Human Signals to Infer Human Proficiency, Trust, and Intent
by: Hulle, Nikhil, et al.
Published: (2024)
by: Hulle, Nikhil, et al.
Published: (2024)
FARPLS: A Feature-Augmented Robot Trajectory Preference Labeling System to Assist Human Labelers' Preference Elicitation
by: Lyu, Hanfang, et al.
Published: (2024)
by: Lyu, Hanfang, et al.
Published: (2024)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
Learning to Plan with Personalized Preferences
by: Xu, Manjie, et al.
Published: (2025)
by: Xu, Manjie, et al.
Published: (2025)
Adult learners recall and recognition performance and affective feedback when learning from an AI-generated synthetic video
by: Li, Zoe Ruo-Yu, et al.
Published: (2024)
by: Li, Zoe Ruo-Yu, et al.
Published: (2024)
Everyone prefers human writers, including AI
by: Haverals, Wouter, et al.
Published: (2025)
by: Haverals, Wouter, et al.
Published: (2025)
Gaze-informed Signatures of Trust and Collaboration in Human-Autonomy Teams
by: Ries, Anthony J., et al.
Published: (2024)
by: Ries, Anthony J., et al.
Published: (2024)
Unpacking Human Preference for LLMs: Demographically Aware Evaluation with the HUMAINE Framework
by: Petrova, Nora, et al.
Published: (2026)
by: Petrova, Nora, et al.
Published: (2026)
Sampling Preferences Yields Simple Trustworthiness Scores
by: Steinle, Sean
Published: (2025)
by: Steinle, Sean
Published: (2025)
Modeling User Preferences via Brain-Computer Interfacing
by: Leiva, Luis A., et al.
Published: (2024)
by: Leiva, Luis A., et al.
Published: (2024)
Tell me why: Training preferences-based RL with human preferences and step-level explanations
by: Karalus, Jakob
Published: (2024)
by: Karalus, Jakob
Published: (2024)
Generative Framework for Personalized Persuasion: Inferring Causal, Counterfactual, and Latent Knowledge
by: Zeng, Donghuo, et al.
Published: (2025)
by: Zeng, Donghuo, et al.
Published: (2025)
Addressing and Visualizing Misalignments in Human Task-Solving Trajectories
by: Kim, Sejin, et al.
Published: (2024)
by: Kim, Sejin, et al.
Published: (2024)
Steerable Chatbots: Personalizing LLMs with Preference-Based Activation Steering
by: Bo, Jessica Y., et al.
Published: (2025)
by: Bo, Jessica Y., et al.
Published: (2025)
In Pursuit of Predictive Models of Human Preferences Toward AI Teammates
by: Siu, Ho Chit, et al.
Published: (2025)
by: Siu, Ho Chit, et al.
Published: (2025)
Problem Solving Through Human-AI Preference-Based Cooperation
by: Dutta, Subhabrata, et al.
Published: (2024)
by: Dutta, Subhabrata, et al.
Published: (2024)
Humans learn to prefer trustworthy AI over human partners
by: Jiang, Yaomin, et al.
Published: (2025)
by: Jiang, Yaomin, et al.
Published: (2025)
LAPPI: Interactive Optimization with LLM-Assisted Preference-Based Problem Instantiation
by: Kuroki, So, et al.
Published: (2025)
by: Kuroki, So, et al.
Published: (2025)
Through the Judge's Eyes: Inferred Thinking Traces Improve Reliability of LLM Raters
by: Zhang, Xingjian, et al.
Published: (2025)
by: Zhang, Xingjian, et al.
Published: (2025)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
by: Daynauth, Roland, et al.
Published: (2024)
by: Daynauth, Roland, et al.
Published: (2024)
Learning with Digital Agents: An Analysis based on the Activity Theory
by: Dolata, Mateusz, et al.
Published: (2024)
by: Dolata, Mateusz, et al.
Published: (2024)
Communication Styles and Reader Preferences of LLM and Human Experts in Explaining Health Information
by: Zhou, Jiawei, et al.
Published: (2025)
by: Zhou, Jiawei, et al.
Published: (2025)
MapAgent: Trajectory-Constructed Memory-Augmented Planning for Mobile Task Automation
by: Kong, Yi, et al.
Published: (2025)
by: Kong, Yi, et al.
Published: (2025)
Explorer: Scaling Exploration-driven Web Trajectory Synthesis for Multimodal Web Agents
by: Pahuja, Vardaan, et al.
Published: (2025)
by: Pahuja, Vardaan, et al.
Published: (2025)
Progressive Autonomy as Preference Learning: A Formalization of Trust Calibration for Agentic Tool Use
by: Ou, Changkun
Published: (2026)
by: Ou, Changkun
Published: (2026)
Interactive Reasoning: Visualizing and Controlling Chain-of-Thought Reasoning in Large Language Models
by: Pang, Rock Yuren, et al.
Published: (2025)
by: Pang, Rock Yuren, et al.
Published: (2025)
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
by: Vishwarupe, Varad, et al.
Published: (2026)
by: Vishwarupe, Varad, et al.
Published: (2026)
Designing and Evaluating Malinowski's Lens: An AI-Native Educational Game for Ethnographic Learning
by: Hoffmann, Michael, et al.
Published: (2025)
by: Hoffmann, Michael, et al.
Published: (2025)
The Role of AI in Peer Support for Young People: A Study of Preferences for Human- and AI-Generated Responses
by: Young, Jordyn, et al.
Published: (2024)
by: Young, Jordyn, et al.
Published: (2024)
Prioritize Economy or Climate Action? Investigating ChatGPT Response Differences Based on Inferred Political Orientation
by: Karadal, Pelin, et al.
Published: (2025)
by: Karadal, Pelin, et al.
Published: (2025)
GenAI Voice Mode in Programming Education
by: Jacobs, Sven, et al.
Published: (2025)
by: Jacobs, Sven, et al.
Published: (2025)
Learning Reward and Policy Jointly from Demonstration and Preference Improves Alignment
by: Li, Chenliang, et al.
Published: (2024)
by: Li, Chenliang, et al.
Published: (2024)
Lessons in Cooperation: A Qualitative Analysis of Driver Sentiments towards Real-Time Advisory Systems from a Driving Simulator User Study
by: Hasan, Aamir, et al.
Published: (2024)
by: Hasan, Aamir, et al.
Published: (2024)
Flows: Building Blocks of Reasoning and Collaborating AI
by: Josifoski, Martin, et al.
Published: (2023)
by: Josifoski, Martin, et al.
Published: (2023)
Guided Reasoning: A Non-Technical Introduction
by: Betz, Gregor
Published: (2024)
by: Betz, Gregor
Published: (2024)
Replicating Human Motivated Reasoning Studies with LLMs
by: Pate, Neeley, et al.
Published: (2026)
by: Pate, Neeley, et al.
Published: (2026)
Similar Items
-
Aligning LLMs by Predicting Preferences from User Writing Samples
by: Aroca-Ouellette, Stéphane, et al.
Published: (2025) -
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024) -
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024) -
Eyes on the Game: Deciphering Implicit Human Signals to Infer Human Proficiency, Trust, and Intent
by: Hulle, Nikhil, et al.
Published: (2024) -
FARPLS: A Feature-Augmented Robot Trajectory Preference Labeling System to Assist Human Labelers' Preference Elicitation
by: Lyu, Hanfang, et al.
Published: (2024)