SteerEval: A Framework for Evaluating Steerability with Natural Language Profiles for Recommendation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Joyce, Zhou, Weijie, Turnbull, Doug, Joachims, Thorsten |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language-Based User Profiles for Recommendation
by: Zhou, Joyce, et al.
Published: (2024)
by: Zhou, Joyce, et al.
Published: (2024)
FairEval: Evaluating Fairness in LLM-Based Recommendations with Personality Awareness
by: Sah, Chandan Kumar, et al.
Published: (2025)
by: Sah, Chandan Kumar, et al.
Published: (2025)
EHR-MCP: Real-world Evaluation of Clinical Information Retrieval by Large Language Models via Model Context Protocol
by: Masayoshi, Kanato, et al.
Published: (2025)
by: Masayoshi, Kanato, et al.
Published: (2025)
Unequal Opportunities: Examining the Bias in Geographical Recommendations by Large Language Models
by: Dudy, Shiran, et al.
Published: (2025)
by: Dudy, Shiran, et al.
Published: (2025)
Semantic Interaction for Narrative Map Sensemaking: An Insight-based Evaluation
by: Keith-Norambuena, Brian Felipe, et al.
Published: (2026)
by: Keith-Norambuena, Brian Felipe, et al.
Published: (2026)
Language Modelling Approaches to Adaptive Machine Translation
by: Moslem, Yasmin
Published: (2024)
by: Moslem, Yasmin
Published: (2024)
Task-Oriented Dialog Systems for the Senegalese Wolof Language
by: Mbaye, Derguene, et al.
Published: (2024)
by: Mbaye, Derguene, et al.
Published: (2024)
Rewriting Conversational Utterances with Instructed Large Language Models
by: Galimzhanova, Elnara, et al.
Published: (2024)
by: Galimzhanova, Elnara, et al.
Published: (2024)
Leveraging Large Language Models for Hybrid Workplace Decision Support
by: Kim, Yujin, et al.
Published: (2024)
by: Kim, Yujin, et al.
Published: (2024)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
What should I wear to a party in a Greek taverna? Evaluation for Conversational Agents in the Fashion Domain
by: Maronikolakis, Antonis, et al.
Published: (2024)
by: Maronikolakis, Antonis, et al.
Published: (2024)
Large Language Models as Conversational Movie Recommenders: A User Study
by: Sun, Ruixuan, et al.
Published: (2024)
by: Sun, Ruixuan, et al.
Published: (2024)
Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles
by: Haq, Irti, et al.
Published: (2026)
by: Haq, Irti, et al.
Published: (2026)
Decoupled Recommender Systems: Exploring Alternative Recommender Ecosystem Designs
by: Buhayh, Anas, et al.
Published: (2025)
by: Buhayh, Anas, et al.
Published: (2025)
Balancing Domestic and Global Perspectives: Evaluating Dual-Calibration and LLM-Generated Nudges for Diverse News Recommendation
by: Sun, Ruixuan, et al.
Published: (2026)
by: Sun, Ruixuan, et al.
Published: (2026)
EasyInstruct: An Easy-to-use Instruction Processing Framework for Large Language Models
by: Ou, Yixin, et al.
Published: (2024)
by: Ou, Yixin, et al.
Published: (2024)
Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap
by: Panda, Akash Kumar, et al.
Published: (2026)
by: Panda, Akash Kumar, et al.
Published: (2026)
Much of Geospatial Web Search Is Beyond Traditional GIS
by: Ilyankou, Ilya, et al.
Published: (2026)
by: Ilyankou, Ilya, et al.
Published: (2026)
Modelling and Classifying the Components of a Literature Review
by: Bolaños, Francisco, et al.
Published: (2025)
by: Bolaños, Francisco, et al.
Published: (2025)
NExT-Search: Rebuilding User Feedback Ecosystem for Generative AI Search
by: Dai, Sunhao, et al.
Published: (2025)
by: Dai, Sunhao, et al.
Published: (2025)
Improving Multi-Domain Task-Oriented Dialogue System with Offline Reinforcement Learning
by: Prajapat, Dharmendra, et al.
Published: (2024)
by: Prajapat, Dharmendra, et al.
Published: (2024)
Retail-GPT: leveraging Retrieval Augmented Generation (RAG) for building E-commerce Chat Assistants
by: de Freitas, Bruno Amaral Teixeira, et al.
Published: (2024)
by: de Freitas, Bruno Amaral Teixeira, et al.
Published: (2024)
From Surface Learning to Deep Understanding: A Grounded AI Tutoring System for Moodle
by: Ostrowska, Anna, et al.
Published: (2026)
by: Ostrowska, Anna, et al.
Published: (2026)
Causal Autoencoder-like Generation of Feedback Fuzzy Cognitive Maps with an LLM Agent
by: Panda, Akash Kumar, et al.
Published: (2025)
by: Panda, Akash Kumar, et al.
Published: (2025)
The Agentic Leash: Extracting Causal Feedback Fuzzy Cognitive Maps with LLMs
by: Panda, Akash Kumar, et al.
Published: (2025)
by: Panda, Akash Kumar, et al.
Published: (2025)
Are the confidence scores of reviewers consistent with the review content? Evidence from top conference proceedings in AI
by: Wu, Wenqing, et al.
Published: (2025)
by: Wu, Wenqing, et al.
Published: (2025)
ClinicalTrialsHub: Bridging Registries and Literature for Comprehensive Clinical Trial Access
by: Park, Jiwoo, et al.
Published: (2025)
by: Park, Jiwoo, et al.
Published: (2025)
Surfacing Isolated Learners with Outcome-Independent Mediation of Feedback between Teachers and Students Using AI
by: Park, Junsoo, et al.
Published: (2026)
by: Park, Junsoo, et al.
Published: (2026)
Detecting Deceptive Dark Patterns in E-commerce Platforms
by: Ramteke, Arya, et al.
Published: (2024)
by: Ramteke, Arya, et al.
Published: (2024)
Search Engines in an AI Era: The False Promise of Factual and Verifiable Source-Cited Responses
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
Towards End-to-End Open Conversational Machine Reading
by: Zhou, Sizhe, et al.
Published: (2022)
by: Zhou, Sizhe, et al.
Published: (2022)
Interactive Counterfactual Exploration of Algorithmic Harms in Recommender Systems
by: Ahn, Yongsu, et al.
Published: (2024)
by: Ahn, Yongsu, et al.
Published: (2024)
Visualization for Recommendation Explainability: A Survey and New Perspectives
by: Chatti, Mohamed Amine, et al.
Published: (2023)
by: Chatti, Mohamed Amine, et al.
Published: (2023)
RAMO: Retrieval-Augmented Generation for Enhancing MOOCs Recommendations
by: Rao, Jiarui, et al.
Published: (2024)
by: Rao, Jiarui, et al.
Published: (2024)
OKRA: an Explainable, Heterogeneous, Multi-Stakeholder Job Recommender System
by: Schellingerhout, Roan, et al.
Published: (2025)
by: Schellingerhout, Roan, et al.
Published: (2025)
A Novel Behavior-Based Recommendation System for E-commerce
by: Nozari, Reza Barzegar, et al.
Published: (2024)
by: Nozari, Reza Barzegar, et al.
Published: (2024)
Knowledge Sharing in Manufacturing using Large Language Models: User Evaluation and Model Benchmarking
by: Freire, Samuel Kernan, et al.
Published: (2024)
by: Freire, Samuel Kernan, et al.
Published: (2024)
Empathic Responding for Digital Interpersonal Emotion Regulation via Content Recommendation
by: Verma, Akriti, et al.
Published: (2024)
by: Verma, Akriti, et al.
Published: (2024)
Contrastive Learning Method for Sequential Recommendation based on Multi-Intention Disentanglement
by: Hu, Zeyu, et al.
Published: (2024)
by: Hu, Zeyu, et al.
Published: (2024)
Retrieve, Annotate, Evaluate, Repeat: Leveraging Multimodal LLMs for Large-Scale Product Retrieval Evaluation
by: Hosseini, Kasra, et al.
Published: (2024)
by: Hosseini, Kasra, et al.
Published: (2024)
Similar Items
-
Language-Based User Profiles for Recommendation
by: Zhou, Joyce, et al.
Published: (2024) -
FairEval: Evaluating Fairness in LLM-Based Recommendations with Personality Awareness
by: Sah, Chandan Kumar, et al.
Published: (2025) -
EHR-MCP: Real-world Evaluation of Clinical Information Retrieval by Large Language Models via Model Context Protocol
by: Masayoshi, Kanato, et al.
Published: (2025) -
Unequal Opportunities: Examining the Bias in Geographical Recommendations by Large Language Models
by: Dudy, Shiran, et al.
Published: (2025) -
Semantic Interaction for Narrative Map Sensemaking: An Insight-based Evaluation
by: Keith-Norambuena, Brian Felipe, et al.
Published: (2026)