STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Kai, He, Zihao, Shi, Taiwei, Lerman, Kristina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How Susceptible are Large Language Models to Ideological Manipulation?
by: Chen, Kai, et al.
Published: (2024)
by: Chen, Kai, et al.
Published: (2024)
Large Language Models Reveal Information Operation Goals, Tactics, and Narrative Frames
by: Burghardt, Keith, et al.
Published: (2024)
by: Burghardt, Keith, et al.
Published: (2024)
Improving and Assessing the Fidelity of Large Language Models Alignment to Online Communities
by: Chu, Minh Duc, et al.
Published: (2024)
by: Chu, Minh Duc, et al.
Published: (2024)
Large Language Models Help Reveal Unhealthy Diet and Body Concerns in Online Eating Disorders Communities
by: Chu, Minh Duc, et al.
Published: (2024)
by: Chu, Minh Duc, et al.
Published: (2024)
COMMUNITY-CROSS-INSTRUCT: Unsupervised Instruction Generation for Aligning Large Language Models to Online Communities
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Whose Emotions and Moral Sentiments Do Language Models Reflect?
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Safer-Instruct: Aligning Language Models with Automated Preference Data
by: Shi, Taiwei, et al.
Published: (2023)
by: Shi, Taiwei, et al.
Published: (2023)
NEO-BENCH: Evaluating Robustness of Large Language Models with Neologisms
by: Zheng, Jonathan, et al.
Published: (2024)
by: Zheng, Jonathan, et al.
Published: (2024)
STEER-ME: Assessing the Microeconomic Reasoning of Large Language Models
by: Raman, Narun, et al.
Published: (2025)
by: Raman, Narun, et al.
Published: (2025)
CPL-NoViD: Context-Aware Prompt-based Learning for Norm Violation Detection in Online Communities
by: He, Zihao, et al.
Published: (2023)
by: He, Zihao, et al.
Published: (2023)
STEER: Assessing the Economic Rationality of Large Language Models
by: Raman, Narun, et al.
Published: (2024)
by: Raman, Narun, et al.
Published: (2024)
DI-BENCH: Benchmarking Large Language Models on Dependency Inference with Testable Repositories at Scale
by: Zhang, Linghao, et al.
Published: (2025)
by: Zhang, Linghao, et al.
Published: (2025)
Evaluating the Prompt Steerability of Large Language Models
by: Miehling, Erik, et al.
Published: (2024)
by: Miehling, Erik, et al.
Published: (2024)
EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models
by: Wang, Zekun, et al.
Published: (2025)
by: Wang, Zekun, et al.
Published: (2025)
CaT-BENCH: Benchmarking Language Model Understanding of Causal and Temporal Dependencies in Plans
by: Lal, Yash Kumar, et al.
Published: (2024)
by: Lal, Yash Kumar, et al.
Published: (2024)
Aggregation Artifacts in Subjective Tasks Collapse Large Language Models' Posteriors
by: Chochlakis, Georgios, et al.
Published: (2024)
by: Chochlakis, Georgios, et al.
Published: (2024)
UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
by: Ji, Yifan, et al.
Published: (2026)
by: Ji, Yifan, et al.
Published: (2026)
The Strong Pull of Prior Knowledge in Large Language Models and Its Impact on Emotion Recognition
by: Chochlakis, Georgios, et al.
Published: (2024)
by: Chochlakis, Georgios, et al.
Published: (2024)
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models
by: Yoshitake, Michiko, et al.
Published: (2024)
by: Yoshitake, Michiko, et al.
Published: (2024)
Reading Between the Tweets: Deciphering Ideological Stances of Interconnected Mixed-Ideology Communities
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
SEAL: Steerable Reasoning Calibration of Large Language Models for Free
by: Chen, Runjin, et al.
Published: (2025)
by: Chen, Runjin, et al.
Published: (2025)
Harmful Speech Detection by Language Models Exhibits Gender-Queer Dialect Bias
by: Dorn, Rebecca, et al.
Published: (2024)
by: Dorn, Rebecca, et al.
Published: (2024)
UNIDOC-BENCH: A Unified Benchmark for Document-Centric Multimodal RAG
by: Peng, Xiangyu, et al.
Published: (2025)
by: Peng, Xiangyu, et al.
Published: (2025)
DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues
by: Jang, Kyochul, et al.
Published: (2025)
by: Jang, Kyochul, et al.
Published: (2025)
CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space
by: Hwang, Yeonjun, et al.
Published: (2026)
by: Hwang, Yeonjun, et al.
Published: (2026)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
by: Rajabi, Navid, et al.
Published: (2024)
by: Rajabi, Navid, et al.
Published: (2024)
ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning
by: Potamitis, Nearchos, et al.
Published: (2025)
by: Potamitis, Nearchos, et al.
Published: (2025)
AI Steerability 360: A Toolkit for Steering Large Language Models
by: Miehling, Erik, et al.
Published: (2026)
by: Miehling, Erik, et al.
Published: (2026)
CoAnnotating: Uncertainty-Guided Work Allocation between Human and Large Language Models for Data Annotation
by: Li, Minzhi, et al.
Published: (2023)
by: Li, Minzhi, et al.
Published: (2023)
Assessing the Impact of Conspiracy Theories Using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?
by: Jiang, Xue, et al.
Published: (2026)
by: Jiang, Xue, et al.
Published: (2026)
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
by: Chen, Kedi, et al.
Published: (2024)
by: Chen, Kedi, et al.
Published: (2024)
AR-BENCH: Benchmarking Legal Reasoning with Judgment Error Detection, Classification and Correction
by: Li, Yifei, et al.
Published: (2026)
by: Li, Yifei, et al.
Published: (2026)
The Hallucination Tax of Reinforcement Finetuning
by: Song, Linxin, et al.
Published: (2025)
by: Song, Linxin, et al.
Published: (2025)
Poor Alignment and Steerability of Large Language Models: Evidence from College Admission Essays
by: Lee, Jinsook, et al.
Published: (2025)
by: Lee, Jinsook, et al.
Published: (2025)
Don't Blame the Data, Blame the Model: Understanding Noise and Bias When Learning from Subjective Annotations
by: Anand, Abhishek, et al.
Published: (2024)
by: Anand, Abhishek, et al.
Published: (2024)
SciEval: A Multi-Level Large Language Model Evaluation Benchmark for Scientific Research
by: Sun, Liangtai, et al.
Published: (2023)
by: Sun, Liangtai, et al.
Published: (2023)
Leveraging Machine Learning to Identify Gendered Stereotypes and Body Image Concerns on Diet and Fitness Online Forums
by: Chu, Minh Duc, et al.
Published: (2024)
by: Chu, Minh Duc, et al.
Published: (2024)
PROFASR-BENCH: A Benchmark for Context-Conditioned ASR in High-Stakes Professional Speech
by: Piskala, Deepak Babu
Published: (2025)
by: Piskala, Deepak Babu
Published: (2025)
Larger Language Models Don't Care How You Think: Why Chain-of-Thought Prompting Fails in Subjective Tasks
by: Chochlakis, Georgios, et al.
Published: (2024)
by: Chochlakis, Georgios, et al.
Published: (2024)
Similar Items
-
How Susceptible are Large Language Models to Ideological Manipulation?
by: Chen, Kai, et al.
Published: (2024) -
Large Language Models Reveal Information Operation Goals, Tactics, and Narrative Frames
by: Burghardt, Keith, et al.
Published: (2024) -
Improving and Assessing the Fidelity of Large Language Models Alignment to Online Communities
by: Chu, Minh Duc, et al.
Published: (2024) -
Large Language Models Help Reveal Unhealthy Diet and Body Concerns in Online Eating Disorders Communities
by: Chu, Minh Duc, et al.
Published: (2024) -
COMMUNITY-CROSS-INSTRUCT: Unsupervised Instruction Generation for Aligning Large Language Models to Online Communities
by: He, Zihao, et al.
Published: (2024)