Accounting for Sycophancy in Language Model Uncertainty Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Sicilia, Anthony, Inan, Mert, Alikhani, Malihe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Generating Signed Language Instructions in Large-Scale Dialogue Systems
por: İnan, Mert, et al.
Publicado: (2024)
por: İnan, Mert, et al.
Publicado: (2024)
Identifying & Interactively Refining Ambiguous User Goals for Data Visualization Code Generation
por: İnan, Mert, et al.
Publicado: (2025)
por: İnan, Mert, et al.
Publicado: (2025)
SiLVERScore: Semantically-Aware Embeddings for Sign Language Generation Evaluation
por: Imai, Saki, et al.
Publicado: (2025)
por: Imai, Saki, et al.
Publicado: (2025)
Measuring How (Not Just Whether) VLMs Build Common Ground
por: Imai, Saki, et al.
Publicado: (2025)
por: Imai, Saki, et al.
Publicado: (2025)
Better Slow than Sorry: Introducing Positive Friction for Reliable Dialogue Systems
por: İnan, Mert, et al.
Publicado: (2025)
por: İnan, Mert, et al.
Publicado: (2025)
BASIL: Bayesian Assessment of Sycophancy in LLMs
por: Atwell, Katherine, et al.
Publicado: (2025)
por: Atwell, Katherine, et al.
Publicado: (2025)
Eliciting Uncertainty in Chain-of-Thought to Mitigate Bias against Forecasting Harmful User Behaviors
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
Evaluating Theory of (an uncertain) Mind: Predicting the Uncertain Beliefs of Others in Conversation Forecasting
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
An Evaluation of Estimative Uncertainty in Large Language Models
por: Tang, Zhisheng, et al.
Publicado: (2024)
por: Tang, Zhisheng, et al.
Publicado: (2024)
Learning Multimodal Cues of Children's Uncertainty
por: Cheng, Qi, et al.
Publicado: (2024)
por: Cheng, Qi, et al.
Publicado: (2024)
HumBEL: A Human-in-the-Loop Approach for Evaluating Demographic Factors of Language Models in Human-Machine Conversations
por: Sicilia, Anthony, et al.
Publicado: (2023)
por: Sicilia, Anthony, et al.
Publicado: (2023)
Human-centered explanation does not fit all: The interplay of sociotechnical, cognitive, and individual factors in the effect AI explanations in algorithmic decision-making
por: Ahn, Yongsu, et al.
Publicado: (2025)
por: Ahn, Yongsu, et al.
Publicado: (2025)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
por: Zhou, Kaitlyn, et al.
Publicado: (2024)
por: Zhou, Kaitlyn, et al.
Publicado: (2024)
Modeling Intensification for Sign Language Generation: A Computational Approach
por: İnan, Mert, et al.
Publicado: (2022)
por: İnan, Mert, et al.
Publicado: (2022)
Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks
por: Kasneci, Enkelejda, et al.
Publicado: (2026)
por: Kasneci, Enkelejda, et al.
Publicado: (2026)
Deal, or no deal (or who knows)? Forecasting Uncertainty in Conversations using Large Language Models
por: Sicilia, Anthony, et al.
Publicado: (2024)
por: Sicilia, Anthony, et al.
Publicado: (2024)
Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models
por: Maltbie, Benjamin, et al.
Publicado: (2026)
por: Maltbie, Benjamin, et al.
Publicado: (2026)
SycoEval-EM: Sycophancy Evaluation of Large Language Models in Simulated Clinical Encounters for Emergency Care
por: Peng, Dongshen, et al.
Publicado: (2026)
por: Peng, Dongshen, et al.
Publicado: (2026)
A Scalable Framework for Evaluating Health Language Models
por: Mallinar, Neil, et al.
Publicado: (2025)
por: Mallinar, Neil, et al.
Publicado: (2025)
Generative Interfaces for Language Models
por: Chen, Jiaqi, et al.
Publicado: (2025)
por: Chen, Jiaqi, et al.
Publicado: (2025)
Status Hierarchies in Language Models
por: Barkett, Emilio
Publicado: (2026)
por: Barkett, Emilio
Publicado: (2026)
Epistemic Integrity in Large Language Models
por: Ghafouri, Bijean, et al.
Publicado: (2024)
por: Ghafouri, Bijean, et al.
Publicado: (2024)
Evaluating the Prompt Steerability of Large Language Models
por: Miehling, Erik, et al.
Publicado: (2024)
por: Miehling, Erik, et al.
Publicado: (2024)
The Art of Saying No: Contextual Noncompliance in Language Models
por: Brahman, Faeze, et al.
Publicado: (2024)
por: Brahman, Faeze, et al.
Publicado: (2024)
Autonomous Prompt Engineering in Large Language Models
por: Kepel, Daan, et al.
Publicado: (2024)
por: Kepel, Daan, et al.
Publicado: (2024)
Multi-agent KTO: Reinforcing Strategic Interactions of Large Language Model in Language Game
por: Ye, Rong, et al.
Publicado: (2025)
por: Ye, Rong, et al.
Publicado: (2025)
Prompt2DeModel: Declarative Neuro-Symbolic Modeling with Natural Language
por: Faghihi, Hossein Rajaby, et al.
Publicado: (2024)
por: Faghihi, Hossein Rajaby, et al.
Publicado: (2024)
Inertia in Moral and Value Judgments of Large Language Models
por: Lee, Bruce W., et al.
Publicado: (2024)
por: Lee, Bruce W., et al.
Publicado: (2024)
Large Language Models and Games: A Survey and Roadmap
por: Gallotta, Roberto, et al.
Publicado: (2024)
por: Gallotta, Roberto, et al.
Publicado: (2024)
(Ir)rationality and Cognitive Biases in Large Language Models
por: Macmillan-Scott, Olivia, et al.
Publicado: (2024)
por: Macmillan-Scott, Olivia, et al.
Publicado: (2024)
Evaluating Large Language Models in Analysing Classroom Dialogue
por: Long, Yun, et al.
Publicado: (2024)
por: Long, Yun, et al.
Publicado: (2024)
Automated Interpretability and Feature Discovery in Language Models with Agents
por: Marin-Llobet, Arnau, et al.
Publicado: (2026)
por: Marin-Llobet, Arnau, et al.
Publicado: (2026)
Large Language Model Use Impact Locus of Control
por: Fu, Jenny Xiyu, et al.
Publicado: (2025)
por: Fu, Jenny Xiyu, et al.
Publicado: (2025)
Empowering Private Tutoring by Chaining Large Language Models
por: Chen, Yulin, et al.
Publicado: (2023)
por: Chen, Yulin, et al.
Publicado: (2023)
Chain of Empathy: Enhancing Empathetic Response of Large Language Models Based on Psychotherapy Models
por: Lee, Yoon Kyung, et al.
Publicado: (2023)
por: Lee, Yoon Kyung, et al.
Publicado: (2023)
RNR: Teaching Large Language Models to Follow Roles and Rules
por: Wang, Kuan, et al.
Publicado: (2024)
por: Wang, Kuan, et al.
Publicado: (2024)
Automatic Histograms: Leveraging Language Models for Text Dataset Exploration
por: Reif, Emily, et al.
Publicado: (2024)
por: Reif, Emily, et al.
Publicado: (2024)
Large Language Models for Automatic Milestone Detection in Group Discussions
por: Duan, Zhuoxu, et al.
Publicado: (2024)
por: Duan, Zhuoxu, et al.
Publicado: (2024)
Meta-Prompting: Enhancing Language Models with Task-Agnostic Scaffolding
por: Suzgun, Mirac, et al.
Publicado: (2024)
por: Suzgun, Mirac, et al.
Publicado: (2024)
Language Models in Dialogue: Conversational Maxims for Human-AI Interactions
por: Miehling, Erik, et al.
Publicado: (2024)
por: Miehling, Erik, et al.
Publicado: (2024)
Ejemplares similares
-
Generating Signed Language Instructions in Large-Scale Dialogue Systems
por: İnan, Mert, et al.
Publicado: (2024) -
Identifying & Interactively Refining Ambiguous User Goals for Data Visualization Code Generation
por: İnan, Mert, et al.
Publicado: (2025) -
SiLVERScore: Semantically-Aware Embeddings for Sign Language Generation Evaluation
por: Imai, Saki, et al.
Publicado: (2025) -
Measuring How (Not Just Whether) VLMs Build Common Ground
por: Imai, Saki, et al.
Publicado: (2025) -
Better Slow than Sorry: Introducing Positive Friction for Reliable Dialogue Systems
por: İnan, Mert, et al.
Publicado: (2025)