"Check My Work?": Measuring Sycophancy in a Simulated Educational Context
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Arvin, Chuck |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identifying Legal Holdings with LLMs: A Systematic Study of Performance, Scale, and Memorization
von: Arvin, Chuck
Veröffentlicht: (2025)
von: Arvin, Chuck
Veröffentlicht: (2025)
SWAY: A Counterfactual Computational Linguistic Approach to Measuring and Mitigating Sycophancy
von: Bhalla, Joy, et al.
Veröffentlicht: (2026)
von: Bhalla, Joy, et al.
Veröffentlicht: (2026)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
von: Christian, Brian, et al.
Veröffentlicht: (2026)
von: Christian, Brian, et al.
Veröffentlicht: (2026)
Sycophancy Claims about Language Models: The Missing Human-in-the-Loop
von: Batzner, Jan, et al.
Veröffentlicht: (2025)
von: Batzner, Jan, et al.
Veröffentlicht: (2025)
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
von: Batzner, Jan, et al.
Veröffentlicht: (2024)
von: Batzner, Jan, et al.
Veröffentlicht: (2024)
Large Language Models Require Curated Context for Reliable Political Fact-Checking -- Even with Reasoning and Web Search
von: DeVerna, Matthew R., et al.
Veröffentlicht: (2025)
von: DeVerna, Matthew R., et al.
Veröffentlicht: (2025)
Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking
von: Russo, Daniel, et al.
Veröffentlicht: (2024)
von: Russo, Daniel, et al.
Veröffentlicht: (2024)
CheckIfExist: Detecting Citation Hallucinations in the Era of AI-Generated Content
von: Abbonato, Diletta
Veröffentlicht: (2026)
von: Abbonato, Diletta
Veröffentlicht: (2026)
When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio Platforms
von: Chun, Chaewan, et al.
Veröffentlicht: (2026)
von: Chun, Chaewan, et al.
Veröffentlicht: (2026)
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
von: Jin, Yiping, et al.
Veröffentlicht: (2024)
LIBRA: Measuring Bias of Large Language Model from a Local Context
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
Facts are Harder Than Opinions -- A Multilingual, Comparative Analysis of LLM-Based Fact-Checking Reliability
von: Saju, Lorraine, et al.
Veröffentlicht: (2025)
von: Saju, Lorraine, et al.
Veröffentlicht: (2025)
Simulated Adoption: Decoupling Magnitude and Direction in LLM In-Context Conflict Resolution
von: Zhang, Long, et al.
Veröffentlicht: (2026)
von: Zhang, Long, et al.
Veröffentlicht: (2026)
Beyond Translation: LLM-Based Data Generation for Multilingual Fact-Checking
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
"Not in My Backyard": LLMs Uncover Online and Offline Social Biases Against Homelessness
von: Karr Jr., Jonathan A., et al.
Veröffentlicht: (2025)
von: Karr Jr., Jonathan A., et al.
Veröffentlicht: (2025)
Conversational Alignment with Artificial Intelligence in Context
von: Sterken, Rachel Katharine, et al.
Veröffentlicht: (2025)
von: Sterken, Rachel Katharine, et al.
Veröffentlicht: (2025)
Measuring Free-Form Decision-Making Inconsistency of Language Models in Military Crisis Simulations
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
von: Shrivastava, Aryan, et al.
Veröffentlicht: (2024)
Do Llamas Work in English? On the Latent Language of Multilingual Transformers
von: Wendler, Chris, et al.
Veröffentlicht: (2024)
von: Wendler, Chris, et al.
Veröffentlicht: (2024)
The Implications of Open Generative Models in Human-Centered Data Science Work: A Case Study with Fact-Checking Organizations
von: Wolfe, Robert, et al.
Veröffentlicht: (2024)
von: Wolfe, Robert, et al.
Veröffentlicht: (2024)
Measuring Sycophancy of Language Models in Multi-turn Dialogues
von: Hong, Jiseung, et al.
Veröffentlicht: (2025)
von: Hong, Jiseung, et al.
Veröffentlicht: (2025)
Not My Truce: Personality Differences in AI-Mediated Workplace Negotiation
von: Duddu, Veda, et al.
Veröffentlicht: (2026)
von: Duddu, Veda, et al.
Veröffentlicht: (2026)
MedSimAI: Simulation and Formative Feedback Generation to Enhance Deliberate Practice in Medical Education
von: Hicke, Yann, et al.
Veröffentlicht: (2025)
von: Hicke, Yann, et al.
Veröffentlicht: (2025)
Quantifying the Persona Effect in LLM Simulations
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
Simulated Students in Tutoring Dialogues: Substance or Illusion?
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2026)
von: Scarlatos, Alexander, et al.
Veröffentlicht: (2026)
Can Large Language Models Simulate Human Responses? A Case Study of Stated Preference Experiments in the Context of Heating-related Choices
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Counterfactual LLM-based Framework for Measuring Rhetorical Style
von: Qiu, Jingyi, et al.
Veröffentlicht: (2025)
von: Qiu, Jingyi, et al.
Veröffentlicht: (2025)
Simulating Students with Large Language Models: A Review of Architecture, Mechanisms, and Role Modelling in Education with Generative AI
von: Marquez-Carpintero, Luis, et al.
Veröffentlicht: (2025)
von: Marquez-Carpintero, Luis, et al.
Veröffentlicht: (2025)
Change My Frame: Reframing in the Wild in r/ChangeMyView
von: Peguero, Arturo Martínez, et al.
Veröffentlicht: (2024)
von: Peguero, Arturo Martínez, et al.
Veröffentlicht: (2024)
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
Evaluating the Simulation of Human Personality-Driven Susceptibility to Misinformation with LLMs
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
Improving Cross-Cultural Survey Simulation with Calibrated Value Personas
von: Abels, Axel, et al.
Veröffentlicht: (2026)
von: Abels, Axel, et al.
Veröffentlicht: (2026)
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
von: Faulborn, Mats, et al.
Veröffentlicht: (2025)
Measuring Fine-Grained Negotiation Tactics of Humans and LLMs in Diplomacy
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
Measuring Implicit Bias in Explicitly Unbiased Large Language Models
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
von: Bai, Xuechunzi, et al.
Veröffentlicht: (2024)
Measuring Large Language Models Capacity to Annotate Journalistic Sourcing
von: Vincent, Subramaniam, et al.
Veröffentlicht: (2024)
von: Vincent, Subramaniam, et al.
Veröffentlicht: (2024)
Measuring Opinion Bias and Sycophancy via LLM-based Persuasion
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs
von: Wang, Angelina, et al.
Veröffentlicht: (2025)
von: Wang, Angelina, et al.
Veröffentlicht: (2025)
CDEval: A Benchmark for Measuring the Cultural Dimensions of Large Language Models
von: Wang, Yuhang, et al.
Veröffentlicht: (2023)
von: Wang, Yuhang, et al.
Veröffentlicht: (2023)
LLMs for Low-Resource Dialect Translation Using Context-Aware Prompting: A Case Study on Sylheti
von: Prama, Tabia Tanzin, et al.
Veröffentlicht: (2025)
von: Prama, Tabia Tanzin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Identifying Legal Holdings with LLMs: A Systematic Study of Performance, Scale, and Memorization
von: Arvin, Chuck
Veröffentlicht: (2025) -
SWAY: A Counterfactual Computational Linguistic Approach to Measuring and Mitigating Sycophancy
von: Bhalla, Joy, et al.
Veröffentlicht: (2026) -
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026) -
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
von: Christian, Brian, et al.
Veröffentlicht: (2026) -
Sycophancy Claims about Language Models: The Missing Human-in-the-Loop
von: Batzner, Jan, et al.
Veröffentlicht: (2025)