Humans overrely on overconfident language models, across languages
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Rathi, Neil, Jurafsky, Dan, Zhou, Kaitlyn |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
par: Zhou, Kaitlyn, et autres
Publié: (2024)
par: Zhou, Kaitlyn, et autres
Publié: (2024)
Generics in science communication: Misaligned interpretations across laypeople, scientists, and large language models
par: Peters, Uwe, et autres
Publié: (2026)
par: Peters, Uwe, et autres
Publié: (2026)
Medical large language models are easily distracted
par: Vishwanath, Krithik, et autres
Publié: (2025)
par: Vishwanath, Krithik, et autres
Publié: (2025)
Medication counseling with large language models: balancing flexibility and rigidity
par: Sabel, Joar, et autres
Publié: (2025)
par: Sabel, Joar, et autres
Publié: (2025)
Creativity Benchmark: A benchmark for marketing creativity for large language models
par: Bhat, Ninad, et autres
Publié: (2025)
par: Bhat, Ninad, et autres
Publié: (2025)
Can Unconfident LLM Annotations Be Used for Confident Conclusions?
par: Gligorić, Kristina, et autres
Publié: (2024)
par: Gligorić, Kristina, et autres
Publié: (2024)
The production of meaning in the processing of natural language
par: Agostino, Christopher J., et autres
Publié: (2026)
par: Agostino, Christopher J., et autres
Publié: (2026)
The role of large language models in UI/UX design: A systematic literature review
par: Ahmed, Ammar, et autres
Publié: (2025)
par: Ahmed, Ammar, et autres
Publié: (2025)
TouchAI: Exploring human-AI perceptual alignment in touch through language model representations
par: Zhong, Shu, et autres
Publié: (2024)
par: Zhong, Shu, et autres
Publié: (2024)
Automated stereotactic radiosurgery planning using a human-in-the-loop reasoning large language model agent
par: Nusrat, Humza, et autres
Publié: (2025)
par: Nusrat, Humza, et autres
Publié: (2025)
Assessing the nature of large language models: A caution against anthropocentrism
par: Speed, Ann
Publié: (2023)
par: Speed, Ann
Publié: (2023)
A technical curriculum on language-oriented artificial intelligence in translation and specialised communication
par: Krüger, Ralph
Publié: (2026)
par: Krüger, Ralph
Publié: (2026)
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring
par: Seßler, Kathrin, et autres
Publié: (2024)
par: Seßler, Kathrin, et autres
Publié: (2024)
A validity-guided workflow for robust large language model research in psychology
par: Lin, Zhicheng
Publié: (2025)
par: Lin, Zhicheng
Publié: (2025)
Evidence of a log scaling law for political persuasion with large language models
par: Hackenburg, Kobi, et autres
Publié: (2024)
par: Hackenburg, Kobi, et autres
Publié: (2024)
Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis
par: Armitage, Richard
Publié: (2025)
par: Armitage, Richard
Publié: (2025)
Relying on the Unreliable: The Impact of Language Models' Reluctance to Express Uncertainty
par: Zhou, Kaitlyn, et autres
Publié: (2024)
par: Zhou, Kaitlyn, et autres
Publié: (2024)
The opportunities and risks of large language models in mental health
par: Lawrence, Hannah R., et autres
Publié: (2024)
par: Lawrence, Hannah R., et autres
Publié: (2024)
To what extent is ChatGPT useful for language teacher lesson plan creation?
par: Dornburg, Alex, et autres
Publié: (2024)
par: Dornburg, Alex, et autres
Publié: (2024)
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology
par: De Duro, Edoardo Sebastiano, et autres
Publié: (2024)
par: De Duro, Edoardo Sebastiano, et autres
Publié: (2024)
Human-like object concept representations emerge naturally in multimodal large language models
par: Du, Changde, et autres
Publié: (2024)
par: Du, Changde, et autres
Publié: (2024)
VeriLA: A Human-Centered Evaluation Framework for Interpretable Verification of LLM Agent Failures
par: Sung, Yoo Yeon, et autres
Publié: (2025)
par: Sung, Yoo Yeon, et autres
Publié: (2025)
Towards a copilot in BIM authoring tool using a large language model-based agent for intelligent human-machine interaction
par: Du, Changyu, et autres
Publié: (2024)
par: Du, Changyu, et autres
Publié: (2024)
BRAIn: Bayesian Reward-conditioned Amortized Inference for natural language generation from feedback
par: Pandey, Gaurav, et autres
Publié: (2024)
par: Pandey, Gaurav, et autres
Publié: (2024)
Beyond touch-based human-machine interface: Control your machines in natural language by utilizing large language models and OPC UA
par: Hofmann, Bernd, et autres
Publié: (2025)
par: Hofmann, Bernd, et autres
Publié: (2025)
Ask don't tell: Reducing sycophancy in large language models
par: Dubois, Magda, et autres
Publié: (2026)
par: Dubois, Magda, et autres
Publié: (2026)
From keywords to semantics: Perceptions of large language models in data discovery
par: Halstead, Maura E, et autres
Publié: (2025)
par: Halstead, Maura E, et autres
Publié: (2025)
Agentic publications: redesigning scientific publishing in the age of thinking large language models
par: Pugliese, Roberto, et autres
Publié: (2025)
par: Pugliese, Roberto, et autres
Publié: (2025)
Dyadic: A Scalable Platform for Human-Human and Human-AI Conversation Research
par: Markowitz, David M.
Publié: (2026)
par: Markowitz, David M.
Publié: (2026)
Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent Collaboration
par: Yao, Bingsheng, et autres
Publié: (2025)
par: Yao, Bingsheng, et autres
Publié: (2025)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
par: Huq, Faria, et autres
Publié: (2025)
par: Huq, Faria, et autres
Publié: (2025)
Performance Gains of LLMs With Humans in a World of LLMs Versus Humans
par: McCullum, Lucas, et autres
Publié: (2025)
par: McCullum, Lucas, et autres
Publié: (2025)
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
par: Gor, Maharshi, et autres
Publié: (2026)
par: Gor, Maharshi, et autres
Publié: (2026)
Human Bias in the Face of AI: Examining Human Judgment Against Text Labeled as AI Generated
par: Zhu, Tiffany, et autres
Publié: (2024)
par: Zhu, Tiffany, et autres
Publié: (2024)
Planning Ahead with RSA: Efficient Signalling in Dynamic Environments by Projecting User Awareness across Future Timesteps
par: Das, Anwesha, et autres
Publié: (2025)
par: Das, Anwesha, et autres
Publié: (2025)
Persona-E$^2$: A Human-Grounded Dataset for Personality-Shaped Emotional Responses to Textual Events
par: Yang, Yuqin, et autres
Publié: (2026)
par: Yang, Yuqin, et autres
Publié: (2026)
Addressing cognitive bias in medical language models
par: Schmidgall, Samuel, et autres
Publié: (2024)
par: Schmidgall, Samuel, et autres
Publié: (2024)
Game Development as Human-LLM Interaction
par: Hong, Jiale, et autres
Publié: (2024)
par: Hong, Jiale, et autres
Publié: (2024)
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
par: Wang, Zora Zhiruo, et autres
Publié: (2025)
par: Wang, Zora Zhiruo, et autres
Publié: (2025)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
par: Shen, Hua, et autres
Publié: (2025)
par: Shen, Hua, et autres
Publié: (2025)
Documents similaires
-
Rel-A.I.: An Interaction-Centered Approach To Measuring Human-LM Reliance
par: Zhou, Kaitlyn, et autres
Publié: (2024) -
Generics in science communication: Misaligned interpretations across laypeople, scientists, and large language models
par: Peters, Uwe, et autres
Publié: (2026) -
Medical large language models are easily distracted
par: Vishwanath, Krithik, et autres
Publié: (2025) -
Medication counseling with large language models: balancing flexibility and rigidity
par: Sabel, Joar, et autres
Publié: (2025) -
Creativity Benchmark: A benchmark for marketing creativity for large language models
par: Bhat, Ninad, et autres
Publié: (2025)