Invisible failures in human-AI interactions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Potts, Christopher, Sudhof, Moritz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A paradox of AI fluency
von: Potts, Christopher, et al.
Veröffentlicht: (2026)
von: Potts, Christopher, et al.
Veröffentlicht: (2026)
Base Models Beat Aligned Models at Randomness and Creativity
von: West, Peter, et al.
Veröffentlicht: (2025)
von: West, Peter, et al.
Veröffentlicht: (2025)
Language models as tools for investigating the distinction between possible and impossible natural languages
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2023)
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2023)
Counterfactual Simulation Training for Chain-of-Thought Faithfulness
von: Hase, Peter, et al.
Veröffentlicht: (2026)
von: Hase, Peter, et al.
Veröffentlicht: (2026)
Translators as Invisible Teachers of AI: Copyright, Translation Memory, and the Political Economy of Linguistic Data
von: Yamada, Masaru
Veröffentlicht: (2026)
von: Yamada, Masaru
Veröffentlicht: (2026)
Invisible Languages of the LLM Universe
von: Khanna, Saurabh, et al.
Veröffentlicht: (2025)
von: Khanna, Saurabh, et al.
Veröffentlicht: (2025)
InvisibleBench: A Deployment Gate for Caregiving Relationship AI
von: Madad, Ali
Veröffentlicht: (2025)
von: Madad, Ali
Veröffentlicht: (2025)
The Invisible Hand of AI Libraries Shaping Open Source Projects and Communities
von: Esposito, Matteo, et al.
Veröffentlicht: (2026)
von: Esposito, Matteo, et al.
Veröffentlicht: (2026)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
von: Kallini, Julie, et al.
Veröffentlicht: (2025)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
von: Rozner, Josh, et al.
Veröffentlicht: (2021)
von: Rozner, Josh, et al.
Veröffentlicht: (2021)
Improving Pretraining Data Using Perplexity Correlations
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions
von: Boguraev, Sasha, et al.
Veröffentlicht: (2025)
von: Boguraev, Sasha, et al.
Veröffentlicht: (2025)
Demystifying Verbatim Memorization in Large Language Models
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
MQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions
von: Zhong, Zexuan, et al.
Veröffentlicht: (2023)
von: Zhong, Zexuan, et al.
Veröffentlicht: (2023)
Improved Representation Steering for Language Models
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2025)
CommVQA: Situating Visual Question Answering in Communicative Contexts
von: Naik, Nandita Shankar, et al.
Veröffentlicht: (2024)
von: Naik, Nandita Shankar, et al.
Veröffentlicht: (2024)
A foundation model for human-AI collaboration in medical literature mining
von: Wang, Zifeng, et al.
Veröffentlicht: (2025)
von: Wang, Zifeng, et al.
Veröffentlicht: (2025)
Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning
von: Yu, Qinan, et al.
Veröffentlicht: (2026)
von: Yu, Qinan, et al.
Veröffentlicht: (2026)
Fine-Tuning and Prompt Optimization: Two Great Steps that Work Better Together
von: Soylu, Dilara, et al.
Veröffentlicht: (2024)
von: Soylu, Dilara, et al.
Veröffentlicht: (2024)
AI shares emotion with humans across languages and cultures
von: Wu, Xiuwen, et al.
Veröffentlicht: (2025)
von: Wu, Xiuwen, et al.
Veröffentlicht: (2025)
Document Optimization for Black-Box Retrieval via Reinforcement Learning
von: Uzan, Omri, et al.
Veröffentlicht: (2026)
von: Uzan, Omri, et al.
Veröffentlicht: (2026)
Retrieval Augmented Spelling Correction for E-Commerce Applications
von: Guo, Xuan, et al.
Veröffentlicht: (2024)
von: Guo, Xuan, et al.
Veröffentlicht: (2024)
Interpretability at Scale: Identifying Causal Mechanisms in Alpaca
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2023)
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2023)
Clinical knowledge in LLMs does not translate to human interactions
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
Unveiling the Invisible: Captioning Videos with Metaphors
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
von: Kalarani, Abisek Rajakumar, et al.
Veröffentlicht: (2024)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
von: Arora, Aryaman, et al.
Veröffentlicht: (2024)
When AI companions become witty: Can human brain recognize AI-generated irony?
von: Rao, Xiaohui, et al.
Veröffentlicht: (2025)
von: Rao, Xiaohui, et al.
Veröffentlicht: (2025)
Building Efficient and Effective OpenQA Systems for Low-Resource Languages
von: Budur, Emrah, et al.
Veröffentlicht: (2024)
von: Budur, Emrah, et al.
Veröffentlicht: (2024)
Using AI to replicate human experimental results: a motion study
von: Castillo, Rosa Illan, et al.
Veröffentlicht: (2025)
von: Castillo, Rosa Illan, et al.
Veröffentlicht: (2025)
Psychologically Potent, Computationally Invisible: LLMs Generate Social-Comparison-Eliciting Posts They Fail to Detect
von: Zhao, Hua, et al.
Veröffentlicht: (2026)
von: Zhao, Hua, et al.
Veröffentlicht: (2026)
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning
von: She, Jingyuan Selena, et al.
Veröffentlicht: (2023)
von: She, Jingyuan Selena, et al.
Veröffentlicht: (2023)
I am a Strange Dataset: Metalinguistic Tests for Language Models
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
Everyone prefers human writers, including AI
von: Haverals, Wouter, et al.
Veröffentlicht: (2025)
von: Haverals, Wouter, et al.
Veröffentlicht: (2025)
Can AI mimic the human ability to define neologisms?
von: Georgiou, Georgios P.
Veröffentlicht: (2025)
von: Georgiou, Georgios P.
Veröffentlicht: (2025)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
Invisible Entropy: Towards Safe and Efficient Low-Entropy LLM Watermarking
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
MrT5: Dynamic Token Merging for Efficient Byte-level Language Models
von: Kallini, Julie, et al.
Veröffentlicht: (2024)
von: Kallini, Julie, et al.
Veröffentlicht: (2024)
Invisible Filters: Cultural Bias in Hiring Evaluations Using Large Language Models
von: Rao, Pooja S. B., et al.
Veröffentlicht: (2025)
von: Rao, Pooja S. B., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A paradox of AI fluency
von: Potts, Christopher, et al.
Veröffentlicht: (2026) -
Base Models Beat Aligned Models at Randomness and Creativity
von: West, Peter, et al.
Veröffentlicht: (2025) -
Language models as tools for investigating the distinction between possible and impossible natural languages
von: Kallini, Julie, et al.
Veröffentlicht: (2025) -
ReCOGS: How Incidental Details of a Logical Form Overshadow an Evaluation of Semantic Interpretation
von: Wu, Zhengxuan, et al.
Veröffentlicht: (2023) -
Counterfactual Simulation Training for Chain-of-Thought Faithfulness
von: Hase, Peter, et al.
Veröffentlicht: (2026)