Failure of contextual invariance in large language models
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Sagar, Flint, Ariel, Aiello, Luca Maria, Baronchelli, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Emergent social conventions and collective bias in LLM populations
by: Ashery, Ariel Flint, et al.
Published: (2024)
by: Ashery, Ariel Flint, et al.
Published: (2024)
Group size effects and collective misalignment in LLM multi-agent systems
by: Flint, Ariel, et al.
Published: (2025)
by: Flint, Ariel, et al.
Published: (2025)
Reply to "Emergent LLM behaviors are observationally equivalent to data leakage"
by: Ashery, Ariel Flint, et al.
Published: (2025)
by: Ashery, Ariel Flint, et al.
Published: (2025)
Uncovering inequalities in new knowledge learning by large language models across different languages
by: Wang, Chenglong, et al.
Published: (2025)
by: Wang, Chenglong, et al.
Published: (2025)
What can large language models do for sustainable food?
by: Thomas, Anna T., et al.
Published: (2025)
by: Thomas, Anna T., et al.
Published: (2025)
AI-AI Bias: large language models favor communications generated by large language models
by: Laurito, Walter, et al.
Published: (2024)
by: Laurito, Walter, et al.
Published: (2024)
Developmental trajectories of decision making and affective dynamics in large language models
by: Wang, Zhihao, et al.
Published: (2025)
by: Wang, Zhihao, et al.
Published: (2025)
A closer look at how large language models trust humans: patterns and biases
by: Lerman, Valeria, et al.
Published: (2025)
by: Lerman, Valeria, et al.
Published: (2025)
A survey on fairness of large language models in e-commerce: progress, application, and challenge
by: Ren, Qingyang, et al.
Published: (2024)
by: Ren, Qingyang, et al.
Published: (2024)
Do Chinese models speak Chinese languages?
by: Wen-Yi, Andrea W, et al.
Published: (2025)
by: Wen-Yi, Andrea W, et al.
Published: (2025)
Can adversarial attacks by large language models be attributed?
by: Cebrian, Manuel, et al.
Published: (2024)
by: Cebrian, Manuel, et al.
Published: (2024)
Assessing the nature of large language models: A caution against anthropocentrism
by: Speed, Ann
Published: (2023)
by: Speed, Ann
Published: (2023)
A validity-guided workflow for robust large language model research in psychology
by: Lin, Zhicheng
Published: (2025)
by: Lin, Zhicheng
Published: (2025)
Evidence of a log scaling law for political persuasion with large language models
by: Hackenburg, Kobi, et al.
Published: (2024)
by: Hackenburg, Kobi, et al.
Published: (2024)
Retrieval-augmented reasoning with lean language models
by: Chan, Ryan Sze-Yin, et al.
Published: (2025)
by: Chan, Ryan Sze-Yin, et al.
Published: (2025)
Large language models in medicine: the potentials and pitfalls
by: Omiye, Jesutofunmi A., et al.
Published: (2023)
by: Omiye, Jesutofunmi A., et al.
Published: (2023)
The opportunities and risks of large language models in mental health
by: Lawrence, Hannah R., et al.
Published: (2024)
by: Lawrence, Hannah R., et al.
Published: (2024)
Can a large language model be a gaslighter?
by: Li, Wei, et al.
Published: (2024)
by: Li, Wei, et al.
Published: (2024)
Training language models to be warm and empathetic makes them less reliable and more sycophantic
by: Ibrahim, Lujain, et al.
Published: (2025)
by: Ibrahim, Lujain, et al.
Published: (2025)
Large language models can consistently generate high-quality content for election disinformation operations
by: Williams, Angus R., et al.
Published: (2024)
by: Williams, Angus R., et al.
Published: (2024)
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
by: De Duro, Edoardo Sebastiano, et al.
Published: (2024)
Redefining technology for indigenous languages
by: Fernandez-Sabido, Silvia, et al.
Published: (2025)
by: Fernandez-Sabido, Silvia, et al.
Published: (2025)
RealHarm: A Collection of Real-World Language Model Application Failures
by: Jeune, Pierre Le, et al.
Published: (2025)
by: Jeune, Pierre Le, et al.
Published: (2025)
The Price of Prompting: Profiling Energy Use in Large Language Models Inference
by: Husom, Erik Johannes, et al.
Published: (2024)
by: Husom, Erik Johannes, et al.
Published: (2024)
The role of interface design on prompt-mediated creativity in Generative AI
by: Torricelli, Maddalena, et al.
Published: (2023)
by: Torricelli, Maddalena, et al.
Published: (2023)
HumT DumT: Measuring and controlling human-like language in LLMs
by: Cheng, Myra, et al.
Published: (2025)
by: Cheng, Myra, et al.
Published: (2025)
A Path Towards Legal Autonomy: An interoperable and explainable approach to extracting, transforming, loading and computing legal information using large language models, expert systems and Bayesian networks
by: Constant, Axel, et al.
Published: (2024)
by: Constant, Axel, et al.
Published: (2024)
Implicit assessment of language learning during practice as accurate as explicit testing
by: Hou, Jue, et al.
Published: (2024)
by: Hou, Jue, et al.
Published: (2024)
Profiling learners' affective engagement: Emotion AI, intercultural pragmatics, and language learning
by: Godwin-Jones, Robert
Published: (2026)
by: Godwin-Jones, Robert
Published: (2026)
Predicting potentially abusive clauses in Chilean terms of services with natural language processing
by: Loeffler, Christoffer, et al.
Published: (2025)
by: Loeffler, Christoffer, et al.
Published: (2025)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
Bringing order into the realm of Transformer-based language models for artificial intelligence and law
by: Greco, Candida M., et al.
Published: (2023)
by: Greco, Candida M., et al.
Published: (2023)
Shaping New Norms for AI
by: Baronchelli, Andrea
Published: (2023)
by: Baronchelli, Andrea
Published: (2023)
First, do NOHARM: towards clinically safe large language models
by: Wu, David, et al.
Published: (2025)
by: Wu, David, et al.
Published: (2025)
Large Language Models are Zero-Shot Next Location Predictors
by: Beneduce, Ciro, et al.
Published: (2024)
by: Beneduce, Ciro, et al.
Published: (2024)
Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?
by: Fontana, Nicoló, et al.
Published: (2024)
by: Fontana, Nicoló, et al.
Published: (2024)
On the attribution of confidence to large language models
by: Keeling, Geoff, et al.
Published: (2024)
by: Keeling, Geoff, et al.
Published: (2024)
Emergent evaluation hubs in a decentralizing large language model ecosystem
by: Cebrian, Manuel, et al.
Published: (2025)
by: Cebrian, Manuel, et al.
Published: (2025)
The LLM Wears Prada: Analysing Gender Bias and Stereotypes through Online Shopping Data
by: Luca, Massimiliano, et al.
Published: (2025)
by: Luca, Massimiliano, et al.
Published: (2025)
RoMathExam: A Longitudinal Dataset of Romanian Math Exams (1895-2025) with a Seven-Decade Core (1957-2025)
by: Cuclea, Luca-Ncolae, et al.
Published: (2026)
by: Cuclea, Luca-Ncolae, et al.
Published: (2026)
Similar Items
-
Emergent social conventions and collective bias in LLM populations
by: Ashery, Ariel Flint, et al.
Published: (2024) -
Group size effects and collective misalignment in LLM multi-agent systems
by: Flint, Ariel, et al.
Published: (2025) -
Reply to "Emergent LLM behaviors are observationally equivalent to data leakage"
by: Ashery, Ariel Flint, et al.
Published: (2025) -
Uncovering inequalities in new knowledge learning by large language models across different languages
by: Wang, Chenglong, et al.
Published: (2025) -
What can large language models do for sustainable food?
by: Thomas, Anna T., et al.
Published: (2025)