Saved in:
| Main Authors: | Kim, Junsol, Lai, Shiyang, Scherrer, Nino, Arcas, Blaise Agüera y, Evans, James |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.10825 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Linear Representations of Political Perspective Emerge in Large Language Models
by: Kim, Junsol, et al.
Published: (2025)
by: Kim, Junsol, et al.
Published: (2025)
Hidden Persuaders: LLMs' Political Leaning and Their Influence on Voters
by: Potter, Yujin, et al.
Published: (2024)
by: Potter, Yujin, et al.
Published: (2024)
Evolving AI Collectives to Enhance Human Diversity and Enable Self-Regulation
by: Lai, Shiyang, et al.
Published: (2024)
by: Lai, Shiyang, et al.
Published: (2024)
Social Learning: Towards Collaborative Learning with Large Language Models
by: Mohtashami, Amirkeivan, et al.
Published: (2023)
by: Mohtashami, Amirkeivan, et al.
Published: (2023)
The unreasonable effectiveness of pattern matching
by: Lupyan, Gary, et al.
Published: (2026)
by: Lupyan, Gary, et al.
Published: (2026)
Biased AI improves human decision-making but reduces trust
by: Lai, Shiyang, et al.
Published: (2025)
by: Lai, Shiyang, et al.
Published: (2025)
MesaNet: Sequence Modeling by Locally Optimal Test-Time Training
by: von Oswald, Johannes, et al.
Published: (2025)
by: von Oswald, Johannes, et al.
Published: (2025)
AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction
by: Kim, Junsol, et al.
Published: (2023)
by: Kim, Junsol, et al.
Published: (2023)
Can LLMs make trade-offs involving stipulated pain and pleasure states?
by: Keeling, Geoff, et al.
Published: (2024)
by: Keeling, Geoff, et al.
Published: (2024)
Inducing Group Fairness in Prompt-Based Language Model Decisions
by: Atwood, James, et al.
Published: (2024)
by: Atwood, James, et al.
Published: (2024)
Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models
by: Vishwamitra, Nishant, et al.
Published: (2023)
by: Vishwamitra, Nishant, et al.
Published: (2023)
Addressing Both Statistical and Causal Gender Fairness in NLP Models
by: Chen, Hannah, et al.
Published: (2024)
by: Chen, Hannah, et al.
Published: (2024)
What Lives? A meta-analysis of diverse opinions on the definition of life
by: Bender, Reed, et al.
Published: (2025)
by: Bender, Reed, et al.
Published: (2025)
Thought Crime: Backdoors and Emergent Misalignment in Reasoning Models
by: Chua, James, et al.
Published: (2025)
by: Chua, James, et al.
Published: (2025)
White-Box Sensitivity Auditing with Steering Vectors
by: Cyberey, Hannah, et al.
Published: (2026)
by: Cyberey, Hannah, et al.
Published: (2026)
M3Hop-CoT: Misogynous Meme Identification with Multimodal Multi-hop Chain-of-Thought
by: Kumari, Gitanjali, et al.
Published: (2024)
by: Kumari, Gitanjali, et al.
Published: (2024)
Social Determinants of Health Prediction for ICD-9 Code with Reasoning Models
by: Khan, Sharim, et al.
Published: (2025)
by: Khan, Sharim, et al.
Published: (2025)
KPoEM: A Human-Annotated Dataset for Emotion Classification and RAG-Based Poetry Generation in Korean Modern Poetry
by: Lim, Iro, et al.
Published: (2025)
by: Lim, Iro, et al.
Published: (2025)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
by: Dumitran, Adrian-Marius, et al.
Published: (2025)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
by: Rahman, Salman, et al.
Published: (2024)
by: Rahman, Salman, et al.
Published: (2024)
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring
by: Mohammadkhani, Ali Ghiasvand
Published: (2024)
by: Mohammadkhani, Ali Ghiasvand
Published: (2024)
Deliberative Alignment: Reasoning Enables Safer Language Models
by: Guan, Melody Y., et al.
Published: (2024)
by: Guan, Melody Y., et al.
Published: (2024)
Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communities
by: Jarvis, Devon, et al.
Published: (2026)
by: Jarvis, Devon, et al.
Published: (2026)
Personalized Decision Modeling: Utility Optimization or Textualized-Symbolic Reasoning
by: Zhao, Yibo, et al.
Published: (2025)
by: Zhao, Yibo, et al.
Published: (2025)
Augmenting Human-Annotated Training Data with Large Language Model Generation and Distillation in Open-Response Assessment
by: Borchers, Conrad, et al.
Published: (2025)
by: Borchers, Conrad, et al.
Published: (2025)
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
by: Liu, Ryan, et al.
Published: (2024)
by: Liu, Ryan, et al.
Published: (2024)
Evidence Is All You Need: Ordering Imaging Studies via Language Model Alignment with the ACR Appropriateness Criteria
by: Yao, Michael S., et al.
Published: (2024)
by: Yao, Michael S., et al.
Published: (2024)
The Reasoning Trap -- Logical Reasoning as a Mechanistic Pathway to Situational Awareness
by: Sahoo, Subramanyam, et al.
Published: (2026)
by: Sahoo, Subramanyam, et al.
Published: (2026)
GreedLlama: Performance of Financial Value-Aligned Large Language Models in Moral Reasoning
by: Yu, Jeffy, et al.
Published: (2024)
by: Yu, Jeffy, et al.
Published: (2024)
Fair Representation in Parliamentary Summaries: Measuring and Mitigating Inclusion Bias
by: Cunningham, Eoghan, et al.
Published: (2025)
by: Cunningham, Eoghan, et al.
Published: (2025)
A Practical Method for Generating String Counterfactuals
by: Avitan, Matan, et al.
Published: (2024)
by: Avitan, Matan, et al.
Published: (2024)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
by: Bohacek, Maty, et al.
Published: (2025)
by: Bohacek, Maty, et al.
Published: (2025)
Hypothesis Generation with Large Language Models
by: Zhou, Yangqiaoyu, et al.
Published: (2024)
by: Zhou, Yangqiaoyu, et al.
Published: (2024)
AppellateGen: A Benchmark for Appellate Legal Judgment Generation
by: Yang, Hongkun, et al.
Published: (2026)
by: Yang, Hongkun, et al.
Published: (2026)
Correlated Errors in Large Language Models
by: Kim, Elliot, et al.
Published: (2025)
by: Kim, Elliot, et al.
Published: (2025)
Towards Modeling Learner Performance with Large Language Models
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Improving Socratic Question Generation using Data Augmentation and Preference Optimization
by: Kumar, Nischal Ashok, et al.
Published: (2024)
by: Kumar, Nischal Ashok, et al.
Published: (2024)
The Moral Foundations Reddit Corpus
by: Trager, Jackson, et al.
Published: (2022)
by: Trager, Jackson, et al.
Published: (2022)
Reflecting in the Reflection: Integrating a Socratic Questioning Framework into Automated AI-Based Question Generation
by: Holub, Ondřej, et al.
Published: (2026)
by: Holub, Ondřej, et al.
Published: (2026)
DiVERT: Distractor Generation with Variational Errors Represented as Text for Math Multiple-choice Questions
by: Fernandez, Nigel, et al.
Published: (2024)
by: Fernandez, Nigel, et al.
Published: (2024)
Similar Items
-
Linear Representations of Political Perspective Emerge in Large Language Models
by: Kim, Junsol, et al.
Published: (2025) -
Hidden Persuaders: LLMs' Political Leaning and Their Influence on Voters
by: Potter, Yujin, et al.
Published: (2024) -
Evolving AI Collectives to Enhance Human Diversity and Enable Self-Regulation
by: Lai, Shiyang, et al.
Published: (2024) -
Social Learning: Towards Collaborative Learning with Large Language Models
by: Mohtashami, Amirkeivan, et al.
Published: (2023) -
The unreasonable effectiveness of pattern matching
by: Lupyan, Gary, et al.
Published: (2026)