How do Humans and LLMs Process Confusing Code?
Fuente:
arXiv
Saved in:
| Main Authors: | Abdelsalam, Youssef, Peitek, Norman, Maurer, Anna-Maria, Toneva, Mariya, Apel, Sven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Restructuring to Stabilization: A Large-Scale Experiment on Iterative Code Readability Refactoring with Large Language Models
by: Peitek, Norman, et al.
Published: (2026)
by: Peitek, Norman, et al.
Published: (2026)
Harnessing Hype to Teach Empirical Thinking: An Experience With AI Coding Assistants
by: Wyrich, Marvin, et al.
Published: (2026)
by: Wyrich, Marvin, et al.
Published: (2026)
Fixation-related potentials reveal that confusing program code elicits a late frontal positivity
by: Bergum, Annabelle, et al.
Published: (2024)
by: Bergum, Annabelle, et al.
Published: (2024)
Hints Help Finding and Fixing Bugs Differently in Python and Text-based Program Representations
by: Rawal, Ruchit, et al.
Published: (2024)
by: Rawal, Ruchit, et al.
Published: (2024)
Evidence Tetris in the Pixelated World of Validity Threats
by: Wyrich, Marvin, et al.
Published: (2024)
by: Wyrich, Marvin, et al.
Published: (2024)
Multi-Location Software Model Completion
by: Welter, Alisa, et al.
Published: (2026)
by: Welter, Alisa, et al.
Published: (2026)
Programming Language Confusion: When Code LLMs Can't Keep their Languages Straight
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
Pragmatic Reasoning improves LLM Code Generation
by: Cao, Zhuchen, et al.
Published: (2025)
by: Cao, Zhuchen, et al.
Published: (2025)
The Silent Scientist: When Software Research Fails to Reach Its Audience
by: Wyrich, Marvin, et al.
Published: (2025)
by: Wyrich, Marvin, et al.
Published: (2025)
Detecting Performance-Relevant Changes in Configurable Software Systems
by: Böhm, Sebastian, et al.
Published: (2025)
by: Böhm, Sebastian, et al.
Published: (2025)
Software Engineering Podcasts: An Empirical Study of Their Potential as a Research Resource
by: Wyrich, Marvin, et al.
Published: (2026)
by: Wyrich, Marvin, et al.
Published: (2026)
Do AI Coding Agents Log Like Humans? An Empirical Study
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
Automata Learning -- Expect Delays!
by: Dengler, Gabriel, et al.
Published: (2025)
by: Dengler, Gabriel, et al.
Published: (2025)
How Quantization Impacts Privacy Risk on LLMs for Code?
by: Haque, Md Nazmul, et al.
Published: (2025)
by: Haque, Md Nazmul, et al.
Published: (2025)
Stalled, Biased, and Confused: Uncovering Reasoning Failures in LLMs for Cloud-Based Root Cause Analysis
by: Riddell, Evelien, et al.
Published: (2026)
by: Riddell, Evelien, et al.
Published: (2026)
The Impact of Large Language Models (LLMs) on Code Review Process
by: Collante, Antonio, et al.
Published: (2025)
by: Collante, Antonio, et al.
Published: (2025)
From Developer Pairs to AI Copilots: A Comparative Study on Knowledge Transfer
by: Welter, Alisa, et al.
Published: (2025)
by: Welter, Alisa, et al.
Published: (2025)
Studying How Configurations Impact Code Generation in LLMs: the Case of ChatGPT
by: Donato, Benedetta, et al.
Published: (2025)
by: Donato, Benedetta, et al.
Published: (2025)
Perplexed: Understanding When Large Language Models are Confused
by: Cooper, Nathan, et al.
Published: (2024)
by: Cooper, Nathan, et al.
Published: (2024)
Model-Assisted and Human-Guided: Perceptions and Practices of Software Professionals Using LLMs for Coding
by: Santos, Italo, et al.
Published: (2025)
by: Santos, Italo, et al.
Published: (2025)
HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent
by: Wu, Jie JW, et al.
Published: (2024)
by: Wu, Jie JW, et al.
Published: (2024)
Humanity's Last Code Exam: Can Advanced LLMs Conquer Human's Hardest Code Competition?
by: Li, Xiangyang, et al.
Published: (2025)
by: Li, Xiangyang, et al.
Published: (2025)
Software Model Evolution with Large Language Models: Experiments on Simulated, Public, and Industrial Datasets
by: Tinnes, Christof, et al.
Published: (2024)
by: Tinnes, Christof, et al.
Published: (2024)
Can Causality Cure Confusion Caused By Correlation (in Software Analytics)?
by: Rayegan, Amirali, et al.
Published: (2026)
by: Rayegan, Amirali, et al.
Published: (2026)
How to Compare the Security of Code Written by Humans to LLM-generated Code
by: Balebako, Rebecca, et al.
Published: (2026)
by: Balebako, Rebecca, et al.
Published: (2026)
Can LLMs Replace Humans During Code Chunking?
by: Glasz, Christopher, et al.
Published: (2025)
by: Glasz, Christopher, et al.
Published: (2025)
Neuron-Guided Interpretation of Code LLMs: Where, Why, and How?
by: Yin, Zhe, et al.
Published: (2025)
by: Yin, Zhe, et al.
Published: (2025)
Model Editing for LLMs4Code: How Far are We?
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
Human and Machine: How Software Engineers Perceive and Engage with AI-Assisted Code Reviews Compared to Their Peers
by: Alami, Adam, et al.
Published: (2025)
by: Alami, Adam, et al.
Published: (2025)
Compiling Code LLMs into Lightweight Executables
by: Shi, Jieke, et al.
Published: (2026)
by: Shi, Jieke, et al.
Published: (2026)
Migrating Code At Scale With LLMs At Google
by: Ziftci, Celal, et al.
Published: (2025)
by: Ziftci, Celal, et al.
Published: (2025)
Don't Confuse! Redrawing GUI Navigation Flow in Mobile Apps for Visually Impaired Users
by: Zhang, Mengxi, et al.
Published: (2025)
by: Zhang, Mengxi, et al.
Published: (2025)
Capturing the Effects of Quantization on Trojans in Code LLMs
by: Hussain, Aftab, et al.
Published: (2025)
by: Hussain, Aftab, et al.
Published: (2025)
Do Code LLMs Do Static Analysis?
by: Su, Chia-Yi, et al.
Published: (2025)
by: Su, Chia-Yi, et al.
Published: (2025)
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
On Evaluating the Efficiency of Source Code Generated by LLMs
by: Niu, Changan, et al.
Published: (2024)
by: Niu, Changan, et al.
Published: (2024)
MUCOCO: Automated Consistency Testing of Code LLMs
by: Chou, Chua Jin, et al.
Published: (2026)
by: Chou, Chua Jin, et al.
Published: (2026)
SWE-Bench+: Enhanced Coding Benchmark for LLMs
by: Aleithan, Reem, et al.
Published: (2024)
by: Aleithan, Reem, et al.
Published: (2024)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
by: Manh, Dung Nguyen, et al.
Published: (2024)
by: Manh, Dung Nguyen, et al.
Published: (2024)
Codev-Bench: How Do LLMs Understand Developer-Centric Code Completion?
by: Pan, Zhenyu, et al.
Published: (2024)
by: Pan, Zhenyu, et al.
Published: (2024)
Similar Items
-
From Restructuring to Stabilization: A Large-Scale Experiment on Iterative Code Readability Refactoring with Large Language Models
by: Peitek, Norman, et al.
Published: (2026) -
Harnessing Hype to Teach Empirical Thinking: An Experience With AI Coding Assistants
by: Wyrich, Marvin, et al.
Published: (2026) -
Fixation-related potentials reveal that confusing program code elicits a late frontal positivity
by: Bergum, Annabelle, et al.
Published: (2024) -
Hints Help Finding and Fixing Bugs Differently in Python and Text-based Program Representations
by: Rawal, Ruchit, et al.
Published: (2024) -
Evidence Tetris in the Pixelated World of Validity Threats
by: Wyrich, Marvin, et al.
Published: (2024)