Humanity's Last Code Exam: Can Advanced LLMs Conquer Human's Hardest Code Competition?
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Xiangyang, Li, Xiaopeng, Dong, Kuicai, Zhang, Quanhu, Ruan, Rongju, Dai, Xinyi, Liu, Xiaoshuang, Xu, Shengchun, Wang, Yasheng, Tang, Ruiming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
by: Li, Qingyao, et al.
Published: (2024)
by: Li, Qingyao, et al.
Published: (2024)
ATGen: Adversarial Reinforcement Learning for Test Case Generation
by: Li, Qingyao, et al.
Published: (2025)
by: Li, Qingyao, et al.
Published: (2025)
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024)
by: Li, Xiangyang, et al.
Published: (2024)
Can LLMs Replace Humans During Code Chunking?
by: Glasz, Christopher, et al.
Published: (2025)
by: Glasz, Christopher, et al.
Published: (2025)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
by: Thillen, Alex, et al.
Published: (2026)
by: Thillen, Alex, et al.
Published: (2026)
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations
by: Rani, Pooja, et al.
Published: (2025)
by: Rani, Pooja, et al.
Published: (2025)
Human to Document, AI to Code: Comparing GenAI for Notebook Competitions
by: Settewong, Tasha, et al.
Published: (2025)
by: Settewong, Tasha, et al.
Published: (2025)
How do Humans and LLMs Process Confusing Code?
by: Abdelsalam, Youssef, et al.
Published: (2025)
by: Abdelsalam, Youssef, et al.
Published: (2025)
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation
by: Chen, Jingchang, et al.
Published: (2024)
by: Chen, Jingchang, et al.
Published: (2024)
CodeGRAG: Bridging the Gap between Natural Language and Programming Language via Graphical Retrieval Augmented Generation
by: Du, Kounianhua, et al.
Published: (2024)
by: Du, Kounianhua, et al.
Published: (2024)
DePro: Understanding the Role of LLMs in Debugging Competitive Programming Code
by: Parvez, Nabiha, et al.
Published: (2026)
by: Parvez, Nabiha, et al.
Published: (2026)
Model Editing for LLMs4Code: How Far are We?
by: Li, Xiaopeng, et al.
Published: (2024)
by: Li, Xiaopeng, et al.
Published: (2024)
SpecRover: Code Intent Extraction via LLMs
by: Ruan, Haifeng, et al.
Published: (2024)
by: Ruan, Haifeng, et al.
Published: (2024)
AutoCode: LLMs as Problem Setters for Competitive Programming
by: Zhou, Shang, et al.
Published: (2025)
by: Zhou, Shang, et al.
Published: (2025)
HumanEvalComm: Benchmarking the Communication Competence of Code Generation for LLMs and LLM Agent
by: Wu, Jie JW, et al.
Published: (2024)
by: Wu, Jie JW, et al.
Published: (2024)
Model-Assisted and Human-Guided: Perceptions and Practices of Software Professionals Using LLMs for Coding
by: Santos, Italo, et al.
Published: (2025)
by: Santos, Italo, et al.
Published: (2025)
AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
An Iterative Test-and-Repair Framework for Competitive Code Generation
by: Tang, Lingxiao, et al.
Published: (2026)
by: Tang, Lingxiao, et al.
Published: (2026)
RepoZero: Can LLMs Generate a Code Repository from Scratch?
by: Zhang, Zhaoxi, et al.
Published: (2026)
by: Zhang, Zhaoxi, et al.
Published: (2026)
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision?
by: Fan, Lishui, et al.
Published: (2026)
by: Fan, Lishui, et al.
Published: (2026)
Automatically Generating UI Code from Screenshot: A Divide-and-Conquer-Based Approach
by: Wan, Yuxuan, et al.
Published: (2024)
by: Wan, Yuxuan, et al.
Published: (2024)
Is LLM-Generated Code More Maintainable \& Reliable than Human-Written Code?
by: Molison, Alfred Santa, et al.
Published: (2025)
by: Molison, Alfred Santa, et al.
Published: (2025)
Do AI Coding Agents Log Like Humans? An Empirical Study
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
by: Ouatiti, Youssef Esseddiq, et al.
Published: (2026)
Human-Aligned Code Readability Assessment with Large Language Models
by: Ouédraogo, Wendkûuni C., et al.
Published: (2025)
by: Ouédraogo, Wendkûuni C., et al.
Published: (2025)
Human-AI Synergy in Agentic Code Review
by: Zhong, Suzhen, et al.
Published: (2026)
by: Zhong, Suzhen, et al.
Published: (2026)
The Effect of Code Obfuscation on Human Program Comprehension
by: Nguyen, Anh H. N., et al.
Published: (2026)
by: Nguyen, Anh H. N., et al.
Published: (2026)
PseudoBridge: Pseudo Code as the Bridge for Better Semantic and Logic Alignment in Code Retrieval
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
Can LLMs Deobfuscate Binary Code? A Systematic Analysis of Large Language Models into Pseudocode Deobfuscation
by: Hu, Li, et al.
Published: (2026)
by: Hu, Li, et al.
Published: (2026)
Automated and Context-Aware Code Documentation Leveraging Advanced LLMs
by: Sarker, Swapnil Sharma, et al.
Published: (2025)
by: Sarker, Swapnil Sharma, et al.
Published: (2025)
Learning to Align Human Code Preferences
by: Yin, Xin, et al.
Published: (2025)
by: Yin, Xin, et al.
Published: (2025)
How to Compare the Security of Code Written by Humans to LLM-generated Code
by: Balebako, Rebecca, et al.
Published: (2026)
by: Balebako, Rebecca, et al.
Published: (2026)
InsightQL: Advancing Human-Assisted Fuzzing with a Unified Code Database and Parameterized Query Interface
by: Gao, Wentao, et al.
Published: (2025)
by: Gao, Wentao, et al.
Published: (2025)
Do Machines and Humans Focus on Similar Code? Exploring Explainability of Large Language Models in Code Summarization
by: Li, Jiliang, et al.
Published: (2024)
by: Li, Jiliang, et al.
Published: (2024)
Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
Can LLMs be Effective Code Contributors? A Study on Open-source Projects
by: Chong, Chun Jie, et al.
Published: (2026)
by: Chong, Chun Jie, et al.
Published: (2026)
Programming Language Confusion: When Code LLMs Can't Keep their Languages Straight
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
by: Moumoula, Micheline Bénédicte, et al.
Published: (2025)
Enhancing Code LLMs with Reinforcement Learning in Code Generation: A Survey
by: Wang, Junqiao, et al.
Published: (2024)
by: Wang, Junqiao, et al.
Published: (2024)
On Evaluating the Efficiency of Source Code Generated by LLMs
by: Niu, Changan, et al.
Published: (2024)
by: Niu, Changan, et al.
Published: (2024)
Code for Machines, Not Just Humans: Quantifying AI-Friendliness with Code Health Metrics
by: Borg, Markus, et al.
Published: (2026)
by: Borg, Markus, et al.
Published: (2026)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
by: Zhu, Yuqi, et al.
Published: (2025)
by: Zhu, Yuqi, et al.
Published: (2025)
Similar Items
-
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
by: Li, Qingyao, et al.
Published: (2024) -
ATGen: Adversarial Reinforcement Learning for Test Case Generation
by: Li, Qingyao, et al.
Published: (2025) -
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
by: Li, Xiangyang, et al.
Published: (2024) -
Can LLMs Replace Humans During Code Chunking?
by: Glasz, Christopher, et al.
Published: (2025) -
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
by: Thillen, Alex, et al.
Published: (2026)