Can AI Assist in Olympiad Coding
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Ren, Samuel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OODEval: Evaluating Large Language Models on Object-Oriented Design
von: Xiao, Bingxu, et al.
Veröffentlicht: (2026)
von: Xiao, Bingxu, et al.
Veröffentlicht: (2026)
AdaDec: A Uncertainty-Guided Lookahead Decoding Framework for LLM-Based Code Generation
von: He, Kaifeng, et al.
Veröffentlicht: (2025)
von: He, Kaifeng, et al.
Veröffentlicht: (2025)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026)
von: Souza, Débora, et al.
Veröffentlicht: (2026)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
REPOT: Recoverable Program-of-Thought via Checkpoint Repair
von: Mazaheri, Parsa
Veröffentlicht: (2026)
von: Mazaheri, Parsa
Veröffentlicht: (2026)
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2026)
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2026)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2025)
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2025)
RelRepair: Enhancing Automated Program Repair by Retrieving Relevant Code
von: Liu, Shunyu, et al.
Veröffentlicht: (2025)
von: Liu, Shunyu, et al.
Veröffentlicht: (2025)
How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval
von: Ashrafi, Nazmus
Veröffentlicht: (2026)
von: Ashrafi, Nazmus
Veröffentlicht: (2026)
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
von: Trooskens, Geert, et al.
Veröffentlicht: (2026)
Contrastive Learning-Enhanced Large Language Models for Monolith-to-Microservice Decomposition
von: Sellami, Khaled, et al.
Veröffentlicht: (2025)
von: Sellami, Khaled, et al.
Veröffentlicht: (2025)
Developer Challenges on Large Language Models: A Study of Stack Overflow and OpenAI Developer Forum Posts
von: Alam, Khairul, et al.
Veröffentlicht: (2024)
von: Alam, Khairul, et al.
Veröffentlicht: (2024)
Improving Existing Optimization Algorithms with LLMs
von: Sartori, Camilo Chacón, et al.
Veröffentlicht: (2025)
von: Sartori, Camilo Chacón, et al.
Veröffentlicht: (2025)
Plan with Code: Comparing approaches for robust NL to DSL generation
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
CIFE: Code Instruction-Following Evaluation
von: Gunnu, Sravani, et al.
Veröffentlicht: (2025)
von: Gunnu, Sravani, et al.
Veröffentlicht: (2025)
A Comparative Study of DSL Code Generation: Fine-Tuning vs. Optimized Retrieval Augmentation
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
von: Iscan, Mehmet
Veröffentlicht: (2026)
von: Iscan, Mehmet
Veröffentlicht: (2026)
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
von: Li, Xu, et al.
Veröffentlicht: (2026)
von: Li, Xu, et al.
Veröffentlicht: (2026)
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges
von: Darshan, Parth, et al.
Veröffentlicht: (2026)
von: Darshan, Parth, et al.
Veröffentlicht: (2026)
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023)
von: Da, Longchao, et al.
Veröffentlicht: (2023)
LLM-Assisted Translation of Legacy FORTRAN Codes to C++: A Cross-Platform Study
von: Ranasinghe, Nishath Rajiv, et al.
Veröffentlicht: (2025)
von: Ranasinghe, Nishath Rajiv, et al.
Veröffentlicht: (2025)
REMoH: A Reflective Evolution of Multi-objective Heuristics approach via Large Language Models
von: Forniés-Tabuenca, Diego, et al.
Veröffentlicht: (2025)
von: Forniés-Tabuenca, Diego, et al.
Veröffentlicht: (2025)
GALA: Multimodal Graph Alignment for Bug Localization in Automated Program Repair
von: Liu, Zhuoyao, et al.
Veröffentlicht: (2026)
von: Liu, Zhuoyao, et al.
Veröffentlicht: (2026)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
von: Bradbury, Jeremy S., et al.
Veröffentlicht: (2024)
von: Bradbury, Jeremy S., et al.
Veröffentlicht: (2024)
Automated Bug Triaging using Instruction-Tuned Large Language Models
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
von: Kiashemshaki, Kiana, et al.
Veröffentlicht: (2025)
TRACE: Temporal Rule-Anchored Chain-of-Evidence on Knowledge Graphs for Interpretable Stock Movement Prediction
von: Ding, Qianggang, et al.
Veröffentlicht: (2026)
von: Ding, Qianggang, et al.
Veröffentlicht: (2026)
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
SiliconMind-V1: Multi-Agent Distillation and Debug-Reasoning Workflows for Verilog Code Generation
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2026)
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2026)
Critical Insights into Leading Conversational AI Models
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
Towards Observation Lakehouses: Living, Interactive Archives of Software Behavior
von: Kessel, Marcus
Veröffentlicht: (2025)
von: Kessel, Marcus
Veröffentlicht: (2025)
Random-Key Algorithms for Optimizing Integrated Operating Room Scheduling
von: Vieira, Bruno Salezze, et al.
Veröffentlicht: (2025)
von: Vieira, Bruno Salezze, et al.
Veröffentlicht: (2025)
Survey Transfer Learning: Recycling Data with Silicon Responses
von: Amini, Ali
Veröffentlicht: (2025)
von: Amini, Ali
Veröffentlicht: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
GEML: A Grammar-based Evolutionary Machine Learning Approach for Design-Pattern Detection
von: Barbudo, Rafael, et al.
Veröffentlicht: (2024)
von: Barbudo, Rafael, et al.
Veröffentlicht: (2024)
DISC: Dynamic Decomposition Improves LLM Inference Scaling
von: Light, Jonathan, et al.
Veröffentlicht: (2025)
von: Light, Jonathan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OODEval: Evaluating Large Language Models on Object-Oriented Design
von: Xiao, Bingxu, et al.
Veröffentlicht: (2026) -
AdaDec: A Uncertainty-Guided Lookahead Decoding Framework for LLM-Based Code Generation
von: He, Kaifeng, et al.
Veröffentlicht: (2025) -
Toward Architecture-Aware Evaluation Metrics for LLM Agents
von: Souza, Débora, et al.
Veröffentlicht: (2026) -
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
von: Zolduoarrati, Elijah, et al.
Veröffentlicht: (2025) -
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)