mcdok at SemEval-2026 Task 13: Finetuning LLMs for Detection of Machine-Generated Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Skurla, Adam, Macko, Dominik, Simko, Jakub |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
mdok-style at SemEval-2026 Task 9: Finetuning LLMs for Multilingual Polarization Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
von: Macko, Dominik, et al.
Veröffentlicht: (2026)
mdok-style at SemEval-2026 Task 10: Finetuning LLMs for Conspiracy Detection
von: Macko, Dominik
Veröffentlicht: (2026)
von: Macko, Dominik
Veröffentlicht: (2026)
Interpretable Predictability-Based AI Text Detection: A Replication Study
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
von: Skurla, Adam, et al.
Veröffentlicht: (2026)
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
von: Spiegel, Michal, et al.
Veröffentlicht: (2024)
von: Spiegel, Michal, et al.
Veröffentlicht: (2024)
UCSC-NLP at SemEval-2026 Task 13: Multi-View Generalization and Diagnostic Analysis of Machine-Generated Code Detection
von: Chauhan, Kargi, et al.
Veröffentlicht: (2026)
von: Chauhan, Kargi, et al.
Veröffentlicht: (2026)
Evaluation of LLMs on Syntax-Aware Code Fill-in-the-Middle Tasks
von: Gong, Linyuan, et al.
Veröffentlicht: (2024)
von: Gong, Linyuan, et al.
Veröffentlicht: (2024)
LiCoEval: Evaluating LLMs on License Compliance in Code Generation
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
von: Xu, Weiwei, et al.
Veröffentlicht: (2024)
Rethinking Repetition Problems of LLMs in Code Generation
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
DevEval: Evaluating Code Generation in Practical Software Projects
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
StructEval: Benchmarking LLMs' Capabilities to Generate Structural Outputs
von: Yang, Jialin, et al.
Veröffentlicht: (2025)
von: Yang, Jialin, et al.
Veröffentlicht: (2025)
NoFunEval: Funny How Code LMs Falter on Requirements Beyond Functional Correctness
von: Singhal, Manav, et al.
Veröffentlicht: (2024)
von: Singhal, Manav, et al.
Veröffentlicht: (2024)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
von: Tambon, Florian, et al.
Veröffentlicht: (2024)
VHDL-Eval: A Framework for Evaluating Large Language Models in VHDL Code Generation
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2024)
von: Vijayaraghavan, Prashanth, et al.
Veröffentlicht: (2024)
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
von: Ding, Yangruibo, et al.
Veröffentlicht: (2024)
Code to Think, Think to Code: A Survey on Code-Enhanced Reasoning and Reasoning-Driven Code Intelligence in LLMs
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
von: Yang, Dayu, et al.
Veröffentlicht: (2025)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
von: Kadosh, Tal, et al.
Veröffentlicht: (2023)
von: Kadosh, Tal, et al.
Veröffentlicht: (2023)
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
von: Kammakomati, Mehant, et al.
Veröffentlicht: (2024)
Linguacodus: A Synergistic Framework for Transformative Code Generation in Machine Learning Pipelines
von: Trofimova, Ekaterina, et al.
Veröffentlicht: (2024)
von: Trofimova, Ekaterina, et al.
Veröffentlicht: (2024)
LyS at SemEval 2025 Task 8: Zero-Shot Code Generation for Tabular QA
von: Gude, Adrián, et al.
Veröffentlicht: (2025)
von: Gude, Adrián, et al.
Veröffentlicht: (2025)
MERA Code: A Unified Framework for Evaluating Code Generation Across Tasks
von: Chervyakov, Artem, et al.
Veröffentlicht: (2025)
von: Chervyakov, Artem, et al.
Veröffentlicht: (2025)
How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data
von: Wang, Yejie, et al.
Veröffentlicht: (2024)
von: Wang, Yejie, et al.
Veröffentlicht: (2024)
Code-Vision: Evaluating Multimodal LLMs Logic Understanding and Code Generation Capabilities
von: Wang, Hanbin, et al.
Veröffentlicht: (2025)
von: Wang, Hanbin, et al.
Veröffentlicht: (2025)
LLMs for Science: Usage for Code Generation and Data Analysis
von: Nejjar, Mohamed, et al.
Veröffentlicht: (2023)
von: Nejjar, Mohamed, et al.
Veröffentlicht: (2023)
ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
von: Fan, Lishui, et al.
Veröffentlicht: (2025)
CodeGeeX: A Pre-Trained Model for Code Generation with Multilingual Benchmarking on HumanEval-X
von: Zheng, Qinkai, et al.
Veröffentlicht: (2023)
von: Zheng, Qinkai, et al.
Veröffentlicht: (2023)
CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision
von: Lu, Yifei, et al.
Veröffentlicht: (2025)
von: Lu, Yifei, et al.
Veröffentlicht: (2025)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
Selective Prompt Anchoring for Code Generation
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
von: Dandamudi, Rohit, et al.
Veröffentlicht: (2024)
CodeVisionary: An Agent-based Framework for Evaluating Large Language Models in Code Generation
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
von: Wang, Xinchen, et al.
Veröffentlicht: (2025)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
CodeScope: An Execution-based Multilingual Multitask Multidimensional Benchmark for Evaluating LLMs on Code Understanding and Generation
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
von: Yan, Weixiang, et al.
Veröffentlicht: (2023)
To See is Not to Master: Teaching LLMs to Use Private Libraries for Code Generation
von: Zhang, Yitong, et al.
Veröffentlicht: (2026)
von: Zhang, Yitong, et al.
Veröffentlicht: (2026)
Investigating the Efficacy of Large Language Models for Code Clone Detection
von: Khajezade, Mohamad, et al.
Veröffentlicht: (2024)
von: Khajezade, Mohamad, et al.
Veröffentlicht: (2024)
StackEval: Benchmarking LLMs in Coding Assistance
von: Shah, Nidhish, et al.
Veröffentlicht: (2024)
von: Shah, Nidhish, et al.
Veröffentlicht: (2024)
Assessing Code Understanding in LLMs
von: Laneve, Cosimo, et al.
Veröffentlicht: (2025)
von: Laneve, Cosimo, et al.
Veröffentlicht: (2025)
Operational Robustness of LLMs on Code Generation
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
von: Paul, Debalina Ghosh, et al.
Veröffentlicht: (2026)
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
Improving Code Generation by Training with Natural Language Feedback
von: Chen, Angelica, et al.
Veröffentlicht: (2023)
von: Chen, Angelica, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
mdok-style at SemEval-2026 Task 9: Finetuning LLMs for Multilingual Polarization Detection
von: Macko, Dominik, et al.
Veröffentlicht: (2026) -
mdok-style at SemEval-2026 Task 10: Finetuning LLMs for Conspiracy Detection
von: Macko, Dominik
Veröffentlicht: (2026) -
Interpretable Predictability-Based AI Text Detection: A Replication Study
von: Skurla, Adam, et al.
Veröffentlicht: (2026) -
KInIT at SemEval-2024 Task 8: Fine-tuned LLMs for Multilingual Machine-Generated Text Detection
von: Spiegel, Michal, et al.
Veröffentlicht: (2024) -
UCSC-NLP at SemEval-2026 Task 13: Multi-View Generalization and Diagnostic Analysis of Machine-Generated Code Detection
von: Chauhan, Kargi, et al.
Veröffentlicht: (2026)