Is This You, LLM? Recognizing AI-written Programs with Multilingual Code Stylometry
Fuente:
arXiv
Saved in:
| Main Authors: | Gurioli, Andrea, Gabbrielli, Maurizio, Zacchiroli, Stefano |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
by: Gurioli, Andrea, et al.
Published: (2026)
by: Gurioli, Andrea, et al.
Published: (2026)
Do not copy and paste! Rewriting strategies for code retrieval
by: Gurioli, Andrea, et al.
Published: (2026)
by: Gurioli, Andrea, et al.
Published: (2026)
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
by: Gurioli, Andrea, et al.
Published: (2025)
by: Gurioli, Andrea, et al.
Published: (2025)
Wild SBOMs: a Large-scale Dataset of Software Bills of Materials from Public Code
by: Soeiro, Luıs, et al.
Published: (2025)
by: Soeiro, Luıs, et al.
Published: (2025)
On the Informativeness of Security Commit Messages: A Large-scale Replication Study
by: Islam, Syful, et al.
Published: (2026)
by: Islam, Syful, et al.
Published: (2026)
On the Use of Commit Messages for Corrective Software Maintenance: A Systematic Mapping Study
by: Islam, Syful, et al.
Published: (2026)
by: Islam, Syful, et al.
Published: (2026)
Source Code Archiving to the Rescue of Reproducible Deployment
by: Courtès, Ludovic, et al.
Published: (2024)
by: Courtès, Ludovic, et al.
Published: (2024)
University Rents Enabling Corporate Innovation: Mapping Academic Researcher Coding and Discursive Labour in the R Language Ecosystem
by: Cai, Xiaolan, et al.
Published: (2025)
by: Cai, Xiaolan, et al.
Published: (2025)
Agentic Much? Adoption of Coding Agents on GitHub
by: Robbes, Romain, et al.
Published: (2026)
by: Robbes, Romain, et al.
Published: (2026)
Promises, Perils, and (Timely) Heuristics for Mining Coding Agent Activity
by: Robbes, Romain, et al.
Published: (2026)
by: Robbes, Romain, et al.
Published: (2026)
Reproducibility of Build Environments through Space and Time
by: Malka, Julien, et al.
Published: (2024)
by: Malka, Julien, et al.
Published: (2024)
Bridging Behavioral Biometrics and Source Code Stylometry: A Survey of Programmer Attribution
by: Horvath, Marek, et al.
Published: (2026)
by: Horvath, Marek, et al.
Published: (2026)
Does Functional Package Management Enable Reproducible Builds at Scale? Yes
by: Malka, Julien, et al.
Published: (2025)
by: Malka, Julien, et al.
Published: (2025)
Docker Does Not Guarantee Reproducibility
by: Malka, Julien, et al.
Published: (2026)
by: Malka, Julien, et al.
Published: (2026)
The Impact of the COVID-19 Pandemic on Women's Contribution to Public Code
by: Casanueva, Annalí, et al.
Published: (2024)
by: Casanueva, Annalí, et al.
Published: (2024)
Did You Forkget It? Detecting One-Day Vulnerabilities in Open-source ForksWith Global History Analysis
by: Lefeuvre, Romain, et al.
Published: (2025)
by: Lefeuvre, Romain, et al.
Published: (2025)
Distinguishing LLM-generated from Human-written Code by Contrastive Learning
by: Xu, Xiaodan, et al.
Published: (2024)
by: Xu, Xiaodan, et al.
Published: (2024)
Altered Histories in Version Control System Repositories: Evidence from the Trenches
by: Rapaport, Solal, et al.
Published: (2025)
by: Rapaport, Solal, et al.
Published: (2025)
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
by: Bisztray, Tamas, et al.
Published: (2025)
by: Bisztray, Tamas, et al.
Published: (2025)
ROSA: Finding Backdoors with Fuzzing
by: Kokkonis, Dimitri, et al.
Published: (2025)
by: Kokkonis, Dimitri, et al.
Published: (2025)
What You Need is What You Get: Theory of Mind for an LLM-Based Code Understanding Assistant
by: Richards, Jonan, et al.
Published: (2024)
by: Richards, Jonan, et al.
Published: (2024)
Context-Aware CodeLLM Eviction for AI-assisted Coding
by: Thangarajah, Kishanthan, et al.
Published: (2025)
by: Thangarajah, Kishanthan, et al.
Published: (2025)
DRAGON: Robust Classification for Very Large Collections of Software Repositories
by: Balla, Stefano, et al.
Published: (2026)
by: Balla, Stefano, et al.
Published: (2026)
Development and Benchmarking of Multilingual Code Clone Detector
by: Zhu, Wenqing, et al.
Published: (2024)
by: Zhu, Wenqing, et al.
Published: (2024)
Error Understanding in Program Code With LLM-DL for Multi-label Classification
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
CodeCSE: A Simple Multilingual Model for Code and Comment Sentence Embeddings
by: Varkey, Anthony, et al.
Published: (2024)
by: Varkey, Anthony, et al.
Published: (2024)
AI-Powered, But Power-Hungry? Energy Efficiency of LLM-Generated Code
by: Solovyeva, Lola, et al.
Published: (2025)
by: Solovyeva, Lola, et al.
Published: (2025)
BitsAI-CR: Automated Code Review via LLM in Practice
by: Sun, Tao, et al.
Published: (2025)
by: Sun, Tao, et al.
Published: (2025)
Position Paper: Programming Language Techniques for Bridging LLM Code Generation Semantic Gaps
by: Du, Yalong, et al.
Published: (2025)
by: Du, Yalong, et al.
Published: (2025)
Experience converting a large mathematical software package written in C++ to C++20 modules
by: Bangerth, Wolfgang
Published: (2025)
by: Bangerth, Wolfgang
Published: (2025)
A first look at ROS 2 applications written in asynchronous Rust
by: Škoudlil, Martin, et al.
Published: (2025)
by: Škoudlil, Martin, et al.
Published: (2025)
Generative AI for Object-Oriented Programming: Writing the Right Code and Reasoning the Right Logic
by: Xu, Gang, et al.
Published: (2025)
by: Xu, Gang, et al.
Published: (2025)
A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics
by: Katzy, Jonathan, et al.
Published: (2025)
by: Katzy, Jonathan, et al.
Published: (2025)
Requirements are All You Need: From Requirements to Code with LLMs
by: Wei, Bingyang
Published: (2024)
by: Wei, Bingyang
Published: (2024)
CodeImprove: Program Adaptation for Deep Code Models
by: Rathnasuriya, Ravishka, et al.
Published: (2025)
by: Rathnasuriya, Ravishka, et al.
Published: (2025)
No Man is an Island: Towards Fully Automatic Programming by Code Search, Code Generation and Program Repair
by: Zhang, Quanjun, et al.
Published: (2024)
by: Zhang, Quanjun, et al.
Published: (2024)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
Challenges of Multilingual Program Specification and Analysis
by: Furia, Carlo A., et al.
Published: (2024)
by: Furia, Carlo A., et al.
Published: (2024)
Optimization is Better than Generation: Optimizing Commit Message Leveraging Human-written Commit Message
by: Li, Jiawei, et al.
Published: (2025)
by: Li, Jiawei, et al.
Published: (2025)
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics
by: Pathak, Aditya, et al.
Published: (2025)
by: Pathak, Aditya, et al.
Published: (2025)
Similar Items
-
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
by: Gurioli, Andrea, et al.
Published: (2026) -
Do not copy and paste! Rewriting strategies for code retrieval
by: Gurioli, Andrea, et al.
Published: (2026) -
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
by: Gurioli, Andrea, et al.
Published: (2025) -
Wild SBOMs: a Large-scale Dataset of Software Bills of Materials from Public Code
by: Soeiro, Luıs, et al.
Published: (2025) -
On the Informativeness of Security Commit Messages: A Large-scale Replication Study
by: Islam, Syful, et al.
Published: (2026)