GitChameleon: Unmasking the Version-Switching Capabilities of Code Generation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Islah, Nizar, Gehring, Justine, Misra, Diganta, Muller, Eilif, Rish, Irina, Zhuo, Terry Yue, Caccia, Massimo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
by: Misra, Diganta, et al.
Published: (2025)
by: Misra, Diganta, et al.
Published: (2025)
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
by: May, Victor, et al.
Published: (2025)
by: May, Victor, et al.
Published: (2025)
Unmasking the Genuine Type Inference Capabilities of LLMs for Java Code Snippets
by: Dong, Yiwen, et al.
Published: (2025)
by: Dong, Yiwen, et al.
Published: (2025)
ICE-Score: Instructing Large Language Models to Evaluate Code
by: Zhuo, Terry Yue
Published: (2023)
by: Zhuo, Terry Yue
Published: (2023)
GitEvo: Code Evolution Analysis for Git Repositories
by: Hora, Andre
Published: (2026)
by: Hora, Andre
Published: (2026)
Fingerprinting AI Coding Agents on GitHub
by: Ghaleb, Taher A.
Published: (2026)
by: Ghaleb, Taher A.
Published: (2026)
Robustness, Security, Privacy, Explainability, Efficiency, and Usability of Large Language Models for Code
by: Yang, Zhou, et al.
Published: (2024)
by: Yang, Zhou, et al.
Published: (2024)
Agentic Much? Adoption of Coding Agents on GitHub
by: Robbes, Romain, et al.
Published: (2026)
by: Robbes, Romain, et al.
Published: (2026)
Chain-of-Thought in Neural Code Generation: From and For Lightweight Language Models
by: Yang, Guang, et al.
Published: (2023)
by: Yang, Guang, et al.
Published: (2023)
Automated Code Review Assignments: An Alternative Perspective of Code Ownership on GitHub
by: Lulla, Jai Lal, et al.
Published: (2025)
by: Lulla, Jai Lal, et al.
Published: (2025)
Measuring Code Efficiency Optimization Capabilities with ACEOB
by: Pan, Yue, et al.
Published: (2024)
by: Pan, Yue, et al.
Published: (2024)
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery
by: Zhao, Jiale, et al.
Published: (2026)
by: Zhao, Jiale, et al.
Published: (2026)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
by: Du, Mingzhe, et al.
Published: (2025)
by: Du, Mingzhe, et al.
Published: (2025)
On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
by: Watanabe, Miku, et al.
Published: (2025)
by: Watanabe, Miku, et al.
Published: (2025)
An Empirical Analysis of Git Commit Logs for Potential Inconsistency in Code Clones
by: Yokomori, Reishi, et al.
Published: (2024)
by: Yokomori, Reishi, et al.
Published: (2024)
From Code to Courtroom: LLMs as the New Software Judges
by: He, Junda, et al.
Published: (2025)
by: He, Junda, et al.
Published: (2025)
Does AI Code Review Lead to Code Changes? A Case Study of GitHub Actions
by: Sun, Kexin, et al.
Published: (2025)
by: Sun, Kexin, et al.
Published: (2025)
Where Is Self-admitted Code Generated by Large Language Models on GitHub?
by: Yu, Xiao, et al.
Published: (2024)
by: Yu, Xiao, et al.
Published: (2024)
Test Code Review in the Era of GitHub Actions: A Replication Study
by: Sun, Hui, et al.
Published: (2026)
by: Sun, Hui, et al.
Published: (2026)
Measuring the Runtime Performance of C++ Code Written by Humans using GitHub Copilot
by: Erhabor, Daniel, et al.
Published: (2023)
by: Erhabor, Daniel, et al.
Published: (2023)
Exploring the Effect of Multiple Natural Languages on Code Suggestion Using GitHub Copilot
by: Koyanagi, Kei, et al.
Published: (2024)
by: Koyanagi, Kei, et al.
Published: (2024)
Git Context Controller: Manage the Context of LLM-based Agents like Git
by: Wu, Junde, et al.
Published: (2025)
by: Wu, Junde, et al.
Published: (2025)
LLMAID: Identifying AI Capabilities in Android Apps with LLMs
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
GitHub Copilot: the perfect Code compLeeter?
by: Siroš, Ilja, et al.
Published: (2024)
by: Siroš, Ilja, et al.
Published: (2024)
AIDev: Studying AI Coding Agents on GitHub
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
GitBug-Actions: Building Reproducible Bug-Fix Benchmarks with GitHub Actions
by: Saavedra, Nuno, et al.
Published: (2023)
by: Saavedra, Nuno, et al.
Published: (2023)
Less is More: DocString Compression in Code Generation
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
Not One to Rule Them All: Mining Meaningful Code Review Orders From GitHub
by: Bouraffa, Abir, et al.
Published: (2025)
by: Bouraffa, Abir, et al.
Published: (2025)
RevMine: An LLM-Assisted Tool for Code Review Mining and Analysis Across Git Platforms
by: Kansab, Samah, et al.
Published: (2025)
by: Kansab, Samah, et al.
Published: (2025)
Uncovering Code Insights: Leveraging GitHub Artifacts for Deeper Code Understanding
by: Nevo, Ziv, et al.
Published: (2025)
by: Nevo, Ziv, et al.
Published: (2025)
GitHub Proxy Server: A tool for supporting massive data collection on GitHub
by: Borges, Hudson Silva, et al.
Published: (2025)
by: Borges, Hudson Silva, et al.
Published: (2025)
Is GitHub's Copilot as Bad as Humans at Introducing Vulnerabilities in Code?
by: Asare, Owura, et al.
Published: (2022)
by: Asare, Owura, et al.
Published: (2022)
LadyBug: A GitHub Bot for UI-Enhanced Bug Localization in Mobile Apps
by: Mahmud, Junayed, et al.
Published: (2025)
by: Mahmud, Junayed, et al.
Published: (2025)
Guidelines for Developing Bots for GitHub
by: Wessel, Mairieli, et al.
Published: (2022)
by: Wessel, Mairieli, et al.
Published: (2022)
GitOps for Capture the Flag Platforms
by: Albrechtsen, Mikkel Bengtson, et al.
Published: (2026)
by: Albrechtsen, Mikkel Bengtson, et al.
Published: (2026)
Prioritising GitHub Priority Labels
by: Caddy, James, et al.
Published: (2024)
by: Caddy, James, et al.
Published: (2024)
"My GitHub Sponsors profile is live!" Investigating the Impact of Twitter/X Mentions on GitHub Sponsors
by: Fan, Youmei, et al.
Published: (2024)
by: Fan, Youmei, et al.
Published: (2024)
Code Comprehension with GitHub Copilot: Performance Gains, Comprehension Trade-offs, and Behavioral Predictors in Brownfield Programming
by: Qiao, Yunhan, et al.
Published: (2025)
by: Qiao, Yunhan, et al.
Published: (2025)
Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code
by: Kim, Myeongsoo, et al.
Published: (2026)
by: Kim, Myeongsoo, et al.
Published: (2026)
Defending Code Language Models against Backdoor Attacks with Deceptive Cross-Entropy Loss
by: Yang, Guang, et al.
Published: (2024)
by: Yang, Guang, et al.
Published: (2024)
Similar Items
-
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
by: Misra, Diganta, et al.
Published: (2025) -
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
by: May, Victor, et al.
Published: (2025) -
Unmasking the Genuine Type Inference Capabilities of LLMs for Java Code Snippets
by: Dong, Yiwen, et al.
Published: (2025) -
ICE-Score: Instructing Large Language Models to Evaluate Code
by: Zhuo, Terry Yue
Published: (2023) -
GitEvo: Code Evolution Analysis for Git Repositories
by: Hora, Andre
Published: (2026)