GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
Fuente:
arXiv
Saved in:
| Main Authors: | Misra, Diganta, Islah, Nizar, May, Victor, Rauby, Brice, Wang, Zihan, Gehring, Justine, Orvieto, Antonio, Chaudhary, Muawiz, Muller, Eilif B., Rish, Irina, Kahou, Samira Ebrahimi, Caccia, Massimo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GitChameleon: Unmasking the Version-Switching Capabilities of Code Generation Models
by: Islah, Nizar, et al.
Published: (2024)
by: Islah, Nizar, et al.
Published: (2024)
Uncovering the Hidden Cost of Model Compression
by: Misra, Diganta, et al.
Published: (2023)
by: Misra, Diganta, et al.
Published: (2023)
Handling Delay in Real-Time Reinforcement Learning
by: Anokhin, Ivan, et al.
Published: (2025)
by: Anokhin, Ivan, et al.
Published: (2025)
Zero-Shot Anomaly Detection with Dual-Branch Prompt Selection
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
Explaining Grokking in Transformers through the Lens of Inductive Bias
by: Singh, Jaisidh, et al.
Published: (2026)
by: Singh, Jaisidh, et al.
Published: (2026)
On the low-shot transferability of [V]-Mamba
by: Misra, Diganta, et al.
Published: (2024)
by: Misra, Diganta, et al.
Published: (2024)
Learning to combine top-down context and feed-forward representations under ambiguity with apical and basal dendrites
by: Islah, Nizar, et al.
Published: (2023)
by: Islah, Nizar, et al.
Published: (2023)
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
by: May, Victor, et al.
Published: (2025)
by: May, Victor, et al.
Published: (2025)
(Almost) Free Modality Stitching of Foundation Models
by: Singh, Jaisidh, et al.
Published: (2025)
by: Singh, Jaisidh, et al.
Published: (2025)
GRASP: Deterministic argument ranking in interaction graphs
by: Misra, Diganta, et al.
Published: (2026)
by: Misra, Diganta, et al.
Published: (2026)
CAMMARL: Conformal Action Modeling in Multi Agent Reinforcement Learning
by: Gupta, Nikunj, et al.
Published: (2023)
by: Gupta, Nikunj, et al.
Published: (2023)
Estimation of Head Motion in Structural MRI and its Impact on Cortical Thickness Measurements in Retrospective Data
by: Bricout, Charles, et al.
Published: (2025)
by: Bricout, Charles, et al.
Published: (2025)
Adaptive Group Robust Ensemble Knowledge Distillation
by: Kenfack, Patrik, et al.
Published: (2024)
by: Kenfack, Patrik, et al.
Published: (2024)
Towards Fair In-Context Learning with Tabular Foundation Models
by: Kenfack, Patrik, et al.
Published: (2025)
by: Kenfack, Patrik, et al.
Published: (2025)
Locally Constrained Representations in Reinforcement Learning
by: Nath, Somjit, et al.
Published: (2022)
by: Nath, Somjit, et al.
Published: (2022)
Learning to Play Atari in a World of Tokens
by: Agarwal, Pranav, et al.
Published: (2024)
by: Agarwal, Pranav, et al.
Published: (2024)
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
by: Abbes, Istabrak, et al.
Published: (2025)
by: Abbes, Istabrak, et al.
Published: (2025)
Fairness Under Demographic Scarce Regime
by: Kenfack, Patrik Joslin, et al.
Published: (2023)
by: Kenfack, Patrik Joslin, et al.
Published: (2023)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
by: Yu, Xuemin, et al.
Published: (2026)
by: Yu, Xuemin, et al.
Published: (2026)
Cross-Layer Discrete Concept Discovery for Interpreting Language Models
by: Garg, Ankur, et al.
Published: (2025)
by: Garg, Ankur, et al.
Published: (2025)
Behaviour Discovery and Attribution for Explainable Reinforcement Learning
by: Rishav, Rishav, et al.
Published: (2025)
by: Rishav, Rishav, et al.
Published: (2025)
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation
by: Azimi, Rambod, et al.
Published: (2024)
by: Azimi, Rambod, et al.
Published: (2024)
Learning Multi-agent Multi-machine Tending by Mobile Robots
by: Abdalwhab, Abdalwhab, et al.
Published: (2024)
by: Abdalwhab, Abdalwhab, et al.
Published: (2024)
Pruning Sparse Tensor Neural Networks Enables Deep Learning for 3D Ultrasound Localization Microscopy
by: Rauby, Brice, et al.
Published: (2024)
by: Rauby, Brice, et al.
Published: (2024)
Agents Learn Their Runtime: Interpreter Persistence as Training-Time Semantics
by: May, Victor, et al.
Published: (2026)
by: May, Victor, et al.
Published: (2026)
Survey on AI Ethics: A Socio-technical Perspective
by: Mbiazi, Dave, et al.
Published: (2023)
by: Mbiazi, Dave, et al.
Published: (2023)
Solving Models of Economic Dynamics with Ridgeless Kernel Regressions
by: Kahou, Mahdi Ebrahimi, et al.
Published: (2024)
by: Kahou, Mahdi Ebrahimi, et al.
Published: (2024)
Prediction of Final Phosphorus Content of Steel in a Scrap-Based Electric Arc Furnace Using Artificial Neural Networks
by: Azzaz, Riadh, et al.
Published: (2024)
by: Azzaz, Riadh, et al.
Published: (2024)
Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment
by: Rahman, Aamer Abdul, et al.
Published: (2024)
by: Rahman, Aamer Abdul, et al.
Published: (2024)
Keeping in Place After the Storm-Emergency Assistance and Evictions
by: Islah, Bilal, et al.
Published: (2025)
by: Islah, Bilal, et al.
Published: (2025)
Bounded Foresight Equilibrium in Large Dynamic Economies with Heterogeneous Agents and Aggregate Shocks
by: Islah, Bilal, et al.
Published: (2025)
by: Islah, Bilal, et al.
Published: (2025)
Learning From the Past with Cascading Eligibility Traces
by: Ralambomihanta, Tokiniaina Raharison, et al.
Published: (2025)
by: Ralambomihanta, Tokiniaina Raharison, et al.
Published: (2025)
Reinforcement Learning for Sequence Design Leveraging Protein Language Models
by: Subramanian, Jithendaraa, et al.
Published: (2024)
by: Subramanian, Jithendaraa, et al.
Published: (2024)
Comparative Analysis of Diffusion Generative Models in Computational Pathology
by: Thakkar, Denisha, et al.
Published: (2024)
by: Thakkar, Denisha, et al.
Published: (2024)
Pollen profile DONVOLD, Donvold, Norway
by: Nilssen, Eilif J
Published: (2010)
by: Nilssen, Eilif J
Published: (2010)
Age determination of sediment core DONVOLD, Donvold, Norway
by: Nilssen, Eilif J
Published: (2010)
by: Nilssen, Eilif J
Published: (2010)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
by: Legate, Gwen, et al.
Published: (2025)
by: Legate, Gwen, et al.
Published: (2025)
Towards ethical multimodal systems
by: Roger, Alexis, et al.
Published: (2023)
by: Roger, Alexis, et al.
Published: (2023)
On the Limits of Multi-modal Meta-Learning with Auxiliary Task Modulation Using Conditional Batch Normalization
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
by: Armengol-Estapé, Jordi, et al.
Published: (2024)
Towards a Neural Debugger for Python
by: Beck, Maximilian, et al.
Published: (2026)
by: Beck, Maximilian, et al.
Published: (2026)
Similar Items
-
GitChameleon: Unmasking the Version-Switching Capabilities of Code Generation Models
by: Islah, Nizar, et al.
Published: (2024) -
Uncovering the Hidden Cost of Model Compression
by: Misra, Diganta, et al.
Published: (2023) -
Handling Delay in Real-Time Reinforcement Learning
by: Anokhin, Ivan, et al.
Published: (2025) -
Zero-Shot Anomaly Detection with Dual-Branch Prompt Selection
by: Wang, Zihan, et al.
Published: (2025) -
Explaining Grokking in Transformers through the Lens of Inductive Bias
by: Singh, Jaisidh, et al.
Published: (2026)