Assessing GPT-4-Vision's Capabilities in UML-Based Code Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Antal, Gábor, Vozár, Richárd, Ferenc, Rudolf |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging GPT-4 for Vulnerability-Witnessing Unit Test Generation
por: Antal, Gábor, et al.
Publicado: (2025)
por: Antal, Gábor, et al.
Publicado: (2025)
Transforming C++11 Code to C++03 to Support Legacy Compilation Environments
por: Antal, Gábor, et al.
Publicado: (2024)
por: Antal, Gábor, et al.
Publicado: (2024)
Identifying Helpful Context for LLM-based Vulnerability Repair: A Preliminary Study
por: Antal, Gábor, et al.
Publicado: (2025)
por: Antal, Gábor, et al.
Publicado: (2025)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
por: Wang, Peiding, et al.
Publicado: (2025)
por: Wang, Peiding, et al.
Publicado: (2025)
Assessing Code Understanding in LLMs
por: Laneve, Cosimo, et al.
Publicado: (2025)
por: Laneve, Cosimo, et al.
Publicado: (2025)
Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
por: Chen, Le, et al.
Publicado: (2025)
por: Chen, Le, et al.
Publicado: (2025)
From Code Generation to Software Testing: AI Copilot with Context-Based RAG
por: Wang, Yuchen, et al.
Publicado: (2025)
por: Wang, Yuchen, et al.
Publicado: (2025)
Dynamic Stability of LLM-Generated Code
por: Rajput, Prateek, et al.
Publicado: (2025)
por: Rajput, Prateek, et al.
Publicado: (2025)
Code-Vision: Evaluating Multimodal LLMs Logic Understanding and Code Generation Capabilities
por: Wang, Hanbin, et al.
Publicado: (2025)
por: Wang, Hanbin, et al.
Publicado: (2025)
Is Functional Correctness Enough to Evaluate Code Language Models? Exploring Diversity of Generated Codes
por: Chon, Heejae, et al.
Publicado: (2024)
por: Chon, Heejae, et al.
Publicado: (2024)
Effective LLM-Driven Code Generation with Pythoness
por: Levin, Kyla H., et al.
Publicado: (2025)
por: Levin, Kyla H., et al.
Publicado: (2025)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
por: Sun, Zhensu, et al.
Publicado: (2026)
por: Sun, Zhensu, et al.
Publicado: (2026)
A Preliminary Study of Multilingual Code Language Models for Code Generation Task Using Translated Benchmarks
por: Dandamudi, Rohit, et al.
Publicado: (2024)
por: Dandamudi, Rohit, et al.
Publicado: (2024)
Smaller = Weaker? Benchmarking Robustness of Quantized LLMs in Code Generation
por: Fang, Sen, et al.
Publicado: (2025)
por: Fang, Sen, et al.
Publicado: (2025)
Insights from the Usage of the Ansible Lightspeed Code Completion Service
por: Sahoo, Priyam, et al.
Publicado: (2024)
por: Sahoo, Priyam, et al.
Publicado: (2024)
Hydra: Efficient, Correct Code Generation via Checkpoint-and-Rollback Support
por: Du, Alexander, et al.
Publicado: (2026)
por: Du, Alexander, et al.
Publicado: (2026)
Self-Improving Code Generation via Semantic Entropy and Behavioral Consensus
por: Zhang, Huan, et al.
Publicado: (2026)
por: Zhang, Huan, et al.
Publicado: (2026)
AutoMCQ -- Automatically Generate Code Comprehension Questions using GenAI
por: Goodfellow, Martin, et al.
Publicado: (2025)
por: Goodfellow, Martin, et al.
Publicado: (2025)
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
por: Sadik, Ahmed R., et al.
Publicado: (2025)
por: Sadik, Ahmed R., et al.
Publicado: (2025)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
por: Peng, Yun, et al.
Publicado: (2024)
por: Peng, Yun, et al.
Publicado: (2024)
From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging
por: Shi, Yuling, et al.
Publicado: (2024)
por: Shi, Yuling, et al.
Publicado: (2024)
Learning to Guarantee Type Correctness in Code Generation through Type-Guided Program Synthesis
por: Huang, Zhechong, et al.
Publicado: (2025)
por: Huang, Zhechong, et al.
Publicado: (2025)
AI Coders Are Among Us: Rethinking Programming Language Grammar Towards Efficient Code Generation
por: Sun, Zhensu, et al.
Publicado: (2024)
por: Sun, Zhensu, et al.
Publicado: (2024)
Perish or Flourish? A Holistic Evaluation of Large Language Models for Code Generation in Functional Programming
por: Lang, Nguyet-Anh H., et al.
Publicado: (2026)
por: Lang, Nguyet-Anh H., et al.
Publicado: (2026)
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
por: Wallraven, Stephan, et al.
Publicado: (2026)
por: Wallraven, Stephan, et al.
Publicado: (2026)
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
por: Misra, Diganta, et al.
Publicado: (2025)
por: Misra, Diganta, et al.
Publicado: (2025)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
por: Dai, Hankun, et al.
Publicado: (2025)
por: Dai, Hankun, et al.
Publicado: (2025)
Agentic Code Reasoning
por: Ugare, Shubham, et al.
Publicado: (2026)
por: Ugare, Shubham, et al.
Publicado: (2026)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
por: Saieva, Anthony, et al.
Publicado: (2023)
por: Saieva, Anthony, et al.
Publicado: (2023)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
por: Kim, Su-Hyeon, et al.
Publicado: (2025)
por: Kim, Su-Hyeon, et al.
Publicado: (2025)
Assessing the Interpretability of Programmatic Policies with Large Language Models
por: Bashir, Zahra, et al.
Publicado: (2023)
por: Bashir, Zahra, et al.
Publicado: (2023)
Is Self-Repair a Silver Bullet for Code Generation?
por: Olausson, Theo X., et al.
Publicado: (2023)
por: Olausson, Theo X., et al.
Publicado: (2023)
PPM: Automated Generation of Diverse Programming Problems for Benchmarking Code Generation Models
por: Chen, Simin, et al.
Publicado: (2024)
por: Chen, Simin, et al.
Publicado: (2024)
EnvTrace: Simulation-Based Semantic Evaluation of LLM Code via Execution Trace Alignment -- Demonstrated at Synchrotron Beamlines
por: van der Vleuten, Noah, et al.
Publicado: (2025)
por: van der Vleuten, Noah, et al.
Publicado: (2025)
AI-Mediated Code Comment Improvement
por: Dhakal, Maria, et al.
Publicado: (2025)
por: Dhakal, Maria, et al.
Publicado: (2025)
Benchmarking LLM Code Generation for Audio Programming with Visual Dataflow Languages
por: Zhang, William, et al.
Publicado: (2024)
por: Zhang, William, et al.
Publicado: (2024)
Incoherence as Oracle-less Measure of Error in LLM-Based Code Generation
por: Valentin, Thomas, et al.
Publicado: (2025)
por: Valentin, Thomas, et al.
Publicado: (2025)
AInsteinBench: Benchmarking Coding Agents on Scientific Repositories
por: Duston, Titouan, et al.
Publicado: (2025)
por: Duston, Titouan, et al.
Publicado: (2025)
AI-Assisted Fixes to Code Review Comments at Scale
por: Maddila, Chandra, et al.
Publicado: (2025)
por: Maddila, Chandra, et al.
Publicado: (2025)
A Problem-Oriented Perspective and Anchor Verification for Code Optimization
por: Ye, Tong, et al.
Publicado: (2024)
por: Ye, Tong, et al.
Publicado: (2024)
Ejemplares similares
-
Leveraging GPT-4 for Vulnerability-Witnessing Unit Test Generation
por: Antal, Gábor, et al.
Publicado: (2025) -
Transforming C++11 Code to C++03 to Support Legacy Compilation Environments
por: Antal, Gábor, et al.
Publicado: (2024) -
Identifying Helpful Context for LLM-based Vulnerability Repair: A Preliminary Study
por: Antal, Gábor, et al.
Publicado: (2025) -
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
por: Wang, Peiding, et al.
Publicado: (2025) -
Assessing Code Understanding in LLMs
por: Laneve, Cosimo, et al.
Publicado: (2025)