A framework for assessing the capabilities of code generation of constraint domain-specific languages with large language models
Fuente:
arXiv
Saved in:
| Main Authors: | Delgado, David, Burgueño, Lola, Clarisó, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating LLM-generated code for domain-specific languages: molecular dynamics with LAMMPS
by: Holbrook, Ethan, et al.
Published: (2026)
by: Holbrook, Ethan, et al.
Published: (2026)
Certus: A domain specific language for confidence assessment in assurance cases
by: Diemert, Simon, et al.
Published: (2025)
by: Diemert, Simon, et al.
Published: (2025)
Evaluation of large language models for assessing code maintainability
by: Dillmann, Marc, et al.
Published: (2024)
by: Dillmann, Marc, et al.
Published: (2024)
A Benchmarking Framework for Model Datasets
by: Glaser, Philipp-Lorenz, et al.
Published: (2026)
by: Glaser, Philipp-Lorenz, et al.
Published: (2026)
ALPINE: An adaptive language-agnostic pruning method for language models for code
by: Saad, Mootez, et al.
Published: (2024)
by: Saad, Mootez, et al.
Published: (2024)
GeoAnalystBench: A GeoAI benchmark for assessing large language models for spatial analysis workflow and code generation
by: Zhang, Qianheng, et al.
Published: (2025)
by: Zhang, Qianheng, et al.
Published: (2025)
Comparing large language models and human programmers for generating programming code
by: Hou, Wenpin, et al.
Published: (2024)
by: Hou, Wenpin, et al.
Published: (2024)
LLM-based Generation of Semantically Diverse and Realistic Domain Model Instances
by: Coman, Andrei, et al.
Published: (2026)
by: Coman, Andrei, et al.
Published: (2026)
Detecting Semantic Alignments between Textual Specifications and Domain Models
by: Shimangaud, Shwetali, et al.
Published: (2026)
by: Shimangaud, Shwetali, et al.
Published: (2026)
Exploring Actions, Interactions and Challenges in Software Modelling Tasks: An Empirical Investigation with Students
by: Chakraborty, Shalini, et al.
Published: (2024)
by: Chakraborty, Shalini, et al.
Published: (2024)
Automation in Model-Driven Engineering: A look back, and ahead
by: Burgueño, Lola, et al.
Published: (2024)
by: Burgueño, Lola, et al.
Published: (2024)
An approach for API synthesis using large language models
by: Zhong, Hua, et al.
Published: (2025)
by: Zhong, Hua, et al.
Published: (2025)
What a diff makes: automating code migration with large language models
by: Rosenfeld, Katherine A., et al.
Published: (2025)
by: Rosenfeld, Katherine A., et al.
Published: (2025)
Large language models for generating rules, yay or nay?
by: Sivasothy, Shangeetha, et al.
Published: (2024)
by: Sivasothy, Shangeetha, et al.
Published: (2024)
LangBiTe: A Platform for Testing Bias in Large Language Models
by: Morales, Sergio, et al.
Published: (2024)
by: Morales, Sergio, et al.
Published: (2024)
Mind the Ethics! The Overlooked Ethical Dimensions of GenAI in Software Modeling Education
by: Chakraborty, Shalini, et al.
Published: (2025)
by: Chakraborty, Shalini, et al.
Published: (2025)
An empirical study of LoRA-based fine-tuning of large language models for automated test case generation
by: Moradi, Milad, et al.
Published: (2026)
by: Moradi, Milad, et al.
Published: (2026)
SPVR: syntax-to-prompt vulnerability repair based on large language models
by: Wang, Ruoke, et al.
Published: (2024)
by: Wang, Ruoke, et al.
Published: (2024)
The importance of visual modelling languages in generative software engineering
by: Rossi, Roberto
Published: (2024)
by: Rossi, Roberto
Published: (2024)
A Framework to Model ML Engineering Processes
by: Morales, Sergio, et al.
Published: (2024)
by: Morales, Sergio, et al.
Published: (2024)
EnseSmells: Deep ensemble and programming language models for automated code smells detection
by: Ho, Anh, et al.
Published: (2025)
by: Ho, Anh, et al.
Published: (2025)
DeepQuali: Initial results of a study on the use of large language models for assessing the quality of user stories
by: Trendowicz, Adam, et al.
Published: (2026)
by: Trendowicz, Adam, et al.
Published: (2026)
LastMerge: A language-agnostic structured tool for code integration
by: Duarte, Joao Pedro, et al.
Published: (2025)
by: Duarte, Joao Pedro, et al.
Published: (2025)
CFD-copilot: leveraging domain-adapted large language model and model context protocol to enhance simulation automation
by: Dong, Zhehao, et al.
Published: (2025)
by: Dong, Zhehao, et al.
Published: (2025)
Retrieval-augmented code completion for local projects using large language models
by: Hostnik, Marko, et al.
Published: (2024)
by: Hostnik, Marko, et al.
Published: (2024)
Usefulness of data flow diagrams and large language models for security threat validation: a registered report
by: Mbaka, Winnie Bahati, et al.
Published: (2024)
by: Mbaka, Winnie Bahati, et al.
Published: (2024)
Large language models for behavioral modeling: A literature survey
by: Laiq, Muhammad
Published: (2025)
by: Laiq, Muhammad
Published: (2025)
On the Utility of Domain Modeling Assistance with Large Language Models
by: Chaaben, Meriem Ben, et al.
Published: (2024)
by: Chaaben, Meriem Ben, et al.
Published: (2024)
BUGFIX: towards a common language and framework for the AutomaticProgram Repair community
by: Meyer, Bertrand, et al.
Published: (2024)
by: Meyer, Bertrand, et al.
Published: (2024)
SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
by: Rashid, Muhammad Shihab, et al.
Published: (2025)
by: Rashid, Muhammad Shihab, et al.
Published: (2025)
A domain-specific language for describing machine learning datasets
by: Giner-Miguelez, Joan, et al.
Published: (2022)
by: Giner-Miguelez, Joan, et al.
Published: (2022)
Reflections on the design, applications and implementations of the normative specification language eFLINT
by: van Binsbergen, L. Thomas, et al.
Published: (2025)
by: van Binsbergen, L. Thomas, et al.
Published: (2025)
An evaluation of LLM code generation capabilities through graded exercises
by: Jiménez, Álvaro Barbero
Published: (2024)
by: Jiménez, Álvaro Barbero
Published: (2024)
Architecting software monitors for control-flow anomaly detection through large language models and conformance checking
by: Vitale, Francesco, et al.
Published: (2025)
by: Vitale, Francesco, et al.
Published: (2025)
David vs. Goliath: A comparative study of different-sized LLMs for code generation in the domain of automotive scenario generation
by: Bauerfeind, Philipp, et al.
Published: (2025)
by: Bauerfeind, Philipp, et al.
Published: (2025)
Agent-based code generation for the Gammapy framework
by: Kostunin, Dmitriy, et al.
Published: (2025)
by: Kostunin, Dmitriy, et al.
Published: (2025)
Towards a unified user modeling language for engineering human centered AI systems
by: Conrardy, Aaron, et al.
Published: (2025)
by: Conrardy, Aaron, et al.
Published: (2025)
On the synchronization between Hugging Face pre-trained language models and their upstream GitHub repository
by: Ajibode, Adekunle, et al.
Published: (2025)
by: Ajibode, Adekunle, et al.
Published: (2025)
Beyond the model: Key differentiators in large language models and multi-agent services
by: Goyal, Muskaan, et al.
Published: (2025)
by: Goyal, Muskaan, et al.
Published: (2025)
Insights into resource utilization of code small language models serving with runtime engines and execution providers
by: Durán, Francisco, et al.
Published: (2024)
by: Durán, Francisco, et al.
Published: (2024)
Similar Items
-
Evaluating LLM-generated code for domain-specific languages: molecular dynamics with LAMMPS
by: Holbrook, Ethan, et al.
Published: (2026) -
Certus: A domain specific language for confidence assessment in assurance cases
by: Diemert, Simon, et al.
Published: (2025) -
Evaluation of large language models for assessing code maintainability
by: Dillmann, Marc, et al.
Published: (2024) -
A Benchmarking Framework for Model Datasets
by: Glaser, Philipp-Lorenz, et al.
Published: (2026) -
ALPINE: An adaptive language-agnostic pruning method for language models for code
by: Saad, Mootez, et al.
Published: (2024)