DSL or Code? Evaluating the Quality of LLM-Generated Algebraic Specifications: A Case Study in Optimization at Kinaxis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ayoughi, Negin, Dewar, David, Nejati, Shiva, Sabetzadeh, Mehrdad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Automata Learning with Statistical Machine Learning: A Network Security Case Study
von: Ayoughi, Negin, et al.
Veröffentlicht: (2024)
von: Ayoughi, Negin, et al.
Veröffentlicht: (2024)
Genetic Programming for Self-Adaptive Auto-Scaling of Microservices
von: Li, Jia, et al.
Veröffentlicht: (2026)
von: Li, Jia, et al.
Veröffentlicht: (2026)
Requirements-driven Slicing of Simulink Models Using LLMs
von: Luitel, Dipeeka, et al.
Veröffentlicht: (2024)
von: Luitel, Dipeeka, et al.
Veröffentlicht: (2024)
Simulink Mutation Testing using CodeBERT
von: Zhang, Jingfan, et al.
Veröffentlicht: (2025)
von: Zhang, Jingfan, et al.
Veröffentlicht: (2025)
Automated Test Validators for Flaky Cyber-Physical System Simulators: Approach and Evaluation
von: Jodat, Baharin A., et al.
Veröffentlicht: (2025)
von: Jodat, Baharin A., et al.
Veröffentlicht: (2025)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2025)
Developing a Llama-Based Chatbot for CI/CD Question Answering: A Case Study at Ericsson
von: Chaudhary, Daksh, et al.
Veröffentlicht: (2024)
von: Chaudhary, Daksh, et al.
Veröffentlicht: (2024)
Question Answering for Multi-Release Systems: A Case Study at Ciena
von: Khamsepour, Parham, et al.
Veröffentlicht: (2026)
von: Khamsepour, Parham, et al.
Veröffentlicht: (2026)
The Impact of Critique on LLM-Based Model Generation from Natural Language: The Case of Activity Diagrams
von: Khamsepour, Parham, et al.
Veröffentlicht: (2025)
von: Khamsepour, Parham, et al.
Veröffentlicht: (2025)
From Law to Gherkin: A Human-Centred Quasi-Experiment on the Quality of LLM-Generated Behavioural Specifications from Food-Safety Regulations
von: Hassani, Shabnam, et al.
Veröffentlicht: (2025)
von: Hassani, Shabnam, et al.
Veröffentlicht: (2025)
A Lean Simulation Framework for Stress Testing IoT Cloud Systems
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Test Input Validation for Vision-based DL Systems: An Active Learning Approach
von: Ghobari, Delaram, et al.
Veröffentlicht: (2025)
von: Ghobari, Delaram, et al.
Veröffentlicht: (2025)
Practical Guidelines for the Selection and Evaluation of Natural Language Processing Techniques in Requirements Engineering
von: Sabetzadeh, Mehrdad, et al.
Veröffentlicht: (2024)
von: Sabetzadeh, Mehrdad, et al.
Veröffentlicht: (2024)
An Empirical Study on LLM-based Classification of Requirements-related Provisions in Food-safety Regulations
von: Hassani, Shabnam, et al.
Veröffentlicht: (2025)
von: Hassani, Shabnam, et al.
Veröffentlicht: (2025)
Improving Requirements Completeness: Automated Assistance through Large Language Models
von: Luitel, Dipeeka, et al.
Veröffentlicht: (2023)
von: Luitel, Dipeeka, et al.
Veröffentlicht: (2023)
Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study
von: Chand, Sivajeet, et al.
Veröffentlicht: (2026)
von: Chand, Sivajeet, et al.
Veröffentlicht: (2026)
Rethinking Legal Compliance Automation: Opportunities with Large Language Models
von: Hassani, Shabnam, et al.
Veröffentlicht: (2024)
von: Hassani, Shabnam, et al.
Veröffentlicht: (2024)
Bridging the Gap between Real-world and Synthetic Images for Testing Autonomous Driving Systems
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2024)
von: Amini, Mohammad Hossein, et al.
Veröffentlicht: (2024)
Can Search-Based Testing with Pareto Optimization Effectively Cover Failure-Revealing Test Inputs?
von: Sorokin, Lev, et al.
Veröffentlicht: (2024)
von: Sorokin, Lev, et al.
Veröffentlicht: (2024)
From Text to DSL: Evaluating Grammar-Based Model Generation Using Open LLMs
von: Baber, Junaid, et al.
Veröffentlicht: (2026)
von: Baber, Junaid, et al.
Veröffentlicht: (2026)
Evaluating LLM-Generated Code: A Benchmark and Developer Study
von: Szych, Joanna, et al.
Veröffentlicht: (2026)
von: Szych, Joanna, et al.
Veröffentlicht: (2026)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
von: Zhang, Binquan, et al.
Veröffentlicht: (2025)
von: Zhang, Binquan, et al.
Veröffentlicht: (2025)
An LLM-driven Scenario Generation Pipeline Using an Extended Scenic DSL for Autonomous Driving Safety Validation
von: Safa, Fida Khandaker, et al.
Veröffentlicht: (2026)
von: Safa, Fida Khandaker, et al.
Veröffentlicht: (2026)
Still Manual? Automated Linter Configuration via DSL-Based LLM Compilation of Coding Standards
von: Zhang, Zejun, et al.
Veröffentlicht: (2026)
von: Zhang, Zejun, et al.
Veröffentlicht: (2026)
An Event-Driven Tool for Context-Aware Code Smell Detection Using SmellDSL
von: Viegas, Matheus dos Santos, et al.
Veröffentlicht: (2026)
von: Viegas, Matheus dos Santos, et al.
Veröffentlicht: (2026)
CONCORD: Towards a DSL for Configurable Graph Code Representation
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
von: Saad, Mootez, et al.
Veröffentlicht: (2024)
Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design
von: Rosa, Giovanni, et al.
Veröffentlicht: (2026)
von: Rosa, Giovanni, et al.
Veröffentlicht: (2026)
Evaluating Efficiency and Novelty of LLM-Generated Code for Graph Analysis
von: Nia, Atieh Barati, et al.
Veröffentlicht: (2025)
von: Nia, Atieh Barati, et al.
Veröffentlicht: (2025)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
von: Du, Mingzhe, et al.
Veröffentlicht: (2025)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
von: Kuang, Shiqi, et al.
Veröffentlicht: (2025)
von: Kuang, Shiqi, et al.
Veröffentlicht: (2025)
Towards a DSL to Formalize Multimodal Requirements
von: Gomez-Vazquez, Marcos, et al.
Veröffentlicht: (2025)
von: Gomez-Vazquez, Marcos, et al.
Veröffentlicht: (2025)
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
von: AKLI, Amal, et al.
Veröffentlicht: (2026)
von: AKLI, Amal, et al.
Veröffentlicht: (2026)
Quality Requirements for Code: On the Untapped Potential in Maintainability Specifications
von: Borg, Markus
Veröffentlicht: (2024)
von: Borg, Markus
Veröffentlicht: (2024)
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
von: McIntyre-Garcia, Cristopher, et al.
Veröffentlicht: (2024)
von: McIntyre-Garcia, Cristopher, et al.
Veröffentlicht: (2024)
Beyond Basic Specifications? A Systematic Study of Logical Constructs in LLM-based Specification Generation
von: Chen, Zehan, et al.
Veröffentlicht: (2026)
von: Chen, Zehan, et al.
Veröffentlicht: (2026)
Automated Repair of TEE Partitioning Issues via DSL-Guided and LLM-Assisted Patching
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
On the Quality of AI-Generated Source Code Comments: A Comprehensive Evaluation
von: Guelman, Ian, et al.
Veröffentlicht: (2024)
von: Guelman, Ian, et al.
Veröffentlicht: (2024)
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
von: Fakhoury, Sarah, et al.
Veröffentlicht: (2024)
von: Fakhoury, Sarah, et al.
Veröffentlicht: (2024)
Hallucination in LLM-Based Code Generation: An Automotive Case Study
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
von: Pavel, Marc, et al.
Veröffentlicht: (2025)
LLM Assisted Coding with Metamorphic Specification Mutation Agent
von: Akhond, Mostafijur Rahman, et al.
Veröffentlicht: (2025)
von: Akhond, Mostafijur Rahman, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Enhancing Automata Learning with Statistical Machine Learning: A Network Security Case Study
von: Ayoughi, Negin, et al.
Veröffentlicht: (2024) -
Genetic Programming for Self-Adaptive Auto-Scaling of Microservices
von: Li, Jia, et al.
Veröffentlicht: (2026) -
Requirements-driven Slicing of Simulink Models Using LLMs
von: Luitel, Dipeeka, et al.
Veröffentlicht: (2024) -
Simulink Mutation Testing using CodeBERT
von: Zhang, Jingfan, et al.
Veröffentlicht: (2025) -
Automated Test Validators for Flaky Cyber-Physical System Simulators: Approach and Evaluation
von: Jodat, Baharin A., et al.
Veröffentlicht: (2025)