Can LLMs Generate User Stories and Assess Their Quality?
Fuente:
arXiv
Saved in:
| Main Authors: | Quattrocchi, Giovanni, Pasquale, Liliana, Spoletini, Paola, Baresi, Luciano |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging LLMs for User Stories in AI Systems: UStAI Dataset
by: Yamani, Asma, et al.
Published: (2025)
by: Yamani, Asma, et al.
Published: (2025)
Variability-Driven User-Story Generation using LLM and Triadic Concept Analysis
by: Bazin, Alexandre, et al.
Published: (2025)
by: Bazin, Alexandre, et al.
Published: (2025)
Can LLMs Solve Science or Just Write Code? Evaluating Quantum Solver Generation
by: Baresi, Luciano, et al.
Published: (2026)
by: Baresi, Luciano, et al.
Published: (2026)
Automated User Story Generation with Test Case Specification Using Large Language Model
by: Rahman, Tajmilur, et al.
Published: (2024)
by: Rahman, Tajmilur, et al.
Published: (2024)
User Story Tutor (UST) to Support Agile Software Developers
by: Neo, Giseldo da Silva, et al.
Published: (2024)
by: Neo, Giseldo da Silva, et al.
Published: (2024)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
by: Acharya, Jagrit, et al.
Published: (2025)
by: Acharya, Jagrit, et al.
Published: (2025)
Extracting Knowledge Graphs from User Stories using LangChain
by: da Silva, Thayná Camargo
Published: (2025)
by: da Silva, Thayná Camargo
Published: (2025)
Reverse Engineering User Stories from Code using Large Language Models
by: Ouf, Mohamed, et al.
Published: (2025)
by: Ouf, Mohamed, et al.
Published: (2025)
Generating Privacy Stories From Software Documentation
by: Baldwin, Wilder, et al.
Published: (2025)
by: Baldwin, Wilder, et al.
Published: (2025)
Efficient Domain Augmentation for Autonomous Driving Testing Using Diffusion Models
by: Baresi, Luciano, et al.
Published: (2024)
by: Baresi, Luciano, et al.
Published: (2024)
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code
by: He, Kaifeng, et al.
Published: (2026)
by: He, Kaifeng, et al.
Published: (2026)
DILLEMA: Diffusion and Large Language Models for Multi-Modal Augmentation
by: Baresi, Luciano, et al.
Published: (2025)
by: Baresi, Luciano, et al.
Published: (2025)
Can LLMs Replace Humans During Code Chunking?
by: Glasz, Christopher, et al.
Published: (2025)
by: Glasz, Christopher, et al.
Published: (2025)
Goal2Story: A Multi-Agent Fleet based on Privately Enabled sLLMs for Impacting Mapping on Requirements Elicitation
by: Zou, Xinkai, et al.
Published: (2025)
by: Zou, Xinkai, et al.
Published: (2025)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
by: Wang, Yibin, et al.
Published: (2025)
by: Wang, Yibin, et al.
Published: (2025)
Application Modernization with LLMs: Addressing Core Challenges in Reliability, Security, and Quality
by: Ponnusamy, Ahilan Ayyachamy Nadar
Published: (2025)
by: Ponnusamy, Ahilan Ayyachamy Nadar
Published: (2025)
A Note on Code Quality Score: LLMs for Maintainable Large Codebases
by: Wong, Sherman, et al.
Published: (2025)
by: Wong, Sherman, et al.
Published: (2025)
GBQA: A Game Benchmark for Evaluating LLMs as Quality Assurance Engineers
by: Jiang, Shufan, et al.
Published: (2026)
by: Jiang, Shufan, et al.
Published: (2026)
User Centric Evaluation of Code Generation Tools
by: Miah, Tanha, et al.
Published: (2024)
by: Miah, Tanha, et al.
Published: (2024)
Assessing Code Understanding in LLMs
by: Laneve, Cosimo, et al.
Published: (2025)
by: Laneve, Cosimo, et al.
Published: (2025)
Efficient Story Point Estimation With Comparative Learning
by: Khan, Monoshiz Mahbub, et al.
Published: (2025)
by: Khan, Monoshiz Mahbub, et al.
Published: (2025)
WebDevJudge: Evaluating (M)LLMs as Critiques for Web Development Quality
by: Li, Chunyang, et al.
Published: (2025)
by: Li, Chunyang, et al.
Published: (2025)
On Assessing the Relevance of Code Reviews Authored by Generative Models
by: Heumüller, Robert, et al.
Published: (2025)
by: Heumüller, Robert, et al.
Published: (2025)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
by: Thillen, Alex, et al.
Published: (2026)
by: Thillen, Alex, et al.
Published: (2026)
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study
by: Dhar, Rudra, et al.
Published: (2024)
by: Dhar, Rudra, et al.
Published: (2024)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
by: Wang, Ruiqi, et al.
Published: (2025)
by: Wang, Ruiqi, et al.
Published: (2025)
Generating Energy-efficient code with LLMs
by: Cappendijk, Tom, et al.
Published: (2024)
by: Cappendijk, Tom, et al.
Published: (2024)
Can LLM Generate Regression Tests for Software Commits?
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
by: Garg, Spandan, et al.
Published: (2026)
by: Garg, Spandan, et al.
Published: (2026)
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
by: Hu, Haichuan, et al.
Published: (2024)
by: Hu, Haichuan, et al.
Published: (2024)
Assessing AI Detectors in Identifying AI-Generated Code: Implications for Education
by: Pan, Wei Hung, et al.
Published: (2024)
by: Pan, Wei Hung, et al.
Published: (2024)
Generating Structured Plan Representation of Procedures with LLMs
by: Garg, Deepeka, et al.
Published: (2025)
by: Garg, Deepeka, et al.
Published: (2025)
Evaluating the Energy-Efficiency of the Code Generated by LLMs
by: Islam, Md Arman, et al.
Published: (2025)
by: Islam, Md Arman, et al.
Published: (2025)
LLMs for Test Input Generation for Semantic Caches
by: Rasool, Zafaryab, et al.
Published: (2024)
by: Rasool, Zafaryab, et al.
Published: (2024)
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems
by: Cao, Yuhan, et al.
Published: (2025)
by: Cao, Yuhan, et al.
Published: (2025)
Search-based Optimisation of LLM Learning Shots for Story Point Estimation
by: Tawosi, Vali, et al.
Published: (2024)
by: Tawosi, Vali, et al.
Published: (2024)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
by: Tambon, Florian, et al.
Published: (2024)
by: Tambon, Florian, et al.
Published: (2024)
Holistic Evaluation of State-of-the-Art LLMs for Code Generation
by: Zhang, Le, et al.
Published: (2025)
by: Zhang, Le, et al.
Published: (2025)
Towards a General Framework for HTN Modeling with LLMs
by: Puerta-Merino, Israel, et al.
Published: (2025)
by: Puerta-Merino, Israel, et al.
Published: (2025)
RuleFlow : Generating Reusable Program Optimizations with LLMs
by: Singh, Avaljot, et al.
Published: (2026)
by: Singh, Avaljot, et al.
Published: (2026)
Similar Items
-
Leveraging LLMs for User Stories in AI Systems: UStAI Dataset
by: Yamani, Asma, et al.
Published: (2025) -
Variability-Driven User-Story Generation using LLM and Triadic Concept Analysis
by: Bazin, Alexandre, et al.
Published: (2025) -
Can LLMs Solve Science or Just Write Code? Evaluating Quantum Solver Generation
by: Baresi, Luciano, et al.
Published: (2026) -
Automated User Story Generation with Test Case Specification Using Large Language Model
by: Rahman, Tajmilur, et al.
Published: (2024) -
User Story Tutor (UST) to Support Agile Software Developers
by: Neo, Giseldo da Silva, et al.
Published: (2024)