Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Karpurapu, Shanthi, Myneni, Sravanthy, Nettur, Unnati, Gajja, Likhit Sagar, Burke, Dave, Stiehm, Tom, Payne, Jeffery |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Role of GitHub Copilot on Software Development: A Perspective on Productivity, Security, Best Practices and Future Directions
par: Nettur, Suresh Babu, et autres
Publié: (2025)
par: Nettur, Suresh Babu, et autres
Publié: (2025)
UltraLightSqueezeNet: A Deep Learning Architecture for Malaria Classification with up to 54x fewer trainable parameters for resource constrained devices
par: Nettur, Suresh Babu, et autres
Publié: (2025)
par: Nettur, Suresh Babu, et autres
Publié: (2025)
Lightweight Weighted Average Ensemble Model for Pneumonia Detection in Chest X-Ray Images
par: Nettur, Suresh Babu, et autres
Publié: (2025)
par: Nettur, Suresh Babu, et autres
Publié: (2025)
A Hybrid Deep Learning CNN Model for Enhanced COVID-19 Detection from Computed Tomography (CT) Scan Images
par: Nettur, Suresh Babu, et autres
Publié: (2025)
par: Nettur, Suresh Babu, et autres
Publié: (2025)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
par: Ferenczi, Bryce, et autres
Publié: (2023)
par: Ferenczi, Bryce, et autres
Publié: (2023)
Automated Circuit Interpretation via Probe Prompting
par: Birardi, Giuseppe
Publié: (2025)
par: Birardi, Giuseppe
Publié: (2025)
Carefully Structured Compression: Efficiently Managing StarCraft II Data
par: Ferenczi, Bryce, et autres
Publié: (2024)
par: Ferenczi, Bryce, et autres
Publié: (2024)
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations
par: Ferenczi, Bryce, et autres
Publié: (2024)
par: Ferenczi, Bryce, et autres
Publié: (2024)
Toward Constraint Compliant Goal Formulation and Planning
par: Jones, Steven J., et autres
Publié: (2024)
par: Jones, Steven J., et autres
Publié: (2024)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
par: Endo, Masafumi, et autres
Publié: (2024)
par: Endo, Masafumi, et autres
Publié: (2024)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
par: Hashemi, Helia, et autres
Publié: (2024)
par: Hashemi, Helia, et autres
Publié: (2024)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
par: Wu, Dekun, et autres
Publié: (2023)
par: Wu, Dekun, et autres
Publié: (2023)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
par: Wei, Kaiwen, et autres
Publié: (2025)
par: Wei, Kaiwen, et autres
Publié: (2025)
Balancing the Scales: A Comprehensive Study on Tackling Class Imbalance in Binary Classification
par: Abdelhamid, Mohamed, et autres
Publié: (2024)
par: Abdelhamid, Mohamed, et autres
Publié: (2024)
EvoCUA: Evolving Computer Use Agents via Learning from Scalable Synthetic Experience
par: Xue, Taofeng, et autres
Publié: (2026)
par: Xue, Taofeng, et autres
Publié: (2026)
BudgetMLAgent: A Cost-Effective LLM Multi-Agent system for Automating Machine Learning Tasks
par: Gandhi, Shubham, et autres
Publié: (2024)
par: Gandhi, Shubham, et autres
Publié: (2024)
PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents
par: Iscan, Mehmet
Publié: (2026)
par: Iscan, Mehmet
Publié: (2026)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
par: Jeevan, Pranav, et autres
Publié: (2024)
par: Jeevan, Pranav, et autres
Publié: (2024)
Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight
par: Kaur, Kirandeep, et autres
Publié: (2026)
par: Kaur, Kirandeep, et autres
Publié: (2026)
Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback
par: Juzek, Tom S., et autres
Publié: (2025)
par: Juzek, Tom S., et autres
Publié: (2025)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
par: Huang, Yunpeng, et autres
Publié: (2023)
par: Huang, Yunpeng, et autres
Publié: (2023)
Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems
par: Albiero, Daniel, et autres
Publié: (2026)
par: Albiero, Daniel, et autres
Publié: (2026)
RefiningGPT: Specialized language Models for Automated Refinery Unit-level Process Diagram Synthesis
par: Liu, Dongxiao, et autres
Publié: (2026)
par: Liu, Dongxiao, et autres
Publié: (2026)
An Explainable Collaborative Dialogue System using a Theory of Mind
par: Cohen, Philip R., et autres
Publié: (2023)
par: Cohen, Philip R., et autres
Publié: (2023)
Training Language Models to Use Prolog as a Tool
par: Mellgren, Niklas, et autres
Publié: (2025)
par: Mellgren, Niklas, et autres
Publié: (2025)
The Limits of Obliviate: Evaluating Unlearning in LLMs via Stimulus-Knowledge Entanglement-Behavior Framework
par: Shah, Aakriti, et autres
Publié: (2025)
par: Shah, Aakriti, et autres
Publié: (2025)
Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation
par: Chen, Mu-Chi, et autres
Publié: (2026)
par: Chen, Mu-Chi, et autres
Publié: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
par: Wang, Yuchen, et autres
Publié: (2026)
par: Wang, Yuchen, et autres
Publié: (2026)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
par: Yousaf, Iqra
Publié: (2024)
par: Yousaf, Iqra
Publié: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
par: Raoufi, Behnam, et autres
Publié: (2025)
par: Raoufi, Behnam, et autres
Publié: (2025)
Test Case Features as Hyper-heuristics for Inductive Programming
par: McDaid, Edward, et autres
Publié: (2024)
par: McDaid, Edward, et autres
Publié: (2024)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
par: Li, Bowen, et autres
Publié: (2026)
par: Li, Bowen, et autres
Publié: (2026)
Learning to Select Goals in Automated Planning with Deep-Q Learning
par: Núñez-Molina, Carlos, et autres
Publié: (2024)
par: Núñez-Molina, Carlos, et autres
Publié: (2024)
Automated Generation of MDPs Using Logic Programming and LLMs for Robotic Applications
par: Saccon, Enrico, et autres
Publié: (2025)
par: Saccon, Enrico, et autres
Publié: (2025)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
par: Reddy, Sandeep, et autres
Publié: (2025)
par: Reddy, Sandeep, et autres
Publié: (2025)
Efficient Action-Constrained Reinforcement Learning via Acceptance-Rejection Method and Augmented MDPs
par: Hung, Wei, et autres
Publié: (2025)
par: Hung, Wei, et autres
Publié: (2025)
Deployment-Time Reliability of Learned Robot Policies
par: Agia, Christopher
Publié: (2026)
par: Agia, Christopher
Publié: (2026)
SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing
par: Meng, Zi, et autres
Publié: (2026)
par: Meng, Zi, et autres
Publié: (2026)
Gyan: An Explainable Neuro-Symbolic Language Model
par: Srinivasan, Venkat, et autres
Publié: (2026)
par: Srinivasan, Venkat, et autres
Publié: (2026)
Classification of Cattle Behavior and Detection of Heat (Estrus) using Sensor Data
par: Dhakshinamoorthy, Druva, et autres
Publié: (2025)
par: Dhakshinamoorthy, Druva, et autres
Publié: (2025)
Documents similaires
-
The Role of GitHub Copilot on Software Development: A Perspective on Productivity, Security, Best Practices and Future Directions
par: Nettur, Suresh Babu, et autres
Publié: (2025) -
UltraLightSqueezeNet: A Deep Learning Architecture for Malaria Classification with up to 54x fewer trainable parameters for resource constrained devices
par: Nettur, Suresh Babu, et autres
Publié: (2025) -
Lightweight Weighted Average Ensemble Model for Pneumonia Detection in Chest X-Ray Images
par: Nettur, Suresh Babu, et autres
Publié: (2025) -
A Hybrid Deep Learning CNN Model for Enhanced COVID-19 Detection from Computed Tomography (CT) Scan Images
par: Nettur, Suresh Babu, et autres
Publié: (2025) -
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
par: Ferenczi, Bryce, et autres
Publié: (2023)