Gespeichert in:
| Hauptverfasser: | Shah, Faiz Ali, Sabir, Ahmed, Sharma, Rajesh, Pfahl, Dietmar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2409.07162 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An LSTM-based Test Selection Method for Self-Driving Cars
von: Güllü, Ali, et al.
Veröffentlicht: (2025)
von: Güllü, Ali, et al.
Veröffentlicht: (2025)
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction
von: Motger, Quim, et al.
Veröffentlicht: (2024)
von: Motger, Quim, et al.
Veröffentlicht: (2024)
Evaluating the impact of code smell refactoring on the energy consumption of Android applications
von: Anwar, Hina, et al.
Veröffentlicht: (2025)
von: Anwar, Hina, et al.
Veröffentlicht: (2025)
Zero-shot Bilingual App Reviews Mining with Large Language Models
von: Wei, Jialiang, et al.
Veröffentlicht: (2023)
von: Wei, Jialiang, et al.
Veröffentlicht: (2023)
Effective Black Box Testing of Sentiment Analysis Classification Networks
von: Karbasizadeh, Parsa, et al.
Veröffentlicht: (2024)
von: Karbasizadeh, Parsa, et al.
Veröffentlicht: (2024)
Model Editing for LLMs4Code: How Far are We?
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
von: Li, Xiaopeng, et al.
Veröffentlicht: (2024)
Exploring a Test Data-Driven Method for Selecting and Constraining Metamorphic Relations
von: Duque-Torres, Alejandra, et al.
Veröffentlicht: (2023)
von: Duque-Torres, Alejandra, et al.
Veröffentlicht: (2023)
Teaching Simulation as a Research Method in Empirical Software Engineering
von: de França, Breno Bernard Nicolau, et al.
Veröffentlicht: (2025)
von: de França, Breno Bernard Nicolau, et al.
Veröffentlicht: (2025)
Leveraging LLMs for Grammar Adaptation: A Study on Metamodel-Grammar Co-Evolution
von: Zhang, Weixing, et al.
Veröffentlicht: (2026)
von: Zhang, Weixing, et al.
Veröffentlicht: (2026)
Beyond Keywords: A Context-based Hybrid Approach to Mining Ethical Concern-related App Reviews
von: Sorathiya, Aakash, et al.
Veröffentlicht: (2024)
von: Sorathiya, Aakash, et al.
Veröffentlicht: (2024)
StackEval: Benchmarking LLMs in Coding Assistance
von: Shah, Nidhish, et al.
Veröffentlicht: (2024)
von: Shah, Nidhish, et al.
Veröffentlicht: (2024)
Demystifying and Extracting Fault-indicating Information from Logs for Failure Diagnosis
von: Huang, Junjie, et al.
Veröffentlicht: (2024)
von: Huang, Junjie, et al.
Veröffentlicht: (2024)
A Critical Study of What Code-LLMs (Do Not) Learn
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
von: Anand, Abhinav, et al.
Veröffentlicht: (2024)
RePair: Automated Program Repair with Process-based Feedback
von: Zhao, Yuze, et al.
Veröffentlicht: (2024)
von: Zhao, Yuze, et al.
Veröffentlicht: (2024)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data
von: Wang, Yejie, et al.
Veröffentlicht: (2024)
von: Wang, Yejie, et al.
Veröffentlicht: (2024)
Towards Automatic Generation of Amplified Regression Test Oracles
von: Duque-Torres, Alejandra, et al.
Veröffentlicht: (2023)
von: Duque-Torres, Alejandra, et al.
Veröffentlicht: (2023)
How Effective are Generative Large Language Models in Performing Requirements Classification?
von: Alhoshan, Waad, et al.
Veröffentlicht: (2025)
von: Alhoshan, Waad, et al.
Veröffentlicht: (2025)
A Case Study of Web App Coding with OpenAI Reasoning Models
von: Cui, Yi
Veröffentlicht: (2024)
von: Cui, Yi
Veröffentlicht: (2024)
Do You Understand How I Feel?: Towards Verified Empathy in Therapy Chatbots
von: Dettori, Francesco, et al.
Veröffentlicht: (2026)
von: Dettori, Francesco, et al.
Veröffentlicht: (2026)
Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches
von: Shah, Syed Mehtab Hussain, et al.
Veröffentlicht: (2026)
von: Shah, Syed Mehtab Hussain, et al.
Veröffentlicht: (2026)
Sentiment Analysis in Software Engineering: Evaluating Generative Pre-trained Transformers
von: Saifullah, KM Khalid, et al.
Veröffentlicht: (2025)
von: Saifullah, KM Khalid, et al.
Veröffentlicht: (2025)
Benchmarking LLMs for Unit Test Generation from Real-World Functions
von: Huang, Dong, et al.
Veröffentlicht: (2025)
von: Huang, Dong, et al.
Veröffentlicht: (2025)
Multi-Programming Language Sandbox for LLMs
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
von: Dou, Shihan, et al.
Veröffentlicht: (2024)
SERA: Soft-Verified Efficient Repository Agents
von: Shen, Ethan, et al.
Veröffentlicht: (2026)
von: Shen, Ethan, et al.
Veröffentlicht: (2026)
EM-Assist: Safe Automated ExtractMethod Refactoring with LLMs
von: Pomian, Dorin, et al.
Veröffentlicht: (2024)
von: Pomian, Dorin, et al.
Veröffentlicht: (2024)
Isolating Language-Coding from Problem-Solving: Benchmarking LLMs with PseudoEval
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
von: Wu, Jiarong, et al.
Veröffentlicht: (2025)
Evaluation of Code LLMs on Geospatial Code Generation
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
von: Gramacki, Piotr, et al.
Veröffentlicht: (2024)
LLMs in Mobile Apps: Practices, Challenges, and Opportunities
von: Hau, Kimberly, et al.
Veröffentlicht: (2025)
von: Hau, Kimberly, et al.
Veröffentlicht: (2025)
Towards Extracting Ethical Concerns-related Software Requirements from App Reviews
von: Sorathiya, Aakash, et al.
Veröffentlicht: (2024)
von: Sorathiya, Aakash, et al.
Veröffentlicht: (2024)
CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding
von: Shi, Yuling, et al.
Veröffentlicht: (2026)
von: Shi, Yuling, et al.
Veröffentlicht: (2026)
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved?
von: Ahmed, Toufique, et al.
Veröffentlicht: (2024)
von: Ahmed, Toufique, et al.
Veröffentlicht: (2024)
Log Summarisation for Defect Evolution Analysis
von: Dolga, Rares, et al.
Veröffentlicht: (2024)
von: Dolga, Rares, et al.
Veröffentlicht: (2024)
MathDuels: Evaluating LLMs as Problem Posers and Solvers
von: Xu, Zhiqiu, et al.
Veröffentlicht: (2026)
von: Xu, Zhiqiu, et al.
Veröffentlicht: (2026)
DependEval: Benchmarking LLMs for Repository Dependency Understanding
von: Du, Junjia, et al.
Veröffentlicht: (2025)
von: Du, Junjia, et al.
Veröffentlicht: (2025)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
CodeReviewQA: The Code Review Comprehension Assessment for Large Language Models
von: Lin, Hong Yi, et al.
Veröffentlicht: (2025)
von: Lin, Hong Yi, et al.
Veröffentlicht: (2025)
NLPerturbator: Studying the Robustness of Code LLMs to Natural Language Variations
von: Chen, Junkai, et al.
Veröffentlicht: (2024)
von: Chen, Junkai, et al.
Veröffentlicht: (2024)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
von: Du, Yongkang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An LSTM-based Test Selection Method for Self-Driving Cars
von: Güllü, Ali, et al.
Veröffentlicht: (2025) -
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction
von: Motger, Quim, et al.
Veröffentlicht: (2024) -
Evaluating the impact of code smell refactoring on the energy consumption of Android applications
von: Anwar, Hina, et al.
Veröffentlicht: (2025) -
Zero-shot Bilingual App Reviews Mining with Large Language Models
von: Wei, Jialiang, et al.
Veröffentlicht: (2023) -
Effective Black Box Testing of Sentiment Analysis Classification Networks
von: Karbasizadeh, Parsa, et al.
Veröffentlicht: (2024)