Hamster: A Large-Scale Study and Characterization of Developer-Written Tests
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Rangeet, Stennett, Tyler, Pavuluri, Raju, Levin, Nate, Orso, Alessandro, Sinha, Saurabh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAINT: Service-level Integration Test Generation with Program Analysis and LLM-based Agents
by: Pan, Rangeet, et al.
Published: (2025)
by: Pan, Rangeet, et al.
Published: (2025)
Sakura: An Approach for Generating Complex Tests from Natural Language Test Descriptions
by: Stennett, Tyler, et al.
Published: (2026)
by: Stennett, Tyler, et al.
Published: (2026)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025)
by: Stennett, Tyler, et al.
Published: (2025)
Leveraging Large Language Models to Improve REST API Testing
by: Kim, Myeongsoo, et al.
Published: (2023)
by: Kim, Myeongsoo, et al.
Published: (2023)
A Multi-Agent Approach for REST API Testing with Semantic Graphs and LLM-Driven Inputs
by: Kim, Myeongsoo, et al.
Published: (2024)
by: Kim, Myeongsoo, et al.
Published: (2024)
ASTER: Natural and Multi-language Unit Test Generation with LLMs
by: Pan, Rangeet, et al.
Published: (2024)
by: Pan, Rangeet, et al.
Published: (2024)
Codellm-Devkit: A Framework for Contextualizing Code LLMs with Program Analysis Insights
by: Krishna, Rahul, et al.
Published: (2024)
by: Krishna, Rahul, et al.
Published: (2024)
LlamaRestTest: Effective REST API Testing with Small Language Models
by: Kim, Myeongsoo, et al.
Published: (2025)
by: Kim, Myeongsoo, et al.
Published: (2025)
Lost in Translation: A Study of Bugs Introduced by Large Language Models while Translating Code
by: Pan, Rangeet, et al.
Published: (2023)
by: Pan, Rangeet, et al.
Published: (2023)
Usage, Effects and Requirements for AI Coding Assistants in the Enterprise: An Empirical Study
by: Vukovic, Maja, et al.
Published: (2026)
by: Vukovic, Maja, et al.
Published: (2026)
Otter: Generating Tests from Issues to Validate SWE Patches
by: Ahmed, Toufique, et al.
Published: (2025)
by: Ahmed, Toufique, et al.
Published: (2025)
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved?
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
Advancing Automated In-Isolation Validation in Repository-Level Code Translation
by: Ke, Kaiyao, et al.
Published: (2025)
by: Ke, Kaiyao, et al.
Published: (2025)
AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validation
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
by: Pavuluri, Advait, et al.
Published: (2026)
by: Pavuluri, Advait, et al.
Published: (2026)
Improving Program Debloating with 1-DU Chain Minimality
by: Kim, Myeongsoo, et al.
Published: (2024)
by: Kim, Myeongsoo, et al.
Published: (2024)
Human-Written vs. AI-Generated Code: A Large-Scale Study of Defects, Vulnerabilities, and Complexity
by: Cotroneo, Domenico, et al.
Published: (2025)
by: Cotroneo, Domenico, et al.
Published: (2025)
CRANE: Constrained Reasoning Injection for Code Agents via Nullspace Editing
by: Zhu, Mingzhi, et al.
Published: (2026)
by: Zhu, Mingzhi, et al.
Published: (2026)
Characterizing Bugs and Quality Attributes in Quantum Software: A Large-Scale Empirical Study
by: Yousuf, Mir Mohammad, et al.
Published: (2025)
by: Yousuf, Mir Mohammad, et al.
Published: (2025)
A Large-Scale Study of Call Graph-based Impact Prediction using Mutation Testing
by: Musco, Vincenzo, et al.
Published: (2018)
by: Musco, Vincenzo, et al.
Published: (2018)
A Large-Scale Study on Developer Engagement and Expertise in Configurable Software System Projects
by: Milano, Karolina M., et al.
Published: (2025)
by: Milano, Karolina M., et al.
Published: (2025)
Guidelines to Prompt Large Language Models for Code Generation: An Empirical Characterization
by: Midolo, Alessandro, et al.
Published: (2026)
by: Midolo, Alessandro, et al.
Published: (2026)
Generative AI in Simulation-Based Test Environments for Large-Scale Cyber-Physical Systems: An Industrial Study
by: Sadrnezhaad, Masoud, et al.
Published: (2025)
by: Sadrnezhaad, Masoud, et al.
Published: (2025)
What Makes Software Bugs Escape Testing? Evidence from a Large-Scale Empirical Study
by: Cotroneo, Domenico, et al.
Published: (2026)
by: Cotroneo, Domenico, et al.
Published: (2026)
Deriving and Validating Requirements Engineering Principles for Large-Scale Agile Development: An Industrial Longitudinal Study
by: Saeeda, Hina, et al.
Published: (2026)
by: Saeeda, Hina, et al.
Published: (2026)
Testing Is Not Boring: Characterizing Challenge in Software Testing Tasks
by: Hardman, Davi Gama, et al.
Published: (2025)
by: Hardman, Davi Gama, et al.
Published: (2025)
VISOR: A Vision-Language Model-based Test Oracle for Testing Robots
by: Saurabh, Prasun, et al.
Published: (2026)
by: Saurabh, Prasun, et al.
Published: (2026)
A Large-Scale Study on the Development and Issues of Multi-Agent AI Systems
by: Liu, Daniel, et al.
Published: (2026)
by: Liu, Daniel, et al.
Published: (2026)
A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing
by: Shang, Ye, et al.
Published: (2024)
by: Shang, Ye, et al.
Published: (2024)
Customer Validation, Feedback and Collaboration in Large-Scale Continuous Software Development
by: Molamphy, David
Published: (2025)
by: Molamphy, David
Published: (2025)
Towards a Taxonomy for Autonomy in Large-Scale Agile Software Development
by: Lassenius, Casper, et al.
Published: (2025)
by: Lassenius, Casper, et al.
Published: (2025)
Software Testing with Large Language Models: An Interview Study with Practitioners
by: Santana, Maria Deolinda, et al.
Published: (2025)
by: Santana, Maria Deolinda, et al.
Published: (2025)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
by: Liang, Yunhao, et al.
Published: (2026)
by: Liang, Yunhao, et al.
Published: (2026)
Is LLM-Generated Code More Maintainable \& Reliable than Human-Written Code?
by: Molison, Alfred Santa, et al.
Published: (2025)
by: Molison, Alfred Santa, et al.
Published: (2025)
Understanding and Characterizing Mock Assertions in Unit Tests
by: Zhu, Hengcheng, et al.
Published: (2025)
by: Zhu, Hengcheng, et al.
Published: (2025)
Governing the Commons: Code Ownership and Code-Clones in Large-Scale Software Development
by: Sundelin, Anders, et al.
Published: (2024)
by: Sundelin, Anders, et al.
Published: (2024)
Harnessing Large Language Model for Virtual Reality Exploration Testing: A Case Study
by: Qi, Zhenyu, et al.
Published: (2025)
by: Qi, Zhenyu, et al.
Published: (2025)
Acceptance Test Generation with Large Language Models: An Industrial Case Study
by: Ferreira, Margarida, et al.
Published: (2025)
by: Ferreira, Margarida, et al.
Published: (2025)
Measuring the Runtime Performance of C++ Code Written by Humans using GitHub Copilot
by: Erhabor, Daniel, et al.
Published: (2023)
by: Erhabor, Daniel, et al.
Published: (2023)
Scaling Mobile Chaos Testing with AI-Driven Test Execution
by: Marcano, Juan, et al.
Published: (2026)
by: Marcano, Juan, et al.
Published: (2026)
Similar Items
-
SAINT: Service-level Integration Test Generation with Program Analysis and LLM-based Agents
by: Pan, Rangeet, et al.
Published: (2025) -
Sakura: An Approach for Generating Complex Tests from Natural Language Test Descriptions
by: Stennett, Tyler, et al.
Published: (2026) -
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
by: Stennett, Tyler, et al.
Published: (2025) -
Leveraging Large Language Models to Improve REST API Testing
by: Kim, Myeongsoo, et al.
Published: (2023) -
A Multi-Agent Approach for REST API Testing with Semantic Graphs and LLM-Driven Inputs
by: Kim, Myeongsoo, et al.
Published: (2024)