Test It Before You Trust It: Applying Software Testing for Trustworthy In-context Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Racharak, Teeradaj, Ragkhitwetsagul, Chaiyong, Sontesadisai, Chommakorn, Sunetnanta, Thanwadee |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Human-interpretable Explanation in Code Clone Detection using LLM-based Post Hoc Explainer
by: Racharak, Teeradaj, et al.
Published: (2025)
by: Racharak, Teeradaj, et al.
Published: (2025)
The Impact of COVID-19 and Remote Work on Software Development in Thailand
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2025)
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2025)
On Categorizing Open Source Software Security Vulnerability Reporting Mechanisms on GitHub
by: Kancharoendee, Sushawapak, et al.
Published: (2025)
by: Kancharoendee, Sushawapak, et al.
Published: (2025)
Exploring the SECURITY.md in the Dependency Chain: Preliminary Analysis of the PyPI Ecosystem
by: Termphaiboon, Chayanid, et al.
Published: (2025)
by: Termphaiboon, Chayanid, et al.
Published: (2025)
See to Believe: Using Visualization To Motivate Updating Third-party Dependencies
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2024)
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2024)
PyGress: Tool for Analyzing the Progression of Code Proficiency in Python OSS Projects
by: Charatvaraphan, Rujiphart, et al.
Published: (2025)
by: Charatvaraphan, Rujiphart, et al.
Published: (2025)
jscefr: A Framework to Evaluate the Code Proficiency for JavaScript
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2024)
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2024)
Social Media Reactions to Open Source Promotions: AI-Powered GitHub Projects on Hacker News
by: Meakpaiboonwattana, Prachnachai, et al.
Published: (2025)
by: Meakpaiboonwattana, Prachnachai, et al.
Published: (2025)
Mining the Characteristics of Jupyter Notebooks in Data Science Projects
by: Choetkiertikul, Morakot, et al.
Published: (2023)
by: Choetkiertikul, Morakot, et al.
Published: (2023)
Quantifying Competitive Relationships Among Open-Source Software Projects
by: Takei, Yuki, et al.
Published: (2026)
by: Takei, Yuki, et al.
Published: (2026)
Generative AI for Requirements Engineering: A Systematic Literature Review
by: Cheng, Haowei, et al.
Published: (2024)
by: Cheng, Haowei, et al.
Published: (2024)
How the Misuse of a Dataset Harmed Semantic Clone Detection
by: Krinke, Jens, et al.
Published: (2025)
by: Krinke, Jens, et al.
Published: (2025)
Test Before You Deploy: Governing Updates in the LLM Supply Chain
by: Chishti, Mohd Sameen, et al.
Published: (2026)
by: Chishti, Mohd Sameen, et al.
Published: (2026)
You Don't Know Until You Click:Automated GUI Testing for Production-Ready Software Evaluation
by: Bian, Yutong, et al.
Published: (2025)
by: Bian, Yutong, et al.
Published: (2025)
ArgRE: Formal Argumentation for Conflict Resolution in Multi-Agent Requirements Negotiation
by: Cheng, Haowei, et al.
Published: (2026)
by: Cheng, Haowei, et al.
Published: (2026)
Repair-R1: Better Test Before Repair
by: Hu, Haichuan, et al.
Published: (2025)
by: Hu, Haichuan, et al.
Published: (2025)
The Role of Artificial Intelligence and Machine Learning in Software Testing
by: Ramadan, Ahmed, et al.
Published: (2024)
by: Ramadan, Ahmed, et al.
Published: (2024)
Reasoning-Based Software Testing
by: Giamattei, Luca, et al.
Published: (2023)
by: Giamattei, Luca, et al.
Published: (2023)
Simulator Ensembles for Trustworthy Autonomous Driving Testing
by: Sorokin, Lev, et al.
Published: (2025)
by: Sorokin, Lev, et al.
Published: (2025)
Fuzzy Inference System for Test Case Prioritization in Software Testing
by: Karatayev, Aron, et al.
Published: (2024)
by: Karatayev, Aron, et al.
Published: (2024)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
by: Gao, Shuzheng, et al.
Published: (2025)
by: Gao, Shuzheng, et al.
Published: (2025)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
by: Baqar, Mohammad, et al.
Published: (2024)
by: Baqar, Mohammad, et al.
Published: (2024)
Reinforcement Learning Integrated Agentic RAG for Software Test Cases Authoring
by: Hariharan, Mohanakrishnan
Published: (2025)
by: Hariharan, Mohanakrishnan
Published: (2025)
Can LLM Generate Regression Tests for Software Commits?
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Breaking Barriers in Software Testing: The Power of AI-Driven Automation
by: Naqvi, Saba, et al.
Published: (2025)
by: Naqvi, Saba, et al.
Published: (2025)
Large Language Models for Software Testing: A Research Roadmap
by: Augusto, Cristian, et al.
Published: (2025)
by: Augusto, Cristian, et al.
Published: (2025)
Evaluating LLM-Based Test Generation Under Software Evolution
by: Haroon, Sabaat, et al.
Published: (2026)
by: Haroon, Sabaat, et al.
Published: (2026)
The Potential of LLMs in Automating Software Testing: From Generation to Reporting
by: Sherifi, Betim, et al.
Published: (2024)
by: Sherifi, Betim, et al.
Published: (2024)
You Name It, I Run It: An LLM Agent to Execute Tests of Arbitrary Projects
by: Bouzenia, Islem, et al.
Published: (2024)
by: Bouzenia, Islem, et al.
Published: (2024)
Agentic AI Software Engineers: Programming with Trust
by: Roychoudhury, Abhik, et al.
Published: (2025)
by: Roychoudhury, Abhik, et al.
Published: (2025)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
by: Huang, Linghan, et al.
Published: (2025)
by: Huang, Linghan, et al.
Published: (2025)
Expectations vs Reality -- A Secondary Study on AI Adoption in Software Testing
by: Karhu, Katja, et al.
Published: (2025)
by: Karhu, Katja, et al.
Published: (2025)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
by: Trae Research Team, et al.
Published: (2025)
by: Trae Research Team, et al.
Published: (2025)
Agentic RAG for Software Testing with Hybrid Vector-Graph and Multi-Agent Orchestration
by: Hariharan, Mohanakrishnan, et al.
Published: (2025)
by: Hariharan, Mohanakrishnan, et al.
Published: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
by: Dobslaw, Felix, et al.
Published: (2025)
by: Dobslaw, Felix, et al.
Published: (2025)
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
by: Naqvi, Saba, et al.
Published: (2026)
by: Naqvi, Saba, et al.
Published: (2026)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
by: Chen, Zhi, et al.
Published: (2026)
by: Chen, Zhi, et al.
Published: (2026)
The Impact of Software Testing with Quantum Optimization Meets Machine Learning
by: Bandarupalli, Gopichand
Published: (2025)
by: Bandarupalli, Gopichand
Published: (2025)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
by: Fu, Jia, et al.
Published: (2025)
by: Fu, Jia, et al.
Published: (2025)
Automating a Complete Software Test Process Using LLMs: An Automotive Case Study
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Similar Items
-
Towards Human-interpretable Explanation in Code Clone Detection using LLM-based Post Hoc Explainer
by: Racharak, Teeradaj, et al.
Published: (2025) -
The Impact of COVID-19 and Remote Work on Software Development in Thailand
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2025) -
On Categorizing Open Source Software Security Vulnerability Reporting Mechanisms on GitHub
by: Kancharoendee, Sushawapak, et al.
Published: (2025) -
Exploring the SECURITY.md in the Dependency Chain: Preliminary Analysis of the PyPI Ecosystem
by: Termphaiboon, Chayanid, et al.
Published: (2025) -
See to Believe: Using Visualization To Motivate Updating Third-party Dependencies
by: Ragkhitwetsagul, Chaiyong, et al.
Published: (2024)