Gespeichert in:
| Hauptverfasser: | Chishti, Mohd Sameen, Oyinloye, Damilare Peter, Li, Jingyue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.27789 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Feature-Centric Methodology for Analyzing Cross-Chain NFT Migration Compatibility
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026)
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026)
AgentReputation: A Decentralized Agentic AI Reputation Framework
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026)
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026)
A Proof of Success and Reward Distribution Protocol for Multi-bridge Architecture in Cross-chain Communication
von: Oyinloye, Damilare Peter, et al.
Veröffentlicht: (2025)
von: Oyinloye, Damilare Peter, et al.
Veröffentlicht: (2025)
Test It Before You Trust It: Applying Software Testing for Trustworthy In-context Learning
von: Racharak, Teeradaj, et al.
Veröffentlicht: (2025)
von: Racharak, Teeradaj, et al.
Veröffentlicht: (2025)
Parameter-Efficient Fine-Tuning of Large Language Models for Unit Test Generation: An Empirical Study
von: Storhaug, André, et al.
Veröffentlicht: (2024)
von: Storhaug, André, et al.
Veröffentlicht: (2024)
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
von: Li, Jingyue, et al.
Veröffentlicht: (2026)
von: Li, Jingyue, et al.
Veröffentlicht: (2026)
An LLM-based Quantitative Framework for Evaluating High-Stealthy Backdoor Risks in OSS Supply Chains
von: Yan, Zihe, et al.
Veröffentlicht: (2025)
von: Yan, Zihe, et al.
Veröffentlicht: (2025)
Repair-R1: Better Test Before Repair
von: Hu, Haichuan, et al.
Veröffentlicht: (2025)
von: Hu, Haichuan, et al.
Veröffentlicht: (2025)
You Name It, I Run It: An LLM Agent to Execute Tests of Arbitrary Projects
von: Bouzenia, Islem, et al.
Veröffentlicht: (2024)
von: Bouzenia, Islem, et al.
Veröffentlicht: (2024)
Call-Chain-Aware LLM-Based Test Generation for Java Projects
von: Wang, Guancheng, et al.
Veröffentlicht: (2026)
von: Wang, Guancheng, et al.
Veröffentlicht: (2026)
GitHub's Copilot Code Review: Can AI Spot Security Flaws Before You Commit?
von: Amro, Amena, et al.
Veröffentlicht: (2025)
von: Amro, Amena, et al.
Veröffentlicht: (2025)
Favia: Forensic Agent for Vulnerability-fix Identification and Analysis
von: Storhaug, André, et al.
Veröffentlicht: (2026)
von: Storhaug, André, et al.
Veröffentlicht: (2026)
You Don't Know Until You Click:Automated GUI Testing for Production-Ready Software Evaluation
von: Bian, Yutong, et al.
Veröffentlicht: (2025)
von: Bian, Yutong, et al.
Veröffentlicht: (2025)
LLM-Driven Kernel Evolution: Automating Driver Updates in Linux
von: Kharlamova, Arina, et al.
Veröffentlicht: (2025)
von: Kharlamova, Arina, et al.
Veröffentlicht: (2025)
Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models
von: Huang, Yuheng, et al.
Veröffentlicht: (2023)
von: Huang, Yuheng, et al.
Veröffentlicht: (2023)
Improving Code Translation with Syntax-Guided and Semantic-aware Preference Optimization
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
Bootstrapping Code Translation with Weighted Multilanguage Exploration
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
von: Wu, Yuhan, et al.
Veröffentlicht: (2026)
Identifying the Supply Chain of AI for Trustworthiness and Risk Management in Critical Applications
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
von: Sheh, Raymond K., et al.
Veröffentlicht: (2025)
Harden and Catch for Just-in-Time Assured LLM-Based Software Testing: Open Research Challenges
von: Harman, Mark, et al.
Veröffentlicht: (2025)
von: Harman, Mark, et al.
Veröffentlicht: (2025)
VerilogReader: LLM-Aided Hardware Test Generation
von: Ma, Ruiyang, et al.
Veröffentlicht: (2024)
von: Ma, Ruiyang, et al.
Veröffentlicht: (2024)
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
von: Arrieta, Aitor, et al.
Veröffentlicht: (2025)
von: Arrieta, Aitor, et al.
Veröffentlicht: (2025)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
CircuChain: Disentangling Competence and Compliance in LLM Circuit Analysis
von: Ravishankara, Mayank
Veröffentlicht: (2026)
von: Ravishankara, Mayank
Veröffentlicht: (2026)
Model-Enhanced LLM-Driven VUI Testing of VPA Apps
von: Li, Suwan, et al.
Veröffentlicht: (2024)
von: Li, Suwan, et al.
Veröffentlicht: (2024)
Large Language Model Supply Chain: Open Problems From the Security Perspective
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
Unveiling the Landscape of LLM Deployment in the Wild: An Empirical Study
von: Hou, Xinyi, et al.
Veröffentlicht: (2025)
von: Hou, Xinyi, et al.
Veröffentlicht: (2025)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
von: Cui, Yi
Veröffentlicht: (2025)
von: Cui, Yi
Veröffentlicht: (2025)
Deploy-Master: Automating the Deployment of 50,000+ Agent-Ready Scientific Tools in One Day
von: Wang, Yi, et al.
Veröffentlicht: (2026)
von: Wang, Yi, et al.
Veröffentlicht: (2026)
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics
von: Pathak, Aditya, et al.
Veröffentlicht: (2025)
von: Pathak, Aditya, et al.
Veröffentlicht: (2025)
DataGovBench: Benchmarking LLM Agents for Real-World Data Governance Workflows
von: Liu, Zhou, et al.
Veröffentlicht: (2025)
von: Liu, Zhou, et al.
Veröffentlicht: (2025)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyin, et al.
Veröffentlicht: (2024)
Can LLM Generate Regression Tests for Software Commits?
von: Liu, Jing, et al.
Veröffentlicht: (2025)
von: Liu, Jing, et al.
Veröffentlicht: (2025)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
LLM-Empowered Event-Chain Driven Code Generation for ADAS in SDV systems
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
von: Petrovic, Nenad, et al.
Veröffentlicht: (2025)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
von: Sun, Zhensu, et al.
Veröffentlicht: (2026)
RESTestBench: A Benchmark for Evaluating the Effectiveness of LLM-Generated REST API Test Cases from NL Requirements
von: Kogler, Leon, et al.
Veröffentlicht: (2026)
von: Kogler, Leon, et al.
Veröffentlicht: (2026)
Rethinking Testing for LLM Applications: Characteristics, Challenges, and a Lightweight Interaction Protocol
von: Ma, Wei, et al.
Veröffentlicht: (2025)
von: Ma, Wei, et al.
Veröffentlicht: (2025)
LLM-Based Robustness Testing of Microservice Applications: An Empirical Study
von: Tigulla, Hrushitha Goud, et al.
Veröffentlicht: (2026)
von: Tigulla, Hrushitha Goud, et al.
Veröffentlicht: (2026)
LLM-Based Automated Diagnosis Of Integration Test Failures At Google
von: Ziftci, Celal, et al.
Veröffentlicht: (2026)
von: Ziftci, Celal, et al.
Veröffentlicht: (2026)
Evaluating LLM-Based Test Generation Under Software Evolution
von: Haroon, Sabaat, et al.
Veröffentlicht: (2026)
von: Haroon, Sabaat, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Feature-Centric Methodology for Analyzing Cross-Chain NFT Migration Compatibility
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026) -
AgentReputation: A Decentralized Agentic AI Reputation Framework
von: Chishti, Mohd Sameen, et al.
Veröffentlicht: (2026) -
A Proof of Success and Reward Distribution Protocol for Multi-bridge Architecture in Cross-chain Communication
von: Oyinloye, Damilare Peter, et al.
Veröffentlicht: (2025) -
Test It Before You Trust It: Applying Software Testing for Trustworthy In-context Learning
von: Racharak, Teeradaj, et al.
Veröffentlicht: (2025) -
Parameter-Efficient Fine-Tuning of Large Language Models for Unit Test Generation: An Empirical Study
von: Storhaug, André, et al.
Veröffentlicht: (2024)