ABFS: Natural Robustness Testing for LLM-based NLP Software
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Mingxuan, Xiao, Yan, Ji, Shunhui, Li, Yunhe, Xue, Lei, Zhang, Pengcheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Assessing the Robustness of LLM-based NLP Software via Automated Testing
por: Xiao, Mingxuan, et al.
Publicado: (2024)
por: Xiao, Mingxuan, et al.
Publicado: (2024)
BASFuzz: Towards Robustness Evaluation of LLM-based NLP Software via Automated Fuzz Testing
por: Xiao, Mingxuan, et al.
Publicado: (2025)
por: Xiao, Mingxuan, et al.
Publicado: (2025)
RITFIS: Robust input testing framework for LLMs-based intelligent software
por: Xiao, Mingxuan, et al.
Publicado: (2024)
por: Xiao, Mingxuan, et al.
Publicado: (2024)
HGA: Heuristic Black‐Box Test Case Generation for NLP Intelligent Software
por: Mingxuan Xiao, et al.
Publicado: (2026)
por: Mingxuan Xiao, et al.
Publicado: (2026)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
por: Xue, Pengyu, et al.
Publicado: (2024)
por: Xue, Pengyu, et al.
Publicado: (2024)
Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
por: Trae Research Team, et al.
Publicado: (2025)
por: Trae Research Team, et al.
Publicado: (2025)
Fine-grained Testing for Autonomous Driving Software: a Study on Autoware with LLM-driven Unit Testing
por: Wang, Wenhan, et al.
Publicado: (2025)
por: Wang, Wenhan, et al.
Publicado: (2025)
Empirical Insights of Test Selection Metrics under Multiple Testing Objectives and Distribution Shifts
por: Zhang, Jingyu, et al.
Publicado: (2026)
por: Zhang, Jingyu, et al.
Publicado: (2026)
Search-based Robustness Testing of Laptop Refurbishing Robotic Software
por: Isaku, Erblin, et al.
Publicado: (2026)
por: Isaku, Erblin, et al.
Publicado: (2026)
Detecting Flaky Tests in Quantum Software: A Dynamic Approach
por: Kim, Dongchan, et al.
Publicado: (2025)
por: Kim, Dongchan, et al.
Publicado: (2025)
Testing the Untestable? An Empirical Study on the Testing Process of LLM-Powered Software Systems
por: Magalhaes, Cleyton, et al.
Publicado: (2025)
por: Magalhaes, Cleyton, et al.
Publicado: (2025)
SGDL: Smart contract vulnerability generation via deep learning
por: Hanting Chu, et al.
Publicado: (2024)
por: Hanting Chu, et al.
Publicado: (2024)
Software Testing in the Quantum World
por: Abreu, Rui, et al.
Publicado: (2026)
por: Abreu, Rui, et al.
Publicado: (2026)
Benchmarking and Studying the LLM-based Agent System in End-to-End Software Development
por: Zeng, Zhengran, et al.
Publicado: (2025)
por: Zeng, Zhengran, et al.
Publicado: (2025)
SGAgent: Suggestion-Guided LLM-Based Multi-Agent Framework for Repository-Level Software Repair
por: Zhang, Quanjun, et al.
Publicado: (2026)
por: Zhang, Quanjun, et al.
Publicado: (2026)
The Nature of Technical Debt in Research Software
por: Ernst, Neil A., et al.
Publicado: (2026)
por: Ernst, Neil A., et al.
Publicado: (2026)
Test vs Mutant: Adversarial LLM Agents for Robust Unit Test Generation
por: Chang, Pengyu, et al.
Publicado: (2026)
por: Chang, Pengyu, et al.
Publicado: (2026)
S3LLM: Large-Scale Scientific Software Understanding with LLMs using Source, Metadata, and Document
por: Shaik, Kareem, et al.
Publicado: (2024)
por: Shaik, Kareem, et al.
Publicado: (2024)
Walk the Talk: Is Your Log-based Software Reliability Maintenance System Really Reliable?
por: He, Minghua, et al.
Publicado: (2025)
por: He, Minghua, et al.
Publicado: (2025)
From Requirements to Test Cases: An NLP-Based Approach for High-Performance ECU Test Case Automation
por: Medeshetty, Nikitha, et al.
Publicado: (2025)
por: Medeshetty, Nikitha, et al.
Publicado: (2025)
Self-Organizing Multi-Agent Systems for Continuous Software Development
por: Lyu, Wenhan, et al.
Publicado: (2026)
por: Lyu, Wenhan, et al.
Publicado: (2026)
KTester: Leveraging Domain and Testing Knowledge for More Effective LLM-based Test Generation
por: Li, Anji, et al.
Publicado: (2025)
por: Li, Anji, et al.
Publicado: (2025)
Test Smell: A Parasitic Energy Consumer in Software Testing
por: Misu, Md Rakib Hossain, et al.
Publicado: (2023)
por: Misu, Md Rakib Hossain, et al.
Publicado: (2023)
Combining Tests and Proofs for Better Software Verification
por: Huang, Li, et al.
Publicado: (2026)
por: Huang, Li, et al.
Publicado: (2026)
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
por: Xiao, Yisong, et al.
Publicado: (2023)
por: Xiao, Yisong, et al.
Publicado: (2023)
Multitask-based Evaluation of Open-Source LLM on Software Vulnerability
por: Yin, Xin, et al.
Publicado: (2024)
por: Yin, Xin, et al.
Publicado: (2024)
IntrinTrans: LLM-based Intrinsic Code Translator for RISC-V Vector
por: Han, Liutong, et al.
Publicado: (2025)
por: Han, Liutong, et al.
Publicado: (2025)
Software Fairness Testing in Practice
por: Santos, Ronnie de Souza, et al.
Publicado: (2025)
por: Santos, Ronnie de Souza, et al.
Publicado: (2025)
Can LLM Generate Regression Tests for Software Commits?
por: Liu, Jing, et al.
Publicado: (2025)
por: Liu, Jing, et al.
Publicado: (2025)
Bridging Natural Language and Formal Specification--Automated Translation of Software Requirements to LTL via Hierarchical Semantics Decomposition Using LLMs
por: Ma, Zhi, et al.
Publicado: (2025)
por: Ma, Zhi, et al.
Publicado: (2025)
TestART: Improving LLM-based Unit Testing via Co-evolution of Automated Generation and Repair Iteration
por: Gu, Siqi, et al.
Publicado: (2024)
por: Gu, Siqi, et al.
Publicado: (2024)
SemOpt: LLM-Driven Code Optimization via Rule-Based Analysis
por: Zhao, Yuwei, et al.
Publicado: (2025)
por: Zhao, Yuwei, et al.
Publicado: (2025)
LLM-based Unit Test Generation via Property Retrieval
por: Zhang, Zhe, et al.
Publicado: (2024)
por: Zhang, Zhe, et al.
Publicado: (2024)
Testing Is Not Boring: Characterizing Challenge in Software Testing Tasks
por: Hardman, Davi Gama, et al.
Publicado: (2025)
por: Hardman, Davi Gama, et al.
Publicado: (2025)
Investigating Software Aging in LLM-Generated Software Systems
por: Santos, César, et al.
Publicado: (2025)
por: Santos, César, et al.
Publicado: (2025)
LLMs Are Not a Silver Bullet: A Case Study on Software Fairness
por: Li, Xinyue, et al.
Publicado: (2026)
por: Li, Xinyue, et al.
Publicado: (2026)
CREF: An LLM-based Conversational Software Repair Framework for Programming Tutors
por: Yang, Boyang, et al.
Publicado: (2024)
por: Yang, Boyang, et al.
Publicado: (2024)
Fuzzing-based Mutation Testing of C/C++ Software in Cyber-Physical Systems
por: Lee, Jaekwon, et al.
Publicado: (2025)
por: Lee, Jaekwon, et al.
Publicado: (2025)
Search-based Software Testing Driven by Domain Knowledge: Reflections and New Perspectives
por: Formica, Federico, et al.
Publicado: (2025)
por: Formica, Federico, et al.
Publicado: (2025)
From Natural Language to Executable Properties for Property-based Testing of Mobile Apps
por: Xiong, Yiheng, et al.
Publicado: (2026)
por: Xiong, Yiheng, et al.
Publicado: (2026)
Ejemplares similares
-
Assessing the Robustness of LLM-based NLP Software via Automated Testing
por: Xiao, Mingxuan, et al.
Publicado: (2024) -
BASFuzz: Towards Robustness Evaluation of LLM-based NLP Software via Automated Fuzz Testing
por: Xiao, Mingxuan, et al.
Publicado: (2025) -
RITFIS: Robust input testing framework for LLMs-based intelligent software
por: Xiao, Mingxuan, et al.
Publicado: (2024) -
HGA: Heuristic Black‐Box Test Case Generation for NLP Intelligent Software
por: Mingxuan Xiao, et al.
Publicado: (2026) -
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
por: Xue, Pengyu, et al.
Publicado: (2024)