Effective Black Box Testing of Sentiment Analysis Classification Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Karbasizadeh, Parsa, Faghih, Fathiyeh, Golshanrad, Pouria |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effective Targeted Testing of Smart Contracts
by: Fooladgar, Mahdi, et al.
Published: (2024)
by: Fooladgar, Mahdi, et al.
Published: (2024)
DeepCover: Advancing RNN Test Coverage and Online Error Prediction using State Machine Extraction
by: Golshanrad, Pouria, et al.
Published: (2024)
by: Golshanrad, Pouria, et al.
Published: (2024)
How Effective are Generative Large Language Models in Performing Requirements Classification?
by: Alhoshan, Waad, et al.
Published: (2025)
by: Alhoshan, Waad, et al.
Published: (2025)
Path Analysis for Effective Fault Localization in Deep Neural Networks
by: Hashemifar, Soroush, et al.
Published: (2023)
by: Hashemifar, Soroush, et al.
Published: (2023)
Effective Harness Engineering for Algorithm Discovery with Coding Agents
by: Ishibashi, Yoichi, et al.
Published: (2026)
by: Ishibashi, Yoichi, et al.
Published: (2026)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
by: Sutawika, Lintang, et al.
Published: (2026)
by: Sutawika, Lintang, et al.
Published: (2026)
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
by: Amayuelas, Alfonso, et al.
Published: (2026)
by: Amayuelas, Alfonso, et al.
Published: (2026)
Testing the Effect of Code Documentation on Large Language Model Code Understanding
by: Macke, William, et al.
Published: (2024)
by: Macke, William, et al.
Published: (2024)
Testing and Evaluation of Large Language Models: Correctness, Non-Toxicity, and Fairness
by: Wang, Wenxuan
Published: (2024)
by: Wang, Wenxuan
Published: (2024)
ACECODER: Acing Coder RL via Automated Test-Case Synthesis
by: Zeng, Huaye, et al.
Published: (2025)
by: Zeng, Huaye, et al.
Published: (2025)
Evaluating LLMs on Sequential API Call Through Automated Test Generation
by: Huang, Yuheng, et al.
Published: (2025)
by: Huang, Yuheng, et al.
Published: (2025)
aiXcoder-7B: A Lightweight and Effective Large Language Model for Code Processing
by: Jiang, Siyuan, et al.
Published: (2024)
by: Jiang, Siyuan, et al.
Published: (2024)
B4: Towards Optimal Assessment of Plausible Code Solutions with Plausible Tests
by: Chen, Mouxiang, et al.
Published: (2024)
by: Chen, Mouxiang, et al.
Published: (2024)
Text2Scenario: Text-Driven Scenario Generation for Autonomous Driving Test
by: Cai, Xuan, et al.
Published: (2025)
by: Cai, Xuan, et al.
Published: (2025)
Satori-SWE: Evolutionary Test-Time Scaling for Sample-Efficient Software Engineering
by: Zeng, Guangtao, et al.
Published: (2025)
by: Zeng, Guangtao, et al.
Published: (2025)
The Prompt Alchemist: Automated LLM-Tailored Prompt Optimization for Test Case Generation
by: Gao, Shuzheng, et al.
Published: (2025)
by: Gao, Shuzheng, et al.
Published: (2025)
Does Model Size Matter? A Comparison of Small and Large Language Models for Requirements Classification
by: Zadenoori, Mohammad Amin, et al.
Published: (2025)
by: Zadenoori, Mohammad Amin, et al.
Published: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
by: Giramata, Suavis, et al.
Published: (2025)
by: Giramata, Suavis, et al.
Published: (2025)
How Toxic Can You Get? Search-based Toxicity Testing for Large Language Models
by: Corbo, Simone, et al.
Published: (2025)
by: Corbo, Simone, et al.
Published: (2025)
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems
by: Cao, Yuhan, et al.
Published: (2025)
by: Cao, Yuhan, et al.
Published: (2025)
Bias Testing and Mitigation in Black Box LLMs using Metamorphic Relations
by: Salimian, Sina, et al.
Published: (2025)
by: Salimian, Sina, et al.
Published: (2025)
Process-Centric Analysis of Agentic Software Systems
by: Liu, Shuyang, et al.
Published: (2025)
by: Liu, Shuyang, et al.
Published: (2025)
LLMs for Science: Usage for Code Generation and Data Analysis
by: Nejjar, Mohamed, et al.
Published: (2023)
by: Nejjar, Mohamed, et al.
Published: (2023)
Metamorphic Testing for Audio Content Moderation Software
by: Wang, Wenxuan, et al.
Published: (2025)
by: Wang, Wenxuan, et al.
Published: (2025)
Source Code Foundation Models are Transferable Binary Analysis Knowledge Bases
by: Su, Zian, et al.
Published: (2024)
by: Su, Zian, et al.
Published: (2024)
Automated Business Process Analysis: An LLM-Based Approach to Value Assessment
by: De Michele, William, et al.
Published: (2025)
by: De Michele, William, et al.
Published: (2025)
Large Language Models (LLMs) for Source Code Analysis: applications, models and datasets
by: Jelodar, Hamed, et al.
Published: (2025)
by: Jelodar, Hamed, et al.
Published: (2025)
AdaptiveLog: An Adaptive Log Analysis Framework with the Collaboration of Large and Small Language Model
by: Ma, Lipeng, et al.
Published: (2025)
by: Ma, Lipeng, et al.
Published: (2025)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
by: Asgari, Ali, et al.
Published: (2025)
by: Asgari, Ali, et al.
Published: (2025)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
REPOT: Recoverable Program-of-Thought via Checkpoint Repair
by: Mazaheri, Parsa
Published: (2026)
by: Mazaheri, Parsa
Published: (2026)
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
by: De Koning, Milan, et al.
Published: (2026)
by: De Koning, Milan, et al.
Published: (2026)
Looking into Black Box Code Language Models
by: Haider, Muhammad Umair, et al.
Published: (2024)
by: Haider, Muhammad Umair, et al.
Published: (2024)
DeepLL: Considering Linear Logic for the Analysis of Deep Learning Experiments
by: Papoulias, Nick
Published: (2024)
by: Papoulias, Nick
Published: (2024)
WebTestBench: Evaluating Computer-Use Agents towards End-to-End Automated Web Testing
by: Kong, Fanheng, et al.
Published: (2026)
by: Kong, Fanheng, et al.
Published: (2026)
Scaling Test-Time Compute for Agentic Coding
by: Kim, Joongwon, et al.
Published: (2026)
by: Kim, Joongwon, et al.
Published: (2026)
Learning to Generate Unit Tests for Automated Debugging
by: Prasad, Archiki, et al.
Published: (2025)
by: Prasad, Archiki, et al.
Published: (2025)
Applying Large Language Models API to Issue Classification Problem
by: Aracena, Gabriel, et al.
Published: (2024)
by: Aracena, Gabriel, et al.
Published: (2024)
Semantic-Preserving Transformations as Mutation Operators: A Study on Their Effectiveness in Defect Detection
by: Hort, Max, et al.
Published: (2025)
by: Hort, Max, et al.
Published: (2025)
ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal
by: Zhang, Haonan, et al.
Published: (2025)
by: Zhang, Haonan, et al.
Published: (2025)
Similar Items
-
Effective Targeted Testing of Smart Contracts
by: Fooladgar, Mahdi, et al.
Published: (2024) -
DeepCover: Advancing RNN Test Coverage and Online Error Prediction using State Machine Extraction
by: Golshanrad, Pouria, et al.
Published: (2024) -
How Effective are Generative Large Language Models in Performing Requirements Classification?
by: Alhoshan, Waad, et al.
Published: (2025) -
Path Analysis for Effective Fault Localization in Deep Neural Networks
by: Hashemifar, Soroush, et al.
Published: (2023) -
Effective Harness Engineering for Algorithm Discovery with Coding Agents
by: Ishibashi, Yoichi, et al.
Published: (2026)