LangBiTe: A Platform for Testing Bias in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Morales, Sergio, Clarisó, Robert, Cabot, Jordi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Framework to Model ML Engineering Processes
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
von: Morales, Sergio, et al.
Veröffentlicht: (2024)
Vibe-driven model-based engineering
von: Cabot, Jordi
Veröffentlicht: (2026)
von: Cabot, Jordi
Veröffentlicht: (2026)
AgentSLA : Towards a Service Level Agreement for AI Agents
von: Jouneaux, Gwendal, et al.
Veröffentlicht: (2025)
von: Jouneaux, Gwendal, et al.
Veröffentlicht: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025)
RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment
von: Fuster-Pena, Marcos, et al.
Veröffentlicht: (2025)
von: Fuster-Pena, Marcos, et al.
Veröffentlicht: (2025)
Metamorphic Testing of Large Language Models for Natural Language Processing
von: Cho, Steven, et al.
Veröffentlicht: (2025)
von: Cho, Steven, et al.
Veröffentlicht: (2025)
Large Language Models for Software Testing: A Research Roadmap
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
von: Augusto, Cristian, et al.
Veröffentlicht: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
von: Giramata, Suavis, et al.
Veröffentlicht: (2025)
The Software Diversity Card: A Framework for Reporting Diversity in Software Projects
von: Giner-Miguelez, Joan, et al.
Veröffentlicht: (2025)
von: Giner-Miguelez, Joan, et al.
Veröffentlicht: (2025)
A Tool for Generating Exceptional Behavior Tests With Large Language Models
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
von: Zhong, Linghan, et al.
Veröffentlicht: (2025)
EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages
von: Sharma, Aman, et al.
Veröffentlicht: (2026)
von: Sharma, Aman, et al.
Veröffentlicht: (2026)
Automating Structural Analysis Across Multiple Software Platforms Using Large Language Models
von: Geng, Ziheng, et al.
Veröffentlicht: (2026)
von: Geng, Ziheng, et al.
Veröffentlicht: (2026)
Can Language Models Pass Software Testing Certification Exams? a case study
von: Haq, Fitash Ul, et al.
Veröffentlicht: (2026)
von: Haq, Fitash Ul, et al.
Veröffentlicht: (2026)
A System for Automated Unit Test Generation Using Large Language Models and Assessment of Generated Test Suites
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
von: Lops, Andrea, et al.
Veröffentlicht: (2024)
Large Language Models as Test Case Generators: Performance Evaluation and Enhancement
von: Li, Kefan, et al.
Veröffentlicht: (2024)
von: Li, Kefan, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Enhancing the Understandability of Generated Unit Tests
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Deljouyi, Amirhossein, et al.
Veröffentlicht: (2024)
exLong: Generating Exceptional Behavior Tests with Large Language Models
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
von: Zhang, Jiyang, et al.
Veröffentlicht: (2024)
Navigating the Labyrinth: Path-Sensitive Unit Test Generation with Large Language Models
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
von: Liao, Dianshu, et al.
Veröffentlicht: (2025)
Turbulence: Systematically and Automatically Testing Instruction-Tuned Large Language Models for Code
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
von: Honarvar, Shahin, et al.
Veröffentlicht: (2023)
LangProp: A code optimization framework using Large Language Models applied to driving
von: Ishida, Shu, et al.
Veröffentlicht: (2024)
von: Ishida, Shu, et al.
Veröffentlicht: (2024)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
Automated User Story Generation with Test Case Specification Using Large Language Model
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
von: Rahman, Tajmilur, et al.
Veröffentlicht: (2024)
HFuzzer: Testing Large Language Models for Package Hallucinations via Phrase-based Fuzzing
von: Zhao, Yukai, et al.
Veröffentlicht: (2025)
von: Zhao, Yukai, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for the Generation of Unit Tests with Equivalence Partitions and Boundary Values
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
von: Rodríguez, Martín, et al.
Veröffentlicht: (2025)
LLMs are All You Need? Improving Fuzz Testing for MOJO with Large Language Models
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
von: Huang, Linghan, et al.
Veröffentlicht: (2025)
Low-Modeling of Software Systems
von: Cabot, Jordi
Veröffentlicht: (2024)
von: Cabot, Jordi
Veröffentlicht: (2024)
Vibe Modeling: Challenges and Opportunities
von: Cabot, Jordi
Veröffentlicht: (2025)
von: Cabot, Jordi
Veröffentlicht: (2025)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
DriveTester: A Unified Platform for Simulation-Based Autonomous Driving Testing
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
Comparison of Static Application Security Testing Tools and Large Language Models for Repo-level Vulnerability Detection
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
von: Zhou, Xin, et al.
Veröffentlicht: (2024)
Mitigating Gender Bias in Code Large Language Models via Model Editing
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
von: Qin, Zhanyue, et al.
Veröffentlicht: (2024)
LlamaRestTest: Effective REST API Testing with Small Language Models
von: Kim, Myeongsoo, et al.
Veröffentlicht: (2025)
von: Kim, Myeongsoo, et al.
Veröffentlicht: (2025)
Extracting Knowledge Graphs from User Stories using LangChain
von: da Silva, Thayná Camargo
Veröffentlicht: (2025)
von: da Silva, Thayná Camargo
Veröffentlicht: (2025)
The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
Impact of Code Context and Prompting Strategies on Automated Unit Test Generation with Modern General-Purpose Large Language Models
von: Walczak, Jakub, et al.
Veröffentlicht: (2025)
von: Walczak, Jakub, et al.
Veröffentlicht: (2025)
Assertion-Aware Test Code Summarization with Large Language Models
von: Mollah, Anamul Haque, et al.
Veröffentlicht: (2025)
von: Mollah, Anamul Haque, et al.
Veröffentlicht: (2025)
PromptPex: Automatic Test Generation for Language Model Prompts
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
von: Sharma, Reshabh K, et al.
Veröffentlicht: (2025)
SHERPA: A Model-Driven Framework for Large Language Model Execution
von: Chen, Boqi, et al.
Veröffentlicht: (2025)
von: Chen, Boqi, et al.
Veröffentlicht: (2025)
StackTrans: From Large Language Model to Large Pushdown Automata Model
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair
von: de-Fitero-Dominguez, David, et al.
Veröffentlicht: (2025)
von: de-Fitero-Dominguez, David, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Framework to Model ML Engineering Processes
von: Morales, Sergio, et al.
Veröffentlicht: (2024) -
Vibe-driven model-based engineering
von: Cabot, Jordi
Veröffentlicht: (2026) -
AgentSLA : Towards a Service Level Agreement for AI Agents
von: Jouneaux, Gwendal, et al.
Veröffentlicht: (2025) -
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
von: Dobslaw, Felix, et al.
Veröffentlicht: (2025) -
RePaCA: Leveraging Reasoning Large Language Models for Static Automated Patch Correctness Assessment
von: Fuster-Pena, Marcos, et al.
Veröffentlicht: (2025)