Meta-Fair: AI-Assisted Fairness Testing of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Romero-Arjona, Miguel, Parejo, José A., Alonso, Juan C., Sánchez, Ana B., Arrieta, Aitor, Segura, Sergio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ASTRAL: Automated Safety Testing of Large Language Models
by: Ugarte, Miriam, et al.
Published: (2025)
by: Ugarte, Miriam, et al.
Published: (2025)
Red Teaming Contemporary AI Models: Insights from Spanish and Basque Perspectives
by: Romero-Arjona, Miguel, et al.
Published: (2025)
by: Romero-Arjona, Miguel, et al.
Published: (2025)
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
by: Arrieta, Aitor, et al.
Published: (2025)
by: Arrieta, Aitor, et al.
Published: (2025)
The Rise of Language Models in Mining Software Repositories: A Survey
by: Romero-Arjona, Miguel, et al.
Published: (2026)
by: Romero-Arjona, Miguel, et al.
Published: (2026)
Metamorphic Testing of Vision-Language Action-Enabled Robots
by: Valle, Pablo, et al.
Published: (2026)
by: Valle, Pablo, et al.
Published: (2026)
o3-mini vs DeepSeek-R1: Which One is Safer?
by: Arrieta, Aitor, et al.
Published: (2025)
by: Arrieta, Aitor, et al.
Published: (2025)
An Empirical Evaluation of White-box and Black-box Test Case Prioritization Techniques in CPSs Modeled in Simulink
by: Arrieta, Aitor
Published: (2025)
by: Arrieta, Aitor
Published: (2025)
GenFair: Systematic Test Generation for Fairness Fault Detection in Large Language Models
by: Srinivasan, Madhusudan, et al.
Published: (2025)
by: Srinivasan, Madhusudan, et al.
Published: (2025)
Exploring the Potential of Large Language Models in Simulink-Stateflow Mutant Generation
by: Valle, Pablo, et al.
Published: (2026)
by: Valle, Pablo, et al.
Published: (2026)
How Fair is Software Fairness Testing?
by: Barcomb, Ann, et al.
Published: (2026)
by: Barcomb, Ann, et al.
Published: (2026)
A Tool for Benchmarking Large Language Models' Robustness in Assessing the Realism of Driving Scenarios
by: Wu, Jiahui, et al.
Published: (2025)
by: Wu, Jiahui, et al.
Published: (2025)
VISOR: A Vision-Language Model-based Test Oracle for Testing Robots
by: Saurabh, Prasun, et al.
Published: (2026)
by: Saurabh, Prasun, et al.
Published: (2026)
Reality Bites: Assessing the Realism of Driving Scenarios with Large Language Models
by: Wu, Jiahui, et al.
Published: (2024)
by: Wu, Jiahui, et al.
Published: (2024)
Vision Language Model-based Testing of Industrial Autonomous Mobile Robots
by: Wu, Jiahui, et al.
Published: (2025)
by: Wu, Jiahui, et al.
Published: (2025)
Software Fairness Testing in Practice
by: Santos, Ronnie de Souza, et al.
Published: (2025)
by: Santos, Ronnie de Souza, et al.
Published: (2025)
Foundation Models for the Digital Twin Creation of Cyber-Physical Systems
by: Ali, Shaukat, et al.
Published: (2024)
by: Ali, Shaukat, et al.
Published: (2024)
Search-based Automated Program Repair of CPS Controllers Modeled in Simulink-Stateflow
by: Arrieta, Aitor, et al.
Published: (2024)
by: Arrieta, Aitor, et al.
Published: (2024)
Testing and Evaluation of Large Language Models: Correctness, Non-Toxicity, and Fairness
by: Wang, Wenxuan
Published: (2024)
by: Wang, Wenxuan
Published: (2024)
Bias Ahead: Sensitive Prompts as Early Warnings for Fairness in Large Language Models
by: Voria, Gianmario, et al.
Published: (2026)
by: Voria, Gianmario, et al.
Published: (2026)
Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models
by: Xiao, Yisong, et al.
Published: (2025)
by: Xiao, Yisong, et al.
Published: (2025)
Toward Systematic Counterfactual Fairness Evaluation of Large Language Models: The CAFFE Framework
by: Parziale, Alessandra, et al.
Published: (2025)
by: Parziale, Alessandra, et al.
Published: (2025)
Bita: A Conversational Assistant for Fairness Testing
by: Johnson, Keeryn, et al.
Published: (2025)
by: Johnson, Keeryn, et al.
Published: (2025)
SATORI: Static Test Oracle Generation for REST APIs
by: Alonso, Juan C., et al.
Published: (2025)
by: Alonso, Juan C., et al.
Published: (2025)
Causally Perturbed Fairness Testing
by: Du, Chengwen, et al.
Published: (2025)
by: Du, Chengwen, et al.
Published: (2025)
Fairness Testing: A Comprehensive Survey and Analysis of Trends
by: Chen, Zhenpeng, et al.
Published: (2022)
by: Chen, Zhenpeng, et al.
Published: (2022)
Assessing Vision-Language Models for Perception in Autonomous Underwater Robotic Software
by: Yousaf, Muhammad, et al.
Published: (2026)
by: Yousaf, Muhammad, et al.
Published: (2026)
Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots
by: Valle, Pablo, et al.
Published: (2025)
by: Valle, Pablo, et al.
Published: (2025)
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection
by: Giramata, Suavis, et al.
Published: (2025)
by: Giramata, Suavis, et al.
Published: (2025)
Concolic Testing on Individual Fairness of Neural Network Models
by: Huang, Ming-I, et al.
Published: (2025)
by: Huang, Ming-I, et al.
Published: (2025)
Using Large Language Models to Develop Requirements Elicitation Skills
by: Lojo, Nelson, et al.
Published: (2025)
by: Lojo, Nelson, et al.
Published: (2025)
Automated Unit Test Improvement using Large Language Models at Meta
by: Alshahwan, Nadia, et al.
Published: (2024)
by: Alshahwan, Nadia, et al.
Published: (2024)
Reinforcement Learning for Testing Interdependent Requirements in Autonomous Vehicles: An Empirical Study
by: Wu, Jiahui, et al.
Published: (2025)
by: Wu, Jiahui, et al.
Published: (2025)
FairRF: Multi-Objective Search for Single and Intersectional Software Fairness
by: d'Alosio, Giordano, et al.
Published: (2026)
by: d'Alosio, Giordano, et al.
Published: (2026)
Search-based Generation of Waypoints for Triggering Self-Adaptations in Maritime Autonomous Vessels
by: Nylænder, Karoline, et al.
Published: (2025)
by: Nylænder, Karoline, et al.
Published: (2025)
Software Fairness Debt
by: Santos, Ronnie de Souza, et al.
Published: (2024)
by: Santos, Ronnie de Souza, et al.
Published: (2024)
Team Diversity Promotes Software Fairness: An Experiment on Fairness-Aware Requirements Prioritization
by: Magalhes, Cleyton, et al.
Published: (2026)
by: Magalhes, Cleyton, et al.
Published: (2026)
From Literature to Practice: Exploring Fairness Testing Tools for the Software Industry Adoption
by: Nguyen, Thanh, et al.
Published: (2024)
by: Nguyen, Thanh, et al.
Published: (2024)
Towards User-Focused Cross-Domain Testing: Disentangling Accessibility, Usability, and Fairness
by: Leça, Matheus de Morais, et al.
Published: (2025)
by: Leça, Matheus de Morais, et al.
Published: (2025)
Should AI Optimize Your Code? A Comparative Study of Classical Optimizing Compilers Versus Current Large Language Models
by: Rosas, Miguel Romero, et al.
Published: (2024)
by: Rosas, Miguel Romero, et al.
Published: (2024)
Quantum software experiments: A reporting and laboratory package structure guidelines
by: Moguel, Enrique, et al.
Published: (2024)
by: Moguel, Enrique, et al.
Published: (2024)
Similar Items
-
ASTRAL: Automated Safety Testing of Large Language Models
by: Ugarte, Miriam, et al.
Published: (2025) -
Red Teaming Contemporary AI Models: Insights from Spanish and Basque Perspectives
by: Romero-Arjona, Miguel, et al.
Published: (2025) -
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
by: Arrieta, Aitor, et al.
Published: (2025) -
The Rise of Language Models in Mining Software Repositories: A Survey
by: Romero-Arjona, Miguel, et al.
Published: (2026) -
Metamorphic Testing of Vision-Language Action-Enabled Robots
by: Valle, Pablo, et al.
Published: (2026)