humancompatible.detect: a Python Toolkit for Detecting Bias in AI Models
Fuente:
arXiv
Saved in:
| Main Authors: | Matilla, German M., Nemecek, Jiri, Kryvoviaz, Illia, Marecek, Jakub |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias Detection via Maximum Subgroup Discrepancy
by: Němeček, Jiří, et al.
Published: (2025)
by: Němeček, Jiří, et al.
Published: (2025)
Intersectional Fairness via Mixed-Integer Optimization
by: Němeček, Jiří, et al.
Published: (2026)
by: Němeček, Jiří, et al.
Published: (2026)
Sample Complexity of Bias Detection with Subsampled Point-to-Subspace Distances
by: Matilla, German Martinez, et al.
Published: (2025)
by: Matilla, German Martinez, et al.
Published: (2025)
Fast, close, non-singular and property-preserving approximations of entropic measures
by: Horenko, Illia, et al.
Published: (2025)
by: Horenko, Illia, et al.
Published: (2025)
seqme: a Python library for evaluating biological sequence design
by: Møller-Larsen, Rasmus, et al.
Published: (2025)
by: Møller-Larsen, Rasmus, et al.
Published: (2025)
Creativity in the Age of AI: Rethinking the Role of Intentional Agency
by: Pearson, James S., et al.
Published: (2026)
by: Pearson, James S., et al.
Published: (2026)
Toward a Dynamic Stackelberg Game-Theoretic Framework for Agentic AI Defense Against LLM Jailbreaking
by: Han, Zhengye, et al.
Published: (2025)
by: Han, Zhengye, et al.
Published: (2025)
Modeling Clinical Concern Trajectories in Language Model Agents
by: Subaharan, Sukesh, et al.
Published: (2026)
by: Subaharan, Sukesh, et al.
Published: (2026)
Benchmarking PNW Model for MedMNIST to 100% Accuracy
by: Deng, Bo
Published: (2026)
by: Deng, Bo
Published: (2026)
humancompatible.interconnect: Testing Properties of Repeated Uses of Interconnections of AI Systems
by: Nazarov, Rodion, et al.
Published: (2025)
by: Nazarov, Rodion, et al.
Published: (2025)
The Hidden Costs of AI: A Review of Energy, E-Waste, and Inequality in Model Development
by: Winsta, Jenis
Published: (2025)
by: Winsta, Jenis
Published: (2025)
How well can a large language model explain business processes as perceived by users?
by: Fahland, Dirk, et al.
Published: (2024)
by: Fahland, Dirk, et al.
Published: (2024)
Mathematical reasoning and the computer
by: Buzzard, Kevin
Published: (2025)
by: Buzzard, Kevin
Published: (2025)
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
by: Sadjoli, Nicholas, et al.
Published: (2026)
by: Sadjoli, Nicholas, et al.
Published: (2026)
BernGraph: Probabilistic Graph Neural Networks for EHR-based Medication Recommendations
by: Piao, Xihao, et al.
Published: (2024)
by: Piao, Xihao, et al.
Published: (2024)
Generating Likely Counterfactuals Using Sum-Product Networks
by: Nemecek, Jiri, et al.
Published: (2024)
by: Nemecek, Jiri, et al.
Published: (2024)
Improving the Validity of Decision Trees as Explanations
by: Nemecek, Jiri, et al.
Published: (2023)
by: Nemecek, Jiri, et al.
Published: (2023)
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
by: Wasenmüller, Robert, et al.
Published: (2024)
by: Wasenmüller, Robert, et al.
Published: (2024)
Beyond the Org Chart: AI and the Transformation of Invisible Work
by: Rosenthal, Stephanie, et al.
Published: (2026)
by: Rosenthal, Stephanie, et al.
Published: (2026)
Modelling Human Values for AI Reasoning
by: Osman, Nardine, et al.
Published: (2024)
by: Osman, Nardine, et al.
Published: (2024)
Resilient Federated Chain: Transforming Blockchain Consensus into an Active Defense Layer for Federated Learning
by: García-Márquez, Mario, et al.
Published: (2026)
by: García-Márquez, Mario, et al.
Published: (2026)
Authenticated Delegation and Authorized AI Agents
by: South, Tobin, et al.
Published: (2025)
by: South, Tobin, et al.
Published: (2025)
From Language Models to Practical Self-Improving Computer Agents
by: Sheng, Alex
Published: (2024)
by: Sheng, Alex
Published: (2024)
AI Consciousness is Inevitable: A Theoretical Computer Science Perspective
by: Blum, Lenore, et al.
Published: (2024)
by: Blum, Lenore, et al.
Published: (2024)
Adaptive Orchestration for Large-Scale Inference on Heterogeneous Accelerator Systems Balancing Cost, Performance, and Resilience
by: Biran, Yahav, et al.
Published: (2025)
by: Biran, Yahav, et al.
Published: (2025)
Tape: A Cellular Automata Benchmark for Evaluating Rule-Shift Generalization in Reinforcement Learning
by: Pan, Enze
Published: (2026)
by: Pan, Enze
Published: (2026)
RASP-Tuner: Retrieval-Augmented Soft Prompts for Context-Aware Black-Box Optimization in Non-Stationary Environments
by: Pan, Enze
Published: (2026)
by: Pan, Enze
Published: (2026)
A Theoretical Analysis of Soft-Label vs Hard-Label Training in Neural Networks
by: Mandal, Saptarshi, et al.
Published: (2024)
by: Mandal, Saptarshi, et al.
Published: (2024)
Leveraging Diversity in Online Interactions
by: Osman, Nardine, et al.
Published: (2023)
by: Osman, Nardine, et al.
Published: (2023)
ConSensus: Multi-Agent Collaboration for Multimodal Sensing
by: Yoon, Hyungjun, et al.
Published: (2026)
by: Yoon, Hyungjun, et al.
Published: (2026)
AI Governance InternationaL Evaluation Index (AGILE Index) 2024
by: Zeng, Yi, et al.
Published: (2025)
by: Zeng, Yi, et al.
Published: (2025)
Auditable Homomorphic-based Decentralized Collaborative AI with Attribute-based Differential Privacy
by: Yeh, Lo-Yao, et al.
Published: (2024)
by: Yeh, Lo-Yao, et al.
Published: (2024)
Towards a Reliable Offline Personal AI Assistant for Long Duration Spaceflight
by: Bensch, Oliver, et al.
Published: (2024)
by: Bensch, Oliver, et al.
Published: (2024)
Charting the Future of Scholarly Knowledge with AI: A Community Perspective
by: Jiomekong, Azanzi, et al.
Published: (2025)
by: Jiomekong, Azanzi, et al.
Published: (2025)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
by: Aksoy, Sinan G., et al.
Published: (2026)
by: Aksoy, Sinan G., et al.
Published: (2026)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
by: Abhishek, Alok, et al.
Published: (2025)
by: Abhishek, Alok, et al.
Published: (2025)
Watermarking Large Language Models in Europe: Interpreting the AI Act in Light of Technology
by: Souverain, Thomas
Published: (2025)
by: Souverain, Thomas
Published: (2025)
Application of machine learning for infrastructure reconstruction programs management
by: Khudiakov, Illia, et al.
Published: (2025)
by: Khudiakov, Illia, et al.
Published: (2025)
AI-FLARES: Artificial Intelligence for the Analysis of Solar Flares Data
by: Piana, Michele, et al.
Published: (2024)
by: Piana, Michele, et al.
Published: (2024)
Piecewise Polynomial Regression of Tame Functions via Integer Programming
by: Bareilles, Gilles, et al.
Published: (2023)
by: Bareilles, Gilles, et al.
Published: (2023)
Similar Items
-
Bias Detection via Maximum Subgroup Discrepancy
by: Němeček, Jiří, et al.
Published: (2025) -
Intersectional Fairness via Mixed-Integer Optimization
by: Němeček, Jiří, et al.
Published: (2026) -
Sample Complexity of Bias Detection with Subsampled Point-to-Subspace Distances
by: Matilla, German Martinez, et al.
Published: (2025) -
Fast, close, non-singular and property-preserving approximations of entropic measures
by: Horenko, Illia, et al.
Published: (2025) -
seqme: a Python library for evaluating biological sequence design
by: Møller-Larsen, Rasmus, et al.
Published: (2025)