Textual Data Bias Detection and Mitigation -- An Extensible Pipeline with Experimental Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Görge, Rebekka, Gannamaneni, Sujan Sai, Naeven, Tabea, Abdelwahab, Hammam, Allende-Cid, Héctor, Cremers, Armin B., Helmer, Lennard, Mock, Michael, Schmitz, Anna, Xue, Songkai, Yildirir, Elif, Poretschkin, Maximilian, Wrobel, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Developing trustworthy AI applications with foundation models
by: Mock, Michael, et al.
Published: (2024)
by: Mock, Michael, et al.
Published: (2024)
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
by: Mueller, Felix B, et al.
Published: (2024)
by: Mueller, Felix B, et al.
Published: (2024)
Detecting Linguistic Indicators for Stereotype Assessment with Large Language Models
by: Görge, Rebekka, et al.
Published: (2025)
by: Görge, Rebekka, et al.
Published: (2025)
Enabling Extensible Embodied Capabilities with Tools
by: Zhou, Xueyang, et al.
Published: (2026)
by: Zhou, Xueyang, et al.
Published: (2026)
A Lightweight and Extensible Cell Segmentation and Classification Model for Whole Slide Images
by: Shvetsov, Nikita, et al.
Published: (2025)
by: Shvetsov, Nikita, et al.
Published: (2025)
Storing Less, Finding More: How Novelty Filtering Improves Cross-Modal Retrieval on Edge Cameras
by: Abdelwahab, Sherif
Published: (2026)
by: Abdelwahab, Sherif
Published: (2026)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
by: Zhang, Li, et al.
Published: (2026)
by: Zhang, Li, et al.
Published: (2026)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
by: Radosky, Lukas, et al.
Published: (2026)
by: Radosky, Lukas, et al.
Published: (2026)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
by: Guo, Willis, et al.
Published: (2024)
by: Guo, Willis, et al.
Published: (2024)
Tiny models from tiny data: Textual and null-text inversion for few-shot distillation
by: Landolsi, Erik, et al.
Published: (2024)
by: Landolsi, Erik, et al.
Published: (2024)
A Scalable Pipeline Combining Procedural 3D Graphics and Guided Diffusion for Photorealistic Synthetic Training Data Generation in White Button Mushroom Segmentation
by: Károly, Artúr I., et al.
Published: (2025)
by: Károly, Artúr I., et al.
Published: (2025)
Pipeline and Dataset Generation for Automated Fact-checking in Almost Any Language
by: Drchal, Jan, et al.
Published: (2023)
by: Drchal, Jan, et al.
Published: (2023)
Composition of rocks and minerals from the Mid-Atlantic Ridge 5-7°N
by: Pushcharovsky, Yury M, et al.
Published: (2004)
by: Pushcharovsky, Yury M, et al.
Published: (2004)
Chemical composition of basalts and basaltic glasses from the Sierra Leone Fracture Zone region
by: Skolotnev, Sergey G, et al.
Published: (2003)
by: Skolotnev, Sergey G, et al.
Published: (2003)
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
by: Wang, Renxi, et al.
Published: (2025)
by: Wang, Renxi, et al.
Published: (2025)
Physical oceanography during L' Atalante cruise Almofront-1
by: Prieur, Louis Marie
Published: (2012)
by: Prieur, Louis Marie
Published: (2012)
Neural Reasoning Networks: Efficient Interpretable Neural Networks With Automatic Textual Explanations
by: Carrow, Stephen, et al.
Published: (2024)
by: Carrow, Stephen, et al.
Published: (2024)
A Computer Vision Pipeline for Iterative Bullet Hole Tracking in Rifle Zeroing
by: Belcher, Robert M., et al.
Published: (2026)
by: Belcher, Robert M., et al.
Published: (2026)
Data Processing for the OpenGPT-X Model Family
by: Brandizzi, Nicolo', et al.
Published: (2024)
by: Brandizzi, Nicolo', et al.
Published: (2024)
CWM: Contrastive World Models for Action Feasibility Learning in Embodied Agent Pipelines
by: Banerjee, Chayan
Published: (2026)
by: Banerjee, Chayan
Published: (2026)
Chemical and isotopic compositions of basalts and glasses from lavas of the Sierra Leone fault site
by: Sharkov, E V, et al.
Published: (2008)
by: Sharkov, E V, et al.
Published: (2008)
(Table 1) Rock types dredged at stations of R/V Akademik Nikolaj Strakhov (Cruise 22) and R/V Akademik Ioffe (Cruise 10), Mid-Atlantic Ridge 5-7°N
by: Pushcharovsky, Yury M, et al.
Published: (2004)
by: Pushcharovsky, Yury M, et al.
Published: (2004)
Nutrients measured on water bottle samples during L'Atalante cruise Almofront-1
by: Prieur, Louis Marie
Published: (2012)
by: Prieur, Louis Marie
Published: (2012)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
REVOLVE: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization
by: Zhang, Peiyan, et al.
Published: (2024)
by: Zhang, Peiyan, et al.
Published: (2024)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
by: Maryanskyy, Artem
Published: (2026)
by: Maryanskyy, Artem
Published: (2026)
Evaluating Voice Command Pipelines for Drone Control: From STT and LLM to Direct Classification and Siamese Networks
by: Simões, Lucca Emmanuel Pineli, et al.
Published: (2024)
by: Simões, Lucca Emmanuel Pineli, et al.
Published: (2024)
Reenact Anything: Semantic Video Motion Transfer Using Motion-Textual Inversion
by: Kansy, Manuel, et al.
Published: (2024)
by: Kansy, Manuel, et al.
Published: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
by: Gupta, Sunny, et al.
Published: (2024)
by: Gupta, Sunny, et al.
Published: (2024)
FLD+: Data-efficient Evaluation Metric for Generative Models
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
Normalizing Flow-Based Metric for Image Generation
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
by: Kashyap, Pankhi, et al.
Published: (2024)
by: Kashyap, Pankhi, et al.
Published: (2024)
Exploring Uncertainty Visualization for Degenerate Tensors in 3D Symmetric Second-Order Tensor Field Ensembles
by: Schmitz, Tadea, et al.
Published: (2024)
by: Schmitz, Tadea, et al.
Published: (2024)
Pigments measured on water bottle samples during cruise ALMOFRONT-1
by: Prieur, Louis Marie, et al.
Published: (2012)
by: Prieur, Louis Marie, et al.
Published: (2012)
Unsupervised Acquisition of Discrete Grammatical Categories
by: Shakouri, David Ph., et al.
Published: (2025)
by: Shakouri, David Ph., et al.
Published: (2025)
A Knowledge-Based Language Model: Deducing Grammatical Knowledge in a Multi-Agent Language Acquisition Simulation
by: Shakouri, David Ph., et al.
Published: (2025)
by: Shakouri, David Ph., et al.
Published: (2025)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
by: Semenov, Andrei, et al.
Published: (2024)
by: Semenov, Andrei, et al.
Published: (2024)
Synthetic-Child: An AIGC-Based Synthetic Data Pipeline for Privacy-Preserving Child Posture Estimation
by: Zeng, Taowen
Published: (2026)
by: Zeng, Taowen
Published: (2026)
IMUVIE: Pickup Timeline Action Localization via Motion Movies
by: Clapham, John, et al.
Published: (2024)
by: Clapham, John, et al.
Published: (2024)
Similar Items
-
Developing trustworthy AI applications with foundation models
by: Mock, Michael, et al.
Published: (2024) -
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
by: Mueller, Felix B, et al.
Published: (2024) -
Detecting Linguistic Indicators for Stereotype Assessment with Large Language Models
by: Görge, Rebekka, et al.
Published: (2025) -
Enabling Extensible Embodied Capabilities with Tools
by: Zhou, Xueyang, et al.
Published: (2026) -
A Lightweight and Extensible Cell Segmentation and Classification Model for Whole Slide Images
by: Shvetsov, Nikita, et al.
Published: (2025)