FairDeDup: Detecting and Mitigating Vision-Language Fairness Disparities in Semantic Dataset Deduplication
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Slyman, Eric, Lee, Stefan, Cohen, Scott, Kafle, Kushal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLSlice: Interactive Vision-and-Language Slice Discovery
von: Slyman, Eric, et al.
Veröffentlicht: (2023)
von: Slyman, Eric, et al.
Veröffentlicht: (2023)
Blocks Architecture (BloArk): Efficient, Cost-Effective, and Incremental Dataset Architecture for Wikipedia Revision History
von: Li, Lingxi, et al.
Veröffentlicht: (2024)
von: Li, Lingxi, et al.
Veröffentlicht: (2024)
Amazon Nova AI Challenge -- Trusted AI: Advancing secure, AI-assisted software development
von: Sahai, Sattvik, et al.
Veröffentlicht: (2025)
von: Sahai, Sattvik, et al.
Veröffentlicht: (2025)
Breaking the HISCO Barrier: Automatic Occupational Standardization with OccCANINE
von: Dahl, Christian Møller, et al.
Veröffentlicht: (2024)
von: Dahl, Christian Møller, et al.
Veröffentlicht: (2024)
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
LLMs and the Human Condition
von: Wallis, Peter
Veröffentlicht: (2024)
von: Wallis, Peter
Veröffentlicht: (2024)
Semantic Decomposition and Selective Context Filtering -- Text Processing Techniques for Context-Aware NLP-Based Systems
von: Villardar, Karl John
Veröffentlicht: (2025)
von: Villardar, Karl John
Veröffentlicht: (2025)
Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content
von: Koushik, Girish A., et al.
Veröffentlicht: (2025)
von: Koushik, Girish A., et al.
Veröffentlicht: (2025)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
Neural Machine Translation for Malayalam Paraphrase Generation
von: Varghese, Christeena, et al.
Veröffentlicht: (2024)
von: Varghese, Christeena, et al.
Veröffentlicht: (2024)
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
von: Lassche, Herman, et al.
Veröffentlicht: (2024)
von: Lassche, Herman, et al.
Veröffentlicht: (2024)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
von: Iida, Kurando, et al.
Veröffentlicht: (2024)
von: Iida, Kurando, et al.
Veröffentlicht: (2024)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
Self-Supervised Borrowing Detection on Multilingual Wordlists
von: Wientzek, Tim
Veröffentlicht: (2025)
von: Wientzek, Tim
Veröffentlicht: (2025)
Identifying Fairness Issues in Automatically Generated Testing Content
von: Stowe, Kevin, et al.
Veröffentlicht: (2024)
von: Stowe, Kevin, et al.
Veröffentlicht: (2024)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
von: Du, Bangde, et al.
Veröffentlicht: (2025)
von: Du, Bangde, et al.
Veröffentlicht: (2025)
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
von: Deng, Minghang, et al.
Veröffentlicht: (2025)
von: Deng, Minghang, et al.
Veröffentlicht: (2025)
FarsEval-PKBETS: A new diverse benchmark for evaluating Persian large language models
von: Shamsfard, Mehrnoush, et al.
Veröffentlicht: (2025)
von: Shamsfard, Mehrnoush, et al.
Veröffentlicht: (2025)
Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviews
von: Barkhordar, Ehsan, et al.
Veröffentlicht: (2026)
von: Barkhordar, Ehsan, et al.
Veröffentlicht: (2026)
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models
von: Gupta, Abhay, et al.
Veröffentlicht: (2025)
von: Gupta, Abhay, et al.
Veröffentlicht: (2025)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
Tversky Neural Networks: Psychologically Plausible Deep Learning with Differentiable Tversky Similarity
von: Doumbouya, Moussa Koulako Bala, et al.
Veröffentlicht: (2025)
von: Doumbouya, Moussa Koulako Bala, et al.
Veröffentlicht: (2025)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
von: Venkatasubramanian, Venkat, et al.
Veröffentlicht: (2024)
von: Venkatasubramanian, Venkat, et al.
Veröffentlicht: (2024)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
von: Blair-Stanek, Andrew, et al.
Veröffentlicht: (2023)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
von: Teixeira, Tiago, et al.
Veröffentlicht: (2026)
von: Teixeira, Tiago, et al.
Veröffentlicht: (2026)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
von: Banthia, Saumya, et al.
Veröffentlicht: (2020)
von: Banthia, Saumya, et al.
Veröffentlicht: (2020)
On Explaining with Attention Matrices
von: Naim, Omar, et al.
Veröffentlicht: (2024)
von: Naim, Omar, et al.
Veröffentlicht: (2024)
ALISON: Fast and Effective Stylometric Authorship Obfuscation
von: Xing, Eric, et al.
Veröffentlicht: (2024)
von: Xing, Eric, et al.
Veröffentlicht: (2024)
Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition
von: Goyal, Prasoon, et al.
Veröffentlicht: (2026)
von: Goyal, Prasoon, et al.
Veröffentlicht: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Automated Circuit Interpretation via Probe Prompting
von: Birardi, Giuseppe
Veröffentlicht: (2025)
von: Birardi, Giuseppe
Veröffentlicht: (2025)
Quaternion Convolutional Neural Networks: Current Advances and Future Directions
von: Altamirano-Gomez, Gerardo, et al.
Veröffentlicht: (2023)
von: Altamirano-Gomez, Gerardo, et al.
Veröffentlicht: (2023)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024)
von: Guo, Willis, et al.
Veröffentlicht: (2024)
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
von: Toker, Michael, et al.
Veröffentlicht: (2024)
von: Toker, Michael, et al.
Veröffentlicht: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VLSlice: Interactive Vision-and-Language Slice Discovery
von: Slyman, Eric, et al.
Veröffentlicht: (2023) -
Blocks Architecture (BloArk): Efficient, Cost-Effective, and Incremental Dataset Architecture for Wikipedia Revision History
von: Li, Lingxi, et al.
Veröffentlicht: (2024) -
Amazon Nova AI Challenge -- Trusted AI: Advancing secure, AI-assisted software development
von: Sahai, Sattvik, et al.
Veröffentlicht: (2025) -
Breaking the HISCO Barrier: Automatic Occupational Standardization with OccCANINE
von: Dahl, Christian Møller, et al.
Veröffentlicht: (2024) -
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
von: Williamson, Dane, et al.
Veröffentlicht: (2025)