IndiCASA: A Dataset and Bias Evaluation Framework in LLMs Using Contrastive Embedding Similarity in the Indian Context
Fuente:
arXiv
Saved in:
| Main Authors: | S, Santhosh G, S, Akshay Govind, Krishnan, Gokul S, Ravindran, Balaraman, Natarajan, Sriraam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
mFARM: Towards Multi-Faceted Fairness Assessment based on HARMs in Clinical Decision Support
by: Adappanavar, Shreyash, et al.
Published: (2025)
by: Adappanavar, Shreyash, et al.
Published: (2025)
LExT: Towards Evaluating Trustworthiness of Natural Language Explanations
by: Shailya, Krithi, et al.
Published: (2025)
by: Shailya, Krithi, et al.
Published: (2025)
AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs
by: S, Santhosh G, et al.
Published: (2025)
by: S, Santhosh G, et al.
Published: (2025)
Where Should I Study? Biased Language Models Decide! Evaluating Fairness in LMs for Academic Recommendations
by: Shailya, Krithi, et al.
Published: (2025)
by: Shailya, Krithi, et al.
Published: (2025)
SWAN: Sparse Winnowed Attention for Reduced Inference Memory via Decompression-Free KV-Cache Compression
by: S, Santhosh G, et al.
Published: (2025)
by: S, Santhosh G, et al.
Published: (2025)
Are Models Trained on Indian Legal Data Fair?
by: Girhepuje, Sahil, et al.
Published: (2023)
by: Girhepuje, Sahil, et al.
Published: (2023)
InSaAF: Incorporating Safety through Accuracy and Fairness | Are LLMs ready for the Indian Legal Domain?
by: Tripathi, Yogesh, et al.
Published: (2024)
by: Tripathi, Yogesh, et al.
Published: (2024)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
Participatory Approaches in AI Development and Governance: Case Studies
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
Unified Threat Detection and Mitigation Framework (UTDMF): Combating Prompt Injection, Deception, and Bias in Enterprise-Scale Transformers
by: KumarRavindran, Santhosh
Published: (2025)
by: KumarRavindran, Santhosh
Published: (2025)
Contradiction Detection in RAG Systems: Evaluating LLMs as Context Validators for Improved Information Consistency
by: Gokul, Vignesh, et al.
Published: (2025)
by: Gokul, Vignesh, et al.
Published: (2025)
Participatory Approaches in AI Development and Governance: A Principled Approach
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
by: Parthasarathy, Ambreesh, et al.
Published: (2024)
Learning Interpretable Models Using Uncertainty Oracles
by: Ghose, Abhishek, et al.
Published: (2019)
by: Ghose, Abhishek, et al.
Published: (2019)
FactEHR: A Dataset for Evaluating Factuality in Clinical Notes Using LLMs
by: Munnangi, Monica, et al.
Published: (2024)
by: Munnangi, Monica, et al.
Published: (2024)
FairI Tales: Evaluation of Fairness in Indian Contexts with a Focus on Bias and Stereotypes
by: Nawale, Janki Atul, et al.
Published: (2025)
by: Nawale, Janki Atul, et al.
Published: (2025)
Bias Mitigation or Cultural Commonsense? Evaluating LLMs with a Japanese Dataset
by: Yamamoto, Taisei, et al.
Published: (2025)
by: Yamamoto, Taisei, et al.
Published: (2025)
Enhanced Sign Language Translation between American Sign Language (ASL) and Indian Sign Language (ISL) Using LLMs
by: Kumar, Malay, et al.
Published: (2024)
by: Kumar, Malay, et al.
Published: (2024)
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs
by: Ao, Shuang, et al.
Published: (2024)
by: Ao, Shuang, et al.
Published: (2024)
Building Expressive and Tractable Probabilistic Generative Models: A Review
by: Sidheekh, Sahil, et al.
Published: (2024)
by: Sidheekh, Sahil, et al.
Published: (2024)
Geometry-Aware Probabilistic Circuits via Voronoi Tessellations
by: Sidheekh, Sahil, et al.
Published: (2026)
by: Sidheekh, Sahil, et al.
Published: (2026)
Moral Anchor System: A Predictive Framework for AI Value Alignment and Drift Prevention
by: Ravindran, Santhosh Kumar
Published: (2025)
by: Ravindran, Santhosh Kumar
Published: (2025)
IndiVec: An Exploration of Leveraging Large Language Models for Media Bias Detection with Fine-Grained Bias Indicators
by: Lin, Luyang, et al.
Published: (2024)
by: Lin, Luyang, et al.
Published: (2024)
Evaluating Metrics for Bias in Word Embeddings
by: Schröder, Sarah, et al.
Published: (2021)
by: Schröder, Sarah, et al.
Published: (2021)
Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers
by: Ravindran, Santhosh Kumar
Published: (2025)
by: Ravindran, Santhosh Kumar
Published: (2025)
MITL Model Checking via Generalized Timed Automata and a New Liveness Algorithm
by: Akshay, S., et al.
Published: (2024)
by: Akshay, S., et al.
Published: (2024)
Simulations for Event-Clock Automata
by: Akshay, S, et al.
Published: (2022)
by: Akshay, S, et al.
Published: (2022)
Leveraging In-Context Learning for Political Bias Testing of LLMs
by: Haller, Patrick, et al.
Published: (2025)
by: Haller, Patrick, et al.
Published: (2025)
Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics
by: Acharjee, Jashaswimalya, et al.
Published: (2026)
by: Acharjee, Jashaswimalya, et al.
Published: (2026)
Profiling Bias in LLMs: Stereotype Dimensions in Contextual Word Embeddings
by: Schuster, Carolin M., et al.
Published: (2024)
by: Schuster, Carolin M., et al.
Published: (2024)
SentiCSE: A Sentiment-aware Contrastive Sentence Embedding Framework with Sentiment-guided Textual Similarity
by: Kim, Jaemin, et al.
Published: (2024)
by: Kim, Jaemin, et al.
Published: (2024)
BharatBBQ: A Multilingual Bias Benchmark for Question Answering in the Indian Context
by: Tomar, Aditya, et al.
Published: (2025)
by: Tomar, Aditya, et al.
Published: (2025)
Building Trust in Clinical LLMs: Bias Analysis and Dataset Transparency
by: Maslenkova, Svetlana, et al.
Published: (2025)
by: Maslenkova, Svetlana, et al.
Published: (2025)
Program Synthesis for Non-Linear Real Arithmetic: Going Beyond Realizability
by: Akshay, S., et al.
Published: (2026)
by: Akshay, S., et al.
Published: (2026)
Assessing the Reliability of LLMs Annotations in the Context of Demographic Bias and Model Explanation
by: Mohammadi, Hadi, et al.
Published: (2025)
by: Mohammadi, Hadi, et al.
Published: (2025)
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
by: Jeong, Hyejun, et al.
Published: (2024)
by: Jeong, Hyejun, et al.
Published: (2024)
TRACE: TRansformer-based Attribution using Contrastive Embeddings in LLMs
by: Wang, Cheng, et al.
Published: (2024)
by: Wang, Cheng, et al.
Published: (2024)
TEMPORA: Efficient Verification of Metric Temporal Properties with Past in Pointwise Semantics
by: Akshay, S., et al.
Published: (2025)
by: Akshay, S., et al.
Published: (2025)
Language Models can Subtly Deceive Without Lying: A Case Study on Strategic Phrasing in Legislation
by: Dogra, Atharvan, et al.
Published: (2024)
by: Dogra, Atharvan, et al.
Published: (2024)
Extending the Context of Pretrained LLMs by Dropping Their Positional Embeddings
by: Gelberg, Yoav, et al.
Published: (2025)
by: Gelberg, Yoav, et al.
Published: (2025)
Using Similarity to Evaluate Factual Consistency in Summaries
by: Ye, Yuxuan, et al.
Published: (2024)
by: Ye, Yuxuan, et al.
Published: (2024)
Similar Items
-
mFARM: Towards Multi-Faceted Fairness Assessment based on HARMs in Clinical Decision Support
by: Adappanavar, Shreyash, et al.
Published: (2025) -
LExT: Towards Evaluating Trustworthiness of Natural Language Explanations
by: Shailya, Krithi, et al.
Published: (2025) -
AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs
by: S, Santhosh G, et al.
Published: (2025) -
Where Should I Study? Biased Language Models Decide! Evaluating Fairness in LMs for Academic Recommendations
by: Shailya, Krithi, et al.
Published: (2025) -
SWAN: Sparse Winnowed Attention for Reduced Inference Memory via Decompression-Free KV-Cache Compression
by: S, Santhosh G, et al.
Published: (2025)