On The Role of Reasoning in the Identification of Subtle Stereotypes in Natural Language
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Jacob-Junqi, Dige, Omkar, Emerson, D. B., Khattak, Faiza Khan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
by: Kohankhaki, Farnaz, et al.
Published: (2024)
by: Kohankhaki, Farnaz, et al.
Published: (2024)
Fairness Certification for Natural Language Processing and Large Language Models
by: Freiberger, Vincent, et al.
Published: (2024)
by: Freiberger, Vincent, et al.
Published: (2024)
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
by: Delaval, Axel, et al.
Published: (2025)
by: Delaval, Axel, et al.
Published: (2025)
Mitigating Social Biases in Language Models through Unlearning
by: Dige, Omkar, et al.
Published: (2024)
by: Dige, Omkar, et al.
Published: (2024)
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
by: Truong, Sang T., et al.
Published: (2024)
by: Truong, Sang T., et al.
Published: (2024)
A Primer on Large Language Models and their Limitations
by: Johnson, Sandra, et al.
Published: (2024)
by: Johnson, Sandra, et al.
Published: (2024)
Acceptable Use Policies for Foundation Models
by: Klyman, Kevin
Published: (2024)
by: Klyman, Kevin
Published: (2024)
CritiSense: Critical Digital Literacy and Resilience Against Misinformation
by: Alam, Firoj, et al.
Published: (2026)
by: Alam, Firoj, et al.
Published: (2026)
Quantifying Fairness in LLMs Beyond Tokens: A Semantic and Statistical Perspective
by: Xu, Weijie, et al.
Published: (2025)
by: Xu, Weijie, et al.
Published: (2025)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Change My Frame: Reframing in the Wild in r/ChangeMyView
by: Peguero, Arturo Martínez, et al.
Published: (2024)
by: Peguero, Arturo Martínez, et al.
Published: (2024)
HInter: Exposing Hidden Intersectional Bias in Large Language Models
by: Souani, Badr, et al.
Published: (2025)
by: Souani, Badr, et al.
Published: (2025)
Transforming and Combining Rewards for Aligning Large Language Models
by: Wang, Zihao, et al.
Published: (2024)
by: Wang, Zihao, et al.
Published: (2024)
Curriculum Recommendations Using Transformer Base Model with InfoNCE Loss And Language Switching Method
by: Xu, Xiaonan, et al.
Published: (2024)
by: Xu, Xiaonan, et al.
Published: (2024)
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering
by: Zong, Chang, et al.
Published: (2025)
by: Zong, Chang, et al.
Published: (2025)
Watermarking Large Language Models in Europe: Interpreting the AI Act in Light of Technology
by: Souverain, Thomas
Published: (2025)
by: Souverain, Thomas
Published: (2025)
Value Lens: Using Large Language Models to Understand Human Values
by: Fernández, Eduardo de la Cruz, et al.
Published: (2025)
by: Fernández, Eduardo de la Cruz, et al.
Published: (2025)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
by: Miliani, Martina, et al.
Published: (2025)
by: Miliani, Martina, et al.
Published: (2025)
A Study on Bias Detection and Classification in Natural Language Processing
by: Evans, Ana Sofia, et al.
Published: (2024)
by: Evans, Ana Sofia, et al.
Published: (2024)
The Illusion of Role Separation: Hidden Shortcuts in LLM Role Learning (and How to Fix Them)
by: Wang, Zihao, et al.
Published: (2025)
by: Wang, Zihao, et al.
Published: (2025)
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis
by: Petrov, Nikolay B, et al.
Published: (2024)
by: Petrov, Nikolay B, et al.
Published: (2024)
The Impact of Role Design in In-Context Learning for Large Language Models
by: Rouzegar, Hamidreza, et al.
Published: (2025)
by: Rouzegar, Hamidreza, et al.
Published: (2025)
Soft-prompt Tuning for Large Language Models to Evaluate Bias
by: Tian, Jacob-Junqi, et al.
Published: (2023)
by: Tian, Jacob-Junqi, et al.
Published: (2023)
LLMs as Deceptive Agents: How Role-Based Prompting Induces Semantic Ambiguity in Puzzle Tasks
by: Yoo, Seunghyun
Published: (2025)
by: Yoo, Seunghyun
Published: (2025)
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
by: Rakshit, Supantho, et al.
Published: (2025)
by: Rakshit, Supantho, et al.
Published: (2025)
T-VEC: A Telecom-Specific Vectorization Model with Enhanced Semantic Understanding via Deep Triplet Loss Fine-Tuning
by: Ethiraj, Vignesh, et al.
Published: (2025)
by: Ethiraj, Vignesh, et al.
Published: (2025)
Auto prompt sql: a resource-efficient architecture for text-to-sql translation in constrained environments
by: Tang, Zetong, et al.
Published: (2025)
by: Tang, Zetong, et al.
Published: (2025)
Japanese Tort-case Dataset for Rationale-supported Legal Judgment Prediction
by: Yamada, Hiroaki, et al.
Published: (2023)
by: Yamada, Hiroaki, et al.
Published: (2023)
One SPACE to Rule Them All: Jointly Mitigating Factuality and Faithfulness Hallucinations in LLMs
by: Wang, Pengbo, et al.
Published: (2025)
by: Wang, Pengbo, et al.
Published: (2025)
OPENXRD: A Comprehensive Benchmark Framework for LLM/MLLM XRD Question Answering
by: Vosoughi, Ali, et al.
Published: (2025)
by: Vosoughi, Ali, et al.
Published: (2025)
Pay Attention to What You Need
by: Gao, Yifei, et al.
Published: (2023)
by: Gao, Yifei, et al.
Published: (2023)
Multi-chain Graph Refinement and Selection for Reliable Reasoning in Large Language Models
by: Yang, Yujiao, et al.
Published: (2025)
by: Yang, Yujiao, et al.
Published: (2025)
Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL
by: Tu, Songjun, et al.
Published: (2025)
by: Tu, Songjun, et al.
Published: (2025)
Language corpora for the Dutch medical domain
by: van Es, B.
Published: (2026)
by: van Es, B.
Published: (2026)
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
by: Yazdanjue, Navid, et al.
Published: (2025)
by: Yazdanjue, Navid, et al.
Published: (2025)
One Law, Many Languages: Benchmarking Multilingual Legal Reasoning for Judicial Support
by: Stern, Ronja, et al.
Published: (2023)
by: Stern, Ronja, et al.
Published: (2023)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
by: Otal, Hakan T., et al.
Published: (2024)
by: Otal, Hakan T., et al.
Published: (2024)
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2024)
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2024)
Lightweight Transformers for Clinical Natural Language Processing
by: Rohanian, Omid, et al.
Published: (2023)
by: Rohanian, Omid, et al.
Published: (2023)
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
by: Fang, Xi, et al.
Published: (2025)
by: Fang, Xi, et al.
Published: (2025)
Similar Items
-
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
by: Kohankhaki, Farnaz, et al.
Published: (2024) -
Fairness Certification for Natural Language Processing and Large Language Models
by: Freiberger, Vincent, et al.
Published: (2024) -
ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
by: Delaval, Axel, et al.
Published: (2025) -
Mitigating Social Biases in Language Models through Unlearning
by: Dige, Omkar, et al.
Published: (2024) -
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
by: Truong, Sang T., et al.
Published: (2024)