NLPGuard: A Framework for Mitigating the Use of Protected Attributes by NLP Classifiers
Fuente:
arXiv
Saved in:
| Main Authors: | Greco, Salvatore, Zhou, Ke, Capra, Licia, Cerquitelli, Tania, Quercia, Daniele |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Frictionless Love: Associations Between AI Companion Roles and Behavioral Addiction
by: Agarwal, Vibhor, et al.
Published: (2026)
by: Agarwal, Vibhor, et al.
Published: (2026)
Concept-based Explainable Artificial Intelligence: A Survey
by: Poeta, Eleonora, et al.
Published: (2023)
by: Poeta, Eleonora, et al.
Published: (2023)
WEIRD ICWSM: How Western, Educated, Industrialized, Rich, and Democratic is Social Computing Research?
by: Septiandri, Ali Akbar, et al.
Published: (2024)
by: Septiandri, Ali Akbar, et al.
Published: (2024)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024)
How to Enable Effective Cooperation Between Humans and NLP Models: A Survey of Principles, Formalizations, and Beyond
by: Huang, Chen, et al.
Published: (2025)
by: Huang, Chen, et al.
Published: (2025)
Human-Centric NLP or AI-Centric Illusion?: A Critical Investigation
by: Spencer, Piyapath T
Published: (2024)
by: Spencer, Piyapath T
Published: (2024)
LLM Based Multi-Agent Generation of Semi-structured Documents from Semantic Templates in the Public Administration Domain
by: Musumeci, Emanuele, et al.
Published: (2024)
by: Musumeci, Emanuele, et al.
Published: (2024)
NLP Privacy Risk Identification in Social Media (NLP-PRISM): A Survey
by: Goswami, Dhiman, et al.
Published: (2026)
by: Goswami, Dhiman, et al.
Published: (2026)
Dehumanizing Machines: Mitigating Anthropomorphic Behaviors in Text Generation Systems
by: Cheng, Myra, et al.
Published: (2025)
by: Cheng, Myra, et al.
Published: (2025)
Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance
by: Srinivasan, Tejas, et al.
Published: (2025)
by: Srinivasan, Tejas, et al.
Published: (2025)
Benchmarking LLM Tool-Use in the Wild
by: Yu, Peijie, et al.
Published: (2026)
by: Yu, Peijie, et al.
Published: (2026)
CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigation
by: Huq, Faria, et al.
Published: (2025)
by: Huq, Faria, et al.
Published: (2025)
Creating General User Models from Computer Use
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
Large Language Model Use Impact Locus of Control
by: Fu, Jenny Xiyu, et al.
Published: (2025)
by: Fu, Jenny Xiyu, et al.
Published: (2025)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
by: Daynauth, Roland, et al.
Published: (2024)
by: Daynauth, Roland, et al.
Published: (2024)
Mitigating Harmful Erraticism in LLMs Through Dialectical Behavior Therapy Based De-Escalation Strategies
by: Rangarajan, Pooja, et al.
Published: (2025)
by: Rangarajan, Pooja, et al.
Published: (2025)
An Evaluation of Estimative Uncertainty in Large Language Models
by: Tang, Zhisheng, et al.
Published: (2024)
by: Tang, Zhisheng, et al.
Published: (2024)
ChatSUMO: Large Language Model for Automating Traffic Scenario Generation in Simulation of Urban MObility
by: Li, Shuyang, et al.
Published: (2024)
by: Li, Shuyang, et al.
Published: (2024)
Exploring Human Perceptions of AI Responses: Insights from a Mixed-Methods Study on Risk Mitigation in Generative Models
by: Candello, Heloisa, et al.
Published: (2025)
by: Candello, Heloisa, et al.
Published: (2025)
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation
by: Saka, Abdullahi, et al.
Published: (2023)
by: Saka, Abdullahi, et al.
Published: (2023)
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI
by: Krause, Stefanie, et al.
Published: (2024)
by: Krause, Stefanie, et al.
Published: (2024)
A Scalable Framework for Evaluating Health Language Models
by: Mallinar, Neil, et al.
Published: (2025)
by: Mallinar, Neil, et al.
Published: (2025)
PatientHub: A Unified Framework for Patient Simulation
by: Sabour, Sahand, et al.
Published: (2026)
by: Sabour, Sahand, et al.
Published: (2026)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
by: Cohen, Myke C., et al.
Published: (2026)
by: Cohen, Myke C., et al.
Published: (2026)
Understanding How Paper Writers Use AI-Generated Captions in Figure Caption Writing
by: Yin, Ho, et al.
Published: (2025)
by: Yin, Ho, et al.
Published: (2025)
Modelling and Classifying the Components of a Literature Review
by: Bolaños, Francisco, et al.
Published: (2025)
by: Bolaños, Francisco, et al.
Published: (2025)
Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks
by: Greco, Candida M., et al.
Published: (2026)
by: Greco, Candida M., et al.
Published: (2026)
Epistemic Alignment: A Mediating Framework for User-LLM Knowledge Delivery
by: Clark, Nicholas, et al.
Published: (2025)
by: Clark, Nicholas, et al.
Published: (2025)
Collaborative Gym: A Framework for Enabling and Evaluating Human-Agent Collaboration
by: Shao, Yijia, et al.
Published: (2024)
by: Shao, Yijia, et al.
Published: (2024)
ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents
by: Liu, Tianjian, et al.
Published: (2025)
by: Liu, Tianjian, et al.
Published: (2025)
"What Are You Really Trying to Do?": Co-Creating Life Goals from Everyday Computer Use
by: Sapkota, Shardul, et al.
Published: (2026)
by: Sapkota, Shardul, et al.
Published: (2026)
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
by: Soni, Nikita, et al.
Published: (2025)
by: Soni, Nikita, et al.
Published: (2025)
ViSP: A PPO-Driven Framework for Sarcasm Generation with Contrastive Learning
by: Wang, Changli, et al.
Published: (2025)
by: Wang, Changli, et al.
Published: (2025)
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
by: Shen, Hua, et al.
Published: (2024)
by: Shen, Hua, et al.
Published: (2024)
SAC: A Framework for Measuring and Inducing Personality Traits in LLMs with Dynamic Intensity Control
by: Chittem, Adithya, et al.
Published: (2025)
by: Chittem, Adithya, et al.
Published: (2025)
A Generalized LLM-Augmented BIM Framework: Application to a Speech-to-BIM system
by: Lee, Ghang, et al.
Published: (2024)
by: Lee, Ghang, et al.
Published: (2024)
On the Psychology of GPT-4: Moderately anxious, slightly masculine, honest, and humble
by: Barua, Adrita, et al.
Published: (2024)
by: Barua, Adrita, et al.
Published: (2024)
A Directed Graph Model and Experimental Framework for Design and Study of Time-Dependent Text Visualisation
by: Fan, Songhai, et al.
Published: (2026)
by: Fan, Songhai, et al.
Published: (2026)
Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain
by: Chiang, Shang-Hsuan, et al.
Published: (2026)
by: Chiang, Shang-Hsuan, et al.
Published: (2026)
VeriLA: A Human-Centered Evaluation Framework for Interpretable Verification of LLM Agent Failures
by: Sung, Yoo Yeon, et al.
Published: (2025)
by: Sung, Yoo Yeon, et al.
Published: (2025)
Similar Items
-
Frictionless Love: Associations Between AI Companion Roles and Behavioral Addiction
by: Agarwal, Vibhor, et al.
Published: (2026) -
Concept-based Explainable Artificial Intelligence: A Survey
by: Poeta, Eleonora, et al.
Published: (2023) -
WEIRD ICWSM: How Western, Educated, Industrialized, Rich, and Democratic is Social Computing Research?
by: Septiandri, Ali Akbar, et al.
Published: (2024) -
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
by: Chaleshtori, Fateme Hashemi, et al.
Published: (2024) -
How to Enable Effective Cooperation Between Humans and NLP Models: A Survey of Principles, Formalizations, and Beyond
by: Huang, Chen, et al.
Published: (2025)