Privacy Evaluation Benchmarks for NLP Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Wei, Wang, Yinggui, Chen, Cen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
von: Huang, Wei, et al.
Veröffentlicht: (2026)
von: Huang, Wei, et al.
Veröffentlicht: (2026)
NLP-ADBench: NLP Anomaly Detection Benchmark
von: Li, Yuangang, et al.
Veröffentlicht: (2024)
von: Li, Yuangang, et al.
Veröffentlicht: (2024)
Does Differential Privacy Impact Bias in Pretrained NLP Models?
von: Islam, Md. Khairul, et al.
Veröffentlicht: (2024)
von: Islam, Md. Khairul, et al.
Veröffentlicht: (2024)
What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference
von: Fan, Mingyuan, et al.
Veröffentlicht: (2026)
von: Fan, Mingyuan, et al.
Veröffentlicht: (2026)
EvalxNLP: A Framework for Benchmarking Post-Hoc Explainability Methods on NLP Models
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
von: Dhaini, Mahdi, et al.
Veröffentlicht: (2025)
Indian Legal NLP Benchmarks : A Survey
von: Kalamkar, Prathamesh, et al.
Veröffentlicht: (2021)
von: Kalamkar, Prathamesh, et al.
Veröffentlicht: (2021)
We Need to Talk About Classification Evaluation Metrics in NLP
von: Vickers, Peter, et al.
Veröffentlicht: (2024)
von: Vickers, Peter, et al.
Veröffentlicht: (2024)
Addressing Both Statistical and Causal Gender Fairness in NLP Models
von: Chen, Hannah, et al.
Veröffentlicht: (2024)
von: Chen, Hannah, et al.
Veröffentlicht: (2024)
FedMCP: Parameter-Efficient Federated Learning with Model-Contrastive Personalization
von: Zhao, Qianyi, et al.
Veröffentlicht: (2024)
von: Zhao, Qianyi, et al.
Veröffentlicht: (2024)
CEB: Compositional Evaluation Benchmark for Fairness in Large Language Models
von: Wang, Song, et al.
Veröffentlicht: (2024)
von: Wang, Song, et al.
Veröffentlicht: (2024)
Transferable Adversarial Examples with Bayes Approach
von: Fan, Mingyuan, et al.
Veröffentlicht: (2022)
von: Fan, Mingyuan, et al.
Veröffentlicht: (2022)
AECBench: A Hierarchical Benchmark for Knowledge Evaluation of Large Language Models in the AEC Field
von: Liang, Chen, et al.
Veröffentlicht: (2025)
von: Liang, Chen, et al.
Veröffentlicht: (2025)
Assessing the Capabilities and Limitations of FinGPT Model in Financial NLP Applications
von: Djagba, Prudence, et al.
Veröffentlicht: (2025)
von: Djagba, Prudence, et al.
Veröffentlicht: (2025)
From Word Sequences to Behavioral Sequences: Adapting Modeling and Evaluation Paradigms for Longitudinal NLP
von: Ganesan, Adithya V, et al.
Veröffentlicht: (2026)
von: Ganesan, Adithya V, et al.
Veröffentlicht: (2026)
Privacy-Preserving End-to-End Spoken Language Understanding
von: Wang, Yinggui, et al.
Veröffentlicht: (2024)
von: Wang, Yinggui, et al.
Veröffentlicht: (2024)
PrivacyMind: Large Language Models Can Be Contextual Privacy Protection Learners
von: Xiao, Yijia, et al.
Veröffentlicht: (2023)
von: Xiao, Yijia, et al.
Veröffentlicht: (2023)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
CLIBE: Detecting Dynamic Backdoors in Transformer-based NLP Models
von: Zeng, Rui, et al.
Veröffentlicht: (2024)
von: Zeng, Rui, et al.
Veröffentlicht: (2024)
The NLP Task Effectiveness of Long-Range Transformers
von: Qin, Guanghui, et al.
Veröffentlicht: (2022)
von: Qin, Guanghui, et al.
Veröffentlicht: (2022)
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks
von: Hu, Jiacheng, et al.
Veröffentlicht: (2024)
von: Hu, Jiacheng, et al.
Veröffentlicht: (2024)
Benchmarking Generation and Evaluation Capabilities of Large Language Models for Instruction Controllable Summarization
von: Liu, Yixin, et al.
Veröffentlicht: (2023)
von: Liu, Yixin, et al.
Veröffentlicht: (2023)
Classification of kinetic-related injury in hospital triage data using NLP
von: Shyam, Midhun, et al.
Veröffentlicht: (2025)
von: Shyam, Midhun, et al.
Veröffentlicht: (2025)
Performance of diverse evaluation metrics in NLP-based assessment and text generation of consumer complaints
von: Gao, Peiheng, et al.
Veröffentlicht: (2025)
von: Gao, Peiheng, et al.
Veröffentlicht: (2025)
Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators
von: Zhou, Yilun, et al.
Veröffentlicht: (2025)
von: Zhou, Yilun, et al.
Veröffentlicht: (2025)
ConvNLP: Image-based AI Text Detection
von: Jambunathan, Suriya Prakash, et al.
Veröffentlicht: (2024)
von: Jambunathan, Suriya Prakash, et al.
Veröffentlicht: (2024)
On Importance of Pruning and Distillation for Efficient Low Resource NLP
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2024)
von: Mirashi, Aishwarya, et al.
Veröffentlicht: (2024)
One Model is All You Need: ByT5-Sanskrit, a Unified Model for Sanskrit NLP Tasks
von: Nehrdich, Sebastian, et al.
Veröffentlicht: (2024)
von: Nehrdich, Sebastian, et al.
Veröffentlicht: (2024)
Pretraining Language Models with Subword Regularization: An Empirical Study of BPE Dropout in Low-Resource NLP
von: Visser, Ruan, et al.
Veröffentlicht: (2026)
von: Visser, Ruan, et al.
Veröffentlicht: (2026)
ArenaBencher: Automatic Benchmark Evolution via Multi-Model Competitive Evaluation
von: Liu, Qin, et al.
Veröffentlicht: (2025)
von: Liu, Qin, et al.
Veröffentlicht: (2025)
Can I understand what I create? Self-Knowledge Evaluation of Large Language Models
von: Tan, Zhiquan, et al.
Veröffentlicht: (2024)
von: Tan, Zhiquan, et al.
Veröffentlicht: (2024)
PBa-LLM: Privacy- and Bias-aware NLP using Named-Entity Recognition (NER)
von: Mancera, Gonzalo, et al.
Veröffentlicht: (2025)
von: Mancera, Gonzalo, et al.
Veröffentlicht: (2025)
Inference Attacks Against Face Recognition Model without Classification Layers
von: Huang, Yuanqing, et al.
Veröffentlicht: (2024)
von: Huang, Yuanqing, et al.
Veröffentlicht: (2024)
Towards Efficient Active Learning in NLP via Pretrained Representations
von: Vysogorets, Artem, et al.
Veröffentlicht: (2024)
von: Vysogorets, Artem, et al.
Veröffentlicht: (2024)
TiME: Tiny Monolingual Encoders for Efficient NLP Pipelines
von: Schulmeister, David, et al.
Veröffentlicht: (2025)
von: Schulmeister, David, et al.
Veröffentlicht: (2025)
Boosting classification reliability of NLP transformer models in the long run
von: Kmetty, Zoltán, et al.
Veröffentlicht: (2023)
von: Kmetty, Zoltán, et al.
Veröffentlicht: (2023)
A Survey of Early Exit Deep Neural Networks in NLP
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
TARAZ: Persian Short-Answer Question Benchmark for Cultural Evaluation of Language Models
von: Iranmanesh, Reihaneh, et al.
Veröffentlicht: (2026)
von: Iranmanesh, Reihaneh, et al.
Veröffentlicht: (2026)
SIDU-TXT: An XAI Algorithm for NLP with a Holistic Assessment Approach
von: Jahromi, Mohammad N. S., et al.
Veröffentlicht: (2024)
von: Jahromi, Mohammad N. S., et al.
Veröffentlicht: (2024)
Detecting AI Generated Text Based on NLP and Machine Learning Approaches
von: Prova, Nuzhat
Veröffentlicht: (2024)
von: Prova, Nuzhat
Veröffentlicht: (2024)
Ähnliche Einträge
-
DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression
von: Huang, Wei, et al.
Veröffentlicht: (2025) -
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
von: Huang, Wei, et al.
Veröffentlicht: (2026) -
NLP-ADBench: NLP Anomaly Detection Benchmark
von: Li, Yuangang, et al.
Veröffentlicht: (2024) -
Does Differential Privacy Impact Bias in Pretrained NLP Models?
von: Islam, Md. Khairul, et al.
Veröffentlicht: (2024) -
What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference
von: Fan, Mingyuan, et al.
Veröffentlicht: (2026)