MultiGraSCCo: A Multilingual Anonymization Benchmark with Annotations of Personal Identifiers
Fuente:
arXiv
Salvato in:
| Autori principali: | Baroud, Ibrahim, Otto, Christoph, Czehmann, Vera, Hovhannisyan, Christine, Raithel, Lisa, Möller, Sebastian, Roller, Roland |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond De-Identification: A Structured Approach for Defining and Detecting Indirect Identifiers in Medical Texts
di: Baroud, Ibrahim, et al.
Pubblicazione: (2025)
di: Baroud, Ibrahim, et al.
Pubblicazione: (2025)
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
di: Frei, Johann, et al.
Pubblicazione: (2025)
di: Frei, Johann, et al.
Pubblicazione: (2025)
A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages
di: Raithel, Lisa, et al.
Pubblicazione: (2024)
di: Raithel, Lisa, et al.
Pubblicazione: (2024)
Evaluation of a Sign Language Avatar on Comprehensibility, User Experience \& Acceptability
di: Wasserroth, Fenya, et al.
Pubblicazione: (2025)
di: Wasserroth, Fenya, et al.
Pubblicazione: (2025)
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
di: MacPhail, Dorothea, et al.
Pubblicazione: (2024)
di: MacPhail, Dorothea, et al.
Pubblicazione: (2024)
DFKI-NLP at SemEval-2024 Task 2: Towards Robust LLMs Using Data Perturbations and MinMax Training
di: Verma, Bhuvanesh, et al.
Pubblicazione: (2024)
di: Verma, Bhuvanesh, et al.
Pubblicazione: (2024)
Hybrid Annotation for Propaganda Detection: Integrating LLM Pre-Annotations with Human Intelligence
di: Sahitaj, Ariana, et al.
Pubblicazione: (2025)
di: Sahitaj, Ariana, et al.
Pubblicazione: (2025)
A Benchmark for Multi-speaker Anonymization
di: Miao, Xiaoxiao, et al.
Pubblicazione: (2024)
di: Miao, Xiaoxiao, et al.
Pubblicazione: (2024)
Probing the Feasibility of Multilingual Speaker Anonymization
di: Meyer, Sarina, et al.
Pubblicazione: (2024)
di: Meyer, Sarina, et al.
Pubblicazione: (2024)
The TUB Sign Language Corpus Collection
di: Avramidis, Eleftherios, et al.
Pubblicazione: (2025)
di: Avramidis, Eleftherios, et al.
Pubblicazione: (2025)
Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects
di: Peng, Ji-Lun, et al.
Pubblicazione: (2026)
di: Peng, Ji-Lun, et al.
Pubblicazione: (2026)
PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection
di: Jha, Pritesh
Pubblicazione: (2026)
di: Jha, Pritesh
Pubblicazione: (2026)
GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models
di: Yuan, Zike, et al.
Pubblicazione: (2024)
di: Yuan, Zike, et al.
Pubblicazione: (2024)
Multilingual Datasets for Custom Input Extraction and Explanation Requests Parsing in Conversational XAI Systems
di: Wang, Qianli, et al.
Pubblicazione: (2025)
di: Wang, Qianli, et al.
Pubblicazione: (2025)
Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following
di: He, Yun, et al.
Pubblicazione: (2024)
di: He, Yun, et al.
Pubblicazione: (2024)
xMEN: A Modular Toolkit for Cross-Lingual Medical Entity Normalization
di: Borchert, Florian, et al.
Pubblicazione: (2023)
di: Borchert, Florian, et al.
Pubblicazione: (2023)
Integrating Text and Time-Series into (Large) Language Models to Predict Medical Outcomes
di: Larbi, Iyadh Ben Cheikh, et al.
Pubblicazione: (2025)
di: Larbi, Iyadh Ben Cheikh, et al.
Pubblicazione: (2025)
GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction
di: Zaratiana, Urchade, et al.
Pubblicazione: (2026)
di: Zaratiana, Urchade, et al.
Pubblicazione: (2026)
Do Multilingual LLMs Think In English?
di: Schut, Lisa, et al.
Pubblicazione: (2025)
di: Schut, Lisa, et al.
Pubblicazione: (2025)
The Conundrum of Trustworthy Research on Attacking Personally Identifiable Information Removal Techniques
di: Ochs, Sebastian, et al.
Pubblicazione: (2026)
di: Ochs, Sebastian, et al.
Pubblicazione: (2026)
Dependency Annotation of Ottoman Turkish with Multilingual BERT
di: Özateş, Şaziye Betül, et al.
Pubblicazione: (2024)
di: Özateş, Şaziye Betül, et al.
Pubblicazione: (2024)
POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization
di: Naseem, Usman, et al.
Pubblicazione: (2025)
di: Naseem, Usman, et al.
Pubblicazione: (2025)
Fine-tuning with Hierarchical Prompting for Robust Propaganda Classification Across Annotation Schemas
di: Stähelin, Lukas, et al.
Pubblicazione: (2026)
di: Stähelin, Lukas, et al.
Pubblicazione: (2026)
Bias in LLMs as Annotators: The Effect of Party Cues on Labelling Decision by Large Language Models
di: Vera, Sebastian Vallejo, et al.
Pubblicazione: (2024)
di: Vera, Sebastian Vallejo, et al.
Pubblicazione: (2024)
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
di: Hyben, Martin, et al.
Pubblicazione: (2023)
di: Hyben, Martin, et al.
Pubblicazione: (2023)
SceneGraMMi: Scene Graph-boosted Hybrid-fusion for Multi-Modal Misinformation Veracity Prediction
di: Joshi, Swarang, et al.
Pubblicazione: (2024)
di: Joshi, Swarang, et al.
Pubblicazione: (2024)
GraSP: Graph-Structured Skill Compositions for LLM Agents
di: Xia, Tianle, et al.
Pubblicazione: (2026)
di: Xia, Tianle, et al.
Pubblicazione: (2026)
Towards Personalized Evaluation of Large Language Models with An Anonymous Crowd-Sourcing Platform
di: Cheng, Mingyue, et al.
Pubblicazione: (2024)
di: Cheng, Mingyue, et al.
Pubblicazione: (2024)
MedErrBench: A Fine-Grained Multilingual Benchmark for Medical Error Detection and Correction with Clinical Expert Annotations
di: Ma, Congbo, et al.
Pubblicazione: (2026)
di: Ma, Congbo, et al.
Pubblicazione: (2026)
Parallel Universes, Parallel Languages: A Comprehensive Study on LLM-based Multilingual Counterfactual Example Generation
di: Wang, Qianli, et al.
Pubblicazione: (2026)
di: Wang, Qianli, et al.
Pubblicazione: (2026)
GPTs Are Multilingual Annotators for Sequence Generation Tasks
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Contamination Report for Multilingual Benchmarks
di: Ahuja, Sanchit, et al.
Pubblicazione: (2024)
di: Ahuja, Sanchit, et al.
Pubblicazione: (2024)
Memory Is All You Need: Testing How Model Memory Affects LLM Performance in Annotation Tasks
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
di: Timoneda, Joan C., et al.
Pubblicazione: (2025)
MultiZebraLogic: A Multilingual Logical Reasoning Benchmark
di: Bruun, Sofie Helene, et al.
Pubblicazione: (2025)
di: Bruun, Sofie Helene, et al.
Pubblicazione: (2025)
Personal Microcomputers in the Library Environment.
di: Raithel, Frederick J.
Pubblicazione: (1980)
di: Raithel, Frederick J.
Pubblicazione: (1980)
Will Annotators Disagree? Identifying Subjectivity in Value-Laden Arguments
di: Homayounirad, Amir, et al.
Pubblicazione: (2025)
di: Homayounirad, Amir, et al.
Pubblicazione: (2025)
Bias in, Bias out: Annotation Bias in Multilingual Large Language Models
di: Cui, Xia, et al.
Pubblicazione: (2025)
di: Cui, Xia, et al.
Pubblicazione: (2025)
On Crowdsourcing Task Design for Discourse Relation Annotation
di: Yung, Frances, et al.
Pubblicazione: (2024)
di: Yung, Frances, et al.
Pubblicazione: (2024)
ConGraT: Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
di: Brannon, William, et al.
Pubblicazione: (2023)
di: Brannon, William, et al.
Pubblicazione: (2023)
An Annotation Scheme and Classifier for Personal Facts in Dialogue
di: Zaitsev, Konstantin
Pubblicazione: (2026)
di: Zaitsev, Konstantin
Pubblicazione: (2026)
Documenti analoghi
-
Beyond De-Identification: A Structured Approach for Defining and Detecting Indirect Identifiers in Medical Texts
di: Baroud, Ibrahim, et al.
Pubblicazione: (2025) -
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
di: Frei, Johann, et al.
Pubblicazione: (2025) -
A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages
di: Raithel, Lisa, et al.
Pubblicazione: (2024) -
Evaluation of a Sign Language Avatar on Comprehensibility, User Experience \& Acceptability
di: Wasserroth, Fenya, et al.
Pubblicazione: (2025) -
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
di: MacPhail, Dorothea, et al.
Pubblicazione: (2024)