Behind Closed Words: Creating and Investigating the forePLay Annotated Dataset for Polish Erotic Discourse
Fuente:
arXiv
Salvato in:
| Autori principali: | Kołos, Anna, Lorenc, Katarzyna, Wiśnios, Emilia, Karlińska, Agnieszka |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BAN-PL: a Novel Polish Dataset of Banned Harmful and Offensive Content from Wykop.pl web service
di: Kołos, Anna, et al.
Pubblicazione: (2023)
di: Kołos, Anna, et al.
Pubblicazione: (2023)
Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2025)
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2025)
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
di: Góral, Gracjan, et al.
Pubblicazione: (2024)
di: Góral, Gracjan, et al.
Pubblicazione: (2024)
From Words to Wisdom: Discourse Annotation and Baseline Models for Student Dialogue Understanding
di: Mim, Farjana Sultana, et al.
Pubblicazione: (2025)
di: Mim, Farjana Sultana, et al.
Pubblicazione: (2025)
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
di: Statkiewicz, Grzegorz, et al.
Pubblicazione: (2026)
di: Statkiewicz, Grzegorz, et al.
Pubblicazione: (2026)
Automatic Alignment of Discourse Relations of Different Discourse Annotation Frameworks
di: Fu, Yingxue
Pubblicazione: (2024)
di: Fu, Yingxue
Pubblicazione: (2024)
OpenGVL -- Benchmarking Visual Temporal Progress for Data Curation
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
PLLuM: A Family of Polish Large Language Models
di: Kocoń, Jan, et al.
Pubblicazione: (2025)
di: Kocoń, Jan, et al.
Pubblicazione: (2025)
POLygraph: Polish Fake News Dataset
di: Dzienisiewicz, Daniel, et al.
Pubblicazione: (2024)
di: Dzienisiewicz, Daniel, et al.
Pubblicazione: (2024)
Polish-ASTE: Aspect-Sentiment Triplet Extraction Datasets for Polish
di: Lango, Marta, et al.
Pubblicazione: (2025)
di: Lango, Marta, et al.
Pubblicazione: (2025)
AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse
di: Sharqawi, Esra'a, et al.
Pubblicazione: (2026)
di: Sharqawi, Esra'a, et al.
Pubblicazione: (2026)
On Crowdsourcing Task Design for Discourse Relation Annotation
di: Yung, Frances, et al.
Pubblicazione: (2024)
di: Yung, Frances, et al.
Pubblicazione: (2024)
PolQA: Polish Question Answering Dataset
di: Rybak, Piotr, et al.
Pubblicazione: (2022)
di: Rybak, Piotr, et al.
Pubblicazione: (2022)
Prompting Implicit Discourse Relation Annotation
di: Yung, Frances, et al.
Pubblicazione: (2024)
di: Yung, Frances, et al.
Pubblicazione: (2024)
Big Tech influence over AI research revisited: memetic analysis of attribution of ideas to affiliation
di: Giziński, Stanisław, et al.
Pubblicazione: (2023)
di: Giziński, Stanisław, et al.
Pubblicazione: (2023)
Sandpiper: Orchestrated AI-Annotation for Educational Discourse at Scale
di: Hedley, Daryl, et al.
Pubblicazione: (2026)
di: Hedley, Daryl, et al.
Pubblicazione: (2026)
The Voice Behind the Words: Quantifying Intersectional Bias in SpeechLLMs
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
di: Satish, Shree Harsha Bokkahalli, et al.
Pubblicazione: (2026)
WhatsApp Vaccine Discourse (WhaVax): An Expert-Annotated Dataset and Benchmark for Health Misinformation Detection
di: Santos, Jônatas H. dos, et al.
Pubblicazione: (2026)
di: Santos, Jônatas H. dos, et al.
Pubblicazione: (2026)
Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences
di: Modzelewski, Arkadiusz, et al.
Pubblicazione: (2026)
di: Modzelewski, Arkadiusz, et al.
Pubblicazione: (2026)
ParlaSpeech 3.0: Richly Annotated Spoken Parliamentary Corpora of Croatian, Czech, Polish, and Serbian
di: Ljubešić, Nikola, et al.
Pubblicazione: (2025)
di: Ljubešić, Nikola, et al.
Pubblicazione: (2025)
101 Billion Arabic Words Dataset
di: Aloui, Manel, et al.
Pubblicazione: (2024)
di: Aloui, Manel, et al.
Pubblicazione: (2024)
KoWit-24: A Richly Annotated Dataset of Wordplay in News Headlines
di: Baranov, Alexander, et al.
Pubblicazione: (2025)
di: Baranov, Alexander, et al.
Pubblicazione: (2025)
nEMO: Dataset of Emotional Speech in Polish
di: Christop, Iwona
Pubblicazione: (2024)
di: Christop, Iwona
Pubblicazione: (2024)
Behind the Screen: Investigating ChatGPT's Dark Personality Traits and Conspiracy Beliefs
di: Weber, Erik, et al.
Pubblicazione: (2024)
di: Weber, Erik, et al.
Pubblicazione: (2024)
Are LLMs Good Annotators for Discourse-level Event Relation Extraction?
di: Wei, Kangda, et al.
Pubblicazione: (2024)
di: Wei, Kangda, et al.
Pubblicazione: (2024)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
di: Christ, Bryan R, et al.
Pubblicazione: (2024)
Refining Word-Based Grammatical Error Annotation for L2 Korean
di: Park, Jungyeul, et al.
Pubblicazione: (2026)
di: Park, Jungyeul, et al.
Pubblicazione: (2026)
Investigating Idiomaticity in Word Representations
di: He, Wei, et al.
Pubblicazione: (2024)
di: He, Wei, et al.
Pubblicazione: (2024)
Investigating Low-Cost LLM Annotation for~Spoken Dialogue Understanding Datasets
di: Druart, Lucas, et al.
Pubblicazione: (2024)
di: Druart, Lucas, et al.
Pubblicazione: (2024)
Understanding and Analyzing Inappropriately Targeting Language in Online Discourse: A Comparative Annotation Study
di: Barbarestani, Baran, et al.
Pubblicazione: (2025)
di: Barbarestani, Baran, et al.
Pubblicazione: (2025)
Large Language Models in Legislative Content Analysis: A Dataset from the Polish Parliament
di: Bryłkowski, Arkadiusz, et al.
Pubblicazione: (2025)
di: Bryłkowski, Arkadiusz, et al.
Pubblicazione: (2025)
Understanding the Dataset Practitioners Behind Large Language Model Development
di: Qian, Crystal, et al.
Pubblicazione: (2024)
di: Qian, Crystal, et al.
Pubblicazione: (2024)
SLURG: Investigating the Feasibility of Generating Synthetic Online Fallacious Discourse
di: Blanco, Cal, et al.
Pubblicazione: (2025)
di: Blanco, Cal, et al.
Pubblicazione: (2025)
LLMs as Data Annotators: How Close Are We to Human Performance
di: Haq, Muhammad Uzair Ul, et al.
Pubblicazione: (2025)
di: Haq, Muhammad Uzair Ul, et al.
Pubblicazione: (2025)
SACRED: A Faithful Annotated Multimedia Multimodal Multilingual Dataset for Classifying Connectedness Types in Online Spirituality
di: Guan, Qinghao, et al.
Pubblicazione: (2026)
di: Guan, Qinghao, et al.
Pubblicazione: (2026)
Peeking Behind Closed Doors: Risks of LLM Evaluation by Private Data Curators
di: Bansal, Hritik, et al.
Pubblicazione: (2025)
di: Bansal, Hritik, et al.
Pubblicazione: (2025)
A digital perspective on the role of a stemma in material-philological transmission studies
di: Kapitan, Katarzyna Anna
Pubblicazione: (2025)
di: Kapitan, Katarzyna Anna
Pubblicazione: (2025)
Investigating Expert-in-the-Loop LLM Discourse Patterns for Ancient Intertextual Analysis
di: Umphrey, Ray, et al.
Pubblicazione: (2024)
di: Umphrey, Ray, et al.
Pubblicazione: (2024)
Diverse Word Choices, Same Reference: Annotating Lexically-Rich Cross-Document Coreference
di: Zhukova, Anastasia, et al.
Pubblicazione: (2026)
di: Zhukova, Anastasia, et al.
Pubblicazione: (2026)
Human-Annotated NER Dataset for the Kyrgyz Language
di: Turatali, Timur, et al.
Pubblicazione: (2025)
di: Turatali, Timur, et al.
Pubblicazione: (2025)
Documenti analoghi
-
BAN-PL: a Novel Polish Dataset of Banned Harmful and Offensive Content from Wykop.pl web service
di: Kołos, Anna, et al.
Pubblicazione: (2023) -
Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2025) -
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
di: Góral, Gracjan, et al.
Pubblicazione: (2024) -
From Words to Wisdom: Discourse Annotation and Baseline Models for Student Dialogue Understanding
di: Mim, Farjana Sultana, et al.
Pubblicazione: (2025) -
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
di: Statkiewicz, Grzegorz, et al.
Pubblicazione: (2026)