A Data-driven Investigation of Euphemistic Language: Comparing the usage of "slave" and "servant" in 19th century US newspapers
Fuente:
arXiv
Saved in:
| Main Authors: | Park, Jaihyun, Cordell, Ryan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Quantitative Discourse Analysis of Asian Workers in the US Historical Newspapers
by: Park, Jaihyun, et al.
Published: (2024)
by: Park, Jaihyun, et al.
Published: (2024)
Reading the unreadable: Creating a dataset of 19th century English newspapers using image-to-text language models
by: Bourne, Jonathan
Published: (2025)
by: Bourne, Jonathan
Published: (2025)
MEDs for PETs: Multilingual Euphemism Disambiguation for Potentially Euphemistic Terms
by: Lee, Patrick, et al.
Published: (2024)
by: Lee, Patrick, et al.
Published: (2024)
CATs are Fuzzy PETs: A Corpus and Analysis of Potentially Euphemistic Terms
by: Gavidia, Martha, et al.
Published: (2022)
by: Gavidia, Martha, et al.
Published: (2022)
Locating the Leading Edge of Cultural Change
by: Griebel, Sarah, et al.
Published: (2024)
by: Griebel, Sarah, et al.
Published: (2024)
Large language models for newspaper sentiment analysis during COVID-19: The Guardian
by: Chandra, Rohitash, et al.
Published: (2024)
by: Chandra, Rohitash, et al.
Published: (2024)
AI use in American newspapers is widespread, uneven, and rarely disclosed
by: Russell, Jenna, et al.
Published: (2025)
by: Russell, Jenna, et al.
Published: (2025)
How do media talk about the Covid-19 pandemic? Metaphorical thematic clustering in Italian online newspapers
by: Busso, Lucia, et al.
Published: (2022)
by: Busso, Lucia, et al.
Published: (2022)
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts
by: Gokceoglu, Gokcen, et al.
Published: (2024)
by: Gokceoglu, Gokcen, et al.
Published: (2024)
No Language Data Left Behind: A Comparative Study of CJK Language Datasets in the Hugging Face Ecosystem
by: Choi, Dasol, et al.
Published: (2025)
by: Choi, Dasol, et al.
Published: (2025)
Investigating Language Preference of Multilingual RAG Systems
by: Park, Jeonghyun, et al.
Published: (2025)
by: Park, Jeonghyun, et al.
Published: (2025)
LP Data Pipeline: Lightweight, Purpose-driven Data Pipeline for Large Language Models
by: Kim, Yungi, et al.
Published: (2024)
by: Kim, Yungi, et al.
Published: (2024)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
by: Constantinescu, Ionut, et al.
Published: (2024)
by: Constantinescu, Ionut, et al.
Published: (2024)
Match, Compare, or Select? An Investigation of Large Language Models for Entity Matching
by: Wang, Tianshu, et al.
Published: (2024)
by: Wang, Tianshu, et al.
Published: (2024)
A Comparative Study of Translation Bias and Accuracy in Multilingual Large Language Models for Cross-Language Claim Verification
by: Singhal, Aryan, et al.
Published: (2024)
by: Singhal, Aryan, et al.
Published: (2024)
Reflecting the Male Gaze: Quantifying Female Objectification in 19th and 20th Century Novels
by: Luo, Kexin, et al.
Published: (2024)
by: Luo, Kexin, et al.
Published: (2024)
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs
by: Dutta, Sujan, et al.
Published: (2024)
by: Dutta, Sujan, et al.
Published: (2024)
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
by: Lee, Taewhoo, et al.
Published: (2025)
by: Lee, Taewhoo, et al.
Published: (2025)
A Hybrid Theory and Data-driven Approach to Persuasion Detection with Large Language Models
by: Hoang, Gia Bao, et al.
Published: (2025)
by: Hoang, Gia Bao, et al.
Published: (2025)
A Cookbook for Community-driven Data Collection of Impaired Speech in LowResource Languages
by: Salihs, Sumaya Ahmed, et al.
Published: (2025)
by: Salihs, Sumaya Ahmed, et al.
Published: (2025)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
by: Li, Haodong, et al.
Published: (2024)
by: Li, Haodong, et al.
Published: (2024)
Investigating the Impact of Semi-Supervised Methods with Data Augmentation on Offensive Language Detection in Romanian Language
by: Nicola, Elena-Beatrice, et al.
Published: (2024)
by: Nicola, Elena-Beatrice, et al.
Published: (2024)
SpokenUS: A Spoken User Simulator for Task-Oriented Dialogue
by: Lee, Jonggeun, et al.
Published: (2026)
by: Lee, Jonggeun, et al.
Published: (2026)
ASC analyzer: A Python package for measuring argument structure construction usage in English texts
by: Sung, Hakyung, et al.
Published: (2025)
by: Sung, Hakyung, et al.
Published: (2025)
A Comparative Investigation of Compositional Syntax and Semantics in DALL-E 2
by: Murphy, Elliot, et al.
Published: (2024)
by: Murphy, Elliot, et al.
Published: (2024)
Investigating the Effect of Parallel Data in the Cross-Lingual Transfer for Vision-Language Encoders
by: Manea, Andrei-Alexandru, et al.
Published: (2025)
by: Manea, Andrei-Alexandru, et al.
Published: (2025)
BERTtime Stories: Investigating the Role of Synthetic Story Data in Language Pre-training
by: Theodoropoulos, Nikitas, et al.
Published: (2024)
by: Theodoropoulos, Nikitas, et al.
Published: (2024)
The Mediomatix Corpus: Parallel Data for Romansh Language Varieties via Comparable Schoolbooks
by: Hopton, Zachary, et al.
Published: (2025)
by: Hopton, Zachary, et al.
Published: (2025)
Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
by: Tu, Shangqing, et al.
Published: (2024)
by: Tu, Shangqing, et al.
Published: (2024)
Investigating the Impact of Data Selection Strategies on Language Model Performance
by: Gu, Jiayao, et al.
Published: (2025)
by: Gu, Jiayao, et al.
Published: (2025)
Investigating Data Contamination in Modern Benchmarks for Large Language Models
by: Deng, Chunyuan, et al.
Published: (2023)
by: Deng, Chunyuan, et al.
Published: (2023)
Framing in the Presence of Supporting Data: A Case Study in U.S. Economic News
by: Leto, Alexandria, et al.
Published: (2024)
by: Leto, Alexandria, et al.
Published: (2024)
Data-driven Coreference-based Ontology Building
by: Ashury-Tahan, Shir, et al.
Published: (2024)
by: Ashury-Tahan, Shir, et al.
Published: (2024)
Learning to Reduce: Optimal Representations of Structured Data in Prompting Large Language Models
by: Lee, Younghun, et al.
Published: (2024)
by: Lee, Younghun, et al.
Published: (2024)
Learning to Reduce: Towards Improving Performance of Large Language Models on Structured Data
by: Lee, Younghun, et al.
Published: (2024)
by: Lee, Younghun, et al.
Published: (2024)
Investigating Consistency in Query-Based Meeting Summarization: A Comparative Study of Different Embedding Methods
by: Jia-Chen, Chen, et al.
Published: (2024)
by: Jia-Chen, Chen, et al.
Published: (2024)
Blackbird Language Matrices: A Framework to Investigate the Linguistic Competence of Language Models
by: Merlo, Paola, et al.
Published: (2026)
by: Merlo, Paola, et al.
Published: (2026)
9th Workshop on Sign Language Translation and Avatar Technologies (SLTAT 2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
Investigating Data Contamination for Pre-training Language Models
by: Jiang, Minhao, et al.
Published: (2024)
by: Jiang, Minhao, et al.
Published: (2024)
Comparative Analysis of Extrinsic Factors for NER in French
by: Yang, Grace, et al.
Published: (2024)
by: Yang, Grace, et al.
Published: (2024)
Similar Items
-
A Quantitative Discourse Analysis of Asian Workers in the US Historical Newspapers
by: Park, Jaihyun, et al.
Published: (2024) -
Reading the unreadable: Creating a dataset of 19th century English newspapers using image-to-text language models
by: Bourne, Jonathan
Published: (2025) -
MEDs for PETs: Multilingual Euphemism Disambiguation for Potentially Euphemistic Terms
by: Lee, Patrick, et al.
Published: (2024) -
CATs are Fuzzy PETs: A Corpus and Analysis of Potentially Euphemistic Terms
by: Gavidia, Martha, et al.
Published: (2022) -
Locating the Leading Edge of Cultural Change
by: Griebel, Sarah, et al.
Published: (2024)