Leveraging Machine Learning to Detect Data Curation Activities
Fuente:
arXiv
Guardado en:
| Autores principales: | Lafia, Sara, Thomer, Andrea, Bleckley, David, Akmon, Dharma, Hemphill, Libby |
|---|---|
| Formato: | Preprint |
| Publicado: |
2021
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating how LLM annotations represent diverse views on contentious topics
por: Brown, Megan A., et al.
Publicado: (2025)
por: Brown, Megan A., et al.
Publicado: (2025)
The Effects of Moral Framing on Online Fundraising Outcomes: Evidence from GoFundMe Campaigns
por: Kim, Ji Eun, et al.
Publicado: (2025)
por: Kim, Ji Eun, et al.
Publicado: (2025)
War and Peace (WarAgent): Large Language Model-based Multi-Agent Simulation of World Wars
por: Hua, Wenyue, et al.
Publicado: (2023)
por: Hua, Wenyue, et al.
Publicado: (2023)
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
por: Najjar, Ayat A., et al.
Publicado: (2025)
por: Najjar, Ayat A., et al.
Publicado: (2025)
Crowdsourced reviews reveal substantial disparities in public perceptions of parking
por: Li, Lingyao, et al.
Publicado: (2024)
por: Li, Lingyao, et al.
Publicado: (2024)
AppealMod: Inducing Friction to Reduce Moderator Workload of Handling User Appeals
por: Atreja, Shubham, et al.
Publicado: (2023)
por: Atreja, Shubham, et al.
Publicado: (2023)
Curating corpora with classifiers: A case study of clean energy sentiment online
por: Arnold, Michael V., et al.
Publicado: (2023)
por: Arnold, Michael V., et al.
Publicado: (2023)
The Spread of Virtual Gifting in Live Streaming: The Case of Twitch
por: Kim, Ji Eun, et al.
Publicado: (2025)
por: Kim, Ji Eun, et al.
Publicado: (2025)
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs
por: Birkenmaier, Lukas, et al.
Publicado: (2024)
por: Birkenmaier, Lukas, et al.
Publicado: (2024)
Leveraging Large Language Models to Measure Gender Representation Bias in Gendered Language Corpora
por: Derner, Erik, et al.
Publicado: (2024)
por: Derner, Erik, et al.
Publicado: (2024)
SoK: Machine Learning for Misinformation Detection
por: Xiao, Madelyne, et al.
Publicado: (2023)
por: Xiao, Madelyne, et al.
Publicado: (2023)
“Unnecessarily cumbersome”: Researchers' Opinions on Restricted Data Access Systems
por: Megan A Brown, et al.
Publicado: (2025)
por: Megan A Brown, et al.
Publicado: (2025)
Machine Learning for Detection and Analysis of Novel LLM Jailbreaks
por: Hawkins, John, et al.
Publicado: (2025)
por: Hawkins, John, et al.
Publicado: (2025)
OpenTuringBench: An Open-Model-based Benchmark and Framework for Machine-Generated Text Detection and Attribution
por: La Cava, Lucio, et al.
Publicado: (2025)
por: La Cava, Lucio, et al.
Publicado: (2025)
Peeking Behind Closed Doors: Risks of LLM Evaluation by Private Data Curators
por: Bansal, Hritik, et al.
Publicado: (2025)
por: Bansal, Hritik, et al.
Publicado: (2025)
Characterizing Online Toxicity During the 2022 Mpox Outbreak: A Computational Analysis of Topical and Network Dynamics
por: Fan, Lizhou, et al.
Publicado: (2024)
por: Fan, Lizhou, et al.
Publicado: (2024)
LM$^2$otifs : An Explainable Framework for Machine-Generated Texts Detection
por: Zheng, Xu, et al.
Publicado: (2025)
por: Zheng, Xu, et al.
Publicado: (2025)
Luminol-AIDetect: Fast Zero-shot Machine-Generated Text Detection based on Perplexity under Text Shuffling
por: La Cava, Lucio, et al.
Publicado: (2026)
por: La Cava, Lucio, et al.
Publicado: (2026)
ChatEd: A Chatbot Leveraging ChatGPT for an Enhanced Learning Experience in Higher Education
por: Wang, Kevin, et al.
Publicado: (2023)
por: Wang, Kevin, et al.
Publicado: (2023)
Leveraging Machine Learning to Identify Gendered Stereotypes and Body Image Concerns on Diet and Fitness Online Forums
por: Chu, Minh Duc, et al.
Publicado: (2024)
por: Chu, Minh Duc, et al.
Publicado: (2024)
Landscape of Generative AI in Global News: Topics, Sentiments, and Spatiotemporal Analysis
por: Xian, Lu, et al.
Publicado: (2024)
por: Xian, Lu, et al.
Publicado: (2024)
Machine Learning Data Practices through a Data Curation Lens: An Evaluation Framework
por: Bhardwaj, Eshta, et al.
Publicado: (2024)
por: Bhardwaj, Eshta, et al.
Publicado: (2024)
LangLingual: A Personalised, Exercise-oriented English Language Learning Tool Leveraging Large Language Models
por: Gupta, Sammriddh, et al.
Publicado: (2025)
por: Gupta, Sammriddh, et al.
Publicado: (2025)
Machines in the Crowd? Measuring the Footprint of Machine-Generated Text on Reddit
por: La Cava, Lucio, et al.
Publicado: (2025)
por: La Cava, Lucio, et al.
Publicado: (2025)
Evaluating Machine Translation Datasets for Low-Web Data Languages: A Gendered Lens
por: Nigatu, Hellina Hailu, et al.
Publicado: (2025)
por: Nigatu, Hellina Hailu, et al.
Publicado: (2025)
Leveraging Prompts in LLMs to Overcome Imbalances in Complex Educational Text Data
por: McClure, Jeanne, et al.
Publicado: (2024)
por: McClure, Jeanne, et al.
Publicado: (2024)
Leveraging Large Language Models for Predictive Analysis of Human Misery
por: Seal, Bishanka, et al.
Publicado: (2025)
por: Seal, Bishanka, et al.
Publicado: (2025)
Prompt Design Matters for Computational Social Science Tasks but in Unpredictable Ways
por: Atreja, Shubham, et al.
Publicado: (2024)
por: Atreja, Shubham, et al.
Publicado: (2024)
"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
por: Li, Lingyao, et al.
Publicado: (2023)
por: Li, Lingyao, et al.
Publicado: (2023)
Is Contrasting All You Need? Contrastive Learning for the Detection and Attribution of AI-generated Text
por: La Cava, Lucio, et al.
Publicado: (2024)
por: La Cava, Lucio, et al.
Publicado: (2024)
The ProLiFIC dataset: Leveraging LLMs to Unveil the Italian Lawmaking Process
por: Contestabile, Matilde, et al.
Publicado: (2025)
por: Contestabile, Matilde, et al.
Publicado: (2025)
Leveraging Large Language Models for Actionable Course Evaluation Student Feedback to Lecturers
por: Zhang, Mike, et al.
Publicado: (2024)
por: Zhang, Mike, et al.
Publicado: (2024)
Large Language Models Require Curated Context for Reliable Political Fact-Checking -- Even with Reasoning and Web Search
por: DeVerna, Matthew R., et al.
Publicado: (2025)
por: DeVerna, Matthew R., et al.
Publicado: (2025)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
por: Haider, Batool, et al.
Publicado: (2025)
por: Haider, Batool, et al.
Publicado: (2025)
AIMSCheck: Leveraging LLMs for AI-Assisted Review of Modern Slavery Statements Across Jurisdictions
por: Bora, Adriana Eufrosina, et al.
Publicado: (2025)
por: Bora, Adriana Eufrosina, et al.
Publicado: (2025)
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
por: Thomas, Danielle R., et al.
Publicado: (2025)
por: Thomas, Danielle R., et al.
Publicado: (2025)
A Big Data-empowered System for Real-time Detection of Regional Discriminatory Comments on Vietnamese Social Media
por: Huynh, An Nghiep, et al.
Publicado: (2024)
por: Huynh, An Nghiep, et al.
Publicado: (2024)
Detecting a Proxy for Potential Comorbid ADHD in People Reporting Anxiety Symptoms from Social Media Data
por: Lee, Claire S., et al.
Publicado: (2024)
por: Lee, Claire S., et al.
Publicado: (2024)
People Make Better Edits: Measuring the Efficacy of LLM-Generated Counterfactually Augmented Data for Harmful Language Detection
por: Sen, Indira, et al.
Publicado: (2023)
por: Sen, Indira, et al.
Publicado: (2023)
When a Nation Speaks: Machine Learning and NLP in People's Sentiment Analysis During Bangladesh's 2024 Mass Uprising
por: Alim, Md. Samiul, et al.
Publicado: (2025)
por: Alim, Md. Samiul, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating how LLM annotations represent diverse views on contentious topics
por: Brown, Megan A., et al.
Publicado: (2025) -
The Effects of Moral Framing on Online Fundraising Outcomes: Evidence from GoFundMe Campaigns
por: Kim, Ji Eun, et al.
Publicado: (2025) -
War and Peace (WarAgent): Large Language Model-based Multi-Agent Simulation of World Wars
por: Hua, Wenyue, et al.
Publicado: (2023) -
Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity
por: Najjar, Ayat A., et al.
Publicado: (2025) -
Crowdsourced reviews reveal substantial disparities in public perceptions of parking
por: Li, Lingyao, et al.
Publicado: (2024)