Digital Gatekeepers: Google's Role in Curating Hashtags and Subreddits
Fuente:
arXiv
Saved in:
| Main Authors: | Poudel, Amrit, Ding, Yifan, Pfeffer, Jurgen, Weninger, Tim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Social and Political Framing in Search Engine Results
by: Poudel, Amrit, et al.
Published: (2025)
by: Poudel, Amrit, et al.
Published: (2025)
Navigating the Post-API Dilemma | Search Engine Results Pages Present a Biased View of Social Media Data
by: Poudel, Amrit, et al.
Published: (2024)
by: Poudel, Amrit, et al.
Published: (2024)
The Power of Framing: How News Headlines Guide Search Behavior
by: Poudel, Amrit, et al.
Published: (2025)
by: Poudel, Amrit, et al.
Published: (2025)
EntGPT: Entity Linking with Generative Large Language Models
by: Ding, Yifan, et al.
Published: (2024)
by: Ding, Yifan, et al.
Published: (2024)
Citations and Trust in LLM Generated Responses
by: Ding, Yifan, et al.
Published: (2025)
by: Ding, Yifan, et al.
Published: (2025)
ChatEL: Entity Linking with Chatbots
by: Ding, Yifan, et al.
Published: (2024)
by: Ding, Yifan, et al.
Published: (2024)
Span-Oriented Information Extraction -- A Unifying Perspective on Information Extraction
by: Ding, Yifan, et al.
Published: (2024)
by: Ding, Yifan, et al.
Published: (2024)
The Role of Network and Identity in the Diffusion of Hashtags
by: Ananthasubramaniam, Aparna, et al.
Published: (2024)
by: Ananthasubramaniam, Aparna, et al.
Published: (2024)
Digital Gatekeepers: Exploring Large Language Model's Role in Immigration Decisions
by: Mao, Yicheng, et al.
Published: (2025)
by: Mao, Yicheng, et al.
Published: (2025)
Evaluating LLMs as Human Surrogates in Controlled Experiments
by: Hoq, Adnan, et al.
Published: (2026)
by: Hoq, Adnan, et al.
Published: (2026)
HICL: Hashtag-Driven In-Context Learning for Social Media Natural Language Understanding
by: Tan, Hanzhuo, et al.
Published: (2023)
by: Tan, Hanzhuo, et al.
Published: (2023)
From #Dr00gtiktok to #harmreduction: Exploring Substance Use Hashtags on TikTok
by: Bouzoubaa, Layla, et al.
Published: (2025)
by: Bouzoubaa, Layla, et al.
Published: (2025)
Tasks and Roles in Legal AI: Data Curation, Annotation, and Verification
by: Koenecke, Allison, et al.
Published: (2025)
by: Koenecke, Allison, et al.
Published: (2025)
The Role of Data Curation in Image Captioning
by: Li, Wenyan, et al.
Published: (2023)
by: Li, Wenyan, et al.
Published: (2023)
From Pixels to Posts: Retrieval-Augmented Fashion Captioning and Hashtag Generation
by: Gondal, Moazzam Umer, et al.
Published: (2025)
by: Gondal, Moazzam Umer, et al.
Published: (2025)
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI
by: Schirmer, Miriam, et al.
Published: (2024)
by: Schirmer, Miriam, et al.
Published: (2024)
Two-Stage Stance Labeling: User-Hashtag Heuristics with Graph Neural Networks
by: Melton, Joshua, et al.
Published: (2024)
by: Melton, Joshua, et al.
Published: (2024)
On LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A Survey
by: Long, Lin, et al.
Published: (2024)
by: Long, Lin, et al.
Published: (2024)
From Quotes to Concepts: Axial Coding of Political Debates with Ensemble LMs
by: Parfenova, Angelina, et al.
Published: (2026)
by: Parfenova, Angelina, et al.
Published: (2026)
Stateful KV Cache Management for LLMs: Balancing Space, Time, Accuracy, and Positional Fidelity
by: Poudel, Pratik
Published: (2025)
by: Poudel, Pratik
Published: (2025)
One Jump Is All You Need: Short-Cutting Transformers for Early Exit Prediction with One Jump to Fit All Exit Levels
by: Seshadri, Amrit Diggavi
Published: (2025)
by: Seshadri, Amrit Diggavi
Published: (2025)
Seed-Coder: Let the Code Model Curate Data for Itself
by: Seed, ByteDance, et al.
Published: (2025)
by: Seed, ByteDance, et al.
Published: (2025)
Emergent Convergence in Multi-Agent LLM Annotation
by: Parfenova, Angelina, et al.
Published: (2025)
by: Parfenova, Angelina, et al.
Published: (2025)
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
by: Uzor, GodsGift, et al.
Published: (2025)
by: Uzor, GodsGift, et al.
Published: (2025)
Cleaner Pretraining Corpus Curation with Neural Web Scraping
by: Xu, Zhipeng, et al.
Published: (2024)
by: Xu, Zhipeng, et al.
Published: (2024)
Leveraging Machine Learning to Detect Data Curation Activities
by: Lafia, Sara, et al.
Published: (2021)
by: Lafia, Sara, et al.
Published: (2021)
On the Role of Feedback in Test-Time Scaling of Agentic AI Workflows
by: Chakraborty, Souradip, et al.
Published: (2025)
by: Chakraborty, Souradip, et al.
Published: (2025)
Ask, Answer, and Detect: Role-Playing LLMs for Personality Detection with Question-Conditioned Mixture-of-Experts
by: Lyu, Yifan, et al.
Published: (2025)
by: Lyu, Yifan, et al.
Published: (2025)
Efficient Alignment of Large Language Models via Data Sampling
by: Khera, Amrit, et al.
Published: (2024)
by: Khera, Amrit, et al.
Published: (2024)
Toxicity of the Commons: Curating Open-Source Pre-Training Data
by: Arnett, Catherine, et al.
Published: (2024)
by: Arnett, Catherine, et al.
Published: (2024)
Automated Data Curation for Robust Language Model Fine-Tuning
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
Fine-Tuning DialoGPT on Common Diseases in Rural Nepal for Medical Conversations
by: Poudel, Birat, et al.
Published: (2025)
by: Poudel, Birat, et al.
Published: (2025)
NLPineers@ NLU of Devanagari Script Languages 2025: Hate Speech Detection using Ensembling of BERT-based models
by: Guragain, Anmol, et al.
Published: (2024)
by: Guragain, Anmol, et al.
Published: (2024)
Organize the Web: Constructing Domains Enhances Pre-Training Data Curation
by: Wettig, Alexander, et al.
Published: (2025)
by: Wettig, Alexander, et al.
Published: (2025)
CRAB: A Benchmark for Evaluating Curation of Retrieval-Augmented LLMs in Biomedicine
by: Zhong, Hanmeng, et al.
Published: (2025)
by: Zhong, Hanmeng, et al.
Published: (2025)
Text Annotation via Inductive Coding: Comparing Human Experts to LLMs in Qualitative Data Analysis
by: Parfenova, Angelina, et al.
Published: (2025)
by: Parfenova, Angelina, et al.
Published: (2025)
Curation of a Palaeohispanic Dataset for Machine Learning
by: Martínez-Fernández, Gonzalo, et al.
Published: (2026)
by: Martínez-Fernández, Gonzalo, et al.
Published: (2026)
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
by: Nguyen, Tuc, et al.
Published: (2025)
by: Nguyen, Tuc, et al.
Published: (2025)
Machine-Assisted Script Curation
by: Ciosici, Manuel R., et al.
Published: (2021)
by: Ciosici, Manuel R., et al.
Published: (2021)
SPIN: Sparsifying and Integrating Internal Neurons in Large Language Models for Text Classification
by: Jiao, Difan, et al.
Published: (2023)
by: Jiao, Difan, et al.
Published: (2023)
Similar Items
-
Social and Political Framing in Search Engine Results
by: Poudel, Amrit, et al.
Published: (2025) -
Navigating the Post-API Dilemma | Search Engine Results Pages Present a Biased View of Social Media Data
by: Poudel, Amrit, et al.
Published: (2024) -
The Power of Framing: How News Headlines Guide Search Behavior
by: Poudel, Amrit, et al.
Published: (2025) -
EntGPT: Entity Linking with Generative Large Language Models
by: Ding, Yifan, et al.
Published: (2024) -
Citations and Trust in LLM Generated Responses
by: Ding, Yifan, et al.
Published: (2025)