YT-30M: A multi-lingual multi-category dataset of YouTube comments
Fuente:
arXiv
Saved in:
| Main Author: | Dutta, Hridoy Sankar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
YTCommentVerse: A Multi-Category Multi-Lingual YouTube Comment Corpus
by: Dutta, Hridoy Sankar, et al.
Published: (2025)
by: Dutta, Hridoy Sankar, et al.
Published: (2025)
A Social Data-Driven System for Identifying Estate-related Events and Topics
by: Mu, Wenchuan, et al.
Published: (2025)
by: Mu, Wenchuan, et al.
Published: (2025)
A "Perspectival" Mirror of the Elephant: Investigating Language Bias on Google, ChatGPT, YouTube, and Wikipedia
by: Luo, Queenie, et al.
Published: (2023)
by: Luo, Queenie, et al.
Published: (2023)
Utilizing Large Language Models to Identify Reddit Users Considering Vaping Cessation for Digital Interventions
by: Vuruma, Sai Krishna Revanth, et al.
Published: (2024)
by: Vuruma, Sai Krishna Revanth, et al.
Published: (2024)
Explainable Identification of Hate Speech towards Islam using Graph Neural Networks
by: Wasi, Azmine Toushik
Published: (2023)
by: Wasi, Azmine Toushik
Published: (2023)
Generative Engine Optimization: How to Dominate AI Search
by: Chen, Mahe, et al.
Published: (2025)
by: Chen, Mahe, et al.
Published: (2025)
Entity Insertion in Multilingual Linked Corpora: The Case of Wikipedia
by: Feith, Tomás, et al.
Published: (2024)
by: Feith, Tomás, et al.
Published: (2024)
Who Shapes Brazil's Vaccine Debate? Semi-Supervised Modeling of Stance and Polarization in YouTube's Media Ecosystem
by: de Oliveira, Geovana S., et al.
Published: (2026)
by: de Oliveira, Geovana S., et al.
Published: (2026)
Cross-Platform Digital Discourse Analysis of the Israel-Hamas Conflict: Sentiment, Topics, and Event Dynamics
by: Antonakaki, Despoina, et al.
Published: (2025)
by: Antonakaki, Despoina, et al.
Published: (2025)
Monitoring the evolution of antisemitic discourse on extremist social media using BERT
by: Mustafa, Raza Ul, et al.
Published: (2024)
by: Mustafa, Raza Ul, et al.
Published: (2024)
Transit Pulse: Utilizing Social Media as a Source for Customer Feedback and Information Extraction with Large Language Model
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
GLiNER multi-task: Generalist Lightweight Model for Various Information Extraction Tasks
by: Stepanov, Ihor, et al.
Published: (2024)
by: Stepanov, Ihor, et al.
Published: (2024)
USE: Dynamic User Modeling with Stateful Sequence Models
by: Zhou, Zhihan, et al.
Published: (2024)
by: Zhou, Zhihan, et al.
Published: (2024)
Multimodal Analysis of State-Funded News Coverage of the Israel-Hamas War on YouTube Shorts
by: Miehling, Daniel, et al.
Published: (2026)
by: Miehling, Daniel, et al.
Published: (2026)
Planning and Editing What You Retrieve for Enhanced Tool Learning
by: Huang, Tenghao, et al.
Published: (2024)
by: Huang, Tenghao, et al.
Published: (2024)
Can't Remember Details in Long Documents? You Need Some R&R
by: Agrawal, Devanshu, et al.
Published: (2024)
by: Agrawal, Devanshu, et al.
Published: (2024)
Utilizing AI and Social Media Analytics to Discover Adverse Side Effects of GLP-1 Receptor Agonists
by: Bartal, Alon, et al.
Published: (2024)
by: Bartal, Alon, et al.
Published: (2024)
Forgetful by Design? A Critical Audit of YouTube's Search API for Academic Research
by: Rieder, Bernhard, et al.
Published: (2025)
by: Rieder, Bernhard, et al.
Published: (2025)
On Bob Dylan: A Computational Perspective
by: Garg, Prashant
Published: (2025)
by: Garg, Prashant
Published: (2025)
ArXivBench: When You Should Avoid Using ChatGPT for Academic Writing
by: Li, Ning, et al.
Published: (2025)
by: Li, Ning, et al.
Published: (2025)
Public Profile Matters: A Scalable Integrated Approach to Recommend Citations in the Wild
by: Goyal, Karan, et al.
Published: (2026)
by: Goyal, Karan, et al.
Published: (2026)
Large Language Models for Causal Relations Extraction in Social Media: A Validation Framework for Disaster Intelligence
by: Jeong, Ujun, et al.
Published: (2026)
by: Jeong, Ujun, et al.
Published: (2026)
Auditing health-related recommendations in social media: A Case Study of Abortion on YouTube
by: Lahsaini, Mohammed, et al.
Published: (2024)
by: Lahsaini, Mohammed, et al.
Published: (2024)
Graph Neural Networks for Tabular Data Learning: A Survey with Taxonomy and Directions
by: Li, Cheng-Te, et al.
Published: (2024)
by: Li, Cheng-Te, et al.
Published: (2024)
COOL: A Conjoint Perspective on Spatio-Temporal Graph Neural Network for Traffic Forecasting
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
A Comparative Study on Enhancing Prediction in Social Network Advertisement through Data Augmentation
by: Yang, Qikai, et al.
Published: (2024)
by: Yang, Qikai, et al.
Published: (2024)
A Scalable Inter-edge Correlation Modeling in CopulaGNN for Link Sign Prediction
by: Sung, Jinkyu, et al.
Published: (2026)
by: Sung, Jinkyu, et al.
Published: (2026)
A Survey of Graph Neural Networks in Real world: Imbalance, Noise, Privacy and OOD Challenges
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
Modeling Social Media Recommendation Impacts Using Academic Networks: A Graph Neural Network Approach
by: Guidotti, Sabrina, et al.
Published: (2024)
by: Guidotti, Sabrina, et al.
Published: (2024)
Could Small Language Models Serve as Recommenders? Towards Data-centric Cold-start Recommendations
by: Wu, Xuansheng, et al.
Published: (2023)
by: Wu, Xuansheng, et al.
Published: (2023)
AI Approaches to Qualitative and Quantitative News Analytics on NATO Unity
by: Pavlyshenko, Bohdan M.
Published: (2025)
by: Pavlyshenko, Bohdan M.
Published: (2025)
Monitoring Critical Infrastructure Facilities During Disasters Using Large Language Models
by: Ziaullah, Abdul Wahab, et al.
Published: (2024)
by: Ziaullah, Abdul Wahab, et al.
Published: (2024)
Investigating disaster response through social media data and the Susceptible-Infected-Recovered (SIR) model: A case study of 2020 Western U.S. wildfire season
by: Ma, Zihui, et al.
Published: (2023)
by: Ma, Zihui, et al.
Published: (2023)
Chi-Square Wavelet Graph Neural Networks for Heterogeneous Graph Anomaly Detection
by: Li, Xiping, et al.
Published: (2025)
by: Li, Xiping, et al.
Published: (2025)
Graph Representation Learning via Causal Diffusion for Out-of-Distribution Recommendation
by: Zhao, Chu, et al.
Published: (2024)
by: Zhao, Chu, et al.
Published: (2024)
CausalMob: Causal Human Mobility Prediction with LLMs-derived Human Intentions toward Public Events
by: Yang, Xiaojie, et al.
Published: (2024)
by: Yang, Xiaojie, et al.
Published: (2024)
Influence Maximization via Graph Neural Bandits
by: Feng, Yuting, et al.
Published: (2024)
by: Feng, Yuting, et al.
Published: (2024)
Hypergraph-enhanced Dual Semi-supervised Graph Classification
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
Cluster-guided Contrastive Class-imbalanced Graph Classification
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
RAT: Retrieval-Augmented Transformer for Click-Through Rate Prediction
by: Li, Yushen, et al.
Published: (2024)
by: Li, Yushen, et al.
Published: (2024)
Similar Items
-
YTCommentVerse: A Multi-Category Multi-Lingual YouTube Comment Corpus
by: Dutta, Hridoy Sankar, et al.
Published: (2025) -
A Social Data-Driven System for Identifying Estate-related Events and Topics
by: Mu, Wenchuan, et al.
Published: (2025) -
A "Perspectival" Mirror of the Elephant: Investigating Language Bias on Google, ChatGPT, YouTube, and Wikipedia
by: Luo, Queenie, et al.
Published: (2023) -
Utilizing Large Language Models to Identify Reddit Users Considering Vaping Cessation for Digital Interventions
by: Vuruma, Sai Krishna Revanth, et al.
Published: (2024) -
Explainable Identification of Hate Speech towards Islam using Graph Neural Networks
by: Wasi, Azmine Toushik
Published: (2023)