Hierarchical Level-Wise News Article Clustering via Multilingual Matryoshka Embeddings
Fuente:
arXiv
Saved in:
| Main Authors: | Hanley, Hans W. A., Durumeric, Zakir |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Partial Mobilization: Tracking Multilingual Information Flows Amongst Russian Media Outlets and Telegram
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
Machine-Made Media: Monitoring the Mobilization of Machine-Generated Articles on Misinformation and Mainstream News Websites
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
Tracking the Takes and Trajectories of English-Language News Narratives across Trustworthy and Worrisome Websites
by: Hanley, Hans W. A., et al.
Published: (2025)
by: Hanley, Hans W. A., et al.
Published: (2025)
Sub-Standards and Mal-Practices: Misinformation's Role in Insular, Polarized, and Toxic Interactions on Reddit
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
Twits, Toxic Tweets, and Tribal Tendencies: Trends in Politically Polarized Posts on Twitter
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
Watch Your Language: Investigating Content Moderation with Large Language Models
by: Kumar, Deepak, et al.
Published: (2023)
by: Kumar, Deepak, et al.
Published: (2023)
Specious Sites: Tracking the Spread and Sway of Spurious News Stories at Scale
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
A Novel Method for News Article Event-Based Embedding
by: Ishlach, Koren, et al.
Published: (2024)
by: Ishlach, Koren, et al.
Published: (2024)
TATA: Stance Detection via Topic-Agnostic and Topic-Aware Embeddings
by: Hanley, Hans W. A., et al.
Published: (2023)
by: Hanley, Hans W. A., et al.
Published: (2023)
User-Aware Multilingual Abusive Content Detection in Social Media
by: Rehman, Mohammad Zia Ur, et al.
Published: (2024)
by: Rehman, Mohammad Zia Ur, et al.
Published: (2024)
The 2021 Tokyo Olympics Multilingual News Article Dataset
by: Novak, Erik, et al.
Published: (2025)
by: Novak, Erik, et al.
Published: (2025)
Style-News: Incorporating Stylized News Generation and Adversarial Verification for Neural Fake News Detection
by: Wang, Wei-Yao, et al.
Published: (2024)
by: Wang, Wei-Yao, et al.
Published: (2024)
Exposing and Explaining Fake News On-the-Fly
by: de Arriba-Pérez, Francisco, et al.
Published: (2024)
by: de Arriba-Pérez, Francisco, et al.
Published: (2024)
Entity Insertion in Multilingual Linked Corpora: The Case of Wikipedia
by: Feith, Tomás, et al.
Published: (2024)
by: Feith, Tomás, et al.
Published: (2024)
Can Large Language Models Detect Misinformation in Scientific News Reporting?
by: Cao, Yupeng, et al.
Published: (2024)
by: Cao, Yupeng, et al.
Published: (2024)
From Skepticism to Acceptance: Simulating the Attitude Dynamics Toward Fake News
by: Liu, Yuhan, et al.
Published: (2024)
by: Liu, Yuhan, et al.
Published: (2024)
Evolving to the Future: Unseen Event Adaptive Fake News Detection on Social Media
by: Zhang, Jiajun, et al.
Published: (2024)
by: Zhang, Jiajun, et al.
Published: (2024)
Revisiting Fake News Detection: Towards Temporality-aware Evaluation by Leveraging Engagement Earliness
by: Kim, Junghoon, et al.
Published: (2024)
by: Kim, Junghoon, et al.
Published: (2024)
Multimodal Analysis of State-Funded News Coverage of the Israel-Hamas War on YouTube Shorts
by: Miehling, Daniel, et al.
Published: (2026)
by: Miehling, Daniel, et al.
Published: (2026)
Transferring Structure Knowledge: A New Task to Fake news Detection Towards Cold-Start Propagation
by: Wei, Lingwei, et al.
Published: (2024)
by: Wei, Lingwei, et al.
Published: (2024)
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts
by: Askari, Hadi, et al.
Published: (2024)
by: Askari, Hadi, et al.
Published: (2024)
AI Approaches to Qualitative and Quantitative News Analytics on NATO Unity
by: Pavlyshenko, Bohdan M.
Published: (2025)
by: Pavlyshenko, Bohdan M.
Published: (2025)
HumVI: A Multilingual Dataset for Detecting Violent Incidents Impacting Humanitarian Aid
by: Lamba, Hemank, et al.
Published: (2024)
by: Lamba, Hemank, et al.
Published: (2024)
Mental Disorder Classification via Temporal Representation of Text
by: Kumar, Raja, et al.
Published: (2024)
by: Kumar, Raja, et al.
Published: (2024)
MASim: Multilingual Agent-Based Simulation for Social Science
by: Zhang, Xuan, et al.
Published: (2025)
by: Zhang, Xuan, et al.
Published: (2025)
GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing
by: Wang, Ming, et al.
Published: (2026)
by: Wang, Ming, et al.
Published: (2026)
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models
by: Nghiem, Huy, et al.
Published: (2024)
by: Nghiem, Huy, et al.
Published: (2024)
GuideWalk: A Novel Graph-Based Word Embedding for Enhanced Text Classification
by: Mohammed, Sarmad N., et al.
Published: (2024)
by: Mohammed, Sarmad N., et al.
Published: (2024)
Hate Speech Detection Using Cross-Platform Social Media Data In English and German Language
by: Shahi, Gautam Kishore, et al.
Published: (2024)
by: Shahi, Gautam Kishore, et al.
Published: (2024)
Decoding Multilingual Topic Dynamics and Trend Identification through ARIMA Time Series Analysis on Social Networks: A Novel Data Translation Framework Enhanced by LDA/HDP Models
by: Jaballi, Samawel, et al.
Published: (2024)
by: Jaballi, Samawel, et al.
Published: (2024)
Less is More: Unseen Domain Fake News Detection via Causal Propagation Substructures
by: Gong, Shuzhi, et al.
Published: (2024)
by: Gong, Shuzhi, et al.
Published: (2024)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
by: Jenny, David F., et al.
Published: (2023)
by: Jenny, David F., et al.
Published: (2023)
Large Language Models for Pedestrian Safety: An Application to Predicting Driver Yielding Behavior at Unsignalized Intersections
by: Yang, Yicheng, et al.
Published: (2025)
by: Yang, Yicheng, et al.
Published: (2025)
Patients Speak, AI Listens: LLM-based Analysis of Online Reviews Uncovers Key Drivers for Urgent Care Satisfaction
by: Xu, Xiaoran, et al.
Published: (2025)
by: Xu, Xiaoran, et al.
Published: (2025)
LLMTaxo: Leveraging Large Language Models for Constructing Taxonomy of Factual Claims from Social Media
by: Zhang, Haiqi, et al.
Published: (2025)
by: Zhang, Haiqi, et al.
Published: (2025)
Network-informed Prompt Engineering against Organized Astroturf Campaigns under Extreme Class Imbalance
by: Kanakaris, Nikos, et al.
Published: (2025)
by: Kanakaris, Nikos, et al.
Published: (2025)
How Social is It? A Benchmark for LLMs' Capabilities in Multi-user Multi-turn Social Agent Tasks
by: Wu, Yusen, et al.
Published: (2025)
by: Wu, Yusen, et al.
Published: (2025)
Simulating Misinformation Vulnerabilities With Agent Personas
by: Farr, David, et al.
Published: (2025)
by: Farr, David, et al.
Published: (2025)
BEYONDWORDS is All You Need: Agentic Generative AI based Social Media Themes Extractor
by: Ghali, Mohammed-Khalil, et al.
Published: (2025)
by: Ghali, Mohammed-Khalil, et al.
Published: (2025)
Human Mobility Datasets Enriched With Contextual and Social Dimensions
by: Pugliese, Chiara, et al.
Published: (2025)
by: Pugliese, Chiara, et al.
Published: (2025)
Similar Items
-
Partial Mobilization: Tracking Multilingual Information Flows Amongst Russian Media Outlets and Telegram
by: Hanley, Hans W. A., et al.
Published: (2023) -
Machine-Made Media: Monitoring the Mobilization of Machine-Generated Articles on Misinformation and Mainstream News Websites
by: Hanley, Hans W. A., et al.
Published: (2023) -
Tracking the Takes and Trajectories of English-Language News Narratives across Trustworthy and Worrisome Websites
by: Hanley, Hans W. A., et al.
Published: (2025) -
Sub-Standards and Mal-Practices: Misinformation's Role in Insular, Polarized, and Toxic Interactions on Reddit
by: Hanley, Hans W. A., et al.
Published: (2023) -
Twits, Toxic Tweets, and Tribal Tendencies: Trends in Politically Polarized Posts on Twitter
by: Hanley, Hans W. A., et al.
Published: (2023)