NewsCaption: Named-Entity aware Captioning for Out-of-Context Media
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Anurag, Aneja, Shivangi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SMTPD: A New Benchmark for Temporal Prediction of Social Media Popularity
di: Xu, Yijie, et al.
Pubblicazione: (2025)
di: Xu, Yijie, et al.
Pubblicazione: (2025)
MSM-BD: Multimodal Social Media Bot Detection Using Heterogeneous Information
di: Wu, Tingxuan, et al.
Pubblicazione: (2024)
di: Wu, Tingxuan, et al.
Pubblicazione: (2024)
Enhancing Fake News Detection in Social Media via Label Propagation on Cross-modal Tweet Graph
di: Zhao, Wanqing, et al.
Pubblicazione: (2024)
di: Zhao, Wanqing, et al.
Pubblicazione: (2024)
Shorts on the Rise: Assessing the Effects of YouTube Shorts on Long-Form Video Content
di: Rajendran, Prajit T., et al.
Pubblicazione: (2024)
di: Rajendran, Prajit T., et al.
Pubblicazione: (2024)
YTLive: A Dataset of Real-World YouTube Live Streaming Sessions
di: Mozhganfar, Mojtaba, et al.
Pubblicazione: (2025)
di: Mozhganfar, Mojtaba, et al.
Pubblicazione: (2025)
Temporal Dynamics of Emotions in Italian Online Soccer Fandoms
di: Citraro, Salvatore, et al.
Pubblicazione: (2025)
di: Citraro, Salvatore, et al.
Pubblicazione: (2025)
A New Dataset and Benchmark for Grounding Multimodal Misinformation
di: Yang, Bingjian, et al.
Pubblicazione: (2025)
di: Yang, Bingjian, et al.
Pubblicazione: (2025)
Video Summarization: Towards Entity-Aware Captions
di: Ayyubi, Hammad A., et al.
Pubblicazione: (2023)
di: Ayyubi, Hammad A., et al.
Pubblicazione: (2023)
On the Robustness of Cover Version Identification Models: A Study Using Cover Versions from YouTube
di: Hachmeier, Simon, et al.
Pubblicazione: (2025)
di: Hachmeier, Simon, et al.
Pubblicazione: (2025)
Multi-Modal Discussion Transformer: Integrating Text, Images and Graph Transformers to Detect Hate Speech on Social Media
di: Hebert, Liam, et al.
Pubblicazione: (2023)
di: Hebert, Liam, et al.
Pubblicazione: (2023)
Television Discourse Decoded: Comprehensive Multimodal Analytics at Scale
di: Agarwal, Anmol, et al.
Pubblicazione: (2024)
di: Agarwal, Anmol, et al.
Pubblicazione: (2024)
Social sustainability through engagement in a training context with tools such as the Native Podcast and Facebook social network
di: Bebey, Danielle Mbambe
Pubblicazione: (2025)
di: Bebey, Danielle Mbambe
Pubblicazione: (2025)
Seeing Sarcasm Through Different Eyes: Analyzing Multimodal Sarcasm Perception in Large Vision-Language Models
di: Chen, Junjie, et al.
Pubblicazione: (2025)
di: Chen, Junjie, et al.
Pubblicazione: (2025)
Stickers on Facebook: Multifunctionality and face-enhancing politeness in everyday social interaction
di: Porrino-Moscoso, Laura M.
Pubblicazione: (2026)
di: Porrino-Moscoso, Laura M.
Pubblicazione: (2026)
Semantic Codebook Learning for Dynamic Recommendation Models
di: Lv, Zheqi, et al.
Pubblicazione: (2024)
di: Lv, Zheqi, et al.
Pubblicazione: (2024)
Hallucination Localization in Video Captioning
di: Nakada, Shota, et al.
Pubblicazione: (2025)
di: Nakada, Shota, et al.
Pubblicazione: (2025)
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
di: Wu, Bo, et al.
Pubblicazione: (2024)
di: Wu, Bo, et al.
Pubblicazione: (2024)
Visual Lifelog Retrieval through Captioning-Enhanced Interpretation
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
An Avalanche of Images on Telegram Preceded Russia's Full-Scale Invasion of Ukraine
di: Theisen, William, et al.
Pubblicazione: (2024)
di: Theisen, William, et al.
Pubblicazione: (2024)
Effects of Limited Field of View on Musical Collaboration Experience with Avatars in Extended Reality
di: Weng, Suibi Che-Chuan, et al.
Pubblicazione: (2026)
di: Weng, Suibi Che-Chuan, et al.
Pubblicazione: (2026)
Visual Authority and the Rhetoric of Health Misinformation: A Multimodal Analysis of Social Media Videos
di: Zarei, Mohammad Reza, et al.
Pubblicazione: (2025)
di: Zarei, Mohammad Reza, et al.
Pubblicazione: (2025)
TriMod Fusion for Multimodal Named Entity Recognition in Social Media
di: Alfaqeeh, Mosab
Pubblicazione: (2025)
di: Alfaqeeh, Mosab
Pubblicazione: (2025)
Chasing RATs: Tracing Reading for and as Creative Activity
di: Liu, Sophia, et al.
Pubblicazione: (2026)
di: Liu, Sophia, et al.
Pubblicazione: (2026)
MagnetDB: A Longitudinal Torrent Discovery Dataset with IMDb-Matched Movies and TV Shows
di: Seidenberger, Scott, et al.
Pubblicazione: (2025)
di: Seidenberger, Scott, et al.
Pubblicazione: (2025)
HateSieve: A Contrastive Learning Framework for Detecting and Segmenting Hateful Content in Multimodal Memes
di: Su, Xuanyu, et al.
Pubblicazione: (2024)
di: Su, Xuanyu, et al.
Pubblicazione: (2024)
Integration of Policy and Reputation based Trust Mechanisms in e-Commerce Industry
di: Siddiqui, Muhammad Yasir, et al.
Pubblicazione: (2024)
di: Siddiqui, Muhammad Yasir, et al.
Pubblicazione: (2024)
Delving Deep into Engagement Prediction of Short Videos
di: Li, Dasong, et al.
Pubblicazione: (2024)
di: Li, Dasong, et al.
Pubblicazione: (2024)
Let Community Rules Be Reflected in Online Content Moderation
di: Xin, Wangjiaxuan, et al.
Pubblicazione: (2024)
di: Xin, Wangjiaxuan, et al.
Pubblicazione: (2024)
VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
di: Li, Dasong, et al.
Pubblicazione: (2025)
di: Li, Dasong, et al.
Pubblicazione: (2025)
VLM as Policy: Common-Law Content Moderation Framework for Short Video Platform
di: Lu, Xingyu, et al.
Pubblicazione: (2025)
di: Lu, Xingyu, et al.
Pubblicazione: (2025)
A Split-Window Transformer for Multi-Model Sequence Spammer Detection using Multi-Model Variational Autoencoder
di: Yang, Zhou, et al.
Pubblicazione: (2025)
di: Yang, Zhou, et al.
Pubblicazione: (2025)
Cap2Sum: Learning to Summarize Videos by Generating Captions
di: Zhao, Cairong, et al.
Pubblicazione: (2024)
di: Zhao, Cairong, et al.
Pubblicazione: (2024)
Verifying Cross-modal Entity Consistency in News using Vision-language Models
di: Tahmasebi, Sahar, et al.
Pubblicazione: (2025)
di: Tahmasebi, Sahar, et al.
Pubblicazione: (2025)
Unveiling Voices: A Co-Hashtag Analysis of TikTok Discourse on the 2023 Israel-Palestine Crisis
di: Hasin, Rozin
Pubblicazione: (2025)
di: Hasin, Rozin
Pubblicazione: (2025)
Muse-it: A Tool for Analyzing Music Discourse on Reddit
di: Agarwala, Jatin, et al.
Pubblicazione: (2025)
di: Agarwala, Jatin, et al.
Pubblicazione: (2025)
Agency Among Agents: Designing with Hypertextual Friction in the Algorithmic Web
di: Liu, Sophia, et al.
Pubblicazione: (2025)
di: Liu, Sophia, et al.
Pubblicazione: (2025)
Short-video Propagation Influence Rating: A New Real-world Dataset and A New Large Graph Model
di: Xue, Dizhan, et al.
Pubblicazione: (2025)
di: Xue, Dizhan, et al.
Pubblicazione: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
di: Mo, Xian, et al.
Pubblicazione: (2025)
di: Mo, Xian, et al.
Pubblicazione: (2025)
Differentially Processed Optimized Collaborative Rich Text Editor
di: Jatana, Nishtha, et al.
Pubblicazione: (2024)
di: Jatana, Nishtha, et al.
Pubblicazione: (2024)
Voices, Faces, and Feelings: Multi-modal Emotion-Cognition Captioning for Mental Health Understanding
di: Zhou, Zhiyuan, et al.
Pubblicazione: (2026)
di: Zhou, Zhiyuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SMTPD: A New Benchmark for Temporal Prediction of Social Media Popularity
di: Xu, Yijie, et al.
Pubblicazione: (2025) -
MSM-BD: Multimodal Social Media Bot Detection Using Heterogeneous Information
di: Wu, Tingxuan, et al.
Pubblicazione: (2024) -
Enhancing Fake News Detection in Social Media via Label Propagation on Cross-modal Tweet Graph
di: Zhao, Wanqing, et al.
Pubblicazione: (2024) -
Shorts on the Rise: Assessing the Effects of YouTube Shorts on Long-Form Video Content
di: Rajendran, Prajit T., et al.
Pubblicazione: (2024) -
YTLive: A Dataset of Real-World YouTube Live Streaming Sessions
di: Mozhganfar, Mojtaba, et al.
Pubblicazione: (2025)