Salvato in:
| Autori principali: | Abdali, Sara, shaham, Sina, Krishnamachari, Bhaskar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2203.13883 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
di: Wu, Bo, et al.
Pubblicazione: (2024)
di: Wu, Bo, et al.
Pubblicazione: (2024)
CrisisViT: A Robust Vision Transformer for Crisis Image Classification
di: Long, Zijun, et al.
Pubblicazione: (2024)
di: Long, Zijun, et al.
Pubblicazione: (2024)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
di: Qi, Peng, et al.
Pubblicazione: (2024)
di: Qi, Peng, et al.
Pubblicazione: (2024)
Personality Analysis from Online Short Video Platforms with Multi-domain Adaptation
di: An, Sixu, et al.
Pubblicazione: (2024)
di: An, Sixu, et al.
Pubblicazione: (2024)
Visual Authority and the Rhetoric of Health Misinformation: A Multimodal Analysis of Social Media Videos
di: Zarei, Mohammad Reza, et al.
Pubblicazione: (2025)
di: Zarei, Mohammad Reza, et al.
Pubblicazione: (2025)
NativE: Multi-modal Knowledge Graph Completion in the Wild
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
di: Li, Dasong, et al.
Pubblicazione: (2025)
di: Li, Dasong, et al.
Pubblicazione: (2025)
Detecting Misinformation in Multimedia Content through Cross-Modal Entity Consistency: A Dual Learning Approach
di: Fu, Zhe, et al.
Pubblicazione: (2024)
di: Fu, Zhe, et al.
Pubblicazione: (2024)
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
di: Jo, Claire Wonjeong, et al.
Pubblicazione: (2024)
di: Jo, Claire Wonjeong, et al.
Pubblicazione: (2024)
Integration of Policy and Reputation based Trust Mechanisms in e-Commerce Industry
di: Siddiqui, Muhammad Yasir, et al.
Pubblicazione: (2024)
di: Siddiqui, Muhammad Yasir, et al.
Pubblicazione: (2024)
VGA: Vision and Graph Fused Attention Network for Rumor Detection
di: Bai, Lin, et al.
Pubblicazione: (2024)
di: Bai, Lin, et al.
Pubblicazione: (2024)
More than Memes: A Multimodal Topic Modeling Approach to Conspiracy Theories on Telegram
di: Steffen, Elisabeth
Pubblicazione: (2024)
di: Steffen, Elisabeth
Pubblicazione: (2024)
Resolving Sentiment Discrepancy for Multimodal Sentiment Detection via Semantics Completion and Decomposition
di: Wu, Daiqing, et al.
Pubblicazione: (2024)
di: Wu, Daiqing, et al.
Pubblicazione: (2024)
Can LLMs Create Legally Relevant Summaries and Analyses of Videos?
di: Hoeben-Kuil, Lyra, et al.
Pubblicazione: (2025)
di: Hoeben-Kuil, Lyra, et al.
Pubblicazione: (2025)
KI-Bilder und die Widerständigkeit der Medienkonvergenz: Von primärer zu sekundärer Intermedialität?
di: Wilde, Lukas R. A.
Pubblicazione: (2024)
di: Wilde, Lukas R. A.
Pubblicazione: (2024)
Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study
di: Bačić, Boris, et al.
Pubblicazione: (2024)
di: Bačić, Boris, et al.
Pubblicazione: (2024)
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
di: Shi, Chuancheng, et al.
Pubblicazione: (2026)
di: Shi, Chuancheng, et al.
Pubblicazione: (2026)
AI-based System for Transforming text and sound to Educational Videos
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
di: Chen, Hongruixuan, et al.
Pubblicazione: (2023)
di: Chen, Hongruixuan, et al.
Pubblicazione: (2023)
Delving Deep into Engagement Prediction of Short Videos
di: Li, Dasong, et al.
Pubblicazione: (2024)
di: Li, Dasong, et al.
Pubblicazione: (2024)
A Rate-Distortion-Classification Approach for Lossy Image Compression
di: Zhang, Yuefeng
Pubblicazione: (2024)
di: Zhang, Yuefeng
Pubblicazione: (2024)
Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech
di: Liu, Rui, et al.
Pubblicazione: (2024)
di: Liu, Rui, et al.
Pubblicazione: (2024)
LLM-based Fusion of Multi-modal Features for Commercial Memorability Prediction
di: Pramov, Aleksandar
Pubblicazione: (2025)
di: Pramov, Aleksandar
Pubblicazione: (2025)
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
di: Fang, Xiang, et al.
Pubblicazione: (2022)
di: Fang, Xiang, et al.
Pubblicazione: (2022)
FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
di: Bhaskar, Paramananda, et al.
Pubblicazione: (2026)
Multi-modal Segment Assemblage Network for Ad Video Editing with Importance-Coherence Reward
di: Tang, Yolo Yunlong, et al.
Pubblicazione: (2022)
di: Tang, Yolo Yunlong, et al.
Pubblicazione: (2022)
PointCoT: A Multi-modal Benchmark for Explicit 3D Geometric Reasoning
di: Zhang, Dongxu, et al.
Pubblicazione: (2026)
di: Zhang, Dongxu, et al.
Pubblicazione: (2026)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
di: Yan, Zehong, et al.
Pubblicazione: (2025)
di: Yan, Zehong, et al.
Pubblicazione: (2025)
Distilling Neuro-Symbolic Programs into 3D Multi-modal LLMs
di: Mo, Wentao, et al.
Pubblicazione: (2026)
di: Mo, Wentao, et al.
Pubblicazione: (2026)
LazyVLM: Neuro-Symbolic Approach to Video Analytics
di: Jian, Xiangru, et al.
Pubblicazione: (2025)
di: Jian, Xiangru, et al.
Pubblicazione: (2025)
Toxic Memes: A Survey of Computational Perspectives on the Detection and Explanation of Meme Toxicities
di: Pandiani, Delfina Sol Martinez, et al.
Pubblicazione: (2024)
di: Pandiani, Delfina Sol Martinez, et al.
Pubblicazione: (2024)
MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
di: Lin, Xiao, et al.
Pubblicazione: (2025)
di: Lin, Xiao, et al.
Pubblicazione: (2025)
Cross-modal Causal Intervention for Alzheimer's Disease Prediction
di: Jin, Yutao, et al.
Pubblicazione: (2025)
di: Jin, Yutao, et al.
Pubblicazione: (2025)
CalliffusionV2: Personalized Natural Calligraphy Generation with Flexible Multi-modal Control
di: Liao, Qisheng, et al.
Pubblicazione: (2024)
di: Liao, Qisheng, et al.
Pubblicazione: (2024)
CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection
di: Li, Fanxiao, et al.
Pubblicazione: (2025)
di: Li, Fanxiao, et al.
Pubblicazione: (2025)
GeoLocator: a location-integrated large multimodal model for inferring geo-privacy
di: Yang, Yifan, et al.
Pubblicazione: (2023)
di: Yang, Yifan, et al.
Pubblicazione: (2023)
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
di: Tian, Zeyue, et al.
Pubblicazione: (2026)
di: Tian, Zeyue, et al.
Pubblicazione: (2026)
Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs
di: Song, Dingjie, et al.
Pubblicazione: (2024)
di: Song, Dingjie, et al.
Pubblicazione: (2024)
PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval
di: Xu, Tianyi, et al.
Pubblicazione: (2026)
di: Xu, Tianyi, et al.
Pubblicazione: (2026)
Cross-Modal Retrieval with Cauchy-Schwarz Divergence
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
di: Zhang, Jiahao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
di: Wu, Bo, et al.
Pubblicazione: (2024) -
CrisisViT: A Robust Vision Transformer for Crisis Image Classification
di: Long, Zijun, et al.
Pubblicazione: (2024) -
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
di: Qi, Peng, et al.
Pubblicazione: (2024) -
Personality Analysis from Online Short Video Platforms with Multi-domain Adaptation
di: An, Sixu, et al.
Pubblicazione: (2024) -
Visual Authority and the Rhetoric of Health Misinformation: A Multimodal Analysis of Social Media Videos
di: Zarei, Mohammad Reza, et al.
Pubblicazione: (2025)