SoMeLVLM: A Large Vision Language Model for Social Media Processing
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Xinnong, Kuang, Haoyu, Mou, Xinyi, Lyu, Hanjia, Wu, Kun, Chen, Siming, Luo, Jiebo, Huang, Xuanjing, Wei, Zhongyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Holistic Visual-Textual Sentiment Analysis with Prior Models
di: Chen, Junyu, et al.
Pubblicazione: (2022)
di: Chen, Junyu, et al.
Pubblicazione: (2022)
EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
di: Du, Mengfei, et al.
Pubblicazione: (2024)
di: Du, Mengfei, et al.
Pubblicazione: (2024)
ElectionSim: Massive Population Election Simulation Powered by Large Language Model Driven Agents
di: Zhang, Xinnong, et al.
Pubblicazione: (2024)
di: Zhang, Xinnong, et al.
Pubblicazione: (2024)
E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection
di: Wu, Junjie, et al.
Pubblicazione: (2025)
di: Wu, Junjie, et al.
Pubblicazione: (2025)
Revisiting Vision-Language Features Adaptation and Inconsistency for Social Media Popularity Prediction
di: Hsu, Chih-Chung, et al.
Pubblicazione: (2024)
di: Hsu, Chih-Chung, et al.
Pubblicazione: (2024)
Unveiling the Truth and Facilitating Change: Towards Agent-based Large-scale Social Movement Simulation
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge
di: Wu, Bo, et al.
Pubblicazione: (2024)
di: Wu, Bo, et al.
Pubblicazione: (2024)
AI-Press: A Multi-Agent News Generating and Feedback Simulation System Powered by Large Language Models
di: Liu, Xiawei, et al.
Pubblicazione: (2024)
di: Liu, Xiawei, et al.
Pubblicazione: (2024)
MORE-R1: Guiding LVLM for Multimodal Object-Entity Relation Extraction via Stepwise Reasoning with Reinforcement Learning
di: Yuan, Xiang, et al.
Pubblicazione: (2026)
di: Yuan, Xiang, et al.
Pubblicazione: (2026)
ConvBench: A Multi-Turn Conversation Evaluation Benchmark with Hierarchical Capability for Large Vision-Language Models
di: Liu, Shuo, et al.
Pubblicazione: (2024)
di: Liu, Shuo, et al.
Pubblicazione: (2024)
Representation Bias in Political Sample Simulations with Large Language Models
di: Qi, Weihong, et al.
Pubblicazione: (2024)
di: Qi, Weihong, et al.
Pubblicazione: (2024)
AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
di: Mou, Xinyi, et al.
Pubblicazione: (2024)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
PopSim: Social Network Simulation for Social Media Popularity Prediction
di: Liu, Yijun, et al.
Pubblicazione: (2025)
di: Liu, Yijun, et al.
Pubblicazione: (2025)
Dual Attribute-Spatial Relation Alignment for 3D Visual Grounding
di: Xu, Yue, et al.
Pubblicazione: (2024)
di: Xu, Yue, et al.
Pubblicazione: (2024)
Human vs. LMMs: Exploring the Discrepancy in Emoji Interpretation and Usage in Digital Communication
di: Lyu, Hanjia, et al.
Pubblicazione: (2024)
di: Lyu, Hanjia, et al.
Pubblicazione: (2024)
Self-Comparison for Dataset-Level Membership Inference in Large (Vision-)Language Models
di: Ren, Jie, et al.
Pubblicazione: (2024)
di: Ren, Jie, et al.
Pubblicazione: (2024)
Content-Adaptive Rate-Quality Curve Prediction Model in Media Processing System
di: Yin, Shibo, et al.
Pubblicazione: (2024)
di: Yin, Shibo, et al.
Pubblicazione: (2024)
PureKV: Plug-and-Play KV Cache Optimization with Spatial-Temporal Sparse Attention for Vision-Language Large Models
di: Jiang, Zhonghua, et al.
Pubblicazione: (2025)
di: Jiang, Zhonghua, et al.
Pubblicazione: (2025)
ChartAdapter: Large Vision-Language Model for Chart Summarization
di: Xu, Peixin, et al.
Pubblicazione: (2024)
di: Xu, Peixin, et al.
Pubblicazione: (2024)
Fact-Checking with Contextual Narratives: Leveraging Retrieval-Augmented LLMs for Social Media Analysis
di: Dey, Arka Ujjal, et al.
Pubblicazione: (2025)
di: Dey, Arka Ujjal, et al.
Pubblicazione: (2025)
Listen, Pause, and Reason: Toward Perception-Grounded Hybrid Reasoning for Audio Understanding
di: Wang, Jieyi, et al.
Pubblicazione: (2026)
di: Wang, Jieyi, et al.
Pubblicazione: (2026)
Anchoring Trends: Mitigating Social Media Popularity Prediction Drift via Feature Clustering and Expansion
di: Lee, Chia-Ming, et al.
Pubblicazione: (2025)
di: Lee, Chia-Ming, et al.
Pubblicazione: (2025)
Can LLMs Simulate Social Media Engagement? A Study on Action-Guided Response Generation
di: Qiu, Zhongyi, et al.
Pubblicazione: (2025)
di: Qiu, Zhongyi, et al.
Pubblicazione: (2025)
COPA: Efficient Vision-Language Pre-training Through Collaborative Object- and Patch-Text Alignment
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models
di: Huang, Yuhang, et al.
Pubblicazione: (2024)
di: Huang, Yuhang, et al.
Pubblicazione: (2024)
Rethinking Vision Transformer for Large-Scale Fine-Grained Image Retrieval
di: Jiang, Xin, et al.
Pubblicazione: (2025)
di: Jiang, Xin, et al.
Pubblicazione: (2025)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
Segmentation-Based Attention Entropy: Detecting and Mitigating Object Hallucinations in Large Vision-Language Models
di: Song, Jiale, et al.
Pubblicazione: (2026)
di: Song, Jiale, et al.
Pubblicazione: (2026)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
di: Xu, Junhao, et al.
Pubblicazione: (2025)
di: Xu, Junhao, et al.
Pubblicazione: (2025)
CalliReader: Contextualizing Chinese Calligraphy via an Embedding-Aligned Vision-Language Model
di: Luo, Yuxuan, et al.
Pubblicazione: (2025)
di: Luo, Yuxuan, et al.
Pubblicazione: (2025)
Gaze and Glow: Exploring Editing Processes on Social Media through Interactive Exhibition
di: Hong, Yang, et al.
Pubblicazione: (2025)
di: Hong, Yang, et al.
Pubblicazione: (2025)
Multimodal Emotion Recognition with Large Language Models
di: Zhang, Hongrui, et al.
Pubblicazione: (2026)
di: Zhang, Hongrui, et al.
Pubblicazione: (2026)
Identity-Preserving Text-to-Video Generation by Frequency Decomposition
di: Yuan, Shenghai, et al.
Pubblicazione: (2024)
di: Yuan, Shenghai, et al.
Pubblicazione: (2024)
SVLA: A Unified Speech-Vision-Language Assistant with Multimodal Reasoning and Speech Generation
di: Huynh, Ngoc Dung, et al.
Pubblicazione: (2025)
di: Huynh, Ngoc Dung, et al.
Pubblicazione: (2025)
Large Language Models (LLMs): Deployment, Tokenomics and Sustainability
di: Dong, Haiwei, et al.
Pubblicazione: (2024)
di: Dong, Haiwei, et al.
Pubblicazione: (2024)
AI-Integrated Decision Support System for Real-Time Market Growth Forecasting and Multi-Source Content Diffusion Analytics
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning
di: Wang, Youze, et al.
Pubblicazione: (2023)
di: Wang, Youze, et al.
Pubblicazione: (2023)
EmoVLM-KD: Fusing Distilled Expertise with Vision-Language Models for Visual Emotion Analysis
di: Lee, SangEun, et al.
Pubblicazione: (2025)
di: Lee, SangEun, et al.
Pubblicazione: (2025)
SocialDF: Benchmark Dataset and Detection Model for Mitigating Harmful Deepfake Content on Social Media Platforms
di: Batra, Arnesh, et al.
Pubblicazione: (2025)
di: Batra, Arnesh, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Holistic Visual-Textual Sentiment Analysis with Prior Models
di: Chen, Junyu, et al.
Pubblicazione: (2022) -
EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
di: Du, Mengfei, et al.
Pubblicazione: (2024) -
ElectionSim: Massive Population Election Simulation Powered by Large Language Model Driven Agents
di: Zhang, Xinnong, et al.
Pubblicazione: (2024) -
E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection
di: Wu, Junjie, et al.
Pubblicazione: (2025) -
Revisiting Vision-Language Features Adaptation and Inconsistency for Social Media Popularity Prediction
di: Hsu, Chih-Chung, et al.
Pubblicazione: (2024)