Detecting Multimedia Generated by Large AI Models: A Survey
Fuente:
arXiv
Salvato in:
| Autori principali: | Lin, Li, Gupta, Neeraj, Zhang, Yue, Ren, Hainan, Liu, Chun-Hao, Ding, Feng, Wang, Xin, Li, Xin, Verdoliva, Luisa, Hu, Shu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs
di: Desai, Shail, et al.
Pubblicazione: (2025)
di: Desai, Shail, et al.
Pubblicazione: (2025)
Generative AI-enabled Mobile Tactical Multimedia Networks: Distribution, Generation, and Perception
di: Xu, Minrui, et al.
Pubblicazione: (2024)
di: Xu, Minrui, et al.
Pubblicazione: (2024)
Identity-Driven Multimedia Forgery Detection via Reference Assistance
di: Xu, Junhao, et al.
Pubblicazione: (2024)
di: Xu, Junhao, et al.
Pubblicazione: (2024)
A Survey on Multimodal Benchmarks: In the Era of Large AI Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Digital Fingerprinting on Multimedia: A Survey
di: Chen, Wendi, et al.
Pubblicazione: (2024)
di: Chen, Wendi, et al.
Pubblicazione: (2024)
Rethinking Vision Transformer for Large-Scale Fine-Grained Image Retrieval
di: Jiang, Xin, et al.
Pubblicazione: (2025)
di: Jiang, Xin, et al.
Pubblicazione: (2025)
Improving Generalization for AI-Synthesized Voice Detection
di: Ren, Hainan, et al.
Pubblicazione: (2024)
di: Ren, Hainan, et al.
Pubblicazione: (2024)
Don't Guess, Escalate: Towards Explainable Uncertainty-Calibrated AI Forensic Agents
di: Boato, Giulia, et al.
Pubblicazione: (2025)
di: Boato, Giulia, et al.
Pubblicazione: (2025)
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
di: Jin, Yili, et al.
Pubblicazione: (2025)
di: Jin, Yili, et al.
Pubblicazione: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
di: Mo, Xian, et al.
Pubblicazione: (2025)
di: Mo, Xian, et al.
Pubblicazione: (2025)
A Multimedia Framework for Continuum Robots: Systematic, Computational, and Control Perspectives
di: Hsieh, Po-Yu, et al.
Pubblicazione: (2024)
di: Hsieh, Po-Yu, et al.
Pubblicazione: (2024)
Stepwise Schema-Guided Prompting Framework with Parameter Efficient Instruction Tuning for Multimedia Event Extraction
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
Harmful Visual Content Manipulation Matters in Misinformation Detection Under Multimedia Scenarios
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
Privacy-Preserving Multimedia Mobile Cloud Computing Using Protective Perturbation
di: Tang, Zhongze, et al.
Pubblicazione: (2024)
di: Tang, Zhongze, et al.
Pubblicazione: (2024)
ConvBench: A Multi-Turn Conversation Evaluation Benchmark with Hierarchical Capability for Large Vision-Language Models
di: Liu, Shuo, et al.
Pubblicazione: (2024)
di: Liu, Shuo, et al.
Pubblicazione: (2024)
Inclusion 2024 Global Multimedia Deepfake Detection Challenge: Towards Multi-dimensional Face Forgery Detection
di: Zhang, Yi, et al.
Pubblicazione: (2024)
di: Zhang, Yi, et al.
Pubblicazione: (2024)
Harnessing Multimodal Large Language Models for Personalized Product Search with Query-aware Refinement
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
di: Zhang, Beibei, et al.
Pubblicazione: (2025)
Deepfake Detection: A Comprehensive Survey from the Reliability Perspective
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
Nagare Media Ingest: A System for Multimedia Ingest Workflows
di: Neugebauer, Matthias
Pubblicazione: (2025)
di: Neugebauer, Matthias
Pubblicazione: (2025)
Reducing Latency for Multimedia Broadcast Services Over Mobile Networks
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
OOD-GraphLLM: Graph Large Language Model for Out-of-Distribution Generalized Drug Synergy Prediction
di: Wang, Xin, et al.
Pubblicazione: (2026)
di: Wang, Xin, et al.
Pubblicazione: (2026)
Performance Evaluation in Multimedia Retrieval
di: Sauter, Loris, et al.
Pubblicazione: (2024)
di: Sauter, Loris, et al.
Pubblicazione: (2024)
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
di: Zhou, Pengyuan, et al.
Pubblicazione: (2024)
di: Zhou, Pengyuan, et al.
Pubblicazione: (2024)
Block Erasure-Aware Semantic Multimedia Compression via JSCC Autoencoder
di: Esfahanizadeh, Homa, et al.
Pubblicazione: (2026)
di: Esfahanizadeh, Homa, et al.
Pubblicazione: (2026)
BeFA: A General Behavior-driven Feature Adapter for Multimedia Recommendation
di: Fan, Qile, et al.
Pubblicazione: (2024)
di: Fan, Qile, et al.
Pubblicazione: (2024)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
di: Jin, Yili, et al.
Pubblicazione: (2025)
di: Jin, Yili, et al.
Pubblicazione: (2025)
DeepTextMark: A Deep Learning-Driven Text Watermarking Approach for Identifying Large Language Model Generated Text
di: Munyer, Travis, et al.
Pubblicazione: (2023)
di: Munyer, Travis, et al.
Pubblicazione: (2023)
DeepStream: Prototyping Deep Joint Source-Channel Coding for Real-Time Multimedia Transmissions
di: Chi, Kaiyi, et al.
Pubblicazione: (2025)
di: Chi, Kaiyi, et al.
Pubblicazione: (2025)
Characterizing Multimedia Information Environment through Multi-modal Clustering of YouTube Videos
di: Yousefi, Niloofar, et al.
Pubblicazione: (2024)
di: Yousefi, Niloofar, et al.
Pubblicazione: (2024)
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation
di: Zhang, Bo, et al.
Pubblicazione: (2024)
di: Zhang, Bo, et al.
Pubblicazione: (2024)
Detecting Misinformation in Multimedia Content through Cross-Modal Entity Consistency: A Dual Learning Approach
di: Fu, Zhe, et al.
Pubblicazione: (2024)
di: Fu, Zhe, et al.
Pubblicazione: (2024)
Mining the Social Fabric: Unveiling Communities for Fake News Detection in Short Videos
di: Gong, Haisong, et al.
Pubblicazione: (2025)
di: Gong, Haisong, et al.
Pubblicazione: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
di: Wang, Sen, et al.
Pubblicazione: (2024)
di: Wang, Sen, et al.
Pubblicazione: (2024)
RoboTron-Mani: All-in-One Multimodal Large Model for Robotic Manipulation
di: Yan, Feng, et al.
Pubblicazione: (2024)
di: Yan, Feng, et al.
Pubblicazione: (2024)
Has Multimodal Learning Delivered Universal Intelligence in Healthcare? A Comprehensive Survey
di: Lin, Qika, et al.
Pubblicazione: (2024)
di: Lin, Qika, et al.
Pubblicazione: (2024)
Human Motion Video Generation: A Survey
di: Xue, Haiwei, et al.
Pubblicazione: (2025)
di: Xue, Haiwei, et al.
Pubblicazione: (2025)
Retrieval-Augmented Multimodal Model for Fake News Detection
di: Li, Yiheng, et al.
Pubblicazione: (2026)
di: Li, Yiheng, et al.
Pubblicazione: (2026)
Design of a 5G Multimedia Broadcast Application Function Supporting Adaptive Error Recovery
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
di: Lentisco, C. M., et al.
Pubblicazione: (2024)
Nagare Media Engine: A System for Cloud- and Edge-Native Network-based Multimedia Workflows
di: Neugebauer, Matthias
Pubblicazione: (2025)
di: Neugebauer, Matthias
Pubblicazione: (2025)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
di: Xu, Junhao, et al.
Pubblicazione: (2025)
di: Xu, Junhao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs
di: Desai, Shail, et al.
Pubblicazione: (2025) -
Generative AI-enabled Mobile Tactical Multimedia Networks: Distribution, Generation, and Perception
di: Xu, Minrui, et al.
Pubblicazione: (2024) -
Identity-Driven Multimedia Forgery Detection via Reference Assistance
di: Xu, Junhao, et al.
Pubblicazione: (2024) -
A Survey on Multimodal Benchmarks: In the Era of Large AI Models
di: Li, Lin, et al.
Pubblicazione: (2024) -
Digital Fingerprinting on Multimedia: A Survey
di: Chen, Wendi, et al.
Pubblicazione: (2024)