KI-Bilder und die Widerständigkeit der Medienkonvergenz: Von primärer zu sekundärer Intermedialität?
Fuente:
arXiv
Salvato in:
| Autore principale: | Wilde, Lukas R. A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AI-based System for Transforming text and sound to Educational Videos
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
di: Jo, Claire Wonjeong, et al.
Pubblicazione: (2024)
di: Jo, Claire Wonjeong, et al.
Pubblicazione: (2024)
Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study
di: Bačić, Boris, et al.
Pubblicazione: (2024)
di: Bačić, Boris, et al.
Pubblicazione: (2024)
Can LLMs Create Legally Relevant Summaries and Analyses of Videos?
di: Hoeben-Kuil, Lyra, et al.
Pubblicazione: (2025)
di: Hoeben-Kuil, Lyra, et al.
Pubblicazione: (2025)
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
di: Shi, Chuancheng, et al.
Pubblicazione: (2026)
di: Shi, Chuancheng, et al.
Pubblicazione: (2026)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
di: Chen, Hongruixuan, et al.
Pubblicazione: (2023)
di: Chen, Hongruixuan, et al.
Pubblicazione: (2023)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
di: Qi, Peng, et al.
Pubblicazione: (2024)
di: Qi, Peng, et al.
Pubblicazione: (2024)
MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
di: Lin, Xiao, et al.
Pubblicazione: (2025)
di: Lin, Xiao, et al.
Pubblicazione: (2025)
Embedding an Ethical Mind: Aligning Text-to-Image Synthesis via Lightweight Value Optimization
di: Wang, Xingqi, et al.
Pubblicazione: (2024)
di: Wang, Xingqi, et al.
Pubblicazione: (2024)
Unmasking Illusions: Understanding Human Perception of Audiovisual Deepfakes
di: Hashmi, Ammarah, et al.
Pubblicazione: (2024)
di: Hashmi, Ammarah, et al.
Pubblicazione: (2024)
Exploring the latent space of diffusion models directly through singular value decomposition
di: Wang, Li, et al.
Pubblicazione: (2025)
di: Wang, Li, et al.
Pubblicazione: (2025)
Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
di: Yu, Lijun, et al.
Pubblicazione: (2023)
di: Yu, Lijun, et al.
Pubblicazione: (2023)
Question-Answering Dense Video Events
di: Qin, Hangyu, et al.
Pubblicazione: (2024)
di: Qin, Hangyu, et al.
Pubblicazione: (2024)
VidCtx: Context-aware Video Question Answering with Image Models
di: Goulas, Andreas, et al.
Pubblicazione: (2024)
di: Goulas, Andreas, et al.
Pubblicazione: (2024)
Cross-domain Multi-step Thinking: Zero-shot Fine-grained Traffic Sign Recognition in the Wild
di: Gan, Yaozong, et al.
Pubblicazione: (2024)
di: Gan, Yaozong, et al.
Pubblicazione: (2024)
Multimodal Markup Document Models for Graphic Design Completion
di: Kikuchi, Kotaro, et al.
Pubblicazione: (2024)
di: Kikuchi, Kotaro, et al.
Pubblicazione: (2024)
When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering
di: Mandelli, Sara, et al.
Pubblicazione: (2024)
di: Mandelli, Sara, et al.
Pubblicazione: (2024)
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
di: Park, Kyu Ri, et al.
Pubblicazione: (2024)
di: Park, Kyu Ri, et al.
Pubblicazione: (2024)
Rethink Predicting the Optical Flow with the Kinetics Perspective
di: Cheng, Yuhao, et al.
Pubblicazione: (2024)
di: Cheng, Yuhao, et al.
Pubblicazione: (2024)
Boosting Audio Visual Question Answering via Key Semantic-Aware Cues
di: Li, Guangyao, et al.
Pubblicazione: (2024)
di: Li, Guangyao, et al.
Pubblicazione: (2024)
Case-based reasoning approach for diagnostic screening of children with developmental delays
di: Song, Zichen, et al.
Pubblicazione: (2024)
di: Song, Zichen, et al.
Pubblicazione: (2024)
Robust Latent Representation Tuning for Image-text Classification
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
SSTFB: Leveraging self-supervised pretext learning and temporal self-attention with feature branching for real-time video polyp segmentation
di: Xu, Ziang, et al.
Pubblicazione: (2024)
di: Xu, Ziang, et al.
Pubblicazione: (2024)
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
di: Mahon, Louis, et al.
Pubblicazione: (2024)
di: Mahon, Louis, et al.
Pubblicazione: (2024)
Kandinsky 3: Text-to-Image Synthesis for Multifunctional Generative Framework
di: Arkhipkin, Vladimir, et al.
Pubblicazione: (2024)
di: Arkhipkin, Vladimir, et al.
Pubblicazione: (2024)
Grounding is All You Need? Dual Temporal Grounding for Video Dialog
di: Qin, You, et al.
Pubblicazione: (2024)
di: Qin, You, et al.
Pubblicazione: (2024)
ENCLIP: Ensembling and Clustering-Based Contrastive Language-Image Pretraining for Fashion Multimodal Search with Limited Data and Low-Quality Images
di: Naik, Prithviraj Purushottam, et al.
Pubblicazione: (2024)
di: Naik, Prithviraj Purushottam, et al.
Pubblicazione: (2024)
Exploiting LMM-based knowledge for image classification tasks
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
di: Zeng, Xiangyu, et al.
Pubblicazione: (2024)
di: Zeng, Xiangyu, et al.
Pubblicazione: (2024)
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation
di: Cai, Minghong, et al.
Pubblicazione: (2024)
di: Cai, Minghong, et al.
Pubblicazione: (2024)
Can Large Language Models Grasp Event Signals? Exploring Pure Zero-Shot Event-based Recognition
di: Yu, Zongyou, et al.
Pubblicazione: (2024)
di: Yu, Zongyou, et al.
Pubblicazione: (2024)
Efficient Low-Resolution Face Recognition via Bridge Distillation
di: Ge, Shiming, et al.
Pubblicazione: (2024)
di: Ge, Shiming, et al.
Pubblicazione: (2024)
Saliency-Based diversity and fairness Metric and FaceKeepOriginalAugment: A Novel Approach for Enhancing Fairness and Diversity
di: Kumar, Teerath, et al.
Pubblicazione: (2024)
di: Kumar, Teerath, et al.
Pubblicazione: (2024)
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
di: Wu, Xiongwei, et al.
Pubblicazione: (2024)
di: Wu, Xiongwei, et al.
Pubblicazione: (2024)
Prompt-Guided Generation of Structured Chest X-Ray Report Using a Pre-trained LLM
di: Li, Hongzhao, et al.
Pubblicazione: (2024)
di: Li, Hongzhao, et al.
Pubblicazione: (2024)
Knowledge-enhanced Multi-perspective Video Representation Learning for Scene Recognition
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
Style-Preserving Lip Sync via Audio-Aware Style Reference
di: Zhong, Weizhi, et al.
Pubblicazione: (2024)
di: Zhong, Weizhi, et al.
Pubblicazione: (2024)
Compressed Deepfake Video Detection Based on 3D Spatiotemporal Trajectories
di: Chen, Zongmei, et al.
Pubblicazione: (2024)
di: Chen, Zongmei, et al.
Pubblicazione: (2024)
MoPE-CLIP: Structured Pruning for Efficient Vision-Language Models with Module-wise Pruning Error Metric
di: Lin, Haokun, et al.
Pubblicazione: (2024)
di: Lin, Haokun, et al.
Pubblicazione: (2024)
UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos
di: Mei, Yuting, et al.
Pubblicazione: (2024)
di: Mei, Yuting, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AI-based System for Transforming text and sound to Educational Videos
di: ElAlami, M. E., et al.
Pubblicazione: (2026) -
Harmful YouTube Video Detection: A Taxonomy of Online Harm and MLLMs as Alternative Annotators
di: Jo, Claire Wonjeong, et al.
Pubblicazione: (2024) -
Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study
di: Bačić, Boris, et al.
Pubblicazione: (2024) -
Can LLMs Create Legally Relevant Summaries and Analyses of Videos?
di: Hoeben-Kuil, Lyra, et al.
Pubblicazione: (2025) -
TraceRouter: Robust Safety for Large Foundation Models via Path-Level Intervention
di: Shi, Chuancheng, et al.
Pubblicazione: (2026)