A Multimedia Analytics Model for the Foundation Model Era
Fuente:
arXiv
Saved in:
| Main Authors: | Worring, Marcel, Zahálka, Jan, Elzen, Stef van den, Fischer, Maximilian T., Keim, Daniel A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Rhetorical Relations-Based Framework for Tailored Multimedia Document Summarization
by: Maredj, Azze-Eddine, et al.
Published: (2024)
by: Maredj, Azze-Eddine, et al.
Published: (2024)
InfoCIR: Multimedia Analysis for Composed Image Retrieval
by: Dravilas, Ioannis, et al.
Published: (2026)
by: Dravilas, Ioannis, et al.
Published: (2026)
MULTI-CASE: A Transformer-based Ethics-aware Multimodal Investigative Intelligence Framework
by: Fischer, Maximilian T., et al.
Published: (2024)
by: Fischer, Maximilian T., et al.
Published: (2024)
IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models
by: Rahimi, Hamed, et al.
Published: (2026)
by: Rahimi, Hamed, et al.
Published: (2026)
Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example
by: Zhou, Aven-Le, et al.
Published: (2024)
by: Zhou, Aven-Le, et al.
Published: (2024)
DreamLLM-3D: Affective Dream Reliving using Large Language Model and 3D Generative AI
by: Liu, Pinyao, et al.
Published: (2025)
by: Liu, Pinyao, et al.
Published: (2025)
Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey
by: Kumar, Puneet, et al.
Published: (2024)
by: Kumar, Puneet, et al.
Published: (2024)
Virtual Reality and Artificial Intelligence as Psychological Countermeasures in Space and Other Isolated and Confined Environments: A Scoping Review
by: Sharp, Jennifer, et al.
Published: (2025)
by: Sharp, Jennifer, et al.
Published: (2025)
Focus360: Guiding User Attention in Immersive Videos for VR
by: Silva, Paulo Vitor S., et al.
Published: (2026)
by: Silva, Paulo Vitor S., et al.
Published: (2026)
Digital Simulations to Enhance Military Medical Evacuation Decision-Making
by: Fischer, Jeremy, et al.
Published: (2025)
by: Fischer, Jeremy, et al.
Published: (2025)
Self-supervised Spatio-Temporal Graph Mask-Passing Attention Network for Perceptual Importance Prediction of Multi-point Tactility
by: He, Dazhong, et al.
Published: (2024)
by: He, Dazhong, et al.
Published: (2024)
Archiving Body Movements: Collective Generation of Chinese Calligraphy
by: Zhou, Aven Le, et al.
Published: (2023)
by: Zhou, Aven Le, et al.
Published: (2023)
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
by: Kyaw, Alexander Htet, et al.
Published: (2025)
by: Kyaw, Alexander Htet, et al.
Published: (2025)
Livia: An Emotion-Aware AR Companion Powered by Modular AI Agents and Progressive Memory Compression
by: Xi, Rui, et al.
Published: (2025)
by: Xi, Rui, et al.
Published: (2025)
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis
by: He, Jun-Yan, et al.
Published: (2024)
by: He, Jun-Yan, et al.
Published: (2024)
Through the Looking-Glass: AI-Mediated Video Communication Reduces Interpersonal Trust and Confidence in Judgments
by: Fernández, Nelson Navajas, et al.
Published: (2026)
by: Fernández, Nelson Navajas, et al.
Published: (2026)
An Empirical Evaluation of AI-Powered Non-Player Characters' Perceived Realism and Performance in Virtual Reality Environments
by: Korkiakoski, Mikko, et al.
Published: (2025)
by: Korkiakoski, Mikko, et al.
Published: (2025)
Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video
by: Yeh, Catherine, et al.
Published: (2026)
by: Yeh, Catherine, et al.
Published: (2026)
Simulacra Naturae: Generative Ecosystem driven by Agent-Based Simulations and Brain Organoid Collective Intelligence
by: Manoudaki, Nefeli, et al.
Published: (2025)
by: Manoudaki, Nefeli, et al.
Published: (2025)
AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents
by: Chai, Yuxiang, et al.
Published: (2024)
by: Chai, Yuxiang, et al.
Published: (2024)
Signals of Provenance: Practices & Challenges of Navigating Indicators in AI-Generated Media for Sighted and Blind Individuals
by: Ide, Ayae, et al.
Published: (2025)
by: Ide, Ayae, et al.
Published: (2025)
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
by: Taneja, Karan, et al.
Published: (2025)
by: Taneja, Karan, et al.
Published: (2025)
SimInterview: Transforming Business Education through Large Language Model-Based Simulated Multilingual Interview Training System
by: Nguyen, Truong Thanh Hung, et al.
Published: (2025)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2025)
Human-Data Interaction, Exploration, and Visualization in the AI Era: Challenges and Opportunities
by: Fekete, Jean-Daniel, et al.
Published: (2026)
by: Fekete, Jean-Daniel, et al.
Published: (2026)
Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era
by: Nguyen, Thanh Tam, et al.
Published: (2024)
by: Nguyen, Thanh Tam, et al.
Published: (2024)
One Size Doesn't Fit All: Age-Aware Gamification Mechanics for Multimedia Learning Environments
by: Kaißer, Sarah, et al.
Published: (2025)
by: Kaißer, Sarah, et al.
Published: (2025)
Proceedings of The third international workshop on eXplainable AI for the Arts (XAIxArts)
by: Ford, Corey, et al.
Published: (2025)
by: Ford, Corey, et al.
Published: (2025)
Towards Difficulty-Aware Analysis of Deep Neural Networks
by: Meng, Linhao, et al.
Published: (2025)
by: Meng, Linhao, et al.
Published: (2025)
DailyLLM: Context-Aware Activity Log Generation Using Multi-Modal Sensors and LLMs
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Counterfactual Reasoning Using Predicted Latent Personality Dimensions for Optimizing Persuasion Outcome
by: Zeng, Donghuo, et al.
Published: (2024)
by: Zeng, Donghuo, et al.
Published: (2024)
LAVE: LLM-Powered Agent Assistance and Language Augmentation for Video Editing
by: Wang, Bryan, et al.
Published: (2024)
by: Wang, Bryan, et al.
Published: (2024)
FeedQUAC: Quick Unobtrusive AI-Generated Commentary
by: Long, Tao, et al.
Published: (2025)
by: Long, Tao, et al.
Published: (2025)
MetaDecorator: Generating Immersive Virtual Tours through Multimodality
by: Xie, Shuang, et al.
Published: (2025)
by: Xie, Shuang, et al.
Published: (2025)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
A(I)nimism: Re-enchanting the World Through AI-Mediated Object Interaction
by: Mykhaylychenko, Diana, et al.
Published: (2025)
by: Mykhaylychenko, Diana, et al.
Published: (2025)
SkinGEN: an Explainable Dermatology Diagnosis-to-Generation Framework with Interactive Vision-Language Models
by: Lin, Bo, et al.
Published: (2024)
by: Lin, Bo, et al.
Published: (2024)
Human-Machine Ritual: Synergic Performance through Real-Time Motion Recognition
by: Cai, Zhuodi, et al.
Published: (2025)
by: Cai, Zhuodi, et al.
Published: (2025)
Leveraging LLMs to Create a Haptic Devices' Recommendation System
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Dynamic and Super-Personalized Media Ecosystem Driven by Generative AI: Unpredictable Plays Never Repeating The Same
by: Ahn, Sungjun, et al.
Published: (2024)
by: Ahn, Sungjun, et al.
Published: (2024)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
by: Li, Haobo, et al.
Published: (2024)
by: Li, Haobo, et al.
Published: (2024)
Similar Items
-
A Rhetorical Relations-Based Framework for Tailored Multimedia Document Summarization
by: Maredj, Azze-Eddine, et al.
Published: (2024) -
InfoCIR: Multimedia Analysis for Composed Image Retrieval
by: Dravilas, Ioannis, et al.
Published: (2026) -
MULTI-CASE: A Transformer-based Ethics-aware Multimodal Investigative Intelligence Framework
by: Fischer, Maximilian T., et al.
Published: (2024) -
IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models
by: Rahimi, Hamed, et al.
Published: (2026) -
Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example
by: Zhou, Aven-Le, et al.
Published: (2024)