Context-aware Multimodal AI Reveals Hidden Pathways in Five Centuries of Art Evolution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jin, Lee, Byunghwee, You, Taekho, Yun, Jinhyuk |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating the diversity and stylization of contemporary user generated visual arts in the complexity entropy plane
von: Kim, Seunghwan, et al.
Veröffentlicht: (2024)
von: Kim, Seunghwan, et al.
Veröffentlicht: (2024)
Aligning AI with Public Values: Deliberation and Decision-Making for Governing Multimodal LLMs in Political Video Analysis
von: Sharma, Tanusree, et al.
Veröffentlicht: (2024)
von: Sharma, Tanusree, et al.
Veröffentlicht: (2024)
Social Links vs. Language Barriers: Decoding the Global Spread of Streaming Content
von: Park, Seoyoung, et al.
Veröffentlicht: (2024)
von: Park, Seoyoung, et al.
Veröffentlicht: (2024)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
von: Qi, Peng, et al.
Veröffentlicht: (2024)
von: Qi, Peng, et al.
Veröffentlicht: (2024)
Direction-Flipped Influence Audits Reveal Hidden Structure in Moral Choices of LLMs
von: Blandfort, Phil, et al.
Veröffentlicht: (2026)
von: Blandfort, Phil, et al.
Veröffentlicht: (2026)
Are Multimodal LLMs Ready for Clinical Dermatology? A Real-World Evaluation in Dermatology
von: Jiang, Roy, et al.
Veröffentlicht: (2026)
von: Jiang, Roy, et al.
Veröffentlicht: (2026)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
von: Um, Sung Jin, et al.
Veröffentlicht: (2025)
von: Um, Sung Jin, et al.
Veröffentlicht: (2025)
Map the Flow: Revealing Hidden Pathways of Information in VideoLLMs
von: Kim, Minji, et al.
Veröffentlicht: (2025)
von: Kim, Minji, et al.
Veröffentlicht: (2025)
Decoding Tourist Perception in Historic Urban Quarters with Multimodal Social Media Data: An AI-Based Framework and Evidence from Shanghai
von: Tan, Kaizhen, et al.
Veröffentlicht: (2025)
von: Tan, Kaizhen, et al.
Veröffentlicht: (2025)
Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views
von: Lee, Inseo, et al.
Veröffentlicht: (2026)
von: Lee, Inseo, et al.
Veröffentlicht: (2026)
Bridging the Gap: Doubles Badminton Analysis with Singles-Trained Models
von: Baek, Seungheon, et al.
Veröffentlicht: (2025)
von: Baek, Seungheon, et al.
Veröffentlicht: (2025)
Multimodal Political Bias Identification and Neutralization
von: Bernard, Cedric, et al.
Veröffentlicht: (2025)
von: Bernard, Cedric, et al.
Veröffentlicht: (2025)
Hidden Bias in the Machine: Stereotypes in Text-to-Image Models
von: Porikli, Sedat, et al.
Veröffentlicht: (2025)
von: Porikli, Sedat, et al.
Veröffentlicht: (2025)
Visual Chronicles: Using Multimodal LLMs to Analyze Massive Collections of Images
von: Deng, Boyang, et al.
Veröffentlicht: (2025)
von: Deng, Boyang, et al.
Veröffentlicht: (2025)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
von: Becerra, Alvaro, et al.
Veröffentlicht: (2025)
von: Becerra, Alvaro, et al.
Veröffentlicht: (2025)
Rainbow Noise: Stress-Testing Multimodal Harmful-Meme Detectors on LGBTQ Content
von: Tong, Ran, et al.
Veröffentlicht: (2025)
von: Tong, Ran, et al.
Veröffentlicht: (2025)
From Content to Audience: A Multimodal Annotation Framework for Broadcast Television Analytics
von: Cupini, Paolo, et al.
Veröffentlicht: (2026)
von: Cupini, Paolo, et al.
Veröffentlicht: (2026)
Emergent AI Surveillance: Overlearned Person Re-Identification and Its Mitigation in Law Enforcement Context
von: Nguyen, An Thi, et al.
Veröffentlicht: (2025)
von: Nguyen, An Thi, et al.
Veröffentlicht: (2025)
ECMF: Enhanced Cross-Modal Fusion for Multimodal Emotion Recognition in MER-SEMI Challenge
von: Hu, Juewen, et al.
Veröffentlicht: (2025)
von: Hu, Juewen, et al.
Veröffentlicht: (2025)
Can Multimodal LLMs See Science Instruction? Benchmarking Pedagogical Reasoning in K-12 Classroom Videos
von: Shen, Yixuan, et al.
Veröffentlicht: (2026)
von: Shen, Yixuan, et al.
Veröffentlicht: (2026)
Two Stage Context Learning with Large Language Models for Multimodal Stance Detection on Climate Change
von: Pangtey, Lata, et al.
Veröffentlicht: (2025)
von: Pangtey, Lata, et al.
Veröffentlicht: (2025)
BuildingView: Constructing Urban Building Exteriors Databases with Street View Imagery and Multimodal Large Language Mode
von: Li, Zongrong, et al.
Veröffentlicht: (2024)
von: Li, Zongrong, et al.
Veröffentlicht: (2024)
Intelligent Systems in Neuroimaging: Pioneering AI Techniques for Brain Tumor Detection
von: Islam, Md. Mohaiminul, et al.
Veröffentlicht: (2025)
von: Islam, Md. Mohaiminul, et al.
Veröffentlicht: (2025)
Safer Prompts: Reducing Risks from Memorization in Visual Generative AI
von: Reissinger, Lena, et al.
Veröffentlicht: (2025)
von: Reissinger, Lena, et al.
Veröffentlicht: (2025)
An AI-Enabled Framework Within Reach for Enhancing Healthcare Sustainability and Fairness
von: Huang, Bin, et al.
Veröffentlicht: (2024)
von: Huang, Bin, et al.
Veröffentlicht: (2024)
EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions
von: Sun, Weiyu, et al.
Veröffentlicht: (2026)
von: Sun, Weiyu, et al.
Veröffentlicht: (2026)
Deepfake-Eval-2024: A Multi-Modal In-the-Wild Benchmark of Deepfakes Circulated in 2024
von: Chandra, Nuria Alina, et al.
Veröffentlicht: (2025)
von: Chandra, Nuria Alina, et al.
Veröffentlicht: (2025)
AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario
von: Beneduce, Ciro, et al.
Veröffentlicht: (2025)
von: Beneduce, Ciro, et al.
Veröffentlicht: (2025)
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings
von: Joshi, Harsh
Veröffentlicht: (2024)
von: Joshi, Harsh
Veröffentlicht: (2024)
Climatic & Anthropogenic Hazards to the Nasca World Heritage: Application of Remote Sensing, AI, and Flood Modelling
von: Sakai, Masato, et al.
Veröffentlicht: (2024)
von: Sakai, Masato, et al.
Veröffentlicht: (2024)
Silicon Minds versus Human Hearts: The Wisdom of Crowds Beats the Wisdom of AI in Emotion Recognition
von: Akben, Mustafa, et al.
Veröffentlicht: (2025)
von: Akben, Mustafa, et al.
Veröffentlicht: (2025)
Smiling Women Pitching Down: Auditing Representational and Presentational Gender Biases in Image Generative AI
von: Sun, Luhang, et al.
Veröffentlicht: (2023)
von: Sun, Luhang, et al.
Veröffentlicht: (2023)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
von: Malekzadeh, Milad, et al.
Veröffentlicht: (2025)
von: Malekzadeh, Milad, et al.
Veröffentlicht: (2025)
Multimodal Learning with Augmentation Techniques for Natural Disaster Assessment
von: Urse, Adrian-Dinu, et al.
Veröffentlicht: (2025)
von: Urse, Adrian-Dinu, et al.
Veröffentlicht: (2025)
No One Knows the State of the Art in Geospatial Foundation Models
von: Corley, Isaac, et al.
Veröffentlicht: (2026)
von: Corley, Isaac, et al.
Veröffentlicht: (2026)
AI Art Neural Constellation: Revealing the Collective and Contrastive State of AI-Generated and Human Art
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2024)
von: Khan, Faizan Farooq, et al.
Veröffentlicht: (2024)
RULE: Reliable Multimodal RAG for Factuality in Medical Vision Language Models
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
Egocentric Co-Pilot: Web-Native Smart-Glasses Agents for Assistive Egocentric AI
von: Yang, Sicheng, et al.
Veröffentlicht: (2026)
von: Yang, Sicheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Investigating the diversity and stylization of contemporary user generated visual arts in the complexity entropy plane
von: Kim, Seunghwan, et al.
Veröffentlicht: (2024) -
Aligning AI with Public Values: Deliberation and Decision-Making for Governing Multimodal LLMs in Political Video Analysis
von: Sharma, Tanusree, et al.
Veröffentlicht: (2024) -
Social Links vs. Language Barriers: Decoding the Global Spread of Streaming Content
von: Park, Seoyoung, et al.
Veröffentlicht: (2024) -
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
von: Qi, Peng, et al.
Veröffentlicht: (2024) -
Direction-Flipped Influence Audits Reveal Hidden Structure in Moral Choices of LLMs
von: Blandfort, Phil, et al.
Veröffentlicht: (2026)