Large Language Models and Provenance Metadata for Determining the Relevance of Images and Videos in News Stories
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Peterka, Tomas, Bohacek, Matyas |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Dataset of News Articles with Provenance Metadata for Media Relevance Assessment
par: Peterka, Tomas, et autres
Publié: (2025)
par: Peterka, Tomas, et autres
Publié: (2025)
GenAI Confessions: Black-box Membership Inference for Generative Image Models
par: Bohacek, Matyas, et autres
Publié: (2025)
par: Bohacek, Matyas, et autres
Publié: (2025)
Synthetic Human Action Video Data Generation with Pose Transfer
par: Knapp, Vaclav, et autres
Publié: (2025)
par: Knapp, Vaclav, et autres
Publié: (2025)
Nepotistically Trained Generative-AI Models Collapse
par: Bohacek, Matyas, et autres
Publié: (2023)
par: Bohacek, Matyas, et autres
Publié: (2023)
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
par: Ashqar, Huthaifa I., et autres
Publié: (2024)
par: Ashqar, Huthaifa I., et autres
Publié: (2024)
Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images
par: Fraser, Kathleen C., et autres
Publié: (2024)
par: Fraser, Kathleen C., et autres
Publié: (2024)
MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms
par: Jin, Yiqiao, et autres
Publié: (2024)
par: Jin, Yiqiao, et autres
Publié: (2024)
Can Pose Transfer Models Generate Realistic Human Motion?
par: Knapp, Vaclav, et autres
Publié: (2025)
par: Knapp, Vaclav, et autres
Publié: (2025)
VideoNorms: Benchmarking Cultural Awareness of Video Language Models
par: Varimalla, Nikhil Reddy, et autres
Publié: (2025)
par: Varimalla, Nikhil Reddy, et autres
Publié: (2025)
A Longitudinal Analysis of Racial and Gender Bias in New York Times and Fox News Images and Articles
par: Ibrahim, Hazem, et autres
Publié: (2024)
par: Ibrahim, Hazem, et autres
Publié: (2024)
Human Action CLIPs: Detecting AI-generated Human Motion
par: Bohacek, Matyas, et autres
Publié: (2024)
par: Bohacek, Matyas, et autres
Publié: (2024)
Stable Signer: Hierarchical Sign Language Generative Model
par: Fang, Sen, et autres
Publié: (2025)
par: Fang, Sen, et autres
Publié: (2025)
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
par: Nazi, Zabir Al, et autres
Publié: (2025)
par: Nazi, Zabir Al, et autres
Publié: (2025)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
par: Sathe, Ashutosh, et autres
Publié: (2024)
par: Sathe, Ashutosh, et autres
Publié: (2024)
The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors
par: Lucy, Li, et autres
Publié: (2026)
par: Lucy, Li, et autres
Publié: (2026)
Uncovering Conceptual Blindspots in Generative Image Models Using Sparse Autoencoders
par: Bohacek, Matyas, et autres
Publié: (2025)
par: Bohacek, Matyas, et autres
Publié: (2025)
Describing Differences in Image Sets with Natural Language
par: Dunlap, Lisa, et autres
Publié: (2023)
par: Dunlap, Lisa, et autres
Publié: (2023)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
par: Jha, Akshita, et autres
Publié: (2024)
par: Jha, Akshita, et autres
Publié: (2024)
Semantic and Expressive Variation in Image Captions Across Languages
par: Ye, Andre, et autres
Publié: (2023)
par: Ye, Andre, et autres
Publié: (2023)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
par: Chen, Liuliu, et autres
Publié: (2026)
par: Chen, Liuliu, et autres
Publié: (2026)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
par: Qi, Peng, et autres
Publié: (2024)
par: Qi, Peng, et autres
Publié: (2024)
Video Understanding with Large Language Models: A Survey
par: Tang, Yolo Y., et autres
Publié: (2023)
par: Tang, Yolo Y., et autres
Publié: (2023)
Compliance Rating Scheme: A Data Provenance Framework for Generative AI Datasets
par: Bohacek, Matyas, et autres
Publié: (2025)
par: Bohacek, Matyas, et autres
Publié: (2025)
Auditing Gender Presentation Differences in Text-to-Image Models
par: Zhang, Yanzhe, et autres
Publié: (2023)
par: Zhang, Yanzhe, et autres
Publié: (2023)
Investigating Disability Representations in Text-to-Image Models
par: Tian, Yang, et autres
Publié: (2026)
par: Tian, Yang, et autres
Publié: (2026)
Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models
par: Yoon, Eunseop, et autres
Publié: (2025)
par: Yoon, Eunseop, et autres
Publié: (2025)
Identifying Implicit Social Biases in Vision-Language Models
par: Hamidieh, Kimia, et autres
Publié: (2024)
par: Hamidieh, Kimia, et autres
Publié: (2024)
Vision-Language Models under Cultural and Inclusive Considerations
par: Karamolegkou, Antonia, et autres
Publié: (2024)
par: Karamolegkou, Antonia, et autres
Publié: (2024)
Frame-Voyager: Learning to Query Frames for Video Large Language Models
par: Yu, Sicheng, et autres
Publié: (2024)
par: Yu, Sicheng, et autres
Publié: (2024)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
par: Madasu, Avinash, et autres
Publié: (2025)
par: Madasu, Avinash, et autres
Publié: (2025)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
par: Zhao, Yuming, et autres
Publié: (2026)
par: Zhao, Yuming, et autres
Publié: (2026)
CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models
par: Xia, Peng, et autres
Publié: (2024)
par: Xia, Peng, et autres
Publié: (2024)
Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
par: Li, Yun, et autres
Publié: (2025)
par: Li, Yun, et autres
Publié: (2025)
HieraVid: Hierarchical Token Pruning for Fast Video Large Language Models
par: Guo, Yansong, et autres
Publié: (2026)
par: Guo, Yansong, et autres
Publié: (2026)
Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting
par: Cai, Chen, et autres
Publié: (2024)
par: Cai, Chen, et autres
Publié: (2024)
The Impact of Image Resolution on Biomedical Multimodal Large Language Models
par: Chen, Liangyu, et autres
Publié: (2025)
par: Chen, Liangyu, et autres
Publié: (2025)
Benchmarking Large Language Models for Image Classification of Marine Mammals
par: Qi, Yijiashun, et autres
Publié: (2024)
par: Qi, Yijiashun, et autres
Publié: (2024)
Closing the Gap: Data-Centric Fine-Tuning of Vision Language Models for the Standardized Exam Questions
par: Sert, Egemen, et autres
Publié: (2025)
par: Sert, Egemen, et autres
Publié: (2025)
Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models
par: Li, Changqun, et autres
Publié: (2024)
par: Li, Changqun, et autres
Publié: (2024)
DEEM: Diffusion Models Serve as the Eyes of Large Language Models for Image Perception
par: Luo, Run, et autres
Publié: (2024)
par: Luo, Run, et autres
Publié: (2024)
Documents similaires
-
Dataset of News Articles with Provenance Metadata for Media Relevance Assessment
par: Peterka, Tomas, et autres
Publié: (2025) -
GenAI Confessions: Black-box Membership Inference for Generative Image Models
par: Bohacek, Matyas, et autres
Publié: (2025) -
Synthetic Human Action Video Data Generation with Pose Transfer
par: Knapp, Vaclav, et autres
Publié: (2025) -
Nepotistically Trained Generative-AI Models Collapse
par: Bohacek, Matyas, et autres
Publié: (2023) -
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
par: Ashqar, Huthaifa I., et autres
Publié: (2024)