TT-BLIP: Enhancing Fake News Detection Using BLIP and Tri-Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Choi, Eunjee, Kim, Jong-Kook |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
von: Choi, Eunjee, et al.
Veröffentlicht: (2025)
von: Choi, Eunjee, et al.
Veröffentlicht: (2025)
DriveBLIP2: Attention-Guided Explanation Generation for Complex Driving Scenarios
von: Ling, Shihong, et al.
Veröffentlicht: (2025)
von: Ling, Shihong, et al.
Veröffentlicht: (2025)
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
von: Limbu, Manshi, et al.
Veröffentlicht: (2025)
von: Limbu, Manshi, et al.
Veröffentlicht: (2025)
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
von: Ryoo, Michael S., et al.
Veröffentlicht: (2024)
von: Ryoo, Michael S., et al.
Veröffentlicht: (2024)
BLIP-FusePPO: A Vision-Language Deep Reinforcement Learning Framework for Lane Keeping in Autonomous Vehicles
von: Miangoleh, Seyed Ahmad Hosseini, et al.
Veröffentlicht: (2025)
von: Miangoleh, Seyed Ahmad Hosseini, et al.
Veröffentlicht: (2025)
BLIP3o-NEXT: Next Frontier of Native Image Generation
von: Chen, Jiuhai, et al.
Veröffentlicht: (2025)
von: Chen, Jiuhai, et al.
Veröffentlicht: (2025)
BLIP3-KALE: Knowledge Augmented Large-Scale Dense Captions
von: Awadalla, Anas, et al.
Veröffentlicht: (2024)
von: Awadalla, Anas, et al.
Veröffentlicht: (2024)
Reframing Image Difference Captioning with BLIP2IDC and Synthetic Augmentation
von: Evennou, Gautier, et al.
Veröffentlicht: (2024)
von: Evennou, Gautier, et al.
Veröffentlicht: (2024)
HemBLIP: A Vision-Language Model for Interpretable Leukemia Cell Morphology Analysis
von: van Logtestijn, Julie, et al.
Veröffentlicht: (2026)
von: van Logtestijn, Julie, et al.
Veröffentlicht: (2026)
Edge-Optimized Multimodal Learning for UAV Video Understanding via BLIP-2
von: Feng, Yizhan, et al.
Veröffentlicht: (2026)
von: Feng, Yizhan, et al.
Veröffentlicht: (2026)
M4-BLIP: Advancing Multi-Modal Media Manipulation Detection through Face-Enhanced Local Analysis
von: Wu, Hang, et al.
Veröffentlicht: (2025)
von: Wu, Hang, et al.
Veröffentlicht: (2025)
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models
von: Xue, Le, et al.
Veröffentlicht: (2024)
von: Xue, Le, et al.
Veröffentlicht: (2024)
ContextBLIP: Doubly Contextual Alignment for Contrastive Image Retrieval from Linguistically Complex Descriptions
von: Lin, Honglin, et al.
Veröffentlicht: (2024)
von: Lin, Honglin, et al.
Veröffentlicht: (2024)
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2025)
Detect Fake with Fake: Leveraging Synthetic Data-driven Representation for Synthetic Image Detection
von: Otake, Hina, et al.
Veröffentlicht: (2024)
von: Otake, Hina, et al.
Veröffentlicht: (2024)
SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
von: Wu, Xue, et al.
Veröffentlicht: (2026)
von: Wu, Xue, et al.
Veröffentlicht: (2026)
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
von: Chen, Jiuhai, et al.
Veröffentlicht: (2025)
von: Chen, Jiuhai, et al.
Veröffentlicht: (2025)
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
von: Panagopoulou, Artemis, et al.
Veröffentlicht: (2023)
von: Panagopoulou, Artemis, et al.
Veröffentlicht: (2023)
Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
von: Selvakumar, Anith, et al.
Veröffentlicht: (2025)
von: Selvakumar, Anith, et al.
Veröffentlicht: (2025)
Automated Detection of Multiple Sclerosis Lesions on 7-tesla MRI Using U-net and Transformer-based Segmentation
von: Maynord, Michael, et al.
Veröffentlicht: (2026)
von: Maynord, Michael, et al.
Veröffentlicht: (2026)
X2CT-CLIP: Enable Multi-Abnormality Detection in Computed Tomography from Chest Radiography via Tri-Modal Contrastive Learning
von: You, Jianzhong, et al.
Veröffentlicht: (2025)
von: You, Jianzhong, et al.
Veröffentlicht: (2025)
RealStats: A Rigorous Real-Only Statistical Framework for Fake Image Detection
von: Zisman, Haim, et al.
Veröffentlicht: (2026)
von: Zisman, Haim, et al.
Veröffentlicht: (2026)
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
von: Lee, Seok Hwan, et al.
Veröffentlicht: (2024)
von: Lee, Seok Hwan, et al.
Veröffentlicht: (2024)
Abn-BLIP: Abnormality-aligned Bootstrapping Language-Image Pre-training for Pulmonary Embolism Diagnosis and Report Generation from CTPA
von: Zhong, Zhusi, et al.
Veröffentlicht: (2025)
von: Zhong, Zhusi, et al.
Veröffentlicht: (2025)
TagFog: Textual Anchor Guidance and Fake Outlier Generation for Visual Out-of-Distribution Detection
von: Chen, Jiankang, et al.
Veröffentlicht: (2024)
von: Chen, Jiankang, et al.
Veröffentlicht: (2024)
Two-Step Data Augmentation for Masked Face Detection and Recognition: Turning Fake Masks to Real
von: Yang, Yan, et al.
Veröffentlicht: (2025)
von: Yang, Yan, et al.
Veröffentlicht: (2025)
Enhancing Social Media Post Popularity Prediction with Visual Content
von: Jeong, Dahyun, et al.
Veröffentlicht: (2024)
von: Jeong, Dahyun, et al.
Veröffentlicht: (2024)
Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
von: Lee, Seungyong, et al.
Veröffentlicht: (2025)
von: Lee, Seungyong, et al.
Veröffentlicht: (2025)
An Explainable Transformer Model for Alzheimer's Disease Detection Using Retinal Imaging
von: Jamshidiha, Saeed, et al.
Veröffentlicht: (2025)
von: Jamshidiha, Saeed, et al.
Veröffentlicht: (2025)
Rethinking the Use of Vision Transformers for AI-Generated Image Detection
von: Park, NaHyeon, et al.
Veröffentlicht: (2025)
von: Park, NaHyeon, et al.
Veröffentlicht: (2025)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
von: Choi, Kanghyun, et al.
Veröffentlicht: (2024)
von: Choi, Kanghyun, et al.
Veröffentlicht: (2024)
MUST: Modality-Specific Representation-Aware Transformer for Diffusion-Enhanced Survival Prediction with Missing Modality
von: Kim, Kyungwon, et al.
Veröffentlicht: (2026)
von: Kim, Kyungwon, et al.
Veröffentlicht: (2026)
VTON-IT: Virtual Try-On using Image Translation
von: Adhikari, Santosh, et al.
Veröffentlicht: (2023)
von: Adhikari, Santosh, et al.
Veröffentlicht: (2023)
EHCTNet: Enhanced Hybrid of CNN and Transformer Network for Remote Sensing Image Change Detection
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
Enhancing Object Detection Accuracy in Autonomous Vehicles Using Synthetic Data
von: Voronin, Sergei, et al.
Veröffentlicht: (2024)
von: Voronin, Sergei, et al.
Veröffentlicht: (2024)
ContextMRI: Enhancing Compressed Sensing MRI through Metadata Conditioning
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
Dual Precision Deep Neural Network
von: Park, Jae Hyun, et al.
Veröffentlicht: (2020)
von: Park, Jae Hyun, et al.
Veröffentlicht: (2020)
Sequence Length Scaling in Vision Transformers for Scientific Images on Frontier
von: Tsaris, Aristeidis, et al.
Veröffentlicht: (2024)
von: Tsaris, Aristeidis, et al.
Veröffentlicht: (2024)
Fake or JPEG? Revealing Common Biases in Generated Image Detection Datasets
von: Grommelt, Patrick, et al.
Veröffentlicht: (2024)
von: Grommelt, Patrick, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
von: Choi, Eunjee, et al.
Veröffentlicht: (2025) -
DriveBLIP2: Attention-Guided Explanation Generation for Complex Driving Scenarios
von: Ling, Shihong, et al.
Veröffentlicht: (2025) -
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
von: Limbu, Manshi, et al.
Veröffentlicht: (2025) -
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
von: Geigle, Gregor, et al.
Veröffentlicht: (2023) -
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
von: Ryoo, Michael S., et al.
Veröffentlicht: (2024)