TT-BLIP: Enhancing Fake News Detection Using BLIP and Tri-Transformer
Fuente:
arXiv
Salvato in:
| Autori principali: | Choi, Eunjee, Kim, Jong-Kook |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
di: Choi, Eunjee, et al.
Pubblicazione: (2025)
di: Choi, Eunjee, et al.
Pubblicazione: (2025)
DriveBLIP2: Attention-Guided Explanation Generation for Complex Driving Scenarios
di: Ling, Shihong, et al.
Pubblicazione: (2025)
di: Ling, Shihong, et al.
Pubblicazione: (2025)
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
di: Limbu, Manshi, et al.
Pubblicazione: (2025)
di: Limbu, Manshi, et al.
Pubblicazione: (2025)
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
di: Geigle, Gregor, et al.
Pubblicazione: (2023)
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
di: Ryoo, Michael S., et al.
Pubblicazione: (2024)
di: Ryoo, Michael S., et al.
Pubblicazione: (2024)
BLIP-FusePPO: A Vision-Language Deep Reinforcement Learning Framework for Lane Keeping in Autonomous Vehicles
di: Miangoleh, Seyed Ahmad Hosseini, et al.
Pubblicazione: (2025)
di: Miangoleh, Seyed Ahmad Hosseini, et al.
Pubblicazione: (2025)
BLIP3o-NEXT: Next Frontier of Native Image Generation
di: Chen, Jiuhai, et al.
Pubblicazione: (2025)
di: Chen, Jiuhai, et al.
Pubblicazione: (2025)
BLIP3-KALE: Knowledge Augmented Large-Scale Dense Captions
di: Awadalla, Anas, et al.
Pubblicazione: (2024)
di: Awadalla, Anas, et al.
Pubblicazione: (2024)
Reframing Image Difference Captioning with BLIP2IDC and Synthetic Augmentation
di: Evennou, Gautier, et al.
Pubblicazione: (2024)
di: Evennou, Gautier, et al.
Pubblicazione: (2024)
HemBLIP: A Vision-Language Model for Interpretable Leukemia Cell Morphology Analysis
di: van Logtestijn, Julie, et al.
Pubblicazione: (2026)
di: van Logtestijn, Julie, et al.
Pubblicazione: (2026)
Edge-Optimized Multimodal Learning for UAV Video Understanding via BLIP-2
di: Feng, Yizhan, et al.
Pubblicazione: (2026)
di: Feng, Yizhan, et al.
Pubblicazione: (2026)
M4-BLIP: Advancing Multi-Modal Media Manipulation Detection through Face-Enhanced Local Analysis
di: Wu, Hang, et al.
Pubblicazione: (2025)
di: Wu, Hang, et al.
Pubblicazione: (2025)
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models
di: Xue, Le, et al.
Pubblicazione: (2024)
di: Xue, Le, et al.
Pubblicazione: (2024)
ContextBLIP: Doubly Contextual Alignment for Contrastive Image Retrieval from Linguistically Complex Descriptions
di: Lin, Honglin, et al.
Pubblicazione: (2024)
di: Lin, Honglin, et al.
Pubblicazione: (2024)
MemeBLIP2: A novel lightweight multimodal system to detect harmful memes
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
Detect Fake with Fake: Leveraging Synthetic Data-driven Representation for Synthetic Image Detection
di: Otake, Hina, et al.
Pubblicazione: (2024)
di: Otake, Hina, et al.
Pubblicazione: (2024)
SatBLIP: Context Understanding and Feature Identification from Satellite Imagery with Vision-Language Learning
di: Wu, Xue, et al.
Pubblicazione: (2026)
di: Wu, Xue, et al.
Pubblicazione: (2026)
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset
di: Chen, Jiuhai, et al.
Pubblicazione: (2025)
di: Chen, Jiuhai, et al.
Pubblicazione: (2025)
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
di: Panagopoulou, Artemis, et al.
Pubblicazione: (2023)
di: Panagopoulou, Artemis, et al.
Pubblicazione: (2023)
Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
di: Selvakumar, Anith, et al.
Pubblicazione: (2025)
di: Selvakumar, Anith, et al.
Pubblicazione: (2025)
Automated Detection of Multiple Sclerosis Lesions on 7-tesla MRI Using U-net and Transformer-based Segmentation
di: Maynord, Michael, et al.
Pubblicazione: (2026)
di: Maynord, Michael, et al.
Pubblicazione: (2026)
X2CT-CLIP: Enable Multi-Abnormality Detection in Computed Tomography from Chest Radiography via Tri-Modal Contrastive Learning
di: You, Jianzhong, et al.
Pubblicazione: (2025)
di: You, Jianzhong, et al.
Pubblicazione: (2025)
RealStats: A Rigorous Real-Only Statistical Framework for Fake Image Detection
di: Zisman, Haim, et al.
Pubblicazione: (2026)
di: Zisman, Haim, et al.
Pubblicazione: (2026)
JARViS: Detecting Actions in Video Using Unified Actor-Scene Context Relation Modeling
di: Lee, Seok Hwan, et al.
Pubblicazione: (2024)
di: Lee, Seok Hwan, et al.
Pubblicazione: (2024)
Abn-BLIP: Abnormality-aligned Bootstrapping Language-Image Pre-training for Pulmonary Embolism Diagnosis and Report Generation from CTPA
di: Zhong, Zhusi, et al.
Pubblicazione: (2025)
di: Zhong, Zhusi, et al.
Pubblicazione: (2025)
TagFog: Textual Anchor Guidance and Fake Outlier Generation for Visual Out-of-Distribution Detection
di: Chen, Jiankang, et al.
Pubblicazione: (2024)
di: Chen, Jiankang, et al.
Pubblicazione: (2024)
Two-Step Data Augmentation for Masked Face Detection and Recognition: Turning Fake Masks to Real
di: Yang, Yan, et al.
Pubblicazione: (2025)
di: Yang, Yan, et al.
Pubblicazione: (2025)
Enhancing Social Media Post Popularity Prediction with Visual Content
di: Jeong, Dahyun, et al.
Pubblicazione: (2024)
di: Jeong, Dahyun, et al.
Pubblicazione: (2024)
Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
di: Lee, Seungyong, et al.
Pubblicazione: (2025)
di: Lee, Seungyong, et al.
Pubblicazione: (2025)
An Explainable Transformer Model for Alzheimer's Disease Detection Using Retinal Imaging
di: Jamshidiha, Saeed, et al.
Pubblicazione: (2025)
di: Jamshidiha, Saeed, et al.
Pubblicazione: (2025)
Rethinking the Use of Vision Transformers for AI-Generated Image Detection
di: Park, NaHyeon, et al.
Pubblicazione: (2025)
di: Park, NaHyeon, et al.
Pubblicazione: (2025)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
di: Choi, Kanghyun, et al.
Pubblicazione: (2024)
di: Choi, Kanghyun, et al.
Pubblicazione: (2024)
MUST: Modality-Specific Representation-Aware Transformer for Diffusion-Enhanced Survival Prediction with Missing Modality
di: Kim, Kyungwon, et al.
Pubblicazione: (2026)
di: Kim, Kyungwon, et al.
Pubblicazione: (2026)
VTON-IT: Virtual Try-On using Image Translation
di: Adhikari, Santosh, et al.
Pubblicazione: (2023)
di: Adhikari, Santosh, et al.
Pubblicazione: (2023)
EHCTNet: Enhanced Hybrid of CNN and Transformer Network for Remote Sensing Image Change Detection
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
Enhancing Object Detection Accuracy in Autonomous Vehicles Using Synthetic Data
di: Voronin, Sergei, et al.
Pubblicazione: (2024)
di: Voronin, Sergei, et al.
Pubblicazione: (2024)
ContextMRI: Enhancing Compressed Sensing MRI through Metadata Conditioning
di: Chung, Hyungjin, et al.
Pubblicazione: (2025)
di: Chung, Hyungjin, et al.
Pubblicazione: (2025)
Dual Precision Deep Neural Network
di: Park, Jae Hyun, et al.
Pubblicazione: (2020)
di: Park, Jae Hyun, et al.
Pubblicazione: (2020)
Sequence Length Scaling in Vision Transformers for Scientific Images on Frontier
di: Tsaris, Aristeidis, et al.
Pubblicazione: (2024)
di: Tsaris, Aristeidis, et al.
Pubblicazione: (2024)
Fake or JPEG? Revealing Common Biases in Generated Image Detection Datasets
di: Grommelt, Patrick, et al.
Pubblicazione: (2024)
di: Grommelt, Patrick, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
di: Choi, Eunjee, et al.
Pubblicazione: (2025) -
DriveBLIP2: Attention-Guided Explanation Generation for Complex Driving Scenarios
di: Ling, Shihong, et al.
Pubblicazione: (2025) -
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
di: Limbu, Manshi, et al.
Pubblicazione: (2025) -
mBLIP: Efficient Bootstrapping of Multilingual Vision-LLMs
di: Geigle, Gregor, et al.
Pubblicazione: (2023) -
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
di: Ryoo, Michael S., et al.
Pubblicazione: (2024)