PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Tuan, Khan, Naseem, Tran, Khang, Phan, NhatHai, Khalil, Issa |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025)
by: ALBarqawi, Ahmad, et al.
Published: (2025)
CapsFake: A Multimodal Capsule Network for Detecting Instruction-Guided Deepfakes
by: Nguyen, Tuan, et al.
Published: (2025)
by: Nguyen, Tuan, et al.
Published: (2025)
CAMME: Adaptive Deepfake Image Detection with Multi-Modal Cross-Attention
by: Khan, Naseem, et al.
Published: (2025)
by: Khan, Naseem, et al.
Published: (2025)
Gradient Transformer: Learning to Generate Updates for LLMs
by: Nguyen, Binh-Nguyen, et al.
Published: (2026)
by: Nguyen, Binh-Nguyen, et al.
Published: (2026)
Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems
by: Khan, Naseem, et al.
Published: (2025)
by: Khan, Naseem, et al.
Published: (2025)
SGFusion: Stochastic Geographic Gradient Fusion in Federated Learning
by: Nguyen, Khoa, et al.
Published: (2025)
by: Nguyen, Khoa, et al.
Published: (2025)
PromSec: Prompt Optimization for Secure Generation of Functional Source Code with Large Language Models (LLMs)
by: Nazzal, Mahmoud, et al.
Published: (2024)
by: Nazzal, Mahmoud, et al.
Published: (2024)
Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics
by: Tran, Khang, et al.
Published: (2026)
by: Tran, Khang, et al.
Published: (2026)
Poison with Style: A Practical Poisoning Attack on Code Large Language Models
by: Tran, Khang, et al.
Published: (2026)
by: Tran, Khang, et al.
Published: (2026)
FairDP: Certified Fairness with Differential Privacy
by: Tran, Khang, et al.
Published: (2023)
by: Tran, Khang, et al.
Published: (2023)
CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection
by: Khan, Sohail Ahmed, et al.
Published: (2024)
by: Khan, Sohail Ahmed, et al.
Published: (2024)
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning
by: Van Vo, Tuan, et al.
Published: (2026)
by: Van Vo, Tuan, et al.
Published: (2026)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Lifelong Whole Slide Image Analysis: Online Vision-Language Adaptation and Past-to-Present Gradient Distillation
by: Bui, Doanh C., et al.
Published: (2025)
by: Bui, Doanh C., et al.
Published: (2025)
Passive Deepfake Detection Across Multi-modalities: A Comprehensive Survey
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
by: Nguyen-Le, Hong-Hanh, et al.
Published: (2024)
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
by: Pham, Hieu Dinh Trung, et al.
Published: (2025)
by: Pham, Hieu Dinh Trung, et al.
Published: (2025)
A Client-level Assessment of Collaborative Backdoor Poisoning in Non-IID Federated Learning
by: Lai, Phung, et al.
Published: (2025)
by: Lai, Phung, et al.
Published: (2025)
Adversarially Robust Deepfake Detection via Adversarial Feature Similarity Learning
by: Khan, Sarwar
Published: (2024)
by: Khan, Sarwar
Published: (2024)
Calibrated and Robust Foundation Models for Vision-Language and Medical Image Tasks Under Distribution Shift
by: Khan, Behraj, et al.
Published: (2025)
by: Khan, Behraj, et al.
Published: (2025)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
SOS: A Shuffle Order Strategy for Data Augmentation in Industrial Human Activity Recognition
by: Ha, Anh Tuan, et al.
Published: (2025)
by: Ha, Anh Tuan, et al.
Published: (2025)
Volumetric Mapping with Panoptic Refinement via Kernel Density Estimation for Mobile Robots
by: Nguyen, Khang, et al.
Published: (2024)
by: Nguyen, Khang, et al.
Published: (2024)
DoRAN: Stabilizing Weight-Decomposed Low-Rank Adaptation via Noise Injection and Auxiliary Networks
by: Diep, Nghiem T., et al.
Published: (2025)
by: Diep, Nghiem T., et al.
Published: (2025)
Unleashing Vision-Language Semantics for Deepfake Video Detection
by: Zhu, Jiawen, et al.
Published: (2026)
by: Zhu, Jiawen, et al.
Published: (2026)
Age-Diverse Deepfake Dataset: Bridging the Age Gap in Deepfake Detection
by: Joshi, Unisha
Published: (2025)
by: Joshi, Unisha
Published: (2025)
WildDeepfake: A Challenging Real-World Dataset for Deepfake Detection
by: Zi, Bojia, et al.
Published: (2021)
by: Zi, Bojia, et al.
Published: (2021)
Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
Distortion-Aware Adversarial Attacks on Bounding Boxes of Object Detectors
by: Phuc, Pham, et al.
Published: (2024)
by: Phuc, Pham, et al.
Published: (2024)
Unveiling Deepfakes: A Frequency-Aware Triple Branch Network for Deepfake Detection
by: Shen, Qihao, et al.
Published: (2026)
by: Shen, Qihao, et al.
Published: (2026)
Energy-Based Sliced Wasserstein Distance
by: Nguyen, Khai, et al.
Published: (2023)
by: Nguyen, Khai, et al.
Published: (2023)
Sliced Wasserstein Estimation with Control Variates
by: Nguyen, Khai, et al.
Published: (2023)
by: Nguyen, Khai, et al.
Published: (2023)
MGPATH: Vision-Language Model with Multi-Granular Prompt Learning for Few-Shot WSI Classification
by: Nguyen, Anh-Tien, et al.
Published: (2025)
by: Nguyen, Anh-Tien, et al.
Published: (2025)
Fair and Interpretable Deepfake Detection in Videos
by: Yoshii, Akihito, et al.
Published: (2025)
by: Yoshii, Akihito, et al.
Published: (2025)
V3D-SLAM: Robust RGB-D SLAM in Dynamic Environments with 3D Semantic Geometry Voting
by: Dang, Tuan, et al.
Published: (2024)
by: Dang, Tuan, et al.
Published: (2024)
GatedLexiconNet: A Comprehensive End-to-End Handwritten Paragraph Text Recognition System
by: Kumari, Lalita, et al.
Published: (2024)
by: Kumari, Lalita, et al.
Published: (2024)
Vision-Language Models Encode Clinical Guidelines for Concept-Based Medical Reasoning
by: Harmanani, Mohamed, et al.
Published: (2026)
by: Harmanani, Mohamed, et al.
Published: (2026)
Stratified Domain Adaptation: A Progressive Self-Training Approach for Scene Text Recognition
by: Le, Kha Nhat, et al.
Published: (2024)
by: Le, Kha Nhat, et al.
Published: (2024)
AI-Powered Deepfake Detection Using CNN and Vision Transformer Architectures
by: Urmi, Sifatullah Sheikh, et al.
Published: (2026)
by: Urmi, Sifatullah Sheikh, et al.
Published: (2026)
A Vision-Language Foundation Model for Leaf Disease Identification
by: Quoc, Khang Nguyen, et al.
Published: (2025)
by: Quoc, Khang Nguyen, et al.
Published: (2025)
Similar Items
-
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025) -
CapsFake: A Multimodal Capsule Network for Detecting Instruction-Guided Deepfakes
by: Nguyen, Tuan, et al.
Published: (2025) -
CAMME: Adaptive Deepfake Image Detection with Multi-Modal Cross-Attention
by: Khan, Naseem, et al.
Published: (2025) -
Gradient Transformer: Learning to Generate Updates for LLMs
by: Nguyen, Binh-Nguyen, et al.
Published: (2026) -
Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems
by: Khan, Naseem, et al.
Published: (2025)