Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Das, Dabbrata, Yahan, Mahshar, Zaman, Md Tareq, Bayesh, Md Rishadul |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Dual-Layer Image Encryption Framework Using Chaotic AES with Dynamic S-Boxes and Steganographic QR Codes
di: Bayesh, Md Rishadul, et al.
Pubblicazione: (2025)
di: Bayesh, Md Rishadul, et al.
Pubblicazione: (2025)
MobileDenseAttn:A Dual-Stream Architecture for Accurate and Interpretable Brain Tumor Detection
di: Banik, Shudipta, et al.
Pubblicazione: (2025)
di: Banik, Shudipta, et al.
Pubblicazione: (2025)
AI-Powered Deepfake Detection Using CNN and Vision Transformer Architectures
di: Urmi, Sifatullah Sheikh, et al.
Pubblicazione: (2026)
di: Urmi, Sifatullah Sheikh, et al.
Pubblicazione: (2026)
LOOPE: Learnable Optimal Patch Order in Positional Embeddings for Vision Transformers
di: Chowdhury, Md Abtahi Majeed, et al.
Pubblicazione: (2025)
di: Chowdhury, Md Abtahi Majeed, et al.
Pubblicazione: (2025)
An Explainable Vision-Language Model Framework with Adaptive PID-Tversky Loss for Lumbar Spinal Stenosis Diagnosis
di: Sk., Md. Sajeebul Islam, et al.
Pubblicazione: (2026)
di: Sk., Md. Sajeebul Islam, et al.
Pubblicazione: (2026)
Enhanced Encoder-Decoder Architecture for Accurate Monocular Depth Estimation
di: Das, Dabbrata, et al.
Pubblicazione: (2024)
di: Das, Dabbrata, et al.
Pubblicazione: (2024)
Modular Deep Active Learning Framework for Image Annotation: A Technical Report for the Ophthalmo-AI Project
di: Kadir, Md Abdul, et al.
Pubblicazione: (2024)
di: Kadir, Md Abdul, et al.
Pubblicazione: (2024)
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
di: Bhuiyan, Hasnat Jamil, et al.
Pubblicazione: (2024)
di: Bhuiyan, Hasnat Jamil, et al.
Pubblicazione: (2024)
DL$^3$M: A Vision-to-Language Framework for Expert-Level Medical Reasoning through Deep Learning and Large Language Models
di: Hasan, Md. Najib, et al.
Pubblicazione: (2025)
di: Hasan, Md. Najib, et al.
Pubblicazione: (2025)
ViTs are Everywhere: A Comprehensive Study Showcasing Vision Transformers in Different Domain
di: Mia, Md Sohag, et al.
Pubblicazione: (2023)
di: Mia, Md Sohag, et al.
Pubblicazione: (2023)
DANet: Enhancing Small Object Detection through an Efficient Deformable Attention Network
di: Mia, Md Sohag, et al.
Pubblicazione: (2023)
di: Mia, Md Sohag, et al.
Pubblicazione: (2023)
Scene Graph-Guided Generative AI Framework for Synthesizing and Evaluating Industrial Hazard Scenarios
di: Acharjee, Sanjay, et al.
Pubblicazione: (2025)
di: Acharjee, Sanjay, et al.
Pubblicazione: (2025)
WaveFormer: A 3D Transformer with Wavelet-Driven Feature Representation for Efficient Medical Image Segmentation
di: Hasan, Md Mahfuz Al, et al.
Pubblicazione: (2025)
di: Hasan, Md Mahfuz Al, et al.
Pubblicazione: (2025)
The Visual Counter Turing Test (VCT2): A Benchmark for Evaluating AI-Generated Image Detection and the Visual AI Index (VAI)
di: Imanpour, Nasrin, et al.
Pubblicazione: (2024)
di: Imanpour, Nasrin, et al.
Pubblicazione: (2024)
EdgeNAT: Transformer for Efficient Edge Detection
di: Jie, Jinghuai, et al.
Pubblicazione: (2024)
di: Jie, Jinghuai, et al.
Pubblicazione: (2024)
An Explainable Agentic AI Framework for Uncertainty-Aware and Abstention-Enabled Acute Ischemic Stroke Imaging Decisions
di: Islam, Md Rashadul
Pubblicazione: (2026)
di: Islam, Md Rashadul
Pubblicazione: (2026)
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
di: Mia, Shakil, et al.
Pubblicazione: (2026)
di: Mia, Shakil, et al.
Pubblicazione: (2026)
FUSED-Net: Detecting Traffic Signs with Limited Data
di: Rahman, Md. Atiqur, et al.
Pubblicazione: (2024)
di: Rahman, Md. Atiqur, et al.
Pubblicazione: (2024)
ReHARK: Refined Hybrid Adaptive RBF Kernels for Robust One-Shot Vision-Language Adaptation
di: Islam, Md Jahidul
Pubblicazione: (2026)
di: Islam, Md Jahidul
Pubblicazione: (2026)
Refining Text-to-Image Generation: Towards Accurate Training-Free Glyph-Enhanced Image Generation
di: Lakhanpal, Sanyam, et al.
Pubblicazione: (2024)
di: Lakhanpal, Sanyam, et al.
Pubblicazione: (2024)
A Modified VGG19-Based Framework for Accurate and Interpretable Real-Time Bone Fracture Detection
di: Haque, Md. Ehsanul, et al.
Pubblicazione: (2025)
di: Haque, Md. Ehsanul, et al.
Pubblicazione: (2025)
VFM-VLM: Vision Foundation Model and Vision Language Model based Visual Comparison for 3D Pose Estimation
di: Sarowar, Md Selim, et al.
Pubblicazione: (2025)
di: Sarowar, Md Selim, et al.
Pubblicazione: (2025)
Intelligent Systems in Neuroimaging: Pioneering AI Techniques for Brain Tumor Detection
di: Islam, Md. Mohaiminul, et al.
Pubblicazione: (2025)
di: Islam, Md. Mohaiminul, et al.
Pubblicazione: (2025)
RA-CMF: Region-Adaptive Conditional MeanFlow for CT Image Reconstruction
di: Apurba, Md Shifatul Ahsan, et al.
Pubblicazione: (2026)
di: Apurba, Md Shifatul Ahsan, et al.
Pubblicazione: (2026)
FUSE: Unifying Spectral and Semantic Cues for Robust AI-Generated Image Detection
di: Hossain, Md. Zahid, et al.
Pubblicazione: (2025)
di: Hossain, Md. Zahid, et al.
Pubblicazione: (2025)
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models
di: Rahman, Md Ashikur, et al.
Pubblicazione: (2026)
di: Rahman, Md Ashikur, et al.
Pubblicazione: (2026)
VLAgeBench: Benchmarking Large Vision-Language Models for Zero-Shot Human Age Estimation
di: Sajib, Rakib Hossain, et al.
Pubblicazione: (2026)
di: Sajib, Rakib Hossain, et al.
Pubblicazione: (2026)
CountFormer: A Transformer Framework for Learning Visual Repetition and Structure in Class-Agnostic Object Counting
di: Hossain, Md Tanvir, et al.
Pubblicazione: (2025)
di: Hossain, Md Tanvir, et al.
Pubblicazione: (2025)
Generalized Single-Image-Based Morphing Attack Detection Using Deep Representations from Vision Transformer
di: Zhang, Haoyu, et al.
Pubblicazione: (2025)
di: Zhang, Haoyu, et al.
Pubblicazione: (2025)
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
di: Kodavanti, Sravanth, et al.
Pubblicazione: (2026)
di: Kodavanti, Sravanth, et al.
Pubblicazione: (2026)
Intriguing Equivalence Structures of the Embedding Space of Vision Transformers
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
Anomaly Detection Using Computer Vision: A Comparative Analysis of Class Distinction and Performance Metrics
di: Tusher, Md. Barkat Ullah, et al.
Pubblicazione: (2025)
di: Tusher, Md. Barkat Ullah, et al.
Pubblicazione: (2025)
Dynamic Meta-Ensemble Framework for Efficient and Accurate Deep Learning in Plant Leaf Disease Detection on Resource-Constrained Edge Devices
di: Moges, Weloday Fikadu, et al.
Pubblicazione: (2026)
di: Moges, Weloday Fikadu, et al.
Pubblicazione: (2026)
EdgeEar: Efficient and Accurate Ear Recognition for Edge Devices
di: Lendering, Camile, et al.
Pubblicazione: (2025)
di: Lendering, Camile, et al.
Pubblicazione: (2025)
Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning
di: Hu, Rui, et al.
Pubblicazione: (2024)
di: Hu, Rui, et al.
Pubblicazione: (2024)
Enhancing Breast Cancer Detection with Vision Transformers and Graph Neural Networks
di: Cai, Yeming, et al.
Pubblicazione: (2025)
di: Cai, Yeming, et al.
Pubblicazione: (2025)
Hierarchical Vision Transformer Enhanced by Graph Convolutional Network for Image Classification
di: Jiao, Haibin
Pubblicazione: (2026)
di: Jiao, Haibin
Pubblicazione: (2026)
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks
di: Hossain, Md Zarif, et al.
Pubblicazione: (2024)
di: Hossain, Md Zarif, et al.
Pubblicazione: (2024)
Confidence-Guided Diffusion Augmentation for Enhanced Bangla Compound Character Recognition
di: Rayhan, Md. Sultan Al
Pubblicazione: (2026)
di: Rayhan, Md. Sultan Al
Pubblicazione: (2026)
Efficient Partitioning Vision Transformer on Edge Devices for Distributed Inference
di: Liu, Xiang, et al.
Pubblicazione: (2024)
di: Liu, Xiang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Dual-Layer Image Encryption Framework Using Chaotic AES with Dynamic S-Boxes and Steganographic QR Codes
di: Bayesh, Md Rishadul, et al.
Pubblicazione: (2025) -
MobileDenseAttn:A Dual-Stream Architecture for Accurate and Interpretable Brain Tumor Detection
di: Banik, Shudipta, et al.
Pubblicazione: (2025) -
AI-Powered Deepfake Detection Using CNN and Vision Transformer Architectures
di: Urmi, Sifatullah Sheikh, et al.
Pubblicazione: (2026) -
LOOPE: Learnable Optimal Patch Order in Positional Embeddings for Vision Transformers
di: Chowdhury, Md Abtahi Majeed, et al.
Pubblicazione: (2025) -
An Explainable Vision-Language Model Framework with Adaptive PID-Tversky Loss for Lumbar Spinal Stenosis Diagnosis
di: Sk., Md. Sajeebul Islam, et al.
Pubblicazione: (2026)