TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Abdullah, Ahmed, Ebert, Nikolas, Wasenmüller, Oliver |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
by: Abdullah, Ahmed, et al.
Published: (2024)
by: Abdullah, Ahmed, et al.
Published: (2024)
GenFormer -- Generated Images are All You Need to Improve Robustness of Transformers on Small Datasets
by: Oehri, Sven, et al.
Published: (2024)
by: Oehri, Sven, et al.
Published: (2024)
SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification
by: Musiat, Alexander, et al.
Published: (2026)
by: Musiat, Alexander, et al.
Published: (2026)
PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms
by: Reichardt, Laurenz, et al.
Published: (2026)
by: Reichardt, Laurenz, et al.
Published: (2026)
Classifier Ensemble for Efficient Uncertainty Calibration of Deep Neural Networks for Image Classification
by: Schulze, Michael, et al.
Published: (2025)
by: Schulze, Michael, et al.
Published: (2025)
IonMorphNet: Generalizable Learning of Ion Image Morphologies for Peak Picking in Mass Spectrometry Imaging
by: Weigand, Philipp, et al.
Published: (2026)
by: Weigand, Philipp, et al.
Published: (2026)
D-PLS: Decoupled Semantic Segmentation for 4D-Panoptic-LiDAR-Segmentation
by: Steinhauser, Maik, et al.
Published: (2025)
by: Steinhauser, Maik, et al.
Published: (2025)
Generative Texture Diversification of 3D Pedestrians for Robust Autonomous Driving Perception
by: Bhowmick, Arka, et al.
Published: (2026)
by: Bhowmick, Arka, et al.
Published: (2026)
Spatial self-supervised Peak Learning and correlation-based Evaluation of peak picking in Mass Spectrometry Imaging
by: Weigand, Philipp, et al.
Published: (2026)
by: Weigand, Philipp, et al.
Published: (2026)
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)
by: Zheng, Anlin, et al.
Published: (2026)
TAP-SLF: Parameter-Efficient Adaptation of Vision Foundation Models for Multi-Task Ultrasound Image Analysis
by: Wan, Hui, et al.
Published: (2026)
by: Wan, Hui, et al.
Published: (2026)
RadarPillars: Efficient Object Detection from 4D Radar Point Clouds
by: Musiat, Alexander, et al.
Published: (2024)
by: Musiat, Alexander, et al.
Published: (2024)
Understanding and Improving Training-Free AI-Generated Image Detections with Vision Foundation Models
by: Tsai, Chung-Ting, et al.
Published: (2024)
by: Tsai, Chung-Ting, et al.
Published: (2024)
TAPNext: Tracking Any Point (TAP) as Next Token Prediction
by: Zholus, Artem, et al.
Published: (2025)
by: Zholus, Artem, et al.
Published: (2025)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
by: Zheng, Anlin, et al.
Published: (2025)
by: Zheng, Anlin, et al.
Published: (2025)
Leveraging Semantic Cues from Foundation Vision Models for Enhanced Local Feature Correspondence
by: Cadar, Felipe, et al.
Published: (2024)
by: Cadar, Felipe, et al.
Published: (2024)
All Patches Matter, More Patches Better: Enhance AI-Generated Image Detection via Panoptic Patch Learning
by: Yang, Zheng, et al.
Published: (2025)
by: Yang, Zheng, et al.
Published: (2025)
PatchFlow: Leveraging a Flow-Based Model with Patch Features
by: Zhang, Boxiang, et al.
Published: (2026)
by: Zhang, Boxiang, et al.
Published: (2026)
Text3DAug -- Prompted Instance Augmentation for LiDAR Perception
by: Reichardt, Laurenz, et al.
Published: (2024)
by: Reichardt, Laurenz, et al.
Published: (2024)
PatchCraft: Exploring Texture Patch for Efficient AI-generated Image Detection
by: Zhong, Nan, et al.
Published: (2023)
by: Zhong, Nan, et al.
Published: (2023)
Leveraging Vision-Language Foundation Models to Reveal Hidden Image-Attribute Relationships in Medical Imaging
by: Kumar, Amar, et al.
Published: (2025)
by: Kumar, Amar, et al.
Published: (2025)
Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
Fusion of Foundation and Vision Transformer Model Features for Dermatoscopic Image Classification
by: Mahbod, Amirreza, et al.
Published: (2025)
by: Mahbod, Amirreza, et al.
Published: (2025)
Vision-Language Semantic Aggregation Leveraging Foundation Model for Generalizable Medical Image Segmentation
by: Yu, Wenjun, et al.
Published: (2025)
by: Yu, Wenjun, et al.
Published: (2025)
Leveraging Visual Signals for Robust Token-Level Uncertainty in Vision-Language Generation
by: Hoche, Joseph, et al.
Published: (2026)
by: Hoche, Joseph, et al.
Published: (2026)
Leveraging Medical Foundation Model Features in Graph Neural Network-Based Retrieval of Breast Histopathology Images
by: Saeidi, Nematollah, et al.
Published: (2024)
by: Saeidi, Nematollah, et al.
Published: (2024)
Leveraging Arbitrary Data Sources for AI-Generated Image Detection Without Sacrificing Generalization
by: He, Qinghui, et al.
Published: (2026)
by: He, Qinghui, et al.
Published: (2026)
Handcrafted Feature Fusion for Reliable Detection of AI-Generated Images
by: Nirob, Syed Mehedi Hasan, et al.
Published: (2026)
by: Nirob, Syed Mehedi Hasan, et al.
Published: (2026)
Multi-Feature Fusion Approach for Generative AI Images Detection
by: Sendjasni, Abderrezzaq, et al.
Published: (2026)
by: Sendjasni, Abderrezzaq, et al.
Published: (2026)
Leveraging Vision-Language Models for Improving Domain Generalization in Image Classification
by: Addepalli, Sravanti, et al.
Published: (2023)
by: Addepalli, Sravanti, et al.
Published: (2023)
TAP-CT: 3D Task-Agnostic Pretraining of Computed Tomography Foundation Models
by: Veenboer, Tim, et al.
Published: (2025)
by: Veenboer, Tim, et al.
Published: (2025)
DART: Differentiable Dynamic Adaptive Region Tokenizer for Vision Foundation Models
by: Yin, Shicheng, et al.
Published: (2025)
by: Yin, Shicheng, et al.
Published: (2025)
TAP: A Token-Adaptive Predictor Framework for Training-Free Diffusion Acceleration
by: Zhu, Haowei, et al.
Published: (2026)
by: Zhu, Haowei, et al.
Published: (2026)
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models
by: Samin, Niamul Hassan, et al.
Published: (2026)
by: Samin, Niamul Hassan, et al.
Published: (2026)
Patch-as-Decodable-Token: Towards Unified Multi-Modal Vision Tasks in MLLMs
by: Su, Yongyi, et al.
Published: (2025)
by: Su, Yongyi, et al.
Published: (2025)
Beyond Boundaries: Leveraging Vision Foundation Models for Source-Free Object Detection
by: Yao, Huizai, et al.
Published: (2025)
by: Yao, Huizai, et al.
Published: (2025)
No Tokens Wasted: Leveraging Long Context in Biomedical Vision-Language Models
by: Sun, Min Woo, et al.
Published: (2025)
by: Sun, Min Woo, et al.
Published: (2025)
PeftCD: Leveraging Vision Foundation Models with Parameter-Efficient Fine-Tuning for Remote Sensing Change Detection
by: Dong, Sijun, et al.
Published: (2025)
by: Dong, Sijun, et al.
Published: (2025)
SpectraIrisPAD: Leveraging Vision Foundation Models for Spectrally Conditioned Multispectral Iris Presentation Attack Detection
by: Ramachandra, Raghavendra, et al.
Published: (2025)
by: Ramachandra, Raghavendra, et al.
Published: (2025)
Are Vision Foundation Models Foundational for Electron Microscopy Image Segmentation?
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
Similar Items
-
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
by: Abdullah, Ahmed, et al.
Published: (2024) -
GenFormer -- Generated Images are All You Need to Improve Robustness of Transformers on Small Datasets
by: Oehri, Sven, et al.
Published: (2024) -
SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification
by: Musiat, Alexander, et al.
Published: (2026) -
PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms
by: Reichardt, Laurenz, et al.
Published: (2026) -
Classifier Ensemble for Efficient Uncertainty Calibration of Deep Neural Networks for Image Classification
by: Schulze, Michael, et al.
Published: (2025)