Understanding and Improving Training-Free AI-Generated Image Detections with Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tsai, Chung-Ting, Ko, Ching-Yun, Chung, I-Hsin, Wang, Yu-Chiang Frank, Chen, Pin-Yu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Overload: Latency Attacks on Object Detection for Edge Devices
by: Chen, Erh-Chung, et al.
Published: (2023)
by: Chen, Erh-Chung, et al.
Published: (2023)
Steal Now and Attack Later: Evaluating Robustness of Object Detection against Black-box Adversarial Attacks
by: Chen, Erh-Chung, et al.
Published: (2024)
by: Chen, Erh-Chung, et al.
Published: (2024)
Receler: Reliable Concept Erasing of Text-to-Image Diffusion Models via Lightweight Erasers
by: Huang, Chi-Pin, et al.
Published: (2023)
by: Huang, Chi-Pin, et al.
Published: (2023)
RIGID: A Training-free and Model-Agnostic Framework for Robust AI-Generated Image Detection
by: He, Zhiyuan, et al.
Published: (2024)
by: He, Zhiyuan, et al.
Published: (2024)
PromptHSI: Universal Hyperspectral Image Restoration with Vision-Language Modulated Frequency Adaptation
by: Lee, Chia-Ming, et al.
Published: (2024)
by: Lee, Chia-Ming, et al.
Published: (2024)
OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding
by: Huang, Sheng-Yu, et al.
Published: (2026)
by: Huang, Sheng-Yu, et al.
Published: (2026)
VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection
by: Qiu, Bo-Cheng, et al.
Published: (2026)
by: Qiu, Bo-Cheng, et al.
Published: (2026)
MENTOR: Multilingual tExt detectioN TOward leaRning by analogy
by: Lin, Hsin-Ju, et al.
Published: (2024)
by: Lin, Hsin-Ju, et al.
Published: (2024)
Improving Generalization Ability for 3D Object Detection by Learning Sparsity-invariant Features
by: Lu, Hsin-Cheng, et al.
Published: (2025)
by: Lu, Hsin-Cheng, et al.
Published: (2025)
Histopathology Image Report Generation by Vision Language Model with Multimodal In-Context Learning
by: Liu, Shih-Wen, et al.
Published: (2025)
by: Liu, Shih-Wen, et al.
Published: (2025)
Modular Prompt Learning Improves Vision-Language Models
by: Huang, Zhenhan, et al.
Published: (2025)
by: Huang, Zhenhan, et al.
Published: (2025)
VideoMage: Multi-Subject and Motion Customization of Text-to-Video Diffusion Models
by: Huang, Chi-Pin, et al.
Published: (2025)
by: Huang, Chi-Pin, et al.
Published: (2025)
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)
by: Zheng, Anlin, et al.
Published: (2026)
TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection
by: Abdullah, Ahmed, et al.
Published: (2026)
by: Abdullah, Ahmed, et al.
Published: (2026)
Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking
by: Chung, Sangyun, et al.
Published: (2024)
by: Chung, Sangyun, et al.
Published: (2024)
LithoHoD: A Litho Simulator-Powered Framework for IC Layout Hotspot Detection
by: Shao, Hao-Chiang, et al.
Published: (2024)
by: Shao, Hao-Chiang, et al.
Published: (2024)
ChangeDINO: DINOv3-Driven Building Change Detection in Optical Remote Sensing Imagery
by: Cheng, Ching-Heng, et al.
Published: (2025)
by: Cheng, Ching-Heng, et al.
Published: (2025)
Improving Generalization of Medical Image Registration Foundation Model
by: Hu, Jing, et al.
Published: (2025)
by: Hu, Jing, et al.
Published: (2025)
Safeguarding Generative AI Applications in Preclinical Imaging through Hybrid Anomaly Detection
by: Binda, Jakub, et al.
Published: (2025)
by: Binda, Jakub, et al.
Published: (2025)
Curriculum Fine-tuning of Vision Foundation Model for Medical Image Classification Under Label Noise
by: Yu, Yeonguk, et al.
Published: (2024)
by: Yu, Yeonguk, et al.
Published: (2024)
Image Outlier Detection Without Training using RANSAC
by: Tsai, Chen-Han, et al.
Published: (2023)
by: Tsai, Chen-Han, et al.
Published: (2023)
Select and Distill: Selective Dual-Teacher Knowledge Transfer for Continual Learning on Vision-Language Models
by: Yu, Yu-Chu, et al.
Published: (2024)
by: Yu, Yu-Chu, et al.
Published: (2024)
Reasoning under Vision: Understanding Visual-Spatial Cognition in Vision-Language Models for CAPTCHA
by: Song, Python, et al.
Published: (2025)
by: Song, Python, et al.
Published: (2025)
Team NYCU at Defactify4: Robust Detection and Source Identification of AI-Generated Images Using CNN and CLIP-Based Models
by: Yang, Tsan-Tsung, et al.
Published: (2025)
by: Yang, Tsan-Tsung, et al.
Published: (2025)
RobustVisRAG: Causality-Aware Vision-Based Retrieval-Augmented Generation under Visual Degradations
by: Chen, I-Hsiang, et al.
Published: (2026)
by: Chen, I-Hsiang, et al.
Published: (2026)
GSNeRF: Generalizable Semantic Neural Radiance Fields with Enhanced 3D Scene Understanding
by: Chou, Zi-Ting, et al.
Published: (2024)
by: Chou, Zi-Ting, et al.
Published: (2024)
A Bias-Free Training Paradigm for More General AI-generated Image Detection
by: Guillaro, Fabrizio, et al.
Published: (2024)
by: Guillaro, Fabrizio, et al.
Published: (2024)
Robust and Annotation-Free Wound Segmentation on Noisy Real-World Pressure Ulcer Images: Towards Automated DESIGN-R\textsuperscript{\textregistered} Assessment
by: Tsai, Yun-Cheng
Published: (2025)
by: Tsai, Yun-Cheng
Published: (2025)
Intermediate Representations are Strong AI-Generated Image Detectors
by: Huang, Zhenhan, et al.
Published: (2026)
by: Huang, Zhenhan, et al.
Published: (2026)
SPARK: Multi-Vision Sensor Perception and Reasoning Benchmark for Large-scale Vision-Language Models
by: Yu, Youngjoon, et al.
Published: (2024)
by: Yu, Youngjoon, et al.
Published: (2024)
Adapting Vision Foundation Models for Robust Cloud Segmentation in Remote Sensing Images
by: Zou, Xuechao, et al.
Published: (2024)
by: Zou, Xuechao, et al.
Published: (2024)
A Novel Grouping-Based Hybrid Color Correction Algorithm for Color Point Clouds
by: Chung, Kuo-Liang, et al.
Published: (2025)
by: Chung, Kuo-Liang, et al.
Published: (2025)
Generative Digital Twins: Vision-Language Simulation Models for Executable Industrial Systems
by: Hsu, YuChe, et al.
Published: (2025)
by: Hsu, YuChe, et al.
Published: (2025)
Turns Out I'm Not Real: Towards Robust Detection of AI-Generated Videos
by: Liu, Qingyuan, et al.
Published: (2024)
by: Liu, Qingyuan, et al.
Published: (2024)
Towards Training-free Anomaly Detection with Vision and Language Foundation Models
by: Zhang, Jinjin, et al.
Published: (2025)
by: Zhang, Jinjin, et al.
Published: (2025)
Color Me Correctly: Bridging Perceptual Color Spaces and Text Embeddings for Improved Diffusion Generation
by: Tsai, Sung-Lin, et al.
Published: (2025)
by: Tsai, Sung-Lin, et al.
Published: (2025)
Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation
by: Xu, Zhiyang, et al.
Published: (2025)
by: Xu, Zhiyang, et al.
Published: (2025)
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
by: Zheng, Anlin, et al.
Published: (2025)
by: Zheng, Anlin, et al.
Published: (2025)
LocateBench: Evaluating the Locating Ability of Vision Language Models
by: Chiang, Ting-Rui, et al.
Published: (2024)
by: Chiang, Ting-Rui, et al.
Published: (2024)
Similar Items
-
Overload: Latency Attacks on Object Detection for Edge Devices
by: Chen, Erh-Chung, et al.
Published: (2023) -
Steal Now and Attack Later: Evaluating Robustness of Object Detection against Black-box Adversarial Attacks
by: Chen, Erh-Chung, et al.
Published: (2024) -
Receler: Reliable Concept Erasing of Text-to-Image Diffusion Models via Lightweight Erasers
by: Huang, Chi-Pin, et al.
Published: (2023) -
RIGID: A Training-free and Model-Agnostic Framework for Robust AI-Generated Image Detection
by: He, Zhiyuan, et al.
Published: (2024) -
PromptHSI: Universal Hyperspectral Image Restoration with Vision-Language Modulated Frequency Adaptation
by: Lee, Chia-Ming, et al.
Published: (2024)