Can a Second-View Image Be a Language? Geometric and Semantic Cross-Modal Reasoning for X-ray Prohibited Item Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Chuang, Tao, Renshuai, Ren, Zhongwei, Liu, Xianglong, Wei, Yunchao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dual-view X-ray Detection: Can AI Detect Prohibited Items from Dual-view X-ray Images like Humans?
by: Tao, Renshuai, et al.
Published: (2024)
by: Tao, Renshuai, et al.
Published: (2024)
BGM: Background Mixup for X-ray Prohibited Items Detection
by: Liu, Weizhe, et al.
Published: (2024)
by: Liu, Weizhe, et al.
Published: (2024)
PAD-F: Prior-Aware Debiasing Framework for Long-Tailed X-ray Prohibited Item Detection
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
Taming Generative Synthetic Data for X-ray Prohibited Item Detection
by: Sun, Jialong, et al.
Published: (2025)
by: Sun, Jialong, et al.
Published: (2025)
Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
by: Tan, Chuangchuang, et al.
Published: (2025)
by: Tan, Chuangchuang, et al.
Published: (2025)
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
by: Tan, Chuangchuang, et al.
Published: (2025)
by: Tan, Chuangchuang, et al.
Published: (2025)
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP
by: Lin, Shuyang, et al.
Published: (2024)
by: Lin, Shuyang, et al.
Published: (2024)
AO-DETR: Anti-Overlapping DETR for X-Ray Prohibited Items Detection
by: Li, Mingyuan, et al.
Published: (2024)
by: Li, Mingyuan, et al.
Published: (2024)
Prohibited Items Segmentation via Occlusion-aware Bilayer Modeling
by: Ren, Yunhan, et al.
Published: (2025)
by: Ren, Yunhan, et al.
Published: (2025)
Behavior Backdoor for Deep Learning Models
by: Wang, Jiakai, et al.
Published: (2024)
by: Wang, Jiakai, et al.
Published: (2024)
From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation
by: Peng, Chuang, et al.
Published: (2026)
by: Peng, Chuang, et al.
Published: (2026)
Augmentation Matters: A Mix-Paste Method for X-Ray Prohibited Item Detection under Noisy Annotations
by: Chen, Ruikang, et al.
Published: (2025)
by: Chen, Ruikang, et al.
Published: (2025)
CSPCL: Category Semantic Prior Contrastive Learning for Deformable DETR-Based Prohibited Item Detectors
by: Li, Mingyuan, et al.
Published: (2025)
by: Li, Mingyuan, et al.
Published: (2025)
C2P-CLIP: Injecting Category Common Prompt in CLIP to Enhance Generalization in Deepfake Detection
by: Tan, Chuangchuang, et al.
Published: (2024)
by: Tan, Chuangchuang, et al.
Published: (2024)
BCR-Net: Boundary-Category Refinement Network for Weakly Semi-Supervised X-Ray Prohibited Item Detection with Points
by: Wong, Sanjoeng
Published: (2024)
by: Wong, Sanjoeng
Published: (2024)
CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer
by: Nie, Wenbo, et al.
Published: (2026)
by: Nie, Wenbo, et al.
Published: (2026)
I$^2$OL-Net: Intra-Inter Objectness Learning Network for Point-Supervised X-Ray Prohibited Item Detection
by: Wong, Sanjoeng, et al.
Published: (2024)
by: Wong, Sanjoeng, et al.
Published: (2024)
Pay Less Attention to Deceptive Artifacts: Robust Detection of Compressed Deepfakes on Online Social Networks
by: Li, Manyi, et al.
Published: (2025)
by: Li, Manyi, et al.
Published: (2025)
PixelLM: Pixel Reasoning with Large Multimodal Model
by: Ren, Zhongwei, et al.
Published: (2023)
by: Ren, Zhongwei, et al.
Published: (2023)
AlignGen: Boosting Personalized Image Generation with Cross-Modality Prior Alignment
by: Lin, Yiheng, et al.
Published: (2025)
by: Lin, Yiheng, et al.
Published: (2025)
Eye Movements, Item Modality, and Multimodal Second Language Vocabulary Learning: Processing and Outcomes
by: Jonathan Malone, et al.
Published: (2025)
by: Jonathan Malone, et al.
Published: (2025)
LLMs Can Evolve Continually on Modality for X-Modal Reasoning
by: Yu, Jiazuo, et al.
Published: (2024)
by: Yu, Jiazuo, et al.
Published: (2024)
Evaluating Cross-Modal Reasoning Ability and Problem Characteristics with Multimodal Item Response Theory
by: Uebayashi, Shunki, et al.
Published: (2026)
by: Uebayashi, Shunki, et al.
Published: (2026)
Self-Enhanced Image Clustering with Cross-Modal Semantic Consistency
by: Li, Zihan, et al.
Published: (2025)
by: Li, Zihan, et al.
Published: (2025)
IPSeg: Image Posterior Mitigates Semantic Drift in Class-Incremental Segmentation
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
X-VILA: Cross-Modality Alignment for Large Language Model
by: Ye, Hanrong, et al.
Published: (2024)
by: Ye, Hanrong, et al.
Published: (2024)
Prohibition and Percolation: The Roaring Success of Coffee During US Alcohol Prohibition
by: Zachary Bartsch
Published: (2025)
by: Zachary Bartsch
Published: (2025)
Cross-Modal Scene Semantic Alignment for Image Complexity Assessment
by: Luo, Yuqing, et al.
Published: (2025)
by: Luo, Yuqing, et al.
Published: (2025)
Scaling Up AI-Generated Image Detection with Generator-Aware Prototypes
by: Qin, Ziheng, et al.
Published: (2025)
by: Qin, Ziheng, et al.
Published: (2025)
Semantic-Guided Natural Language and Visual Fusion for Cross-Modal Interaction Based on Tiny Object Detection
by: Huang, Xian-Hong, et al.
Published: (2025)
by: Huang, Xian-Hong, et al.
Published: (2025)
Unsupervised Region-Based Image Editing of Denoising Diffusion Models
by: Li, Zixiang, et al.
Published: (2024)
by: Li, Zixiang, et al.
Published: (2024)
Object Detection as an Optional Basis: A Graph Matching Network for Cross-View UAV Localization
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
Semantic-Enhanced Feature Matching with Learnable Geometric Verification for Cross-Modal Neuron Registration
by: Li, Wenwei, et al.
Published: (2025)
by: Li, Wenwei, et al.
Published: (2025)
Food Taboos and Biblical Prohibitions
Published: (2020)
Published: (2020)
Vision-fused Attack: Advancing Aggressive and Stealthy Adversarial Text against Neural Machine Translation
by: Xue, Yanni, et al.
Published: (2024)
by: Xue, Yanni, et al.
Published: (2024)
VideoWorld: Exploring Knowledge Learning from Unlabeled Videos
by: Ren, Zhongwei, et al.
Published: (2025)
by: Ren, Zhongwei, et al.
Published: (2025)
Leveraging Failed Samples: A Few-Shot and Training-Free Framework for Generalized Deepfake Detection
by: Yao, Shibo, et al.
Published: (2025)
by: Yao, Shibo, et al.
Published: (2025)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
by: Zhao, Shifang, et al.
Published: (2025)
by: Zhao, Shifang, et al.
Published: (2025)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Similar Items
-
Dual-view X-ray Detection: Can AI Detect Prohibited Items from Dual-view X-ray Images like Humans?
by: Tao, Renshuai, et al.
Published: (2024) -
BGM: Background Mixup for X-ray Prohibited Items Detection
by: Liu, Weizhe, et al.
Published: (2024) -
PAD-F: Prior-Aware Debiasing Framework for Long-Tailed X-ray Prohibited Item Detection
by: Wang, Haoyu, et al.
Published: (2024) -
Taming Generative Synthetic Data for X-ray Prohibited Item Detection
by: Sun, Jialong, et al.
Published: (2025) -
Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
by: Tan, Chuangchuang, et al.
Published: (2025)