Object-oriented backdoor attack against image captioning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Meiling, Zhong, Nan, Zhang, Xinpeng, Qian, Zhenxing, Li, Sheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
by: Li, Meiling, et al.
Published: (2024)
by: Li, Meiling, et al.
Published: (2024)
PatchCraft: Exploring Texture Patch for Efficient AI-generated Image Detection
by: Zhong, Nan, et al.
Published: (2023)
by: Zhong, Nan, et al.
Published: (2023)
Cognitive resilience: Unraveling the proficiency of image-captioning models to interpret masked visual content
by: Du, Zhicheng, et al.
Published: (2024)
by: Du, Zhicheng, et al.
Published: (2024)
Partial train and isolate, mitigate backdoor attack
by: Li, Yong, et al.
Published: (2024)
by: Li, Yong, et al.
Published: (2024)
Multi-Modal interpretable automatic video captioning
by: Hanna-Asaad, Antoine, et al.
Published: (2024)
by: Hanna-Asaad, Antoine, et al.
Published: (2024)
An Exceptional Dataset For Rare Pancreatic Tumor Segmentation
by: Li, Wenqi, et al.
Published: (2025)
by: Li, Wenqi, et al.
Published: (2025)
Evaluating authenticity and quality of image captions via sentiment and semantic analyses
by: Krotov, Aleksei, et al.
Published: (2024)
by: Krotov, Aleksei, et al.
Published: (2024)
Attention-based transformer models for image captioning across languages: An in-depth survey and evaluation
by: Albadarneh, Israa A., et al.
Published: (2025)
by: Albadarneh, Israa A., et al.
Published: (2025)
Image captioning for Brazilian Portuguese using GRIT model
by: de Alencar, Rafael Silva, et al.
Published: (2024)
by: de Alencar, Rafael Silva, et al.
Published: (2024)
Revocable Backdoor for Deep Model Trading
by: Xu, Yiran, et al.
Published: (2024)
by: Xu, Yiran, et al.
Published: (2024)
Are handcrafted filters helpful for attributing AI-generated images?
by: Li, Jialiang, et al.
Published: (2024)
by: Li, Jialiang, et al.
Published: (2024)
Cover-separable Fixed Neural Network Steganography via Deep Generative Models
by: Li, Guobiao, et al.
Published: (2024)
by: Li, Guobiao, et al.
Published: (2024)
Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data
by: Zhang, Chenhui, et al.
Published: (2024)
by: Zhang, Chenhui, et al.
Published: (2024)
PBCAT: Patch-based composite adversarial training against physically realizable attacks on object detection
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
Multi-style conversion for semantic segmentation of lesions in fundus images by adversarial attacks
by: Playout, Clément, et al.
Published: (2024)
by: Playout, Clément, et al.
Published: (2024)
Semantic search for 100M+ galaxy images using AI-generated captions
by: Koblischke, Nolan, et al.
Published: (2025)
by: Koblischke, Nolan, et al.
Published: (2025)
Purified and Unified Steganographic Network
by: Li, Guobiao, et al.
Published: (2024)
by: Li, Guobiao, et al.
Published: (2024)
MITracker: Multi-View Integration for Visual Object Tracking
by: Xu, Mengjie, et al.
Published: (2025)
by: Xu, Mengjie, et al.
Published: (2025)
SimuFreeMark: A Noise-Simulation-Free Robust Watermarking Against Image Editing
by: Tang, Yichao, et al.
Published: (2025)
by: Tang, Yichao, et al.
Published: (2025)
Dynamic Object Queries for Transformer-based Incremental Object Detection
by: Zhang, Jichuan, et al.
Published: (2024)
by: Zhang, Jichuan, et al.
Published: (2024)
Dynamic Multi-Target Fusion for Efficient Audio-Visual Navigation
by: Yu, Yinfeng, et al.
Published: (2025)
by: Yu, Yinfeng, et al.
Published: (2025)
FNBench: Benchmarking Robust Federated Learning against Noisy Labels
by: Jiang, Xuefeng, et al.
Published: (2025)
by: Jiang, Xuefeng, et al.
Published: (2025)
Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
by: Liu, Xiaoyang, et al.
Published: (2024)
by: Liu, Xiaoyang, et al.
Published: (2024)
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting
by: Ding, Jiayu, et al.
Published: (2025)
by: Ding, Jiayu, et al.
Published: (2025)
PROMPT-IML: Image Manipulation Localization with Pre-trained Foundation Models Through Prompt Tuning
by: Liu, Xuntao, et al.
Published: (2024)
by: Liu, Xuntao, et al.
Published: (2024)
Exploring Depth Information for Detecting Manipulated Face Videos
by: Wang, Haoyue, et al.
Published: (2024)
by: Wang, Haoyue, et al.
Published: (2024)
Highly Efficient and Unsupervised Framework for Moving Object Detection in Satellite Videos
by: Xiao, C., et al.
Published: (2024)
by: Xiao, C., et al.
Published: (2024)
PAM: A Propagation-Based Model for Segmenting Any 3D Objects across Multi-Modal Medical Images
by: Chen, Zifan, et al.
Published: (2024)
by: Chen, Zifan, et al.
Published: (2024)
SlowBA: An efficiency backdoor attack towards VLM-based GUI agents
by: Li, Junxian, et al.
Published: (2026)
by: Li, Junxian, et al.
Published: (2026)
TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2026)
by: Ming, Zhenxing, et al.
Published: (2026)
Strategic Preys Make Acute Predators: Enhancing Camouflaged Object Detectors by Generating Camouflaged Objects
by: He, Chunming, et al.
Published: (2023)
by: He, Chunming, et al.
Published: (2023)
Self-Supervised AI-Generated Image Detection: A Camera Metadata Perspective
by: Zhong, Nan, et al.
Published: (2025)
by: Zhong, Nan, et al.
Published: (2025)
Revisiting Physically Realizable Adversarial Object Attack against LiDAR-based Detection: Clarifying Problem Formulation and Experimental Protocols
by: Cheng, Luo, et al.
Published: (2025)
by: Cheng, Luo, et al.
Published: (2025)
DualTeacher: Bridging Coexistence of Unlabelled Classes for Semi-supervised Incremental Object Detection
by: Yuan, Ziqi, et al.
Published: (2023)
by: Yuan, Ziqi, et al.
Published: (2023)
PathReasoning: A multimodal reasoning agent for query-based ROI navigation on whole-slide images
by: Zhang, Kunpeng, et al.
Published: (2025)
by: Zhang, Kunpeng, et al.
Published: (2025)
MaS-VQA: A Mask-and-Select Framework for Knowledge-Based Visual Question Answering
by: Mao, Xianwei, et al.
Published: (2026)
by: Mao, Xianwei, et al.
Published: (2026)
Shadow defense against gradient inversion attack in federated learning
by: Jiang, Le, et al.
Published: (2025)
by: Jiang, Le, et al.
Published: (2025)
GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects
by: Li, Shujia, et al.
Published: (2025)
by: Li, Shujia, et al.
Published: (2025)
Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding
by: Luo, Bingjun, et al.
Published: (2026)
by: Luo, Bingjun, et al.
Published: (2026)
ST-SimDiff: Balancing Spatiotemporal Similarity and Difference for Efficient Video Understanding with MLLMs
by: Luo, Bingjun, et al.
Published: (2026)
by: Luo, Bingjun, et al.
Published: (2026)
Similar Items
-
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
by: Li, Meiling, et al.
Published: (2024) -
PatchCraft: Exploring Texture Patch for Efficient AI-generated Image Detection
by: Zhong, Nan, et al.
Published: (2023) -
Cognitive resilience: Unraveling the proficiency of image-captioning models to interpret masked visual content
by: Du, Zhicheng, et al.
Published: (2024) -
Partial train and isolate, mitigate backdoor attack
by: Li, Yong, et al.
Published: (2024) -
Multi-Modal interpretable automatic video captioning
by: Hanna-Asaad, Antoine, et al.
Published: (2024)