One Image is Worth a Thousand Words: A Usability Preservable Text-Image Collaborative Erasing Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Feiran, Xu, Qianqian, Bao, Shilong, Yang, Zhiyong, Cao, Xiaochun, Huang, Qingming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BlackMirror: Black-Box Backdoor Detection for Text-to-Image Models via Instruction-Response Deviation
by: Li, Feiran, et al.
Published: (2026)
by: Li, Feiran, et al.
Published: (2026)
Size-invariance Matters: Rethinking Metrics and Losses for Imbalanced Multi-object Salient Object Detection
by: Li, Feiran, et al.
Published: (2024)
by: Li, Feiran, et al.
Published: (2024)
Hybrid Generative Fusion for Efficient and Privacy-Preserving Face Recognition Dataset Generation
by: Li, Feiran, et al.
Published: (2025)
by: Li, Feiran, et al.
Published: (2025)
Towards Size-invariant Salient Object Detection: A Generic Evaluation and Optimization Approach
by: Bao, Shilong, et al.
Published: (2025)
by: Bao, Shilong, et al.
Published: (2025)
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
by: Luo, Zhi, et al.
Published: (2025)
by: Luo, Zhi, et al.
Published: (2025)
Closing the Approximation Gap of Partial AUC Optimization: A Tale of Two Formulations
by: Jiang, Yangbangyan, et al.
Published: (2025)
by: Jiang, Yangbangyan, et al.
Published: (2025)
Bidirectional Logits Tree: Pursuing Granularity Reconcilement in Fine-Grained Classification
by: Lu, Zhiguang, et al.
Published: (2024)
by: Lu, Zhiguang, et al.
Published: (2024)
Vript: A Video Is Worth Thousands of Words
by: Yang, Dongjie, et al.
Published: (2024)
by: Yang, Dongjie, et al.
Published: (2024)
Understanding-Enhanced Model Collaboration for Long-Tailed Egocentric Mistake Detection
by: Han, Boyu, et al.
Published: (2026)
by: Han, Boyu, et al.
Published: (2026)
Words Worth a Thousand Pictures: Measuring and Understanding Perceptual Variability in Text-to-Image Generation
by: Tang, Raphael, et al.
Published: (2024)
by: Tang, Raphael, et al.
Published: (2024)
LightFair: Towards an Efficient Alternative for Fair T2I Diffusion via Debiasing Pre-trained Text Encoders
by: Han, Boyu, et al.
Published: (2025)
by: Han, Boyu, et al.
Published: (2025)
A Video Is Not Worth a Thousand Words
by: Pollard, Sam, et al.
Published: (2025)
by: Pollard, Sam, et al.
Published: (2025)
ReconBoost: Boosting Can Achieve Modality Reconcilement
by: Hua, Cong, et al.
Published: (2024)
by: Hua, Cong, et al.
Published: (2024)
Not Every Image is Worth a Thousand Words: Quantifying Originality in Stable Diffusion
by: Haviv, Adi, et al.
Published: (2024)
by: Haviv, Adi, et al.
Published: (2024)
Suppress Content Shift: Better Diffusion Features via Off-the-Shelf Generation Techniques
by: Meng, Benyuan, et al.
Published: (2024)
by: Meng, Benyuan, et al.
Published: (2024)
WordVIS: A Color Worth A Thousand Words
by: Khan, Umar, et al.
Published: (2024)
by: Khan, Umar, et al.
Published: (2024)
Dual-Stage Reweighted MoE for Long-Tailed Egocentric Mistake Detection
by: Han, Boyu, et al.
Published: (2025)
by: Han, Boyu, et al.
Published: (2025)
DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Vartiations
by: Yang, Zhiyong, et al.
Published: (2024)
by: Yang, Zhiyong, et al.
Published: (2024)
Top-K Pairwise Ranking: Bridging the Gap Among Ranking-Based Measures for Multi-Label Classification
by: Wang, Zitai, et al.
Published: (2024)
by: Wang, Zitai, et al.
Published: (2024)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
by: Seo, Hoigi, et al.
Published: (2026)
by: Seo, Hoigi, et al.
Published: (2026)
Guiding Diffusion-based Reconstruction with Contrastive Signals for Balanced Visual Representation
by: Han, Boyu, et al.
Published: (2026)
by: Han, Boyu, et al.
Published: (2026)
AUCSeg: AUC-oriented Pixel-level Long-tail Semantic Segmentation
by: Han, Boyu, et al.
Published: (2024)
by: Han, Boyu, et al.
Published: (2024)
A Unified Perspective for Loss-Oriented Imbalanced Learning via Localization
by: Wang, Zitai, et al.
Published: (2023)
by: Wang, Zitai, et al.
Published: (2023)
A Label is Worth a Thousand Images in Dataset Distillation
by: Qin, Tian, et al.
Published: (2024)
by: Qin, Tian, et al.
Published: (2024)
Making Training-Free Diffusion Segmentors Scale with the Generative Power
by: Meng, Benyuan, et al.
Published: (2026)
by: Meng, Benyuan, et al.
Published: (2026)
Not All Diffusion Model Activations Have Been Evaluated as Discriminative Features
by: Meng, Benyuan, et al.
Published: (2024)
by: Meng, Benyuan, et al.
Published: (2024)
Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs
by: Xu, Zhikang, et al.
Published: (2026)
by: Xu, Zhikang, et al.
Published: (2026)
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
by: Dai, Siran, et al.
Published: (2025)
by: Dai, Siran, et al.
Published: (2025)
A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting
by: Zhuang, Junhao, et al.
Published: (2023)
by: Zhuang, Junhao, et al.
Published: (2023)
Malware Detection in Docker Containers: An Image is Worth a Thousand Logs
by: Nousias, Akis, et al.
Published: (2025)
by: Nousias, Akis, et al.
Published: (2025)
Video Is Worth a Thousand Images: Exploring the Latest Trends in Long Video Generation
by: Waseem, Faraz, et al.
Published: (2024)
by: Waseem, Faraz, et al.
Published: (2024)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models
by: Wang, Jiayu, et al.
Published: (2024)
by: Wang, Jiayu, et al.
Published: (2024)
A LoRA is Worth a Thousand Pictures
by: Liu, Chenxi, et al.
Published: (2024)
by: Liu, Chenxi, et al.
Published: (2024)
Channel Vision Transformers: An Image Is Worth 1 x 16 x 16 Words
by: Bao, Yujia, et al.
Published: (2023)
by: Bao, Yujia, et al.
Published: (2023)
A Scene is Worth a Thousand Features: Feed-Forward Camera Localization from a Collection of Image Features
by: Barroso-Laguna, Axel, et al.
Published: (2025)
by: Barroso-Laguna, Axel, et al.
Published: (2025)
Modeling Thousands of Human Annotators for Generalizable Text-to-Image Person Re-identification
by: Jiang, Jiayu, et al.
Published: (2025)
by: Jiang, Jiayu, et al.
Published: (2025)
A Picture is Worth a Thousand Words? An Empirical Study of Aggregation Strategies for Visual Financial Document Retrieval
by: Lim, Ho Hung, et al.
Published: (2026)
by: Lim, Ho Hung, et al.
Published: (2026)
Pick-and-Draw: Training-free Semantic Guidance for Text-to-Image Personalization
by: Lv, Henglei, et al.
Published: (2024)
by: Lv, Henglei, et al.
Published: (2024)
STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
by: Srivatsan, Koushik, et al.
Published: (2024)
by: Srivatsan, Koushik, et al.
Published: (2024)
Similar Items
-
BlackMirror: Black-Box Backdoor Detection for Text-to-Image Models via Instruction-Response Deviation
by: Li, Feiran, et al.
Published: (2026) -
Size-invariance Matters: Rethinking Metrics and Losses for Imbalanced Multi-object Salient Object Detection
by: Li, Feiran, et al.
Published: (2024) -
Hybrid Generative Fusion for Efficient and Privacy-Preserving Face Recognition Dataset Generation
by: Li, Feiran, et al.
Published: (2025) -
Towards Size-invariant Salient Object Detection: A Generic Evaluation and Optimization Approach
by: Bao, Shilong, et al.
Published: (2025) -
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
by: Luo, Zhi, et al.
Published: (2025)