Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Saehyung, Mok, Jisoo, Park, Sangha, Shin, Yongho, Jung, Dahuin, Yoon, Sungroh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
by: Park, Sangha, et al.
Published: (2025)
by: Park, Sangha, et al.
Published: (2025)
Entropy is not Enough for Test-Time Adaptation: From the Perspective of Disentangled Factors
by: Lee, Jonghyun, et al.
Published: (2024)
by: Lee, Jonghyun, et al.
Published: (2024)
Normality Addition via Normality Detection in Industrial Image Anomaly Detection Models
by: Yi, Jihun, et al.
Published: (2024)
by: Yi, Jihun, et al.
Published: (2024)
RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
by: Oh, Yeongtak, et al.
Published: (2025)
by: Oh, Yeongtak, et al.
Published: (2025)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
by: Jung, Mingi, et al.
Published: (2025)
by: Jung, Mingi, et al.
Published: (2025)
Sample-efficient Adversarial Imitation Learning
by: Jung, Dahuin, et al.
Published: (2023)
by: Jung, Dahuin, et al.
Published: (2023)
Self-Supervised Time-Series Anomaly Detection Using Learnable Data Augmentation
by: Choi, Kukjin, et al.
Published: (2024)
by: Choi, Kukjin, et al.
Published: (2024)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
CANDI: Curated Test-Time Adaptation for Multivariate Time-Series Anomaly Detection Under Distribution Shift
by: Kim, HyunGi, et al.
Published: (2026)
by: Kim, HyunGi, et al.
Published: (2026)
Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning
by: Park, Sangkwon, et al.
Published: (2026)
by: Park, Sangkwon, et al.
Published: (2026)
Know "No" Better: A Data-Driven Approach for Enhancing Negation Awareness in CLIP
by: Park, Junsung, et al.
Published: (2025)
by: Park, Junsung, et al.
Published: (2025)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
by: Lee, Saehyung, et al.
Published: (2024)
by: Lee, Saehyung, et al.
Published: (2024)
Balancing Saliency and Coverage: Semantic Prominence-Aware Budgeting for Visual Token Compression in VLMs
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
Exploring the Potential of LLMs as Personalized Assistants: Dataset, Evaluation, and Analysis
by: Mok, Jisoo, et al.
Published: (2025)
by: Mok, Jisoo, et al.
Published: (2025)
Disentangled Motion Modeling for Video Frame Interpolation
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
DAFA: Distance-Aware Fair Adversarial Training
by: Lee, Hyungyu, et al.
Published: (2024)
by: Lee, Hyungyu, et al.
Published: (2024)
Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage
by: Lee, Saehyung, et al.
Published: (2024)
by: Lee, Saehyung, et al.
Published: (2024)
Contextualized Visual Personalization in Vision-Language Models
by: Oh, Yeongtak, et al.
Published: (2026)
by: Oh, Yeongtak, et al.
Published: (2026)
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss
by: Oh, Yeongtak, et al.
Published: (2024)
by: Oh, Yeongtak, et al.
Published: (2024)
Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
by: Kim, HyunGi, et al.
Published: (2025)
by: Kim, HyunGi, et al.
Published: (2025)
Causality-Aware Contrastive Learning for Robust Multivariate Time-Series Anomaly Detection
by: Kim, HyunGi, et al.
Published: (2025)
by: Kim, HyunGi, et al.
Published: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
by: Park, Sangha, et al.
Published: (2025)
by: Park, Sangha, et al.
Published: (2025)
Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
by: Oh, Yeongtak, et al.
Published: (2024)
by: Oh, Yeongtak, et al.
Published: (2024)
STAG: Structural Test-time Alignment of Gradients for Online Adaptation
by: Shin, Juhyeon, et al.
Published: (2024)
by: Shin, Juhyeon, et al.
Published: (2024)
Unleashing Multi-Hop Reasoning Potential in Large Language Models through Repetition of Misordered Context
by: Yu, Sangwon, et al.
Published: (2024)
by: Yu, Sangwon, et al.
Published: (2024)
LLM-based Frameworks for API Argument Filling in Task-Oriented Conversational Systems
by: Mok, Jisoo, et al.
Published: (2024)
by: Mok, Jisoo, et al.
Published: (2024)
SF(DA)$^2$: Source-free Domain Adaptation Through the Lens of Data Augmentation
by: Hwang, Uiwon, et al.
Published: (2024)
by: Hwang, Uiwon, et al.
Published: (2024)
Theme-Explanation Structure for Table Summarization using Large Language Models: A Case Study on Korean Tabular Data
by: Kwack, TaeYoon, et al.
Published: (2025)
by: Kwack, TaeYoon, et al.
Published: (2025)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
by: Baek, Kanghyun, et al.
Published: (2025)
by: Baek, Kanghyun, et al.
Published: (2025)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
by: Shin, Chaehun, et al.
Published: (2024)
by: Shin, Chaehun, et al.
Published: (2024)
Rethinking Training for De-biasing Text-to-Image Generation: Unlocking the Potential of Stable Diffusion
by: Kim, Eunji, et al.
Published: (2024)
by: Kim, Eunji, et al.
Published: (2024)
Authors and Artists as Speakers: Suggestions for Hassle-Free Visits
by: Botham, Jane, et al.
Published: (1977)
by: Botham, Jane, et al.
Published: (1977)
CKNN: Cleansed k-Nearest Neighbor for Unsupervised Video Anomaly Detection
by: Yi, Jihun, et al.
Published: (2024)
by: Yi, Jihun, et al.
Published: (2024)
Automatic Textual Normalization for Hate Speech Detection
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
by: Nguyen, Anh Thi-Hoang, et al.
Published: (2023)
FEAT: Fashion Editing and Try-On from Any Design
by: Kwon, Soye, et al.
Published: (2026)
by: Kwon, Soye, et al.
Published: (2026)
Unsupervised Layer-wise Score Aggregation for Textual OOD Detection
by: Darrin, Maxime, et al.
Published: (2023)
by: Darrin, Maxime, et al.
Published: (2023)
Align before Attend: Aligning Visual and Textual Features for Multimodal Hateful Content Detection
by: Hossain, Eftekhar, et al.
Published: (2024)
by: Hossain, Eftekhar, et al.
Published: (2024)
FUN-AD: Fully Unsupervised Learning for Anomaly Detection with Noisy Training Data
by: Im, Jiin, et al.
Published: (2024)
by: Im, Jiin, et al.
Published: (2024)
mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval
by: Kim, Kyeong Seon, et al.
Published: (2026)
by: Kim, Kyeong Seon, et al.
Published: (2026)
DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
by: Song, Jaewoo, et al.
Published: (2025)
by: Song, Jaewoo, et al.
Published: (2025)
Similar Items
-
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
by: Park, Sangha, et al.
Published: (2025) -
Entropy is not Enough for Test-Time Adaptation: From the Perspective of Disentangled Factors
by: Lee, Jonghyun, et al.
Published: (2024) -
Normality Addition via Normality Detection in Industrial Image Anomaly Detection Models
by: Yi, Jihun, et al.
Published: (2024) -
RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
by: Oh, Yeongtak, et al.
Published: (2025) -
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
by: Jung, Mingi, et al.
Published: (2025)