UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Wójcik, Piotr, Petrenko, Maksym, Gromski, Wojciech, Spurek, Przemysław, Zieba, Maciej
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866911418446512128
author Wójcik, Piotr
Petrenko, Maksym
Gromski, Wojciech
Spurek, Przemysław
Zieba, Maciej
author_facet Wójcik, Piotr
Petrenko, Maksym
Gromski, Wojciech
Spurek, Przemysław
Zieba, Maciej
contents Recent advances in large-scale diffusion models have intensified concerns about their potential misuse, particularly in generating realistic yet harmful or socially disruptive content. This challenge has spurred growing interest in effective machine unlearning, the process of selectively removing specific knowledge or concepts from a model without compromising its overall generative capabilities. Among various approaches, Low-Rank Adaptation (LoRA) has emerged as an effective and efficient method for fine-tuning models toward targeted unlearning. However, LoRA-based methods often exhibit limited adaptability to concept semantics and struggle to balance removing closely related concepts with maintaining generalization across broader meanings. Moreover, these methods face scalability challenges when multiple concepts must be erased simultaneously. To address these limitations, we introduce UnHype, a framework that incorporates hypernetworks into single- and multi-concept LoRA training. The proposed architecture can be directly plugged into Stable Diffusion as well as modern flow-based text-to-image models, where it demonstrates stable training behavior and effective concept control. During inference, the hypernetwork dynamically generates adaptive LoRA weights based on the CLIP embedding, enabling more context-aware, scalable unlearning. We evaluate UnHype across several challenging tasks, including object erasure, celebrity erasure, and explicit content removal, demonstrating its effectiveness and versatility. Repository: https://github.com/gmum/UnHype.
format Preprint
id arxiv_https___arxiv_org_abs_2602_03410
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning
Wójcik, Piotr
Petrenko, Maksym
Gromski, Wojciech
Spurek, Przemysław
Zieba, Maciej
Computer Vision and Pattern Recognition
Recent advances in large-scale diffusion models have intensified concerns about their potential misuse, particularly in generating realistic yet harmful or socially disruptive content. This challenge has spurred growing interest in effective machine unlearning, the process of selectively removing specific knowledge or concepts from a model without compromising its overall generative capabilities. Among various approaches, Low-Rank Adaptation (LoRA) has emerged as an effective and efficient method for fine-tuning models toward targeted unlearning. However, LoRA-based methods often exhibit limited adaptability to concept semantics and struggle to balance removing closely related concepts with maintaining generalization across broader meanings. Moreover, these methods face scalability challenges when multiple concepts must be erased simultaneously. To address these limitations, we introduce UnHype, a framework that incorporates hypernetworks into single- and multi-concept LoRA training. The proposed architecture can be directly plugged into Stable Diffusion as well as modern flow-based text-to-image models, where it demonstrates stable training behavior and effective concept control. During inference, the hypernetwork dynamically generates adaptive LoRA weights based on the CLIP embedding, enabling more context-aware, scalable unlearning. We evaluate UnHype across several challenging tasks, including object erasure, celebrity erasure, and explicit content removal, demonstrating its effectiveness and versatility. Repository: https://github.com/gmum/UnHype.
title UnHype: CLIP-Guided Hypernetworks for Dynamic LoRA Unlearning
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2602.03410