Enhancing Variational Autoencoders with Smooth Robust Latent Encoding

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Lee, Hyomin, Kim, Minseon, Jang, Sangwon, Jeong, Jongheon, Hwang, Sung Ju
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917997325582336
author Lee, Hyomin
Kim, Minseon
Jang, Sangwon
Jeong, Jongheon
Hwang, Sung Ju
author_facet Lee, Hyomin
Kim, Minseon
Jang, Sangwon
Jeong, Jongheon
Hwang, Sung Ju
contents Variational Autoencoders (VAEs) have played a key role in scaling up diffusion-based generative models, as in Stable Diffusion, yet questions regarding their robustness remain largely underexplored. Although adversarial training has been an established technique for enhancing robustness in predictive models, it has been overlooked for generative models due to concerns about potential fidelity degradation by the nature of trade-offs between performance and robustness. In this work, we challenge this presumption, introducing Smooth Robust Latent VAE (SRL-VAE), a novel adversarial training framework that boosts both generation quality and robustness. In contrast to conventional adversarial training, which focuses on robustness only, our approach smooths the latent space via adversarial perturbations, promoting more generalizable representations while regularizing with originality representation to sustain original fidelity. Applied as a post-training step on pre-trained VAEs, SRL-VAE improves image robustness and fidelity with minimal computational overhead. Experiments show that SRL-VAE improves both generation quality, in image reconstruction and text-guided image editing, and robustness, against Nightshade attacks and image editing attacks. These results establish a new paradigm, showing that adversarial training, once thought to be detrimental to generative models, can instead enhance both fidelity and robustness.
format Preprint
id arxiv_https___arxiv_org_abs_2504_17219
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Enhancing Variational Autoencoders with Smooth Robust Latent Encoding
Lee, Hyomin
Kim, Minseon
Jang, Sangwon
Jeong, Jongheon
Hwang, Sung Ju
Machine Learning
Artificial Intelligence
Cryptography and Security
Variational Autoencoders (VAEs) have played a key role in scaling up diffusion-based generative models, as in Stable Diffusion, yet questions regarding their robustness remain largely underexplored. Although adversarial training has been an established technique for enhancing robustness in predictive models, it has been overlooked for generative models due to concerns about potential fidelity degradation by the nature of trade-offs between performance and robustness. In this work, we challenge this presumption, introducing Smooth Robust Latent VAE (SRL-VAE), a novel adversarial training framework that boosts both generation quality and robustness. In contrast to conventional adversarial training, which focuses on robustness only, our approach smooths the latent space via adversarial perturbations, promoting more generalizable representations while regularizing with originality representation to sustain original fidelity. Applied as a post-training step on pre-trained VAEs, SRL-VAE improves image robustness and fidelity with minimal computational overhead. Experiments show that SRL-VAE improves both generation quality, in image reconstruction and text-guided image editing, and robustness, against Nightshade attacks and image editing attacks. These results establish a new paradigm, showing that adversarial training, once thought to be detrimental to generative models, can instead enhance both fidelity and robustness.
title Enhancing Variational Autoencoders with Smooth Robust Latent Encoding
topic Machine Learning
Artificial Intelligence
Cryptography and Security
url https://arxiv.org/abs/2504.17219