Isometric Representation Learning for Disentangled Latent Space of Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hahm, Jaehoon, Lee, Junho, Kim, Sunghyun, Lee, Joonseok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Equivariant Latent Alignment via Flow Matching under Group Symmetries
von: Kim, Sunghyun, et al.
Veröffentlicht: (2026)
von: Kim, Sunghyun, et al.
Veröffentlicht: (2026)
Latent Diffusion Models with Masked AutoEncoders
von: Lee, Junho, et al.
Veröffentlicht: (2025)
von: Lee, Junho, et al.
Veröffentlicht: (2025)
Is There a Better Source Distribution than Gaussian? Exploring Source Distributions for Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2025)
von: Lee, Junho, et al.
Veröffentlicht: (2025)
Geometry-Aware Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2026)
von: Lee, Junho, et al.
Veröffentlicht: (2026)
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
von: Jun, Youngjun, et al.
Veröffentlicht: (2024)
von: Jun, Youngjun, et al.
Veröffentlicht: (2024)
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2025)
Disentangled Representation Learning via Modular Compositional Bias
von: Jung, Whie, et al.
Veröffentlicht: (2025)
von: Jung, Whie, et al.
Veröffentlicht: (2025)
Learning Latent Space Hierarchical EBM Diffusion Models
von: Cui, Jiali, et al.
Veröffentlicht: (2024)
von: Cui, Jiali, et al.
Veröffentlicht: (2024)
VG3T: Visual Geometry Grounded Gaussian Transformer
von: Kim, Junho, et al.
Veröffentlicht: (2025)
von: Kim, Junho, et al.
Veröffentlicht: (2025)
Learning 3D Scene Analogies with Neural Contextual Scene Maps
von: Kim, Junho, et al.
Veröffentlicht: (2025)
von: Kim, Junho, et al.
Veröffentlicht: (2025)
Self-Guided Masked Autoencoder
von: Shin, Jeongwoo, et al.
Veröffentlicht: (2025)
von: Shin, Jeongwoo, et al.
Veröffentlicht: (2025)
LatentGAN Autoencoder: Learning Disentangled Latent Distribution
von: Kalwar, Sanket, et al.
Veröffentlicht: (2022)
von: Kalwar, Sanket, et al.
Veröffentlicht: (2022)
Canonical Latent Representations in Conditional Diffusion Models
von: Xu, Yitao, et al.
Veröffentlicht: (2025)
von: Xu, Yitao, et al.
Veröffentlicht: (2025)
Disentanglement in T-space for Faster and Distributed Training of Diffusion Models with Fewer Latent-states
von: Gupta, Samarth, et al.
Veröffentlicht: (2025)
von: Gupta, Samarth, et al.
Veröffentlicht: (2025)
Shortcut Invariance: Targeted Jacobian Regularization in Disentangled Latent Space
von: Pal, Shivam, et al.
Veröffentlicht: (2025)
von: Pal, Shivam, et al.
Veröffentlicht: (2025)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
von: Choi, Seokeon, et al.
Veröffentlicht: (2025)
Steering Guidance for Personalized Text-to-Image Diffusion Models
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
Scalable Frame Sampling for Video Classification: A Semi-Optimal Policy Approach with Reduced Search Space
von: Lee, Junho, et al.
Veröffentlicht: (2024)
von: Lee, Junho, et al.
Veröffentlicht: (2024)
HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images
von: Choi, Sungik, et al.
Veröffentlicht: (2024)
von: Choi, Sungik, et al.
Veröffentlicht: (2024)
Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval
von: Lyou, Eunyi, et al.
Veröffentlicht: (2024)
von: Lyou, Eunyi, et al.
Veröffentlicht: (2024)
CAPA: Contribution-Aware Pruning and FFN Approximation for Efficient Large Vision-Language Models
von: Jha, Samyak, et al.
Veröffentlicht: (2026)
von: Jha, Samyak, et al.
Veröffentlicht: (2026)
Latent Diffusion Inversion Requires Understanding the Latent Space
von: Rao, Mingxing, et al.
Veröffentlicht: (2025)
von: Rao, Mingxing, et al.
Veröffentlicht: (2025)
Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image Segmentation
von: Ha, Seongsu, et al.
Veröffentlicht: (2024)
von: Ha, Seongsu, et al.
Veröffentlicht: (2024)
Gradient-free Decoder Inversion in Latent Diffusion Models
von: Hong, Seongmin, et al.
Veröffentlicht: (2024)
von: Hong, Seongmin, et al.
Veröffentlicht: (2024)
Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model
von: Huang, Lin-Chun, et al.
Veröffentlicht: (2025)
von: Huang, Lin-Chun, et al.
Veröffentlicht: (2025)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
von: Haas, René, et al.
Veröffentlicht: (2023)
von: Haas, René, et al.
Veröffentlicht: (2023)
Disentangled Sparse Representations for Concept-Separated Diffusion Unlearning
von: Kim, Hyeonjin, et al.
Veröffentlicht: (2026)
von: Kim, Hyeonjin, et al.
Veröffentlicht: (2026)
What's in a Latent? Leveraging Diffusion Latent Space for Domain Generalization
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models
von: Sun, Mengyu, et al.
Veröffentlicht: (2026)
von: Sun, Mengyu, et al.
Veröffentlicht: (2026)
Disentangled Representation Learning with the Gromov-Monge Gap
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
Disentangled Representation Learning with Transmitted Information Bottleneck
von: Dang, Zhuohang, et al.
Veröffentlicht: (2023)
von: Dang, Zhuohang, et al.
Veröffentlicht: (2023)
TripleSumm: Adaptive Triple-Modality Fusion for Video Summarization
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
von: Kim, Sumin, et al.
Veröffentlicht: (2026)
A More Word-like Image Tokenization for MLLMs
von: Lee, Hyun, et al.
Veröffentlicht: (2026)
von: Lee, Hyun, et al.
Veröffentlicht: (2026)
CapeLLM: Support-Free Category-Agnostic Pose Estimation with Multimodal Large Language Models
von: Kim, Junho, et al.
Veröffentlicht: (2024)
von: Kim, Junho, et al.
Veröffentlicht: (2024)
Density-based Isometric Mapping
von: Yousefi, Bardia, et al.
Veröffentlicht: (2024)
von: Yousefi, Bardia, et al.
Veröffentlicht: (2024)
Global Context-aware Representation Learning for Spatially Resolved Transcriptomics
von: Oh, Yunhak, et al.
Veröffentlicht: (2025)
von: Oh, Yunhak, et al.
Veröffentlicht: (2025)
FedSOL: Stabilized Orthogonal Learning with Proximal Restrictions in Federated Learning
von: Lee, Gihun, et al.
Veröffentlicht: (2023)
von: Lee, Gihun, et al.
Veröffentlicht: (2023)
Generalization of Diffusion Models Arises with a Balanced Representation Space
von: Zhang, Zekai, et al.
Veröffentlicht: (2025)
von: Zhang, Zekai, et al.
Veröffentlicht: (2025)
Constructing Fair Latent Space for Intersection of Fairness and Explainability
von: Joo, Hyungjun, et al.
Veröffentlicht: (2024)
von: Joo, Hyungjun, et al.
Veröffentlicht: (2024)
Erase at the Core: Representation Unlearning for Machine Unlearning
von: Lee, Jaewon, et al.
Veröffentlicht: (2026)
von: Lee, Jaewon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Equivariant Latent Alignment via Flow Matching under Group Symmetries
von: Kim, Sunghyun, et al.
Veröffentlicht: (2026) -
Latent Diffusion Models with Masked AutoEncoders
von: Lee, Junho, et al.
Veröffentlicht: (2025) -
Is There a Better Source Distribution than Gaussian? Exploring Source Distributions for Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2025) -
Geometry-Aware Image Flow Matching
von: Lee, Junho, et al.
Veröffentlicht: (2026) -
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
von: Jun, Youngjun, et al.
Veröffentlicht: (2024)