ShaLa: Multimodal Shared Latent Space Modelling
Fuente:
arXiv
Saved in:
| Main Authors: | Cui, Jiali, Chen, Yan-Ying, Zhang, Yanxia, Klenk, Matthew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Latent Space Hierarchical EBM Diffusion Models
by: Cui, Jiali, et al.
Published: (2024)
by: Cui, Jiali, et al.
Published: (2024)
Learning Multimodal Latent Generative Models with Energy-Based Prior
by: Yuan, Shiyu, et al.
Published: (2024)
by: Yuan, Shiyu, et al.
Published: (2024)
LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation
by: Nehme, Ghadi, et al.
Published: (2025)
by: Nehme, Ghadi, et al.
Published: (2025)
Learning From Design Procedure To Generate CAD Programs for Data Augmentation
by: Chen, Yan-Ying, et al.
Published: (2026)
by: Chen, Yan-Ying, et al.
Published: (2026)
Stylish and Functional: Guided Interpolation Subject to Physical Constraints
by: Chen, Yan-Ying, et al.
Published: (2024)
by: Chen, Yan-Ying, et al.
Published: (2024)
Learning Multimodal Latent Space with EBM Prior and MCMC Inference
by: Yuan, Shiyu, et al.
Published: (2024)
by: Yuan, Shiyu, et al.
Published: (2024)
Aligning Latent Spaces with Flow Priors
by: Li, Yizhuo, et al.
Published: (2025)
by: Li, Yizhuo, et al.
Published: (2025)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
by: Liu, Shengqi, et al.
Published: (2024)
by: Liu, Shengqi, et al.
Published: (2024)
Att-Adapter: A Robust and Precise Domain-Specific Multi-Attributes T2I Diffusion Adapter via Conditional Variational Autoencoder
by: Cho, Wonwoong, et al.
Published: (2025)
by: Cho, Wonwoong, et al.
Published: (2025)
Beyond DAGs: A Latent Partial Causal Model for Multimodal Learning
by: Liu, Yuhang, et al.
Published: (2024)
by: Liu, Yuhang, et al.
Published: (2024)
DocShaDiffusion: Diffusion Model in Latent Space for Document Image Shadow Removal
by: Liu, Wenjie, et al.
Published: (2025)
by: Liu, Wenjie, et al.
Published: (2025)
LatentUM: Unleashing the Potential of Interleaved Cross-Modal Reasoning via a Latent-Space Unified Model
by: Jin, Jiachun, et al.
Published: (2026)
by: Jin, Jiachun, et al.
Published: (2026)
LURE: Latent Space Unblocking for Multi-Concept Reawakening in Diffusion Models
by: Sun, Mengyu, et al.
Published: (2026)
by: Sun, Mengyu, et al.
Published: (2026)
Pixel-Space Post-Training of Latent Diffusion Models
by: Zhang, Christina, et al.
Published: (2024)
by: Zhang, Christina, et al.
Published: (2024)
Latent Diffusion Inversion Requires Understanding the Latent Space
by: Rao, Mingxing, et al.
Published: (2025)
by: Rao, Mingxing, et al.
Published: (2025)
CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space
by: Ding, Tianxingjian, et al.
Published: (2025)
by: Ding, Tianxingjian, et al.
Published: (2025)
OlmoEarth: Stable Latent Image Modeling for Multimodal Earth Observation
by: Herzog, Henry, et al.
Published: (2025)
by: Herzog, Henry, et al.
Published: (2025)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
by: Haas, René, et al.
Published: (2023)
by: Haas, René, et al.
Published: (2023)
Isometric Representation Learning for Disentangled Latent Space of Diffusion Models
by: Hahm, Jaehoon, et al.
Published: (2024)
by: Hahm, Jaehoon, et al.
Published: (2024)
EM-Net: Gaze Estimation with Expectation Maximization Algorithm
by: Cheng, Zhang, et al.
Published: (2024)
by: Cheng, Zhang, et al.
Published: (2024)
What's in a Latent? Leveraging Diffusion Latent Space for Domain Generalization
by: Thomas, Xavier, et al.
Published: (2025)
by: Thomas, Xavier, et al.
Published: (2025)
ReLaX: Reasoning with Latent Exploration for Large Reasoning Models
by: Zhang, Shimin, et al.
Published: (2025)
by: Zhang, Shimin, et al.
Published: (2025)
Bridging Compressed Image Latents and Multimodal Large Language Models
by: Kao, Chia-Hao, et al.
Published: (2024)
by: Kao, Chia-Hao, et al.
Published: (2024)
LatentExplainer: Explaining Latent Representations in Deep Generative Models with Multimodal Large Language Models
by: Zhu, Mengdan, et al.
Published: (2024)
by: Zhu, Mengdan, et al.
Published: (2024)
VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models
by: Bi, Tianci, et al.
Published: (2025)
by: Bi, Tianci, et al.
Published: (2025)
Multimodal Latent Language Modeling with Next-Token Diffusion
by: Sun, Yutao, et al.
Published: (2024)
by: Sun, Yutao, et al.
Published: (2024)
Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation
by: Baade, Alan, et al.
Published: (2026)
by: Baade, Alan, et al.
Published: (2026)
MultiDelete for Multimodal Machine Unlearning
by: Cheng, Jiali, et al.
Published: (2023)
by: Cheng, Jiali, et al.
Published: (2023)
Gradient-Guided Exploration of Generative Model's Latent Space for Controlled Iris Image Augmentations
by: Mitcheff, Mahsa, et al.
Published: (2025)
by: Mitcheff, Mahsa, et al.
Published: (2025)
LLVD: LSTM-based Explicit Motion Modeling in Latent Space for Blind Video Denoising
by: Rashid, Loay, et al.
Published: (2025)
by: Rashid, Loay, et al.
Published: (2025)
Evaluating the Efficiency of Latent Spaces via the Coupling-Matrix
by: Yavuz, Mehmet Can, et al.
Published: (2025)
by: Yavuz, Mehmet Can, et al.
Published: (2025)
DragGANSpace: Latent Space Exploration and Control for GANs
by: Odendaal, Kirsten, et al.
Published: (2025)
by: Odendaal, Kirsten, et al.
Published: (2025)
Multimodal Structure Learning: Disentangling Shared and Specific Topology via Cross-Modal Graphical Lasso
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
by: Li, Chengzu, et al.
Published: (2025)
by: Li, Chengzu, et al.
Published: (2025)
Segment to Focus: Guiding Latent Action Models in the Presence of Distractors
by: Fechner, Marcus, et al.
Published: (2026)
by: Fechner, Marcus, et al.
Published: (2026)
Shortcut Invariance: Targeted Jacobian Regularization in Disentangled Latent Space
by: Pal, Shivam, et al.
Published: (2025)
by: Pal, Shivam, et al.
Published: (2025)
Fast Autoregressive Models for Continuous Latent Generation
by: Hang, Tiankai, et al.
Published: (2025)
by: Hang, Tiankai, et al.
Published: (2025)
Unifying Image Counterfactuals and Feature Attributions with Latent-Space Adversarial Attacks
by: Goldwasser, Jeremy, et al.
Published: (2025)
by: Goldwasser, Jeremy, et al.
Published: (2025)
Efficient Long-Tail Learning in Latent Space by sampling Synthetic Data
by: Sharma, Nakul
Published: (2025)
by: Sharma, Nakul
Published: (2025)
Canonical Latent Representations in Conditional Diffusion Models
by: Xu, Yitao, et al.
Published: (2025)
by: Xu, Yitao, et al.
Published: (2025)
Similar Items
-
Learning Latent Space Hierarchical EBM Diffusion Models
by: Cui, Jiali, et al.
Published: (2024) -
Learning Multimodal Latent Generative Models with Energy-Based Prior
by: Yuan, Shiyu, et al.
Published: (2024) -
LAMP: Data-Efficient Linear Affine Weight-Space Models for Parameter-Controlled 3D Shape Generation and Extrapolation
by: Nehme, Ghadi, et al.
Published: (2025) -
Learning From Design Procedure To Generate CAD Programs for Data Augmentation
by: Chen, Yan-Ying, et al.
Published: (2026) -
Stylish and Functional: Guided Interpolation Subject to Physical Constraints
by: Chen, Yan-Ying, et al.
Published: (2024)