Learning Multimodal Latent Generative Models with Energy-Based Prior
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Shiyu, Cui, Jiali, Li, Hanao, Han, Tian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Multimodal Latent Space with EBM Prior and MCMC Inference
von: Yuan, Shiyu, et al.
Veröffentlicht: (2024)
von: Yuan, Shiyu, et al.
Veröffentlicht: (2024)
Learning Latent Space Hierarchical EBM Diffusion Models
von: Cui, Jiali, et al.
Veröffentlicht: (2024)
von: Cui, Jiali, et al.
Veröffentlicht: (2024)
ShaLa: Multimodal Shared Latent Space Modelling
von: Cui, Jiali, et al.
Veröffentlicht: (2025)
von: Cui, Jiali, et al.
Veröffentlicht: (2025)
Learning Energy-based Variational Latent Prior for VAEs
von: Dutta, Debottam, et al.
Veröffentlicht: (2025)
von: Dutta, Debottam, et al.
Veröffentlicht: (2025)
Neural Prior Estimation: Learning Class Priors from Latent Representations
von: Yavari, Masoud, et al.
Veröffentlicht: (2026)
von: Yavari, Masoud, et al.
Veröffentlicht: (2026)
Aligning Latent Spaces with Flow Priors
von: Li, Yizhuo, et al.
Veröffentlicht: (2025)
von: Li, Yizhuo, et al.
Veröffentlicht: (2025)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
von: Liu, Shengqi, et al.
Veröffentlicht: (2024)
von: Liu, Shengqi, et al.
Veröffentlicht: (2024)
Beyond DAGs: A Latent Partial Causal Model for Multimodal Learning
von: Liu, Yuhang, et al.
Veröffentlicht: (2024)
von: Liu, Yuhang, et al.
Veröffentlicht: (2024)
LatentExplainer: Explaining Latent Representations in Deep Generative Models with Multimodal Large Language Models
von: Zhu, Mengdan, et al.
Veröffentlicht: (2024)
von: Zhu, Mengdan, et al.
Veröffentlicht: (2024)
Improving Adversarial Energy-Based Model via Diffusion Process
von: Geng, Cong, et al.
Veröffentlicht: (2024)
von: Geng, Cong, et al.
Veröffentlicht: (2024)
Designing a Conditional Prior Distribution for Flow-Based Generative Models
von: Issachar, Noam, et al.
Veröffentlicht: (2025)
von: Issachar, Noam, et al.
Veröffentlicht: (2025)
Bridging Compressed Image Latents and Multimodal Large Language Models
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
von: Kao, Chia-Hao, et al.
Veröffentlicht: (2024)
Multimodal Latent Language Modeling with Next-Token Diffusion
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
von: Maduabuchi, Chika, et al.
Veröffentlicht: (2025)
von: Maduabuchi, Chika, et al.
Veröffentlicht: (2025)
OlmoEarth: Stable Latent Image Modeling for Multimodal Earth Observation
von: Herzog, Henry, et al.
Veröffentlicht: (2025)
von: Herzog, Henry, et al.
Veröffentlicht: (2025)
Learning Single Index Models with Diffusion Priors
von: Tang, Anqi, et al.
Veröffentlicht: (2025)
von: Tang, Anqi, et al.
Veröffentlicht: (2025)
Residual Prior Diffusion: A Probabilistic Framework Integrating Coarse Latent Priors with Diffusion Models
von: Kutsuna, Takuro
Veröffentlicht: (2025)
von: Kutsuna, Takuro
Veröffentlicht: (2025)
Fast Autoregressive Models for Continuous Latent Generation
von: Hang, Tiankai, et al.
Veröffentlicht: (2025)
von: Hang, Tiankai, et al.
Veröffentlicht: (2025)
Taming Feed-forward Reconstruction Models as Latent Encoders for 3D Generative Models
von: Wizadwongsa, Suttisak, et al.
Veröffentlicht: (2024)
von: Wizadwongsa, Suttisak, et al.
Veröffentlicht: (2024)
AGMA: Adaptive Gaussian Mixture Anchors for Prior-Guided Multimodal Human Trajectory Forecasting
von: Li, Chao, et al.
Veröffentlicht: (2026)
von: Li, Chao, et al.
Veröffentlicht: (2026)
MultiDelete for Multimodal Machine Unlearning
von: Cheng, Jiali, et al.
Veröffentlicht: (2023)
von: Cheng, Jiali, et al.
Veröffentlicht: (2023)
A Transformer-based Multimodal Fusion Model for Efficient Crowd Counting Using Visual and Wireless Signals
von: Cui, Zhe, et al.
Veröffentlicht: (2025)
von: Cui, Zhe, et al.
Veröffentlicht: (2025)
EnergyLens: Interpretable Closed-Form Energy Models for Multimodal LLM Inference Serving
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
Foundation Model Makes Clustering A Better Initialization For Cold-Start Active Learning
von: Yuan, Han, et al.
Veröffentlicht: (2024)
von: Yuan, Han, et al.
Veröffentlicht: (2024)
TIME: TabPFN-Integrated Multimodal Engine for Robust Tabular-Image Learning
von: Luo, Jiaqi, et al.
Veröffentlicht: (2025)
von: Luo, Jiaqi, et al.
Veröffentlicht: (2025)
Continuity-Preserving Convolutional Autoencoders for Learning Continuous Latent Dynamical Models from Images
von: Zhu, Aiqing, et al.
Veröffentlicht: (2025)
von: Zhu, Aiqing, et al.
Veröffentlicht: (2025)
LatentGAN Autoencoder: Learning Disentangled Latent Distribution
von: Kalwar, Sanket, et al.
Veröffentlicht: (2022)
von: Kalwar, Sanket, et al.
Veröffentlicht: (2022)
Research on Image Recognition Technology Based on Multimodal Deep Learning
von: Wang, Jinyin, et al.
Veröffentlicht: (2024)
von: Wang, Jinyin, et al.
Veröffentlicht: (2024)
Prior-free Balanced Replay: Uncertainty-guided Reservoir Sampling for Long-Tailed Continual Learning
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
VEDIT: Latent Prediction Architecture For Procedural Video Representation Learning
von: Lin, Han, et al.
Veröffentlicht: (2024)
von: Lin, Han, et al.
Veröffentlicht: (2024)
Generalized Contrastive Learning for Universal Multimodal Retrieval
von: Lee, Jungsoo, et al.
Veröffentlicht: (2025)
von: Lee, Jungsoo, et al.
Veröffentlicht: (2025)
Exploring Structured Semantic Priors Underlying Diffusion Score for Test-time Adaptation
von: Li, Mingjia, et al.
Veröffentlicht: (2025)
von: Li, Mingjia, et al.
Veröffentlicht: (2025)
Generative Latent Diffusion for Efficient Spatiotemporal Data Reduction
von: Li, Xiao, et al.
Veröffentlicht: (2025)
von: Li, Xiao, et al.
Veröffentlicht: (2025)
What's in a Latent? Leveraging Diffusion Latent Space for Domain Generalization
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
von: Thomas, Xavier, et al.
Veröffentlicht: (2025)
Isometric Representation Learning for Disentangled Latent Space of Diffusion Models
von: Hahm, Jaehoon, et al.
Veröffentlicht: (2024)
von: Hahm, Jaehoon, et al.
Veröffentlicht: (2024)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
von: Shen, Ying, et al.
Veröffentlicht: (2026)
von: Shen, Ying, et al.
Veröffentlicht: (2026)
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space
von: Ding, Tianxingjian, et al.
Veröffentlicht: (2025)
von: Ding, Tianxingjian, et al.
Veröffentlicht: (2025)
Correcting Diffusion Generation through Resampling
von: Liu, Yujian, et al.
Veröffentlicht: (2023)
von: Liu, Yujian, et al.
Veröffentlicht: (2023)
Consistent3D: Towards Consistent High-Fidelity Text-to-3D Generation with Deterministic Sampling Prior
von: Wu, Zike, et al.
Veröffentlicht: (2024)
von: Wu, Zike, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Multimodal Latent Space with EBM Prior and MCMC Inference
von: Yuan, Shiyu, et al.
Veröffentlicht: (2024) -
Learning Latent Space Hierarchical EBM Diffusion Models
von: Cui, Jiali, et al.
Veröffentlicht: (2024) -
ShaLa: Multimodal Shared Latent Space Modelling
von: Cui, Jiali, et al.
Veröffentlicht: (2025) -
Learning Energy-based Variational Latent Prior for VAEs
von: Dutta, Debottam, et al.
Veröffentlicht: (2025) -
Neural Prior Estimation: Learning Class Priors from Latent Representations
von: Yavari, Masoud, et al.
Veröffentlicht: (2026)