Render, Don't Decode: Weight-Space World Models with Latent Structural Disentanglement
Fuente:
arXiv
Saved in:
| Main Authors: | Nzoyem, Roussel Desmond, Comi, Mauro |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
DiLA: Disentangled Latent Action World Models
by: Zhang, Tianqiu, et al.
Published: (2026)
by: Zhang, Tianqiu, et al.
Published: (2026)
Vision Transformers Don't Need Trained Registers
by: Jiang, Nick, et al.
Published: (2025)
by: Jiang, Nick, et al.
Published: (2025)
Don't Fear Peculiar Activation Functions: EUAF and Beyond
by: Wang, Qianchao, et al.
Published: (2024)
by: Wang, Qianchao, et al.
Published: (2024)
Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
by: Woo, Sangmin, et al.
Published: (2024)
by: Woo, Sangmin, et al.
Published: (2024)
I Detect What I Don't Know: Incremental Anomaly Learning with Stochastic Weight Averaging-Gaussian for Oracle-Free Medical Imaging
by: Yadav, Nand Kumar, et al.
Published: (2025)
by: Yadav, Nand Kumar, et al.
Published: (2025)
Don't Play Favorites: Minority Guidance for Diffusion Models
by: Um, Soobin, et al.
Published: (2023)
by: Um, Soobin, et al.
Published: (2023)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
by: Jiao, Pengkun, et al.
Published: (2025)
by: Jiao, Pengkun, et al.
Published: (2025)
You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models
by: Zhao, Kairan, et al.
Published: (2026)
by: Zhao, Kairan, et al.
Published: (2026)
RenderWorld: World Model with Self-Supervised 3D Label
by: Yan, Ziyang, et al.
Published: (2024)
by: Yan, Ziyang, et al.
Published: (2024)
Focus, Don't Prune: Identifying Instruction-Relevant Regions for Information-Rich Image Understanding
by: Kwon, Mincheol, et al.
Published: (2026)
by: Kwon, Mincheol, et al.
Published: (2026)
Large Pre-Training Datasets Don't Always Guarantee Robustness after Fine-Tuning
by: Hwang, Jaedong, et al.
Published: (2024)
by: Hwang, Jaedong, et al.
Published: (2024)
When Text and Images Don't Mix: Bias-Correcting Language-Image Similarity Scores for Anomaly Detection
by: Goodge, Adam, et al.
Published: (2024)
by: Goodge, Adam, et al.
Published: (2024)
DC-AE 1.5: Accelerating Diffusion Model Convergence with Structured Latent Space
by: Chen, Junyu, et al.
Published: (2025)
by: Chen, Junyu, et al.
Published: (2025)
Don't Judge by the Look: Towards Motion Coherent Video Representation
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
by: Jun, Youngjun, et al.
Published: (2024)
by: Jun, Youngjun, et al.
Published: (2024)
Learning Latent Action World Models In The Wild
by: Garrido, Quentin, et al.
Published: (2026)
by: Garrido, Quentin, et al.
Published: (2026)
MVCustom: Multi-View Customized Diffusion via Geometric Latent Rendering and Completion
by: Shin, Minjung, et al.
Published: (2025)
by: Shin, Minjung, et al.
Published: (2025)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
by: Kalluri, Tarun, et al.
Published: (2024)
by: Kalluri, Tarun, et al.
Published: (2024)
Conditioning Latent-Space Clusters for Real-World Anomaly Classification
by: Bogdoll, Daniel, et al.
Published: (2023)
by: Bogdoll, Daniel, et al.
Published: (2023)
GMapLatent: Geometric Mapping in Latent Space
by: Zeng, Wei, et al.
Published: (2025)
by: Zeng, Wei, et al.
Published: (2025)
Latent Video Prediction Learns Better World Models
by: Alrasheed, Ali J, et al.
Published: (2026)
by: Alrasheed, Ali J, et al.
Published: (2026)
Color encoding in Latent Space of Stable Diffusion Models
by: Arias, Guillem, et al.
Published: (2025)
by: Arias, Guillem, et al.
Published: (2025)
DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks
by: Deng, Yueci, et al.
Published: (2026)
by: Deng, Yueci, et al.
Published: (2026)
Chain of World: World Model Thinking in Latent Motion
by: Yang, Fuxiang, et al.
Published: (2026)
by: Yang, Fuxiang, et al.
Published: (2026)
Robust Embodied Perception in Dynamic Environments via Disentangled Weight Fusion
by: Guo, Juncen, et al.
Published: (2026)
by: Guo, Juncen, et al.
Published: (2026)
Out-of-Support Generalisation via Weight-Space Sequence Modelling
by: Nzoyem, Roussel Desmond
Published: (2026)
by: Nzoyem, Roussel Desmond
Published: (2026)
LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs
by: Krojer, Benno, et al.
Published: (2026)
by: Krojer, Benno, et al.
Published: (2026)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
by: Parmar, Mihir, et al.
Published: (2022)
by: Parmar, Mihir, et al.
Published: (2022)
Render-FM: A Foundation Model for Real-time Photorealistic Volumetric Rendering
by: Gao, Zhongpai, et al.
Published: (2025)
by: Gao, Zhongpai, et al.
Published: (2025)
Many-Worlds Inverse Rendering
by: Zhang, Ziyi, et al.
Published: (2024)
by: Zhang, Ziyi, et al.
Published: (2024)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
by: Rykov, Elisei, et al.
Published: (2025)
by: Rykov, Elisei, et al.
Published: (2025)
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latent Space
by: Zhu, Jian, et al.
Published: (2025)
by: Zhu, Jian, et al.
Published: (2025)
Diffusion Model in Latent Space for Medical Image Segmentation Task
by: Ngoc, Huynh Trinh, et al.
Published: (2025)
by: Ngoc, Huynh Trinh, et al.
Published: (2025)
Diversify, Don't Fine-Tune: Scaling Up Visual Recognition Training with Synthetic Images
by: Yu, Zhuoran, et al.
Published: (2023)
by: Yu, Zhuoran, et al.
Published: (2023)
Don't Get Me Wrong: How to Apply Deep Visual Interpretations to Time Series
by: Loeffler, Christoffer, et al.
Published: (2022)
by: Loeffler, Christoffer, et al.
Published: (2022)
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
by: Huang, Zhening, et al.
Published: (2025)
by: Huang, Zhening, et al.
Published: (2025)
Towards Kinetic Manipulation of the Latent Space
by: Porres, Diego
Published: (2024)
by: Porres, Diego
Published: (2024)
Latent Space Analysis for Melanoma Prevention
by: Listone, Ciro, et al.
Published: (2025)
by: Listone, Ciro, et al.
Published: (2025)
Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective
by: Zhu, Yongxin, et al.
Published: (2024)
by: Zhu, Yongxin, et al.
Published: (2024)
Similar Items
-
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025) -
DiLA: Disentangled Latent Action World Models
by: Zhang, Tianqiu, et al.
Published: (2026) -
Vision Transformers Don't Need Trained Registers
by: Jiang, Nick, et al.
Published: (2025) -
Don't Fear Peculiar Activation Functions: EUAF and Beyond
by: Wang, Qianchao, et al.
Published: (2024) -
Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
by: Woo, Sangmin, et al.
Published: (2024)