RELIC: Interactive Video World Model with Long-Horizon Memory
Fuente:
arXiv
Salvato in:
| Autori principali: | Hong, Yicong, Mei, Yiqun, Ge, Chongjian, Xu, Yiran, Zhou, Yang, Bi, Sai, Hold-Geoffroy, Yannick, Roberts, Mike, Fisher, Matthew, Shechtman, Eli, Sunkavalli, Kalyan, Liu, Feng, Li, Zhengqi, Tan, Hao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation
di: Liu, Yuheng, et al.
Pubblicazione: (2026)
di: Liu, Yuheng, et al.
Pubblicazione: (2026)
LightIt: Illumination Modeling and Control for Diffusion Models
di: Kocsis, Peter, et al.
Pubblicazione: (2024)
di: Kocsis, Peter, et al.
Pubblicazione: (2024)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
di: Huang, Xun, et al.
Pubblicazione: (2025)
di: Huang, Xun, et al.
Pubblicazione: (2025)
Test-Time Training Done Right
di: Zhang, Tianyuan, et al.
Pubblicazione: (2025)
di: Zhang, Tianyuan, et al.
Pubblicazione: (2025)
Causality in Video Diffusers is Separable from Denoising
di: Bai, Xingjian, et al.
Pubblicazione: (2026)
di: Bai, Xingjian, et al.
Pubblicazione: (2026)
TokenLight: Precise Lighting Control in Images using Attribute Tokens
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2026)
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2026)
GimbalDiffusion: Gravity-Aware Camera Control for Video Generation
di: Fortier-Chouinard, Frédéric, et al.
Pubblicazione: (2025)
di: Fortier-Chouinard, Frédéric, et al.
Pubblicazione: (2025)
LRM: Large Reconstruction Model for Single Image to 3D
di: Hong, Yicong, et al.
Pubblicazione: (2023)
di: Hong, Yicong, et al.
Pubblicazione: (2023)
GaSLight: Gaussian Splats for Spatially-Varying Lighting in HDR
di: Bolduc, Christophe, et al.
Pubblicazione: (2025)
di: Bolduc, Christophe, et al.
Pubblicazione: (2025)
MatSwap: Light-aware material transfers in images
di: Lopes, Ivan, et al.
Pubblicazione: (2025)
di: Lopes, Ivan, et al.
Pubblicazione: (2025)
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
di: Wang, Chen, et al.
Pubblicazione: (2026)
di: Wang, Chen, et al.
Pubblicazione: (2026)
Long-Context State-Space Video World Models
di: Po, Ryan, et al.
Pubblicazione: (2025)
di: Po, Ryan, et al.
Pubblicazione: (2025)
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
di: Zhang, Kai, et al.
Pubblicazione: (2024)
di: Zhang, Kai, et al.
Pubblicazione: (2024)
NeuManifold: Neural Watertight Manifold Reconstruction with Efficient and High-Quality Rendering Support
di: Wei, Xinyue, et al.
Pubblicazione: (2023)
di: Wei, Xinyue, et al.
Pubblicazione: (2023)
Generating 360° Video is What You Need For a 3D Scene
di: Zhang, Zhaoyang, et al.
Pubblicazione: (2025)
di: Zhang, Zhaoyang, et al.
Pubblicazione: (2025)
MotionStream: Real-Time Video Generation with Interactive Motion Controls
di: Shin, Joonghyuk, et al.
Pubblicazione: (2025)
di: Shin, Joonghyuk, et al.
Pubblicazione: (2025)
Neural Directional Encoding for Efficient and Accurate View-Dependent Appearance Modeling
di: Wu, Liwen, et al.
Pubblicazione: (2024)
di: Wu, Liwen, et al.
Pubblicazione: (2024)
E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training
di: Zhao, Qitao, et al.
Pubblicazione: (2025)
di: Zhao, Qitao, et al.
Pubblicazione: (2025)
Rethinking Training Dynamics in Scale-wise Autoregressive Generation
di: Zhou, Gengze, et al.
Pubblicazione: (2025)
di: Zhou, Gengze, et al.
Pubblicazione: (2025)
Endless World: Real-Time 3D-Aware Long Video Generation
di: Zhang, Ke, et al.
Pubblicazione: (2025)
di: Zhang, Ke, et al.
Pubblicazione: (2025)
SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic Faces
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2025)
di: Chaturvedi, Sumit, et al.
Pubblicazione: (2025)
MeshLRM: Large Reconstruction Model for High-Quality Meshes
di: Wei, Xinyue, et al.
Pubblicazione: (2024)
di: Wei, Xinyue, et al.
Pubblicazione: (2024)
PyraVid: Hierarchical Multimodal Memory for Long-Horizon Video Reasoning
di: Yan, Sikuan, et al.
Pubblicazione: (2026)
di: Yan, Sikuan, et al.
Pubblicazione: (2026)
Reinforcement Learning for Long-Horizon Multi-Turn Search Agents
di: Kalyan, Vivek, et al.
Pubblicazione: (2025)
di: Kalyan, Vivek, et al.
Pubblicazione: (2025)
VideoGigaGAN: Towards Detail-rich Video Super-Resolution
di: Xu, Yiran, et al.
Pubblicazione: (2024)
di: Xu, Yiran, et al.
Pubblicazione: (2024)
UniLight: A Unified Representation for Lighting
di: Zhang, Zitian, et al.
Pubblicazione: (2025)
di: Zhang, Zitian, et al.
Pubblicazione: (2025)
Motion Modes: What Could Happen Next?
di: Pandey, Karran, et al.
Pubblicazione: (2024)
di: Pandey, Karran, et al.
Pubblicazione: (2024)
ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes
di: Chen, Honglin, et al.
Pubblicazione: (2026)
di: Chen, Honglin, et al.
Pubblicazione: (2026)
Generative Video Motion Editing with 3D Point Tracks
di: Lee, Yao-Chih, et al.
Pubblicazione: (2025)
di: Lee, Yao-Chih, et al.
Pubblicazione: (2025)
Towards a Perceptual Evaluation Framework for Lighting Estimation
di: Giroux, Justine, et al.
Pubblicazione: (2023)
di: Giroux, Justine, et al.
Pubblicazione: (2023)
RELIC: Evaluating Complex Reasoning via the Recognition of Languages In-Context
di: Petty, Jackson, et al.
Pubblicazione: (2025)
di: Petty, Jackson, et al.
Pubblicazione: (2025)
PreciseCam: Precise Camera Control for Text-to-Image Generation
di: Bernal-Berdun, Edurne, et al.
Pubblicazione: (2025)
di: Bernal-Berdun, Edurne, et al.
Pubblicazione: (2025)
COMPOSE: Comprehensive Portrait Shadow Editing
di: Hou, Andrew, et al.
Pubblicazione: (2024)
di: Hou, Andrew, et al.
Pubblicazione: (2024)
RELIC: Investigating Large Language Model Responses using Self-Consistency
di: Cheng, Furui, et al.
Pubblicazione: (2023)
di: Cheng, Furui, et al.
Pubblicazione: (2023)
Character Mixing for Video Generation
di: Liao, Tingting, et al.
Pubblicazione: (2025)
di: Liao, Tingting, et al.
Pubblicazione: (2025)
RayZer: A Self-supervised Large View Synthesis Model
di: Jiang, Hanwen, et al.
Pubblicazione: (2025)
di: Jiang, Hanwen, et al.
Pubblicazione: (2025)
Generative Portrait Shadow Removal
di: Yoon, Jae Shin, et al.
Pubblicazione: (2024)
di: Yoon, Jae Shin, et al.
Pubblicazione: (2024)
MagicWorld: Towards Long-Horizon Stability for Interactive Video World Exploration
di: Li, Guangyuan, et al.
Pubblicazione: (2025)
di: Li, Guangyuan, et al.
Pubblicazione: (2025)
WorldWeaver: Generating Long-Horizon Video Worlds via Rich Perception
di: Liu, Zhiheng, et al.
Pubblicazione: (2025)
di: Liu, Zhiheng, et al.
Pubblicazione: (2025)
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats
di: Ziwen, Chen, et al.
Pubblicazione: (2024)
di: Ziwen, Chen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation
di: Liu, Yuheng, et al.
Pubblicazione: (2026) -
LightIt: Illumination Modeling and Control for Diffusion Models
di: Kocsis, Peter, et al.
Pubblicazione: (2024) -
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
di: Huang, Xun, et al.
Pubblicazione: (2025) -
Test-Time Training Done Right
di: Zhang, Tianyuan, et al.
Pubblicazione: (2025) -
Causality in Video Diffusers is Separable from Denoising
di: Bai, Xingjian, et al.
Pubblicazione: (2026)