Teaching Video Generators to Remember: Eliciting Dynamic Memory for Out-of-Sight State Evolution
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Tianshuo, Xie, Yichen, Meng, Depu, Peng, Chensheng, Herau, Quentin, Jiang, Bo, Hu, Yihan, Zhan, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
URoPE: Universal Relative Position Embedding across Geometric Spaces
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
SpectralSplat: Appearance-Disentangled Feed-Forward Gaussian Splatting for Driving Scenes
by: Herau, Quentin, et al.
Published: (2026)
by: Herau, Quentin, et al.
Published: (2026)
LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation
by: Jiang, Bo, et al.
Published: (2026)
by: Jiang, Bo, et al.
Published: (2026)
Learning to Drive is a Free Gift: Large-Scale Label-Free Autonomy Pretraining from Unposed In-The-Wild Videos
by: Strong, Matthew, et al.
Published: (2026)
by: Strong, Matthew, et al.
Published: (2026)
UniQueR: Unified Query-based Feedforward 3D Reconstruction
by: Peng, Chensheng, et al.
Published: (2026)
by: Peng, Chensheng, et al.
Published: (2026)
Out of Sight, Out of Mind? Evaluating State Evolution in Video World Models
by: Ma, Ziqi, et al.
Published: (2026)
by: Ma, Ziqi, et al.
Published: (2026)
Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models
by: Chen, Kaijin, et al.
Published: (2026)
by: Chen, Kaijin, et al.
Published: (2026)
RAYNOVA: Scale-Temporal Autoregressive World Modeling in Ray Space
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
by: Duan, Zicheng, et al.
Published: (2026)
by: Duan, Zicheng, et al.
Published: (2026)
S2GO: Streaming Sparse Gaussian Occupancy Prediction
by: Park, Jinhyung, et al.
Published: (2025)
by: Park, Jinhyung, et al.
Published: (2025)
Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind
by: Plizzari, Chiara, et al.
Published: (2024)
by: Plizzari, Chiara, et al.
Published: (2024)
RT-GS: Gaussian Splatting with Reflection and Transmittance Primitives
by: Zeng, Kunnong, et al.
Published: (2026)
by: Zeng, Kunnong, et al.
Published: (2026)
DeSiRe-GS: 4D Street Gaussians for Static-Dynamic Decomposition and Surface Reconstruction for Urban Driving Scenes
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
Out of Sight, Out of Mind: The State of Government Document Collections and the Need for Awareness.
by: Parker, June D.
Published: (1993)
by: Parker, June D.
Published: (1993)
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
by: Huang, Yixuan, et al.
Published: (2023)
by: Huang, Yixuan, et al.
Published: (2023)
HindSight: Evaluating LLM-Generated Research Ideas via Future Impact
by: Jiang, Bo
Published: (2026)
by: Jiang, Bo
Published: (2026)
Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution
by: Cao, Zouying, et al.
Published: (2025)
by: Cao, Zouying, et al.
Published: (2025)
Grating haptic perception through touchscreen: Sighted vs. Visually Impaired
by: Gao, Yichen, et al.
Published: (2025)
by: Gao, Yichen, et al.
Published: (2025)
IN-Sight: Interactive Navigation through Sight
by: Schoch, Philipp, et al.
Published: (2024)
by: Schoch, Philipp, et al.
Published: (2024)
Seeing, Listening, Remembering, and Reasoning: A Multimodal Agent with Long-Term Memory
by: Long, Lin, et al.
Published: (2025)
by: Long, Lin, et al.
Published: (2025)
VideoSSM: Autoregressive Long Video Generation with Hybrid State-Space Memory
by: Yu, Yifei, et al.
Published: (2025)
by: Yu, Yifei, et al.
Published: (2025)
Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts
by: Elsafoury, Fatma, et al.
Published: (2025)
by: Elsafoury, Fatma, et al.
Published: (2025)
Vietnam Memorial. America Remembers
Published: (1985)
Published: (1985)
Memory and Perception: Remembering Snowflake
by: Jordi FERNÁNDEZ
Published: (2006)
by: Jordi FERNÁNDEZ
Published: (2006)
Optimizing Diffusion Models for Joint Trajectory Prediction and Controllable Generation
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
X-Drive: Cross-modality consistent multi-sensor data synthesis for driving scenarios
by: Xie, Yichen, et al.
Published: (2024)
by: Xie, Yichen, et al.
Published: (2024)
Do You Remember? Dense Video Captioning with Cross-Modal Memory Retrieval
by: Kim, Minkuk, et al.
Published: (2024)
by: Kim, Minkuk, et al.
Published: (2024)
Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics
by: Xu, Tianshuo, et al.
Published: (2026)
by: Xu, Tianshuo, et al.
Published: (2026)
Graph Memory Learning: Imitating Lifelong Remembering and Forgetting of Brain Networks
by: Miao, Jiaxing, et al.
Published: (2024)
by: Miao, Jiaxing, et al.
Published: (2024)
Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions
by: Qu, Yiting, et al.
Published: (2025)
by: Qu, Yiting, et al.
Published: (2025)
Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory
by: Sun, Jingwei, et al.
Published: (2026)
by: Sun, Jingwei, et al.
Published: (2026)
Scoring, Remember, and Reference: Catching Camouflaged Objects in Videos
by: Feng, Yuang, et al.
Published: (2025)
by: Feng, Yuang, et al.
Published: (2025)
Out of Sight, Out of Track: Adversarial Attacks on Propagation-based Multi-Object Trackers via Query State Manipulation
by: Bouzidi, Halima, et al.
Published: (2026)
by: Bouzidi, Halima, et al.
Published: (2026)
A Mechanistic View on Video Generation as World Models: State and Dynamics
by: Wang, Luozhou, et al.
Published: (2026)
by: Wang, Luozhou, et al.
Published: (2026)
Choosing How to Remember: Adaptive Memory Structures for LLM Agents
by: Lu, Mingfei, et al.
Published: (2026)
by: Lu, Mingfei, et al.
Published: (2026)
Attend Locally, Remember Linearly: Linear Attention as Cross-Frame Memory for Autoregressive Video Diffusion
by: Li, Kunyang, et al.
Published: (2026)
by: Li, Kunyang, et al.
Published: (2026)
Semiclassical asymptotics of the Bloch--Torrey operator in two dimensions
by: Hérau, Frédéric, et al.
Published: (2024)
by: Hérau, Frédéric, et al.
Published: (2024)
HiddenDetect: Detecting Jailbreak Attacks against Large Vision-Language Models via Monitoring Hidden States
by: Jiang, Yilei, et al.
Published: (2025)
by: Jiang, Yilei, et al.
Published: (2025)
How Generations Remember
by: Palmberger, Monika
Published: (2017)
by: Palmberger, Monika
Published: (2017)
Out of Sight, Not Out of Context? Egocentric Spatial Reasoning in VLMs Across Disjoint Frames
by: Ravi, Sahithya, et al.
Published: (2025)
by: Ravi, Sahithya, et al.
Published: (2025)
Similar Items
-
URoPE: Universal Relative Position Embedding across Geometric Spaces
by: Xie, Yichen, et al.
Published: (2026) -
SpectralSplat: Appearance-Disentangled Feed-Forward Gaussian Splatting for Driving Scenes
by: Herau, Quentin, et al.
Published: (2026) -
LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation
by: Jiang, Bo, et al.
Published: (2026) -
Learning to Drive is a Free Gift: Large-Scale Label-Free Autonomy Pretraining from Unposed In-The-Wild Videos
by: Strong, Matthew, et al.
Published: (2026) -
UniQueR: Unified Query-based Feedforward 3D Reconstruction
by: Peng, Chensheng, et al.
Published: (2026)