VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Runjia, Torr, Philip, Vedaldi, Andrea, Jakab, Tomas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scene-Conditional 3D Object Stylization and Composition
by: Zhou, Jinghao, et al.
Published: (2023)
by: Zhou, Jinghao, et al.
Published: (2023)
Instant Uncertainty Calibration of NeRFs Using a Meta-Calibrator
by: Amini-Naieni, Niki, et al.
Published: (2023)
by: Amini-Naieni, Niki, et al.
Published: (2023)
Farm3D: Learning Articulated 3D Animals by Distilling 2D Diffusion
by: Jakab, Tomas, et al.
Published: (2023)
by: Jakab, Tomas, et al.
Published: (2023)
Free3D: Consistent Novel View Synthesis without 3D Representation
by: Zheng, Chuanxia, et al.
Published: (2023)
by: Zheng, Chuanxia, et al.
Published: (2023)
DualPM: Dual Posed-Canonical Point Maps for 3D Shape and Pose Reconstruction
by: Kaye, Ben, et al.
Published: (2024)
by: Kaye, Ben, et al.
Published: (2024)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
by: Yu, Jiwen, et al.
Published: (2025)
by: Yu, Jiwen, et al.
Published: (2025)
Flex3D: Feed-Forward 3D Generation with Flexible Reconstruction Model and Input View Curation
by: Han, Junlin, et al.
Published: (2024)
by: Han, Junlin, et al.
Published: (2024)
DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors
by: Zhu, Thomas Hanwen, et al.
Published: (2024)
by: Zhu, Thomas Hanwen, et al.
Published: (2024)
SurfelSplat: Learning Efficient and Generalizable Gaussian Surfel Representations for Sparse-View Surface Reconstruction
by: Dai, Chensheng, et al.
Published: (2026)
by: Dai, Chensheng, et al.
Published: (2026)
Surgical Gaussian Surfels: Highly Accurate Real-time Surgical Scene Rendering using Gaussian Surfels
by: Sunmola, Idris O., et al.
Published: (2025)
by: Sunmola, Idris O., et al.
Published: (2025)
EgoEdit: Dataset, Real-Time Streaming Model, and Benchmark for Egocentric Video Editing
by: Li, Runjia, et al.
Published: (2025)
by: Li, Runjia, et al.
Published: (2025)
Radiometrically Consistent Gaussian Surfels for Inverse Rendering
by: Han, Kyu Beom, et al.
Published: (2026)
by: Han, Kyu Beom, et al.
Published: (2026)
Invisible Stitch: Generating Smooth 3D Scenes with Depth Inpainting
by: Engstler, Paul, et al.
Published: (2024)
by: Engstler, Paul, et al.
Published: (2024)
Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics
by: Li, Ruining, et al.
Published: (2024)
by: Li, Ruining, et al.
Published: (2024)
Semantic Score Distillation Sampling for Compositional Text-to-3D Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
DGE: Direct Gaussian 3D Editing by Consistent Multi-view Editing
by: Chen, Minghao, et al.
Published: (2024)
by: Chen, Minghao, et al.
Published: (2024)
Splatter Image: Ultra-Fast Single-View 3D Reconstruction
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
DreamBeast: Distilling 3D Fantastical Animals with Part-Aware Knowledge Transfer
by: Li, Runjia, et al.
Published: (2024)
by: Li, Runjia, et al.
Published: (2024)
Geometrically Consistent Multi-View Scene Generation from Freehand Sketches
by: Bourouis, Ahmed, et al.
Published: (2026)
by: Bourouis, Ahmed, et al.
Published: (2026)
Learning the 3D Fauna of the Web
by: Li, Zizhang, et al.
Published: (2024)
by: Li, Zizhang, et al.
Published: (2024)
MVSplat360: Feed-Forward 360 Scene Synthesis from Sparse Views
by: Chen, Yuedong, et al.
Published: (2024)
by: Chen, Yuedong, et al.
Published: (2024)
SolidGS: Consolidating Gaussian Surfel Splatting for Sparse-View Surface Reconstruction
by: Shen, Zhuowen, et al.
Published: (2024)
by: Shen, Zhuowen, et al.
Published: (2024)
Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft
by: Huang, Junchao, et al.
Published: (2025)
by: Huang, Junchao, et al.
Published: (2025)
Adaptive and Temporally Consistent Gaussian Surfels for Multi-view Dynamic Reconstruction
by: Chen, Decai, et al.
Published: (2024)
by: Chen, Decai, et al.
Published: (2024)
Nexels: Neurally-Textured Surfels for Real-Time Novel View Synthesis with Sparse Geometries
by: Rong, Victor, et al.
Published: (2025)
by: Rong, Victor, et al.
Published: (2025)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
by: Jiang, Zeren, et al.
Published: (2025)
by: Jiang, Zeren, et al.
Published: (2025)
Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation
by: Wu, Mingqiang, et al.
Published: (2026)
by: Wu, Mingqiang, et al.
Published: (2026)
VideoMemory: Toward Consistent Video Generation via Memory Integration
by: Zhou, Jinsong, et al.
Published: (2026)
by: Zhou, Jinsong, et al.
Published: (2026)
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
by: Lai, Zihang, et al.
Published: (2025)
by: Lai, Zihang, et al.
Published: (2025)
CLIP as RNN: Segment Countless Visual Concepts without Training Endeavor
by: Sun, Shuyang, et al.
Published: (2023)
by: Sun, Shuyang, et al.
Published: (2023)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
by: Danier, Duolikun, et al.
Published: (2025)
by: Danier, Duolikun, et al.
Published: (2025)
I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
GaussCtrl: Multi-View Consistent Text-Driven 3D Gaussian Splatting Editing
by: Wu, Jing, et al.
Published: (2024)
by: Wu, Jing, et al.
Published: (2024)
Generative Photographic Control for Scene-Consistent Video Cinematic Editing
by: Sun, Huiqiang, et al.
Published: (2025)
by: Sun, Huiqiang, et al.
Published: (2025)
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
by: Lee, JoungBin, et al.
Published: (2025)
by: Lee, JoungBin, et al.
Published: (2025)
TeSO: Representing and Compressing 3D Point Cloud Scenes with Textured Surfel Octree
by: Hu, Yueyu, et al.
Published: (2025)
by: Hu, Yueyu, et al.
Published: (2025)
VideoStudio: Generating Consistent-Content and Multi-Scene Videos
by: Long, Fuchen, et al.
Published: (2024)
by: Long, Fuchen, et al.
Published: (2024)
PhyScene3D: Physically Consistent Interactive 3D Tabletop Scene Generation
by: Chen, Weixing, et al.
Published: (2026)
by: Chen, Weixing, et al.
Published: (2026)
V-DPM: 4D Video Reconstruction with Dynamic Point Maps
by: Sucar, Edgar, et al.
Published: (2026)
by: Sucar, Edgar, et al.
Published: (2026)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
by: Tschernezki, Vadim, et al.
Published: (2025)
by: Tschernezki, Vadim, et al.
Published: (2025)
Similar Items
-
Scene-Conditional 3D Object Stylization and Composition
by: Zhou, Jinghao, et al.
Published: (2023) -
Instant Uncertainty Calibration of NeRFs Using a Meta-Calibrator
by: Amini-Naieni, Niki, et al.
Published: (2023) -
Farm3D: Learning Articulated 3D Animals by Distilling 2D Diffusion
by: Jakab, Tomas, et al.
Published: (2023) -
Free3D: Consistent Novel View Synthesis without 3D Representation
by: Zheng, Chuanxia, et al.
Published: (2023) -
DualPM: Dual Posed-Canonical Point Maps for 3D Shape and Pose Reconstruction
by: Kaye, Ben, et al.
Published: (2024)