Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas
Fuente:
arXiv
Saved in:
| Main Authors: | Wimbauer, Felix, Manhardt, Fabian, Oechsle, Michael, Kalischek, Nikolai, Rupprecht, Christian, Cremers, Daniel, Tombari, Federico |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CubeDiff: Repurposing Diffusion-Based Image Models for Panorama Generation
by: Kalischek, Nikolai, et al.
Published: (2025)
by: Kalischek, Nikolai, et al.
Published: (2025)
Epipolar Geometry Improves Video Generation Models
by: Kupyn, Orest, et al.
Published: (2025)
by: Kupyn, Orest, et al.
Published: (2025)
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
by: Jevtić, Aleksandar, et al.
Published: (2025)
by: Jevtić, Aleksandar, et al.
Published: (2025)
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos
by: Wimbauer, Felix, et al.
Published: (2025)
by: Wimbauer, Felix, et al.
Published: (2025)
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
by: Wimmer, Thomas, et al.
Published: (2024)
by: Wimmer, Thomas, et al.
Published: (2024)
Boosting Self-Supervision for Single-View Scene Completion via Knowledge Distillation
by: Han, Keonhee, et al.
Published: (2024)
by: Han, Keonhee, et al.
Published: (2024)
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
Mixed Diffusion for 3D Indoor Scene Synthesis
by: Hu, Siyi, et al.
Published: (2024)
by: Hu, Siyi, et al.
Published: (2024)
Learning Neural Exposure Fields for View Synthesis
by: Niemeyer, Michael, et al.
Published: (2025)
by: Niemeyer, Michael, et al.
Published: (2025)
A Taxonomy and Library for Visualizing Learned Features in Convolutional Neural Networks
by: Grün, Felix, et al.
Published: (2016)
by: Grün, Felix, et al.
Published: (2016)
Dream-to-Recon: Monocular 3D Reconstruction with Diffusion-Depth Distillation from Single Images
by: Wulff, Philipp, et al.
Published: (2025)
by: Wulff, Philipp, et al.
Published: (2025)
Denoising Diffusion via Image-Based Rendering
by: Anciukevičius, Titas, et al.
Published: (2024)
by: Anciukevičius, Titas, et al.
Published: (2024)
RadSplat: Radiance Field-Informed Gaussian Splatting for Robust Real-Time Rendering with 900+ FPS
by: Niemeyer, Michael, et al.
Published: (2024)
by: Niemeyer, Michael, et al.
Published: (2024)
Back on Track: Bundle Adjustment for Dynamic Scene Reconstruction
by: Chen, Weirong, et al.
Published: (2025)
by: Chen, Weirong, et al.
Published: (2025)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
by: Parelli, Maria, et al.
Published: (2025)
by: Parelli, Maria, et al.
Published: (2025)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
Scene-Centric Unsupervised Panoptic Segmentation
by: Hahn, Oliver, et al.
Published: (2025)
by: Hahn, Oliver, et al.
Published: (2025)
KP-RED: Exploiting Semantic Keypoints for Joint 3D Shape Retrieval and Deformation
by: Zhang, Ruida, et al.
Published: (2024)
by: Zhang, Ruida, et al.
Published: (2024)
3D scene generation from scene graphs and self-attention
by: Bonazzi, Pietro, et al.
Published: (2024)
by: Bonazzi, Pietro, et al.
Published: (2024)
LODGE: Level-of-Detail Large-Scale Gaussian Splatting with Efficient Rendering
by: Kulhanek, Jonas, et al.
Published: (2025)
by: Kulhanek, Jonas, et al.
Published: (2025)
TiP4GEN: Text to Immersive Panorama 4D Scene Generation
by: Xing, Ke, et al.
Published: (2025)
by: Xing, Ke, et al.
Published: (2025)
PRISM: Probabilistic Representation for Integrated Shape Modeling and Generation
by: Cheng, Lei, et al.
Published: (2025)
by: Cheng, Lei, et al.
Published: (2025)
LayerPano3D: Layered 3D Panorama for Hyper-Immersive Scene Generation
by: Yang, Shuai, et al.
Published: (2024)
by: Yang, Shuai, et al.
Published: (2024)
Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
by: Sandström, Erik, et al.
Published: (2024)
by: Sandström, Erik, et al.
Published: (2024)
Multiview Scene Graph
by: Zhang, Juexiao, et al.
Published: (2024)
by: Zhang, Juexiao, et al.
Published: (2024)
SecondPose: SE(3)-Consistent Dual-Stream Feature Fusion for Category-Level Pose Estimation
by: Chen, Yamei, et al.
Published: (2023)
by: Chen, Yamei, et al.
Published: (2023)
D-SCo: Dual-Stream Conditional Diffusion for Monocular Hand-Held Object Reconstruction
by: Fu, Bowen, et al.
Published: (2023)
by: Fu, Bowen, et al.
Published: (2023)
Cache Me if You Can: Accelerating Diffusion Models through Block Caching
by: Wimbauer, Felix, et al.
Published: (2023)
by: Wimbauer, Felix, et al.
Published: (2023)
Understanding, Accelerating, and Improving MeanFlow Training
by: Kim, Jin-Young, et al.
Published: (2025)
by: Kim, Jin-Young, et al.
Published: (2025)
TetraDiffusion: Tetrahedral Diffusion Models for 3D Shape Generation
by: Kalischek, Nikolai, et al.
Published: (2022)
by: Kalischek, Nikolai, et al.
Published: (2022)
Invisible Stitch: Generating Smooth 3D Scenes with Depth Inpainting
by: Engstler, Paul, et al.
Published: (2024)
by: Engstler, Paul, et al.
Published: (2024)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
MOHO: Learning Single-view Hand-held Object Reconstruction with Multi-view Occlusion-Aware Supervision
by: Zhang, Chenyangguang, et al.
Published: (2023)
by: Zhang, Chenyangguang, et al.
Published: (2023)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
AnyUp: Universal Feature Upsampling
by: Wimmer, Thomas, et al.
Published: (2025)
by: Wimmer, Thomas, et al.
Published: (2025)
Scene-Conditional 3D Object Stylization and Composition
by: Zhou, Jinghao, et al.
Published: (2023)
by: Zhou, Jinghao, et al.
Published: (2023)
Syn4D: A Multiview Synthetic 4D Dataset
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
Geometry-Aware Diffusion Models for Multiview Scene Inpainting
by: Salimi, Ahmad, et al.
Published: (2025)
by: Salimi, Ahmad, et al.
Published: (2025)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
by: Zheng, Shuhong, et al.
Published: (2026)
by: Zheng, Shuhong, et al.
Published: (2026)
MOSAIC-GS: Monocular Scene Reconstruction via Advanced Initialization for Complex Dynamic Environments
by: Morkva, Svitlana, et al.
Published: (2026)
by: Morkva, Svitlana, et al.
Published: (2026)
Similar Items
-
CubeDiff: Repurposing Diffusion-Based Image Models for Panorama Generation
by: Kalischek, Nikolai, et al.
Published: (2025) -
Epipolar Geometry Improves Video Generation Models
by: Kupyn, Orest, et al.
Published: (2025) -
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
by: Jevtić, Aleksandar, et al.
Published: (2025) -
AnyCam: Learning to Recover Camera Poses and Intrinsics from Casual Videos
by: Wimbauer, Felix, et al.
Published: (2025) -
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
by: Wimmer, Thomas, et al.
Published: (2024)