SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Zhenghao, Liu, Yuxin, Zhou, Bolei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embodied Scene Understanding for Vision Language Models via MetaVQA
by: Wang, Weizhen, et al.
Published: (2025)
by: Wang, Weizhen, et al.
Published: (2025)
Vid2Sim: Realistic and Interactive Simulation from Video for Urban Navigation
by: Xie, Ziyang, et al.
Published: (2025)
by: Xie, Ziyang, et al.
Published: (2025)
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
by: Jia, Xiaosong, et al.
Published: (2024)
by: Jia, Xiaosong, et al.
Published: (2024)
Humanoid Locomotion as Next Token Prediction
by: Radosavovic, Ilija, et al.
Published: (2024)
by: Radosavovic, Ilija, et al.
Published: (2024)
DriveSceneGen: Generating Diverse and Realistic Driving Scenarios from Scratch
by: Sun, Shuo, et al.
Published: (2023)
by: Sun, Shuo, et al.
Published: (2023)
SimGen: Simulator-conditioned Driving Scene Generation
by: Zhou, Yunsong, et al.
Published: (2024)
by: Zhou, Yunsong, et al.
Published: (2024)
SSF-PAN: Semantic Scene Flow-Based Perception for Autonomous Navigation in Traffic Scenarios
by: Chen, Yinqi, et al.
Published: (2025)
by: Chen, Yinqi, et al.
Published: (2025)
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
by: He, Honglin, et al.
Published: (2025)
by: He, Honglin, et al.
Published: (2025)
Scene Completeness-Aware Lidar Depth Completion for Driving Scenario
by: Wu, Cho-Ying, et al.
Published: (2020)
by: Wu, Cho-Ying, et al.
Published: (2020)
Intelligent Spatial Perception by Building Hierarchical 3D Scene Graphs for Indoor Scenarios with the Help of LLMs
by: Cheng, Yao, et al.
Published: (2025)
by: Cheng, Yao, et al.
Published: (2025)
Learning Sidewalk Autopilot from Multi-Scale Imitation with Corrective Behavior Expansion
by: He, Honglin, et al.
Published: (2026)
by: He, Honglin, et al.
Published: (2026)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
by: Zhang, Haiming, et al.
Published: (2026)
by: Zhang, Haiming, et al.
Published: (2026)
ScenarioControl: Vision-Language Controllable Vectorized Latent Scenario Generation
by: Gao, Lili, et al.
Published: (2026)
by: Gao, Lili, et al.
Published: (2026)
SMART: Scalable Multi-agent Real-time Motion Generation via Next-token Prediction
by: Wu, Wei, et al.
Published: (2024)
by: Wu, Wei, et al.
Published: (2024)
Angle Robustness Unmanned Aerial Vehicle Navigation in GNSS-Denied Scenarios
by: Wang, Yuxin, et al.
Published: (2024)
by: Wang, Yuxin, et al.
Published: (2024)
IONext: Unlocking the Next Era of Inertial Odometry
by: Zhang, Shanshan, et al.
Published: (2025)
by: Zhang, Shanshan, et al.
Published: (2025)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
UrbanVerse: Scaling Urban Simulation by Watching City-Tour Videos
by: Liu, Mingxuan, et al.
Published: (2025)
by: Liu, Mingxuan, et al.
Published: (2025)
Perspective from a Higher Dimension: Can 3D Geometric Priors Help Visual Floorplan Localization?
by: Chen, Bolei, et al.
Published: (2025)
by: Chen, Bolei, et al.
Published: (2025)
SemanticFlow: A Self-Supervised Framework for Joint Scene Flow Prediction and Instance Segmentation in Dynamic Environments
by: Chen, Yinqi, et al.
Published: (2025)
by: Chen, Yinqi, et al.
Published: (2025)
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
by: Li, Jiahang, et al.
Published: (2023)
by: Li, Jiahang, et al.
Published: (2023)
MASSTAR: A Multi-Modal and Large-Scale Scene Dataset with a Versatile Toolchain for Surface Prediction and Completion
by: Zheng, Guiyong, et al.
Published: (2024)
by: Zheng, Guiyong, et al.
Published: (2024)
Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs
by: Catalano, Iacopo, et al.
Published: (2026)
by: Catalano, Iacopo, et al.
Published: (2026)
CounterScene: Counterfactual Causal Reasoning in Generative World Models for Safety-Critical Closed-Loop Evaluation
by: Jing, Bowen, et al.
Published: (2026)
by: Jing, Bowen, et al.
Published: (2026)
SmartRefine: A Scenario-Adaptive Refinement Framework for Efficient Motion Prediction
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments
by: Rowe, Luke, et al.
Published: (2025)
by: Rowe, Luke, et al.
Published: (2025)
Perspective from a Broader Context: Can Room Style Knowledge Help Visual Floorplan Localization?
by: Chen, Bolei, et al.
Published: (2025)
by: Chen, Bolei, et al.
Published: (2025)
Recursive Visual Imagination and Adaptive Linguistic Grounding for Vision Language Navigation
by: Chen, Bolei, et al.
Published: (2025)
by: Chen, Bolei, et al.
Published: (2025)
SAGE: Scalable Agentic 3D Scene Generation for Embodied AI
by: Xia, Hongchi, et al.
Published: (2026)
by: Xia, Hongchi, et al.
Published: (2026)
DynamicCity: Large-Scale 4D Occupancy Generation from Dynamic Scenes
by: Bian, Hengwei, et al.
Published: (2024)
by: Bian, Hengwei, et al.
Published: (2024)
Synth It Like KITTI: Synthetic Data Generation for Object Detection in Driving Scenarios
by: Marcus, Richard, et al.
Published: (2025)
by: Marcus, Richard, et al.
Published: (2025)
InspecSafe-V1: A Multimodal Benchmark for Safety Assessment in Industrial Inspection Scenarios
by: Liu, Zeyi, et al.
Published: (2026)
by: Liu, Zeyi, et al.
Published: (2026)
DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding
by: Ge, Luzhou, et al.
Published: (2026)
by: Ge, Luzhou, et al.
Published: (2026)
ReconDrive: Fast Feed-Forward 4D Gaussian Splatting for Autonomous Driving Scene Reconstruction
by: Yu, Haibao, et al.
Published: (2026)
by: Yu, Haibao, et al.
Published: (2026)
Towards Next-Generation SLAM: A Survey on 3DGS-SLAM Focusing on Performance, Robustness, and Future Directions
by: Wang, Li, et al.
Published: (2026)
by: Wang, Li, et al.
Published: (2026)
DeformGS: Scene Flow in Highly Deformable Scenes for Deformable Object Manipulation
by: Duisterhof, Bardienus P., et al.
Published: (2023)
by: Duisterhof, Bardienus P., et al.
Published: (2023)
DragTraffic: Interactive and Controllable Traffic Scene Generation for Autonomous Driving
by: Wang, Sheng, et al.
Published: (2024)
by: Wang, Sheng, et al.
Published: (2024)
Estimating Commonsense Scene Composition on Belief Scene Graphs
by: Saucedo, Mario A. V., et al.
Published: (2025)
by: Saucedo, Mario A. V., et al.
Published: (2025)
SemanticFormer: Holistic and Semantic Traffic Scene Representation for Trajectory Prediction using Knowledge Graphs
by: Sun, Zhigang, et al.
Published: (2024)
by: Sun, Zhigang, et al.
Published: (2024)
MoD-SLAM: Monocular Dense Mapping for Unbounded 3D Scene Reconstruction
by: Zhou, Heng, et al.
Published: (2024)
by: Zhou, Heng, et al.
Published: (2024)
Similar Items
-
Embodied Scene Understanding for Vision Language Models via MetaVQA
by: Wang, Weizhen, et al.
Published: (2025) -
Vid2Sim: Realistic and Interactive Simulation from Video for Urban Navigation
by: Xie, Ziyang, et al.
Published: (2025) -
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
by: Jia, Xiaosong, et al.
Published: (2024) -
Humanoid Locomotion as Next Token Prediction
by: Radosavovic, Ilija, et al.
Published: (2024) -
DriveSceneGen: Generating Diverse and Realistic Driving Scenarios from Scratch
by: Sun, Shuo, et al.
Published: (2023)