PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhan, Jiahao, Li, Zizhang, Yu, Hong-Xing, Wu, Jiajun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RealWonder: Real-Time Physical Action-Conditioned Video Generation
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
von: Li, Zizhang, et al.
Veröffentlicht: (2025)
von: Li, Zizhang, et al.
Veröffentlicht: (2025)
WonderWorld: Interactive 3D Scene Generation from a Single Image
von: Yu, Hong-Xing, et al.
Veröffentlicht: (2024)
von: Yu, Hong-Xing, et al.
Veröffentlicht: (2024)
WonderZoom: Multi-Scale 3D World Generation
von: Cao, Jin, et al.
Veröffentlicht: (2025)
von: Cao, Jin, et al.
Veröffentlicht: (2025)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
von: Li, Hongjie, et al.
Veröffentlicht: (2024)
von: Li, Hongjie, et al.
Veröffentlicht: (2024)
The Scene Language: Representing Scenes with Programs, Words, and Embeddings
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment
von: Xia, Chong, et al.
Veröffentlicht: (2025)
von: Xia, Chong, et al.
Veröffentlicht: (2025)
WonderVerse: Extendable 3D Scene Generation with Video Generative Models
von: Feng, Hao, et al.
Veröffentlicht: (2025)
von: Feng, Hao, et al.
Veröffentlicht: (2025)
WonderJourney: Going from Anywhere to Everywhere
von: Yu, Hong-Xing, et al.
Veröffentlicht: (2023)
von: Yu, Hong-Xing, et al.
Veröffentlicht: (2023)
AutoScape: Geometry-Consistent Long-Horizon Scene Generation
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
Product of Experts for Visual Generation
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2025)
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2025)
3D Congealing: 3D-Aware Image Alignment in the Wild
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
von: Zhang, Yunzhi, et al.
Veröffentlicht: (2024)
Reconstruction and Simulation of Elastic Objects with Spring-Mass 3D Gaussians
von: Zhong, Licheng, et al.
Veröffentlicht: (2024)
von: Zhong, Licheng, et al.
Veröffentlicht: (2024)
Voyaging into Perpetual Dynamic Scenes from a Single View
von: Tian, Fengrui, et al.
Veröffentlicht: (2025)
von: Tian, Fengrui, et al.
Veröffentlicht: (2025)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
Learning the 3D Fauna of the Web
von: Li, Zizhang, et al.
Veröffentlicht: (2024)
von: Li, Zizhang, et al.
Veröffentlicht: (2024)
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
von: Wang, Jiamin, et al.
Veröffentlicht: (2025)
von: Wang, Jiamin, et al.
Veröffentlicht: (2025)
FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video
von: Gao, Yue, et al.
Veröffentlicht: (2025)
von: Gao, Yue, et al.
Veröffentlicht: (2025)
WonderTurbo: Generating Interactive 3D World in 0.72 Seconds
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
Unsupervised Discovery of Object-Centric Neural Fields
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
Learning 4D Panoptic Scene Graph Generation from Rich 2D Visual Scene
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
von: Yu, Chengjun, et al.
Veröffentlicht: (2026)
von: Yu, Chengjun, et al.
Veröffentlicht: (2026)
4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
Sparse4DGS: 4D Gaussian Splatting for Sparse-Frame Dynamic Scene Reconstruction
von: Shi, Changyue, et al.
Veröffentlicht: (2025)
von: Shi, Changyue, et al.
Veröffentlicht: (2025)
General Scene Adaptation for Vision-and-Language Navigation
von: Hong, Haodong, et al.
Veröffentlicht: (2025)
von: Hong, Haodong, et al.
Veröffentlicht: (2025)
FastScene: Text-Driven Fast 3D Indoor Scene Generation via Panoramic Gaussian Splatting
von: Ma, Yikun, et al.
Veröffentlicht: (2024)
von: Ma, Yikun, et al.
Veröffentlicht: (2024)
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
von: Li, Hongjie, et al.
Veröffentlicht: (2026)
von: Li, Hongjie, et al.
Veröffentlicht: (2026)
Efficient 3D Content Reconstruction and Generation
von: Li, Jiahao
Veröffentlicht: (2026)
von: Li, Jiahao
Veröffentlicht: (2026)
HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction
von: Cheng, Chong, et al.
Veröffentlicht: (2026)
von: Cheng, Chong, et al.
Veröffentlicht: (2026)
DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
von: Cai, Kaiwen, et al.
Veröffentlicht: (2025)
TiP4GEN: Text to Immersive Panorama 4D Scene Generation
von: Xing, Ke, et al.
Veröffentlicht: (2025)
von: Xing, Ke, et al.
Veröffentlicht: (2025)
DreamJourney: Perpetual View Generation with Video Diffusion Models
von: Pan, Bo, et al.
Veröffentlicht: (2025)
von: Pan, Bo, et al.
Veröffentlicht: (2025)
The Devil is in Fine-tuning and Long-tailed Problems:A New Benchmark for Scene Text Detection
von: Cao, Tianjiao, et al.
Veröffentlicht: (2025)
von: Cao, Tianjiao, et al.
Veröffentlicht: (2025)
SAS: Segment Any 3D Scene with Integrated 2D Priors
von: Li, Zhuoyuan, et al.
Veröffentlicht: (2025)
von: Li, Zhuoyuan, et al.
Veröffentlicht: (2025)
Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation
von: Yang, Yuanbo, et al.
Veröffentlicht: (2024)
von: Yang, Yuanbo, et al.
Veröffentlicht: (2024)
CoCo4D: Comprehensive and Complex 4D Scene Generation
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
von: Sargent, Kyle, et al.
Veröffentlicht: (2023)
von: Sargent, Kyle, et al.
Veröffentlicht: (2023)
Point'n Move: Interactive Scene Object Manipulation on Gaussian Splatting Radiance Fields
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
von: Huang, Jiajun, et al.
Veröffentlicht: (2023)
4D Panoptic Scene Graph Generation
von: Yang, Jingkang, et al.
Veröffentlicht: (2024)
von: Yang, Jingkang, et al.
Veröffentlicht: (2024)
Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation
von: Li, Ruibin, et al.
Veröffentlicht: (2026)
von: Li, Ruibin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RealWonder: Real-Time Physical Action-Conditioned Video Generation
von: Liu, Wei, et al.
Veröffentlicht: (2026) -
WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions
von: Li, Zizhang, et al.
Veröffentlicht: (2025) -
WonderWorld: Interactive 3D Scene Generation from a Single Image
von: Yu, Hong-Xing, et al.
Veröffentlicht: (2024) -
WonderZoom: Multi-Scale 3D World Generation
von: Cao, Jin, et al.
Veröffentlicht: (2025) -
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
von: Li, Hongjie, et al.
Veröffentlicht: (2024)