PanoWorld: Geometry-Consistent Panoramic Video World Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Le, Bai, Xiangyu, Galoaa, Bishoy, Moezzi, Shayda, Lee, Caleb James, Imtiaz, Tooba, Yeh, Edmund, Dy, Jennifer, Wang, Yanzhi, Ostadabbas, Sarah |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Motion-o: Trajectory-Grounded Video Reasoning
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis
by: Bai, Xiangyu, et al.
Published: (2025)
by: Bai, Xiangyu, et al.
Published: (2025)
Structure Over Scale: Learning Visual Reasoning from Pedagogical Video
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
HORNet: Task-Guided Frame Selection for Video Question Answering with Vision-Language Models
by: Bai, Xiangyu, et al.
Published: (2026)
by: Bai, Xiangyu, et al.
Published: (2026)
Track and Caption Any Motion: Query-Free Motion Discovery and Description in Videos
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Lang2Motion: Bridging Language and Motion through Joint Embedding Spaces
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Look Around and Pay Attention: Multi-camera Point Tracking Reimagined with Transformers
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Broadening View Synthesis of Dynamic Scenes from Constrained Monocular Videos
by: Jiang, Le, et al.
Published: (2025)
by: Jiang, Le, et al.
Published: (2025)
PanoWorld-X: Generating Explorable Panoramic Worlds via Sphere-Aware Video Diffusion
by: Yin, Yuyang, et al.
Published: (2025)
by: Yin, Yuyang, et al.
Published: (2025)
K-Track: Kalman-Enhanced Tracking for Accelerating Deep Point Trackers on Edge Devices
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesis
by: Jia, Jinrang, et al.
Published: (2026)
by: Jia, Jinrang, et al.
Published: (2026)
PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World
by: Wang, Changpeng, et al.
Published: (2026)
by: Wang, Changpeng, et al.
Published: (2026)
UniTrack: Differentiable Graph Representation Learning for Multi-Object Tracking
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
ADAPT to Robustify Prompt Tuning Vision Transformers
by: Eskandar, Masih, et al.
Published: (2024)
by: Eskandar, Masih, et al.
Published: (2024)
PhyGround: Benchmarking Physical Reasoning in Generative World Models
by: Lin, Juyi, et al.
Published: (2026)
by: Lin, Juyi, et al.
Published: (2026)
STAR: Stability-Inducing Weight Perturbation for Continual Learning
by: Eskandar, Masih, et al.
Published: (2025)
by: Eskandar, Masih, et al.
Published: (2025)
Pano360: Perspective to Panoramic Vision with Geometric Consistency
by: Zhu, Zhengdong, et al.
Published: (2026)
by: Zhu, Zhengdong, et al.
Published: (2026)
VidPanos: Generative Panoramic Videos from Casual Panning Videos
by: Ma, Jingwei, et al.
Published: (2024)
by: Ma, Jingwei, et al.
Published: (2024)
Human Cognition in Machines: A Unified Perspective of World Models
by: Rupprecht, Timothy, et al.
Published: (2026)
by: Rupprecht, Timothy, et al.
Published: (2026)
LVT: Large-Scale Scene Reconstruction via Local View Transformers
by: Imtiaz, Tooba, et al.
Published: (2025)
by: Imtiaz, Tooba, et al.
Published: (2025)
WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling
by: Fang, Shaoheng, et al.
Published: (2025)
by: Fang, Shaoheng, et al.
Published: (2025)
PanoAir: A Panoramic Visual-Inertial SLAM with Cross-Time Real-World UAV Dataset
by: Wu, Yiyang, et al.
Published: (2026)
by: Wu, Yiyang, et al.
Published: (2026)
PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation
by: Dong, Zeyu, et al.
Published: (2025)
by: Dong, Zeyu, et al.
Published: (2025)
PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation
by: Yan, Shilin, et al.
Published: (2023)
by: Yan, Shilin, et al.
Published: (2023)
SAIF: Sparse Adversarial and Imperceptible Attack Framework
by: Imtiaz, Tooba, et al.
Published: (2022)
by: Imtiaz, Tooba, et al.
Published: (2022)
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
by: Dai, Yixiang, et al.
Published: (2025)
by: Dai, Yixiang, et al.
Published: (2025)
Pano-NeRF: Synthesizing High Dynamic Range Novel Views with Geometry from Sparse Low Dynamic Range Panoramic Images
by: Lu, Zhan, et al.
Published: (2023)
by: Lu, Zhan, et al.
Published: (2023)
PanoDP: Learning Collision-Free Navigation with Panoramic Depth and Differentiable Physics
by: Zhong, Hao, et al.
Published: (2026)
by: Zhong, Hao, et al.
Published: (2026)
STREAMS: An Assistive Multimodal AI Framework for Empowering Biosignal Based Robotic Controls
by: Rabiee, Ali, et al.
Published: (2024)
by: Rabiee, Ali, et al.
Published: (2024)
VGGT-360: Geometry-Consistent Zero-Shot Panoramic Depth Estimation
by: Yuan, Jiayi, et al.
Published: (2026)
by: Yuan, Jiayi, et al.
Published: (2026)
PanoGAN A Deep Generative Model for Panoramic Dental Radiographs
by: Pedersen, Soren, et al.
Published: (2025)
by: Pedersen, Soren, et al.
Published: (2025)
Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
by: Wu, Haoyu, et al.
Published: (2025)
by: Wu, Haoyu, et al.
Published: (2025)
Geometry-Aware Rotary Position Embedding for Consistent Video World Model
by: Xiang, Chendong, et al.
Published: (2026)
by: Xiang, Chendong, et al.
Published: (2026)
DreamHome-Pano: Design-Aware and Conflict-Free Panoramic Interior Generation
by: Chen, Lulu, et al.
Published: (2026)
by: Chen, Lulu, et al.
Published: (2026)
PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic Imagery
by: Guo, Yijing, et al.
Published: (2026)
by: Guo, Yijing, et al.
Published: (2026)
PanoEnv: Exploring 3D Spatial Intelligence in Panoramic Environments with Reinforcement Learning
by: Lin, Zekai, et al.
Published: (2026)
by: Lin, Zekai, et al.
Published: (2026)
PanoDiff-SR: Synthesizing Dental Panoramic Radiographs using Diffusion and Super-resolution
by: Jain, Sanyam, et al.
Published: (2025)
by: Jain, Sanyam, et al.
Published: (2025)
Assessing Lower Limb Strength using Internet-of-Things Enabled Chair
by: Yeh, Chelsea, et al.
Published: (2022)
by: Yeh, Chelsea, et al.
Published: (2022)
OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation
by: Liu, Yuheng, et al.
Published: (2026)
by: Liu, Yuheng, et al.
Published: (2026)
PanoTPS-Net: Panoramic Room Layout Estimation via Thin Plate Spline Transformation
by: Ibrahem, Hatem, et al.
Published: (2025)
by: Ibrahem, Hatem, et al.
Published: (2025)
Similar Items
-
Motion-o: Trajectory-Grounded Video Reasoning
by: Galoaa, Bishoy, et al.
Published: (2026) -
MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis
by: Bai, Xiangyu, et al.
Published: (2025) -
Structure Over Scale: Learning Visual Reasoning from Pedagogical Video
by: Galoaa, Bishoy, et al.
Published: (2026) -
HORNet: Task-Guided Frame Selection for Video Question Answering with Vision-Language Models
by: Bai, Xiangyu, et al.
Published: (2026) -
Track and Caption Any Motion: Query-Free Motion Discovery and Description in Videos
by: Galoaa, Bishoy, et al.
Published: (2025)