OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Yuheng, Lin, Xin, Li, Xinke, Yang, Baihan, Wang, Chen, Sunkavalli, Kalyan, Hold-Geoffroy, Yannick, Tan, Hao, Zhang, Kai, Xie, Xiaohui, Shi, Zifan, Hu, Yiwei |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
RELIC: Interactive Video World Model with Long-Horizon Memory
par: Hong, Yicong, et autres
Publié: (2025)
par: Hong, Yicong, et autres
Publié: (2025)
LightIt: Illumination Modeling and Control for Diffusion Models
par: Kocsis, Peter, et autres
Publié: (2024)
par: Kocsis, Peter, et autres
Publié: (2024)
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
par: Wang, Chen, et autres
Publié: (2026)
par: Wang, Chen, et autres
Publié: (2026)
Learning Unified Representation of 3D Gaussian Splatting
par: Xin, Yuelin, et autres
Publié: (2025)
par: Xin, Yuelin, et autres
Publié: (2025)
Generating 360° Video is What You Need For a 3D Scene
par: Zhang, Zhaoyang, et autres
Publié: (2025)
par: Zhang, Zhaoyang, et autres
Publié: (2025)
GimbalDiffusion: Gravity-Aware Camera Control for Video Generation
par: Fortier-Chouinard, Frédéric, et autres
Publié: (2025)
par: Fortier-Chouinard, Frédéric, et autres
Publié: (2025)
GaSLight: Gaussian Splats for Spatially-Varying Lighting in HDR
par: Bolduc, Christophe, et autres
Publié: (2025)
par: Bolduc, Christophe, et autres
Publié: (2025)
MatSwap: Light-aware material transfers in images
par: Lopes, Ivan, et autres
Publié: (2025)
par: Lopes, Ivan, et autres
Publié: (2025)
360-GS: Layout-guided Panoramic Gaussian Splatting For Indoor Roaming
par: Bai, Jiayang, et autres
Publié: (2024)
par: Bai, Jiayang, et autres
Publié: (2024)
RGB$\leftrightarrow$X: Image decomposition and synthesis using material- and lighting-aware diffusion models
par: Zeng, Zheng, et autres
Publié: (2024)
par: Zeng, Zheng, et autres
Publié: (2024)
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
par: Zhang, Kai, et autres
Publié: (2024)
par: Zhang, Kai, et autres
Publié: (2024)
SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic Faces
par: Chaturvedi, Sumit, et autres
Publié: (2025)
par: Chaturvedi, Sumit, et autres
Publié: (2025)
VideoPanda: Video Panoramic Diffusion with Multi-view Attention
par: Xie, Kevin, et autres
Publié: (2025)
par: Xie, Kevin, et autres
Publié: (2025)
Owl-1: Omni World Model for Consistent Long Video Generation
par: Huang, Yuanhui, et autres
Publié: (2024)
par: Huang, Yuanhui, et autres
Publié: (2024)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
par: Song, Quanjian, et autres
Publié: (2025)
par: Song, Quanjian, et autres
Publié: (2025)
PreciseCam: Precise Camera Control for Text-to-Image Generation
par: Bernal-Berdun, Edurne, et autres
Publié: (2025)
par: Bernal-Berdun, Edurne, et autres
Publié: (2025)
Reinforcement Learning for Long-Horizon Multi-Turn Search Agents
par: Kalyan, Vivek, et autres
Publié: (2025)
par: Kalyan, Vivek, et autres
Publié: (2025)
MGM-Omni: Scaling Omni LLMs to Personalized Long-Horizon Speech
par: Wang, Chengyao, et autres
Publié: (2025)
par: Wang, Chengyao, et autres
Publié: (2025)
UniLight: A Unified Representation for Lighting
par: Zhang, Zitian, et autres
Publié: (2025)
par: Zhang, Zitian, et autres
Publié: (2025)
Motion Modes: What Could Happen Next?
par: Pandey, Karran, et autres
Publié: (2024)
par: Pandey, Karran, et autres
Publié: (2024)
ViPS: Video-informed Pose Spaces for Auto-Rigged Meshes
par: Chen, Honglin, et autres
Publié: (2026)
par: Chen, Honglin, et autres
Publié: (2026)
Roaming in acetaldehyde
par: Krajňák, Vladimír, et autres
Publié: (2020)
par: Krajňák, Vladimír, et autres
Publié: (2020)
Remote and Roaming
par: Toivanen, Mari
Publié: (2025)
par: Toivanen, Mari
Publié: (2025)
E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training
par: Zhao, Qitao, et autres
Publié: (2025)
par: Zhao, Qitao, et autres
Publié: (2025)
Towards a Perceptual Evaluation Framework for Lighting Estimation
par: Giroux, Justine, et autres
Publié: (2023)
par: Giroux, Justine, et autres
Publié: (2023)
Entropy-Guided k-Guard Sampling for Long-Horizon Autoregressive Video Generation
par: Han, Yizhao, et autres
Publié: (2026)
par: Han, Yizhao, et autres
Publié: (2026)
PanoWorld: Geometry-Consistent Panoramic Video World Modeling
par: Jiang, Le, et autres
Publié: (2026)
par: Jiang, Le, et autres
Publié: (2026)
Panoramic Out-of-Distribution Segmentation
par: Duan, Mengfei, et autres
Publié: (2025)
par: Duan, Mengfei, et autres
Publié: (2025)
COMPOSE: Comprehensive Portrait Shadow Editing
par: Hou, Andrew, et autres
Publié: (2024)
par: Hou, Andrew, et autres
Publié: (2024)
Test-Time Training Done Right
par: Zhang, Tianyuan, et autres
Publié: (2025)
par: Zhang, Tianyuan, et autres
Publié: (2025)
MeshLRM: Large Reconstruction Model for High-Quality Meshes
par: Wei, Xinyue, et autres
Publié: (2024)
par: Wei, Xinyue, et autres
Publié: (2024)
Conversational DNA: A New Visual Language for Understanding Dialogue Structure in Human and AI
par: Lin, Baihan
Publié: (2025)
par: Lin, Baihan
Publié: (2025)
The Gradient of Health Data Privacy
par: Lin, Baihan
Publié: (2024)
par: Lin, Baihan
Publié: (2024)
Topological Representational Similarity Analysis in Brains and Beyond
par: Lin, Baihan
Publié: (2024)
par: Lin, Baihan
Publié: (2024)
The Machine Can't Replace the Human Heart
par: Lin, Baihan
Publié: (2024)
par: Lin, Baihan
Publié: (2024)
Perceptual Reality Transformer: Neural Architectures for Simulating Neurological Perception Conditions
par: Lin, Baihan
Publié: (2025)
par: Lin, Baihan
Publié: (2025)
Generative Portrait Shadow Removal
par: Yoon, Jae Shin, et autres
Publié: (2024)
par: Yoon, Jae Shin, et autres
Publié: (2024)
TokenLight: Precise Lighting Control in Images using Attribute Tokens
par: Chaturvedi, Sumit, et autres
Publié: (2026)
par: Chaturvedi, Sumit, et autres
Publié: (2026)
QuaDreamer: Controllable Panoramic Video Generation for Quadruped Robots
par: Wu, Sheng, et autres
Publié: (2025)
par: Wu, Sheng, et autres
Publié: (2025)
MagicWorld: Towards Long-Horizon Stability for Interactive Video World Exploration
par: Li, Guangyuan, et autres
Publié: (2025)
par: Li, Guangyuan, et autres
Publié: (2025)
Documents similaires
-
RELIC: Interactive Video World Model with Long-Horizon Memory
par: Hong, Yicong, et autres
Publié: (2025) -
LightIt: Illumination Modeling and Control for Diffusion Models
par: Kocsis, Peter, et autres
Publié: (2024) -
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
par: Wang, Chen, et autres
Publié: (2026) -
Learning Unified Representation of 3D Gaussian Splatting
par: Xin, Yuelin, et autres
Publié: (2025) -
Generating 360° Video is What You Need For a 3D Scene
par: Zhang, Zhaoyang, et autres
Publié: (2025)