SLARM: Streaming and Language-Aligned Reconstruction Model for Dynamic Scenes
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiu, Zhicheng, Meng, Jiarui, Luo, Tong-an, Huang, Yican, Feng, Xuan, Li, Xuanfu, Xu, ZHan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SHARE: Scene-Human Aligned Reconstruction
di: Li, Joshua, et al.
Pubblicazione: (2025)
di: Li, Joshua, et al.
Pubblicazione: (2025)
Learning to Robustly Reconstruct Low-light Dynamic Scenes from Spike Streams
di: Hu, Liwen, et al.
Pubblicazione: (2024)
di: Hu, Liwen, et al.
Pubblicazione: (2024)
Instant Gaussian Stream: Fast and Generalizable Streaming of Dynamic Scene Reconstruction via Gaussian Splatting
di: Yan, Jinbo, et al.
Pubblicazione: (2025)
di: Yan, Jinbo, et al.
Pubblicazione: (2025)
RelayGS: Reconstructing Dynamic Scenes with Large-Scale and Complex Motions via Relay Gaussians
di: Gao, Qiankun, et al.
Pubblicazione: (2024)
di: Gao, Qiankun, et al.
Pubblicazione: (2024)
StreamCacheVGGT: Streaming Visual Geometry Transformers with Robust Scoring and Hybrid Cache Compression
di: Liu, Xuanyi, et al.
Pubblicazione: (2026)
di: Liu, Xuanyi, et al.
Pubblicazione: (2026)
Dynamic Scene Reconstruction: Recent Advance in Real-time Rendering and Streaming
di: Zhu, Jiaxuan, et al.
Pubblicazione: (2025)
di: Zhu, Jiaxuan, et al.
Pubblicazione: (2025)
SPKLIP: Aligning Spike Video Streams with Natural Language
di: Gao, Yongchang, et al.
Pubblicazione: (2025)
di: Gao, Yongchang, et al.
Pubblicazione: (2025)
DyGLNet: Hybrid Global-Local Feature Fusion with Dynamic Upsampling for Medical Image Segmentation
di: Zhao, Yican, et al.
Pubblicazione: (2025)
di: Zhao, Yican, et al.
Pubblicazione: (2025)
ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction
di: Liang, Jie, et al.
Pubblicazione: (2026)
di: Liang, Jie, et al.
Pubblicazione: (2026)
SceneScript: Reconstructing Scenes With An Autoregressive Structured Language Model
di: Avetisyan, Armen, et al.
Pubblicazione: (2024)
di: Avetisyan, Armen, et al.
Pubblicazione: (2024)
Dynamic EventNeRF: Reconstructing General Dynamic Scenes from Multi-view RGB and Event Streams
di: Rudnev, Viktor, et al.
Pubblicazione: (2024)
di: Rudnev, Viktor, et al.
Pubblicazione: (2024)
CAST: Component-Aligned 3D Scene Reconstruction from an RGB Image
di: Yao, Kaixin, et al.
Pubblicazione: (2025)
di: Yao, Kaixin, et al.
Pubblicazione: (2025)
ADC-GS: Anchor-Driven Deformable and Compressed Gaussian Splatting for Dynamic Scene Reconstruction
di: Huang, He, et al.
Pubblicazione: (2025)
di: Huang, He, et al.
Pubblicazione: (2025)
HiCoM: Hierarchical Coherent Motion for Streamable Dynamic Scene with 3D Gaussian Splatting
di: Gao, Qiankun, et al.
Pubblicazione: (2024)
di: Gao, Qiankun, et al.
Pubblicazione: (2024)
Prior-Enhanced Gaussian Splatting for Dynamic Scene Reconstruction from Casual Video
di: Shih, Meng-Li, et al.
Pubblicazione: (2025)
di: Shih, Meng-Li, et al.
Pubblicazione: (2025)
Learning Dynamic Scene Reconstruction with Sinusoidal Geometric Priors
di: Guo, Tian, et al.
Pubblicazione: (2025)
di: Guo, Tian, et al.
Pubblicazione: (2025)
Modeling Ambient Scene Dynamics for Free-view Synthesis
di: Shih, Meng-Li, et al.
Pubblicazione: (2024)
di: Shih, Meng-Li, et al.
Pubblicazione: (2024)
Hierarchically-Structured Open-Vocabulary Indoor Scene Synthesis with Pre-trained Large Language Model
di: Sun, Weilin, et al.
Pubblicazione: (2025)
di: Sun, Weilin, et al.
Pubblicazione: (2025)
Generating Multimodal Driving Scenes via Next-Scene Prediction
di: Wu, Yanhao, et al.
Pubblicazione: (2025)
di: Wu, Yanhao, et al.
Pubblicazione: (2025)
DaRePlane: Direction-aware Representations for Dynamic Scene Reconstruction
di: Lou, Ange, et al.
Pubblicazione: (2024)
di: Lou, Ange, et al.
Pubblicazione: (2024)
RF4D:Neural Radar Fields for Novel View Synthesis in Outdoor Dynamic Scenes
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction
di: Fan, Zhiwen, et al.
Pubblicazione: (2025)
di: Fan, Zhiwen, et al.
Pubblicazione: (2025)
Dynamic Gaussian Scene Reconstruction from Unsynchronized Videos
di: Xu, Zhixin, et al.
Pubblicazione: (2025)
di: Xu, Zhixin, et al.
Pubblicazione: (2025)
EndoLRMGS: Complete Endoscopic Scene Reconstruction combining Large Reconstruction Modelling and Gaussian Splatting
di: Wang, Xu, et al.
Pubblicazione: (2025)
di: Wang, Xu, et al.
Pubblicazione: (2025)
SplitGaussian: Reconstructing Dynamic Scenes via Visual Geometry Decomposition
di: Li, Jiahui, et al.
Pubblicazione: (2025)
di: Li, Jiahui, et al.
Pubblicazione: (2025)
Event-boosted Deformable 3D Gaussians for Dynamic Scene Reconstruction
di: Xu, Wenhao, et al.
Pubblicazione: (2024)
di: Xu, Wenhao, et al.
Pubblicazione: (2024)
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model
di: Feng, Qianhan, et al.
Pubblicazione: (2024)
di: Feng, Qianhan, et al.
Pubblicazione: (2024)
ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes
di: Liu, Zhiming, et al.
Pubblicazione: (2026)
di: Liu, Zhiming, et al.
Pubblicazione: (2026)
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
di: Wang, Jiamin, et al.
Pubblicazione: (2025)
di: Wang, Jiamin, et al.
Pubblicazione: (2025)
FireRescue: A UAV-Based Dataset and Enhanced YOLO Model for Object Detection in Fire Rescue Scenes
di: Xu, Qingyu, et al.
Pubblicazione: (2025)
di: Xu, Qingyu, et al.
Pubblicazione: (2025)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
di: Wang, Chuhan, et al.
Pubblicazione: (2026)
di: Wang, Chuhan, et al.
Pubblicazione: (2026)
SpectroMotion: Dynamic 3D Reconstruction of Specular Scenes
di: Fan, Cheng-De, et al.
Pubblicazione: (2024)
di: Fan, Cheng-De, et al.
Pubblicazione: (2024)
Mem4D: Decoupling Static and Dynamic Memory for Dynamic Scene Reconstruction
di: Cai, Xudong, et al.
Pubblicazione: (2025)
di: Cai, Xudong, et al.
Pubblicazione: (2025)
Multi-Stream Keypoint Attention Network for Sign Language Recognition and Translation
di: Guan, Mo, et al.
Pubblicazione: (2024)
di: Guan, Mo, et al.
Pubblicazione: (2024)
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking
di: Feng, X., et al.
Pubblicazione: (2025)
di: Feng, X., et al.
Pubblicazione: (2025)
StreamBridge: Turning Your Offline Video Large Language Model into a Proactive Streaming Assistant
di: Wang, Haibo, et al.
Pubblicazione: (2025)
di: Wang, Haibo, et al.
Pubblicazione: (2025)
OmniStream: Mastering Perception, Reconstruction and Action in Continuous Streams
di: Yan, Yibin, et al.
Pubblicazione: (2026)
di: Yan, Yibin, et al.
Pubblicazione: (2026)
Aligning Neuronal Coding of Dynamic Visual Scenes with Foundation Vision Models
di: Wu, Rining, et al.
Pubblicazione: (2024)
di: Wu, Rining, et al.
Pubblicazione: (2024)
Unified Scene Representation and Reconstruction for 3D Large Language Models
di: Chu, Tao, et al.
Pubblicazione: (2024)
di: Chu, Tao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SHARE: Scene-Human Aligned Reconstruction
di: Li, Joshua, et al.
Pubblicazione: (2025) -
Learning to Robustly Reconstruct Low-light Dynamic Scenes from Spike Streams
di: Hu, Liwen, et al.
Pubblicazione: (2024) -
Instant Gaussian Stream: Fast and Generalizable Streaming of Dynamic Scene Reconstruction via Gaussian Splatting
di: Yan, Jinbo, et al.
Pubblicazione: (2025) -
RelayGS: Reconstructing Dynamic Scenes with Large-Scale and Complex Motions via Relay Gaussians
di: Gao, Qiankun, et al.
Pubblicazione: (2024) -
StreamCacheVGGT: Streaming Visual Geometry Transformers with Robust Scoring and Hybrid Cache Compression
di: Liu, Xuanyi, et al.
Pubblicazione: (2026)