Online Language Splatting
Fuente:
arXiv
Saved in:
| Main Authors: | Katragadda, Saimouli, Wu, Cho-Ying, Guo, Yuliang, Huang, Xinyu, Huang, Guoquan, Ren, Liu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VIMD: Monocular Visual-Inertial Motion and Depth Estimation
by: Katragadda, Saimouli, et al.
Published: (2025)
by: Katragadda, Saimouli, et al.
Published: (2025)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
by: Guo, Yuliang, et al.
Published: (2025)
by: Guo, Yuliang, et al.
Published: (2025)
Large-Scale Gaussian Splatting SLAM
by: Xin, Zhe, et al.
Published: (2025)
by: Xin, Zhe, et al.
Published: (2025)
MapGS: Generalizable Pretraining and Data Augmentation for Online Mapping via Novel View Synthesis
by: Zhang, Hengyuan, et al.
Published: (2025)
by: Zhang, Hengyuan, et al.
Published: (2025)
SLAG: Scalable Language-Augmented Gaussian Splatting
by: Szilagyi, Laszlo, et al.
Published: (2025)
by: Szilagyi, Laszlo, et al.
Published: (2025)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
by: Ji, Yihong, et al.
Published: (2025)
by: Ji, Yihong, et al.
Published: (2025)
Learning IMU Bias with Diffusion Model
by: Zhou, Shenghao, et al.
Published: (2025)
by: Zhou, Shenghao, et al.
Published: (2025)
SeaBird: Segmentation in Bird's View with Dice Loss Improves Monocular 3D Detection of Large Objects
by: Kumar, Abhinav, et al.
Published: (2024)
by: Kumar, Abhinav, et al.
Published: (2024)
Integrating Object Detection Modality into Visual Language Model for Enhanced Autonomous Driving Agent
by: He, Linfeng, et al.
Published: (2024)
by: He, Linfeng, et al.
Published: (2024)
Enhancing Online Road Network Perception and Reasoning with Standard Definition Maps
by: Zhang, Hengyuan, et al.
Published: (2024)
by: Zhang, Hengyuan, et al.
Published: (2024)
Online,Target-Free LiDAR-Camera Extrinsic Calibration via Cross-Modal Mask Matching
by: Huang, Zhiwei, et al.
Published: (2024)
by: Huang, Zhiwei, et al.
Published: (2024)
SGS-SLAM: Semantic Gaussian Splatting For Neural Dense SLAM
by: Li, Mingrui, et al.
Published: (2024)
by: Li, Mingrui, et al.
Published: (2024)
BEACON: Language-Conditioned Navigation Affordance Prediction under Occlusion
by: Gao, Xinyu, et al.
Published: (2026)
by: Gao, Xinyu, et al.
Published: (2026)
Environment-Driven Online LiDAR-Camera Extrinsic Calibration
by: Huang, Zhiwei, et al.
Published: (2025)
by: Huang, Zhiwei, et al.
Published: (2025)
Gradient-Guided Parameter Mask for Multi-Scenario Image Restoration Under Adverse Weather
by: Guo, Jilong, et al.
Published: (2024)
by: Guo, Jilong, et al.
Published: (2024)
Dino-Diffusion Modular Designs Bridge the Cross-Domain Gap in Autonomous Parking
by: Wu, Zixuan, et al.
Published: (2025)
by: Wu, Zixuan, et al.
Published: (2025)
3D Scene Rendering with Multimodal Gaussian Splatting
by: Gau, Chi-Shiang, et al.
Published: (2026)
by: Gau, Chi-Shiang, et al.
Published: (2026)
Efficient Driving Behavior Narration and Reasoning on Edge Device Using Large Language Models
by: Huang, Yizhou, et al.
Published: (2024)
by: Huang, Yizhou, et al.
Published: (2024)
SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2025)
by: Zhang, Haiming, et al.
Published: (2025)
Latent Gaussian Splatting for 4D Panoptic Occupancy Tracking
by: Luz, Maximilian, et al.
Published: (2026)
by: Luz, Maximilian, et al.
Published: (2026)
Language-Conditioned World Modeling for Visual Navigation
by: Dong, Yifei, et al.
Published: (2026)
by: Dong, Yifei, et al.
Published: (2026)
SynHLMA:Synthesizing Hand Language Manipulation for Articulated Object with Discrete Human Object Interaction Representation
by: zhi, Wang, et al.
Published: (2025)
by: zhi, Wang, et al.
Published: (2025)
Complete Gaussian Splats from a Single Image with Denoising Diffusion Models
by: Liao, Ziwei, et al.
Published: (2025)
by: Liao, Ziwei, et al.
Published: (2025)
SIMSplat: Predictive Driving Scene Editing with Language-aligned 4D Gaussian Splatting
by: Park, Sung-Yeon, et al.
Published: (2025)
by: Park, Sung-Yeon, et al.
Published: (2025)
MM3DGS SLAM: Multi-modal 3D Gaussian Splatting for SLAM Using Vision, Depth, and Inertial Measurements
by: Sun, Lisong C., et al.
Published: (2024)
by: Sun, Lisong C., et al.
Published: (2024)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
by: Huang, Zilin, et al.
Published: (2024)
by: Huang, Zilin, et al.
Published: (2024)
SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM
by: Keetha, Nikhil, et al.
Published: (2023)
by: Keetha, Nikhil, et al.
Published: (2023)
SkiP: When to Skip and When to Refine for Efficient Robot Manipulation
by: Dai, Mingtong, et al.
Published: (2026)
by: Dai, Mingtong, et al.
Published: (2026)
FGS-SLAM: Fourier-based Gaussian Splatting for Real-time SLAM with Sparse and Dense Map Fusion
by: Xu, Yansong, et al.
Published: (2025)
by: Xu, Yansong, et al.
Published: (2025)
VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models
by: Gao, Chongkai, et al.
Published: (2025)
by: Gao, Chongkai, et al.
Published: (2025)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
by: Guo, Jianing, et al.
Published: (2025)
by: Guo, Jianing, et al.
Published: (2025)
CurricuVLM: Towards Safe Autonomous Driving via Personalized Safety-Critical Curriculum Learning with Vision-Language Models
by: Sheng, Zihao, et al.
Published: (2025)
by: Sheng, Zihao, et al.
Published: (2025)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
by: Liu, Chenyv, et al.
Published: (2026)
by: Liu, Chenyv, et al.
Published: (2026)
DriveVLM-RL: Neuroscience-Inspired Reinforcement Learning with Vision-Language Models for Safe and Deployable Autonomous Driving
by: Huang, Zilin, et al.
Published: (2026)
by: Huang, Zilin, et al.
Published: (2026)
SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling
by: You, Junwei, et al.
Published: (2025)
by: You, Junwei, et al.
Published: (2025)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
by: Dharmarajan, Karthik, et al.
Published: (2025)
by: Dharmarajan, Karthik, et al.
Published: (2025)
V2X-QA: A Comprehensive Reasoning Dataset and Benchmark for Multimodal Large Language Models in Autonomous Driving Across Ego, Infrastructure, and Cooperative Views
by: You, Junwei, et al.
Published: (2026)
by: You, Junwei, et al.
Published: (2026)
SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning
by: Liu, Yuecheng, et al.
Published: (2025)
by: Liu, Yuecheng, et al.
Published: (2025)
SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation
by: Li, Xin, et al.
Published: (2024)
by: Li, Xin, et al.
Published: (2024)
Multi-modal Situated Reasoning in 3D Scenes
by: Linghu, Xiongkun, et al.
Published: (2024)
by: Linghu, Xiongkun, et al.
Published: (2024)
Similar Items
-
VIMD: Monocular Visual-Inertial Motion and Depth Estimation
by: Katragadda, Saimouli, et al.
Published: (2025) -
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
by: Guo, Yuliang, et al.
Published: (2025) -
Large-Scale Gaussian Splatting SLAM
by: Xin, Zhe, et al.
Published: (2025) -
MapGS: Generalizable Pretraining and Data Augmentation for Online Mapping via Novel View Synthesis
by: Zhang, Hengyuan, et al.
Published: (2025) -
SLAG: Scalable Language-Augmented Gaussian Splatting
by: Szilagyi, Laszlo, et al.
Published: (2025)