$\text{VG}^2$GT: Voxel-Gaussian Splatting Visual Geometry Grounded Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yibin, Pan, Yihan, Nan, Jun, Yang, Wenli, Chen, Liwei, Yi, Jianjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FSFSplatter: Build Surface and Novel Views with Sparse-Views within 2min
von: Zhao, Yibin, et al.
Veröffentlicht: (2025)
von: Zhao, Yibin, et al.
Veröffentlicht: (2025)
VG3S: Visual Geometry Grounded Gaussian Splatting for Semantic Occupancy Prediction
von: Yan, Xiaoyang, et al.
Veröffentlicht: (2026)
von: Yan, Xiaoyang, et al.
Veröffentlicht: (2026)
VG3T: Visual Geometry Grounded Gaussian Transformer
von: Kim, Junho, et al.
Veröffentlicht: (2025)
von: Kim, Junho, et al.
Veröffentlicht: (2025)
GT2-GS: Geometry-aware Texture Transfer for Gaussian Splatting
von: Liu, Wenjie, et al.
Veröffentlicht: (2025)
von: Liu, Wenjie, et al.
Veröffentlicht: (2025)
Geometry-Grounded Gaussian Splatting
von: Zhang, Baowen, et al.
Veröffentlicht: (2026)
von: Zhang, Baowen, et al.
Veröffentlicht: (2026)
4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting
von: Wu, Chun-Tin, et al.
Veröffentlicht: (2025)
von: Wu, Chun-Tin, et al.
Veröffentlicht: (2025)
TrajVG: 3D Trajectory-Coupled Visual Geometry Learning
von: Miao, Xingyu, et al.
Veröffentlicht: (2026)
von: Miao, Xingyu, et al.
Veröffentlicht: (2026)
HiVG: Hierarchical Multimodal Fine-grained Modulation for Visual Grounding
von: Xiao, Linhui, et al.
Veröffentlicht: (2024)
von: Xiao, Linhui, et al.
Veröffentlicht: (2024)
CLIP-VG: Self-paced Curriculum Adapting of CLIP for Visual Grounding
von: Xiao, Linhui, et al.
Veröffentlicht: (2023)
von: Xiao, Linhui, et al.
Veröffentlicht: (2023)
VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow Prediction
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
von: Zhu, Ziyue, et al.
Veröffentlicht: (2025)
ResVG: Enhancing Relation and Semantic Understanding in Multiple Instances for Visual Grounding
von: Zheng, Minghang, et al.
Veröffentlicht: (2024)
von: Zheng, Minghang, et al.
Veröffentlicht: (2024)
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
von: Kang, Weitai, et al.
Veröffentlicht: (2024)
SwimVG: Step-wise Multimodal Fusion and Adaption for Visual Grounding
von: Shi, Liangtao, et al.
Veröffentlicht: (2025)
von: Shi, Liangtao, et al.
Veröffentlicht: (2025)
PathVG: A New Benchmark and Dataset for Pathology Visual Grounding
von: Zhong, Chunlin, et al.
Veröffentlicht: (2025)
von: Zhong, Chunlin, et al.
Veröffentlicht: (2025)
VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
VGGT: Visual Geometry Grounded Transformer
von: Wang, Jianyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jianyuan, et al.
Veröffentlicht: (2025)
QVGGT: Post-Training Quantized Visual Geometry Grounded Transformer
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
von: Pan, Zhizhen, et al.
Veröffentlicht: (2026)
Quantized Visual Geometry Grounded Transformer
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
von: Feng, Weilun, et al.
Veröffentlicht: (2025)
PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination
von: Dai, Ming, et al.
Veröffentlicht: (2025)
von: Dai, Ming, et al.
Veröffentlicht: (2025)
InfiniteVGGT: Visual Geometry Grounded Transformer for Endless Streams
von: Yuan, Shuai, et al.
Veröffentlicht: (2026)
von: Yuan, Shuai, et al.
Veröffentlicht: (2026)
AerialVG: A Challenging Benchmark for Aerial Visual Grounding by Exploring Positional Relations
von: Liu, Junli, et al.
Veröffentlicht: (2025)
von: Liu, Junli, et al.
Veröffentlicht: (2025)
SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion
von: Dai, Ming, et al.
Veröffentlicht: (2024)
von: Dai, Ming, et al.
Veröffentlicht: (2024)
LLM4VG: Large Language Models Evaluation for Video Grounding
von: Feng, Wei, et al.
Veröffentlicht: (2023)
von: Feng, Wei, et al.
Veröffentlicht: (2023)
ProVG: Progressive Visual Grounding via Language Decoupling for Remote Sensing Imagery
von: Li, Ke, et al.
Veröffentlicht: (2026)
von: Li, Ke, et al.
Veröffentlicht: (2026)
UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning
von: Bai, Sule, et al.
Veröffentlicht: (2025)
von: Bai, Sule, et al.
Veröffentlicht: (2025)
GeM-VG: Towards Generalized Multi-image Visual Grounding with Multimodal Large Language Models
von: Zheng, Shurong, et al.
Veröffentlicht: (2026)
von: Zheng, Shurong, et al.
Veröffentlicht: (2026)
G4Splat: Geometry-Guided Gaussian Splatting with Generative Prior
von: Ni, Junfeng, et al.
Veröffentlicht: (2025)
von: Ni, Junfeng, et al.
Veröffentlicht: (2025)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
von: Wu, Guile, et al.
Veröffentlicht: (2026)
von: Wu, Guile, et al.
Veröffentlicht: (2026)
AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding
von: Li, Haocheng, et al.
Veröffentlicht: (2026)
von: Li, Haocheng, et al.
Veröffentlicht: (2026)
VG-TVP: Multimodal Procedural Planning via Visually Grounded Text-Video Prompting
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting
von: Tran, Anh Thuan, et al.
Veröffentlicht: (2026)
von: Tran, Anh Thuan, et al.
Veröffentlicht: (2026)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
Decoupling Motion and Geometry in 4D Gaussian Splatting
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
Emergent Outlier View Rejection in Visual Geometry Grounded Transformers
von: Han, Jisang, et al.
Veröffentlicht: (2025)
von: Han, Jisang, et al.
Veröffentlicht: (2025)
Reloc-VGGT: Visual Re-localization with Geometry Grounded Transformer
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
WaterVG: Waterway Visual Grounding based on Text-Guided Vision and mmWave Radar
von: Guan, Runwei, et al.
Veröffentlicht: (2024)
von: Guan, Runwei, et al.
Veröffentlicht: (2024)
GeoSplatting: Towards Geometry Guided Gaussian Splatting for Physically-based Inverse Rendering
von: Ye, Kai, et al.
Veröffentlicht: (2024)
von: Ye, Kai, et al.
Veröffentlicht: (2024)
SG-Splatting: Accelerating 3D Gaussian Splatting with Spherical Gaussians
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
von: Wang, Yiwen, et al.
Veröffentlicht: (2024)
EG-Gaussian: Epipolar Geometry and Graph Network Enhanced 3D Gaussian Splatting
von: Zhao, Beizhen, et al.
Veröffentlicht: (2025)
von: Zhao, Beizhen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FSFSplatter: Build Surface and Novel Views with Sparse-Views within 2min
von: Zhao, Yibin, et al.
Veröffentlicht: (2025) -
VG3S: Visual Geometry Grounded Gaussian Splatting for Semantic Occupancy Prediction
von: Yan, Xiaoyang, et al.
Veröffentlicht: (2026) -
VG3T: Visual Geometry Grounded Gaussian Transformer
von: Kim, Junho, et al.
Veröffentlicht: (2025) -
GT2-GS: Geometry-aware Texture Transfer for Gaussian Splatting
von: Liu, Wenjie, et al.
Veröffentlicht: (2025) -
Geometry-Grounded Gaussian Splatting
von: Zhang, Baowen, et al.
Veröffentlicht: (2026)