Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Chung-Shien Brian, Schmidt, Christian, Piekenbrinck, Jens, Leibe, Bastian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Look Gauss, No Pose: Novel View Synthesis using Gaussian Splatting without Accurate Pose Initialization
by: Schmidt, Christian, et al.
Published: (2024)
by: Schmidt, Christian, et al.
Published: (2024)
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting
by: Piekenbrinck, Jens, et al.
Published: (2025)
by: Piekenbrinck, Jens, et al.
Published: (2025)
SurGe: Improved Surface Geometry in Point Maps
by: Knaebel, Karim, et al.
Published: (2026)
by: Knaebel, Karim, et al.
Published: (2026)
Mask4Former: Mask Transformer for 4D Panoptic Segmentation
by: Yilmaz, Kadir, et al.
Published: (2023)
by: Yilmaz, Kadir, et al.
Published: (2023)
Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding
by: Yilmaz, Kadir, et al.
Published: (2026)
by: Yilmaz, Kadir, et al.
Published: (2026)
OCCUQ: Exploring Efficient Uncertainty Quantification for 3D Occupancy Prediction
by: Heidrich, Severin, et al.
Published: (2025)
by: Heidrich, Severin, et al.
Published: (2025)
DONUT: A Decoder-Only Model for Trajectory Prediction
by: Knoche, Markus, et al.
Published: (2025)
by: Knoche, Markus, et al.
Published: (2025)
Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think
by: Garcia, Gonzalo Martin, et al.
Published: (2024)
by: Garcia, Gonzalo Martin, et al.
Published: (2024)
AGILE3D: Attention Guided Interactive Multi-object 3D Segmentation
by: Yue, Yuanwen, et al.
Published: (2023)
by: Yue, Yuanwen, et al.
Published: (2023)
Point2Vec for Self-Supervised Representation Learning on Point Clouds
by: Knaebel, Karim, et al.
Published: (2023)
by: Knaebel, Karim, et al.
Published: (2023)
Epipolar Attention Field Transformers for Bird's Eye View Semantic Segmentation
by: Witte, Christian, et al.
Published: (2024)
by: Witte, Christian, et al.
Published: (2024)
Learning Fine-Grained Geometry for Sparse-View Splatting via Cascade Depth Loss
by: Lu, Wenjun, et al.
Published: (2025)
by: Lu, Wenjun, et al.
Published: (2025)
Point-VOS: Pointing Up Video Object Segmentation
by: Zulfikar, Idil Esen, et al.
Published: (2024)
by: Zulfikar, Idil Esen, et al.
Published: (2024)
Efficient Long-Context Modeling in Diffusion Language Models via Block Approximate Sparse Attention
by: Zhang, Wenhu, et al.
Published: (2026)
by: Zhang, Wenhu, et al.
Published: (2026)
FlashVGGT: Efficient and Scalable Visual Geometry Transformers with Compressed Descriptor Attention
by: Wang, Zipeng, et al.
Published: (2025)
by: Wang, Zipeng, et al.
Published: (2025)
Acquisition of high-quality images for camera calibration in robotics applications via speech prompts
by: Linder, Timm, et al.
Published: (2025)
by: Linder, Timm, et al.
Published: (2025)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
by: Miyato, Takeru, et al.
Published: (2023)
by: Miyato, Takeru, et al.
Published: (2023)
Multi-View Large Reconstruction Model via Geometry-Aware Positional Encoding and Attention
by: Li, Mengfei, et al.
Published: (2024)
by: Li, Mengfei, et al.
Published: (2024)
Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers
by: Zhou, Yifan, et al.
Published: (2025)
by: Zhou, Yifan, et al.
Published: (2025)
PISA: Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers
by: Li, Haopeng, et al.
Published: (2026)
by: Li, Haopeng, et al.
Published: (2026)
FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views
by: Zhang, Shangzhan, et al.
Published: (2025)
by: Zhang, Shangzhan, et al.
Published: (2025)
XAttention: Block Sparse Attention with Antidiagonal Scoring
by: Xu, Ruyi, et al.
Published: (2025)
by: Xu, Ruyi, et al.
Published: (2025)
Intrinsic Geometry-Appearance Consistency Optimization for Sparse-View Gaussian Splatting
by: Xiong, Kaiqiang, et al.
Published: (2026)
by: Xiong, Kaiqiang, et al.
Published: (2026)
Sa2VA-i: Improving Sa2VA Results with Consistent Training and Inference
by: Nekrasov, Alexey, et al.
Published: (2025)
by: Nekrasov, Alexey, et al.
Published: (2025)
How Important are Videos for Training Video LLMs?
by: Lydakis, George, et al.
Published: (2025)
by: Lydakis, George, et al.
Published: (2025)
Interactive4D: Interactive 4D LiDAR Segmentation
by: Fradlin, Ilya, et al.
Published: (2024)
by: Fradlin, Ilya, et al.
Published: (2024)
Prism: Spectral-Aware Block-Sparse Attention
by: Wang, Xinghao, et al.
Published: (2026)
by: Wang, Xinghao, et al.
Published: (2026)
GeoQuery: Geometry-Query Diffusion for Sparse-View Reconstruction
by: Cao, Xiao, et al.
Published: (2026)
by: Cao, Xiao, et al.
Published: (2026)
RainFusion2.0: Temporal-Spatial Awareness and Hardware-Efficient Block-wise Sparse Attention
by: Chen, Aiyue, et al.
Published: (2025)
by: Chen, Aiyue, et al.
Published: (2025)
Disentangled Geometry and Appearance for Efficient Multi-View Surface Reconstruction and Rendering
by: Zhang, Qitong, et al.
Published: (2025)
by: Zhang, Qitong, et al.
Published: (2025)
Spotting the Unexpected (STU): A 3D LiDAR Dataset for Anomaly Segmentation in Autonomous Driving
by: Nekrasov, Alexey, et al.
Published: (2025)
by: Nekrasov, Alexey, et al.
Published: (2025)
OoDIS: Anomaly Instance Segmentation and Detection Benchmark
by: Nekrasov, Alexey, et al.
Published: (2024)
by: Nekrasov, Alexey, et al.
Published: (2024)
GaMO: Geometry-aware Multi-view Diffusion Outpainting for Sparse-View 3D Reconstruction
by: Huang, Yi-Chuan, et al.
Published: (2025)
by: Huang, Yi-Chuan, et al.
Published: (2025)
Sparse2DGS: Geometry-Prioritized Gaussian Splatting for Surface Reconstruction from Sparse Views
by: Wu, Jiang, et al.
Published: (2025)
by: Wu, Jiang, et al.
Published: (2025)
CT-MVSNet: Efficient Multi-View Stereo with Cross-scale Transformer
by: Wang, Sicheng, et al.
Published: (2023)
by: Wang, Sicheng, et al.
Published: (2023)
Neural Surface Reconstruction from Sparse Views Using Epipolar Geometry
by: Chang, Xinhai, et al.
Published: (2024)
by: Chang, Xinhai, et al.
Published: (2024)
Emergent Outlier View Rejection in Visual Geometry Grounded Transformers
by: Han, Jisang, et al.
Published: (2025)
by: Han, Jisang, et al.
Published: (2025)
TranSplat: Generalizable 3D Gaussian Splatting from Sparse Multi-View Images with Transformers
by: Zhang, Chuanrui, et al.
Published: (2024)
by: Zhang, Chuanrui, et al.
Published: (2024)
TouchMap-OR: Multi-View 3D Mapping of Hand-Surface Contacts
by: Ktistakis, Sophokles, et al.
Published: (2026)
by: Ktistakis, Sophokles, et al.
Published: (2026)
Sparser Block-Sparse Attention via Token Permutation
by: Wang, Xinghao, et al.
Published: (2025)
by: Wang, Xinghao, et al.
Published: (2025)
Similar Items
-
Look Gauss, No Pose: Novel View Synthesis using Gaussian Splatting without Accurate Pose Initialization
by: Schmidt, Christian, et al.
Published: (2024) -
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting
by: Piekenbrinck, Jens, et al.
Published: (2025) -
SurGe: Improved Surface Geometry in Point Maps
by: Knaebel, Karim, et al.
Published: (2026) -
Mask4Former: Mask Transformer for 4D Panoptic Segmentation
by: Yilmaz, Kadir, et al.
Published: (2023) -
Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding
by: Yilmaz, Kadir, et al.
Published: (2026)