Scenes as Tokens: Multi-Scale Normal Distributions Transform Tokenizer for General 3D Vision-Language Understanding

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Tang, Yutao, Zhao, Cheng, Mittal, Gaurav, Kukkala, Rohith, Chellappa, Rama, Peng, Cheng, Chen, Mei
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!