TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Fengyi, Zhang, Tianjun, Khosoussi, Kasra, Zhang, Zheng, Huang, Zi, Luo, Yadan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GaussianForest: Hierarchical-Hybrid 3D Gaussian Splatting for Compressed Scene Modeling
von: Zhang, Fengyi, et al.
Veröffentlicht: (2024)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2024)
Test-Time 3D Occupancy Prediction
von: Zhang, Fengyi, et al.
Veröffentlicht: (2025)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2025)
CodeMerge: Codebook-Guided Model Merging for Robust Test-Time Adaptation in Autonomous Driving
von: Yang, Huitong, et al.
Veröffentlicht: (2025)
von: Yang, Huitong, et al.
Veröffentlicht: (2025)
VGGT-World: Transforming VGGT into an Autoregressive Geometry World Model
von: Sun, Xiangyu, et al.
Veröffentlicht: (2026)
von: Sun, Xiangyu, et al.
Veröffentlicht: (2026)
Find n' Propagate: Open-Vocabulary 3D Object Detection in Urban Environments
von: Etchegaray, Djamahl, et al.
Veröffentlicht: (2024)
von: Etchegaray, Djamahl, et al.
Veröffentlicht: (2024)
Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model
von: Zhou, Shunkai, et al.
Veröffentlicht: (2026)
von: Zhou, Shunkai, et al.
Veröffentlicht: (2026)
Consistency Diffusion Models for Single-Image 3D Reconstruction with Priors
von: Jiang, Chenru, et al.
Veröffentlicht: (2025)
von: Jiang, Chenru, et al.
Veröffentlicht: (2025)
Provable Ordering and Continuity in Vision-Language Pretraining for Generalizable Embodied Agents
von: Zhang, Zhizhen, et al.
Veröffentlicht: (2025)
von: Zhang, Zhizhen, et al.
Veröffentlicht: (2025)
MOS: Model Synergy for Test-Time Adaptation on LiDAR-Based 3D Object Detection
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
DPO: Dual-Perturbation Optimization for Test-time Adaptation in 3D Object Detection
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2024)
Open-CRB: Towards Open World Active Learning for 3D Object Detection
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2023)
von: Chen, Zhuoxiao, et al.
Veröffentlicht: (2023)
FastEdit: Fast Text-Guided Single-Image Editing via Semantic-Aware Diffusion Fine-Tuning
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
PolaFormer: Polarity-aware Linear Attention for Vision Transformers
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
Language-driven Fine-grained Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
Pushing Auto-regressive Models for 3D Shape Generation at Capacity and Scalability
von: Qian, Xuelin, et al.
Veröffentlicht: (2024)
von: Qian, Xuelin, et al.
Veröffentlicht: (2024)
Towards Training-free Anomaly Detection with Vision and Language Foundation Models
von: Zhang, Jinjin, et al.
Veröffentlicht: (2025)
von: Zhang, Jinjin, et al.
Veröffentlicht: (2025)
CF-PRNet: Coarse-to-Fine Prototype Refining Network for Point Cloud Completion and Reconstruction
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
Towards Generalist Intelligence in Dentistry: Vision Foundation Models for Oral and Maxillofacial Radiology
von: Huang, Xinrui, et al.
Veröffentlicht: (2025)
von: Huang, Xinrui, et al.
Veröffentlicht: (2025)
Box-QAymo: Box-Referring VQA Dataset for Autonomous Driving
von: Etchegaray, Djamahl, et al.
Veröffentlicht: (2025)
von: Etchegaray, Djamahl, et al.
Veröffentlicht: (2025)
ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection
von: Song, Ziying, et al.
Veröffentlicht: (2024)
von: Song, Ziying, et al.
Veröffentlicht: (2024)
Towards Unbiased Source-Free Object Detection via Vision Foundation Models
von: Cai, Zhi, et al.
Veröffentlicht: (2026)
von: Cai, Zhi, et al.
Veröffentlicht: (2026)
Learning to Align Generative Appearance Priors for Fine-grained Image Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
SCORE: Soft Label Compression-Centric Dataset Condensation via Coding Rate Optimization
von: Yuan, Bowen, et al.
Veröffentlicht: (2025)
von: Yuan, Bowen, et al.
Veröffentlicht: (2025)
Color-Oriented Redundancy Reduction in Dataset Distillation
von: Yuan, Bowen, et al.
Veröffentlicht: (2024)
von: Yuan, Bowen, et al.
Veröffentlicht: (2024)
Towards Foundation Models for 3D Vision: How Close Are We?
von: Zuo, Yiming, et al.
Veröffentlicht: (2024)
von: Zuo, Yiming, et al.
Veröffentlicht: (2024)
One for All: Toward Unified Foundation Models for Earth Vision
von: Xiong, Zhitong, et al.
Veröffentlicht: (2024)
von: Xiong, Zhitong, et al.
Veröffentlicht: (2024)
Hierarchical Refinement of Universal Multimodal Attacks on Vision-Language Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2026)
FedStain: Modeling Higher-Order Stain Statistics for Federated Domain Generalization in Computational Pathology
von: Zhang, Fengyi, et al.
Veröffentlicht: (2026)
von: Zhang, Fengyi, et al.
Veröffentlicht: (2026)
Towards Unified 3D Hair Reconstruction from Single-View Portraits
von: Zheng, Yujian, et al.
Veröffentlicht: (2024)
von: Zheng, Yujian, et al.
Veröffentlicht: (2024)
View Transformation Robustness for Multi-View 3D Object Reconstruction with Reconstruction Error-Guided View Selection
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
GCA-3D: Towards Generalized and Consistent Domain Adaptation of 3D Generators
von: Li, Hengjia, et al.
Veröffentlicht: (2024)
von: Li, Hengjia, et al.
Veröffentlicht: (2024)
EMOv2: Pushing 5M Vision Model Frontier
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
Bi-VLM: Pushing Ultra-Low Precision Post-Training Quantization Boundaries in Vision-Language Models
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
Distributed Zero-Shot Learning for Visual Recognition
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
Learning to Synergize Semantic and Geometric Priors for Limited-Data Wheat Disease Segmentation
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
Divide-and-Conquer Approach to Holistic Cognition in High-Similarity Contexts with Limited Data
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
Geometry-Guided Self-Supervision for Ultra-Fine-Grained Recognition with Limited Data
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging
von: Tang, Yijie, et al.
Veröffentlicht: (2025)
von: Tang, Yijie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GaussianForest: Hierarchical-Hybrid 3D Gaussian Splatting for Compressed Scene Modeling
von: Zhang, Fengyi, et al.
Veröffentlicht: (2024) -
Test-Time 3D Occupancy Prediction
von: Zhang, Fengyi, et al.
Veröffentlicht: (2025) -
CodeMerge: Codebook-Guided Model Merging for Robust Test-Time Adaptation in Autonomous Driving
von: Yang, Huitong, et al.
Veröffentlicht: (2025) -
VGGT-World: Transforming VGGT into an Autoregressive Geometry World Model
von: Sun, Xiangyu, et al.
Veröffentlicht: (2026) -
Find n' Propagate: Open-Vocabulary 3D Object Detection in Urban Environments
von: Etchegaray, Djamahl, et al.
Veröffentlicht: (2024)