Bridging the Dimensionality Gap: A Taxonomy and Survey of 2D Vision Model Adaptation for 3D Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pandya, Akshat, Jain, Bhavuk |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SIGMA: Bridging Structural and Distributional Gaps for Vision Foundation Model Adaptation
von: Xiong, Lingyu, et al.
Veröffentlicht: (2026)
von: Xiong, Lingyu, et al.
Veröffentlicht: (2026)
SemanticBridge - A Dataset for 3D Semantic Segmentation of Bridges and Domain Gap Analysis
von: Kellner, Maximilian, et al.
Veröffentlicht: (2025)
von: Kellner, Maximilian, et al.
Veröffentlicht: (2025)
Diffusion Models in 3D Vision: A Survey
von: Wang, Zhen, et al.
Veröffentlicht: (2024)
von: Wang, Zhen, et al.
Veröffentlicht: (2024)
Feature Fusion Attention Network with CycleGAN for Image Dehazing, De-Snowing and De-Raining
von: Jain, Akshat
Veröffentlicht: (2025)
von: Jain, Akshat
Veröffentlicht: (2025)
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
von: Wang, Wei, et al.
Veröffentlicht: (2026)
von: Wang, Wei, et al.
Veröffentlicht: (2026)
Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation
von: Li, Kailing, et al.
Veröffentlicht: (2026)
von: Li, Kailing, et al.
Veröffentlicht: (2026)
X-Dreamer: Creating High-quality 3D Content by Bridging the Domain Gap Between Text-to-2D and Text-to-3D Generation
von: Ma, Yiwei, et al.
Veröffentlicht: (2023)
von: Ma, Yiwei, et al.
Veröffentlicht: (2023)
DPGLA: Bridging the Gap between Synthetic and Real Data for Unsupervised Domain Adaptation in 3D LiDAR Semantic Segmentation
von: Li, Wanmeng, et al.
Veröffentlicht: (2025)
von: Li, Wanmeng, et al.
Veröffentlicht: (2025)
GENA3D: Generative Amodal 3D Modeling by Bridging 2D Priors and 3D Coherence
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
von: Zhou, Junwei, et al.
Veröffentlicht: (2025)
PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models
von: Li, Yuliang, et al.
Veröffentlicht: (2026)
von: Li, Yuliang, et al.
Veröffentlicht: (2026)
Learning 3D Reconstruction with Priors in Test Time
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
Bridging the Skill Gap in Clinical CBCT Interpretation with CBCTRepD
von: Wu, Qinxin, et al.
Veröffentlicht: (2026)
von: Wu, Qinxin, et al.
Veröffentlicht: (2026)
Bridging the Domain Gap for Flight-Ready Spaceborne Vision
von: Park, Tae Ha, et al.
Veröffentlicht: (2024)
von: Park, Tae Ha, et al.
Veröffentlicht: (2024)
Mamba2D: A Natively Multi-Dimensional State-Space Model for Vision Tasks
von: Baty, Enis, et al.
Veröffentlicht: (2024)
von: Baty, Enis, et al.
Veröffentlicht: (2024)
Earth-Adapter: Bridge the Geospatial Domain Gaps with Mixture of Frequency Adaptation
von: Hu, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Hu, Xiaoxing, et al.
Veröffentlicht: (2025)
Prompt-based Adaptation in Large-scale Vision Models: A Survey
von: Xiao, Xi, et al.
Veröffentlicht: (2025)
von: Xiao, Xi, et al.
Veröffentlicht: (2025)
Bridging the Gap: Doubles Badminton Analysis with Singles-Trained Models
von: Baek, Seungheon, et al.
Veröffentlicht: (2025)
von: Baek, Seungheon, et al.
Veröffentlicht: (2025)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
von: T, Mukund Varma, et al.
Veröffentlicht: (2024)
von: T, Mukund Varma, et al.
Veröffentlicht: (2024)
BridgeNet: A Unified Multimodal Framework for Bridging 2D and 3D Industrial Anomaly Detection
von: Xiang, An, et al.
Veröffentlicht: (2025)
von: Xiang, An, et al.
Veröffentlicht: (2025)
Bridging the Vision-Brain Gap with an Uncertainty-Aware Blur Prior
von: Wu, Haitao, et al.
Veröffentlicht: (2025)
von: Wu, Haitao, et al.
Veröffentlicht: (2025)
MITA: Bridging the Gap between Model and Data for Test-time Adaptation
von: Yuan, Yige, et al.
Veröffentlicht: (2024)
von: Yuan, Yige, et al.
Veröffentlicht: (2024)
SpectraDINO: Bridging the Spectral Gap in Vision Foundation Models via Lightweight Adapters
von: Nalcakan, Yagiz, et al.
Veröffentlicht: (2026)
von: Nalcakan, Yagiz, et al.
Veröffentlicht: (2026)
Taxonomy-Aware Evaluation of Vision-Language Models
von: Snæbjarnarson, Vésteinn, et al.
Veröffentlicht: (2025)
von: Snæbjarnarson, Vésteinn, et al.
Veröffentlicht: (2025)
From 2D to 3D Cognition: A Brief Survey of General World Models
von: Xie, Ningwei, et al.
Veröffentlicht: (2025)
von: Xie, Ningwei, et al.
Veröffentlicht: (2025)
Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation
von: Duangprom, Krit, et al.
Veröffentlicht: (2025)
von: Duangprom, Krit, et al.
Veröffentlicht: (2025)
Bridge the Modality and Capability Gaps in Vision-Language Model Selection
von: Yi, Chao, et al.
Veröffentlicht: (2024)
von: Yi, Chao, et al.
Veröffentlicht: (2024)
SSDA: Bridging Spectral and Structural Gaps via Dual Adaptation for Vision-Based Time Series Forecasting
von: Zhang, Mingrui, et al.
Veröffentlicht: (2026)
von: Zhang, Mingrui, et al.
Veröffentlicht: (2026)
Bridging Diffusion Models and 3D Representations: A 3D Consistent Super-Resolution Framework
von: Chen, Yi-Ting, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Ting, et al.
Veröffentlicht: (2025)
Unifying 2D and 3D Vision-Language Understanding
von: Jain, Ayush, et al.
Veröffentlicht: (2025)
von: Jain, Ayush, et al.
Veröffentlicht: (2025)
R3eVision: A Survey on Robust Rendering, Restoration, and Enhancement for 3D Low-Level Vision
von: Kwon, Weeyoung, et al.
Veröffentlicht: (2025)
von: Kwon, Weeyoung, et al.
Veröffentlicht: (2025)
VLPose: Bridging the Domain Gap in Pose Estimation with Language-Vision Tuning
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
3D Question Answering via only 2D Vision-Language Models
von: Wang, Fengyun, et al.
Veröffentlicht: (2025)
von: Wang, Fengyun, et al.
Veröffentlicht: (2025)
Bridging Domain Gaps for Fine-Grained Moth Classification Through Expert-Informed Adaptation and Foundation Model Priors
von: Gardiner, Ross J, et al.
Veröffentlicht: (2025)
von: Gardiner, Ross J, et al.
Veröffentlicht: (2025)
A Survey of Low-shot Vision-Language Model Adaptation via Representer Theorem
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
PartSTAD: 2D-to-3D Part Segmentation Task Adaptation
von: Kim, Hyunjin, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjin, et al.
Veröffentlicht: (2024)
Reasoning in Computer Vision: Taxonomy, Models, Tasks, and Methodologies
von: Sarkar, Ayushman, et al.
Veröffentlicht: (2025)
von: Sarkar, Ayushman, et al.
Veröffentlicht: (2025)
3D and 4D World Modeling: A Survey
von: Kong, Lingdong, et al.
Veröffentlicht: (2025)
von: Kong, Lingdong, et al.
Veröffentlicht: (2025)
From 2D CAD Drawings to 3D Parametric Models: A Vision-Language Approach
von: Wang, Xilin, et al.
Veröffentlicht: (2024)
von: Wang, Xilin, et al.
Veröffentlicht: (2024)
Vision Mamba: A Comprehensive Survey and Taxonomy
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA
von: Mo, Wentao, et al.
Veröffentlicht: (2024)
von: Mo, Wentao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SIGMA: Bridging Structural and Distributional Gaps for Vision Foundation Model Adaptation
von: Xiong, Lingyu, et al.
Veröffentlicht: (2026) -
SemanticBridge - A Dataset for 3D Semantic Segmentation of Bridges and Domain Gap Analysis
von: Kellner, Maximilian, et al.
Veröffentlicht: (2025) -
Diffusion Models in 3D Vision: A Survey
von: Wang, Zhen, et al.
Veröffentlicht: (2024) -
Feature Fusion Attention Network with CycleGAN for Image Dehazing, De-Snowing and De-Raining
von: Jain, Akshat
Veröffentlicht: (2025) -
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
von: Wang, Wei, et al.
Veröffentlicht: (2026)