SPACE-CLIP: Spatial Perception via Adaptive CLIP Embeddings for Monocular Depth Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Cho, Taewan, Kim, Taeryang, Choi, Andrew Jaeyong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PureCLIP-Depth: Prompt-Free and Decoder-Free Monocular Depth Estimation within CLIP Embedding Space
di: Miya, Ryutaro, et al.
Pubblicazione: (2026)
di: Miya, Ryutaro, et al.
Pubblicazione: (2026)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
di: Kim, Donghyeong, et al.
Pubblicazione: (2025)
di: Kim, Donghyeong, et al.
Pubblicazione: (2025)
CLIP Can Understand Depth
di: Kim, Sohee, et al.
Pubblicazione: (2024)
di: Kim, Sohee, et al.
Pubblicazione: (2024)
AnimalMotionCLIP: Embedding motion in CLIP for Animal Behavior Analysis
di: Zhong, Enmin, et al.
Pubblicazione: (2025)
di: Zhong, Enmin, et al.
Pubblicazione: (2025)
TIE-KD: Teacher-Independent and Explainable Knowledge Distillation for Monocular Depth Estimation
di: Choi, Sangwon, et al.
Pubblicazione: (2024)
di: Choi, Sangwon, et al.
Pubblicazione: (2024)
Real-time Monocular Depth Estimation on Embedded Systems
di: Feng, Cheng, et al.
Pubblicazione: (2023)
di: Feng, Cheng, et al.
Pubblicazione: (2023)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
di: Zhu, Wencheng, et al.
Pubblicazione: (2025)
di: Zhu, Wencheng, et al.
Pubblicazione: (2025)
Lightweight Prompt-Guided CLIP Adaptation for Monocular Depth Estimation
di: Manghotay, Reyhaneh Ahani, et al.
Pubblicazione: (2026)
di: Manghotay, Reyhaneh Ahani, et al.
Pubblicazione: (2026)
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
di: Li, Yinqi, et al.
Pubblicazione: (2025)
di: Li, Yinqi, et al.
Pubblicazione: (2025)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
di: Song, Yehun, et al.
Pubblicazione: (2025)
di: Song, Yehun, et al.
Pubblicazione: (2025)
Beyond Semantics: Disentangling Information Scope in Sparse Autoencoders for CLIP
di: Ro, Yusung, et al.
Pubblicazione: (2026)
di: Ro, Yusung, et al.
Pubblicazione: (2026)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
di: Liu, Mushui, et al.
Pubblicazione: (2024)
di: Liu, Mushui, et al.
Pubblicazione: (2024)
Efficiently Disentangling CLIP for Multi-Object Perception
di: Rawlekar, Samyak, et al.
Pubblicazione: (2025)
di: Rawlekar, Samyak, et al.
Pubblicazione: (2025)
UM-Depth : Uncertainty Masked Self-Supervised Monocular Depth Estimation with Visual Odometry
di: Um, Tae-Wook, et al.
Pubblicazione: (2025)
di: Um, Tae-Wook, et al.
Pubblicazione: (2025)
LiteEmbed: Adapting CLIP to Rare Classes
di: Agarwal, Aishwarya, et al.
Pubblicazione: (2026)
di: Agarwal, Aishwarya, et al.
Pubblicazione: (2026)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
di: Yang, Kaicheng, et al.
Pubblicazione: (2024)
di: Yang, Kaicheng, et al.
Pubblicazione: (2024)
TripletCLIP: Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
di: Patel, Maitreya, et al.
Pubblicazione: (2024)
SuperCLIP: CLIP with Simple Classification Supervision
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
Adaptive Depth-converted-Scale Convolution for Self-supervised Monocular Depth Estimation
di: Gao, Yanbo, et al.
Pubblicazione: (2026)
di: Gao, Yanbo, et al.
Pubblicazione: (2026)
Plane2Depth: Hierarchical Adaptive Plane Guidance for Monocular Depth Estimation
di: Liu, Li, et al.
Pubblicazione: (2024)
di: Liu, Li, et al.
Pubblicazione: (2024)
Impression-CLIP: Contrastive Shape-Impression Embedding for Fonts
di: Kubota, Yugo, et al.
Pubblicazione: (2024)
di: Kubota, Yugo, et al.
Pubblicazione: (2024)
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
di: Ma, Sanggyun, et al.
Pubblicazione: (2025)
di: Ma, Sanggyun, et al.
Pubblicazione: (2025)
Stereo-Matching Knowledge Distilled Monocular Depth Estimation Filtered by Multiple Disparity Consistency
di: Ka, Woonghyun, et al.
Pubblicazione: (2024)
di: Ka, Woonghyun, et al.
Pubblicazione: (2024)
LMM-Regularized CLIP Embeddings for Image Classification
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
CLIP-aware Domain-Adaptive Super-Resolution
di: Lu, Zhengyang, et al.
Pubblicazione: (2025)
di: Lu, Zhengyang, et al.
Pubblicazione: (2025)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
di: Zhang, Beichen, et al.
Pubblicazione: (2024)
di: Zhang, Beichen, et al.
Pubblicazione: (2024)
CLIP-KD: An Empirical Study of CLIP Model Distillation
di: Yang, Chuanguang, et al.
Pubblicazione: (2023)
di: Yang, Chuanguang, et al.
Pubblicazione: (2023)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
di: Karimian, Banafsheh, et al.
Pubblicazione: (2025)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
di: Guo, Yufei, et al.
Pubblicazione: (2023)
di: Guo, Yufei, et al.
Pubblicazione: (2023)
AF-CLIP: Zero-Shot Anomaly Detection via Anomaly-Focused CLIP Adaptation
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels
di: Singh, Darshan, et al.
Pubblicazione: (2024)
di: Singh, Darshan, et al.
Pubblicazione: (2024)
The Third Monocular Depth Estimation Challenge
di: Spencer, Jaime, et al.
Pubblicazione: (2024)
di: Spencer, Jaime, et al.
Pubblicazione: (2024)
PTC-Depth: Pose-Refined Monocular Depth Estimation with Temporal Consistency
di: Han, Leezy, et al.
Pubblicazione: (2026)
di: Han, Leezy, et al.
Pubblicazione: (2026)
Spatially-Weighted CLIP for Street-View Geo-localization
di: Han, Ting, et al.
Pubblicazione: (2026)
di: Han, Ting, et al.
Pubblicazione: (2026)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
Enhancing CLIP with CLIP: Exploring Pseudolabeling for Limited-Label Prompt Tuning
di: Menghini, Cristina, et al.
Pubblicazione: (2023)
di: Menghini, Cristina, et al.
Pubblicazione: (2023)
CLIP-VQDiffusion : Langauge Free Training of Text To Image generation using CLIP and vector quantized diffusion model
di: Han, Seungdae, et al.
Pubblicazione: (2024)
di: Han, Seungdae, et al.
Pubblicazione: (2024)
Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings
di: Kim, Bumjun, et al.
Pubblicazione: (2026)
di: Kim, Bumjun, et al.
Pubblicazione: (2026)
VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
di: Giahi, Ramin, et al.
Pubblicazione: (2025)
di: Giahi, Ramin, et al.
Pubblicazione: (2025)
DeCLIP: Decoupled Learning for Open-Vocabulary Dense Perception
di: Wang, Junjie, et al.
Pubblicazione: (2025)
di: Wang, Junjie, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PureCLIP-Depth: Prompt-Free and Decoder-Free Monocular Depth Estimation within CLIP Embedding Space
di: Miya, Ryutaro, et al.
Pubblicazione: (2026) -
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
di: Kim, Donghyeong, et al.
Pubblicazione: (2025) -
CLIP Can Understand Depth
di: Kim, Sohee, et al.
Pubblicazione: (2024) -
AnimalMotionCLIP: Embedding motion in CLIP for Animal Behavior Analysis
di: Zhong, Enmin, et al.
Pubblicazione: (2025) -
TIE-KD: Teacher-Independent and Explainable Knowledge Distillation for Monocular Depth Estimation
di: Choi, Sangwon, et al.
Pubblicazione: (2024)