Test-Time Adaptation for Height Completion via Self-Supervised ViT Features and Monocular Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rafaeli, Osher, Svoray, Tal, Nahlieli, Ariel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024)
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024)
Seamless High-Resolution Terrain Reconstruction: A Prior-Based Vision Transformer Approach
von: Rafaeli, Osher, et al.
Veröffentlicht: (2025)
von: Rafaeli, Osher, et al.
Veröffentlicht: (2025)
Prompt-Based Segmentation at Multiple Resolutions and Lighting Conditions using Segment Anything Model 2
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024)
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024)
On the Effectiveness of Textual Prompting with Lightweight Fine-Tuning for SAM3 Remote Sensing Segmentation
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2025)
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2025)
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
von: Han, Dongchen, et al.
Veröffentlicht: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
von: Kuzucu, Selim, et al.
Veröffentlicht: (2025)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
SegDebias: Test-Time Bias Mitigation for ViT-Based CLIP via Segmentation
von: Wu, Fangyu, et al.
Veröffentlicht: (2025)
von: Wu, Fangyu, et al.
Veröffentlicht: (2025)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
ViT-Split: Unleashing the Power of Vision Foundation Models via Efficient Splitting Heads
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
Rethinking Random Masking in Self-Distillation on ViT
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation
von: Zhao, Wangbo, et al.
Veröffentlicht: (2024)
von: Zhao, Wangbo, et al.
Veröffentlicht: (2024)
Which Direction to Choose? An Analysis on the Representation Power of Self-Supervised ViTs in Downstream Tasks
von: Kaltampanidis, Yannis, et al.
Veröffentlicht: (2025)
von: Kaltampanidis, Yannis, et al.
Veröffentlicht: (2025)
Enhancing Monocular Height Estimation via Weak Supervision from Imperfect Labels
von: Chen, Sining, et al.
Veröffentlicht: (2025)
von: Chen, Sining, et al.
Veröffentlicht: (2025)
Performance of Human Annotators in Object Detection and Segmentation of Remotely Sensed Data
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2024)
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2024)
TFS-ViT: Token-Level Feature Stylization for Domain Generalization
von: Noori, Mehrdad, et al.
Veröffentlicht: (2023)
von: Noori, Mehrdad, et al.
Veröffentlicht: (2023)
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
von: Sosa, Jose, et al.
Veröffentlicht: (2025)
von: Sosa, Jose, et al.
Veröffentlicht: (2025)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Learning CNN on ViT: A Hybrid Model to Explicitly Class-specific Boundaries for Domain Adaptation
von: Ngo, Ba Hung, et al.
Veröffentlicht: (2024)
von: Ngo, Ba Hung, et al.
Veröffentlicht: (2024)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
von: Ma, Yunsheng, et al.
Veröffentlicht: (2022)
von: Ma, Yunsheng, et al.
Veröffentlicht: (2022)
When Test-Time Adaptation Meets Self-Supervised Models
von: Han, Jisu, et al.
Veröffentlicht: (2025)
von: Han, Jisu, et al.
Veröffentlicht: (2025)
Distilling Monocular Foundation Model for Fine-grained Depth Completion
von: Liang, Yingping, et al.
Veröffentlicht: (2025)
von: Liang, Yingping, et al.
Veröffentlicht: (2025)
Few-Shot Class-Incremental Model Attribution Using Learnable Representation From CLIP-ViT Features
von: Lee, Hanbyul, et al.
Veröffentlicht: (2025)
von: Lee, Hanbyul, et al.
Veröffentlicht: (2025)
Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
von: Suri, Saksham, et al.
Veröffentlicht: (2024)
von: Suri, Saksham, et al.
Veröffentlicht: (2024)
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction
von: Chen, Weiming, et al.
Veröffentlicht: (2026)
von: Chen, Weiming, et al.
Veröffentlicht: (2026)
Elastic ViTs from Pretrained Models without Retraining
von: Simoncini, Walter, et al.
Veröffentlicht: (2025)
von: Simoncini, Walter, et al.
Veröffentlicht: (2025)
UniRefiner: Teaching Pre-trained ViTs to Self-Dispose Dross via Contrastive Register
von: Qiu, Congpei, et al.
Veröffentlicht: (2026)
von: Qiu, Congpei, et al.
Veröffentlicht: (2026)
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026)
von: Wang, Feng, et al.
Veröffentlicht: (2026)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
ViT-VS: On the Applicability of Pretrained Vision Transformer Features for Generalizable Visual Servoing
von: Scherl, Alessandro, et al.
Veröffentlicht: (2025)
von: Scherl, Alessandro, et al.
Veröffentlicht: (2025)
UniCT Depth: Event-Image Fusion Based Monocular Depth Estimation with Convolution-Compensated ViT Dual SA Block
von: Jing, Luoxi, et al.
Veröffentlicht: (2025)
von: Jing, Luoxi, et al.
Veröffentlicht: (2025)
Alias-Free ViT: Fractional Shift Invariance via Linear Attention
von: Michaeli, Hagay, et al.
Veröffentlicht: (2025)
von: Michaeli, Hagay, et al.
Veröffentlicht: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
VIVID-Med: LLM-Supervised Structured Pretraining for Deployable Medical ViTs
von: Wang, Xiyao, et al.
Veröffentlicht: (2026)
von: Wang, Xiyao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024) -
Seamless High-Resolution Terrain Reconstruction: A Prior-Based Vision Transformer Approach
von: Rafaeli, Osher, et al.
Veröffentlicht: (2025) -
Prompt-Based Segmentation at Multiple Resolutions and Lighting Conditions using Segment Anything Model 2
von: Rafaeli, Osher, et al.
Veröffentlicht: (2024) -
On the Effectiveness of Textual Prompting with Lightweight Fine-Tuning for SAM3 Remote Sensing Segmentation
von: Blushtein-Livnon, Roni, et al.
Veröffentlicht: (2025) -
ViT$^3$: Unlocking Test-Time Training in Vision
von: Han, Dongchen, et al.
Veröffentlicht: (2025)