Exploring Modality Guidance to Enhance VFM-based Feature Fusion for UDA in 3D Semantic Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Spoecklberger, Johannes, Lin, Wei, Hermosilla, Pedro, Doveh, Sivan, Possegger, Horst, Mirza, M. Jehanzeb |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
Into the Fog: Evaluating Robustness of Multiple Object Tracking
di: Kirillova, Nadezda, et al.
Pubblicazione: (2024)
di: Kirillova, Nadezda, et al.
Pubblicazione: (2024)
TTT-KD: Test-Time Training for 3D Semantic Segmentation through Knowledge Distillation from Foundation Models
di: Weijler, Lisa, et al.
Pubblicazione: (2024)
di: Weijler, Lisa, et al.
Pubblicazione: (2024)
Comparison Visual Instruction Tuning
di: Lin, Wei, et al.
Pubblicazione: (2024)
di: Lin, Wei, et al.
Pubblicazione: (2024)
PRISMM-Bench: A Benchmark of Peer-Review Grounded Multimodal Inconsistencies
di: Selch, Lukas, et al.
Pubblicazione: (2025)
di: Selch, Lukas, et al.
Pubblicazione: (2025)
VFM-UDA++: Improving Network Architectures and Data Strategies for Unsupervised Domain Adaptive Semantic Segmentation
di: Englert, Brunó B., et al.
Pubblicazione: (2025)
di: Englert, Brunó B., et al.
Pubblicazione: (2025)
What is the Added Value of UDA in the VFM Era?
di: Englert, Brunó B., et al.
Pubblicazione: (2025)
di: Englert, Brunó B., et al.
Pubblicazione: (2025)
Vision-Language Guidance for LiDAR-based Unsupervised 3D Object Detection
di: Fruhwirth-Reisinger, Christian, et al.
Pubblicazione: (2024)
di: Fruhwirth-Reisinger, Christian, et al.
Pubblicazione: (2024)
Towards Multimodal In-Context Learning for Vision & Language Models
di: Doveh, Sivan, et al.
Pubblicazione: (2024)
di: Doveh, Sivan, et al.
Pubblicazione: (2024)
LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content
di: Shabtay, Nimrod, et al.
Pubblicazione: (2024)
di: Shabtay, Nimrod, et al.
Pubblicazione: (2024)
GLOV: Guided Large Language Models as Implicit Optimizers for Vision Language Models
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes
di: Gavrikov, Paul, et al.
Pubblicazione: (2025)
di: Gavrikov, Paul, et al.
Pubblicazione: (2025)
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
di: Kniesel, Hannah, et al.
Pubblicazione: (2025)
di: Kniesel, Hannah, et al.
Pubblicazione: (2025)
TTRV: Test-Time Reinforcement Learning for Vision Language Models
di: Singh, Akshit, et al.
Pubblicazione: (2025)
di: Singh, Akshit, et al.
Pubblicazione: (2025)
Multi-Granularity Feature Calibration via VFM for Domain Generalized Semantic Segmentation
di: Li, Xinhui, et al.
Pubblicazione: (2025)
di: Li, Xinhui, et al.
Pubblicazione: (2025)
Unsupervised Semantic Segmentation Through Depth-Guided Feature Correlation and Sampling
di: Sick, Leon, et al.
Pubblicazione: (2023)
di: Sick, Leon, et al.
Pubblicazione: (2023)
Teaching VLMs to Localize Specific Objects from In-context Examples
di: Doveh, Sivan, et al.
Pubblicazione: (2024)
di: Doveh, Sivan, et al.
Pubblicazione: (2024)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
di: Sick, Leon, et al.
Pubblicazione: (2024)
di: Sick, Leon, et al.
Pubblicazione: (2024)
Semantics, Distortion, and Style Matter: Towards Source-free UDA for Panoramic Segmentation
di: Zheng, Xu, et al.
Pubblicazione: (2024)
di: Zheng, Xu, et al.
Pubblicazione: (2024)
OmniSAM: Omnidirectional Segment Anything Model for UDA in Panoramic Semantic Segmentation
di: Zhong, Ding, et al.
Pubblicazione: (2025)
di: Zhong, Ding, et al.
Pubblicazione: (2025)
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
di: Guo, Xiaodong, et al.
Pubblicazione: (2025)
di: Guo, Xiaodong, et al.
Pubblicazione: (2025)
Adapting Segment Anything Model to Multi-modal Salient Object Detection with Semantic Feature Fusion Guidance
di: Wang, Kunpeng, et al.
Pubblicazione: (2024)
di: Wang, Kunpeng, et al.
Pubblicazione: (2024)
Denoise and Align: Towards Source-Free UDA for Robust Panoramic Semantic Segmentation
di: Chang, Yaowen, et al.
Pubblicazione: (2026)
di: Chang, Yaowen, et al.
Pubblicazione: (2026)
Efficient Motion Prediction: A Lightweight & Accurate Trajectory Prediction Model With Fast Training and Inference Speed
di: Prutsch, Alexander, et al.
Pubblicazione: (2024)
di: Prutsch, Alexander, et al.
Pubblicazione: (2024)
GBlobs: Explicit Local Structure via Gaussian Blobs for Improved Cross-Domain LiDAR-based 3D Object Detection
di: Malić, Dušan, et al.
Pubblicazione: (2025)
di: Malić, Dušan, et al.
Pubblicazione: (2025)
NumeroLogic: Number Encoding for Enhanced LLMs' Numerical Reasoning
di: Schwartz, Eli, et al.
Pubblicazione: (2024)
di: Schwartz, Eli, et al.
Pubblicazione: (2024)
ICONIC-444: A 3.1-Million-Image Dataset for OOD Detection Research
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
ConMe: Rethinking Evaluation of Compositional Reasoning for Modern VLMs
di: Huang, Irene, et al.
Pubblicazione: (2024)
di: Huang, Irene, et al.
Pubblicazione: (2024)
Probing the effectiveness of World Models for Spatial Reasoning through Test-time Scaling
di: Jha, Saurav, et al.
Pubblicazione: (2025)
di: Jha, Saurav, et al.
Pubblicazione: (2025)
MAEDAY: MAE for few and zero shot AnomalY-Detection
di: Schwartz, Eli, et al.
Pubblicazione: (2022)
di: Schwartz, Eli, et al.
Pubblicazione: (2022)
ControlUDA: Controllable Diffusion-assisted Unsupervised Domain Adaptation for Cross-Weather Semantic Segmentation
di: Shen, Fengyi, et al.
Pubblicazione: (2024)
di: Shen, Fengyi, et al.
Pubblicazione: (2024)
Early Fusion of Features for Semantic Segmentation
di: Gupta, Anupam, et al.
Pubblicazione: (2024)
di: Gupta, Anupam, et al.
Pubblicazione: (2024)
On the mixed UDA states and additivity
di: Qiu, Xinyu, et al.
Pubblicazione: (2025)
di: Qiu, Xinyu, et al.
Pubblicazione: (2025)
Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic Segmentation
di: Chen, Siyu, et al.
Pubblicazione: (2025)
di: Chen, Siyu, et al.
Pubblicazione: (2025)
One Model, Many Behaviors: Training-Induced Effects on Out-of-Distribution Detection
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
di: Krumpl, Gerhard, et al.
Pubblicazione: (2026)
Streaming Real-Time Trajectory Prediction Using Endpoint-Aware Modeling
di: Prutsch, Alexander, et al.
Pubblicazione: (2026)
di: Prutsch, Alexander, et al.
Pubblicazione: (2026)
ASCENT: Transformer-Based Aircraft Trajectory Prediction in Non-Towered Terminal Airspace
di: Prutsch, Alexander, et al.
Pubblicazione: (2026)
di: Prutsch, Alexander, et al.
Pubblicazione: (2026)
StitchFusion: Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2024)
di: Li, Bingyu, et al.
Pubblicazione: (2024)
Efficient Continuous Group Convolutions for Local SE(3) Equivariance in 3D Point Clouds
di: Weijler, Lisa, et al.
Pubblicazione: (2025)
di: Weijler, Lisa, et al.
Pubblicazione: (2025)
Robust Localization of Key Fob Using Channel Impulse Response of Ultra Wide Band Sensors for Keyless Entry Systems
di: Kolli, Abhiram, et al.
Pubblicazione: (2024)
di: Kolli, Abhiram, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024) -
Into the Fog: Evaluating Robustness of Multiple Object Tracking
di: Kirillova, Nadezda, et al.
Pubblicazione: (2024) -
TTT-KD: Test-Time Training for 3D Semantic Segmentation through Knowledge Distillation from Foundation Models
di: Weijler, Lisa, et al.
Pubblicazione: (2024) -
Comparison Visual Instruction Tuning
di: Lin, Wei, et al.
Pubblicazione: (2024) -
PRISMM-Bench: A Benchmark of Peer-Review Grounded Multimodal Inconsistencies
di: Selch, Lukas, et al.
Pubblicazione: (2025)