Enhancing Weakly Supervised Semantic Segmentation with Multi-modal Foundation Models: An End-to-End Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Ravanbakhsh, Elham, Niu, Cheng, Liang, Yongqing, Ramanujam, J., Li, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep video representation learning: a survey
by: Ravanbakhsh, Elham, et al.
Published: (2024)
by: Ravanbakhsh, Elham, et al.
Published: (2024)
Multi-modality Affinity Inference for Weakly Supervised 3D Semantic Segmentation
by: Li, Xiawei, et al.
Published: (2023)
by: Li, Xiawei, et al.
Published: (2023)
Boundary-Refined Prototype Generation: A General End-to-End Paradigm for Semi-Supervised Semantic Segmentation
by: Dong, Junhao, et al.
Published: (2023)
by: Dong, Junhao, et al.
Published: (2023)
An End-to-End Robust Point Cloud Semantic Segmentation Network with Single-Step Conditional Diffusion Models
by: Qu, Wentao, et al.
Published: (2024)
by: Qu, Wentao, et al.
Published: (2024)
Towards Weakly Supervised End-to-end Learning for Long-video Action Recognition
by: Zhou, Jiaming, et al.
Published: (2023)
by: Zhou, Jiaming, et al.
Published: (2023)
Noise2Map: End-to-End Diffusion Model for Semantic Segmentation and Change Detection
by: Shibli, Ali, et al.
Published: (2026)
by: Shibli, Ali, et al.
Published: (2026)
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
by: Chen, Zhaozheng, et al.
Published: (2023)
by: Chen, Zhaozheng, et al.
Published: (2023)
WPS-SAM: Towards Weakly-Supervised Part Segmentation with Foundation Models
by: Wu, Xinjian, et al.
Published: (2024)
by: Wu, Xinjian, et al.
Published: (2024)
Is Self-Supervision Enough? Benchmarking Foundation Models Against End-to-End Training for Mitotic Figure Classification
by: Ganz, Jonathan, et al.
Published: (2024)
by: Ganz, Jonathan, et al.
Published: (2024)
Towards End-to-End Semi-Supervised Table Detection with Semantic Aligned Matching Transformer
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
Learning from the Web: Language Drives Weakly-Supervised Incremental Learning for Semantic Segmentation
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
Revisiting End-to-End Learning with Slide-level Supervision in Computational Pathology
by: Tang, Wenhao, et al.
Published: (2025)
by: Tang, Wenhao, et al.
Published: (2025)
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025)
by: Qin, Jack, et al.
Published: (2025)
FoundationSLAM: Unleashing the Power of Depth Foundation Models for End-to-End Dense Visual SLAM
by: Wu, Yuchen, et al.
Published: (2025)
by: Wu, Yuchen, et al.
Published: (2025)
Spatial Structure Constraints for Weakly Supervised Semantic Segmentation
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
Weakly Supervised Semantic Segmentation for Driving Scenes
by: Kim, Dongseob, et al.
Published: (2023)
by: Kim, Dongseob, et al.
Published: (2023)
WP-CrackNet: A Collaborative Adversarial Learning Framework for End-to-End Weakly-Supervised Road Crack Detection
by: Ma, Nachuan, et al.
Published: (2025)
by: Ma, Nachuan, et al.
Published: (2025)
EXAONE Path 2.0: Pathology Foundation Model with End-to-End Supervision
by: Pyeon, Myeongjang, et al.
Published: (2025)
by: Pyeon, Myeongjang, et al.
Published: (2025)
DiCLIP: Diffusion Model Enhances CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
by: Yang, Zhiwei, et al.
Published: (2026)
by: Yang, Zhiwei, et al.
Published: (2026)
Auxiliary Tasks Enhanced Dual-affinity Learning for Weakly Supervised Semantic Segmentation
by: Xu, Lian, et al.
Published: (2024)
by: Xu, Lian, et al.
Published: (2024)
Calibrating Undisciplined Over-Smoothing in Transformer for Weakly Supervised Semantic Segmentation
by: Cheng, Lechao, et al.
Published: (2023)
by: Cheng, Lechao, et al.
Published: (2023)
Enhancing End-to-End Autonomous Driving with Latent World Model
by: Li, Yingyan, et al.
Published: (2024)
by: Li, Yingyan, et al.
Published: (2024)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
Image Augmentation Agent for Weakly Supervised Semantic Segmentation
by: Wu, Wangyu, et al.
Published: (2024)
by: Wu, Wangyu, et al.
Published: (2024)
Contrastive Prompt Clustering for Weakly Supervised Semantic Segmentation
by: Wu, Wangyu, et al.
Published: (2025)
by: Wu, Wangyu, et al.
Published: (2025)
Adaptive Patch Contrast for Weakly Supervised Semantic Segmentation
by: Wu, Wangyu, et al.
Published: (2024)
by: Wu, Wangyu, et al.
Published: (2024)
Prompt Categories Cluster for Weakly Supervised Semantic Segmentation
by: Wu, Wangyu, et al.
Published: (2024)
by: Wu, Wangyu, et al.
Published: (2024)
Rethinking Saliency-Guided Weakly-Supervised Semantic Segmentation
by: Kim, Beomyoung, et al.
Published: (2024)
by: Kim, Beomyoung, et al.
Published: (2024)
Lightweight Transformer Framework for Weakly Supervised Semantic Segmentation
by: Torabi, Ali, et al.
Published: (2025)
by: Torabi, Ali, et al.
Published: (2025)
E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models
by: Cong, Wenyan, et al.
Published: (2025)
by: Cong, Wenyan, et al.
Published: (2025)
SynDiff-AD: Improving Semantic Segmentation and End-to-End Autonomous Driving with Synthetic Data from Latent Diffusion Models
by: Goel, Harsh, et al.
Published: (2024)
by: Goel, Harsh, et al.
Published: (2024)
End-to-End Streaming Video Temporal Action Segmentation with Reinforce Learning
by: Zhang, Jinrong, et al.
Published: (2023)
by: Zhang, Jinrong, et al.
Published: (2023)
Towards Segmenting the Invisible: An End-to-End Registration and Segmentation Framework for Weakly Supervised Tumour Analysis
by: Mukhopadhyay, Budhaditya, et al.
Published: (2026)
by: Mukhopadhyay, Budhaditya, et al.
Published: (2026)
End-To-End Underwater Video Enhancement: Dataset and Model
by: Du, Dazhao, et al.
Published: (2024)
by: Du, Dazhao, et al.
Published: (2024)
Learning Robust Correlation with Foundation Model for Weakly-Supervised Few-Shot Segmentation
by: Huang, Xinyang, et al.
Published: (2024)
by: Huang, Xinyang, et al.
Published: (2024)
Multi-modal NeRF Self-Supervision for LiDAR Semantic Segmentation
by: Timoneda, Xavier, et al.
Published: (2024)
by: Timoneda, Xavier, et al.
Published: (2024)
Distribution Guidance Network for Weakly Supervised Point Cloud Semantic Segmentation
by: Pan, Zhiyi, et al.
Published: (2024)
by: Pan, Zhiyi, et al.
Published: (2024)
VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Weakly Supervised Co-training with Swapping Assignments for Semantic Segmentation
by: Yang, Xinyu, et al.
Published: (2024)
by: Yang, Xinyu, et al.
Published: (2024)
Similar Items
-
Deep video representation learning: a survey
by: Ravanbakhsh, Elham, et al.
Published: (2024) -
Multi-modality Affinity Inference for Weakly Supervised 3D Semantic Segmentation
by: Li, Xiawei, et al.
Published: (2023) -
Boundary-Refined Prototype Generation: A General End-to-End Paradigm for Semi-Supervised Semantic Segmentation
by: Dong, Junhao, et al.
Published: (2023) -
An End-to-End Robust Point Cloud Semantic Segmentation Network with Single-Step Conditional Diffusion Models
by: Qu, Wentao, et al.
Published: (2024) -
Towards Weakly Supervised End-to-end Learning for Long-video Action Recognition
by: Zhou, Jiaming, et al.
Published: (2023)