FusionSAM: Visual Multi-Modal Learning with Segment Anything
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Daixun, Xie, Weiying, Cao, Mingxiang, Wang, Yunke, Zhang, Yusi, Fang, Leyuan, Li, Yunsong, Xu, Chang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-scale direction-aware SAR object detection network via global information fusion
by: Cao, Mingxiang, et al.
Published: (2023)
by: Cao, Mingxiang, et al.
Published: (2023)
E2E-MFD: Towards End-to-End Synchronous Multimodal Fusion Detection
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
BSDM: Background Suppression Diffusion Model for Hyperspectral Anomaly Detection
by: Ma, Jitao, et al.
Published: (2023)
by: Ma, Jitao, et al.
Published: (2023)
Reducing Spurious Correlation for Federated Domain Generalization
by: Ma, Shuran, et al.
Published: (2024)
by: Ma, Shuran, et al.
Published: (2024)
RS-DGC: Exploring Neighborhood Statistics for Dynamic Gradient Compression on Remote Sensing Image Interpretation
by: Xie, Weiying, et al.
Published: (2023)
by: Xie, Weiying, et al.
Published: (2023)
FedDiff: Diffusion Model Driven Federated Learning for Multi-Modal and Multi-Clients
by: Li, DaiXun, et al.
Published: (2023)
by: Li, DaiXun, et al.
Published: (2023)
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Physics Inspired Criterion for Pruning-Quantization Joint Learning
by: Xie, Weiying, et al.
Published: (2023)
by: Xie, Weiying, et al.
Published: (2023)
Distribution-aware Interactive Attention Network and Large-scale Cloud Recognition Benchmark on FY-4A Satellite Image
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
M$^3$amba: CLIP-driven Mamba Model for Multi-modal Remote Sensing Classification
by: Cao, Mingxiang, et al.
Published: (2025)
by: Cao, Mingxiang, et al.
Published: (2025)
Hyperspectral Anomaly Detection with Self-Supervised Anomaly Prior
by: Liu, Yidan, et al.
Published: (2024)
by: Liu, Yidan, et al.
Published: (2024)
MemorySAM: Memorize Modalities and Semantics with Segment Anything Model 2 for Multi-modal Semantic Segmentation
by: Liao, Chenfei, et al.
Published: (2025)
by: Liao, Chenfei, et al.
Published: (2025)
Exploring Hyperspectral Anomaly Detection with Human Vision: A Small Target Aware Detector
by: Ma, Jitao, et al.
Published: (2024)
by: Ma, Jitao, et al.
Published: (2024)
FoRA: Low-Rank Adaptation Model beyond Multimodal Siamese Network
by: Xie, Weiying, et al.
Published: (2024)
by: Xie, Weiying, et al.
Published: (2024)
ViRefSAM: Visual Reference-Guided Segment Anything Model for Remote Sensing Segmentation
by: Bi, Hanbo, et al.
Published: (2025)
by: Bi, Hanbo, et al.
Published: (2025)
RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything
by: Xu, Shilin, et al.
Published: (2024)
by: Xu, Shilin, et al.
Published: (2024)
SAM2-LOVE: Segment Anything Model 2 in Language-aided Audio-Visual Scenes
by: Wang, Yuji, et al.
Published: (2025)
by: Wang, Yuji, et al.
Published: (2025)
SAM3-I: Segment Anything with Instructions
by: Li, Jingjing, et al.
Published: (2025)
by: Li, Jingjing, et al.
Published: (2025)
AM-SAM: Automated Prompting and Mask Calibration for Segment Anything Model
by: Li, Yuchen, et al.
Published: (2024)
by: Li, Yuchen, et al.
Published: (2024)
SwiMDiff: Scene-wide Matching Contrastive Learning with Diffusion Constraint for Remote Sensing Image
by: Tian, Jiayuan, et al.
Published: (2024)
by: Tian, Jiayuan, et al.
Published: (2024)
MedSAM3: Delving into Segment Anything with Medical Concepts
by: Liu, Anglin, et al.
Published: (2025)
by: Liu, Anglin, et al.
Published: (2025)
I-MedSAM: Implicit Medical Image Segmentation with Segment Anything
by: Wei, Xiaobao, et al.
Published: (2023)
by: Wei, Xiaobao, et al.
Published: (2023)
Char-SAM: Turning Segment Anything Model into Scene Text Segmentation Annotator with Character-level Visual Prompts
by: Xie, Enze, et al.
Published: (2024)
by: Xie, Enze, et al.
Published: (2024)
Biomedical SAM 2: Segment Anything in Biomedical Images and Videos
by: Yan, Zhiling, et al.
Published: (2024)
by: Yan, Zhiling, et al.
Published: (2024)
TinySAM: Pushing the Envelope for Efficient Segment Anything Model
by: Shu, Han, et al.
Published: (2023)
by: Shu, Han, et al.
Published: (2023)
SAM 3: Segment Anything with Concepts
by: Carion, Nicolas, et al.
Published: (2025)
by: Carion, Nicolas, et al.
Published: (2025)
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
by: Huang, You, et al.
Published: (2024)
by: Huang, You, et al.
Published: (2024)
SlimSAM: 0.1% Data Makes Segment Anything Slim
by: Chen, Zigeng, et al.
Published: (2023)
by: Chen, Zigeng, et al.
Published: (2023)
AlignSAM: Aligning Segment Anything Model to Open Context via Reinforcement Learning
by: Huang, Duojun, et al.
Published: (2024)
by: Huang, Duojun, et al.
Published: (2024)
EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model
by: Zhang, Yuxuan, et al.
Published: (2024)
by: Zhang, Yuxuan, et al.
Published: (2024)
SAQ-SAM: Semantically-Aligned Quantization for Segment Anything Model
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
RemoteSAM: Towards Segment Anything for Earth Observation
by: Yao, Liang, et al.
Published: (2025)
by: Yao, Liang, et al.
Published: (2025)
WSI-SAM: Multi-resolution Segment Anything Model (SAM) for histopathology whole-slide images
by: Liu, Hong, et al.
Published: (2024)
by: Liu, Hong, et al.
Published: (2024)
UrbanSAM: Learning Invariance-Inspired Adapters for Segment Anything Models in Urban Construction
by: Li, Chenyu, et al.
Published: (2025)
by: Li, Chenyu, et al.
Published: (2025)
PointSAM: Pointly-Supervised Segment Anything Model for Remote Sensing Images
by: Liu, Nanqing, et al.
Published: (2024)
by: Liu, Nanqing, et al.
Published: (2024)
Medical SAM Adapter: Adapting Segment Anything Model for Medical Image Segmentation
by: Wu, Junde, et al.
Published: (2023)
by: Wu, Junde, et al.
Published: (2023)
PaveSAM Segment Anything for Pavement Distress
by: Owor, Neema Jakisa, et al.
Published: (2024)
by: Owor, Neema Jakisa, et al.
Published: (2024)
DiffCLIP: Few-shot Language-driven Multimodal Classifier
by: Zhang, Jiaqing, et al.
Published: (2024)
by: Zhang, Jiaqing, et al.
Published: (2024)
Segment Anything Is Not Always Perfect: An Investigation of SAM on Different Real-world Applications
by: Ji, Wei, et al.
Published: (2023)
by: Ji, Wei, et al.
Published: (2023)
SPLF-SAM: Self-Prompting Segment Anything Model for Light Field Salient Object Detection
by: Xu, Qiyao, et al.
Published: (2025)
by: Xu, Qiyao, et al.
Published: (2025)
Similar Items
-
Multi-scale direction-aware SAR object detection network via global information fusion
by: Cao, Mingxiang, et al.
Published: (2023) -
E2E-MFD: Towards End-to-End Synchronous Multimodal Fusion Detection
by: Zhang, Jiaqing, et al.
Published: (2024) -
BSDM: Background Suppression Diffusion Model for Hyperspectral Anomaly Detection
by: Ma, Jitao, et al.
Published: (2023) -
Reducing Spurious Correlation for Federated Domain Generalization
by: Ma, Shuran, et al.
Published: (2024) -
RS-DGC: Exploring Neighborhood Statistics for Dynamic Gradient Compression on Remote Sensing Image Interpretation
by: Xie, Weiying, et al.
Published: (2023)