RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Shilin, Yuan, Haobo, Shi, Qingyu, Qi, Lu, Wang, Jingbo, Yang, Yibo, Li, Yining, Chen, Kai, Tong, Yunhai, Ghanem, Bernard, Li, Xiangtai, Yang, Ming-Hsuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLAVADI: What Matters For Multimodal Large Language Models Distillation
von: Xu, Shilin, et al.
Veröffentlicht: (2024)
von: Xu, Shilin, et al.
Veröffentlicht: (2024)
PanopticPartFormer++: A Unified and Decoupled View for Panoptic Part Segmentation
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
DC-SAM: In-Context Segment Anything in Images and Videos via Dual Consistency
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Towards Open Vocabulary Learning: A Survey
von: Wu, Jianzong, et al.
Veröffentlicht: (2023)
von: Wu, Jianzong, et al.
Veröffentlicht: (2023)
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
Open-Vocabulary SAM: Segment and Recognize Twenty-thousand Classes Interactively
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything Model
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
SemFlow: Binding Semantic Segmentation and Image Synthesis via Rectified Flow
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
DreamRelation: Bridging Customization and Relation Generation
von: Shi, Qingyu, et al.
Veröffentlicht: (2024)
von: Shi, Qingyu, et al.
Veröffentlicht: (2024)
BA-SAM: Scalable Bias-Mode Attention Mask for Segment Anything Model
von: Song, Yiran, et al.
Veröffentlicht: (2024)
von: Song, Yiran, et al.
Veröffentlicht: (2024)
Segment Anything Is Not Always Perfect: An Investigation of SAM on Different Real-world Applications
von: Ji, Wei, et al.
Veröffentlicht: (2023)
von: Ji, Wei, et al.
Veröffentlicht: (2023)
4th PVUW MeViS 3rd Place Report: Sa2VA
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)
Decouple and Track: Benchmarking and Improving Video Diffusion Transformers for Motion Transfer
von: Shi, Qingyu, et al.
Veröffentlicht: (2025)
von: Shi, Qingyu, et al.
Veröffentlicht: (2025)
RepViT-SAM: Towards Real-Time Segmenting Anything
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
RTMO: Towards High-Performance One-Stage Real-Time Multi-Person Pose Estimation
von: Lu, Peng, et al.
Veröffentlicht: (2023)
von: Lu, Peng, et al.
Veröffentlicht: (2023)
TS-SAM: Fine-Tuning Segment-Anything Model for Downstream Tasks
von: Yu, Yang, et al.
Veröffentlicht: (2024)
von: Yu, Yang, et al.
Veröffentlicht: (2024)
GenView: Enhancing View Quality with Pretrained Generative Model for Self-Supervised Learning
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
von: Li, Xiaojie, et al.
Veröffentlicht: (2024)
UMC: Unified Resilient Controller for Legged Robots with Joint Malfunctions
von: Qiu, Yu, et al.
Veröffentlicht: (2025)
von: Qiu, Yu, et al.
Veröffentlicht: (2025)
DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
SAM3-I: Segment Anything with Instructions
von: Li, Jingjing, et al.
Veröffentlicht: (2025)
von: Li, Jingjing, et al.
Veröffentlicht: (2025)
DST-Det: Simple Dynamic Self-Training for Open-Vocabulary Object Detection
von: Xu, Shilin, et al.
Veröffentlicht: (2023)
von: Xu, Shilin, et al.
Veröffentlicht: (2023)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
RecTok: Reconstruction Distillation along Rectified Flow
von: Shi, Qingyu, et al.
Veröffentlicht: (2025)
von: Shi, Qingyu, et al.
Veröffentlicht: (2025)
RemoteSAM: Towards Segment Anything for Earth Observation
von: Yao, Liang, et al.
Veröffentlicht: (2025)
von: Yao, Liang, et al.
Veröffentlicht: (2025)
Biomedical SAM 2: Segment Anything in Biomedical Images and Videos
von: Yan, Zhiling, et al.
Veröffentlicht: (2024)
von: Yan, Zhiling, et al.
Veröffentlicht: (2024)
Towards Language-Driven Video Inpainting via Multimodal Large Language Models
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
von: Wu, Jianzong, et al.
Veröffentlicht: (2024)
SAM-DA: UAV Tracks Anything at Night with SAM-Powered Domain Adaptation
von: Fu, Changhong, et al.
Veröffentlicht: (2023)
von: Fu, Changhong, et al.
Veröffentlicht: (2023)
SAM Audio: Segment Anything in Audio
von: Shi, Bowen, et al.
Veröffentlicht: (2025)
von: Shi, Bowen, et al.
Veröffentlicht: (2025)
SAM 3: Segment Anything with Concepts
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
OMG-Seg: Is One Model Good Enough For All Segmentation?
von: Li, Xiangtai, et al.
Veröffentlicht: (2024)
von: Li, Xiangtai, et al.
Veröffentlicht: (2024)
PlaneSAM: Multimodal Plane Instance Segmentation Using the Segment Anything Model
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
SAM-OCTA: Prompting Segment-Anything for OCTA Image Segmentation
von: Chen, Xinrun, et al.
Veröffentlicht: (2023)
von: Chen, Xinrun, et al.
Veröffentlicht: (2023)
SAM-PD: How Far Can SAM Take Us in Tracking and Segmenting Anything in Videos by Prompt Denoising
von: Zhou, Tao, et al.
Veröffentlicht: (2024)
von: Zhou, Tao, et al.
Veröffentlicht: (2024)
SAM Struggles in Concealed Scenes -- Empirical Study on Segment Anything
von: Ji, Ge-Peng, et al.
Veröffentlicht: (2023)
von: Ji, Ge-Peng, et al.
Veröffentlicht: (2023)
SAM4D: Segment Anything in Camera and LiDAR Streams
von: Xu, Jianyun, et al.
Veröffentlicht: (2025)
von: Xu, Jianyun, et al.
Veröffentlicht: (2025)
CAR-SAM: Cross-Attention Reconstruction for Post-Training Quantization of the Segment Anything Model
von: Wen, Houji, et al.
Veröffentlicht: (2026)
von: Wen, Houji, et al.
Veröffentlicht: (2026)
I-MedSAM: Implicit Medical Image Segmentation with Segment Anything
von: Wei, Xiaobao, et al.
Veröffentlicht: (2023)
von: Wei, Xiaobao, et al.
Veröffentlicht: (2023)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
DarkSAM: Fooling Segment Anything Model to Segment Nothing
von: Zhou, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhou, Ziqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LLAVADI: What Matters For Multimodal Large Language Models Distillation
von: Xu, Shilin, et al.
Veröffentlicht: (2024) -
PanopticPartFormer++: A Unified and Decoupled View for Panoptic Part Segmentation
von: Li, Xiangtai, et al.
Veröffentlicht: (2023) -
DC-SAM: In-Context Segment Anything in Images and Videos via Dual Consistency
von: Qi, Mengshi, et al.
Veröffentlicht: (2025) -
Towards Open Vocabulary Learning: A Survey
von: Wu, Jianzong, et al.
Veröffentlicht: (2023) -
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
von: Yuan, Haobo, et al.
Veröffentlicht: (2025)