SAM-SP: Self-Prompting Makes SAM Great Again
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Chunpeng, Ning, Kangjie, Shen, Qianqian, Zhou, Sheng, Yu, Zhi, Wang, Haishuai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2025)
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2025)
From Images to Insights: Explainable Biodiversity Monitoring with Plain Language Habitat Explanations
von: Zhou, Yutong, et al.
Veröffentlicht: (2025)
von: Zhou, Yutong, et al.
Veröffentlicht: (2025)
Prompt to Protection: A Comparative Study of Multimodal LLMs in Construction Hazard Recognition
von: Chaudhary, Nishi, et al.
Veröffentlicht: (2025)
von: Chaudhary, Nishi, et al.
Veröffentlicht: (2025)
PhytoSynth: Leveraging Multi-modal Generative Models for Crop Disease Data Generation with Novel Benchmarking and Prompt Engineering Approach
von: Rai, Nitin, et al.
Veröffentlicht: (2025)
von: Rai, Nitin, et al.
Veröffentlicht: (2025)
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
von: Jongwiriyanurak, Natchapon, et al.
Veröffentlicht: (2024)
von: Jongwiriyanurak, Natchapon, et al.
Veröffentlicht: (2024)
Progressive Multimodal Interaction Network for Reliable Quantification of Fish Feeding Intensity in Aquaculture
von: Zhang, Shulong, et al.
Veröffentlicht: (2025)
von: Zhang, Shulong, et al.
Veröffentlicht: (2025)
Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems
von: Kyem, Blessing Agyei, et al.
Veröffentlicht: (2026)
von: Kyem, Blessing Agyei, et al.
Veröffentlicht: (2026)
Improving Object Detector Training on Synthetic Data by Starting With a Strong Baseline Methodology
von: Ruis, Frank A., et al.
Veröffentlicht: (2024)
von: Ruis, Frank A., et al.
Veröffentlicht: (2024)
Detecting Multiple Diseases in Multiple Crops Using Deep Learning
von: Yadav, Vivek, et al.
Veröffentlicht: (2025)
von: Yadav, Vivek, et al.
Veröffentlicht: (2025)
Improving watermelon (Citrullus lanatus) disease classification with generative artificial intelligence (GenAI)-based synthetic and real-field images via a custom EfficientNetV2-L model
von: Rai, Nitin, et al.
Veröffentlicht: (2025)
von: Rai, Nitin, et al.
Veröffentlicht: (2025)
Labits: Layered Bidirectional Time Surfaces Representation for Event Camera-based Continuous Dense Trajectory Estimation
von: Zhang, Zhongyang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongyang, et al.
Veröffentlicht: (2024)
Automated Facility Enumeration for Building Compliance Checking using Door Detection and Large Language Models
von: Zhang, Licheng, et al.
Veröffentlicht: (2025)
von: Zhang, Licheng, et al.
Veröffentlicht: (2025)
Human-in-the-Loop: Quantitative Evaluation of 3D Models Generation by Large Language Models
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
von: Sadik, Ahmed R., et al.
Veröffentlicht: (2025)
QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing
von: Mittal, Garvit Kumar, et al.
Veröffentlicht: (2026)
von: Mittal, Garvit Kumar, et al.
Veröffentlicht: (2026)
DoorDet: Semi-Automated Multi-Class Door Detection Dataset via Object Detection and Large Language Models
von: Zhang, Licheng, et al.
Veröffentlicht: (2025)
von: Zhang, Licheng, et al.
Veröffentlicht: (2025)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
von: Goodge, Adam, et al.
Veröffentlicht: (2025)
von: Goodge, Adam, et al.
Veröffentlicht: (2025)
FollowGen: A Scaled Noise Conditional Diffusion Model for Car-Following Trajectory Prediction
von: You, Junwei, et al.
Veröffentlicht: (2024)
von: You, Junwei, et al.
Veröffentlicht: (2024)
Human Cognition in Machines: A Unified Perspective of World Models
von: Rupprecht, Timothy, et al.
Veröffentlicht: (2026)
von: Rupprecht, Timothy, et al.
Veröffentlicht: (2026)
Less is More: A Closer Look at Semantic-based Few-Shot Learning
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2024)
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2024)
Advanced Clustering Framework for Semiconductor Image Analytics Integrating Deep TDA with Self-Supervised and Transfer Learning Techniques
von: Giri, Janhavi, et al.
Veröffentlicht: (2025)
von: Giri, Janhavi, et al.
Veröffentlicht: (2025)
Self-supervised Normality Learning and Divergence Vector-guided Model Merging for Zero-shot Congenital Heart Disease Detection in Fetal Ultrasound Videos
von: Saha, Pramit, et al.
Veröffentlicht: (2025)
von: Saha, Pramit, et al.
Veröffentlicht: (2025)
Toward Accountable AI-Generated Content on Social Platforms: Steganographic Attribution and Multimodal Harm Detection
von: Guan, Xinlei, et al.
Veröffentlicht: (2026)
von: Guan, Xinlei, et al.
Veröffentlicht: (2026)
Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
von: Wang, Dongdong, et al.
Veröffentlicht: (2026)
Resource-Aware Evolutionary Neural Architecture Search for Cardiac MRI Segmentation
von: Yasmin, Farhana, et al.
Veröffentlicht: (2026)
von: Yasmin, Farhana, et al.
Veröffentlicht: (2026)
DeepSeek-Inspired Exploration of RL-based LLMs and Synergy with Wireless Networks: A Survey
von: Qiao, Yu, et al.
Veröffentlicht: (2025)
von: Qiao, Yu, et al.
Veröffentlicht: (2025)
MicroCrackAttentionNeXt: Advancing Microcrack Detection in Wave Field Analysis Using Deep Neural Networks through Feature Visualization
von: Moreh, Fatahlla, et al.
Veröffentlicht: (2024)
von: Moreh, Fatahlla, et al.
Veröffentlicht: (2024)
Assessing the Added Value of Onboard Earth Observation Processing with the IRIDE HEO Service Segment
von: Thind, Parampuneet Kaur, et al.
Veröffentlicht: (2026)
von: Thind, Parampuneet Kaur, et al.
Veröffentlicht: (2026)
MedGrad E-CLIP: Enhancing Trust and Transparency in AI-Driven Skin Lesion Diagnosis
von: Kamal, Sadia, et al.
Veröffentlicht: (2025)
von: Kamal, Sadia, et al.
Veröffentlicht: (2025)
WiFlexFormer: Efficient WiFi-Based Person-Centric Sensing
von: Strohmayer, Julian, et al.
Veröffentlicht: (2024)
von: Strohmayer, Julian, et al.
Veröffentlicht: (2024)
DATTA: Domain-Adversarial Test-Time Adaptation for Cross-Domain WiFi-Based Human Activity Recognition
von: Strohmayer, Julian, et al.
Veröffentlicht: (2024)
von: Strohmayer, Julian, et al.
Veröffentlicht: (2024)
Interpretability-Aware Pruning for Efficient Medical Image Analysis
von: Malik, Nikita, et al.
Veröffentlicht: (2025)
von: Malik, Nikita, et al.
Veröffentlicht: (2025)
Exploring Advanced Techniques for Visual Question Answering: A Comprehensive Comparison
von: Baby, Aiswarya, et al.
Veröffentlicht: (2025)
von: Baby, Aiswarya, et al.
Veröffentlicht: (2025)
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking
von: Alghoraibi, Huda, et al.
Veröffentlicht: (2025)
von: Alghoraibi, Huda, et al.
Veröffentlicht: (2025)
Automating grapevine LAI features estimation with UAV imagery and machine learning
von: Akram, Muhammad Waseem, et al.
Veröffentlicht: (2024)
von: Akram, Muhammad Waseem, et al.
Veröffentlicht: (2024)
Demographic Predictability in 3D CT Foundation Embeddings
von: Zheng, Guangyao, et al.
Veröffentlicht: (2024)
von: Zheng, Guangyao, et al.
Veröffentlicht: (2024)
Haphazard Inputs as Images in Online Learning
von: Agarwal, Rohit, et al.
Veröffentlicht: (2025)
von: Agarwal, Rohit, et al.
Veröffentlicht: (2025)
CNN-Based Automated Parameter Extraction Framework for Modeling Memristive Devices
von: Hamid, Akif, et al.
Veröffentlicht: (2025)
von: Hamid, Akif, et al.
Veröffentlicht: (2025)
An explainable vision transformer with transfer learning based efficient drought stress identification
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2024)
von: Patra, Aswini Kumar, et al.
Veröffentlicht: (2024)
Enhancing Object Detection with Privileged Information: A Model-Agnostic Teacher-Student Approach
von: Bartolo, Matthias, et al.
Veröffentlicht: (2026)
von: Bartolo, Matthias, et al.
Veröffentlicht: (2026)
Self-evolving Embodied AI
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2025) -
From Images to Insights: Explainable Biodiversity Monitoring with Plain Language Habitat Explanations
von: Zhou, Yutong, et al.
Veröffentlicht: (2025) -
Prompt to Protection: A Comparative Study of Multimodal LLMs in Construction Hazard Recognition
von: Chaudhary, Nishi, et al.
Veröffentlicht: (2025) -
PhytoSynth: Leveraging Multi-modal Generative Models for Crop Disease Data Generation with Novel Benchmarking and Prompt Engineering Approach
von: Rai, Nitin, et al.
Veröffentlicht: (2025) -
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
von: Jongwiriyanurak, Natchapon, et al.
Veröffentlicht: (2024)