Tiny-YOLOSAM: Fast Hybrid Image Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Kenneth, Wu, Songhan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
di: Wu, Songhan
Pubblicazione: (2025)
di: Wu, Songhan
Pubblicazione: (2025)
Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning
di: Kunzo, Tomáš, et al.
Pubblicazione: (2023)
di: Kunzo, Tomáš, et al.
Pubblicazione: (2023)
Cross-View-Prediction: Exploring Contrastive Feature for Hyperspectral Image Classification
di: Zhang, Anyu, et al.
Pubblicazione: (2022)
di: Zhang, Anyu, et al.
Pubblicazione: (2022)
TDIP: Tunable Deep Image Processing, a Real Time Melt Pool Monitoring Solution
di: Akhavan, Javid, et al.
Pubblicazione: (2024)
di: Akhavan, Javid, et al.
Pubblicazione: (2024)
Distant Object Localisation from Noisy Image Segmentation Sequences
di: Pesonen, Julius, et al.
Pubblicazione: (2025)
di: Pesonen, Julius, et al.
Pubblicazione: (2025)
On-the-Fly Guidance Training for Medical Image Registration
di: Xin, Yuelin, et al.
Pubblicazione: (2023)
di: Xin, Yuelin, et al.
Pubblicazione: (2023)
Robust Multi-Source Covid-19 Detection in CT Images
di: Pritha, Asmita Yuki, et al.
Pubblicazione: (2026)
di: Pritha, Asmita Yuki, et al.
Pubblicazione: (2026)
Cost Savings from Automatic Quality Assessment of Generated Images
di: Giro-i-Nieto, Xavier, et al.
Pubblicazione: (2025)
di: Giro-i-Nieto, Xavier, et al.
Pubblicazione: (2025)
Exploring Diffusion with Test-Time Training on Efficient Image Restoration
di: Lu, Rongchang, et al.
Pubblicazione: (2025)
di: Lu, Rongchang, et al.
Pubblicazione: (2025)
Dynamic Brightness Adaptation for Robust Multi-modal Image Fusion
di: Sun, Yiming, et al.
Pubblicazione: (2024)
di: Sun, Yiming, et al.
Pubblicazione: (2024)
Rapid Adaptation of Earth Observation Foundation Models for Segmentation
di: Selvam, Karthick Panner, et al.
Pubblicazione: (2024)
di: Selvam, Karthick Panner, et al.
Pubblicazione: (2024)
AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection
di: Figueira, Joaquín, et al.
Pubblicazione: (2026)
di: Figueira, Joaquín, et al.
Pubblicazione: (2026)
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
di: Xiao, Yuqi, et al.
Pubblicazione: (2025)
di: Xiao, Yuqi, et al.
Pubblicazione: (2025)
Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues
di: Willi, Marco, et al.
Pubblicazione: (2026)
di: Willi, Marco, et al.
Pubblicazione: (2026)
VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
MCA-Bench: A Multimodal Benchmark for Evaluating CAPTCHA Robustness Against VLM-based Attacks
di: Wu, Zonglin, et al.
Pubblicazione: (2025)
di: Wu, Zonglin, et al.
Pubblicazione: (2025)
DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
di: Li, Wenhao, et al.
Pubblicazione: (2026)
di: Li, Wenhao, et al.
Pubblicazione: (2026)
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
di: Wang, Junjue, et al.
Pubblicazione: (2025)
di: Wang, Junjue, et al.
Pubblicazione: (2025)
NumeriKontrol: Adding Numeric Control to Diffusion Transformers for Instruction-based Image Editing
di: Xu, Zhenyu, et al.
Pubblicazione: (2025)
di: Xu, Zhenyu, et al.
Pubblicazione: (2025)
AVadCLIP: Audio-Visual Collaboration for Robust Video Anomaly Detection
di: Wu, Peng, et al.
Pubblicazione: (2025)
di: Wu, Peng, et al.
Pubblicazione: (2025)
Scalable and Realistic Virtual Try-on Application for Foundation Makeup with Kubelka-Munk Theory
di: Pang, Hui, et al.
Pubblicazione: (2025)
di: Pang, Hui, et al.
Pubblicazione: (2025)
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models
di: Wang, Zeqing, et al.
Pubblicazione: (2025)
di: Wang, Zeqing, et al.
Pubblicazione: (2025)
SAM Encoder Breach by Adversarial Simplicial Complex Triggers Downstream Model Failures
di: Qin, Yi, et al.
Pubblicazione: (2025)
di: Qin, Yi, et al.
Pubblicazione: (2025)
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation
di: Zhang, Zongye, et al.
Pubblicazione: (2025)
di: Zhang, Zongye, et al.
Pubblicazione: (2025)
A Deep Learning Approach to Identify Rock Bolts in Complex 3D Point Clouds of Underground Mines Captured Using Mobile Laser Scanners
di: Patra, Dibyayan, et al.
Pubblicazione: (2025)
di: Patra, Dibyayan, et al.
Pubblicazione: (2025)
Supersampling of Data from Structured-light Scanner with Deep Learning
di: Melicherčík, Martin, et al.
Pubblicazione: (2023)
di: Melicherčík, Martin, et al.
Pubblicazione: (2023)
Group Activity Recognition using Unreliable Tracked Pose
di: Thilakarathne, Haritha, et al.
Pubblicazione: (2024)
di: Thilakarathne, Haritha, et al.
Pubblicazione: (2024)
Deepfake Detection Generalization with Diffusion Noise
di: Qi, Hongyuan, et al.
Pubblicazione: (2026)
di: Qi, Hongyuan, et al.
Pubblicazione: (2026)
Towards Integrated Rock Support Visualisation in 3D Point Cloud of Underground Mines
di: Patra, Dibyayan, et al.
Pubblicazione: (2026)
di: Patra, Dibyayan, et al.
Pubblicazione: (2026)
Video-Based Human Pose Regression via Decoupled Space-Time Aggregation
di: He, Jijie, et al.
Pubblicazione: (2024)
di: He, Jijie, et al.
Pubblicazione: (2024)
HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping
di: Tushar, Zahid Hassan, et al.
Pubblicazione: (2026)
di: Tushar, Zahid Hassan, et al.
Pubblicazione: (2026)
KNN Transformer with Pyramid Prompts for Few-Shot Learning
di: Li, Wenhao, et al.
Pubblicazione: (2024)
di: Li, Wenhao, et al.
Pubblicazione: (2024)
EUFCC-340K: A Faceted Hierarchical Dataset for Metadata Annotation in GLAM Collections
di: Net, Francesc, et al.
Pubblicazione: (2024)
di: Net, Francesc, et al.
Pubblicazione: (2024)
Facial Spatiotemporal Graphs: Leveraging the 3D Facial Surface for Remote Physiological Measurement
di: Cantrill, Sam, et al.
Pubblicazione: (2026)
di: Cantrill, Sam, et al.
Pubblicazione: (2026)
EarthVL: A Progressive Earth Vision-Language Understanding and Generation Framework
di: Wang, Junjue, et al.
Pubblicazione: (2026)
di: Wang, Junjue, et al.
Pubblicazione: (2026)
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
di: Patra, Dibyayan, et al.
Pubblicazione: (2026)
di: Patra, Dibyayan, et al.
Pubblicazione: (2026)
Skeletonization-Based Adversarial Perturbations on Large Vision Language Model's Mathematical Text Recognition
di: Yoshida, Masatomo, et al.
Pubblicazione: (2026)
di: Yoshida, Masatomo, et al.
Pubblicazione: (2026)
Evaluating the Significance of Outdoor Advertising from Driver's Perspective Using Computer Vision
di: Černeková, Zuzana, et al.
Pubblicazione: (2023)
di: Černeková, Zuzana, et al.
Pubblicazione: (2023)
CMAB: A First National-Scale Multi-Attribute Building Dataset in China Derived from Open Source Data and GeoAI
di: Zhang, Yecheng, et al.
Pubblicazione: (2024)
di: Zhang, Yecheng, et al.
Pubblicazione: (2024)
Orientation-conditioned Facial Texture Mapping for Video-based Facial Remote Photoplethysmography Estimation
di: Cantrill, Sam, et al.
Pubblicazione: (2024)
di: Cantrill, Sam, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
di: Wu, Songhan
Pubblicazione: (2025) -
Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning
di: Kunzo, Tomáš, et al.
Pubblicazione: (2023) -
Cross-View-Prediction: Exploring Contrastive Feature for Hyperspectral Image Classification
di: Zhang, Anyu, et al.
Pubblicazione: (2022) -
TDIP: Tunable Deep Image Processing, a Real Time Melt Pool Monitoring Solution
di: Akhavan, Javid, et al.
Pubblicazione: (2024) -
Distant Object Localisation from Noisy Image Segmentation Sequences
di: Pesonen, Julius, et al.
Pubblicazione: (2025)