Reference Twice: A Simple and Unified Baseline for Few-Shot Instance Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Yue, Zhang, Jiangning, Wang, Yabiao, Wang, Chengjie, Liu, Yong, Qi, Lu, Li, Xiangtai, Yang, Ming-Hsuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
von: Hu, Teng, et al.
Veröffentlicht: (2023)
von: Hu, Teng, et al.
Veröffentlicht: (2023)
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm
von: Zhang, Jiangning, et al.
Veröffentlicht: (2022)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2022)
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
von: He, Qingdong, et al.
Veröffentlicht: (2024)
von: He, Qingdong, et al.
Veröffentlicht: (2024)
EMOv2: Pushing 5M Vision Model Frontier
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
A Generalist FaceX via Learning Unified Facial Representation
von: Han, Yue, et al.
Veröffentlicht: (2023)
von: Han, Yue, et al.
Veröffentlicht: (2023)
Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
Unified Dense Prediction of Video Diffusion
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
von: Yang, Lehan, et al.
Veröffentlicht: (2025)
SemFlow: Binding Semantic Segmentation and Image Synthesis via Rectified Flow
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Few-Shot Learning for Annotation-Efficient Nucleus Instance Segmentation
von: Ming, Yu, et al.
Veröffentlicht: (2024)
von: Ming, Yu, et al.
Veröffentlicht: (2024)
Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning
von: He, Qingdong, et al.
Veröffentlicht: (2025)
von: He, Qingdong, et al.
Veröffentlicht: (2025)
UltraVideo: High-Quality UHD Video Dataset with Comprehensive Captions
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
DynamicControl: Adaptive Condition Selection for Improved Text-to-Image Generation
von: He, Qingdong, et al.
Veröffentlicht: (2024)
von: He, Qingdong, et al.
Veröffentlicht: (2024)
Few-Shot Referring Video Single- and Multi-Object Segmentation via Cross-Modal Affinity with Instance Sequence Matching
von: Liu, Heng, et al.
Veröffentlicht: (2025)
von: Liu, Heng, et al.
Veröffentlicht: (2025)
Referring Expression Instance Retrieval and A Strong End-to-End Baseline
von: Hao, Xiangzhao, et al.
Veröffentlicht: (2025)
von: Hao, Xiangzhao, et al.
Veröffentlicht: (2025)
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
von: Huang, Kuan-Chih, et al.
Veröffentlicht: (2024)
A Simple Baseline with Single-encoder for Referring Image Segmentation
von: Yu, Seonghoon, et al.
Veröffentlicht: (2024)
von: Yu, Seonghoon, et al.
Veröffentlicht: (2024)
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection
von: Chen, Xuhai, et al.
Veröffentlicht: (2023)
von: Chen, Xuhai, et al.
Veröffentlicht: (2023)
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
Decoupling Classifier for Boosting Few-shot Object Detection and Instance Segmentation
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
PVG: Progressive Vision Graph for Vision Recognition
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
PiT: Progressive Diffusion Transformer
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
SwiftVideo: A Unified Framework for Few-Step Video Generation through Trajectory-Distribution Alignment
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
DC-SAM: In-Context Segment Anything in Images and Videos via Dual Consistency
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Learning Unified Reference Representation for Unsupervised Multi-class Anomaly Detection
von: He, Liren, et al.
Veröffentlicht: (2024)
von: He, Liren, et al.
Veröffentlicht: (2024)
SAM-IF: Leveraging SAM for Incremental Few-Shot Instance Segmentation
von: Zhou, Xudong, et al.
Veröffentlicht: (2024)
von: Zhou, Xudong, et al.
Veröffentlicht: (2024)
UIFormer: A Unified Transformer-based Framework for Incremental Few-Shot Object Detection and Instance Segmentation
von: Zhang, Chengyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Chengyuan, et al.
Veröffentlicht: (2024)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
Face Adapter for Pre-Trained Diffusion Models with Fine-Grained ID and Attribute Control
von: Han, Yue, et al.
Veröffentlicht: (2024)
von: Han, Yue, et al.
Veröffentlicht: (2024)
Few-Shot Anomaly-Driven Generation for Anomaly Classification and Segmentation
von: Gui, Guan, et al.
Veröffentlicht: (2025)
von: Gui, Guan, et al.
Veröffentlicht: (2025)
Learning Feature Inversion for Multi-class Anomaly Detection under General-purpose COCO-AD Benchmark
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
SimToken: A Simple Baseline for Referring Audio-Visual Segmentation
von: Jin, Dian, et al.
Veröffentlicht: (2025)
von: Jin, Dian, et al.
Veröffentlicht: (2025)
Unify the Views: View-Consistent Prototype Learning for Few-Shot Segmentation
von: Liu, Hongli, et al.
Veröffentlicht: (2026)
von: Liu, Hongli, et al.
Veröffentlicht: (2026)
Explore In-Context Segmentation via Latent Diffusion Models
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
Video Prediction Transformers without Recurrence or Convolution
von: Tang, Yujin, et al.
Veröffentlicht: (2024)
von: Tang, Yujin, et al.
Veröffentlicht: (2024)
Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything Model
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
PanopticPartFormer++: A Unified and Decoupled View for Panoptic Part Segmentation
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
A Simple and Efficient Baseline for Zero-Shot Generative Classification
von: Qi, Zipeng, et al.
Veröffentlicht: (2024)
von: Qi, Zipeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023) -
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
von: Hu, Teng, et al.
Veröffentlicht: (2023) -
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm
von: Zhang, Jiangning, et al.
Veröffentlicht: (2022) -
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
von: He, Qingdong, et al.
Veröffentlicht: (2024) -
EMOv2: Pushing 5M Vision Model Frontier
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)