MaskDiff: Modeling Mask Distribution with Diffusion Probabilistic Model for Few-Shot Instance Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Minh-Quan, Nguyen, Tam V., Le, Trung-Nghia, Do, Thanh-Toan, Do, Minh N., Tran, Minh-Triet |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CamoFA: A Learnable Fourier-based Augmentation for Camouflage Segmentation
by: Le, Minh-Quan, et al.
Published: (2023)
by: Le, Minh-Quan, et al.
Published: (2023)
CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing
by: Vo, Dinh-Khoi, et al.
Published: (2025)
by: Vo, Dinh-Khoi, et al.
Published: (2025)
GUNNEL: Guided Mixup Augmentation and Multi-Model Fusion for Aquatic Animal Segmentation
by: Le, Minh-Quan, et al.
Published: (2021)
by: Le, Minh-Quan, et al.
Published: (2021)
Interactive Interface For Semantic Segmentation Dataset Synthesis
by: Tran, Ngoc-Do, et al.
Published: (2025)
by: Tran, Ngoc-Do, et al.
Published: (2025)
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
by: Hoang, Trong-Vu, et al.
Published: (2025)
by: Hoang, Trong-Vu, et al.
Published: (2025)
Efficient 3D Brain Tumor Segmentation with Axial-Coronal-Sagittal Embedding
by: Huynh, Tuan-Luc, et al.
Published: (2025)
by: Huynh, Tuan-Luc, et al.
Published: (2025)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
TaleForge: Interactive Multimodal System for Personalized Story Creation
by: Nguyen, Minh-Loi, et al.
Published: (2025)
by: Nguyen, Minh-Loi, et al.
Published: (2025)
PANDORA: Pixel-wise Attention Dissolution and Latent Guidance for Zero-Shot Object Removal
by: Vo, Dinh-Khoi, et al.
Published: (2026)
by: Vo, Dinh-Khoi, et al.
Published: (2026)
Automated Image Recognition Framework
by: Nguyen, Quang-Binh, et al.
Published: (2025)
by: Nguyen, Quang-Binh, et al.
Published: (2025)
Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
EDGER: EDge-Guided with HEatmap Refinement for Generalizable Image Forgery Localization
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
Graphilosophy: Graph-Based Digital Humanities Computing with The Four Books
by: Do, Minh-Thu, et al.
Published: (2026)
by: Do, Minh-Thu, et al.
Published: (2026)
GenKOL: Modular Generative AI Framework For Scalable Virtual KOL Generation
by: To, Tan-Hiep, et al.
Published: (2025)
by: To, Tan-Hiep, et al.
Published: (2025)
KiseKloset for Fashion Retrieval and Recommendation
by: Phan-Nguyen, Thanh-Tung, et al.
Published: (2025)
by: Phan-Nguyen, Thanh-Tung, et al.
Published: (2025)
EventCap
by: Nguyen, Phuc-Tan, et al.
Published: (2024)
by: Nguyen, Phuc-Tan, et al.
Published: (2024)
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
by: Nguyen, Hieu, et al.
Published: (2025)
by: Nguyen, Hieu, et al.
Published: (2025)
Enhancing Dataset Distillation via Non-Critical Region Refinement
by: Tran, Minh-Tuan, et al.
Published: (2025)
by: Tran, Minh-Tuan, et al.
Published: (2025)
MasHeNe: A Benchmark for Head and Neck CT Mass Segmentation using Window-Enhanced Mamba with Frequency-Domain Integration
by: Dao, Thao Thi Phuong, et al.
Published: (2025)
by: Dao, Thao Thi Phuong, et al.
Published: (2025)
Toward Content-based Indexing and Retrieval of Head and Neck CT with Abscess Segmentation
by: Dao, Thao Thi Phuong, et al.
Published: (2025)
by: Dao, Thao Thi Phuong, et al.
Published: (2025)
Conditional Distribution Modelling for Few-Shot Image Synthesis with Diffusion Models
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
Event-Enriched Image Analysis Grand Challenge at ACM Multimedia 2025
by: Tran, Thien-Phuc, et al.
Published: (2025)
by: Tran, Thien-Phuc, et al.
Published: (2025)
Amodal Instance Segmentation with Diffusion Shape Prior Estimation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
SAMURAI: Shape-Aware Multimodal Retrieval for 3D Object Identification
by: Vo, Dinh-Khoi, et al.
Published: (2025)
by: Vo, Dinh-Khoi, et al.
Published: (2025)
GenFlow: Interactive Modular System for Image Generation
by: Nguyen, Duc-Hung, et al.
Published: (2025)
by: Nguyen, Duc-Hung, et al.
Published: (2025)
EVENT-Retriever: Event-Aware Multimodal Image Retrieval for Realistic Captions
by: Vo, Dinh-Khoi, et al.
Published: (2025)
by: Vo, Dinh-Khoi, et al.
Published: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
Ascent and descent of Artinian module structures under flat base changes
by: Chau, Tran Do Minh, et al.
Published: (2025)
by: Chau, Tran Do Minh, et al.
Published: (2025)
Geometric influence of width ratio and contraction ratio on droplet dynamics in microchannel using a 3D numerical simulation
by: Le Hung Toan Do, et al.
Published: (2024)
by: Le Hung Toan Do, et al.
Published: (2024)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
by: Nguyen, Hieu, et al.
Published: (2024)
by: Nguyen, Hieu, et al.
Published: (2024)
iCONTRA: Toward Thematic Collection Design Via Interactive Concept Transfer
by: Vo, Dinh-Khoi, et al.
Published: (2024)
by: Vo, Dinh-Khoi, et al.
Published: (2024)
Learning Structural Causal Models from Ordering: Identifiable Flow Models
by: Le, Minh Khoa, et al.
Published: (2024)
by: Le, Minh Khoa, et al.
Published: (2024)
DiffAugment: Diffusion based Long-Tailed Visual Relationship Recognition
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
Shape2Animal: Creative Animal Generation from Natural Silhouettes
by: Tran, Quoc-Duy, et al.
Published: (2025)
by: Tran, Quoc-Duy, et al.
Published: (2025)
VietMEAgent: Culturally-Aware Few-Shot Multimodal Explanation for Vietnamese Visual Question Answering
by: Nguyen, Hai-Dang, et al.
Published: (2025)
by: Nguyen, Hai-Dang, et al.
Published: (2025)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
by: Nguyen, Phuc, et al.
Published: (2024)
by: Nguyen, Phuc, et al.
Published: (2024)
VisionGuard: Synergistic Framework for Helmet Violation Detection
by: Nguyen, Lam-Huy, et al.
Published: (2025)
by: Nguyen, Lam-Huy, et al.
Published: (2025)
ReCap: Event-Aware Image Captioning with Article Retrieval and Semantic Gaussian Normalization
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
by: Nguyen, Thinh-Phuc, et al.
Published: (2025)
ARtVista: Gateway To Empower Anyone Into Artist
by: Hoang, Trong-Vu, et al.
Published: (2024)
by: Hoang, Trong-Vu, et al.
Published: (2024)
Vietnamese Poem Generation & The Prospect Of Cross-Language Poem-To-Poem Translation
by: Huynh, Triet Minh, et al.
Published: (2024)
by: Huynh, Triet Minh, et al.
Published: (2024)
Similar Items
-
CamoFA: A Learnable Fourier-based Augmentation for Camouflage Segmentation
by: Le, Minh-Quan, et al.
Published: (2023) -
CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing
by: Vo, Dinh-Khoi, et al.
Published: (2025) -
GUNNEL: Guided Mixup Augmentation and Multi-Model Fusion for Aquatic Animal Segmentation
by: Le, Minh-Quan, et al.
Published: (2021) -
Interactive Interface For Semantic Segmentation Dataset Synthesis
by: Tran, Ngoc-Do, et al.
Published: (2025) -
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
by: Hoang, Trong-Vu, et al.
Published: (2025)