DiverGen: Improving Instance Segmentation by Learning Wider Data Distribution with More Diverse Generative Data
Fuente:
arXiv
Saved in:
| Main Authors: | Fan, Chengxiang, Zhu, Muzhi, Chen, Hao, Liu, Yang, Wu, Weijia, Zhang, Huaqi, Shen, Chunhua |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Active Learning for Long-tailed Instance Segmentation
by: Zhu, Muzhi, et al.
Published: (2024)
by: Zhu, Muzhi, et al.
Published: (2024)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
by: Liu, Yang, et al.
Published: (2023)
by: Liu, Yang, et al.
Published: (2023)
A Simple Image Segmentation Framework via In-Context Examples
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
by: Zhu, Muzhi, et al.
Published: (2024)
by: Zhu, Muzhi, et al.
Published: (2024)
Unified Open-World Segmentation with Multi-Modal Prompts
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
by: Chen, Cong, et al.
Published: (2025)
by: Chen, Cong, et al.
Published: (2025)
DiffuMask: Synthesizing Images with Pixel-level Annotations for Semantic Segmentation Using Diffusion Models
by: Wu, Weijia, et al.
Published: (2023)
by: Wu, Weijia, et al.
Published: (2023)
Data Synthesis Improves 3D Myotube Instance Segmentation
by: Exler, David, et al.
Published: (2026)
by: Exler, David, et al.
Published: (2026)
ScribbleGen: Generative Data Augmentation Improves Scribble-supervised Semantic Segmentation
by: Schnell, Jacob, et al.
Published: (2023)
by: Schnell, Jacob, et al.
Published: (2023)
What Matters When Repurposing Diffusion Models for General Dense Perception Tasks?
by: Xu, Guangkai, et al.
Published: (2024)
by: Xu, Guangkai, et al.
Published: (2024)
SegAgent: Exploring Pixel Understanding Capabilities in MLLMs by Imitating Human Annotator Trajectories
by: Zhu, Muzhi, et al.
Published: (2025)
by: Zhu, Muzhi, et al.
Published: (2025)
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
by: Zhong, Hao, et al.
Published: (2025)
by: Zhong, Hao, et al.
Published: (2025)
ContextGen: Contextual Layout Anchoring for Identity-Consistent Multi-Instance Generation
by: Xu, Ruihang, et al.
Published: (2025)
by: Xu, Ruihang, et al.
Published: (2025)
InstanceGen: Image Generation with Instance-level Instructions
by: Sella, Etai, et al.
Published: (2025)
by: Sella, Etai, et al.
Published: (2025)
Diffusion Models are Efficient Data Generators for Human Mesh Recovery
by: Ge, Yongtao, et al.
Published: (2024)
by: Ge, Yongtao, et al.
Published: (2024)
Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?
by: Li, Liyang, et al.
Published: (2026)
by: Li, Liyang, et al.
Published: (2026)
Fourier Boundary Features Network with Wider Catchers for Glass Segmentation
by: Qin, Xiaolin, et al.
Published: (2024)
by: Qin, Xiaolin, et al.
Published: (2024)
Exploring Spatial Intelligence from a Generative Perspective
by: Zhu, Muzhi, et al.
Published: (2026)
by: Zhu, Muzhi, et al.
Published: (2026)
ReMoGen: Real-time Human Interaction-to-Reaction Generation via Modular Learning from Diverse Data
by: Ye, Yaoqin, et al.
Published: (2026)
by: Ye, Yaoqin, et al.
Published: (2026)
Probabilistic Modeling of Multi-rater Medical Image Segmentation for Diversity and Personalization
by: Liu, Ke, et al.
Published: (2025)
by: Liu, Ke, et al.
Published: (2025)
Recurrent Generic Contour-based Instance Segmentation with Progressive Learning
by: Feng, Hao, et al.
Published: (2023)
by: Feng, Hao, et al.
Published: (2023)
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering
by: Jia, Yiduo, et al.
Published: (2026)
by: Jia, Yiduo, et al.
Published: (2026)
Preserving Source Video Realism: High-Fidelity Face Swapping for Cinematic Quality
by: Luo, Zekai, et al.
Published: (2025)
by: Luo, Zekai, et al.
Published: (2025)
Unsupervised Pre-training with Language-Vision Prompts for Low-Data Instance Segmentation
by: Zhang, Dingwen, et al.
Published: (2024)
by: Zhang, Dingwen, et al.
Published: (2024)
InstaDA: Augmenting Instance Segmentation Data with Dual-Agent System
by: Hou, Xianbao, et al.
Published: (2025)
by: Hou, Xianbao, et al.
Published: (2025)
Improving Pixel Embedding Learning through Intermediate Distance Regression Supervision for Instance Segmentation
by: Wu, Yuli, et al.
Published: (2020)
by: Wu, Yuli, et al.
Published: (2020)
DICEPTION: A Generalist Diffusion Model for Visual Perceptual Tasks
by: Zhao, Canyu, et al.
Published: (2025)
by: Zhao, Canyu, et al.
Published: (2025)
Improving Generalized Visual Grounding with Instance-aware Joint Learning
by: Dai, Ming, et al.
Published: (2025)
by: Dai, Ming, et al.
Published: (2025)
InstanceCap: Improving Text-to-Video Generation via Instance-aware Structured Caption
by: Fan, Tiehan, et al.
Published: (2024)
by: Fan, Tiehan, et al.
Published: (2024)
3D Instance Segmentation Using Deep Learning on RGB-D Indoor Data
by: Yasir, Siddiqui Muhammad, et al.
Published: (2024)
by: Yasir, Siddiqui Muhammad, et al.
Published: (2024)
Watch Wider and Think Deeper: Collaborative Cross-modal Chain-of-Thought for Complex Visual Reasoning
by: Lu, Wenting, et al.
Published: (2026)
by: Lu, Wenting, et al.
Published: (2026)
VimTS: A Unified Video and Image Text Spotter for Enhancing the Cross-domain Generalization
by: Liu, Yuliang, et al.
Published: (2024)
by: Liu, Yuliang, et al.
Published: (2024)
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Bridge Thinking and Acting: Unleashing Physical Potential of VLM with Generalizable Action Expert
by: Liu, Mingyu, et al.
Published: (2025)
by: Liu, Mingyu, et al.
Published: (2025)
DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
by: Jiang, Zhenyu, et al.
Published: (2024)
by: Jiang, Zhenyu, et al.
Published: (2024)
GenClaw: Code-Driven Agentic Image Generation
by: Ye, Junyan, et al.
Published: (2026)
by: Ye, Junyan, et al.
Published: (2026)
Active-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO
by: Zhu, Muzhi, et al.
Published: (2025)
by: Zhu, Muzhi, et al.
Published: (2025)
Synthetic Data Augmentation for Enhanced Chicken Carcass Instance Segmentation
by: Feng, Yihong, et al.
Published: (2025)
by: Feng, Yihong, et al.
Published: (2025)
NucleiMix: Realistic Data Augmentation for Nuclei Instance Segmentation
by: Wang, Jiamu, et al.
Published: (2024)
by: Wang, Jiamu, et al.
Published: (2024)
GeoPurify: A Data-Efficient Geometric Distillation Framework for Open-Vocabulary 3D Segmentation
by: Dou, Weijia, et al.
Published: (2025)
by: Dou, Weijia, et al.
Published: (2025)
Similar Items
-
Generative Active Learning for Long-tailed Instance Segmentation
by: Zhu, Muzhi, et al.
Published: (2024) -
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
by: Liu, Yang, et al.
Published: (2023) -
A Simple Image Segmentation Framework via In-Context Examples
by: Liu, Yang, et al.
Published: (2024) -
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
by: Zhu, Muzhi, et al.
Published: (2024) -
Unified Open-World Segmentation with Multi-Modal Prompts
by: Liu, Yang, et al.
Published: (2025)