Generative Active Learning for Long-tailed Instance Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Muzhi, Fan, Chengxiang, Chen, Hao, Liu, Yang, Mao, Weian, Xu, Xiaogang, Shen, Chunhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DiverGen: Improving Instance Segmentation by Learning Wider Data Distribution with More Diverse Generative Data
von: Fan, Chengxiang, et al.
Veröffentlicht: (2024)
von: Fan, Chengxiang, et al.
Veröffentlicht: (2024)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
A Simple Image Segmentation Framework via In-Context Examples
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?
von: Li, Liyang, et al.
Veröffentlicht: (2026)
von: Li, Liyang, et al.
Veröffentlicht: (2026)
Unified Open-World Segmentation with Multi-Modal Prompts
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
What Matters When Repurposing Diffusion Models for General Dense Perception Tasks?
von: Xu, Guangkai, et al.
Veröffentlicht: (2024)
von: Xu, Guangkai, et al.
Veröffentlicht: (2024)
Active-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
SegAgent: Exploring Pixel Understanding Capabilities in MLLMs by Imitating Human Annotator Trajectories
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
von: Zhong, Hao, et al.
Veröffentlicht: (2025)
von: Zhong, Hao, et al.
Veröffentlicht: (2025)
MovieDreamer: Hierarchical Generation for Coherent Long Visual Sequence
von: Zhao, Canyu, et al.
Veröffentlicht: (2024)
von: Zhao, Canyu, et al.
Veröffentlicht: (2024)
Exploring Spatial Intelligence from a Generative Perspective
von: Zhu, Muzhi, et al.
Veröffentlicht: (2026)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2026)
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
von: Chen, Cong, et al.
Veröffentlicht: (2025)
von: Chen, Cong, et al.
Veröffentlicht: (2025)
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering
von: Jia, Yiduo, et al.
Veröffentlicht: (2026)
von: Jia, Yiduo, et al.
Veröffentlicht: (2026)
LTRL: Boosting Long-tail Recognition via Reflective Learning
von: Zhao, Qihao, et al.
Veröffentlicht: (2024)
von: Zhao, Qihao, et al.
Veröffentlicht: (2024)
DICEPTION: A Generalist Diffusion Model for Visual Perceptual Tasks
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model
von: Park, Daehee, et al.
Veröffentlicht: (2025)
von: Park, Daehee, et al.
Veröffentlicht: (2025)
Preserving Source Video Realism: High-Fidelity Face Swapping for Cinematic Quality
von: Luo, Zekai, et al.
Veröffentlicht: (2025)
von: Luo, Zekai, et al.
Veröffentlicht: (2025)
LTGC: Long-tail Recognition via Leveraging LLMs-driven Generated Content
von: Zhao, Qihao, et al.
Veröffentlicht: (2024)
von: Zhao, Qihao, et al.
Veröffentlicht: (2024)
Frequency-based Matcher for Long-tailed Semantic Segmentation
von: Li, Shan, et al.
Veröffentlicht: (2024)
von: Li, Shan, et al.
Veröffentlicht: (2024)
Recurrent Generic Contour-based Instance Segmentation with Progressive Learning
von: Feng, Hao, et al.
Veröffentlicht: (2023)
von: Feng, Hao, et al.
Veröffentlicht: (2023)
GAInS: Gradient Anomaly-aware Biomedical Instance Segmentation
von: Liu, Runsheng, et al.
Veröffentlicht: (2024)
von: Liu, Runsheng, et al.
Veröffentlicht: (2024)
Bridge Thinking and Acting: Unleashing Physical Potential of VLM with Generalizable Action Expert
von: Liu, Mingyu, et al.
Veröffentlicht: (2025)
von: Liu, Mingyu, et al.
Veröffentlicht: (2025)
Instance Brownian Bridge as Texts for Open-vocabulary Video Instance Segmentation
von: Cheng, Zesen, et al.
Veröffentlicht: (2024)
von: Cheng, Zesen, et al.
Veröffentlicht: (2024)
FC-VFI: Faithful and Consistent Video Frame Interpolation for High-FPS Slow Motion Video Generation
von: Ding, Ganggui, et al.
Veröffentlicht: (2026)
von: Ding, Ganggui, et al.
Veröffentlicht: (2026)
Decision Boundary-aware Generation for Long-tailed Learning
von: Yang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Yang, Jiacheng, et al.
Veröffentlicht: (2026)
FreeCompose: Generic Zero-Shot Image Composition with Diffusion Prior
von: Chen, Zhekai, et al.
Veröffentlicht: (2024)
von: Chen, Zhekai, et al.
Veröffentlicht: (2024)
Active Learning Enabled Low-cost Cell Image Segmentation Using Bounding Box Annotation
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
Scale Disparity of Instances in Interactive Point Cloud Segmentation
von: Han, Chenrui, et al.
Veröffentlicht: (2024)
von: Han, Chenrui, et al.
Veröffentlicht: (2024)
Cell Instance Segmentation: The Devil Is in the Boundaries
von: Liang, Peixian, et al.
Veröffentlicht: (2025)
von: Liang, Peixian, et al.
Veröffentlicht: (2025)
Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter
von: Chen, Zhiyang, et al.
Veröffentlicht: (2025)
von: Chen, Zhiyang, et al.
Veröffentlicht: (2025)
Synthetic Instance Segmentation from Semantic Image Segmentation Masks
von: Shen, Yuchen, et al.
Veröffentlicht: (2023)
von: Shen, Yuchen, et al.
Veröffentlicht: (2023)
Uncertainty-aware Sampling for Long-tailed Semi-supervised Learning
von: Yang, Kuo, et al.
Veröffentlicht: (2024)
von: Yang, Kuo, et al.
Veröffentlicht: (2024)
SAM-guided Graph Cut for 3D Instance Segmentation
von: Guo, Haoyu, et al.
Veröffentlicht: (2023)
von: Guo, Haoyu, et al.
Veröffentlicht: (2023)
NoTVLA: Semantics-Preserving Robot Adaptation via Narrative Action Interfaces
von: Huang, Zheng, et al.
Veröffentlicht: (2025)
von: Huang, Zheng, et al.
Veröffentlicht: (2025)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
Towards Instance Segmentation with Polygon Detection Transformers
von: Sun, Jiacheng, et al.
Veröffentlicht: (2026)
von: Sun, Jiacheng, et al.
Veröffentlicht: (2026)
SDI-Paste: Synthetic Dynamic Instance Copy-Paste for Video Instance Segmentation
von: Shrestha, Sahir, et al.
Veröffentlicht: (2024)
von: Shrestha, Sahir, et al.
Veröffentlicht: (2024)
Causal Disentanglement for Robust Long-tail Medical Image Generation
von: Nie, Weizhi, et al.
Veröffentlicht: (2025)
von: Nie, Weizhi, et al.
Veröffentlicht: (2025)
Gradient-Aware Logit Adjustment Loss for Long-tailed Classifier
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DiverGen: Improving Instance Segmentation by Learning Wider Data Distribution with More Diverse Generative Data
von: Fan, Chengxiang, et al.
Veröffentlicht: (2024) -
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
von: Liu, Yang, et al.
Veröffentlicht: (2023) -
A Simple Image Segmentation Framework via In-Context Examples
von: Liu, Yang, et al.
Veröffentlicht: (2024) -
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024) -
Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?
von: Li, Liyang, et al.
Veröffentlicht: (2026)