MixReorg: Cross-Modal Mixed Patch Reorganization is a Good Mask Learner for Open-World Semantic Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Cai, Kaixin, Ren, Pengzhen, Zhu, Yi, Xu, Hang, Liu, Jianzhuang, Li, Changlin, Wang, Guangrun, Liang, Xiaodan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MagicSeg: Open-World Segmentation Pretraining via Counterfactural Diffusion-Based Auto-Generation
di: Cai, Kaixin, et al.
Pubblicazione: (2026)
di: Cai, Kaixin, et al.
Pubblicazione: (2026)
ReorgGS: Equivalent Distribution Reorganization for 3D Gaussian Splatting
di: Wang, Luchao, et al.
Pubblicazione: (2026)
di: Wang, Luchao, et al.
Pubblicazione: (2026)
PIVOT-R: Primitive-Driven Waypoint-Aware World Model for Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2024)
di: Zhang, Kaidong, et al.
Pubblicazione: (2024)
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
Multi-Modal Prototypes for Open-World Semantic Segmentation
di: Yang, Yuhuan, et al.
Pubblicazione: (2023)
di: Yang, Yuhuan, et al.
Pubblicazione: (2023)
Visible-Infrared Person Re-Identification via Patch-Mixed Cross-Modality Learning
di: Qian, Zhihao, et al.
Pubblicazione: (2023)
di: Qian, Zhihao, et al.
Pubblicazione: (2023)
DNA Family: Boosting Weight-Sharing NAS with Block-Wise Supervisions
di: Wang, Guangrun, et al.
Pubblicazione: (2024)
di: Wang, Guangrun, et al.
Pubblicazione: (2024)
MLP Can Be A Good Transformer Learner
di: Lin, Sihao, et al.
Pubblicazione: (2024)
di: Lin, Sihao, et al.
Pubblicazione: (2024)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
di: Ren, Pengzhen, et al.
Pubblicazione: (2023)
di: Ren, Pengzhen, et al.
Pubblicazione: (2023)
Actional Atomic-Concept Learning for Demystifying Vision-Language Navigation
di: Lin, Bingqian, et al.
Pubblicazione: (2023)
di: Lin, Bingqian, et al.
Pubblicazione: (2023)
SCMix: Stochastic Compound Mixing for Open Compound Domain Adaptation in Semantic Segmentation
di: Yao, Kai, et al.
Pubblicazione: (2024)
di: Yao, Kai, et al.
Pubblicazione: (2024)
Semantic Segmentation Prior for Diffusion-Based Real-World Super-Resolution
di: Xiao, Jiahua, et al.
Pubblicazione: (2024)
di: Xiao, Jiahua, et al.
Pubblicazione: (2024)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
di: Zhang, Zelin, et al.
Pubblicazione: (2026)
di: Zhang, Zelin, et al.
Pubblicazione: (2026)
Best-of-Both-Worlds Fair Allocation of Indivisible and Mixed Goods
di: Bu, Xiaolin, et al.
Pubblicazione: (2024)
di: Bu, Xiaolin, et al.
Pubblicazione: (2024)
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2025)
di: Zhang, Kaidong, et al.
Pubblicazione: (2025)
MiPa: Mixed Patch Infrared-Visible Modality Agnostic Object Detection
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
Semi-supervised 3D Object Detection with PatchTeacher and PillarMix
di: Wu, Xiaopei, et al.
Pubblicazione: (2024)
di: Wu, Xiaopei, et al.
Pubblicazione: (2024)
PatchMixer: A Patch-Mixing Architecture for Long-Term Time Series Forecasting
di: Gong, Zeying, et al.
Pubblicazione: (2023)
di: Gong, Zeying, et al.
Pubblicazione: (2023)
Closing the Modality Gap for Mixed Modality Search
di: Li, Binxu, et al.
Pubblicazione: (2025)
di: Li, Binxu, et al.
Pubblicazione: (2025)
OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling
di: Chen, Hongyu, et al.
Pubblicazione: (2026)
di: Chen, Hongyu, et al.
Pubblicazione: (2026)
From Cross-Modal to Mixed-Modal Visible-Infrared Re-Identification
di: Alehdaghi, Mahdi, et al.
Pubblicazione: (2025)
di: Alehdaghi, Mahdi, et al.
Pubblicazione: (2025)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
di: Xie, Pengzhen, et al.
Pubblicazione: (2025)
di: Xie, Pengzhen, et al.
Pubblicazione: (2025)
CM-MaskSD: Cross-Modality Masked Self-Distillation for Referring Image Segmentation
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation
di: Jie, Pengyu, et al.
Pubblicazione: (2025)
di: Jie, Pengyu, et al.
Pubblicazione: (2025)
Multimodal Cross-Document Event Coreference Resolution Using Linear Semantic Transfer and Mixed-Modality Ensembles
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
Correctable Landmark Discovery via Large Models for Vision-Language Navigation
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
di: Wang, Jiaze, et al.
Pubblicazione: (2024)
di: Wang, Jiaze, et al.
Pubblicazione: (2024)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
di: Wang, Ziyi, et al.
Pubblicazione: (2024)
di: Wang, Ziyi, et al.
Pubblicazione: (2024)
GS: Generative Segmentation via Label Diffusion
di: Chen, Yuhao, et al.
Pubblicazione: (2025)
di: Chen, Yuhao, et al.
Pubblicazione: (2025)
Rejection Mixing: Fast Semantic Propagation of Mask Tokens for Efficient DLLM Inference
di: Ye, Yushi, et al.
Pubblicazione: (2026)
di: Ye, Yushi, et al.
Pubblicazione: (2026)
CP2M: Clustered-Patch-Mixed Mosaic Augmentation for Aerial Image Segmentation
di: Li, Yijie, et al.
Pubblicazione: (2025)
di: Li, Yijie, et al.
Pubblicazione: (2025)
MixPolyp: Integrating Mask, Box and Scribble Supervision for Enhanced Polyp Segmentation
di: Hu, Yiwen, et al.
Pubblicazione: (2024)
di: Hu, Yiwen, et al.
Pubblicazione: (2024)
Unified Open-World Segmentation with Multi-Modal Prompts
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Unleashing the Electromechanical Response of Ferroelastic Domain Reorganization in Mixed‐Phase Tetragonal Ferroelectric Multilayers
di: Zishen Tian, et al.
Pubblicazione: (2026)
di: Zishen Tian, et al.
Pubblicazione: (2026)
OmniPatch: A Universal Adversarial Patch for ViT-CNN Cross-Architecture Transfer in Semantic Segmentation
di: Aggarwal, Aarush, et al.
Pubblicazione: (2026)
di: Aggarwal, Aarush, et al.
Pubblicazione: (2026)
Approval-Based Voting with Mixed Goods
di: Lu, Xinhang, et al.
Pubblicazione: (2022)
di: Lu, Xinhang, et al.
Pubblicazione: (2022)
OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation
di: Zhang, Dengke, et al.
Pubblicazione: (2024)
di: Zhang, Dengke, et al.
Pubblicazione: (2024)
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
di: Stojnić, Vladan, et al.
Pubblicazione: (2025)
di: Stojnić, Vladan, et al.
Pubblicazione: (2025)
Semantic-guided Masked Mutual Learning for Multi-modal Brain Tumor Segmentation with Arbitrary Missing Modalities
di: Liang, Guoyan, et al.
Pubblicazione: (2025)
di: Liang, Guoyan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MagicSeg: Open-World Segmentation Pretraining via Counterfactural Diffusion-Based Auto-Generation
di: Cai, Kaixin, et al.
Pubblicazione: (2026) -
ReorgGS: Equivalent Distribution Reorganization for 3D Gaussian Splatting
di: Wang, Luchao, et al.
Pubblicazione: (2026) -
PIVOT-R: Primitive-Driven Waypoint-Aware World Model for Robotic Manipulation
di: Zhang, Kaidong, et al.
Pubblicazione: (2024) -
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
di: Zhang, Zicheng, et al.
Pubblicazione: (2024) -
Multi-Modal Prototypes for Open-World Semantic Segmentation
di: Yang, Yuhuan, et al.
Pubblicazione: (2023)