Rethinking MLLM Itself as a Segmenter with a Single Segmentation Token
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Anqi, Ji, Xiaokang, Gao, Guangyu, Jiao, Jianbo, Liu, Chi Harold, Wei, Yunchao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridge the Points: Graph-based Few-shot Segment Anything Semantically
by: Zhang, Anqi, et al.
Published: (2024)
by: Zhang, Anqi, et al.
Published: (2024)
CoMBO: Conflict Mitigation via Branched Optimization for Class Incremental Segmentation
by: Fang, Kai, et al.
Published: (2025)
by: Fang, Kai, et al.
Published: (2025)
FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation
by: Zhang, Zekang, et al.
Published: (2026)
by: Zhang, Zekang, et al.
Published: (2026)
Background Adaptation with Residual Modeling for Exemplar-Free Class-Incremental Semantic Segmentation
by: Zhang, Anqi, et al.
Published: (2024)
by: Zhang, Anqi, et al.
Published: (2024)
Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation
by: Liang, Chen, et al.
Published: (2021)
by: Liang, Chen, et al.
Published: (2021)
Collaborative Vision-Text Representation Optimizing for Open-Vocabulary Segmentation
by: Jiao, Siyu, et al.
Published: (2024)
by: Jiao, Siyu, et al.
Published: (2024)
Surface-SOS: Self-Supervised Object Segmentation via Neural Surface Representation
by: Zheng, Xiaoyun, et al.
Published: (2025)
by: Zheng, Xiaoyun, et al.
Published: (2025)
One-shot In-context Part Segmentation
by: Dai, Zhenqi, et al.
Published: (2025)
by: Dai, Zhenqi, et al.
Published: (2025)
I-MedSAM: Implicit Medical Image Segmentation with Segment Anything
by: Wei, Xiaobao, et al.
Published: (2023)
by: Wei, Xiaobao, et al.
Published: (2023)
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation
by: Zhang, Bingfeng, et al.
Published: (2024)
by: Zhang, Bingfeng, et al.
Published: (2024)
IPSeg: Image Posterior Mitigates Semantic Drift in Class-Incremental Segmentation
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
Beyond the Label Itself: Latent Labels Enhance Semi-supervised Point Cloud Panoptic Segmentation
by: Chen, Yujun, et al.
Published: (2023)
by: Chen, Yujun, et al.
Published: (2023)
Moment and Highlight Detection via MLLM Frame Segmentation
by: Jiwanta, I Putu Andika Bagas, et al.
Published: (2025)
by: Jiwanta, I Putu Andika Bagas, et al.
Published: (2025)
PyramidMamba: Rethinking Pyramid Feature Fusion with Selective Space State Model for Semantic Segmentation of Remote Sensing Imagery
by: Wang, Libo, et al.
Published: (2024)
by: Wang, Libo, et al.
Published: (2024)
A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis
by: Lin, Dongheng, et al.
Published: (2025)
by: Lin, Dongheng, et al.
Published: (2025)
CADFormer: Fine-Grained Cross-modal Alignment and Decoding Transformer for Referring Remote Sensing Image Segmentation
by: Liu, Maofu, et al.
Published: (2025)
by: Liu, Maofu, et al.
Published: (2025)
Geospatial-Reasoning-Driven Vocabulary-Agnostic Remote Sensing Semantic Segmentation
by: Zhou, Chufeng, et al.
Published: (2026)
by: Zhou, Chufeng, et al.
Published: (2026)
Scalable Video Object Segmentation with Identification Mechanism
by: Yang, Zongxin, et al.
Published: (2022)
by: Yang, Zongxin, et al.
Published: (2022)
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
Robust MLLM Unlearning via Visual Knowledge Distillation
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
Structure-Aware Feature Rectification with Region Adjacency Graphs for Training-Free Open-Vocabulary Semantic Segmentation
by: Huang, Qiming, et al.
Published: (2025)
by: Huang, Qiming, et al.
Published: (2025)
CoT-Seg: Rethinking Segmentation with Chain-of-Thought Reasoning and Self-Correction
by: Kao, Shiu-hong, et al.
Published: (2026)
by: Kao, Shiu-hong, et al.
Published: (2026)
Tendency-driven Mutual Exclusivity for Weakly Supervised Incremental Semantic Segmentation
by: Si, Chongjie, et al.
Published: (2024)
by: Si, Chongjie, et al.
Published: (2024)
Learning Spectral-Decomposed Tokens for Domain Generalized Semantic Segmentation
by: Yi, Jingjun, et al.
Published: (2024)
by: Yi, Jingjun, et al.
Published: (2024)
AINet+: Advancing Superpixel Segmentation via Cascaded Association Implantation
by: Wang, Yaxiong, et al.
Published: (2021)
by: Wang, Yaxiong, et al.
Published: (2021)
Referring Industrial Anomaly Segmentation
by: Yue, Pengfei, et al.
Published: (2026)
by: Yue, Pengfei, et al.
Published: (2026)
Med-Tuning: A New Parameter-Efficient Tuning Framework for Medical Volumetric Segmentation
by: Shen, Jiachen, et al.
Published: (2023)
by: Shen, Jiachen, et al.
Published: (2023)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
RS3Mamba: Visual State Space Model for Remote Sensing Images Semantic Segmentation
by: Ma, Xianping, et al.
Published: (2024)
by: Ma, Xianping, et al.
Published: (2024)
DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation
by: Yin, Bowen, et al.
Published: (2023)
by: Yin, Bowen, et al.
Published: (2023)
Better, Stronger, Faster: Tackling the Trilemma in MLLM-based Segmentation with Simultaneous Textual Mask Prediction
by: Liu, Jiazhen, et al.
Published: (2025)
by: Liu, Jiazhen, et al.
Published: (2025)
Region-Adaptive Transform with Segmentation Prior for Image Compression
by: Liu, Yuxi, et al.
Published: (2024)
by: Liu, Yuxi, et al.
Published: (2024)
Rethinking Prior Information Generation with CLIP for Few-Shot Segmentation
by: Wang, Jin, et al.
Published: (2024)
by: Wang, Jin, et al.
Published: (2024)
Unlocking 3D Affordance Segmentation with 2D Semantic Knowledge
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
CIT: Rethinking Class-incremental Semantic Segmentation with a Class Independent Transformation
by: Ge, Jinchao, et al.
Published: (2024)
by: Ge, Jinchao, et al.
Published: (2024)
A Unified Framework with Multimodal Fine-tuning for Remote Sensing Semantic Segmentation
by: Ma, Xianping, et al.
Published: (2024)
by: Ma, Xianping, et al.
Published: (2024)
Transferable and Principled Efficiency for Open-Vocabulary Segmentation
by: Xu, Jingxuan, et al.
Published: (2024)
by: Xu, Jingxuan, et al.
Published: (2024)
Few Exemplar-Based General Medical Image Segmentation via Domain-Aware Selective Adaptation
by: Xu, Chen, et al.
Published: (2024)
by: Xu, Chen, et al.
Published: (2024)
TASAM: Terrain-and-Aware Segment Anything Model for Temporal-Scale Remote Sensing Segmentation
by: Wang, Tianyang, et al.
Published: (2025)
by: Wang, Tianyang, et al.
Published: (2025)
Rethinking Query-based Transformer for Continual Image Segmentation
by: Zhu, Yuchen, et al.
Published: (2025)
by: Zhu, Yuchen, et al.
Published: (2025)
Similar Items
-
Bridge the Points: Graph-based Few-shot Segment Anything Semantically
by: Zhang, Anqi, et al.
Published: (2024) -
CoMBO: Conflict Mitigation via Branched Optimization for Class Incremental Segmentation
by: Fang, Kai, et al.
Published: (2025) -
FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation
by: Zhang, Zekang, et al.
Published: (2026) -
Background Adaptation with Residual Modeling for Exemplar-Free Class-Incremental Semantic Segmentation
by: Zhang, Anqi, et al.
Published: (2024) -
Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation
by: Liang, Chen, et al.
Published: (2021)