Rethinking Query-based Transformer for Continual Image Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Yuchen, Shi, Cheng, Wang, Dingyou, Tang, Jiajin, Wei, Zhengxuan, Wu, Yu, Li, Guanbin, Yang, Sibei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sim-DETR: Unlock DETR for Temporal Sentence Grounding
by: Tang, Jiajin, et al.
Published: (2025)
by: Tang, Jiajin, et al.
Published: (2025)
Augmenting Moment Retrieval: Zero-Dependency Two-Stage Learning
by: Wei, Zhengxuan, et al.
Published: (2025)
by: Wei, Zhengxuan, et al.
Published: (2025)
Closed-Loop Transfer for Weakly-supervised Affordance Grounding
by: Tang, Jiajin, et al.
Published: (2025)
by: Tang, Jiajin, et al.
Published: (2025)
Part2Object: Hierarchical Unsupervised 3D Instance Segmentation
by: Shi, Cheng, et al.
Published: (2024)
by: Shi, Cheng, et al.
Published: (2024)
Plain-Det: A Plain Multi-Dataset Object Detector
by: Shi, Cheng, et al.
Published: (2024)
by: Shi, Cheng, et al.
Published: (2024)
Vision Transformers Need More Than Registers
by: Shi, Cheng, et al.
Published: (2026)
by: Shi, Cheng, et al.
Published: (2026)
Curriculum Point Prompting for Weakly-Supervised Referring Image Segmentation
by: Dai, Qiyuan, et al.
Published: (2024)
by: Dai, Qiyuan, et al.
Published: (2024)
The devil is in the object boundary: towards annotation-free instance segmentation using Foundation Models
by: Shi, Cheng, et al.
Published: (2024)
by: Shi, Cheng, et al.
Published: (2024)
Why LVLMs Are More Prone to Hallucinations in Longer Responses: The Role of Context
by: Zheng, Ge, et al.
Published: (2025)
by: Zheng, Ge, et al.
Published: (2025)
Vision Function Layer in Multimodal LLMs
by: Shi, Cheng, et al.
Published: (2025)
by: Shi, Cheng, et al.
Published: (2025)
Self-Prophetic Decoding to Unlock Visual Search in LVLMs
by: He, Zhendong, et al.
Published: (2026)
by: He, Zhendong, et al.
Published: (2026)
Chart Deep Research in LVLMs via Parallel Relative Policy Optimization
by: Tang, Jiajin, et al.
Published: (2026)
by: Tang, Jiajin, et al.
Published: (2026)
WeaveTime: Stream from Earlier Frames into Emergent Memory in VideoLLMs
by: Zhang, Yulin, et al.
Published: (2026)
by: Zhang, Yulin, et al.
Published: (2026)
Continual Alignment for SAM: Rethinking Foundation Models for Medical Image Segmentation in Continual Learning
by: Wang, Jiayi, et al.
Published: (2025)
by: Wang, Jiayi, et al.
Published: (2025)
VLDrive: Vision-Augmented Lightweight MLLMs for Efficient Language-grounded Autonomous Driving
by: Zhang, Ruifei, et al.
Published: (2025)
by: Zhang, Ruifei, et al.
Published: (2025)
Diffusion-based Data Augmentation for Nuclei Image Segmentation
by: Yu, Xinyi, et al.
Published: (2023)
by: Yu, Xinyi, et al.
Published: (2023)
Rethinking Early-Fusion Strategies for Improved Multimodal Image Segmentation
by: Shen, Zhengwen, et al.
Published: (2025)
by: Shen, Zhengwen, et al.
Published: (2025)
SAM2-UNet: Segment Anything 2 Makes Strong Encoder for Natural and Medical Image Segmentation
by: Xiong, Xinyu, et al.
Published: (2024)
by: Xiong, Xinyu, et al.
Published: (2024)
Mixed-Query Transformer: A Unified Image Segmentation Architecture
by: Wang, Pei, et al.
Published: (2024)
by: Wang, Pei, et al.
Published: (2024)
Eyes Wide Open: Ego Proactive Video-LLM for Streaming Video
by: Zhang, Yulin, et al.
Published: (2025)
by: Zhang, Yulin, et al.
Published: (2025)
Intervene-All-Paths: Unified Mitigation of LVLM Hallucinations across Alignment Formats
by: Qian, Jiaye, et al.
Published: (2025)
by: Qian, Jiaye, et al.
Published: (2025)
VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction
by: He, Zijian, et al.
Published: (2025)
by: He, Zijian, et al.
Published: (2025)
UltraImage: Rethinking Resolution Extrapolation in Image Diffusion Transformers
by: Zhao, Min, et al.
Published: (2025)
by: Zhao, Min, et al.
Published: (2025)
Cell Graph Transformer for Nuclei Classification
by: Lou, Wei, et al.
Published: (2024)
by: Lou, Wei, et al.
Published: (2024)
Q2A: Querying Implicit Fully Continuous Feature Pyramid to Align Features for Medical Image Segmentation
by: Yu, Jiahao, et al.
Published: (2024)
by: Yu, Jiahao, et al.
Published: (2024)
Open-Vocabulary Segmentation with Semantic-Assisted Calibration
by: Liu, Yong, et al.
Published: (2023)
by: Liu, Yong, et al.
Published: (2023)
Bearing fault diagnosis based on multi-scale spectral images and convolutional neural network
by: Luo, Tongchao, et al.
Published: (2025)
by: Luo, Tongchao, et al.
Published: (2025)
Semi-supervised Medical Image Segmentation via Query Distribution Consistency
by: Wu, Rong, et al.
Published: (2023)
by: Wu, Rong, et al.
Published: (2023)
Dual-domain Adaptation Networks for Realistic Image Super-resolution
by: Fang, Chaowei, et al.
Published: (2025)
by: Fang, Chaowei, et al.
Published: (2025)
Rethinking Transformer for Long Contextual Histopathology Whole Slide Image Analysis
by: Li, Honglin, et al.
Published: (2024)
by: Li, Honglin, et al.
Published: (2024)
Adaptive Part Learning for Fine-Grained Generalized Category Discovery: A Plug-and-Play Enhancement
by: Dai, Qiyuan, et al.
Published: (2025)
by: Dai, Qiyuan, et al.
Published: (2025)
DreamFuse: Adaptive Image Fusion with Diffusion Transformer
by: Huang, Junjia, et al.
Published: (2025)
by: Huang, Junjia, et al.
Published: (2025)
Dynamic Object Queries for Transformer-based Incremental Object Detection
by: Zhang, Jichuan, et al.
Published: (2024)
by: Zhang, Jichuan, et al.
Published: (2024)
Free on the Fly: Enhancing Flexibility in Test-Time Adaptation with Online EM
by: Dai, Qiyuan, et al.
Published: (2025)
by: Dai, Qiyuan, et al.
Published: (2025)
Semi- and Weakly-Supervised Learning for Mammogram Mass Segmentation with Limited Annotations
by: Xiong, Xinyu, et al.
Published: (2024)
by: Xiong, Xinyu, et al.
Published: (2024)
SemiSAM+: Rethinking Semi-Supervised Medical Image Segmentation in the Era of Foundation Models
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
LaSagnA: Language-based Segmentation Assistant for Complex Queries
by: Wei, Cong, et al.
Published: (2024)
by: Wei, Cong, et al.
Published: (2024)
Rethinking Decoders for Transformer-based Semantic Segmentation: A Compression Perspective
by: Wen, Qishuai, et al.
Published: (2024)
by: Wen, Qishuai, et al.
Published: (2024)
Beyond Background Shift: Rethinking Instance Replay in Continual Semantic Segmentation
by: Yin, Hongmei, et al.
Published: (2025)
by: Yin, Hongmei, et al.
Published: (2025)
Aerial Vision-and-Language Navigation with Grid-based View Selection and Map Construction
by: Zhao, Ganlong, et al.
Published: (2025)
by: Zhao, Ganlong, et al.
Published: (2025)
Similar Items
-
Sim-DETR: Unlock DETR for Temporal Sentence Grounding
by: Tang, Jiajin, et al.
Published: (2025) -
Augmenting Moment Retrieval: Zero-Dependency Two-Stage Learning
by: Wei, Zhengxuan, et al.
Published: (2025) -
Closed-Loop Transfer for Weakly-supervised Affordance Grounding
by: Tang, Jiajin, et al.
Published: (2025) -
Part2Object: Hierarchical Unsupervised 3D Instance Segmentation
by: Shi, Cheng, et al.
Published: (2024) -
Plain-Det: A Plain Multi-Dataset Object Detector
by: Shi, Cheng, et al.
Published: (2024)