LLM-Seg: Bridging Image Segmentation and Large Language Model Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Junchi, Ke, Lei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TipSegNet: Fingertip Segmentation in Contactless Fingerprint Imaging
by: Ruzicka, Laurenz, et al.
Published: (2025)
by: Ruzicka, Laurenz, et al.
Published: (2025)
Seg-LSTM: Performance of xLSTM for Semantic Segmentation of Remotely Sensed Images
by: Zhu, Qinfeng, et al.
Published: (2024)
by: Zhu, Qinfeng, et al.
Published: (2024)
UnSegGNet: Unsupervised Image Segmentation using Graph Neural Networks
by: Reddy, Kovvuri Sai Gopal, et al.
Published: (2024)
by: Reddy, Kovvuri Sai Gopal, et al.
Published: (2024)
SegResMamba: An Efficient Architecture for 3D Medical Image Segmentation
by: Das, Badhan Kumar, et al.
Published: (2025)
by: Das, Badhan Kumar, et al.
Published: (2025)
SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation
by: Lu, Zhenyu, et al.
Published: (2026)
by: Lu, Zhenyu, et al.
Published: (2026)
LightMedSeg: Lightweight 3D Medical Image Segmentation with Learned Spatial Anchors
by: Tyagi, Kavyansh, et al.
Published: (2026)
by: Tyagi, Kavyansh, et al.
Published: (2026)
ConvTransSeg: A Multi-resolution Convolution-Transformer Network for Medical Image Segmentation
by: Gong, Zhendi, et al.
Published: (2022)
by: Gong, Zhendi, et al.
Published: (2022)
Bridging Compressed Image Latents and Multimodal Large Language Models
by: Kao, Chia-Hao, et al.
Published: (2024)
by: Kao, Chia-Hao, et al.
Published: (2024)
UnSegMedGAT: Unsupervised Medical Image Segmentation using Graph Attention Networks Clustering
by: Adityaja, A. Mudit, et al.
Published: (2024)
by: Adityaja, A. Mudit, et al.
Published: (2024)
MedSeg-R: Reasoning Segmentation in Medical Images with Multimodal Large Language Models
by: Huang, Yu, et al.
Published: (2025)
by: Huang, Yu, et al.
Published: (2025)
AutoMiSeg: Automatic Medical Image Segmentation via Test-Time Adaptation of Foundation Models
by: Li, Xingjian, et al.
Published: (2025)
by: Li, Xingjian, et al.
Published: (2025)
PRIMA: Multi-Image Vision-Language Models for Reasoning Segmentation
by: Wahed, Muntasir, et al.
Published: (2024)
by: Wahed, Muntasir, et al.
Published: (2024)
AdaSemSeg: An Adaptive Few-shot Semantic Segmentation of Seismic Facies
by: Saha, Surojit, et al.
Published: (2025)
by: Saha, Surojit, et al.
Published: (2025)
DA-SegFormer: Damage-Aware Semantic Segmentation for Fine-Grained Disaster Assessment
by: Zhu, Kevin, et al.
Published: (2026)
by: Zhu, Kevin, et al.
Published: (2026)
SegMate: Asymmetric Attention-Based Lightweight Architecture for Efficient Multi-Organ Segmentation
by: Bunea, Andrei-Alexandru, et al.
Published: (2026)
by: Bunea, Andrei-Alexandru, et al.
Published: (2026)
ImageChain: Advancing Sequential Image-to-Text Reasoning in Multimodal Large Language Models
by: Villegas, Danae Sánchez, et al.
Published: (2025)
by: Villegas, Danae Sánchez, et al.
Published: (2025)
MedCLIP-SAM: Bridging Text and Image Towards Universal Medical Image Segmentation
by: Koleilat, Taha, et al.
Published: (2024)
by: Koleilat, Taha, et al.
Published: (2024)
SegWithU: Uncertainty as Perturbation Energy for Single-Forward-Pass Risk-Aware Medical Image Segmentation
by: Fu, Tianhao, et al.
Published: (2026)
by: Fu, Tianhao, et al.
Published: (2026)
Bayesian SegNet for Semantic Segmentation with Improved Interpretation of Microstructural Evolution During Irradiation of Materials
by: Oostrom, Marjolein, et al.
Published: (2025)
by: Oostrom, Marjolein, et al.
Published: (2025)
UnSeg: One Universal Unlearnable Example Generator is Enough against All Image Segmentation
by: Sun, Ye, et al.
Published: (2024)
by: Sun, Ye, et al.
Published: (2024)
SemSegDepth: A Combined Model for Semantic Segmentation and Depth Completion
by: Lagos, Juan Pablo, et al.
Published: (2022)
by: Lagos, Juan Pablo, et al.
Published: (2022)
MultiverSeg: Scalable Interactive Segmentation of Biomedical Imaging Datasets with In-Context Guidance
by: Wong, Hallee E., et al.
Published: (2024)
by: Wong, Hallee E., et al.
Published: (2024)
BiSeg-SAM: Weakly-Supervised Post-Processing Framework for Boosting Binary Segmentation in Segment Anything Models
by: Su, Encheng, et al.
Published: (2025)
by: Su, Encheng, et al.
Published: (2025)
ImageNet-Think-250K: A Large-Scale Synthetic Dataset for Multimodal Reasoning for Vision Language Models
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
by: Chitty-Venkata, Krishna Teja, et al.
Published: (2025)
SegHeD+: Segmentation of Heterogeneous Data for Multiple Sclerosis Lesions with Anatomical Constraints and Lesion-aware Augmentation
by: Basaran, Berke Doga, et al.
Published: (2024)
by: Basaran, Berke Doga, et al.
Published: (2024)
ε-Seg: Sparsely Supervised Semantic Segmentation of Microscopy Data
by: Kordasiabi, Sheida Rahnamai, et al.
Published: (2025)
by: Kordasiabi, Sheida Rahnamai, et al.
Published: (2025)
GeoReasoner: Geo-localization with Reasoning in Street Views using a Large Vision-Language Model
by: Li, Ling, et al.
Published: (2024)
by: Li, Ling, et al.
Published: (2024)
SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis
by: Ye, Hanrong, et al.
Published: (2023)
by: Ye, Hanrong, et al.
Published: (2023)
Predictive Regularization Against Visual Representation Degradation in Multimodal Large Language Models
by: Wang, Enguang, et al.
Published: (2026)
by: Wang, Enguang, et al.
Published: (2026)
ReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
by: Tian, Jiaxu, et al.
Published: (2025)
by: Tian, Jiaxu, et al.
Published: (2025)
Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models
by: You, Weiqiu, et al.
Published: (2026)
by: You, Weiqiu, et al.
Published: (2026)
Text4Seg++: Advancing Image Segmentation via Generative Language Modeling
by: Lan, Mengcheng, et al.
Published: (2025)
by: Lan, Mengcheng, et al.
Published: (2025)
Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
by: Das, Aryan, et al.
Published: (2026)
by: Das, Aryan, et al.
Published: (2026)
Fully Kolmogorov-Arnold Deep Model in Medical Image Segmentation
by: Qiu, Xingyu, et al.
Published: (2026)
by: Qiu, Xingyu, et al.
Published: (2026)
SegLLM: Multi-round Reasoning Segmentation
by: Wang, XuDong, et al.
Published: (2024)
by: Wang, XuDong, et al.
Published: (2024)
vFusedSeg3D: 3rd Place Solution for 2024 Waymo Open Dataset Challenge in Semantic Segmentation
by: Amjad, Osama, et al.
Published: (2024)
by: Amjad, Osama, et al.
Published: (2024)
DualSwinFusionSeg: Multimodal Martian Landslide Segmentation via Dual Swin Transformer with Multi-Scale Fusion and UNet++
by: Kabir, Shahriar, et al.
Published: (2026)
by: Kabir, Shahriar, et al.
Published: (2026)
Language Integration in Fine-Tuning Multimodal Large Language Models for Image-Based Regression
by: Jennings, Roy H., et al.
Published: (2025)
by: Jennings, Roy H., et al.
Published: (2025)
Quickly Tuning Foundation Models for Image Segmentation
by: Das, Breenda, et al.
Published: (2025)
by: Das, Breenda, et al.
Published: (2025)
Incentivizing Reasoning for Advanced Instruction-Following of Large Language Models
by: Qin, Yulei, et al.
Published: (2025)
by: Qin, Yulei, et al.
Published: (2025)
Similar Items
-
TipSegNet: Fingertip Segmentation in Contactless Fingerprint Imaging
by: Ruzicka, Laurenz, et al.
Published: (2025) -
Seg-LSTM: Performance of xLSTM for Semantic Segmentation of Remotely Sensed Images
by: Zhu, Qinfeng, et al.
Published: (2024) -
UnSegGNet: Unsupervised Image Segmentation using Graph Neural Networks
by: Reddy, Kovvuri Sai Gopal, et al.
Published: (2024) -
SegResMamba: An Efficient Architecture for 3D Medical Image Segmentation
by: Das, Badhan Kumar, et al.
Published: (2025) -
SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation
by: Lu, Zhenyu, et al.
Published: (2026)