Boosting Weakly-Supervised Referring Image Segmentation via Progressive Comprehension
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Zaiquan, Liu, Yuhao, Lin, Jiaying, Hancke, Gerhard, Lau, Rynson W. H. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
di: Yang, Zaiquan, et al.
Pubblicazione: (2025)
di: Yang, Zaiquan, et al.
Pubblicazione: (2025)
Color Shift Estimation-and-Correction for Image Enhancement
di: Li, Yiyu, et al.
Pubblicazione: (2024)
di: Li, Yiyu, et al.
Pubblicazione: (2024)
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
LuSh-NeRF: Lighting up and Sharpening NeRFs for Low-light Scenes
di: Qu, Zefan, et al.
Pubblicazione: (2024)
di: Qu, Zefan, et al.
Pubblicazione: (2024)
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
di: Zhao, Youjun, et al.
Pubblicazione: (2025)
di: Zhao, Youjun, et al.
Pubblicazione: (2025)
MirrorMamba: Towards Scalable and Robust Mirror Detection in Videos
di: Song, Rui, et al.
Pubblicazione: (2025)
di: Song, Rui, et al.
Pubblicazione: (2025)
StyleSculptor: Zero-Shot Style-Controllable 3D Asset Generation with Texture-Geometry Dual Guidance
di: Qu, Zefan, et al.
Pubblicazione: (2025)
di: Qu, Zefan, et al.
Pubblicazione: (2025)
Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface Detection
di: Lin, Jiaying, et al.
Pubblicazione: (2022)
di: Lin, Jiaying, et al.
Pubblicazione: (2022)
HOComp: Interaction-Aware Human-Object Composition
di: Liang, Dong, et al.
Pubblicazione: (2025)
di: Liang, Dong, et al.
Pubblicazione: (2025)
WeakMCN: Multi-task Collaborative Network for Weakly Supervised Referring Expression Comprehension and Segmentation
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Do MLLMs Exhibit Human-like Perceptual Behaviors? HVSBench: A Benchmark for MLLM Alignment with Human Perceptual Behavior
di: Lin, Jiaying, et al.
Pubblicazione: (2024)
di: Lin, Jiaying, et al.
Pubblicazione: (2024)
OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding
di: Zhao, Youjun, et al.
Pubblicazione: (2024)
di: Zhao, Youjun, et al.
Pubblicazione: (2024)
Curriculum Point Prompting for Weakly-Supervised Referring Image Segmentation
di: Dai, Qiyuan, et al.
Pubblicazione: (2024)
di: Dai, Qiyuan, et al.
Pubblicazione: (2024)
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
di: Liu, Yuhao, et al.
Pubblicazione: (2025)
di: Liu, Yuhao, et al.
Pubblicazione: (2025)
Diff-Plugin: Revitalizing Details for Diffusion-based Low-level Tasks
di: Liu, Yuhao, et al.
Pubblicazione: (2024)
di: Liu, Yuhao, et al.
Pubblicazione: (2024)
World-Shaper: A Unified Framework for 360° Panoramic Editing
di: Liang, Dong, et al.
Pubblicazione: (2026)
di: Liang, Dong, et al.
Pubblicazione: (2026)
Recasting Regional Lighting for Shadow Removal
di: Liu, Yuhao, et al.
Pubblicazione: (2024)
di: Liu, Yuhao, et al.
Pubblicazione: (2024)
Delving into Dark Regions for Robust Shadow Detection
di: Guan, Huankang, et al.
Pubblicazione: (2024)
di: Guan, Huankang, et al.
Pubblicazione: (2024)
Inverse Rendering of Glossy Objects via the Neural Plenoptic Function and Radiance Fields
di: Wang, Haoyuan, et al.
Pubblicazione: (2024)
di: Wang, Haoyuan, et al.
Pubblicazione: (2024)
RefSTAR: Blind Facial Image Restoration with Reference Selection, Transfer, and Reconstruction
di: Yin, Zhicun, et al.
Pubblicazione: (2025)
di: Yin, Zhicun, et al.
Pubblicazione: (2025)
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
di: Shi, Miaojing, et al.
Pubblicazione: (2026)
di: Shi, Miaojing, et al.
Pubblicazione: (2026)
ProCNS: Progressive Prototype Calibration and Noise Suppression for Weakly-Supervised Medical Image Segmentation
di: Liu, Y., et al.
Pubblicazione: (2024)
di: Liu, Y., et al.
Pubblicazione: (2024)
Segment, Select, Correct: A Framework for Weakly-Supervised Referring Segmentation
di: Eiras, Francisco, et al.
Pubblicazione: (2023)
di: Eiras, Francisco, et al.
Pubblicazione: (2023)
Prototype-Based Image Prompting for Weakly Supervised Histopathological Image Segmentation
di: Tang, Qingchen, et al.
Pubblicazione: (2025)
di: Tang, Qingchen, et al.
Pubblicazione: (2025)
LIHE: Linguistic Instance-Split Hyperbolic-Euclidean Framework for Generalized Weakly-Supervised Referring Expression Comprehension
di: Shi, Xianglong, et al.
Pubblicazione: (2025)
di: Shi, Xianglong, et al.
Pubblicazione: (2025)
Efficient Universal Models for Medical Image Segmentation via Weakly Supervised In-Context Learning
di: Hu, Jiesi, et al.
Pubblicazione: (2025)
di: Hu, Jiesi, et al.
Pubblicazione: (2025)
Integrating SAM Supervision for 3D Weakly Supervised Point Cloud Segmentation
di: You, Lechun, et al.
Pubblicazione: (2025)
di: You, Lechun, et al.
Pubblicazione: (2025)
Structure-Informed Shadow Removal Networks
di: Liu, Yuhao, et al.
Pubblicazione: (2023)
di: Liu, Yuhao, et al.
Pubblicazione: (2023)
State and Scene Enhanced Prototypes for Weakly Supervised Open-Vocabulary Object Detection
di: Zhou, Jiaying, et al.
Pubblicazione: (2025)
di: Zhou, Jiaying, et al.
Pubblicazione: (2025)
S$^{5}$Mars: Semi-Supervised Learning for Mars Semantic Segmentation
di: Zhang, Jiahang, et al.
Pubblicazione: (2022)
di: Zhang, Jiahang, et al.
Pubblicazione: (2022)
Image Augmentation Agent for Weakly Supervised Semantic Segmentation
di: Wu, Wangyu, et al.
Pubblicazione: (2024)
di: Wu, Wangyu, et al.
Pubblicazione: (2024)
DuPL: Dual Student with Trustworthy Progressive Learning for Robust Weakly Supervised Semantic Segmentation
di: Wu, Yuanchen, et al.
Pubblicazione: (2024)
di: Wu, Yuanchen, et al.
Pubblicazione: (2024)
Weakly Supervised LiDAR Semantic Segmentation via Scatter Image Annotation
di: Chen, Yilong, et al.
Pubblicazione: (2024)
di: Chen, Yilong, et al.
Pubblicazione: (2024)
MDeRainNet: An Efficient Macro-pixel Image Rain Removal Network
di: Yan, Tao, et al.
Pubblicazione: (2024)
di: Yan, Tao, et al.
Pubblicazione: (2024)
PPBoost: Progressive Prompt Boosting for Text-Driven Medical Image Segmentation
di: Li, Xuchen, et al.
Pubblicazione: (2025)
di: Li, Xuchen, et al.
Pubblicazione: (2025)
Image Augmentation with Controlled Diffusion for Weakly-Supervised Semantic Segmentation
di: Wu, Wangyu, et al.
Pubblicazione: (2023)
di: Wu, Wangyu, et al.
Pubblicazione: (2023)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
di: Li, Jiachen, et al.
Pubblicazione: (2026)
di: Li, Jiachen, et al.
Pubblicazione: (2026)
Dense Supervision Propagation for Weakly Supervised Semantic Segmentation on 3D Point Clouds
di: Wei, Jiacheng, et al.
Pubblicazione: (2021)
di: Wei, Jiacheng, et al.
Pubblicazione: (2021)
Masked Image Modeling Boosting Semi-Supervised Semantic Segmentation
di: Li, Yangyang, et al.
Pubblicazione: (2024)
di: Li, Yangyang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
di: Yang, Zaiquan, et al.
Pubblicazione: (2025) -
Color Shift Estimation-and-Correction for Image Enhancement
di: Li, Yiyu, et al.
Pubblicazione: (2024) -
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
di: Wang, Zhenwei, et al.
Pubblicazione: (2024) -
LuSh-NeRF: Lighting up and Sharpening NeRFs for Low-light Scenes
di: Qu, Zefan, et al.
Pubblicazione: (2024) -
ThemeStation: Generating Theme-Aware 3D Assets from Few Exemplars
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)