Saved in:
| Main Authors: | Wang, Chao, Li, Xuanying, Dai, Cheng, Feng, Jinglei, Luo, Yuxiang, Ouyang, Yuqi, Qin, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.18252 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Co-Seg++: Mutual Prompt-Guided Collaborative Learning for Versatile Medical Segmentation
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
Adaptive Dual-Constrained Line Aggregation for Robust Generic and Wireframe Line Segment Detection
by: Liu, Chenguang, et al.
Published: (2025)
by: Liu, Chenguang, et al.
Published: (2025)
MSM-Seg: A Modality-and-Slice Memory Framework with Category-Agnostic Prompting for Multi-Modal Brain Tumor Segmentation
by: Luo, Yuxiang, et al.
Published: (2025)
by: Luo, Yuxiang, et al.
Published: (2025)
Edge Prediction for Roof Wireframe Reconstruction with Transformers
by: Hanning, Gustav, et al.
Published: (2026)
by: Hanning, Gustav, et al.
Published: (2026)
RxnCaption: Reformulating Reaction Diagram Parsing as Visual Prompt Guided Captioning
by: Song, Jiahe, et al.
Published: (2025)
by: Song, Jiahe, et al.
Published: (2025)
Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting
by: Feng, Hao, et al.
Published: (2025)
by: Feng, Hao, et al.
Published: (2025)
Generating 3D House Wireframes with Semantics
by: Ma, Xueqi, et al.
Published: (2024)
by: Ma, Xueqi, et al.
Published: (2024)
Seeing the Unseen: A Frequency Prompt Guided Transformer for Image Restoration
by: Zhou, Shihao, et al.
Published: (2024)
by: Zhou, Shihao, et al.
Published: (2024)
Delaunay Canopy: Building Wireframe Reconstruction from Airborne LiDAR Point Clouds via Delaunay Graph
by: Kim, Donghyun, et al.
Published: (2026)
by: Kim, Donghyun, et al.
Published: (2026)
Co-Seg: Mutual Prompt-Guided Collaborative Learning for Tissue and Nuclei Segmentation
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
Traffic Scene Parsing through the TSP6K Dataset
by: Jiang, Peng-Tao, et al.
Published: (2023)
by: Jiang, Peng-Tao, et al.
Published: (2023)
Spot the Error: Non-autoregressive Graphic Layout Generation with Wireframe Locator
by: Lin, Jieru, et al.
Published: (2024)
by: Lin, Jieru, et al.
Published: (2024)
Dolphin-v2: Universal Document Parsing via Scalable Anchor Prompting
by: Feng, Hao, et al.
Published: (2026)
by: Feng, Hao, et al.
Published: (2026)
NEAT: Distilling 3D Wireframes from Neural Attraction Fields
by: Xue, Nan, et al.
Published: (2023)
by: Xue, Nan, et al.
Published: (2023)
SAVMap: Structure-Aided Visual Mapping of Large-Scale 2.5D Manhattan Wireframes from Panoramic Video
by: Huang, Howard, et al.
Published: (2026)
by: Huang, Howard, et al.
Published: (2026)
VTimeCoT: Thinking by Drawing for Video Temporal Grounding and Reasoning
by: Zhang, Jinglei, et al.
Published: (2025)
by: Zhang, Jinglei, et al.
Published: (2025)
Point or Line? Using Line-based Representation for Panoptic Symbol Spotting in CAD Drawings
by: Wei, Xingguang, et al.
Published: (2025)
by: Wei, Xingguang, et al.
Published: (2025)
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
by: Fan, Jiahe, et al.
Published: (2026)
by: Fan, Jiahe, et al.
Published: (2026)
Knowledge-Guided Prompt Learning for Deepfake Facial Image Detection
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
CoLeaF: A Contrastive-Collaborative Learning Framework for Weakly Supervised Audio-Visual Video Parsing
by: Sardari, Faegheh, et al.
Published: (2024)
by: Sardari, Faegheh, et al.
Published: (2024)
ParseCaps: An Interpretable Parsing Capsule Network for Medical Image Diagnosis
by: Geng, Xinyu, et al.
Published: (2024)
by: Geng, Xinyu, et al.
Published: (2024)
Collaborative Position Reasoning Network for Referring Image Segmentation
by: Cao, Jianjian, et al.
Published: (2024)
by: Cao, Jianjian, et al.
Published: (2024)
CLR-Wire: Towards Continuous Latent Representations for 3D Curve Wireframe Generation
by: Ma, Xueqi, et al.
Published: (2025)
by: Ma, Xueqi, et al.
Published: (2025)
PIG: Prompt Images Guidance for Night-Time Scene Parsing
by: Xie, Zhifeng, et al.
Published: (2024)
by: Xie, Zhifeng, et al.
Published: (2024)
ICFRNet: Image Complexity Prior Guided Feature Refinement for Real-time Semantic Segmentation
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
Molecular Identifier Visual Prompt and Verifiable Reinforcement Learning for Chemical Reaction Diagram Parsing
by: Song, Jiahe, et al.
Published: (2026)
by: Song, Jiahe, et al.
Published: (2026)
Curriculum Point Prompting for Weakly-Supervised Referring Image Segmentation
by: Dai, Qiyuan, et al.
Published: (2024)
by: Dai, Qiyuan, et al.
Published: (2024)
MangaNinja: Line Art Colorization with Precise Reference Following
by: Liu, Zhiheng, et al.
Published: (2025)
by: Liu, Zhiheng, et al.
Published: (2025)
Efficient Multi-Instance Generation with Janus-Pro-Dirven Prompt Parsing
by: Qi, Fan, et al.
Published: (2025)
by: Qi, Fan, et al.
Published: (2025)
Prompt-Guided Adaptive Model Transformation for Whole Slide Image Classification
by: Lin, Yi, et al.
Published: (2024)
by: Lin, Yi, et al.
Published: (2024)
Embracing Collaboration Over Competition: Condensing Multiple Prompts for Visual In-Context Learning
by: Wang, Jinpeng, et al.
Published: (2025)
by: Wang, Jinpeng, et al.
Published: (2025)
Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection
by: Liu, Xinyuan, et al.
Published: (2025)
by: Liu, Xinyuan, et al.
Published: (2025)
HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos
by: Zhang, Jinglei, et al.
Published: (2025)
by: Zhang, Jinglei, et al.
Published: (2025)
Beyond Full Labels: Energy-Double-Guided Single-Point Prompt for Infrared Small Target Label Generation
by: Yuan, Shuai, et al.
Published: (2024)
by: Yuan, Shuai, et al.
Published: (2024)
Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing
by: Chen, Yaru, et al.
Published: (2025)
by: Chen, Yaru, et al.
Published: (2025)
Straightforward Layer-wise Pruning for More Efficient Visual Adaptation
by: Han, Ruizi, et al.
Published: (2024)
by: Han, Ruizi, et al.
Published: (2024)
LoD-Loc: Aerial Visual Localization using LoD 3D Map with Neural Wireframe Alignment
by: Zhu, Juelin, et al.
Published: (2024)
by: Zhu, Juelin, et al.
Published: (2024)
Explore Human Parsing Modality for Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
CasPoinTr: Point Cloud Completion with Cascaded Networks and Knowledge Distillation
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Self-Paced Collaborative and Adversarial Network for Unsupervised Domain Adaptation
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
Similar Items
-
Co-Seg++: Mutual Prompt-Guided Collaborative Learning for Versatile Medical Segmentation
by: Xu, Qing, et al.
Published: (2025) -
Adaptive Dual-Constrained Line Aggregation for Robust Generic and Wireframe Line Segment Detection
by: Liu, Chenguang, et al.
Published: (2025) -
MSM-Seg: A Modality-and-Slice Memory Framework with Category-Agnostic Prompting for Multi-Modal Brain Tumor Segmentation
by: Luo, Yuxiang, et al.
Published: (2025) -
Edge Prediction for Roof Wireframe Reconstruction with Transformers
by: Hanning, Gustav, et al.
Published: (2026) -
RxnCaption: Reformulating Reaction Diagram Parsing as Visual Prompt Guided Captioning
by: Song, Jiahe, et al.
Published: (2025)