VectorLLM: Human-like Extraction of Structured Building Contours vis Multimodal LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Tao, Wei, Shiqing, Chen, Shihao, Yu, Wenling, Luo, Muying, Ji, Shunping |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
P2PFormer: A Primitive-to-polygon Method for Regular Building Contour Extraction from Remote Sensing Images
por: Zhang, Tao, et al.
Publicado: (2024)
por: Zhang, Tao, et al.
Publicado: (2024)
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
por: Zhou, Yikang, et al.
Publicado: (2025)
por: Zhou, Yikang, et al.
Publicado: (2025)
DeH4R: A Decoupled and Hybrid Method for Road Network Graph Extraction
por: Gong, Dengxian, et al.
Publicado: (2025)
por: Gong, Dengxian, et al.
Publicado: (2025)
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
por: Niu, Quanzhu, et al.
Publicado: (2025)
por: Niu, Quanzhu, et al.
Publicado: (2025)
A Novel Shape Guided Transformer Network for Instance Segmentation in Remote Sensing Images
por: Yu, Dawen, et al.
Publicado: (2024)
por: Yu, Dawen, et al.
Publicado: (2024)
Reconstruction of Contour Lines During the Digitization of Contour Maps to Build a Digital Elevation Model
por: Subedi, Aroj, et al.
Publicado: (2024)
por: Subedi, Aroj, et al.
Publicado: (2024)
Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
por: Zhang, Tao, et al.
Publicado: (2025)
por: Zhang, Tao, et al.
Publicado: (2025)
SaSaSaSa2VA: 2nd Place of the 5th PVUW MeViS-Text Track
por: Gong, Dengxian, et al.
Publicado: (2026)
por: Gong, Dengxian, et al.
Publicado: (2026)
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries
por: Zhou, Yikang, et al.
Publicado: (2024)
por: Zhou, Yikang, et al.
Publicado: (2024)
3D Gaussian Splatting for Large-scale Surface Reconstruction from Aerial Images
por: Wu, YuanZheng, et al.
Publicado: (2024)
por: Wu, YuanZheng, et al.
Publicado: (2024)
LLM-driven Multimodal Target Volume Contouring in Radiation Oncology
por: Oh, Yujin, et al.
Publicado: (2023)
por: Oh, Yujin, et al.
Publicado: (2023)
RoIPoly: Vectorized Building Outline Extraction Using Vertex and Logit Embeddings
por: Jiao, Weiqin, et al.
Publicado: (2024)
por: Jiao, Weiqin, et al.
Publicado: (2024)
Prompting DirectSAM for Semantic Contour Extraction in Remote Sensing Images
por: Miao, Shiyu, et al.
Publicado: (2024)
por: Miao, Shiyu, et al.
Publicado: (2024)
Contour Integration Underlies Human-Like Vision
por: Lonnqvist, Ben, et al.
Publicado: (2025)
por: Lonnqvist, Ben, et al.
Publicado: (2025)
Opt3DGS: Optimizing 3D Gaussian Splatting with Adaptive Exploration and Curvature-Aware Exploitation
por: Huang, Ziyang, et al.
Publicado: (2025)
por: Huang, Ziyang, et al.
Publicado: (2025)
Dense360: Dense Understanding from Omnidirectional Panoramas
por: Zhou, Yikang, et al.
Publicado: (2025)
por: Zhou, Yikang, et al.
Publicado: (2025)
The P$^3$ dataset: Pixels, Points and Polygons for Multimodal Building Vectorization
por: Sulzer, Raphael, et al.
Publicado: (2025)
por: Sulzer, Raphael, et al.
Publicado: (2025)
The 1st Solution for 7th LSVOS RVOS Track: SaSaSa2VA
por: Niu, Quanzhu, et al.
Publicado: (2025)
por: Niu, Quanzhu, et al.
Publicado: (2025)
Contour-Native Bridge Defect Detection and Compact Digital Archiving with Frequency-Supervised Fourier Contours
por: Liu, Jin, et al.
Publicado: (2026)
por: Liu, Jin, et al.
Publicado: (2026)
MERLIN: Building Low-SNR Robust Multimodal LLMs for Electromagnetic Signals
por: Shen, Junyu, et al.
Publicado: (2026)
por: Shen, Junyu, et al.
Publicado: (2026)
An Approach for Air Drawing Using Background Subtraction and Contour Extraction
por: Acharya, Ramkrishna
Publicado: (2025)
por: Acharya, Ramkrishna
Publicado: (2025)
ΩSFormer: Dual-Modal Ω-like Super-Resolution Transformer Network for Cross-scale and High-accuracy Terraced Field Vectorization Extraction
por: Li, Chang, et al.
Publicado: (2024)
por: Li, Chang, et al.
Publicado: (2024)
Aligning Multimodal LLM with Human Preference: A Survey
por: Yu, Tao, et al.
Publicado: (2025)
por: Yu, Tao, et al.
Publicado: (2025)
LLM-I: LLMs are Naturally Interleaved Multimodal Creators
por: Guo, Zirun, et al.
Publicado: (2025)
por: Guo, Zirun, et al.
Publicado: (2025)
Deep ContourFlow: Advancing Active Contours with Deep Learning
por: Habis, Antoine, et al.
Publicado: (2024)
por: Habis, Antoine, et al.
Publicado: (2024)
An Active Contour Model for Silhouette Vectorization using Bézier Curves
por: Alvarez, Luis, et al.
Publicado: (2025)
por: Alvarez, Luis, et al.
Publicado: (2025)
Merlin:Empowering Multimodal LLMs with Foresight Minds
por: Yu, En, et al.
Publicado: (2023)
por: Yu, En, et al.
Publicado: (2023)
Building Extraction from Remote Sensing Imagery under Hazy and Low-light Conditions: Benchmark and Baseline
por: Sang, Feifei, et al.
Publicado: (2026)
por: Sang, Feifei, et al.
Publicado: (2026)
A Hybrid Generative and Discriminative PointNet on Unordered Point Sets
por: Ye, Yang, et al.
Publicado: (2024)
por: Ye, Yang, et al.
Publicado: (2024)
Silhouette-to-Contour Registration: Aligning Intraoral Scan Models with Cephalometric Radiographs
por: Miao, Yiyi, et al.
Publicado: (2025)
por: Miao, Yiyi, et al.
Publicado: (2025)
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
por: Ding, Tianyu, et al.
Publicado: (2024)
por: Ding, Tianyu, et al.
Publicado: (2024)
GaitContour: Efficient Gait Recognition based on a Contour-Pose Representation
por: Guo, Yuxiang, et al.
Publicado: (2023)
por: Guo, Yuxiang, et al.
Publicado: (2023)
SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature Alignment
por: Xu, Qi, et al.
Publicado: (2025)
por: Xu, Qi, et al.
Publicado: (2025)
UniVector: Unified Vector Extraction via Instance-Geometry Interaction
por: Yan, Yinglong, et al.
Publicado: (2025)
por: Yan, Yinglong, et al.
Publicado: (2025)
MARL-MambaContour: Unleashing Multi-Agent Deep Reinforcement Learning for Active Contour Optimization in Medical Image Segmentation
por: Zhang, Ruicheng, et al.
Publicado: (2025)
por: Zhang, Ruicheng, et al.
Publicado: (2025)
Point Cloud Mamba: Point Cloud Learning via State Space Model
por: Zhang, Tao, et al.
Publicado: (2024)
por: Zhang, Tao, et al.
Publicado: (2024)
Path and Bone-Contour Regularized Unpaired MRI-to-CT Translation
por: Zhou, Teng, et al.
Publicado: (2025)
por: Zhou, Teng, et al.
Publicado: (2025)
Building a Mind Palace: Structuring Environment-Grounded Semantic Graphs for Effective Long Video Analysis with LLMs
por: Huang, Zeyi, et al.
Publicado: (2025)
por: Huang, Zeyi, et al.
Publicado: (2025)
Self-Supervised Dual Contouring
por: Sundararaman, Ramana, et al.
Publicado: (2024)
por: Sundararaman, Ramana, et al.
Publicado: (2024)
rFaceNet: An End-to-End Network for Enhanced Physiological Signal Extraction through Identity-Specific Facial Contours
por: Zhu, Dali, et al.
Publicado: (2024)
por: Zhu, Dali, et al.
Publicado: (2024)
Ejemplares similares
-
P2PFormer: A Primitive-to-polygon Method for Regular Building Contour Extraction from Remote Sensing Images
por: Zhang, Tao, et al.
Publicado: (2024) -
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
por: Zhou, Yikang, et al.
Publicado: (2025) -
DeH4R: A Decoupled and Hybrid Method for Road Network Graph Extraction
por: Gong, Dengxian, et al.
Publicado: (2025) -
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
por: Niu, Quanzhu, et al.
Publicado: (2025) -
A Novel Shape Guided Transformer Network for Instance Segmentation in Remote Sensing Images
por: Yu, Dawen, et al.
Publicado: (2024)