Box2Poly: Memory-Efficient Polygon Prediction of Arbitrarily Shaped and Rotated Text
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xuyang, Wang, Dong, Schindler, Konrad, Sun, Mingwei, Wang, Yongliang, Savioli, Nicolo, Meng, Liqiu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
Lane Graph Extraction from Aerial Imagery via Lane Segmentation Refinement with Diffusion Models
von: Ruiz, Antonio, et al.
Veröffentlicht: (2024)
von: Ruiz, Antonio, et al.
Veröffentlicht: (2024)
Driving with DINO: Vision Foundation Features as a Unified Bridge for Sim-to-Real Generation in Autonomous Driving
von: Chen, Xuyang, et al.
Veröffentlicht: (2026)
von: Chen, Xuyang, et al.
Veröffentlicht: (2026)
PolyFootNet: Extracting Polygonal Building Footprints in Off-Nadir Remote Sensing Images
von: Li, Kai, et al.
Veröffentlicht: (2024)
von: Li, Kai, et al.
Veröffentlicht: (2024)
HierLoc: Hyperbolic Entity Embeddings for Hierarchical Visual Geolocation
von: Gadi, Hari Krishna, et al.
Veröffentlicht: (2026)
von: Gadi, Hari Krishna, et al.
Veröffentlicht: (2026)
LADB: Latent Aligned Diffusion Bridges for Semi-Supervised Domain Translation
von: Wang, Xuqin, et al.
Veröffentlicht: (2025)
von: Wang, Xuqin, et al.
Veröffentlicht: (2025)
Linear Gaussian Bounding Box Representation and Ring-Shaped Rotated Convolution for Oriented Object Detection
von: Zhou, Zhen, et al.
Veröffentlicht: (2023)
von: Zhou, Zhen, et al.
Veröffentlicht: (2023)
Memory Efficient Transformer Adapter for Dense Predictions
von: Zhang, Dong, et al.
Veröffentlicht: (2025)
von: Zhang, Dong, et al.
Veröffentlicht: (2025)
Consistency^2: Consistent and Fast 3D Painting with Latent Consistency Models
von: Wang, Tianfu, et al.
Veröffentlicht: (2024)
von: Wang, Tianfu, et al.
Veröffentlicht: (2024)
GeodesicNVS: Probability Density Geodesic Flow Matching for Novel View Synthesis
von: Wang, Xuqin, et al.
Veröffentlicht: (2026)
von: Wang, Xuqin, et al.
Veröffentlicht: (2026)
PolyRoof: Precision Roof Polygonization in Urban Residential Building with Graph Neural Networks
von: Amrullah, Chaikal, et al.
Veröffentlicht: (2025)
von: Amrullah, Chaikal, et al.
Veröffentlicht: (2025)
RI-Mamba: Rotation-Invariant Mamba for Robust Text-to-Shape Retrieval
von: Nguyen, Khanh, et al.
Veröffentlicht: (2026)
von: Nguyen, Khanh, et al.
Veröffentlicht: (2026)
Towards Efficient General Feature Prediction in Masked Skeleton Modeling
von: Sun, Shengkai, et al.
Veröffentlicht: (2025)
von: Sun, Shengkai, et al.
Veröffentlicht: (2025)
TetraDiffusion: Tetrahedral Diffusion Models for 3D Shape Generation
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2022)
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2022)
Progressive Evolution from Single-Point to Polygon for Scene Text
von: Deng, Linger, et al.
Veröffentlicht: (2023)
von: Deng, Linger, et al.
Veröffentlicht: (2023)
Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
von: Go, Hyojun, et al.
Veröffentlicht: (2025)
Pix2Poly: A Sequence Prediction Method for End-to-end Polygonal Building Footprint Extraction from Remote Sensing Imagery
von: Adimoolam, Yeshwanth Kumar, et al.
Veröffentlicht: (2024)
von: Adimoolam, Yeshwanth Kumar, et al.
Veröffentlicht: (2024)
Fine-tune Smarter, Not Harder: Parameter-Efficient Fine-Tuning for Geospatial Foundation Models
von: Marti-Escofet, Francesc, et al.
Veröffentlicht: (2025)
von: Marti-Escofet, Francesc, et al.
Veröffentlicht: (2025)
FPDIoU Loss: A Loss Function for Efficient Bounding Box Regression of Rotated Object Detection
von: Ma, Siliang, et al.
Veröffentlicht: (2024)
von: Ma, Siliang, et al.
Veröffentlicht: (2024)
A Probabilistic Rotation Representation for Symmetric Shapes With an Efficiently Computable Bingham Loss Function
von: Sato, Hiroya, et al.
Veröffentlicht: (2023)
von: Sato, Hiroya, et al.
Veröffentlicht: (2023)
Goldfish: Vision-Language Understanding of Arbitrarily Long Videos
von: Ataallah, Kirolos, et al.
Veröffentlicht: (2024)
von: Ataallah, Kirolos, et al.
Veröffentlicht: (2024)
Pairwise Alignment & Compatibility for Arbitrarily Irregular Image Fragments
von: Shahar, Ofir Itzhak, et al.
Veröffentlicht: (2025)
von: Shahar, Ofir Itzhak, et al.
Veröffentlicht: (2025)
VGDiffZero: Text-to-image Diffusion Models Can Be Zero-shot Visual Grounders
von: Liu, Xuyang, et al.
Veröffentlicht: (2023)
von: Liu, Xuyang, et al.
Veröffentlicht: (2023)
GALA: Geometry-Aware Local Adaptive Grids for Detailed 3D Generation
von: Yang, Dingdong, et al.
Veröffentlicht: (2024)
von: Yang, Dingdong, et al.
Veröffentlicht: (2024)
Towards Better Robustness: Pose-Free 3D Gaussian Splatting for Arbitrarily Long Videos
von: Dong, Zhen-Hui, et al.
Veröffentlicht: (2025)
von: Dong, Zhen-Hui, et al.
Veröffentlicht: (2025)
FViT: A Focal Vision Transformer with Gabor Filter
von: Shi, Yulong, et al.
Veröffentlicht: (2024)
von: Shi, Yulong, et al.
Veröffentlicht: (2024)
MambaVF: State Space Model for Efficient Video Fusion
von: Zhao, Zixiang, et al.
Veröffentlicht: (2026)
von: Zhao, Zixiang, et al.
Veröffentlicht: (2026)
Continuous Space-Time Video Super-Resolution with 3D Fourier Fields
von: Becker, Alexander, et al.
Veröffentlicht: (2025)
von: Becker, Alexander, et al.
Veröffentlicht: (2025)
Living Scenes: Multi-object Relocalization and Reconstruction in Changing 3D Environments
von: Zhu, Liyuan, et al.
Veröffentlicht: (2023)
von: Zhu, Liyuan, et al.
Veröffentlicht: (2023)
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
Parametric Point Cloud Completion for Polygonal Surface Reconstruction
von: Chen, Zhaiyu, et al.
Veröffentlicht: (2025)
von: Chen, Zhaiyu, et al.
Veröffentlicht: (2025)
Shape-IoU: More Accurate Metric considering Bounding Box Shape and Scale
von: Zhang, Hao, et al.
Veröffentlicht: (2023)
von: Zhang, Hao, et al.
Veröffentlicht: (2023)
Towards Instance Segmentation with Polygon Detection Transformers
von: Sun, Jiacheng, et al.
Veröffentlicht: (2026)
von: Sun, Jiacheng, et al.
Veröffentlicht: (2026)
Touch2Shape: Touch-Conditioned 3D Diffusion for Shape Exploration and Reconstruction
von: Wang, Yuanbo, et al.
Veröffentlicht: (2025)
von: Wang, Yuanbo, et al.
Veröffentlicht: (2025)
Rethinking Rotation-Invariant Recognition of Fine-grained Shapes from the Perspective of Contour Points
von: Xu, Yanjie, et al.
Veröffentlicht: (2025)
von: Xu, Yanjie, et al.
Veröffentlicht: (2025)
Poly Kernel Inception Network for Remote Sensing Detection
von: Cai, Xinhao, et al.
Veröffentlicht: (2024)
von: Cai, Xinhao, et al.
Veröffentlicht: (2024)
Shape-Guided Clothing Warping for Virtual Try-On
von: Han, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Han, Xiaoyu, et al.
Veröffentlicht: (2025)
TerraCodec: Compressing Optical Earth Observation Data
von: Costa-Watanabe, Julen, et al.
Veröffentlicht: (2025)
von: Costa-Watanabe, Julen, et al.
Veröffentlicht: (2025)
Point2Building: Reconstructing Buildings from Airborne LiDAR Point Clouds
von: Liu, Yujia, et al.
Veröffentlicht: (2024)
von: Liu, Yujia, et al.
Veröffentlicht: (2024)
High-resolution Population Maps Derived from Sentinel-1 and Sentinel-2
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
von: Metzger, Nando, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion
von: Chen, Xuyang, et al.
Veröffentlicht: (2025) -
Lane Graph Extraction from Aerial Imagery via Lane Segmentation Refinement with Diffusion Models
von: Ruiz, Antonio, et al.
Veröffentlicht: (2024) -
Driving with DINO: Vision Foundation Features as a Unified Bridge for Sim-to-Real Generation in Autonomous Driving
von: Chen, Xuyang, et al.
Veröffentlicht: (2026) -
PolyFootNet: Extracting Polygonal Building Footprints in Off-Nadir Remote Sensing Images
von: Li, Kai, et al.
Veröffentlicht: (2024) -
HierLoc: Hyperbolic Entity Embeddings for Hierarchical Visual Geolocation
von: Gadi, Hari Krishna, et al.
Veröffentlicht: (2026)