Inverse-like Antagonistic Scene Text Spotting via Reading-Order Estimation and Dynamic Sampling
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Shi-Xue, Yang, Chun, Zhu, Xiaobin, Zhou, Hongyang, Wang, Hongfa, Yin, Xu-Cheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Video-Language Alignment via Spatio-Temporal Graph Transformer
por: Zhang, Shi-Xue, et al.
Publicado: (2024)
por: Zhang, Shi-Xue, et al.
Publicado: (2024)
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
por: Zhang, Shi-Xue, et al.
Publicado: (2025)
por: Zhang, Shi-Xue, et al.
Publicado: (2025)
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance
por: Lyu, Jiahao, et al.
Publicado: (2024)
por: Lyu, Jiahao, et al.
Publicado: (2024)
Efficiently Leveraging Linguistic Priors for Scene Text Spotting
por: Nguyen, Nguyen, et al.
Publicado: (2024)
por: Nguyen, Nguyen, et al.
Publicado: (2024)
Unsupervised Real-World Super-Resolution via Rectified Flow Degradation Modelling
por: Zhou, Hongyang, et al.
Publicado: (2025)
por: Zhou, Hongyang, et al.
Publicado: (2025)
Hear the Scene: Audio-Enhanced Text Spotting
por: Li, Jing, et al.
Publicado: (2024)
por: Li, Jing, et al.
Publicado: (2024)
Inverse Scene Text Removal
por: Yoshimatsu, Takumi, et al.
Publicado: (2025)
por: Yoshimatsu, Takumi, et al.
Publicado: (2025)
InstructOCR: Instruction Boosting Scene Text Spotting
por: Duan, Chen, et al.
Publicado: (2024)
por: Duan, Chen, et al.
Publicado: (2024)
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
por: Das, Alloy, et al.
Publicado: (2023)
por: Das, Alloy, et al.
Publicado: (2023)
DreamScene: 3D Gaussian-based Text-to-3D Scene Generation via Formation Pattern Sampling
por: Li, Haoran, et al.
Publicado: (2024)
por: Li, Haoran, et al.
Publicado: (2024)
Similarity Matters: A Novel Depth-guided Network for Image Restoration and A New Dataset
por: He, Junyi, et al.
Publicado: (2025)
por: He, Junyi, et al.
Publicado: (2025)
SwinTextSpotter v2: Towards Better Synergy for Scene Text Spotting
por: Huang, Mingxin, et al.
Publicado: (2024)
por: Huang, Mingxin, et al.
Publicado: (2024)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
por: Das, Alloy, et al.
Publicado: (2024)
por: Das, Alloy, et al.
Publicado: (2024)
DNTextSpotter: Arbitrary-Shaped Scene Text Spotting via Improved Denoising Training
por: Xie, Yu, et al.
Publicado: (2024)
por: Xie, Yu, et al.
Publicado: (2024)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
por: Nguyen, Hieu, et al.
Publicado: (2024)
por: Nguyen, Hieu, et al.
Publicado: (2024)
When Semantics Mislead Vision: Mitigating Large Multimodal Models Hallucinations in Scene Text Spotting and Understanding
por: Shu, Yan, et al.
Publicado: (2025)
por: Shu, Yan, et al.
Publicado: (2025)
DPFlow: Adaptive Optical Flow Estimation with a Dual-Pyramid Framework
por: Morimitsu, Henrique, et al.
Publicado: (2025)
por: Morimitsu, Henrique, et al.
Publicado: (2025)
TextBlockV2: Towards Precise-Detection-Free Scene Text Spotting with Pre-trained Language Model
por: Lyu, Jiahao, et al.
Publicado: (2024)
por: Lyu, Jiahao, et al.
Publicado: (2024)
Aggregated Text Transformer for Scene Text Detection
por: Zhou, Zhao, et al.
Publicado: (2022)
por: Zhou, Zhao, et al.
Publicado: (2022)
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
por: Tao, Huaqi, et al.
Publicado: (2025)
por: Tao, Huaqi, et al.
Publicado: (2025)
Reading in the Dark: Low-light Scene Text Recognition
por: Fu, Xuanshuo, et al.
Publicado: (2026)
por: Fu, Xuanshuo, et al.
Publicado: (2026)
ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting
por: Duan, Chen, et al.
Publicado: (2024)
por: Duan, Chen, et al.
Publicado: (2024)
SceneMaker: Open-set 3D Scene Generation with Decoupled De-occlusion and Pose Estimation Model
por: Shi, Yukai, et al.
Publicado: (2025)
por: Shi, Yukai, et al.
Publicado: (2025)
Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting
por: Colombo, Antonio, et al.
Publicado: (2026)
por: Colombo, Antonio, et al.
Publicado: (2026)
DreamText: High Fidelity Scene Text Synthesis
por: Wang, Yibin, et al.
Publicado: (2024)
por: Wang, Yibin, et al.
Publicado: (2024)
Decoupled Diffusion Sparks Adaptive Scene Generation
por: Zhou, Yunsong, et al.
Publicado: (2025)
por: Zhou, Yunsong, et al.
Publicado: (2025)
Event-assisted 12-stop HDR Imaging of Dynamic Scene
por: Guo, Shi, et al.
Publicado: (2024)
por: Guo, Shi, et al.
Publicado: (2024)
4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting
por: Wu, Chun-Tin, et al.
Publicado: (2025)
por: Wu, Chun-Tin, et al.
Publicado: (2025)
Text-Pass Filter: An Efficient Scene Text Detector
por: Yang, Chuang, et al.
Publicado: (2026)
por: Yang, Chuang, et al.
Publicado: (2026)
BARD-GS: Blur-Aware Reconstruction of Dynamic Scenes via Gaussian Splatting
por: Lu, Yiren, et al.
Publicado: (2025)
por: Lu, Yiren, et al.
Publicado: (2025)
Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules
por: Bang, Junseo, et al.
Publicado: (2026)
por: Bang, Junseo, et al.
Publicado: (2026)
Bridging Synthetic and Real Worlds for Pre-training Scene Text Detectors
por: Guan, Tongkun, et al.
Publicado: (2023)
por: Guan, Tongkun, et al.
Publicado: (2023)
AdaptiveAE: An Adaptive Exposure Strategy for HDR Capturing in Dynamic Scenes
por: Xu, Tianyi, et al.
Publicado: (2025)
por: Xu, Tianyi, et al.
Publicado: (2025)
Dynamic Gaussian Scene Reconstruction from Unsynchronized Videos
por: Xu, Zhixin, et al.
Publicado: (2025)
por: Xu, Zhixin, et al.
Publicado: (2025)
Global-Local Aware Scene Text Editing
por: Yang, Fuxiang, et al.
Publicado: (2025)
por: Yang, Fuxiang, et al.
Publicado: (2025)
GIR: 3D Gaussian Inverse Rendering for Relightable Scene Factorization
por: Shi, Yahao, et al.
Publicado: (2023)
por: Shi, Yahao, et al.
Publicado: (2023)
SkyLink: Unifying Street-Satellite Geo-Localization via UAV-Mediated 3D Scene Alignment
por: Zhang, Hongyang, et al.
Publicado: (2025)
por: Zhang, Hongyang, et al.
Publicado: (2025)
Block-level Text Spotting with LLMs
por: Bannur, Ganesh, et al.
Publicado: (2024)
por: Bannur, Ganesh, et al.
Publicado: (2024)
Analyzing and Improving Fast Sampling of Text-to-Image Diffusion Models
por: Zhou, Zhenyu, et al.
Publicado: (2026)
por: Zhou, Zhenyu, et al.
Publicado: (2026)
FDSG: Forecasting Dynamic Scene Graphs
por: Yang, Yi, et al.
Publicado: (2025)
por: Yang, Yi, et al.
Publicado: (2025)
Ejemplares similares
-
Video-Language Alignment via Spatio-Temporal Graph Transformer
por: Zhang, Shi-Xue, et al.
Publicado: (2024) -
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
por: Zhang, Shi-Xue, et al.
Publicado: (2025) -
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance
por: Lyu, Jiahao, et al.
Publicado: (2024) -
Efficiently Leveraging Linguistic Priors for Scene Text Spotting
por: Nguyen, Nguyen, et al.
Publicado: (2024) -
Unsupervised Real-World Super-Resolution via Rectified Flow Degradation Modelling
por: Zhou, Hongyang, et al.
Publicado: (2025)