TADoc: Robust Time-Aware Document Image Dewarping
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Fangmin, Zeng, Weichao, Li, Zhenhang, Yang, Dongbao, Zhou, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
by: Zhao, Fangmin, et al.
Published: (2025)
by: Zhao, Fangmin, et al.
Published: (2025)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
by: Zeng, Weichao, et al.
Published: (2024)
by: Zeng, Weichao, et al.
Published: (2024)
First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending
by: Li, Zhenhang, et al.
Published: (2024)
by: Li, Zhenhang, et al.
Published: (2024)
Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing
by: Shu, Yan, et al.
Published: (2024)
by: Shu, Yan, et al.
Published: (2024)
D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image Dewarping
by: Li, Heng, et al.
Published: (2025)
by: Li, Heng, et al.
Published: (2025)
Axis-Aligned Document Dewarping
by: Wang, Chaoyun, et al.
Published: (2025)
by: Wang, Chaoyun, et al.
Published: (2025)
The Role of Video Generation in Enhancing Data-Limited Action Understanding
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Visual Text Processing: A Comprehensive Review and Unified Evaluation
by: Shu, Yan, et al.
Published: (2025)
by: Shu, Yan, et al.
Published: (2025)
IMTBench: A Multi-Scenario Cross-Modal Collaborative Evaluation Benchmark for In-Image Machine Translation
by: Lyu, Jiahao, et al.
Published: (2026)
by: Lyu, Jiahao, et al.
Published: (2026)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
DvD: Unleashing a Generative Paradigm for Document Dewarping via Coordinates-based Diffusion Model
by: Zhang, Weiguang, et al.
Published: (2025)
by: Zhang, Weiguang, et al.
Published: (2025)
Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation
by: Luo, Minxing, et al.
Published: (2025)
by: Luo, Minxing, et al.
Published: (2025)
Efficient Document Image Dewarping via Hybrid Deep Learning and Cubic Polynomial Geometry Restoration
by: Istomin, Valery, et al.
Published: (2025)
by: Istomin, Valery, et al.
Published: (2025)
Customizing Visual Emotion Evaluation for MLLMs: An Open-vocabulary, Multifaceted, and Scalable Approach
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
EmoCaliber: Advancing Reliable Visual Emotion Comprehension via Confidence Verbalization and Calibration
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
An Empirical Study on Configuring In-Context Learning Demonstrations for Unleashing MLLMs' Sentimental Perception Capability
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
Masked and Permuted Implicit Context Learning for Scene Text Recognition
by: Yang, Xiaomeng, et al.
Published: (2023)
by: Yang, Xiaomeng, et al.
Published: (2023)
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance
by: Lyu, Jiahao, et al.
Published: (2024)
by: Lyu, Jiahao, et al.
Published: (2024)
Real-Time Dynamic Scale-Aware Fusion Detection Network: Take Road Damage Detection as an example
by: Pan, Weichao, et al.
Published: (2024)
by: Pan, Weichao, et al.
Published: (2024)
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
by: Chen, Zeyu, et al.
Published: (2026)
by: Chen, Zeyu, et al.
Published: (2026)
DCA: Dividing and Conquering Amnesia in Incremental Object Detection
by: Zhang, Aoting, et al.
Published: (2025)
by: Zhang, Aoting, et al.
Published: (2025)
The Devil is in Fine-tuning and Long-tailed Problems:A New Benchmark for Scene Text Detection
by: Cao, Tianjiao, et al.
Published: (2025)
by: Cao, Tianjiao, et al.
Published: (2025)
DocR1: Evidence Page-Guided GRPO for Multi-Page Document Understanding
by: Xiong, Junyu, et al.
Published: (2025)
by: Xiong, Junyu, et al.
Published: (2025)
Specifying What You Know or Not for Multi-Label Class-Incremental Learning
by: Zhang, Aoting, et al.
Published: (2025)
by: Zhang, Aoting, et al.
Published: (2025)
YOLO-ROC: A High-Precision and Ultra-Lightweight Model for Real-Time Road Damage Detection
by: Lin, Zicheng, et al.
Published: (2025)
by: Lin, Zicheng, et al.
Published: (2025)
Collapse-Aware Triplet Decoupling for Adversarially Robust Image Retrieval
by: Tian, Qiwei, et al.
Published: (2023)
by: Tian, Qiwei, et al.
Published: (2023)
Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training
by: Li, Gengluo, et al.
Published: (2026)
by: Li, Gengluo, et al.
Published: (2026)
Exploiting Spatial-Temporal Context for Interacting Hand Reconstruction on Monocular RGB Video
by: Zhao, Weichao, et al.
Published: (2023)
by: Zhao, Weichao, et al.
Published: (2023)
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval
by: Zeng, Gangyan, et al.
Published: (2024)
by: Zeng, Gangyan, et al.
Published: (2024)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
TRACE: Structure-Aware Character Encoding for Robust and Generalizable Document Watermarking
by: Meng, Jiale, et al.
Published: (2026)
by: Meng, Jiale, et al.
Published: (2026)
DAGNet: A Dual-View Attention-Guided Network for Efficient X-ray Security Inspection
by: Hong, Shilong, et al.
Published: (2025)
by: Hong, Shilong, et al.
Published: (2025)
Uni-Sign: Toward Unified Sign Language Understanding at Scale
by: Li, Zecheng, et al.
Published: (2025)
by: Li, Zecheng, et al.
Published: (2025)
Scaling up Multimodal Pre-training for Sign Language Understanding
by: Zhou, Wengang, et al.
Published: (2024)
by: Zhou, Wengang, et al.
Published: (2024)
Resolving Sentiment Discrepancy for Multimodal Sentiment Detection via Semantics Completion and Decomposition
by: Wu, Daiqing, et al.
Published: (2024)
by: Wu, Daiqing, et al.
Published: (2024)
IAIFNet: An Illumination-Aware Infrared and Visible Image Fusion Network
by: Yang, Qiao, et al.
Published: (2023)
by: Yang, Qiao, et al.
Published: (2023)
Cascaded Robust Rectification for Arbitrary Document Images
by: Wang, Chaoyun, et al.
Published: (2025)
by: Wang, Chaoyun, et al.
Published: (2025)
Degradation-Robust Fusion: An Efficient Degradation-Aware Diffusion Framework for Multimodal Image Fusion in Arbitrary Degradation Scenarios
by: Shi, Yu, et al.
Published: (2026)
by: Shi, Yu, et al.
Published: (2026)
MASA: Motion-aware Masked Autoencoder with Semantic Alignment for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
RoBiS: Robust Binary Segmentation for High-Resolution Industrial Images
by: Li, Xurui, et al.
Published: (2025)
by: Li, Xurui, et al.
Published: (2025)
Similar Items
-
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
by: Zhao, Fangmin, et al.
Published: (2025) -
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
by: Zeng, Weichao, et al.
Published: (2024) -
First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending
by: Li, Zhenhang, et al.
Published: (2024) -
Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing
by: Shu, Yan, et al.
Published: (2024) -
D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image Dewarping
by: Li, Heng, et al.
Published: (2025)