DvD: Unleashing a Generative Paradigm for Document Dewarping via Coordinates-based Diffusion Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Weiguang, Lu, Huangcheng, Ning, Maizhen, Huang, Xiaowei, Wang, Wei, Huang, Kaizhu, Wang, Qiufeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GeoSDF: Plane Geometry Diagram Synthesis via Signed Distance Field
by: Zhang, Chengrui, et al.
Published: (2025)
by: Zhang, Chengrui, et al.
Published: (2025)
BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
by: Zhao, Weiguang, et al.
Published: (2025)
by: Zhao, Weiguang, et al.
Published: (2025)
Axis-Aligned Document Dewarping
by: Wang, Chaoyun, et al.
Published: (2025)
by: Wang, Chaoyun, et al.
Published: (2025)
D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image Dewarping
by: Li, Heng, et al.
Published: (2025)
by: Li, Heng, et al.
Published: (2025)
Interpret Your Decision: Logical Reasoning Regularization for Generalization in Visual Classification
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
IDEA: Image Description Enhanced CLIP-Adapter
by: Ye, Zhipeng, et al.
Published: (2025)
by: Ye, Zhipeng, et al.
Published: (2025)
TADoc: Robust Time-Aware Document Image Dewarping
by: Zhao, Fangmin, et al.
Published: (2025)
by: Zhao, Fangmin, et al.
Published: (2025)
PO3AD: Predicting Point Offsets toward Better 3D Point Cloud Anomaly Detection
by: Ye, Jianan, et al.
Published: (2024)
by: Ye, Jianan, et al.
Published: (2024)
Towards Training-Free Open-World Classification with 3D Generative Models
by: Xia, Xinzhe, et al.
Published: (2025)
by: Xia, Xinzhe, et al.
Published: (2025)
HOMER: Homography-Based Efficient Multi-view 3D Object Removal
by: Ni, Jingcheng, et al.
Published: (2025)
by: Ni, Jingcheng, et al.
Published: (2025)
Tele-Catch: Adaptive Teleoperation for Dexterous Dynamic 3D Object Catching
by: Zhao, Weiguang, et al.
Published: (2026)
by: Zhao, Weiguang, et al.
Published: (2026)
Unraveling Batch Normalization for Realistic Test-Time Adaptation
by: Su, Zixian, et al.
Published: (2023)
by: Su, Zixian, et al.
Published: (2023)
Hyperbolic Structured Classification for Robust Single Positive Multi-label Learning
by: Lin, Yiming, et al.
Published: (2025)
by: Lin, Yiming, et al.
Published: (2025)
The Demon is in Ambiguity: Revisiting Situation Recognition with Single Positive Multi-Label Learning
by: Lin, Yiming, et al.
Published: (2025)
by: Lin, Yiming, et al.
Published: (2025)
Covariance-based Space Regularization for Few-shot Class Incremental Learning
by: Hu, Yijie, et al.
Published: (2024)
by: Hu, Yijie, et al.
Published: (2024)
3D-CDRGP: Towards Cross-Device Robotic Grasping Policy in 3D Open World
by: Zhao, Weiguang, et al.
Published: (2024)
by: Zhao, Weiguang, et al.
Published: (2024)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
by: Zhang, Tianqi, et al.
Published: (2026)
by: Zhang, Tianqi, et al.
Published: (2026)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
by: Zhou, Zihao, et al.
Published: (2024)
by: Zhou, Zihao, et al.
Published: (2024)
Towards a Universal 3D Medical Multi-modality Generalization via Learning Personalized Invariant Representation
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
A comprehensive survey of oracle character recognition: challenges, benchmarks, and beyond
by: Li, Jing, et al.
Published: (2024)
by: Li, Jing, et al.
Published: (2024)
Efficient Document Image Dewarping via Hybrid Deep Learning and Cubic Polynomial Geometry Restoration
by: Istomin, Valery, et al.
Published: (2025)
by: Istomin, Valery, et al.
Published: (2025)
Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portrait
by: Yang, Chaolong, et al.
Published: (2025)
by: Yang, Chaolong, et al.
Published: (2025)
Open-Pose 3D Zero-Shot Learning: Benchmark and Challenges
by: Zhao, Weiguang, et al.
Published: (2023)
by: Zhao, Weiguang, et al.
Published: (2023)
Pareto-Guided Optimization for Uncertainty-Aware Medical Image Segmentation
by: Zhang, Jinming, et al.
Published: (2026)
by: Zhang, Jinming, et al.
Published: (2026)
From 2D Images to 3D Model:Weakly Supervised Multi-View Face Reconstruction with Deep Fusion
by: Zhao, Weiguang, et al.
Published: (2022)
by: Zhao, Weiguang, et al.
Published: (2022)
Navigating Distribution Shifts in Medical Image Analysis: A Survey
by: Su, Zixian, et al.
Published: (2024)
by: Su, Zixian, et al.
Published: (2024)
Diff-Oracle: Deciphering Oracle Bone Scripts with Controllable Diffusion Model
by: Li, Jing, et al.
Published: (2023)
by: Li, Jing, et al.
Published: (2023)
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
by: Wen, Hao, et al.
Published: (2024)
by: Wen, Hao, et al.
Published: (2024)
Rethinking Multi-domain Generalization with A General Learning Objective
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
by: Wang, Jiangshan, et al.
Published: (2024)
by: Wang, Jiangshan, et al.
Published: (2024)
Generalized W-Net: Arbitrary-style Chinese Character Synthesization
by: Jiang, Haochuan, et al.
Published: (2024)
by: Jiang, Haochuan, et al.
Published: (2024)
Consistency Diffusion Models for Single-Image 3D Reconstruction with Priors
by: Jiang, Chenru, et al.
Published: (2025)
by: Jiang, Chenru, et al.
Published: (2025)
Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation
by: Hu, Mu, et al.
Published: (2024)
by: Hu, Mu, et al.
Published: (2024)
W-Net: One-Shot Arbitrary-Style Chinese Character Generation with Deep Neural Networks
by: Jiang, Haochuan, et al.
Published: (2024)
by: Jiang, Haochuan, et al.
Published: (2024)
Towards Cross-modal Retrieval in Chinese Cultural Heritage Documents: Dataset and Solution
by: Yuan, Junyi, et al.
Published: (2025)
by: Yuan, Junyi, et al.
Published: (2025)
Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception
by: Pang, Ziqi, et al.
Published: (2025)
by: Pang, Ziqi, et al.
Published: (2025)
Material Anything: Generating Materials for Any 3D Object via Diffusion
by: Huang, Xin, et al.
Published: (2024)
by: Huang, Xin, et al.
Published: (2024)
Revisiting Mutual Information Maximization for Generalized Category Discovery
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
A generalizable framework for low-rank tensor completion with numerical priors
by: Yuan, Shiran, et al.
Published: (2023)
by: Yuan, Shiran, et al.
Published: (2023)
FreeScale: Unleashing the Resolution of Diffusion Models via Tuning-Free Scale Fusion
by: Qiu, Haonan, et al.
Published: (2024)
by: Qiu, Haonan, et al.
Published: (2024)
Similar Items
-
GeoSDF: Plane Geometry Diagram Synthesis via Signed Distance Field
by: Zhang, Chengrui, et al.
Published: (2025) -
BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
by: Zhao, Weiguang, et al.
Published: (2025) -
Axis-Aligned Document Dewarping
by: Wang, Chaoyun, et al.
Published: (2025) -
D2Dewarp: Dual Dimensions Geometric Representation Learning Based Document Image Dewarping
by: Li, Heng, et al.
Published: (2025) -
Interpret Your Decision: Logical Reasoning Regularization for Generalization in Visual Classification
by: Tan, Zhaorui, et al.
Published: (2024)