Dual Orthogonal Guidance for Robust Diffusion-based Handwritten Text Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Nikolaidou, Konstantina, Retsinas, George, Sfikas, Giorgos, Cascianelli, Silvia, Cucchiara, Rita, Liwicki, Marcus |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?
by: Pippi, Vittorio, et al.
Published: (2025)
by: Pippi, Vittorio, et al.
Published: (2025)
DiffusionPen: Towards Controlling the Style of Handwritten Text Generation
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
Rethinking HTG Evaluation: Bridging Generation and Recognition
by: Nikolaidou, Konstantina, et al.
Published: (2024)
by: Nikolaidou, Konstantina, et al.
Published: (2024)
Best Practices for a Handwritten Text Recognition System
by: Retsinas, George, et al.
Published: (2024)
by: Retsinas, George, et al.
Published: (2024)
Are Euler angles a useful rotation parameterisation for pose estimation with Normalizing Flows?
by: Sfikas, Giorgos, et al.
Published: (2025)
by: Sfikas, Giorgos, et al.
Published: (2025)
Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime
by: Wraight, Petros Georgoulas, et al.
Published: (2025)
by: Wraight, Petros Georgoulas, et al.
Published: (2025)
VATr++: Choose Your Words Wisely for Handwritten Text Generation
by: Vanherle, Bram, et al.
Published: (2024)
by: Vanherle, Bram, et al.
Published: (2024)
On the Matrix Form of the Quaternion Fourier Transform and Quaternion Convolution
by: Sfikas, Giorgos, et al.
Published: (2023)
by: Sfikas, Giorgos, et al.
Published: (2023)
Zero-Shot Synthetic-to-Real Handwritten Text Recognition via Task Analogies
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
by: Garrido-Munoz, Carlos, et al.
Published: (2026)
Merging and Splitting Diffusion Paths for Semantically Coherent Panoramas
by: Quattrini, Fabio, et al.
Published: (2024)
by: Quattrini, Fabio, et al.
Published: (2024)
Zero-Shot Styled Text Image Generation, but Make It Autoregressive
by: Pippi, Vittorio, et al.
Published: (2025)
by: Pippi, Vittorio, et al.
Published: (2025)
Alfie: Democratising RGBA Image Generation With No $$$
by: Quattrini, Fabio, et al.
Published: (2024)
by: Quattrini, Fabio, et al.
Published: (2024)
Autoregressive Styled Text Image Generation, but Make it Reliable
by: Zaccagnino, Carmine, et al.
Published: (2025)
by: Zaccagnino, Carmine, et al.
Published: (2025)
Binarizing Documents by Leveraging both Space and Frequency
by: Quattrini, Fabio, et al.
Published: (2024)
by: Quattrini, Fabio, et al.
Published: (2024)
μgat: Improving Single-Page Document Parsing by Providing Multi-Page Context
by: Quattrini, Fabio, et al.
Published: (2024)
by: Quattrini, Fabio, et al.
Published: (2024)
ToFu: Visual Tokens Reduction via Fusion for Multi-modal, Multi-patch, Multi-image Task
by: Pippi, Vittorio, et al.
Published: (2025)
by: Pippi, Vittorio, et al.
Published: (2025)
Embodied Agents for Efficient Exploration and Smart Scene Description
by: Bigazzi, Roberto, et al.
Published: (2023)
by: Bigazzi, Roberto, et al.
Published: (2023)
A Systematic Performance Analysis of Deep Perceptual Loss Networks: Breaking Transfer Learning Conventions
by: Pihlgren, Gustav Grund, et al.
Published: (2023)
by: Pihlgren, Gustav Grund, et al.
Published: (2023)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
by: Petsangourakis, Giorgos, et al.
Published: (2025)
by: Petsangourakis, Giorgos, et al.
Published: (2025)
Spot the Difference: A Novel Task for Embodied Agents in Changing Environments
by: Landi, Federico, et al.
Published: (2022)
by: Landi, Federico, et al.
Published: (2022)
Embodied Navigation at the Art Gallery
by: Bigazzi, Roberto, et al.
Published: (2022)
by: Bigazzi, Roberto, et al.
Published: (2022)
Shifting the Breaking Point of Flow Matching for Multi-Instance Editing
by: Zaccagnino, Carmine, et al.
Published: (2026)
by: Zaccagnino, Carmine, et al.
Published: (2026)
Explore and Explain: Self-supervised Navigation and Recounting
by: Bigazzi, Roberto, et al.
Published: (2020)
by: Bigazzi, Roberto, et al.
Published: (2020)
One-Shot Diffusion Mimicker for Handwritten Text Generation
by: Dai, Gang, et al.
Published: (2024)
by: Dai, Gang, et al.
Published: (2024)
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
by: Svyezhentsev, Davyd, et al.
Published: (2024)
by: Svyezhentsev, Davyd, et al.
Published: (2024)
Semi-Supervised Adaptation of Diffusion Models for Handwritten Text Generation
by: Brandenbusch, Kai
Published: (2024)
by: Brandenbusch, Kai
Published: (2024)
Designing Practical Models for Isolated Word Visual Speech Recognition
by: Panagos, Iason Ioannis, et al.
Published: (2025)
by: Panagos, Iason Ioannis, et al.
Published: (2025)
Beyond Isolated Words: Diffusion Brush for Handwritten Text-Line Generation
by: Dai, Gang, et al.
Published: (2025)
by: Dai, Gang, et al.
Published: (2025)
Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation
by: Ko, Jungmin, et al.
Published: (2026)
by: Ko, Jungmin, et al.
Published: (2026)
Lightweight Operations for Visual Speech Recognition
by: Panagos, Iason Ioannis, et al.
Published: (2025)
by: Panagos, Iason Ioannis, et al.
Published: (2025)
Training-Free Open-Vocabulary Segmentation with Offline Diffusion-Augmented Prototype Generation
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
Only-Style: Stylistic Consistency in Image Generation without Content Leakage
by: Aravanis, Tilemachos, et al.
Published: (2025)
by: Aravanis, Tilemachos, et al.
Published: (2025)
Uncovering the Handwritten Text in the Margins: End-to-end Handwritten Text Detection and Recognition
by: Cheng, Liang, et al.
Published: (2023)
by: Cheng, Liang, et al.
Published: (2023)
Mushroom Segmentation and 3D Pose Estimation from Point Clouds using Fully Convolutional Geometric Features and Implicit Pose Encoding
by: Retsinas, George, et al.
Published: (2024)
by: Retsinas, George, et al.
Published: (2024)
Category-Level 6D Object Pose Estimation in Agricultural Settings Using a Lattice-Deformation Framework and Diffusion-Augmented Synthetic Data
by: Glytsos, Marios, et al.
Published: (2025)
by: Glytsos, Marios, et al.
Published: (2025)
Attention Guidance Mechanism for Handwritten Mathematical Expression Recognition
by: Liu, Yutian, et al.
Published: (2024)
by: Liu, Yutian, et al.
Published: (2024)
Masked Generative Story Transformer with Character Guidance and Caption Augmentation
by: Papadimitriou, Christos, et al.
Published: (2024)
by: Papadimitriou, Christos, et al.
Published: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
by: Nazir, Danish, et al.
Published: (2022)
by: Nazir, Danish, et al.
Published: (2022)
Hallucination Early Detection in Diffusion Models
by: Betti, Federico, et al.
Published: (2026)
by: Betti, Federico, et al.
Published: (2026)
CER-HV: A Human-in-the-Loop Framework for Cleaning Datasets Applied to Arabic-Script HTR
by: Al-azzawi, Sana, et al.
Published: (2026)
by: Al-azzawi, Sana, et al.
Published: (2026)
Similar Items
-
Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?
by: Pippi, Vittorio, et al.
Published: (2025) -
DiffusionPen: Towards Controlling the Style of Handwritten Text Generation
by: Nikolaidou, Konstantina, et al.
Published: (2024) -
Rethinking HTG Evaluation: Bridging Generation and Recognition
by: Nikolaidou, Konstantina, et al.
Published: (2024) -
Best Practices for a Handwritten Text Recognition System
by: Retsinas, George, et al.
Published: (2024) -
Are Euler angles a useful rotation parameterisation for pose estimation with Normalizing Flows?
by: Sfikas, Giorgos, et al.
Published: (2025)