Exploring text-to-image generation for historical document image retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cote, Melissa, Albu, Alexandra Branzan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
von: McIntosh, Declan, et al.
Veröffentlicht: (2026)
von: McIntosh, Declan, et al.
Veröffentlicht: (2026)
Learning text-to-video retrieval from image captioning
von: Ventura, Lucas, et al.
Veröffentlicht: (2024)
von: Ventura, Lucas, et al.
Veröffentlicht: (2024)
Consistent text-to-image generation via scene de-contextualization
von: Tang, Song, et al.
Veröffentlicht: (2025)
von: Tang, Song, et al.
Veröffentlicht: (2025)
Exploring scalable medical image encoders beyond text supervision
von: Pérez-García, Fernando, et al.
Veröffentlicht: (2024)
von: Pérez-García, Fernando, et al.
Veröffentlicht: (2024)
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
von: Tiwari, Amritanshu, et al.
Veröffentlicht: (2025)
von: Tiwari, Amritanshu, et al.
Veröffentlicht: (2025)
Visual question answering based evaluation metrics for text-to-image generation
von: Miyamoto, Mizuki, et al.
Veröffentlicht: (2024)
von: Miyamoto, Mizuki, et al.
Veröffentlicht: (2024)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
von: Lei, Youbo, et al.
Veröffentlicht: (2023)
von: Lei, Youbo, et al.
Veröffentlicht: (2023)
PointT2I: LLM-based text-to-image generation via keypoints
von: Lee, Taekyung, et al.
Veröffentlicht: (2025)
von: Lee, Taekyung, et al.
Veröffentlicht: (2025)
Dark Miner: Defend against undesirable generation for text-to-image diffusion models
von: Meng, Zheling, et al.
Veröffentlicht: (2024)
von: Meng, Zheling, et al.
Veröffentlicht: (2024)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
von: Song, Jiawei, et al.
Veröffentlicht: (2024)
von: Song, Jiawei, et al.
Veröffentlicht: (2024)
Three-stage binarization of color document images based on discrete wavelet transform and generative adversarial networks
von: Ju, Rui-Yang, et al.
Veröffentlicht: (2022)
von: Ju, Rui-Yang, et al.
Veröffentlicht: (2022)
Transformation trees -- documentation of multimodal image registration
von: Tomaka, Agnieszka Anna, et al.
Veröffentlicht: (2025)
von: Tomaka, Agnieszka Anna, et al.
Veröffentlicht: (2025)
Hyper-parameter tuning for text guided image editing
von: Zhang, Shiwen
Veröffentlicht: (2024)
von: Zhang, Shiwen
Veröffentlicht: (2024)
LEAST: "Local" text-conditioned image style transfer
von: Singh, Silky, et al.
Veröffentlicht: (2024)
von: Singh, Silky, et al.
Veröffentlicht: (2024)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
von: Rahman, Umaima, et al.
Veröffentlicht: (2024)
Efficient scene text image super-resolution with semantic guidance
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
Ranking-aware adapter for text-driven image ordering with CLIP
von: Yu, Wei-Hsiang, et al.
Veröffentlicht: (2024)
von: Yu, Wei-Hsiang, et al.
Veröffentlicht: (2024)
IMMA: Immunizing text-to-image Models against Malicious Adaptation
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2023)
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2023)
FiCo-ITR: bridging fine-grained and coarse-grained image-text retrieval for comparative performance analysis
von: Williams-Lekuona, Mikel, et al.
Veröffentlicht: (2024)
von: Williams-Lekuona, Mikel, et al.
Veröffentlicht: (2024)
MultiWay-Adapater: Adapting large-scale multi-modal models for scalable image-text retrieval
von: Long, Zijun, et al.
Veröffentlicht: (2023)
von: Long, Zijun, et al.
Veröffentlicht: (2023)
Concept Corrector: Erase concepts on the fly for text-to-image diffusion models
von: Meng, Zheling, et al.
Veröffentlicht: (2025)
von: Meng, Zheling, et al.
Veröffentlicht: (2025)
Enhancement of text recognition for hanja handwritten documents of Ancient Korea
von: Ahna, Joonmo, et al.
Veröffentlicht: (2024)
von: Ahna, Joonmo, et al.
Veröffentlicht: (2024)
Decomposed evaluations of geographic disparities in text-to-image models
von: Sureddy, Abhishek, et al.
Veröffentlicht: (2024)
von: Sureddy, Abhishek, et al.
Veröffentlicht: (2024)
Task-driven single-image super-resolution reconstruction of document scans
von: Zyrek, Maciej, et al.
Veröffentlicht: (2024)
von: Zyrek, Maciej, et al.
Veröffentlicht: (2024)
Direct content-based retrieval from music scores images
von: Luna-Barahona, Noelia, et al.
Veröffentlicht: (2026)
von: Luna-Barahona, Noelia, et al.
Veröffentlicht: (2026)
Have we unified image generation and understanding yet? An empirical study of GPT-4o's image generation ability
von: Li, Ning, et al.
Veröffentlicht: (2025)
von: Li, Ning, et al.
Veröffentlicht: (2025)
Acquisition of interpretable domain information during brain MR image harmonization for content-based image retrieval
von: Abe, Keima, et al.
Veröffentlicht: (2025)
von: Abe, Keima, et al.
Veröffentlicht: (2025)
Dynamic watermarks in images generated by diffusion models
von: Chen, Yunzhuo, et al.
Veröffentlicht: (2025)
von: Chen, Yunzhuo, et al.
Veröffentlicht: (2025)
TurboEdit: Instant text-based image editing
von: Wu, Zongze, et al.
Veröffentlicht: (2024)
von: Wu, Zongze, et al.
Veröffentlicht: (2024)
A large-scale image-text dataset benchmark for farmland segmentation
von: Tao, Chao, et al.
Veröffentlicht: (2025)
von: Tao, Chao, et al.
Veröffentlicht: (2025)
Texture image retrieval using a classification and contourlet-based features
von: Rouhafzay, Asal, et al.
Veröffentlicht: (2024)
von: Rouhafzay, Asal, et al.
Veröffentlicht: (2024)
Are generative models fair? A study of racial bias in dermatological image generation
von: López-Pérez, Miguel, et al.
Veröffentlicht: (2025)
von: López-Pérez, Miguel, et al.
Veröffentlicht: (2025)
Consistency-diversity-realism Pareto fronts of conditional image generative models
von: Astolfi, Pietro, et al.
Veröffentlicht: (2024)
von: Astolfi, Pietro, et al.
Veröffentlicht: (2024)
Elucidating the design space of language models for image generation
von: Liu, Xuantong, et al.
Veröffentlicht: (2024)
von: Liu, Xuantong, et al.
Veröffentlicht: (2024)
Are handcrafted filters helpful for attributing AI-generated images?
von: Li, Jialiang, et al.
Veröffentlicht: (2024)
von: Li, Jialiang, et al.
Veröffentlicht: (2024)
Cardiac fat segmentation using computed tomography and an image-to-image conditional generative adversarial neural network
von: da Silva, Guilherme Santos, et al.
Veröffentlicht: (2026)
von: da Silva, Guilherme Santos, et al.
Veröffentlicht: (2026)
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval
von: Su, Hanwen, et al.
Veröffentlicht: (2024)
von: Su, Hanwen, et al.
Veröffentlicht: (2024)
Zero-shot sketch-based remote sensing image retrieval based on multi-level and attention-guided tokenization
von: Yang, Bo, et al.
Veröffentlicht: (2024)
von: Yang, Bo, et al.
Veröffentlicht: (2024)
Domain-invariant feature learning in brain MR imaging for content-based image retrieval
von: Tobari, Shuya, et al.
Veröffentlicht: (2025)
von: Tobari, Shuya, et al.
Veröffentlicht: (2025)
Unfolder: Fast localization and image rectification of a document with a crease from folding in half
von: Ershov, A. M., et al.
Veröffentlicht: (2023)
von: Ershov, A. M., et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
von: McIntosh, Declan, et al.
Veröffentlicht: (2026) -
Learning text-to-video retrieval from image captioning
von: Ventura, Lucas, et al.
Veröffentlicht: (2024) -
Consistent text-to-image generation via scene de-contextualization
von: Tang, Song, et al.
Veröffentlicht: (2025) -
Exploring scalable medical image encoders beyond text supervision
von: Pérez-García, Fernando, et al.
Veröffentlicht: (2024) -
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
von: Tiwari, Amritanshu, et al.
Veröffentlicht: (2025)