Exploring text-to-image generation for historical document image retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Cote, Melissa, Albu, Alexandra Branzan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
di: McIntosh, Declan, et al.
Pubblicazione: (2026)
di: McIntosh, Declan, et al.
Pubblicazione: (2026)
Learning text-to-video retrieval from image captioning
di: Ventura, Lucas, et al.
Pubblicazione: (2024)
di: Ventura, Lucas, et al.
Pubblicazione: (2024)
Consistent text-to-image generation via scene de-contextualization
di: Tang, Song, et al.
Pubblicazione: (2025)
di: Tang, Song, et al.
Pubblicazione: (2025)
Exploring scalable medical image encoders beyond text supervision
di: Pérez-García, Fernando, et al.
Pubblicazione: (2024)
di: Pérez-García, Fernando, et al.
Pubblicazione: (2024)
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
di: Tiwari, Amritanshu, et al.
Pubblicazione: (2025)
di: Tiwari, Amritanshu, et al.
Pubblicazione: (2025)
Visual question answering based evaluation metrics for text-to-image generation
di: Miyamoto, Mizuki, et al.
Pubblicazione: (2024)
di: Miyamoto, Mizuki, et al.
Pubblicazione: (2024)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
di: Lei, Youbo, et al.
Pubblicazione: (2023)
di: Lei, Youbo, et al.
Pubblicazione: (2023)
PointT2I: LLM-based text-to-image generation via keypoints
di: Lee, Taekyung, et al.
Pubblicazione: (2025)
di: Lee, Taekyung, et al.
Pubblicazione: (2025)
Dark Miner: Defend against undesirable generation for text-to-image diffusion models
di: Meng, Zheling, et al.
Pubblicazione: (2024)
di: Meng, Zheling, et al.
Pubblicazione: (2024)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
di: Song, Jiawei, et al.
Pubblicazione: (2024)
di: Song, Jiawei, et al.
Pubblicazione: (2024)
Three-stage binarization of color document images based on discrete wavelet transform and generative adversarial networks
di: Ju, Rui-Yang, et al.
Pubblicazione: (2022)
di: Ju, Rui-Yang, et al.
Pubblicazione: (2022)
Transformation trees -- documentation of multimodal image registration
di: Tomaka, Agnieszka Anna, et al.
Pubblicazione: (2025)
di: Tomaka, Agnieszka Anna, et al.
Pubblicazione: (2025)
Hyper-parameter tuning for text guided image editing
di: Zhang, Shiwen
Pubblicazione: (2024)
di: Zhang, Shiwen
Pubblicazione: (2024)
LEAST: "Local" text-conditioned image style transfer
di: Singh, Silky, et al.
Pubblicazione: (2024)
di: Singh, Silky, et al.
Pubblicazione: (2024)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
di: Rahman, Umaima, et al.
Pubblicazione: (2024)
di: Rahman, Umaima, et al.
Pubblicazione: (2024)
Efficient scene text image super-resolution with semantic guidance
di: TomyEnrique, LeoWu, et al.
Pubblicazione: (2024)
di: TomyEnrique, LeoWu, et al.
Pubblicazione: (2024)
Ranking-aware adapter for text-driven image ordering with CLIP
di: Yu, Wei-Hsiang, et al.
Pubblicazione: (2024)
di: Yu, Wei-Hsiang, et al.
Pubblicazione: (2024)
IMMA: Immunizing text-to-image Models against Malicious Adaptation
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2023)
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2023)
FiCo-ITR: bridging fine-grained and coarse-grained image-text retrieval for comparative performance analysis
di: Williams-Lekuona, Mikel, et al.
Pubblicazione: (2024)
di: Williams-Lekuona, Mikel, et al.
Pubblicazione: (2024)
MultiWay-Adapater: Adapting large-scale multi-modal models for scalable image-text retrieval
di: Long, Zijun, et al.
Pubblicazione: (2023)
di: Long, Zijun, et al.
Pubblicazione: (2023)
Concept Corrector: Erase concepts on the fly for text-to-image diffusion models
di: Meng, Zheling, et al.
Pubblicazione: (2025)
di: Meng, Zheling, et al.
Pubblicazione: (2025)
Enhancement of text recognition for hanja handwritten documents of Ancient Korea
di: Ahna, Joonmo, et al.
Pubblicazione: (2024)
di: Ahna, Joonmo, et al.
Pubblicazione: (2024)
Decomposed evaluations of geographic disparities in text-to-image models
di: Sureddy, Abhishek, et al.
Pubblicazione: (2024)
di: Sureddy, Abhishek, et al.
Pubblicazione: (2024)
Task-driven single-image super-resolution reconstruction of document scans
di: Zyrek, Maciej, et al.
Pubblicazione: (2024)
di: Zyrek, Maciej, et al.
Pubblicazione: (2024)
Direct content-based retrieval from music scores images
di: Luna-Barahona, Noelia, et al.
Pubblicazione: (2026)
di: Luna-Barahona, Noelia, et al.
Pubblicazione: (2026)
Have we unified image generation and understanding yet? An empirical study of GPT-4o's image generation ability
di: Li, Ning, et al.
Pubblicazione: (2025)
di: Li, Ning, et al.
Pubblicazione: (2025)
Acquisition of interpretable domain information during brain MR image harmonization for content-based image retrieval
di: Abe, Keima, et al.
Pubblicazione: (2025)
di: Abe, Keima, et al.
Pubblicazione: (2025)
Dynamic watermarks in images generated by diffusion models
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
TurboEdit: Instant text-based image editing
di: Wu, Zongze, et al.
Pubblicazione: (2024)
di: Wu, Zongze, et al.
Pubblicazione: (2024)
A large-scale image-text dataset benchmark for farmland segmentation
di: Tao, Chao, et al.
Pubblicazione: (2025)
di: Tao, Chao, et al.
Pubblicazione: (2025)
Texture image retrieval using a classification and contourlet-based features
di: Rouhafzay, Asal, et al.
Pubblicazione: (2024)
di: Rouhafzay, Asal, et al.
Pubblicazione: (2024)
Are generative models fair? A study of racial bias in dermatological image generation
di: López-Pérez, Miguel, et al.
Pubblicazione: (2025)
di: López-Pérez, Miguel, et al.
Pubblicazione: (2025)
Consistency-diversity-realism Pareto fronts of conditional image generative models
di: Astolfi, Pietro, et al.
Pubblicazione: (2024)
di: Astolfi, Pietro, et al.
Pubblicazione: (2024)
Elucidating the design space of language models for image generation
di: Liu, Xuantong, et al.
Pubblicazione: (2024)
di: Liu, Xuantong, et al.
Pubblicazione: (2024)
Are handcrafted filters helpful for attributing AI-generated images?
di: Li, Jialiang, et al.
Pubblicazione: (2024)
di: Li, Jialiang, et al.
Pubblicazione: (2024)
Cardiac fat segmentation using computed tomography and an image-to-image conditional generative adversarial neural network
di: da Silva, Guilherme Santos, et al.
Pubblicazione: (2026)
di: da Silva, Guilherme Santos, et al.
Pubblicazione: (2026)
Cross-Modal Attention Alignment Network with Auxiliary Text Description for zero-shot sketch-based image retrieval
di: Su, Hanwen, et al.
Pubblicazione: (2024)
di: Su, Hanwen, et al.
Pubblicazione: (2024)
Zero-shot sketch-based remote sensing image retrieval based on multi-level and attention-guided tokenization
di: Yang, Bo, et al.
Pubblicazione: (2024)
di: Yang, Bo, et al.
Pubblicazione: (2024)
Domain-invariant feature learning in brain MR imaging for content-based image retrieval
di: Tobari, Shuya, et al.
Pubblicazione: (2025)
di: Tobari, Shuya, et al.
Pubblicazione: (2025)
Unfolder: Fast localization and image rectification of a document with a crease from folding in half
di: Ershov, A. M., et al.
Pubblicazione: (2023)
di: Ershov, A. M., et al.
Pubblicazione: (2023)
Documenti analoghi
-
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
di: McIntosh, Declan, et al.
Pubblicazione: (2026) -
Learning text-to-video retrieval from image captioning
di: Ventura, Lucas, et al.
Pubblicazione: (2024) -
Consistent text-to-image generation via scene de-contextualization
di: Tang, Song, et al.
Pubblicazione: (2025) -
Exploring scalable medical image encoders beyond text supervision
di: Pérez-García, Fernando, et al.
Pubblicazione: (2024) -
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
di: Tiwari, Amritanshu, et al.
Pubblicazione: (2025)