Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Meinecke, Christofer, Guéville, Estelle, Wrisley, David Joseph |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Is Medieval Distant Viewing Possible? : Extending and Enriching Annotation of Legacy Image Collections using Visual Analytics
di: Meinecke, Christofer, et al.
Pubblicazione: (2022)
di: Meinecke, Christofer, et al.
Pubblicazione: (2022)
Labeling of Cultural Heritage Collections on the Intersection of Visual Analytics and Digital Humanities
di: Meinecke, Christofer
Pubblicazione: (2022)
di: Meinecke, Christofer
Pubblicazione: (2022)
Medieval Manuscripts and the Computational Humanities
di: Wrisley, David Joseph, et al.
Pubblicazione: (2026)
di: Wrisley, David Joseph, et al.
Pubblicazione: (2026)
Transcribing Medieval Manuscripts for Machine Learning
di: Guéville, Estelle, et al.
Pubblicazione: (2022)
di: Guéville, Estelle, et al.
Pubblicazione: (2022)
An Interactive Decision Support System for Analyzing Time Related Restrictions in Renaturation and Redevelopment Planning Projects
di: Annanias, Yves, et al.
Pubblicazione: (2023)
di: Annanias, Yves, et al.
Pubblicazione: (2023)
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2025)
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2025)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
di: Voss, Hendric, et al.
Pubblicazione: (2025)
di: Voss, Hendric, et al.
Pubblicazione: (2025)
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
di: Foteinopoulou, Niki Maria, et al.
Pubblicazione: (2023)
di: Foteinopoulou, Niki Maria, et al.
Pubblicazione: (2023)
Zero-Shot Pupil Segmentation with SAM 2: A Case Study of Over 14 Million Images
di: Maquiling, Virmarie, et al.
Pubblicazione: (2024)
di: Maquiling, Virmarie, et al.
Pubblicazione: (2024)
Plug-and-Play Clarifier: A Zero-Shot Multimodal Framework for Egocentric Intent Disambiguation
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
Quantitative Analysis of Objects in Prisoner Artworks
di: Christoffersen, Thea, et al.
Pubblicazione: (2025)
di: Christoffersen, Thea, et al.
Pubblicazione: (2025)
Zero-Shot Segmentation of Eye Features Using the Segment Anything Model (SAM)
di: Maquiling, Virmarie, et al.
Pubblicazione: (2023)
di: Maquiling, Virmarie, et al.
Pubblicazione: (2023)
Extracting Human Attention through Crowdsourced Patch Labeling
di: Chang, Minsuk, et al.
Pubblicazione: (2024)
di: Chang, Minsuk, et al.
Pubblicazione: (2024)
HaDR: Applying Domain Randomization for Generating Synthetic Multimodal Dataset for Hand Instance Segmentation in Cluttered Industrial Environments
di: Grushko, Stefan, et al.
Pubblicazione: (2023)
di: Grushko, Stefan, et al.
Pubblicazione: (2023)
Viewpoint Recommendation for Point Cloud Labeling through Interaction Cost Modeling
di: Zhang, Yu, et al.
Pubblicazione: (2026)
di: Zhang, Yu, et al.
Pubblicazione: (2026)
Self-Calibrating BCIs: Ranking and Recovery of Mental Targets Without Labels
di: Grizou, Jonathan, et al.
Pubblicazione: (2025)
di: Grizou, Jonathan, et al.
Pubblicazione: (2025)
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
di: Zhang, Dongping, et al.
Pubblicazione: (2024)
di: Zhang, Dongping, et al.
Pubblicazione: (2024)
Collection Space Navigator: An Interactive Visualization Interface for Multidimensional Datasets
di: Ohm, Tillmann, et al.
Pubblicazione: (2023)
di: Ohm, Tillmann, et al.
Pubblicazione: (2023)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
Prints in the Magnetic Dust: Robust Similarity Search in Legacy Media Images Using Checksum Count Vectors
di: Grzeszczuk, Maciej, et al.
Pubblicazione: (2026)
di: Grzeszczuk, Maciej, et al.
Pubblicazione: (2026)
Creo: From One-Shot Image Generation to Progressive, Co-Creative Ideation
di: De Simone, Zoe, et al.
Pubblicazione: (2026)
di: De Simone, Zoe, et al.
Pubblicazione: (2026)
UniHands: Unifying Various Wild-Collected Keypoints for Personalized Hand Reconstruction
di: Zhang, Menghe, et al.
Pubblicazione: (2024)
di: Zhang, Menghe, et al.
Pubblicazione: (2024)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
di: Miyoshi, Ryo, et al.
Pubblicazione: (2025)
di: Miyoshi, Ryo, et al.
Pubblicazione: (2025)
Detecting Clues for Skill Levels and Machine Operation Difficulty from Egocentric Vision
di: Long-fei, Chen, et al.
Pubblicazione: (2019)
di: Long-fei, Chen, et al.
Pubblicazione: (2019)
mEBAL: A Multimodal Database for Eye Blink Detection and Attention Level Estimation
di: Daza, Roberto, et al.
Pubblicazione: (2020)
di: Daza, Roberto, et al.
Pubblicazione: (2020)
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation
di: Zhang, He, et al.
Pubblicazione: (2025)
di: Zhang, He, et al.
Pubblicazione: (2025)
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
di: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Pubblicazione: (2024)
di: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Pubblicazione: (2024)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
di: Wong, David C, et al.
Pubblicazione: (2025)
di: Wong, David C, et al.
Pubblicazione: (2025)
Semantic Draw Engineering for Text-to-Image Creation
di: Li, Yang, et al.
Pubblicazione: (2023)
di: Li, Yang, et al.
Pubblicazione: (2023)
QuizRank: Picking Images by Quizzing VLMs
di: Ji, Tenghao, et al.
Pubblicazione: (2025)
di: Ji, Tenghao, et al.
Pubblicazione: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
di: Yamani, Asma, et al.
Pubblicazione: (2024)
di: Yamani, Asma, et al.
Pubblicazione: (2024)
SCHEMA for Gemini 3 Pro Image: A Structured Methodology for Controlled AI Image Generation on Google's Native Multimodal Model
di: Cazzaniga, Luca
Pubblicazione: (2026)
di: Cazzaniga, Luca
Pubblicazione: (2026)
VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation
di: Zhao, Yiming, et al.
Pubblicazione: (2026)
di: Zhao, Yiming, et al.
Pubblicazione: (2026)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
di: Hall, Melissa, et al.
Pubblicazione: (2023)
di: Hall, Melissa, et al.
Pubblicazione: (2023)
Steering Generative Models for Accessibility: EasyRead Image Generation
di: Dickenmann, Nicolas, et al.
Pubblicazione: (2026)
di: Dickenmann, Nicolas, et al.
Pubblicazione: (2026)
Accurate Eye Tracking from Dense 3D Surface Reconstructions using Single-Shot Deflectometry
di: Wang, Jiazhang, et al.
Pubblicazione: (2023)
di: Wang, Jiazhang, et al.
Pubblicazione: (2023)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
Development of a Mobile Application for at-Home Analysis of Retinal Fundus Images
di: Reid, Mattea, et al.
Pubblicazione: (2025)
di: Reid, Mattea, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Is Medieval Distant Viewing Possible? : Extending and Enriching Annotation of Legacy Image Collections using Visual Analytics
di: Meinecke, Christofer, et al.
Pubblicazione: (2022) -
Labeling of Cultural Heritage Collections on the Intersection of Visual Analytics and Digital Humanities
di: Meinecke, Christofer
Pubblicazione: (2022) -
Medieval Manuscripts and the Computational Humanities
di: Wrisley, David Joseph, et al.
Pubblicazione: (2026) -
Transcribing Medieval Manuscripts for Machine Learning
di: Guéville, Estelle, et al.
Pubblicazione: (2022) -
An Interactive Decision Support System for Analyzing Time Related Restrictions in Renaturation and Redevelopment Planning Projects
di: Annanias, Yves, et al.
Pubblicazione: (2023)