A Preliminary Survey of Semantic Descriptive Model for Images
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yan, Chengxi, Jian, Jie, Li, Yang |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A First Ontological Model for the Description of the Art Market in the Semantic Web
par: Veggi, Manuele
Publié: (2024)
par: Veggi, Manuele
Publié: (2024)
Ultrasound Image Generation using Latent Diffusion Models
par: Freiche, Benoit, et autres
Publié: (2025)
par: Freiche, Benoit, et autres
Publié: (2025)
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
par: Beyene, Fitsum Sileshi, et autres
Publié: (2026)
par: Beyene, Fitsum Sileshi, et autres
Publié: (2026)
Callico: a Versatile Open-Source Document Image Annotation Platform
par: Kermorvant, Christopher, et autres
Publié: (2024)
par: Kermorvant, Christopher, et autres
Publié: (2024)
BROTHER: Behavioral Recognition Optimized Through Heterogeneous Ensemble Regularization for Ambivalence and Hesitancy
par: Pereira, Alexandre, et autres
Publié: (2026)
par: Pereira, Alexandre, et autres
Publié: (2026)
Drafting the Landscape of Computational Musicology Tools: a Survey-Based Approach
par: Vega, Jorge Junior Morgado, et autres
Publié: (2025)
par: Vega, Jorge Junior Morgado, et autres
Publié: (2025)
Evolving Thematic Map Design in Academic Cartography: A Thirty-Year Study Based on Multilingual Journals
par: Wei, Zhiwei, et autres
Publié: (2026)
par: Wei, Zhiwei, et autres
Publié: (2026)
A Literature Review of Literature Reviews in Pattern Analysis and Machine Intelligence
par: Zhao, Penghai, et autres
Publié: (2024)
par: Zhao, Penghai, et autres
Publié: (2024)
LEGATO: Large-scale End-to-end Generalizable Approach to Typeset OMR
par: Yang, Guang, et autres
Publié: (2025)
par: Yang, Guang, et autres
Publié: (2025)
Unfolding the Past: A Comprehensive Deep Learning Approach to Analyzing Incunabula Pages
par: Ropel, Klaudia, et autres
Publié: (2025)
par: Ropel, Klaudia, et autres
Publié: (2025)
NeuroPapyri: A Deep Attention Embedding Network for Handwritten Papyri Retrieval
par: De Gregorio, Giuseppe, et autres
Publié: (2024)
par: De Gregorio, Giuseppe, et autres
Publié: (2024)
What Lies Beneath: A Call for Distribution-based Visual Question & Answer Datasets
par: Naiman, Jill P., et autres
Publié: (2026)
par: Naiman, Jill P., et autres
Publié: (2026)
WildDepth: A Multimodal Dataset for 3D Wildlife Perception and Depth Estimation
par: Aamir, Muhammad, et autres
Publié: (2026)
par: Aamir, Muhammad, et autres
Publié: (2026)
Taming Anomalies with Down-Up Sampling Networks: Group Center Preserving Reconstruction for 3D Anomaly Detection
par: Liang, Hanzhe, et autres
Publié: (2025)
par: Liang, Hanzhe, et autres
Publié: (2025)
Hybrid Retrieval-Augmented Generation for Robust Multilingual Document Question Answering
par: Mudet, Anthony, et autres
Publié: (2025)
par: Mudet, Anthony, et autres
Publié: (2025)
Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?
par: Pippi, Vittorio, et autres
Publié: (2025)
par: Pippi, Vittorio, et autres
Publié: (2025)
Knowledge Graphs for Digitized Manuscripts in Jagiellonian Digital Library Application
par: Ignatowicz, Jan, et autres
Publié: (2025)
par: Ignatowicz, Jan, et autres
Publié: (2025)
Automatic Reviewers Assignment to a Research Paper Based on Allied References and Publications Weight
par: Mahmud, Tamim Al, et autres
Publié: (2025)
par: Mahmud, Tamim Al, et autres
Publié: (2025)
Ink Detection from Surface Topography of the Herculaneum Papyri
par: Angelotti, Giorgio, et autres
Publié: (2026)
par: Angelotti, Giorgio, et autres
Publié: (2026)
Structured Analysis and Comparison of Alphabets in Historical Handwritten Ciphers
par: Méndez, Martín, et autres
Publié: (2024)
par: Méndez, Martín, et autres
Publié: (2024)
Automatic Recognition of Learning Resource Category in a Digital Library
par: Banerjee, Soumya, et autres
Publié: (2023)
par: Banerjee, Soumya, et autres
Publié: (2023)
An Intelligent Framework for Real-Time Yoga Pose Detection and Posture Correction
par: Haldar, Chandramouli
Publié: (2026)
par: Haldar, Chandramouli
Publié: (2026)
Productivity profile of CNPq scholarship researchers in computer science from 2017 to 2021
par: Albertini, Marcelo Keese, et autres
Publié: (2024)
par: Albertini, Marcelo Keese, et autres
Publié: (2024)
μgat: Improving Single-Page Document Parsing by Providing Multi-Page Context
par: Quattrini, Fabio, et autres
Publié: (2024)
par: Quattrini, Fabio, et autres
Publié: (2024)
Segmentation of Ink and Parchment in Dead Sea Scroll Fragments
par: Kurar-Barakat, Berat, et autres
Publié: (2024)
par: Kurar-Barakat, Berat, et autres
Publié: (2024)
[Citation needed] Data usage and citation practices in medical imaging conferences
par: Sourget, Théo, et autres
Publié: (2024)
par: Sourget, Théo, et autres
Publié: (2024)
Studying Maps at Scale: A Digital Investigation of Cartography and the Evolution of Figuration
par: Petitpierre, Remi
Publié: (2025)
par: Petitpierre, Remi
Publié: (2025)
Dual-Teacher Ensemble Models with Double-Copy-Paste for 3D Semi-Supervised Medical Image Segmentation
par: Fa, Zhan, et autres
Publié: (2024)
par: Fa, Zhan, et autres
Publié: (2024)
[Re] Network Deconvolution
par: Obadage, Rochana R., et autres
Publié: (2024)
par: Obadage, Rochana R., et autres
Publié: (2024)
Position-Prior-Guided Network for System Matrix Super-Resolution in Magnetic Particle Imaging
par: Geng, Xuqing, et autres
Publié: (2025)
par: Geng, Xuqing, et autres
Publié: (2025)
Automatic staff reconstruction within SIMSSA proect
par: Tardon, Lorenzo J., et autres
Publié: (2024)
par: Tardon, Lorenzo J., et autres
Publié: (2024)
A Semantic Segmentation Algorithm for Pleural Effusion Based on DBIF-AUNet
par: Tang, Ruixiang, et autres
Publié: (2025)
par: Tang, Ruixiang, et autres
Publié: (2025)
An HTR-LLM Workflow for High-Accuracy Transcription and Analysis of Abbreviated Latin Court Hand
par: Isom, Joshua D.
Publié: (2025)
par: Isom, Joshua D.
Publié: (2025)
Object Recognition from Scientific Document based on Compartment Refinement Framework
par: Li, Jinghong, et autres
Publié: (2023)
par: Li, Jinghong, et autres
Publié: (2023)
Automatic Modeling of Social Concepts Evoked by Art Images as Multimodal Frames
par: Pandiani, Delfina Sol Martinez, et autres
Publié: (2021)
par: Pandiani, Delfina Sol Martinez, et autres
Publié: (2021)
A Global Atlas of Digital Dermatology to Map Innovation and Disparities
par: Gröger, Fabian, et autres
Publié: (2025)
par: Gröger, Fabian, et autres
Publié: (2025)
Allure of Craquelure: A Variational-Generative Approach to Crack Detection in Paintings
par: Paul, Laura, et autres
Publié: (2026)
par: Paul, Laura, et autres
Publié: (2026)
Multimodal LLMs for Historical Dataset Construction from Archival Image Scans: German Patents (1877-1918)
par: Griesshaber, Niclas, et autres
Publié: (2025)
par: Griesshaber, Niclas, et autres
Publié: (2025)
Detecting Latin in Historical Books with Large Language Models: A Multimodal Benchmark
par: Wu, Yu, et autres
Publié: (2025)
par: Wu, Yu, et autres
Publié: (2025)
Position: The Artificial Intelligence and Machine Learning Community Should Adopt a More Transparent and Regulated Peer Review Process
par: Yang, Jing
Publié: (2025)
par: Yang, Jing
Publié: (2025)
Documents similaires
-
A First Ontological Model for the Description of the Art Market in the Semantic Web
par: Veggi, Manuele
Publié: (2024) -
Ultrasound Image Generation using Latent Diffusion Models
par: Freiche, Benoit, et autres
Publié: (2025) -
A Survey of OCR Evaluation Methods and Metrics and the Invisibility of Historical Documents
par: Beyene, Fitsum Sileshi, et autres
Publié: (2026) -
Callico: a Versatile Open-Source Document Image Annotation Platform
par: Kermorvant, Christopher, et autres
Publié: (2024) -
BROTHER: Behavioral Recognition Optimized Through Heterogeneous Ensemble Regularization for Ambivalence and Hesitancy
par: Pereira, Alexandre, et autres
Publié: (2026)