Prompt-Based Caption Generation for Single-Tooth Dental Images Using Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sukhanova, Anastasiia, Taylor, Aiden, Myers, Julian, Wang, Zichun, Jammuladinne, Kartha Veerya, Nimmagadda, Satya Sri Rajiteswari, Maiti, Aniruddha, Jana, Ananya |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convergence of Outputs When Two Large Language Models Interact in a Multi-Agentic Setup
by: Maiti, Aniruddha, et al.
Published: (2025)
by: Maiti, Aniruddha, et al.
Published: (2025)
Text to Automata Diagrams: Comparing TikZ Code Generation with Direct Image Synthesis
by: Young, Ethan, et al.
Published: (2026)
by: Young, Ethan, et al.
Published: (2026)
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
by: Nimmagadda, Satya Sri Rajiteswari, et al.
Published: (2026)
by: Nimmagadda, Satya Sri Rajiteswari, et al.
Published: (2026)
Comparative Analysis of OpenAI GPT-4o and DeepSeek R1 for Scientific Text Categorization Using Prompt Engineering
by: Maiti, Aniruddha, et al.
Published: (2025)
by: Maiti, Aniruddha, et al.
Published: (2025)
Evaluating the Suitability of Different Intraoral Scan Resolutions for Deep Learning-Based Tooth Segmentation
by: Weekley, Daron, et al.
Published: (2025)
by: Weekley, Daron, et al.
Published: (2025)
Intraoral Coil Arrays for Single‐Tooth Dental MRI
by: Ali Caglar Özen, et al.
Published: (2025)
by: Ali Caglar Özen, et al.
Published: (2025)
Target-Dependent Multimodal Sentiment Analysis Via Employing Visual-to Emotional-Caption Translation Network using Visual-Caption Pairs
by: Pandey, Ananya, et al.
Published: (2024)
by: Pandey, Ananya, et al.
Published: (2024)
Distortion‐less video wireless transmission in 5G new radio using delay‐distortion‐rate optimization (DDRO)
by: K. Maheswari, et al.
Published: (2024)
by: K. Maheswari, et al.
Published: (2024)
Toward Generalizable Surrogate Models for Molecular Dynamics via Graph Neural Networks
by: Immanuel, Judah, et al.
Published: (2025)
by: Immanuel, Judah, et al.
Published: (2025)
Image Captions are Natural Prompts for Text-to-Image Models
by: Lei, Shiye, et al.
Published: (2023)
by: Lei, Shiye, et al.
Published: (2023)
Knowledge Retention in Tooth Morphology Among Dental Hygienists
by: Anna Tostrup Kristensen, et al.
Published: (2025)
by: Anna Tostrup Kristensen, et al.
Published: (2025)
Quantum Inverse Contextual Vision Transformers (Q-ICVT): A New Frontier in 3D Object Detection for AVs
by: Dharavath, Sanjay Bhargav, et al.
Published: (2024)
by: Dharavath, Sanjay Bhargav, et al.
Published: (2024)
Text Data-Centric Image Captioning with Interactive Prompts
by: Wang, Yiyu, et al.
Published: (2024)
by: Wang, Yiyu, et al.
Published: (2024)
ArchMap: Arch-Flattening and Knowledge-Guided Vision Language Model for Tooth Counting and Structured Dental Understanding
by: Zhang, Bohan, et al.
Published: (2025)
by: Zhang, Bohan, et al.
Published: (2025)
Evolution of suitable type of bottom trawls for the medium size steel trawlers of Orissa Fisheries Department
by: Kartha, K.N.
Published: (1976)
by: Kartha, K.N.
Published: (1976)
RORPCap: Retrieval-based Objects and Relations Prompt for Image Captioning
by: Gu, Jinjing, et al.
Published: (2025)
by: Gu, Jinjing, et al.
Published: (2025)
Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M)
by: Merchant, Nicholas, et al.
Published: (2025)
by: Merchant, Nicholas, et al.
Published: (2025)
Modelling Visual Semantics via Image Captioning to extract Enhanced Multi-Level Cross-Modal Semantic Incongruity Representation with Attention for Multimodal Sarcasm Detection
by: Aggarwal, Sajal, et al.
Published: (2024)
by: Aggarwal, Sajal, et al.
Published: (2024)
Effectively Enhancing Vision Language Large Models by Prompt Augmentation and Caption Utilization
by: Zhao, Minyi, et al.
Published: (2024)
by: Zhao, Minyi, et al.
Published: (2024)
Linear Alignment of Vision-language Models for Image Captioning
by: Paischer, Fabian, et al.
Published: (2023)
by: Paischer, Fabian, et al.
Published: (2023)
Mitigating Image Captioning Hallucinations in Vision-Language Models
by: Zhao, Fei, et al.
Published: (2025)
by: Zhao, Fei, et al.
Published: (2025)
DualPrompt-MedCap: A Dual-Prompt Enhanced Approach for Medical Image Captioning
by: Zhao, Yining, et al.
Published: (2025)
by: Zhao, Yining, et al.
Published: (2025)
RACap: Relation-Aware Prompting for Lightweight Retrieval-Augmented Image Captioning
by: Long, Xiaosheng, et al.
Published: (2025)
by: Long, Xiaosheng, et al.
Published: (2025)
Going over Fine Web with a Fine-Tooth Comb: Technical Report of Indexing Fine Web for Problematic Content Search and Retrieval
by: Marinas, Inés Altemir, et al.
Published: (2025)
by: Marinas, Inés Altemir, et al.
Published: (2025)
DENTEX: Dental Enumeration and Tooth Pathosis Detection Benchmark for Panoramic X-ray
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2023)
ToothForge: Automatic Dental Shape Generation using Synchronized Spectral Embeddings
by: Kubík, Tibor, et al.
Published: (2025)
by: Kubík, Tibor, et al.
Published: (2025)
A Novel Double Arrowhead Auxetic Dental Splint for Avulsed Tooth Reimplantation
by: Mariam Maroof, et al.
Published: (2026)
by: Mariam Maroof, et al.
Published: (2026)
Reliability of the Assessment of Tooth Wear Severity on Dental Hard Tissues and Dental Restorations, Using the TWES 2.0, by Nonexperts
by: Jakob C. Roehl, et al.
Published: (2024)
by: Jakob C. Roehl, et al.
Published: (2024)
Sparse Anatomical Prompt Semi-Supervised Learning with Masked Image Modeling for CBCT Tooth Segmentation
by: Dai, Pengyu, et al.
Published: (2024)
by: Dai, Pengyu, et al.
Published: (2024)
No‐Prep Zirconia Cantilever Resin‐Bonded Fixed Dental Prostheses: A Noninvasive, Simple Approach to Replacing a Single Missing Tooth
by: Amélie Karine Jacques Mainjot
Published: (2025)
by: Amélie Karine Jacques Mainjot
Published: (2025)
OPCap:Object-aware Prompting Captioning
by: Huang, Feiyang
Published: (2024)
by: Huang, Feiyang
Published: (2024)
Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts
by: Özdemir, Övgü, et al.
Published: (2024)
by: Özdemir, Övgü, et al.
Published: (2024)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
by: Kim, Taewhan, et al.
Published: (2024)
by: Kim, Taewhan, et al.
Published: (2024)
Rapid Assessment of Stable Crystal Structures in Single Phase High Entropy Alloys Via Graph Neural Network Based Surrogate Modelling
by: Beaver, Nicholas, et al.
Published: (2024)
by: Beaver, Nicholas, et al.
Published: (2024)
SaViD: Spectravista Aesthetic Vision Integration for Robust and Discerning 3D Object Detection in Challenging Environments
by: Dam, Tanmoy, et al.
Published: (2025)
by: Dam, Tanmoy, et al.
Published: (2025)
RxnCaption: Reformulating Reaction Diagram Parsing as Visual Prompt Guided Captioning
by: Song, Jiahe, et al.
Published: (2025)
by: Song, Jiahe, et al.
Published: (2025)
Tight Gap-Dependent Memory-Regret Trade-Off for Single-Pass Streaming Stochastic Multi-Armed Bandits
by: Ye, Zichun, et al.
Published: (2025)
by: Ye, Zichun, et al.
Published: (2025)
A Picture is Worth a Thousand Prompts? Efficacy of Iterative Human-Driven Prompt Refinement in Image Regeneration Tasks
by: Trinh, Khoi, et al.
Published: (2025)
by: Trinh, Khoi, et al.
Published: (2025)
A Multi-Stage Framework for 3D Individual Tooth Segmentation in Dental CBCT
by: Wang, Chunshi, et al.
Published: (2024)
by: Wang, Chunshi, et al.
Published: (2024)
Effectiveness of Flipped Learning Compared to Conventional Instruction in Teaching Tooth Extraction to Dental Students
by: Emine Asena Singer, et al.
Published: (2025)
by: Emine Asena Singer, et al.
Published: (2025)
Similar Items
-
Convergence of Outputs When Two Large Language Models Interact in a Multi-Agentic Setup
by: Maiti, Aniruddha, et al.
Published: (2025) -
Text to Automata Diagrams: Comparing TikZ Code Generation with Direct Image Synthesis
by: Young, Ethan, et al.
Published: (2026) -
Generating Hierarchical JSON Representations of Scientific Sentences Using LLMs
by: Nimmagadda, Satya Sri Rajiteswari, et al.
Published: (2026) -
Comparative Analysis of OpenAI GPT-4o and DeepSeek R1 for Scientific Text Categorization Using Prompt Engineering
by: Maiti, Aniruddha, et al.
Published: (2025) -
Evaluating the Suitability of Different Intraoral Scan Resolutions for Deep Learning-Based Tooth Segmentation
by: Weekley, Daron, et al.
Published: (2025)