From Scope to Script: An Automated Report Generation Model for Gastrointestinal Endoscopy
Fuente:
arXiv
Saved in:
| Main Authors: | Kaklamanos, Evandros, Kristinsdottir, Kristjana, Huang, Jonathan, Carlson, Dustin, Keswani, Rajesh, Pandolfino, John, Etemadi, Mozziyar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Endoscopic Prediction of Achalasia: Putting the CART Before the CARS
by: Meng Li, et al.
Published: (2025)
by: Meng Li, et al.
Published: (2025)
Analogical Reasoning as a Doctor: A Foundation Model for Gastrointestinal Endoscopy Diagnosis
by: Peng, Peixi, et al.
Published: (2026)
by: Peng, Peixi, et al.
Published: (2026)
Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering
by: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Published: (2026)
by: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Published: (2026)
A Study on Self-Supervised Pretraining for Vision Problems in Gastrointestinal Endoscopy
by: Sanderson, Edward, et al.
Published: (2024)
by: Sanderson, Edward, et al.
Published: (2024)
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy
by: Roth, Marcel, et al.
Published: (2024)
by: Roth, Marcel, et al.
Published: (2024)
Endora: Video Generation Models as Endoscopy Simulators
by: Li, Chenxin, et al.
Published: (2024)
by: Li, Chenxin, et al.
Published: (2024)
Pairing-free Group-level Knowledge Distillation for Robust Gastrointestinal Lesion Classification in White-Light Endoscopy
by: Hu, Qiang, et al.
Published: (2026)
by: Hu, Qiang, et al.
Published: (2026)
Artificial Intelligence in Gastrointestinal Bleeding Analysis for Video Capsule Endoscopy: Insights, Innovations, and Prospects (2008-2023)
by: Singh, Tanisha, et al.
Published: (2024)
by: Singh, Tanisha, et al.
Published: (2024)
MetaScope: Optics-Driven Neural Network for Ultra-Micro Metalens Endoscopy
by: Li, Wuyang, et al.
Published: (2025)
by: Li, Wuyang, et al.
Published: (2025)
Capsule Vision Challenge 2024: Multi-Class Abnormality Classification for Video Capsule Endoscopy
by: Bansal, Aakarsh, et al.
Published: (2024)
by: Bansal, Aakarsh, et al.
Published: (2024)
GI-Bench: A Panoramic Benchmark Revealing the Knowledge-Experience Dissociation of Multimodal Large Language Models in Gastrointestinal Endoscopy Against Clinical Standards
by: Zhu, Yan, et al.
Published: (2026)
by: Zhu, Yan, et al.
Published: (2026)
Exploring Self-Supervised Learning with U-Net Masked Autoencoders and EfficientNet-B7 for Improved Gastrointestinal Abnormality Classification in Video Capsule Endoscopy
by: Kancharla, Vamshi Krishna, et al.
Published: (2024)
by: Kancharla, Vamshi Krishna, et al.
Published: (2024)
Transformer-Enhanced Iterative Feedback Mechanism for Polyp Segmentation
by: Tomar, Nikhil Kumar, et al.
Published: (2024)
by: Tomar, Nikhil Kumar, et al.
Published: (2024)
OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video
by: Pu, Junfu, et al.
Published: (2026)
by: Pu, Junfu, et al.
Published: (2026)
Automated Bleeding Detection and Classification in Wireless Capsule Endoscopy with YOLOv8-X
by: Shekar, Pavan C, et al.
Published: (2024)
by: Shekar, Pavan C, et al.
Published: (2024)
Diff-Oracle: Deciphering Oracle Bone Scripts with Controllable Diffusion Model
by: Li, Jing, et al.
Published: (2023)
by: Li, Jing, et al.
Published: (2023)
A Highlight Removal Method for Capsule Endoscopy Images
by: Zhang, Shaojie, et al.
Published: (2024)
by: Zhang, Shaojie, et al.
Published: (2024)
Capsule Endoscopy Image Enhancement for Small Intestinal Villi Clarity
by: Zhang, Shaojie, et al.
Published: (2024)
by: Zhang, Shaojie, et al.
Published: (2024)
Multimedia Generative Script Learning for Task Planning
by: Wang, Qingyun, et al.
Published: (2022)
by: Wang, Qingyun, et al.
Published: (2022)
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
by: Suzuki, Rena, et al.
Published: (2026)
by: Suzuki, Rena, et al.
Published: (2026)
Learning to Adapt Foundation Model DINOv2 for Capsule Endoscopy Diagnosis
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Vision-Language Models for Automated 3D PET/CT Report Generation
by: Jiao, Wenpei, et al.
Published: (2025)
by: Jiao, Wenpei, et al.
Published: (2025)
GPT4Motion: Scripting Physical Motions in Text-to-Video Generation via Blender-Oriented GPT Planning
by: Lv, Jiaxi, et al.
Published: (2023)
by: Lv, Jiaxi, et al.
Published: (2023)
FakeScope: Large Multimodal Expert Model for Transparent AI-Generated Image Forensics
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
DeepGI: An Automated Approach for Gastrointestinal Tract Segmentation in MRI Scans
by: Zhang, Ye, et al.
Published: (2024)
by: Zhang, Ye, et al.
Published: (2024)
SceneScript: Reconstructing Scenes With An Autoregressive Structured Language Model
by: Avetisyan, Armen, et al.
Published: (2024)
by: Avetisyan, Armen, et al.
Published: (2024)
MDIW-13: a New Multi-Lingual and Multi-Script Database and Benchmark for Script Identification
by: Ferrer, Miguel A., et al.
Published: (2024)
by: Ferrer, Miguel A., et al.
Published: (2024)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
by: Nguyen, Minh Anh, et al.
Published: (2026)
by: Nguyen, Minh Anh, et al.
Published: (2026)
Foundational Models for Pathology and Endoscopy Images: Application for Gastric Inflammation
by: Kerdegari, Hamideh, et al.
Published: (2024)
by: Kerdegari, Hamideh, et al.
Published: (2024)
EndoDINO: A Foundation Model for GI Endoscopy
by: Dermyer, Patrick, et al.
Published: (2025)
by: Dermyer, Patrick, et al.
Published: (2025)
Ancient Script Image Recognition and Processing: A Review
by: Diao, Xiaolei, et al.
Published: (2025)
by: Diao, Xiaolei, et al.
Published: (2025)
Prediction of Rectal Cancer Regrowth from Longitudinal Endoscopy
by: Gomez, Jorge Tapias, et al.
Published: (2026)
by: Gomez, Jorge Tapias, et al.
Published: (2026)
CAVE-Net: Classifying Abnormalities in Video Capsule Endoscopy
by: Harish, Ishita, et al.
Published: (2024)
by: Harish, Ishita, et al.
Published: (2024)
ACM Multimedia Grand Challenge on ENT Endoscopy Analysis
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization
by: Peng, Linkai, et al.
Published: (2026)
by: Peng, Linkai, et al.
Published: (2026)
MeshBrush: Painting the Anatomical Mesh with Neural Stylization for Endoscopy
by: Han, John J., et al.
Published: (2024)
by: Han, John J., et al.
Published: (2024)
Historic Scripts to Modern Vision: A Novel Dataset and A VLM Framework for Transliteration of Modi Script to Devanagari
by: Kausadikar, Harshal, et al.
Published: (2025)
by: Kausadikar, Harshal, et al.
Published: (2025)
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy
by: Li, Bojian, et al.
Published: (2024)
by: Li, Bojian, et al.
Published: (2024)
Foundation Model for Endoscopy Video Analysis via Large-scale Self-supervised Pre-train
by: Wang, Zhao, et al.
Published: (2023)
by: Wang, Zhao, et al.
Published: (2023)
EndoBench: A Comprehensive Evaluation of Multi-Modal Large Language Models for Endoscopy Analysis
by: Liu, Shengyuan, et al.
Published: (2025)
by: Liu, Shengyuan, et al.
Published: (2025)
Similar Items
-
Endoscopic Prediction of Achalasia: Putting the CART Before the CARS
by: Meng Li, et al.
Published: (2025) -
Analogical Reasoning as a Doctor: A Foundation Model for Gastrointestinal Endoscopy Diagnosis
by: Peng, Peixi, et al.
Published: (2026) -
Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering
by: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Published: (2026) -
A Study on Self-Supervised Pretraining for Vision Problems in Gastrointestinal Endoscopy
by: Sanderson, Edward, et al.
Published: (2024) -
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy
by: Roth, Marcel, et al.
Published: (2024)