ACM Multimedia Grand Challenge on ENT Endoscopy Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | , , , , , , , , |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
| _version_ | 1866911095261757440 |
|---|---|
| author | Nguyen, Trong-Thuan Huynh, Viet-Tham Dao, Thao Thi Phuong Thi, Ha Nguyen Thuy, Tien To Vu Tran, Uyen Hanh Nguyen, Tam V. Le, Thanh Dinh Tran, Minh-Triet |
| author_facet | Nguyen, Trong-Thuan Huynh, Viet-Tham Dao, Thao Thi Phuong Thi, Ha Nguyen Thuy, Tien To Vu Tran, Uyen Hanh Nguyen, Tam V. Le, Thanh Dinh Tran, Minh-Triet |
| contents | Automated analysis of endoscopic imagery is a critical yet underdeveloped component of ENT (ear, nose, and throat) care, hindered by variability in devices and operators, subtle and localized findings, and fine-grained distinctions such as laterality and vocal-fold state. In addition to classification, clinicians require reliable retrieval of similar cases, both visually and through concise textual descriptions. These capabilities are rarely supported by existing public benchmarks. To this end, we introduce ENTRep, the ACM Multimedia 2025 Grand Challenge on ENT endoscopy analysis, which integrates fine-grained anatomical classification with image-to-image and text-to-image retrieval under bilingual (Vietnamese and English) clinical supervision. Specifically, the dataset comprises expert-annotated images, labeled for anatomical region and normal or abnormal status, and accompanied by dual-language narrative descriptions. In addition, we define three benchmark tasks, standardize the submission protocol, and evaluate performance on public and private test splits using server-side scoring. Moreover, we report results from the top-performing teams and provide an insight discussion. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2508_04801 |
| institution | arXiv |
| publishDate | 2025 |
| record_format | arxiv |
| spellingShingle | ACM Multimedia Grand Challenge on ENT Endoscopy Analysis Nguyen, Trong-Thuan Huynh, Viet-Tham Dao, Thao Thi Phuong Thi, Ha Nguyen Thuy, Tien To Vu Tran, Uyen Hanh Nguyen, Tam V. Le, Thanh Dinh Tran, Minh-Triet Computer Vision and Pattern Recognition Automated analysis of endoscopic imagery is a critical yet underdeveloped component of ENT (ear, nose, and throat) care, hindered by variability in devices and operators, subtle and localized findings, and fine-grained distinctions such as laterality and vocal-fold state. In addition to classification, clinicians require reliable retrieval of similar cases, both visually and through concise textual descriptions. These capabilities are rarely supported by existing public benchmarks. To this end, we introduce ENTRep, the ACM Multimedia 2025 Grand Challenge on ENT endoscopy analysis, which integrates fine-grained anatomical classification with image-to-image and text-to-image retrieval under bilingual (Vietnamese and English) clinical supervision. Specifically, the dataset comprises expert-annotated images, labeled for anatomical region and normal or abnormal status, and accompanied by dual-language narrative descriptions. In addition, we define three benchmark tasks, standardize the submission protocol, and evaluate performance on public and private test splits using server-side scoring. Moreover, we report results from the top-performing teams and provide an insight discussion. |
| title | ACM Multimedia Grand Challenge on ENT Endoscopy Analysis |
| topic | Computer Vision and Pattern Recognition |
| url | https://arxiv.org/abs/2508.04801 |