Development and multi-center evaluation of domain-adapted speech recognition for human-AI teaming in real-world gastrointestinal endoscopy
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Ruijie, Zhu, Yan, Fu, Peiyao, Luo, Te, Wang, Zhihua, Yang, Xian, Li, Quanlin, Zhou, Pinghong, Wang, Shuo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations
por: Yang, Ruijie, et al.
Publicado: (2025)
por: Yang, Ruijie, et al.
Publicado: (2025)
EndoFinder: Online Image Retrieval for Explainable Colorectal Polyp Diagnosis
por: Yang, Ruijie, et al.
Publicado: (2024)
por: Yang, Ruijie, et al.
Publicado: (2024)
One-shot synthesis of rare gastrointestinal lesions improves diagnostic accuracy and clinical training
por: Yu, Jia, et al.
Publicado: (2025)
por: Yu, Jia, et al.
Publicado: (2025)
Endo-CLIP: Progressive Self-Supervised Pre-training on Raw Colonoscopy Records
por: He, Yili, et al.
Publicado: (2025)
por: He, Yili, et al.
Publicado: (2025)
Robust Polyp Detection and Diagnosis through Compositional Prompt-Guided Diffusion Models
por: Yu, Jia, et al.
Publicado: (2025)
por: Yu, Jia, et al.
Publicado: (2025)
Teaching the Teachers: Boosting unsupervised domain adaptation in speech recognition by ensemble update
por: Ahmad, Rehan, et al.
Publicado: (2026)
por: Ahmad, Rehan, et al.
Publicado: (2026)
UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment
por: Wang, Zheng, et al.
Publicado: (2026)
por: Wang, Zheng, et al.
Publicado: (2026)
asr_eval: Algorithms and tools for multi-reference and streaming speech recognition evaluation
por: Sedukhin, Oleg, et al.
Publicado: (2026)
por: Sedukhin, Oleg, et al.
Publicado: (2026)
Pediatric advanced complex endoscopy team enhances endoscopy quality and provider satisfaction
por: Monique T. Barakat, et al.
Publicado: (2024)
por: Monique T. Barakat, et al.
Publicado: (2024)
The application of the combination between artificial intelligence and endoscopy in gastrointestinal tumors
por: Shen Li, et al.
Publicado: (2024)
por: Shen Li, et al.
Publicado: (2024)
Graph-based multi-Feature fusion method for speech emotion recognition
por: Liu, Xueyu, et al.
Publicado: (2024)
por: Liu, Xueyu, et al.
Publicado: (2024)
LLM-based phoneme-to-grapheme for phoneme-based speech recognition
por: Ma, Te, et al.
Publicado: (2025)
por: Ma, Te, et al.
Publicado: (2025)
Phoneme-based speech recognition driven by large language models and sampling marginalization
por: Ma, Te, et al.
Publicado: (2025)
por: Ma, Te, et al.
Publicado: (2025)
Multi-channel multi-speaker transformer for speech recognition
por: Yifan, Guo, et al.
Publicado: (2026)
por: Yifan, Guo, et al.
Publicado: (2026)
Snooze but don't lose: Remimazolam for sedation in gastrointestinal endoscopy
por: Yousuke Nakai
Publicado: (2025)
por: Yousuke Nakai
Publicado: (2025)
Multi-contrast laser endoscopy for in vivo gastrointestinal imaging
por: Bobrow, Taylor L., et al.
Publicado: (2025)
por: Bobrow, Taylor L., et al.
Publicado: (2025)
Broadband tunable microwave photonic radar for simultaneous detection of human respiration, heartbeat, and speech with deep learning-based speech recognition
por: Gao, Lei, et al.
Publicado: (2025)
por: Gao, Lei, et al.
Publicado: (2025)
Introduction to speech recognition
por: Dauphin, Gabriel
Publicado: (2024)
por: Dauphin, Gabriel
Publicado: (2024)
Video capsule endoscopy in overt and occult obscure gastrointestinal bleeding: Insights from a single‐center, observational study in Japan
por: Anna Tojo, et al.
Publicado: (2024)
por: Anna Tojo, et al.
Publicado: (2024)
learning discriminative features from spectrograms using center loss for speech emotion recognition
por: Dai, Dongyang, et al.
Publicado: (2025)
por: Dai, Dongyang, et al.
Publicado: (2025)
Toward human-centered shared autonomy AI paradigms for human-robot teaming in healthcare
por: Abiri, Reza, et al.
Publicado: (2024)
por: Abiri, Reza, et al.
Publicado: (2024)
Combo: Co-speech holistic 3D human motion generation and efficient customizable adaptation in harmony
por: Xu, Chao, et al.
Publicado: (2024)
por: Xu, Chao, et al.
Publicado: (2024)
The evaluation of a code-switched Sepedi-English automatic speech recognition system
por: Phaladi, Amanda, et al.
Publicado: (2024)
por: Phaladi, Amanda, et al.
Publicado: (2024)
Prominence-aware automatic speech recognition for conversational speech
por: Linke, Julian, et al.
Publicado: (2025)
por: Linke, Julian, et al.
Publicado: (2025)
Can AI detect pain and express pain empathy? A review from emotion recognition and a human-centered AI perspective
por: Cao, Siqi, et al.
Publicado: (2021)
por: Cao, Siqi, et al.
Publicado: (2021)
Zipformer: A faster and better encoder for automatic speech recognition
por: Yao, Zengwei, et al.
Publicado: (2023)
por: Yao, Zengwei, et al.
Publicado: (2023)
Tele‐rehabilitation in COVID‐19 survivors (TERCOV): An investigator‐initiated, prospective, multi‐center, real‐world study
por: Geyi Wen, et al.
Publicado: (2024)
por: Geyi Wen, et al.
Publicado: (2024)
Heterogeneous bimodal attention fusion for speech emotion recognition
por: Luo, Jiachen, et al.
Publicado: (2025)
por: Luo, Jiachen, et al.
Publicado: (2025)
Cross-domain Open-world Discovery
por: Wen, Shuo, et al.
Publicado: (2024)
por: Wen, Shuo, et al.
Publicado: (2024)
Cross-user activity recognition using deep domain adaptation with temporal relation information
por: Ye, Xiaozhou, et al.
Publicado: (2024)
por: Ye, Xiaozhou, et al.
Publicado: (2024)
Improving child speech recognition with augmented child-like speech
por: Zhang, Yuanyuan, et al.
Publicado: (2024)
por: Zhang, Yuanyuan, et al.
Publicado: (2024)
Acoustic and linguistic effects in synthesized speech augmentation for speech recognition
por: Yohan Lim, et al.
Publicado: (2025)
por: Yohan Lim, et al.
Publicado: (2025)
SeMaScore : a new evaluation metric for automatic speech recognition tasks
por: Sasindran, Zitha, et al.
Publicado: (2024)
por: Sasindran, Zitha, et al.
Publicado: (2024)
cantnlp@DravidianLangTech 2026: organic domain adaptation improves multi-class hope speech detection in Tulu
por: Li, Andrew, et al.
Publicado: (2026)
por: Li, Andrew, et al.
Publicado: (2026)
PROCTER: PROnunciation-aware ConTextual adaptER for personalized speech recognition in neural transducers
por: Pandey, Rahul, et al.
Publicado: (2023)
por: Pandey, Rahul, et al.
Publicado: (2023)
Glucagon‐like peptide‐1 receptor agonists and upper endoscopy: a real‐world experience
por: Pichamol Jirapinyo, et al.
Publicado: (2025)
por: Pichamol Jirapinyo, et al.
Publicado: (2025)
Low-resource speech recognition and dialect identification of Irish in a multi-task framework
por: Lonergan, Liam, et al.
Publicado: (2024)
por: Lonergan, Liam, et al.
Publicado: (2024)
Gen-SER: When the generative model meets speech emotion recognition
por: Wang, Taihui, et al.
Publicado: (2026)
por: Wang, Taihui, et al.
Publicado: (2026)
Comparing L2 intelligibility for learners of French: Automatic speech recognition versus human listeners
por: Elena Shimanskaya
Publicado: (2025)
por: Elena Shimanskaya
Publicado: (2025)
Exploring the potential: New chapter in gastrointestinal endoscopy with innovative 3D imaging technology
por: Xiaoqing Lin, et al.
Publicado: (2024)
por: Xiaoqing Lin, et al.
Publicado: (2024)
Ejemplares similares
-
EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations
por: Yang, Ruijie, et al.
Publicado: (2025) -
EndoFinder: Online Image Retrieval for Explainable Colorectal Polyp Diagnosis
por: Yang, Ruijie, et al.
Publicado: (2024) -
One-shot synthesis of rare gastrointestinal lesions improves diagnostic accuracy and clinical training
por: Yu, Jia, et al.
Publicado: (2025) -
Endo-CLIP: Progressive Self-Supervised Pre-training on Raw Colonoscopy Records
por: He, Yili, et al.
Publicado: (2025) -
Robust Polyp Detection and Diagnosis through Compositional Prompt-Guided Diffusion Models
por: Yu, Jia, et al.
Publicado: (2025)