Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Development and testing of a multi-lingual Natural Language Processing-based deep learning system in 10 languages for COVID-19 pandemic crisis: A multi-center study

作者:Lily Wei Yun Yang, Wei Yan Ng, Xiaofeng Lei, Shaun Chern Yuan Tan, Zhaoran Wang, Ming Yan, Mohan Kashyap Pargi, Xiaoman Zhang, Jane Lim, Dinesh Visva Gunasekeran, Franklin Chee Ping Tan, Chen Ee Lee, Khung Keong Yeo, Hiang Khoon Tan, Henry Sun Sien Ho, Benedict Tan, Tien Yin Wong, Kenneth Kwek, Rick Siow Mong Goh, Yong Liu, Daniel Shu Wei Ting · 发表于:Frontiers in Public Health · 年份:2023 · DOI:10.3389/fpubh.2023.1063466 · 被引用次数:36 · 研究领域:AI in Service Interactions、Artificial Intelligence in Healthcare and Education、Misinformation and Its Impacts

Purpose: The COVID-19 pandemic has drastically disrupted global healthcare systems. With the higher demand for healthcare and misinformation related to COVID-19, there is a need to explore alternative models to improve communication. Artificial Intelligence (AI) and Natural Language Processing (NLP) have emerged as promising solutions to improve healthcare delivery. Chatbots could fill a pivotal role in the dissemination and easy accessibility of accurate information in a pandemic. In this study, we developed a multi-lingual NLP-based AI chatbot, DR-COVID, which responds accurately to open-ended, COVID-19 related questions. This was used to facilitate pandemic education and healthcare delivery. Methods: First, we developed DR-COVID with an ensemble NLP model on the Telegram platform (https://t.me/drcovid_nlp_chatbot). Second, we evaluated various performance metrics. Third, we evaluated multi-lingual text-to-text translation to Chinese, Malay, Tamil, Filipino, Thai, Japanese, French, Spanish, and Portuguese. We utilized 2,728 training questions and 821 test questions in English. Primary outcome measurements were (A) overall and top 3 accuracies; (B) Area Under the Curve (AUC), precision, recall, and F1 score. Overall accuracy referred to a correct response for the top answer, whereas top 3 accuracy referred to an appropriate response for any one answer amongst the top 3 answers. AUC and its relevant matrices were obtained from the Receiver Operation Characteristics (ROC) curv...