A pre-trained BERT for Korean medical natural language processing

Kim, Yoojoong; Kim, Jong-Ho; Lee, Jeong Moon; Jang, Moon Joung; Yum, Yun Jin; Kim, Seongtae; Shin, Unsub; Kim, Young-Min; Joo, Hyung Joon; Song, Sanghoun

doi:10.1038/s41598-022-17806-8

Detailed Information

Cited 0 time in webofscience

Cited 0 time in scopus

Metadata Downloads

A pre-trained BERT for Korean medical natural language processing

Full metadata record

DC Field	Value	Language
dc.contributor.author	Kim, Yoojoong	-
dc.contributor.author	Kim, Jong-Ho	-
dc.contributor.author	Lee, Jeong Moon	-
dc.contributor.author	Jang, Moon Joung	-
dc.contributor.author	Yum, Yun Jin	-
dc.contributor.author	Kim, Seongtae	-
dc.contributor.author	Shin, Unsub	-
dc.contributor.author	Kim, Young-Min	-
dc.contributor.author	Joo, Hyung Joon	-
dc.contributor.author	Song, Sanghoun	-
dc.date.accessioned	2022-09-19T12:14:25Z	-
dc.date.available	2022-09-19T12:14:25Z	-
dc.date.created	2022-09-08	-
dc.date.issued	2022-08	-
dc.identifier.issn	2045-2322	-
dc.identifier.uri	https://scholarworks.bwise.kr/hanyang/handle/2021.sw.hanyang/171527	-
dc.description.abstract	With advances in deep learning and natural language processing (NLP), the analysis of medical texts is becoming increasingly important. Nonetheless, despite the importance of processing medical texts, no research on Korean medical-specific language models has been conducted. The Korean medical text is highly difficult to analyze because of the agglutinative characteristics of the language, as well as the complex terminologies in the medical domain. To solve this problem, we collected a Korean medical corpus and used it to train the language models. In this paper, we present a Korean medical language model based on deep learning NLP. The model was trained using the pre-training framework of BERT for the medical context based on a state-of-the-art Korean language model. The pre-trained model showed increased accuracies of 0.147 and 0.148 for the masked language model with next sentence prediction. In the intrinsic evaluation, the next sentence prediction accuracy improved by 0.258, which is a remarkable enhancement. In addition, the extrinsic evaluation of Korean medical semantic textual similarity data showed a 0.046 increase in the Pearson correlation, and the evaluation for the Korean medical named entity recognition showed a 0.053 increase in the F1-score.	-
dc.language	영어	-
dc.language.iso	en	-
dc.publisher	NATURE PORTFOLIO	-
dc.title	A pre-trained BERT for Korean medical natural language processing	-
dc.type	Article	-
dc.contributor.affiliatedAuthor	Kim, Young-Min	-
dc.identifier.doi	10.1038/s41598-022-17806-8	-
dc.identifier.scopusid	2-s2.0-85135987936	-
dc.identifier.wosid	000841397200059	-
dc.identifier.bibliographicCitation	SCIENTIFIC REPORTS, v.12, no.1, pp.1 - 10	-
dc.relation.isPartOf	SCIENTIFIC REPORTS	-
dc.citation.title	SCIENTIFIC REPORTS	-
dc.citation.volume	12	-
dc.citation.number	1	-
dc.citation.startPage	1	-
dc.citation.endPage	10	-
dc.type.rims	ART	-
dc.type.docType	Article	-
dc.description.journalClass	1	-
dc.description.isOpenAccess	Y	-
dc.description.journalRegisteredClass	scie	-
dc.description.journalRegisteredClass	scopus	-
dc.relation.journalResearchArea	Science & Technology - Other Topics	-
dc.relation.journalWebOfScienceCategory	Multidisciplinary Sciences	-
dc.subject.keywordPlus	article	-
dc.subject.keywordPlus	deep learning	-
dc.subject.keywordPlus	human	-
dc.subject.keywordPlus	human experiment	-
dc.subject.keywordPlus	natural language processing	-
dc.subject.keywordPlus	prediction	-
dc.subject.keywordPlus	language	-
dc.subject.keywordPlus	semantics	-
dc.subject.keywordPlus	South Korea	-
dc.identifier.url	https://www.nature.com/articles/s41598-022-17806-8	-

Files in This Item: Go to Link

Appears in Collections: 서울 기술경영전문대학원 > 서울 기술경영학과 > 1. Journal Articles

Show simple item record

qrcode

Related Researcher

Researcher Kim, Young min photo

Kim, Young min: GRADUATE SCHOOL OF TECHNOLOGY & INNOVATION MANAGEMENT (DEPARTMENT OF TECHNOLOGY MANAGEMENT)

Read more

Altmetrics

Total Views & Downloads

STATISTICS: Total View :5,980,979; Today View :6,952

RSS_1.0 RSS_2.0 ATOM_1.0

222, Wangsimni-ro, Seongdong-gu, Seoul, 04763, Korea+82-2-2220-1365

Certain data included herein are derived from the © Web of Science of Clarivate Analytics. All rights reserved.
You may not copy or re-distribute this material in whole or in part without the prior written consent of Clarivate Analytics.

Detailed Information

Related Researcher

Altmetrics

Total Views & Downloads

BROWSE