On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain

On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain

초록

Today, we are exposed to various text-based media such as newspapers, Internet articles, and SNS, and the amount of text data we encounter has increased exponentially due to the recent availability of Internet access using mobile devices such as smartphones. Collecting useful information from a lot of text information is called text analysis, and in order to extract information, it is performed using technologies such as Natural Language Processing (NLP) for processing natural language with the recent development of artificial intelligence. For this purpose, a morpheme analyzer based on everyday language has been disclosed and is being used. Pre-learning language models, which can acquire natural language knowledge through unsupervised learning based on large numbers of corpus, are a very common factor in natural language processing recently, but conventional morpheme analysts are limited in their use in specialized fields. In this paper, as a preliminary work to develop a natural language analysis language model specialized in the railway field, the procedure for construction a corpus specialized in the railway field is presented.

키워드

Natural Language ProcessingMorphologyCorpusArtificial IntelligenceIntelligent Railway and Transportation TechnologiesNLPAI
제목
On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain
제목 (타언어)
On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain
저자
원종운전홍규김민중김백현김영민
DOI
10.7236/IJIBC.2022.14.4.189
발행일
2022-11
저널명
The International Journal of Internet, Broadcasting and Communication
14
4
페이지
189 ~ 197