상세 보기
초록
Today, we are exposed to various text-based media such as newspapers, Internet articles, and SNS, and the amount of text data we encounter has increased exponentially due to the recent availability of Internet access using mobile devices such as smartphones. Collecting useful information from a lot of text information is called text analysis, and in order to extract information, it is performed using technologies such as Natural Language Processing (NLP) for processing natural language with the recent development of artificial intelligence. For this purpose, a morpheme analyzer based on everyday language has been disclosed and is being used. Pre-learning language models, which can acquire natural language knowledge through unsupervised learning based on large numbers of corpus, are a very common factor in natural language processing recently, but conventional morpheme analysts are limited in their use in specialized fields. In this paper, as a preliminary work to develop a natural language analysis language model specialized in the railway field, the procedure for construction a corpus specialized in the railway field is presented.
키워드
- 제목
- On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain
- 제목 (타언어)
- On the Analysis of Natural Language Processing Morphology for the Specialized Corpus in the Railway Domain
- 저자
- 원종운; 전홍규; 김민중; 김백현; 김영민
- 발행일
- 2022-11
- 권
- 14
- 호
- 4
- 페이지
- 189 ~ 197