기록관리 분야에서 한국어 자연어 처리 기술을 적용하기 위한 고려사항

Considerations for Applying Korean Natural Language Processing Technology in Records Management

초록

Records have temporal characteristics, including the past and present; linguistic characteristics not limited to a specific language; and various types categorized in a complex way. Processing records such as text, video, and audio in the life cycle of records’ creation, preservation, and utilization entails exhaustive effort and cost. Primary natural language processing (NLP) technologies, such as machine translation, document summarization, named-entity recognition, and image recognition, can be widely applied to electronic records and analog digitization. In particular, Korean deep learning–based NLP technologies effectively recognize various record types and generate record management metadata. This paper provides an overview of Korean NLP technologies and discusses considerations for applying NLP technology in records management. The process of using NLP technologies, such as machine translation and optical character recognition for digital conversion of records, is introduced as an example implemented in the Python environment. In contrast, a plan to improve environmental factors and record digitization guidelines for applying NLP technology in the records management field is proposed for utilizing NLP technology.

키워드

기록관리자연어 처리인공지능머신러닝딥러닝Records managementNatural language processingArtificial intelligenceMachine learningDeep learning
제목
기록관리 분야에서 한국어 자연어 처리 기술을 적용하기 위한 고려사항
제목 (타언어)
Considerations for Applying Korean Natural Language Processing Technology in Records Management
저자
김학래
DOI
10.14404/JKSARM.2022.22.4.129
발행일
2022-11
저널명
한국기록관리학회지
22
4
페이지
129 ~ 149