Skip to main content

Jeongsun Jang

Korea University · Computer Science

About the Lab

Professor Jeongsun Jang's research lab specializes in applying advanced natural language processing and large language models to historical and cultural texts, with a focus on Korean historical corpora such as the Annals of the Joseon Dynasty and Seungjeongwon Ilgi. The lab pioneers AI-driven approaches for named entity recognition, relation extraction, and semantic search in classical East Asian texts, emphasizing context-aware, low-resource, and domain-specific language modeling. Research directions include semantic retrieval for historical documents, reconstruction of damaged inscriptions using NLP, and the interpretation of linguistic evolution in classical vocabulary.

historical NLPsemantic searchnamed entity recognitionclassical Korean textsAI for humanities

Research Overview

Papers
9
Total Citations
21
Papers (5y)
9
Primary Field
Computer Science

Research Output Trend

Figures are computed from collected data and may differ slightly.

Publications per year (5y)
9total
2016
2019
2024
2025
2026
Citations per year (5y)
21total
20162019202420252026

Selected Papers

9
1
Article|15 citations·2016
A novel density-based clustering method using word embedding features for dialogue intention recognition
Jungsun Jang, Yeonsoo Lee, Seolhwa Lee, Dongwon Shin, Dongjun Kim, Hae‐Chang Rim
SJR Q1Cluster Computing
Artificial IntelligenceComputer Science
2
Article|3 citations·2019
Narrative context-based data-to-text generation for ambient intelligence
Jungsun Jang, Hyungjong Noh, Yeonsoo Lee, Soo-Min Pantel, Hae‐Chang Rim
SJR Q1Journal of Ambient Intelligence and Humanized Computing
Artificial IntelligenceComputer Science
3
Article|1 citations·2026
Evaluating over-empathizing in emotional support conversations: A user-centered framework
Suhyune Son, Seonmin Koo, Evelyn Hayoon Zi, Jungsun Jang, Heuiseok Lim
SJR Q1Expert Systems with Applications
Social PsychologyPsychology
4
Article|1 citations·2024
Analysis of the Effectiveness of Model, Data, and User-Centric Approaches for Chat Application: A Case Study of BlenderBot 2.0
Chanjun Park, Jungseob Lee, Suhyune Son, Kinam Park, Jungsun Jang, Heuiseok Lim
SJR Q2Applied SciencesOA

BlenderBot 2.0 represents a significant advancement in open-domain chatbots by incorporating real-time information and retaining user information across multiple sessions through an internet search module. Despite its innovations, there are still areas for improvement. This paper examines BlenderBot 2.0’s limitations and errors from three perspectives: model, data, and user interaction. From the data perspective, we highlight the challenges associated with the crowdsourcing process, including un

Artificial IntelligenceComputer Science
5
Article|1 citations·2024
Benchmark Study for Evaluating the Korean Comprehension of Large Language Models
Minju Seo, YeonJoo Jeong, Hyunjeong Lee, Hee Yeon Im, Jungsun Jang
The Journal of Korean Association of Computer EducationOA
Artificial IntelligenceComputer Science
6
Article|0 citations·2026
조선시대 한문 사료를 위한의미 기반 검색 모델의 개발과 활용―『조선왕조실록』과 『경국대전』을 중심으로―
정연주, 장정선

조선시대 사료는 대부분 한문으로 이루어져 있으며, 그중 󰡔조선왕조실록󰡕은 조선 전시대의 다양한 내용을 담고 있어 역사 연구의 기본이 된다. 또한 󰡔경국대전󰡕은 조선 초기 제도를 집대성하고 통치 이념을 담고 있는 서적인데, 조문이 간략하고 함축되어 있어 쉽게 이해하기 어렵다. 이를 쉽게 이해하기 위해 󰡔경국대전󰡕 조문의 입법 의도를 실록 속에서 찾아 검토할 필요가 있다. 그러나 󰡔경국대전󰡕 조문의 내용은 실록에서 같은 표현으로 나타나지 않는 경우가 많아 현재 실록 웹사이트의 키워드 검색만으로는 안정적으로 수행하기 어렵다. 따라서 󰡔경국대전󰡕 조문을 질의(query)로 입력하면 문자열의 일치 여부와 상관없이 관련 실록 기사를 검색할 수 있는 의미 기반 검색 모델을 개발하였다. 사고전서를 학습한 SikuRoBERTa 모델을 베이스로 하여 실록과 󰡔승정원일기󰡕로 MLM 학습 후 조선시대 문체와 어휘를 반영하도록 하였다. 이어 비지도 SimCSE와 약지도 대조학습을 통해 문장/청크 임베딩을

7
Article|0 citations·2026
사전학습 언어모델(Pre-trained Language Model) 기반의 고려시대 묘지명에 대한 결락 문자 추정 연구
이현정, 장정선

고려시대 묘지명은 당대의 생활상을 생생하게 보여주는 귀중한 사료이나, 상당수가 마모 및 파손으로 인한 결락을 포함하고 있다. 기존의 결락 복원은 연구자의 지식에 의존한 수작업 교차검증에 국한되어, 객관적 지표 마련과 방대한 데이터 처리에 한계가 있었다. 최근 해외에서는 Ithaca(그리스어), Aeneas(라틴어) 등 AI를 활용한 비문 복원 연구가 활발하나, 표의문자인 한자를 사용하며 독자적인 제도적 배경을 가진 한국 금석문에 대한 인공지능 활용 연구는 미진한 상태이다. 본 연구는 인공지능 모델을 통한 고려 금석문 결락 추정의 가능성을 타진하는 시론적 실험을 목적으로 한다. 본 연구에서는 한문 특화 언어 모델인 SIKU-RoBERTa를 기반으로, 고려시대 인물인 원선지(元善之)의 묘지명과 󰡔고려사󰡕 열전 기록을 실험 대상으로 삼았다. 그리고 원선지 열전의 학습 여부 및 LoRA(Low-Rank Adaptation) 어댑팅 기법 적용 여부에 따라 총 4가지 변주 모델을 구축하여 원문

8
Article|0 citations·2026
개체명 인식과 관계 추출을 활용한『승정원일기』 관인-관직 데이터베이스 구축 방안 연구
서민주, 장정선

본 연구는 조선후기 관료 제도 및 인사 제도를 정밀하게 분석할 수 있는 기초자료를 구축하기 위해 언어모델 기반 개체명 인식(Named Entity Recognition, NER)과 관계 추출(Relation Extraction, RE) 기법을 적용하여 『승정원일기』에 수록된 방대한 관인(官人) 및 관직(官職) 정보를 체계적으로 구조화하는 방법론을 제시한다. 먼저 기존에 구축된 『일성록』의 개체명 정보를 학습 데이터로 활용하여 사전학습 언어모델(RoBERTa)을 파인튜닝(fine-tuning)하였으며, 『승정원일기』 내 인물명과 관직명을 자동 식별하였다. 이후 인물명과 관직명이 동시에 등장하는 문단을 대상으로 관계 추출을 진행하여 관원이 특정 관직과 맺는 인사 관계를 의미 단위로 재구성하였다. 관계 유형은 근무, 재직, 임명, 상환, 전직, 추증, 사망의 7가지로 유형화하였으며 각 관계는 문장 구조와 핵심 어휘에 기반한 규칙을 통해 추출되었다. 인물-관직-인사 행위를 중심으로 한 관인

9
Article|0 citations·2025
‘백수(白手)’의 언어문화사 시고
박주원, 오민석, 김혜령, 도원영, 장정선

본고는 어휘의 의미를 보다 정확하게 기술하기 위해 언어문화사적 측면에서 접근할 필요가 있다는 전제 아래, 한자어 ‘백수(白手)’의 변화 양상을 추적하고 특징을 밝히고자 하였다. ‘백수’는 ‘맨손’과 ‘일하지 않는 사람’ 두 가지 의미로 기술되어 있는데, 둘 사이의 연관성을 추정하기가 어렵다. 중국어에서 유래한 ‘백수’는 ‘빈손, 맨손’ 의미로 사용되었으나 그 쓰임은 현대로 오면서 거의 없어졌다. ‘비노동, 무직’을 뜻하는 ‘백수’는 20세기 초반부터 간접적 기원이 될 만한 용례가 등장하지만 단독형 ‘백수’의 용법은 1990년대 후반부터 시작된 사회경제적 변화로 인해 그 용례가 급격히 증가하면서 자리를 잡게 되었다. 등장 및 발전 시기가 다른 두 의미가 하나의 어휘에 공존하게 된 배경을 밝히는 데에 국어사 문헌뿐만 아니라 한국 한문 문헌을 확인하였고 어휘 사용을 둘러싼 사회적 맥락까지 살펴 어휘 정보를 더 풍성하게 만들 수 있음을 보여 주었다는 점에서 이 연구가 의의를 지닌다.

Research Areas

Artificial IntelligenceSocial Psychology

Dive deeper into Jeongsun Jang's research on Nubint

Open this lab's papers in the app to read with AI, summarize, and cite in your writing.