정서현 교수
Seo-Hyun Jeong
KAIST 전산학부 · 컴퓨터과학
연구실 소개
정서현 교수의 연구실은 대규모 이미지 생성 모델이 유발할 수 있는 윤리적·법적 문제, 특히 불법 콘텐츠나 저작권 침해, 편향 문제를 해결하기 위한 기반 기술을 연구하고 있습니다. 특히 텍스트-이미지 디퓨전 모델에서 유해 콘텐츠 생성을 사전 차단하기 위한 자기학습 기반의 안전성 강화 기법(SDD), 인간 피드백을 통합한 지식 정렬 프레임워크(HFI) 등 인간의 가치관과 모델 지식을 일치시키는 혁신적 접근을 개발하고 있습니다. 또한 베이지안 신경망의 복잡한 사후 분포 탐색을 위한 메타학습 기반의 효율적 추론 기법 등 고차원 모델의 신뢰성과 안정성 향상 기술도 함께 연구하고 있습니다.
연구 현황
연구 성과 추이
표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.
주요 논문
5Large-scale image generation models, with impressive quality made possible by the vast amount of data available on the Internet, raise social concerns that these models may generate harmful or copyrighted content. The biases and harmfulness arise throughout the entire training process and are hard to completely remove, which have become significant hurdles to the safe deployment of these models. In this paper, we propose a method called SDD to prevent problematic content generation in text-to-im
This study analyzed evaluator competency trends and standards to provide guidance for the evaluation society in Korea, as the evaluation of development programs are becoming increasingly important. The study compared and analyzed evaluator competencies that are established by both evaluation associations in the Americas, Canada, Australia, and South Africa, and the development cooperation organizations such as the United Nations. Furthermore, the study analyzed the extent to which the evaluation
Bayesian Neural Networks(BNNs) with high-dimensional parameters pose a challenge for posterior inference due to the multi-modality of the posterior distributions. Stochastic Gradient MCMC(SGMCMC) with cyclical learning rate scheduling is a promising solution, but it requires a large number of sampling steps to explore high-dimensional multi-modal posteriors, making it computationally expensive. In this paper, we propose a meta-learning strategy to build \gls{sgmcmc} which can efficiently explore
This paper addresses the societal concerns arising from large-scale text-to-image diffusion models for generating potentially harmful or copyrighted content. Existing models rely heavily on internet-crawled data, wherein problematic concepts persist due to incomplete filtration processes. While previous approaches somewhat alleviate the issue, they often rely on text-specified concepts, introducing challenges in accurately capturing nuanced concepts and aligning model knowledge with human unders
대표 연구 분야
정서현 교수의 연구를 Nubint에서 더 깊이 살펴보세요
이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.