유환조 교수
Hwanjo Yu
포항공과대학교 컴퓨터공학과 · 컴퓨터과학
연구실 소개
유환조 교수의 연구실은 웹 마이닝과 텍스트 분류 분야에서 핵심적인 기반 기술을 연구하고 있습니다. 특히, 부족한 레이블 데이터 환경에서 효과적으로 분류 모델을 학습하는 데 중점을 두며, 긍정 예시 기반 학습(PEBL), 음성 데이터가 없는 텍스트 분류(TC-WON), 그리고 SVM 기반의 스케일러블 학습 기법 등 고도화된 분류 및 랭킹 학습 기법을 개발하고 있습니다. 연구는 실세계의 대량 데이터 환경에서 효율적이고 일반화 능력이 뛰어난 지도학습 및 준지도학습 기법의 설계에 초점을 맞추고 있습니다.
연구 현황
연구 성과 추이
표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.
주요 논문
15Web page classification is one of the essential techniques for Web mining. Specifically, classifying Web pages of a user-interesting class is the first step of mining interesting information from the Web. However, constructing a classifier for an interesting class requires laborious pre-processing such as collecting positive and negative training examples. For instance, in order to construct a homepage classifier, one needs to collect a sample of homepages (positive examples) and a sample of non
Web page classification is one of the essential techniques for Web mining because classifying Web pages of an interesting class is often the first step of mining the Web. However, constructing a classifier for an interesting class requires laborious preprocessing such as collecting positive and negative training examples. For instance, in order to construct a "homepage" classifier, one needs to collect a sample of homepages (positive examples) and a sample of nonhomepages (negative examples). In
Learning ranking (or preference) functions has been a major issue in the machine learning community and has produced many applications in information retrieval. SVMs (Support Vector Machines) - a classification and regression methodology - have also shown excellent performance in learning ranking functions. They effectively learn ranking functions of high generalization based on the large-margin principle and also systematically support nonlinear ranking by the kernel trick. In this paper, we pr
Support vector machines (SVMs) have been promising methods for classification and regression analysis because of their solid mathematical foundations which convery several salient properties that other methods hardly provide. However, despite the prominent properties of SVMs, they are not as favored for large-scale data mining as for pattern recognition or machine learning because the training complexity of SVMs is highly dependent on the size of a data set. Many real-world data mining applicati
Active sampling (also called active learning or selective sampling) has been extensively researched for classification and rank learning methods, which is to select the most informative samples from unlabeled data such that, once the samples are labeled, the accuracy of the function learned from the samples is maximized. While active sampling methods require learning a function at each iteration to find the most informative samples, this paper proposes passive sampling techniques for regression,
Most existing studies of text classification assume that the training data are completely labeled. In reality, however, many information retrieval problems can be more accurately described as learning a binary classifier from a set of incompletely labeled examples, where we typically have a small number of labeled positive examples and a very large number of unlabeled examples. In this paper, we study such a problem of performing Text Classification WithOut labeled Negative data TC-WON). In this
Adoption of Electronic Health Record (EHR) systems has led to collection of massive healthcare data, which creates oppor- tunities and challenges to study them. Computational phenotyping offers a promising way to convert the sparse and complex data into meaningful concepts that are interpretable to healthcare givers to make use of them. We propose a novel su- pervised nonnegative tensor factorization methodology that derives discriminative and distinct phenotypes. We represented co-occurrence of
The demonstration of diminished or scarred renal parenchyma in children is often the decisive factor in determining the future management of children with urinary tract malformations. Renal scintigraphy using technetium 99m-labelled dimercaptosuccinic acid (DMSA), computed tomography (CT) and intravenous urography (IU) were used to evaluate the renal parenchyma prior to ureter re-implantation in a series of 13 children. Their ages ranged from 5 months to 3 years 8 months. The indication for oper
대표 연구 분야
유환조 교수의 연구를 Nubint에서 더 깊이 살펴보세요
이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.