Skip to main content

유연주 교수

Yeunju Yoo

서울대학교 수학교육과 · 생화학·유전·분자생물학

연구실 소개

유연주 교수의 연구실은 유전체역학과 인구유전학 분야에서 고밀도 유전자 데이터의 구조적 분석과 유전자 연관성 연구를 중심으로 활동하고 있습니다. 특히 연관성 분석의 통계적 검정력 향상과 복합유전자형 분석을 위한 블록 구조화, 링크 disequilibrium(LD) 분석 기법 개발에 초점을 맞추고 있으며, 대규모 유전자 데이터를 효율적으로 처리할 수 있는 소프트웨어 도구 개발도 함께 진행하고 있습니다. 연구는 주로 고밀도 SNP 데이터 기반의 유전자 블록 분할, 허브타입 블록 추정, 다유전자 연속 분석 기법 개발을 포함합니다.

유전자 블록 분할링크 disequilibrium다유전자 연관 분석유전체 연속 분석고밀도 SNP 데이터

연구 현황

논문 수
87
총 인용 수
952
최근 5년 논문
31
주요 분야
생화학·유전·분자생물학

연구 성과 추이

표시된 성과는 수집된 데이터 기준으로 산출되며, 일부 차이가 있을 수 있습니다.

5개년 연도별 논문 게재 수
31총합
2022
2023
2024
2025
2026
5개년 연도별 피인용 수
41총합
20222023202420252026

주요 논문

15
1
논문|인용수 79·2017
A new haplotype block detection method for dense genome sequencing data based on interval graph modeling of clusters of highly correlated SNPs
Sun Ah Kim, Chang-Sung Cho, Suh-Ryung Kim, Shelley B. Bull, Yun Joo Yoo
SJR Q1BioinformaticsOA

Motivation: Linkage disequilibrium (LD) block construction is required for research in population genetics and genetic epidemiology, including specification of sets of single nucleotide polymorphisms (SNPs) for analysis of multi-SNP based association and identification of haplotype blocks in high density sequencing data. Existing methods based on a narrow sense definition do not allow intermediate regions of low LD between strongly associated SNP pairs and tend to split high density SNP data int

GeneticsBiochemistry, Genetics and Molecular Biology
2
논문|인용수 49·2019
gpart: human genome partitioning and visualization of high-density SNP data by identifying haplotype blocks
Sunah Kim, Myriam Brossard, Delnaz Roshandel, Andrew D. Paterson, Shelley B. Bull, Yun Joo Yoo
SJR Q1BioinformaticsOA

SUMMARY: For the analysis of high-throughput genomic data produced by next-generation sequencing (NGS) technologies, researchers need to identify linkage disequilibrium (LD) structure in the genome. In this work, we developed an R package gpart which provides clustering algorithms to define LD blocks or analysis units consisting of SNPs. The visualization tool in gpart can display the LD structure and gene positions for up to 20 000 SNPs in one image. The gpart functions facilitate construction

GeneticsBiochemistry, Genetics and Molecular Biology
3
논문|인용수 36·1989
Synthesis of Peptides as Cloned Ubiquitin Extensions
Yun Joo Yoo, K V Rote, Martin Rechsteiner
SJR Q1Journal of Biological ChemistryOA
Molecular BiologyBiochemistry, Genetics and Molecular Biology
4
논문|인용수 26·2007
Haplotype inference for present–absent genotype data using previously identified haplotypes and haplotype patterns
Yun Joo Yoo, Jianming Tang, Richard A. Kaslow, Kui Zhang
SJR Q1BioinformaticsOA

MOTIVATION: Killer immunoglobulin-like receptor (KIR) genes vary considerably in their presence or absence on a specific regional haplotype. Because presence or absence of these genes is largely detected using locus-specific genotyping technology, the distinction between homozygosity and hemizygosity is often ambiguous. The performance of methods for haplotype inference (e.g. PL-EM, PHASE) for KIR genes may be compromised due to the large portion of ambiguous data. At the same time, many haploty

GeneticsBiochemistry, Genetics and Molecular Biology
5
논문|인용수 26·2009
Were genome‐wide linkage studies a waste of time? Exploiting candidate regions within genome‐wide association studies
Yun Joo Yoo, Shelley B. Bull, Andrew D. Paterson, Daryl Waggott, Lei Sun
SJR Q2Genetic EpidemiologyOA

A central issue in genome-wide association (GWA) studies is assessing statistical significance while adjusting for multiple hypothesis testing. An equally important question is the statistical efficiency of the GWA design as compared to the traditional sequential approach in which genome-wide linkage analysis is followed by region-wise association mapping. Nevertheless, GWA is becoming more popular due in part to cost efficiency: commercially available 1M chips are nearly as inexpensive as a cus

GeneticsBiochemistry, Genetics and Molecular Biology
6
논문|인용수 20·2016
Multiple linear combination (MLC) regression tests for common variants adapted to linkage disequilibrium structure
Yun Joo Yoo, Lei Sun, Julia G. Poirier, Andrew D. Paterson, Shelley B. Bull
SJR Q2Genetic EpidemiologyOA

By jointly analyzing multiple variants within a gene, instead of one at a time, gene-based multiple regression can improve power, robustness, and interpretation in genetic association analysis. We investigate multiple linear combination (MLC) test statistics for analysis of common variants under realistic trait models with linkage disequilibrium (LD) based on HapMap Asian haplotypes. MLC is a directional test that exploits LD structure in a gene to construct clusters of closely correlated varian

GeneticsBiochemistry, Genetics and Molecular Biology
7
논문|인용수 15·2007
Case-control association analysis of rheumatoid arthritis with candidate genes using related cases
Yun Joo Yoo, Guimin Gao, Kui Zhang
SJR Q2BMC ProceedingsOA

We performed a case-control association analysis of rheumatoid arthritis (RA) for several candidate genes using the North American Rheumatoid Arthritis Consortium (NARAC) data provided in Genetic Analysis Workshop 15. We conducted the case-control association analysis using all related cases and unrelated controls and compared the results with those from the analysis of samples using only one randomly selected case from each family and all unrelated controls. For both analyses we used a weighted

RheumatologyMedicine
8
논문|인용수 13·2009
Genome-wide association analyses of North American Rheumatoid Arthritis Consortium and Framingham Heart Study data utilizing genome-wide linkage results
Yun Joo Yoo, Dushanthi Pinnaduwage, Daryl Waggott, Shelley B. Bull, Lei Sun
SJR Q2BMC ProceedingsOA

The power of genome-wide association studies can be improved by incorporating information from previous study findings, for example, results of genome-wide linkage analyses. Weighted false-discovery rate (FDR) control can incorporate genome-wide linkage scan results into the analysis of genome-wide association data by assigning single-nucleotide polymorphism (SNP) specific weights. Stratified FDR control can also be applied by stratifying the SNPs into high and low linkage strata. We applied the

GeneticsMedicine
9
논문|인용수 12·2015
Clique-Based Clustering of Correlated SNPs in a Gene Can Improve Performance of Gene-Based Multi-Bin Linear Combination Test
Yun Joo Yoo, Sunah Kim, Shelley B. Bull
SJR Q2BioMed Research InternationalOA

Gene-based analysis of multiple single nucleotide polymorphisms (SNPs) in a gene region is an alternative to single SNP analysis. The multi-bin linear combination test (MLC) proposed in previous studies utilizes the correlation among SNPs within a gene to construct a gene-based global test. SNPs are partitioned into clusters of highly correlated SNPs, and the MLC test statistic quadratically combines linear combination statistics constructed for each cluster. The test has degrees of freedom equa

Molecular BiologyBiochemistry, Genetics and Molecular Biology
10
논문|인용수 11·2013
Gene-based multiple regression association testing for combined examination of common and low frequency variants in quantitative trait analysis
Yun Joo Yoo, Lei Sun, Shelley B. Bull
SJR Q2Frontiers in GeneticsOA

Multi-marker methods for genetic association analysis can be performed for common and low frequency SNPs to improve power. Regression models are an intuitive way to formulate multi-marker tests. In previous studies we evaluated regression-based multi-marker tests for common SNPs, and through identification of bins consisting of correlated SNPs, developed a multi-bin linear combination (MLC) test that is a compromise between a 1 df linear combination test and a multi-df global test. Bins of SNPs

GeneticsBiochemistry, Genetics and Molecular Biology
11
논문|인용수 10·2024
Can we Use GPT-4 as a Mathematics Evaluator in Education?: Exploring the Efficacy and Limitation of LLM-based Automatic Assessment System for Open-ended Mathematics Question
Unggi Lee, Youngbean Kim, Sangyun Lee, Jaehyeon Park, Jin Mun, Eunseo Lee, Hyeoncheol Kim, Cheolil Lim, Yun Joo Yoo
SJR Q1International Journal of Artificial Intelligence in Education
Artificial IntelligenceComputer Science
12
논문|인용수 9·2020
코로나바이러스감염증-19 유행 상황 속 마스크 관련 소비자불안
유연주, 여정성

2020년 1월 20일, 국내에서 코로나바이러스감염증-19의 첫 확진자가 발생하였다. 신종 감염병의 특성은 사회적으로 불안을 형성하였고, 필수 소비 품목으로 자리 잡은 마스크에 관한 각종 소비자 문제는 소비자불안을 촉발하는 계기가 되었다. 본 연구에서는 코로나19가 장기적인 영향을 미치는 가운데, 시기 및 상황의 변화에 따라 마스크에 관하여 소비자들이 어떠한 불안을 느끼는지 파악하고 이를 줄이기 위한 적절한 위험 커뮤니케이션 전략을 도출하고자 하였다. 이를 위해 트위터의 텍스트 데이터를 분석하였으며, 토픽모델링을 통해 시기별 주요 토픽을 추출하였다. 분석 결과, 마스크 관련 소비자불안에 대한 트윗의 버즈량 및 주요 키워드는 시기별로 차이가 있었다. 특히 집단감염이 발생하며 이른바 마스크 대란이 발생하였던 시기에 불안 관련 버즈량이 급증하여, 코로나19 상황에서 소비자불안에 가장 큰 영향을 준 요인은 마스크의 수급문제였음을 확인하였다. 또한, 토픽모델링을 실시하여 시기별 토픽과 그 내

13
논문|인용수 8·2020
소비자의 상품큐레이션서비스 이용에 관한 탐색적 연구
유연주, 여정성

상품큐레이션서비스의 등장 이래 시장의 성장세에도 불구하고 소비자들의 현실적인 목소리를 들어볼 수 있는 연구는 거의 이루어지지 않았다. 이에 본 연구에서는 기존 문헌을 토대로 상품큐레이션서비스를 ‘큐레이터가 선별하여구성한 상품을 소비자에게 제공하는 전자상거래 서비스’로 정의하고, 누가, 왜, 어떻게 서비스를 이용하는지 포괄적으로 탐색하고자 하였다. 이를 위해 심층면접을 실시하였으며 Glaser의 근거이론 방법론에 따라 자료를 분석한 결과 83개의 개념, 22개의 하위범주, 9개의 범주 및 ‘정보 과잉 환경에서 상품큐레이션서비스 이용을 자신에게 맞추어 나감’이라는 핵심범주가 도출되었다. 연구 결과, 상품큐레이션서비스 이용자들은 새로움과 효율성 추구 성향을 가지며 구매하는 상품에 대한 관여도가높았고, 정보 탐색에 대한 부담과 직접 선택에서의 한계와 더불어 서비스에 대한 다양한 기대로 서비스를 이용하는것으로 나타났다. 또한 이용자들은 서비스 이용 시 다양한 혜택과 문제를 지각했으며 다수는

14
논문|인용수 7·2008
The Power and Robustness of Maximum LOD Score Statistics
Yun Joo Yoo, Nancy R. Mendell
SJR Q3Annals of Human GeneticsOA

The maximum LOD score statistic is extremely powerful for gene mapping when calculated using the correct genetic parameter value. When the mode of genetic transmission is unknown, the maximum of the LOD scores obtained using several genetic parameter values is reported. This latter statistic requires higher critical value than the maximum LOD score statistic calculated from a single genetic parameter value. In this paper, we compare the power of maximum LOD scores based on three fixed sets of ge

GeneticsBiochemistry, Genetics and Molecular Biology
15
논문|인용수 7·2024
Supervised diagnostic classification of cognitive attributes using data augmentation
Ji-Young Yoon, Gahgene Gweon, Yun Joo Yoo
SJR Q1PLoS ONEOA

Over recent decades, machine learning, an integral subfield of artificial intelligence, has revolutionized diverse sectors, enabling data-driven decisions with minimal human intervention. In particular, the field of educational assessment emerges as a promising area for machine learning applications, where students can be classified and diagnosed using their performance data. The objectives of Diagnostic Classification Models (DCMs), which provide a suite of methods for diagnosing students' cogn

Artificial IntelligenceComputer Science

대표 연구 분야

GeneticsMolecular BiologyArtificial IntelligenceStatistics and ProbabilityComputer Networks and CommunicationsInformation Systems

유연주 교수의 연구를 Nubint에서 더 깊이 살펴보세요

이 연구실의 논문을 앱에서 열어 AI와 함께 읽고, 핵심을 요약하고, 내 글에 인용하세요.