Skip to main content

Sung‐Hoon Kim

Korea Advanced Institute of Science and Technology · 情報科学

研究室紹介

Professor Sung-Hoon Kim's research lab specializes in computational science and software engineering, with a focus on improving software quality and reliability through machine learning and data-driven analysis. The lab develops innovative techniques for latent bug detection using change classification and defect prediction models trained on software repository data, while also addressing challenges in simulation setup and free energy calculations for molecular systems. Their work bridges software engineering and computational chemistry, particularly through tools like CHARMM-GUI for ligand modeling and alchemical free energy simulations. The lab emphasizes practical, scalable solutions for both software development and molecular simulation workflows.

software defect predictionchange classificationmolecular simulationfree energy calculationCHARMM-GUI

Research Overview

Papers
405
Total Citations
14,314
Papers (5y)
100
Primary Field
情報科学

Research Output Trend

Figures are computed from collected data and may differ slightly.

Publications per year (5y)
100total
2021
2022
2023
2024
2025
Citations per year (5y)
482total
20212022202320242025

Selected Papers

15
1
Article|656 citations·2008
Classifying Software Changes: Clean or Buggy?
Sunghun Kim, E. James Whitehead, Yi Zhang
SJR Q1IEEE Transactions on Software Engineering

This paper introduces a new technique for finding latent software bugs called change classification. Change classification uses a machine learning classifier to determine whether a new software change is more similar to prior buggy changes, or clean changes. In this manner, change classification predicts the existence of bugs in software changes. The classifier is trained using features (in the machine learning sense) extracted from the revision history of a software project, as stored in its so

Information SystemsComputer Science
2
Article|652 citations·2017
CHARMM-GUI ligand reader and modeler for CHARMM force field generation of small molecules
Seonghoon Kim, Jumin Lee, Sunhwan Jo, Charles L. Brooks, Hui Sun Lee, Wonpil Im
SJR Q1Journal of Computational Chemistry

Reading ligand structures into any simulation program is often nontrivial and time consuming, especially when the force field parameters and/or structure files of the corresponding molecules are not available. To address this problem, we have developed Ligand Reader & Modeler in CHARMM-GUI. Users can upload ligand structure information in various forms (using PDB ID, ligand ID, SMILES, MOL/MOL2/SDF file, or PDB/mmCIF file), and the uploaded structure is displayed on a sketchpad for verification

Molecular BiologyBiochemistry, Genetics and Molecular Biology
3
Article|334 citations·2011
Dealing with noise in defect prediction
Sunghun Kim, Hongyu Zhang, Rongxin Wu, Liang Gong

Many software defect prediction models have been built using historical defect data obtained by mining software repositories (MSR). Recent studies have discovered that data so collected contain noises because current defect collection practices are based on optional bug fix keywords or bug report links in change logs. Automatically collected defect data based on the change logs could include noises.

Information SystemsComputer Science
4
Article|245 citations·2008
Toward an understanding of bug fix patterns
Kai Pan, Sunghun Kim, E. James Whitehead
SJR Q1Empirical Software Engineering
Information SystemsComputer Science
5
Article|174 citations·2006
How long did it take to fix bugs?
Sunghun Kim, E. James Whitehead

The number of bugs (or fixes) is a common factor used to measure the quality of software and assist bug related analysis. For example, if software files have many bugs, they may be unstable. In comparison, the bug-fix time - the time to fix a bug after the bug was introduced - is neglected. We believe that the bug-fix time is an important factor for bug related analysis, such as measuring software quality. For example, if bugs in a file take a relatively long time to be fixed, the file may have

Information SystemsComputer Science
6
Article|113 citations·2011
Surface Scattering via Bulk Continuum States in the 3D Topological InsulatorBi2Se3
Sunghun Kim, Mao Ye, Kenta Kuroda, Y. Yamada, E. E. Krasovskii, Е. В. Чулков, K. Miyamoto, M. Nakatake, Taichi Okuda, Y. Ueda, K. Shimada, H. Namatame
SJR Q1Physical Review LettersOA

We have performed scanning tunneling microscopy and differential tunneling conductance (dI/dV) mapping for the surface of the three-dimensional topological insulator Bi(2)Se(3). The fast Fourier transformation applied to the dI/dV image shows an electron interference pattern near Dirac node despite the general belief that the backscattering is well suppressed in the bulk energy gap region. The comparison of the present experimental result with theoretical surface and bulk band structures shows t

Atomic and Molecular Physics, and OpticsPhysics and Astronomy
7
Article|104 citations·2020
CHARMM-GUI Free Energy Calculator for Absolute and Relative Ligand Solvation and Binding Free Energy Simulations
Seonghoon Kim, Hiraku Oshima, Han Zhang, Nathan R. Kern, Suyong Re, Jumin Lee, Benoı̂t Roux, Yuji Sugita, Wei Jiang, Wonpil Im
SJR Q1Journal of Chemical Theory and ComputationOA

Alchemical free energy simulations have long been utilized to predict free energy changes for binding affinity and solubility of small molecules. However, while the theoretical foundation of these methods is well established, seamlessly handling many of the practical aspects regarding the preparation of the different thermodynamic end states of complex molecular systems and the numerous processing scripts often remains a burden for successful applications. In this work, we present CHARMM-GUI <i>

Molecular BiologyBiochemistry, Genetics and Molecular Biology
8
Article|96 citations·2006
A Comparative Study of IRT Fixed Parameter Calibration Methods
Seonghoon Kim
SJR Q1Journal of Educational Measurement

This article provides technical descriptions of five fixed parameter calibration (FPC) methods, which were based on marginal maximum likelihood estimation via the EM algorithm, and evaluates them through simulation. The five FPC methods described are distinguished from each other by how many times they update the prior ability distribution and by how many EM cycles they use. Specifically, the five FPC methods included no prior weights updating and one EM cycle (NWU‐OEM) or multiple EM cycles (NW

Management Science and Operations ResearchDecision Sciences
9
Article|91 citations·2005
When functions change their names: automatic detection of origin relationships
Sunghun Kim, Kai Pan, E. James Whitehead

It is a common understanding that identifying the same entity such as module, file, and function between revisions is important for software evolution related analysis. Most software evolution researchers use entity names, such as file names and function names, as entity identifiers based on the assumption that each entity is uniquely identifiable by its name. Unfortunately names change over time. In this paper, we propose an automated algorithm that identifies entity mapping at the function lev

Information SystemsComputer Science
10
Article|85 citations·2007
Prioritizing Warning Categories by Analyzing Software History
Sunghun Kim, Michael D. Ernst

Automatic bug finding tools tend to have high false positive rates: most warnings do not indicate real bugs. Usually bug finding tools prioritize each warning category. For example, the priority of "overflow " is 1 and the priority of "jumbled incremental" is 3, but the tools 'prioritization is not very effective. In this paper, we prioritize warning categories by analyzing the software change history. The underlying intuition is that if warnings from a category are resolved quickly by developer

Information SystemsComputer Science
11
Article|84 citations·2008
A Comparison of Tests for Equality of Two or More Independent Alpha Coefficients
Seonghoon Kim, Leonard S. Feldt
SJR Q1Journal of Educational MeasurementOA

This article extends the Bonett (2003a) approach to testing the equality of alpha coefficients from two independent samples to the case of m ≥ 2 independent samples. The extended Fisher‐Bonett test and its competitor, the Hakstian‐Whalen (1976) test, are illustrated with numerical examples of both hypothesis testing and power calculation. Computer simulations are used to compare the performance of the two tests and the Feldt (1969) test (for m = 2) in terms of power and Type I error control. It

Statistics and ProbabilityMathematics
12
Article|62 citations·2014
Robust Protection from Backscattering in the Topological InsulatorBi1.5Sb0.5Te1.7Se1.3
Sunghun Kim, Shunsuke Yoshizawa, Y. Ishida, Kazuma Eto, Kouji Segawa, Yoichi Ando, Shik Shin, Fumio Komori
SJR Q1Physical Review LettersOA

Electron scattering in the topological surface state (TSS) of the topological insulator Bi1.5Sb0.5Te1.7Se1.3 was studied using quasiparticle interference observed by scanning tunneling microscopy. It was found that not only the 180° backscattering but also a wide range of backscattering angles of 100°-180° are effectively prohibited in the TSS. This conclusion was obtained by comparing the observed scattering vectors with the diameters of the constant-energy contours of the TSS, which were measu

Atomic and Molecular Physics, and OpticsPhysics and Astronomy
13
Article|55 citations·2010
The estimation of the IRT reliability coefficient and its lower and upper bounds, with comparisons to CTT reliability statistics
Seonghoon Kim, Leonard S. Feldt
SJR Q1Asia Pacific Education Review
Management Science and Operations ResearchDecision Sciences
14
Article|51 citations·2006
An Extension of Four IRT Linking Methods for Mixed‐Format Tests
Seonghoon Kim, Won‐Chan Lee
SJR Q1Journal of Educational Measurement

Under item response theory (IRT), linking proficiency scales from separate calibrations of multiple forms of a test to achieve a common scale is required in many applications. Four IRT linking methods including the mean/mean, mean/sigma, Haebara, and Stocking‐Lord methods have been presented for use with single‐format tests. This study extends the four linking methods to a mixture of unidimensional IRT models for mixed‐format tests. Each linking method extended is intended to handle mixed‐format

Management Science and Operations ResearchDecision Sciences
15
Article|49 citations·2011
A Note on the Reliability Coefficients for Item Response Model-Based Ability Estimates
Seonghoon Kim
SJR Q1Psychometrika

Assuming item parameters on a test are known constants, the reliability coefficient for item response theory (IRT) ability estimates is defined for a population of examinees in two different ways: as (a) the product-moment correlation between ability estimates on two parallel forms of a test and (b) the squared correlation between the true abilities and estimates. Due to the bias of IRT ability estimates, the parallel-forms reliability coefficient is not generally equal to the squared-correlatio

Management Science and Operations ResearchDecision Sciences

Research Areas

Information SystemsComputer Networks and CommunicationsArtificial IntelligenceElectrical and Electronic EngineeringMaterials ChemistryManagement Science and Operations Research

Sung‐Hoon Kimの研究をNubintでさらに深く

この研究室の論文をアプリで開き、AIと共に読み、要約し、引用しましょう。